<?xml version="1.0" encoding="utf-8"?>
<rss version="2.0">
  <channel>
    <title>Ai Evaluation Engineer Jobs RSS Feed</title>
    <link>https://jobs.co.uk/jobs-results?Keyword=Ai%20Evaluation%20Engineer&amp;RadiusMiles=10</link>
    <description>RSS feed for Ai Evaluation Engineer Jobs.</description>
    <language>en-gb</language>
    <lastBuildDate>Tue, 22 Sep 2026 21:23:25 GMT</lastBuildDate>
    <item>
      <title>AI Evaluation Engineer - TXP</title>
      <link>https://jobs.co.uk/job/ai-evaluation-engineer-txp--ef099cfe-3a30-47e5-8aa0-06db1b3ccd59</link>
      <guid>https://jobs.co.uk/job/ai-evaluation-engineer-txp--ef099cfe-3a30-47e5-8aa0-06db1b3ccd59</guid>
      <pubDate>Mon, 21 Sep 2026 11:50:54 GMT</pubDate>
      <description>Location: London | Salary: 700.00-700.00 Daily | Type: Contract | AI Evaluation EngineerRate / Duration / Location£700-£750/day Inside IR35 Duration: 4 months Clearance: BPSS + SC eligible Location: London, Bristol or Manchester Hybrid: 2 days onsite per week.OverviewWe''re looking for an experienced AI Evaluation Engineer to join a specialist Public Sector Technology team  click apply for full job details</description>
      <category>Contract</category>
    </item>
    <item>
      <title>AI Evaluation Engineer - SmartSourcing Ltd</title>
      <link>https://jobs.co.uk/job/ai-evaluation-engineer-smartsourcing-ltd--db0a9928-30ca-4800-be82-224b6c6b25b5</link>
      <guid>https://jobs.co.uk/job/ai-evaluation-engineer-smartsourcing-ltd--db0a9928-30ca-4800-be82-224b6c6b25b5</guid>
      <pubDate>Thu, 17 Sep 2026 11:52:31 GMT</pubDate>
      <description>Location: London | Salary: 700.00-700.00 Daily | Type: Contract | AI Evaluation Engineer / AI Assurance Engineer / AI QA Engineer / AI Governance Engineer  4 Months 750 per day (Inside IR35)  Location: London, Bristol or Manchester (2 days per week onsite)  Security Clearance: Active SC Clearance preferred / eligible and willing to undergo SC Clearance.  Alternative backgrounds may include SDET, AI Evaluation Engineer, AI Assurance Engineer, Prompt Engineer, Harness Engineer or AI Quality Engineer.  This is a hands-on engineering role for someone with deep expertise in AI/ML evaluation, testing and quality assurance, rather than a cloud, infrastructure or platform engineer. You will play a key role in establishing robust evaluation frameworks for production-grade AI systems, helping assure the quality, safety and effectiveness of data-driven and AI-enabled products.  Responsibilities Design, develop and deploy AI evaluation and risk management tooling. Build and maintain evaluation frameworks for AI, ML and LLM-based products. Create and develop code solutions, harnesses and testing capabilities from the ground up using Python. Define and execute evaluation strategies across model, agentic and application layers. Assess AI system quality, perform...</description>
      <category>Contract</category>
    </item>
    <item>
      <title>AI Evaluation Engineer - TXP</title>
      <link>https://jobs.co.uk/job/ai-evaluation-engineer-txp--ad92f809-a14c-4902-afb4-5d30c3ec8949</link>
      <guid>https://jobs.co.uk/job/ai-evaluation-engineer-txp--ad92f809-a14c-4902-afb4-5d30c3ec8949</guid>
      <pubDate>Tue, 15 Sep 2026 23:00:00 GMT</pubDate>
      <description>Location: London | Salary: &amp;pound;700 - &amp;pound;750/day | Type: Contract | AI Evaluation Engineer    Rate / Duration / Location   -700750/day Inside IR35 | Duration: 4 months | Clearance: BPSS + SC eligible | Location: London, Bristol or Manchester | Hybrid: 2 days onsite per week.   Overview   We''re looking for an experienced AI Evaluation Engineer to join a specialist Public Sector Technology team. This hands-on role focuses on building evaluation frameworks, tooling and harnesses for AI systems, particularly LLM and agentic AI solutions. The role requires ownership, experimentation and rapid delivery.   Key Responsibilities    Design evaluation frameworks and tooling  Develop evaluation harnesses  Evaluate model and agentic AI systems  Define metrics and methodologies  Build repeatable evaluation processes  Investigate AI failures  Work directly with government teams  Prototype and test solutions  Present findings and recommendations  Contribute to strategic direction.    What We Are Looking For    Strong software engineering experience  Experience building and evaluating AI systems  Understanding of LLMs and agentic AI  Experience with AI evaluation frameworks  Ability to code independently  Strong problem-solving and communication skills  Comfortabl...</description>
      <category>Contract</category>
    </item>
    <item>
      <title>AI Evaluation Engineer - SmartSourcing Ltd</title>
      <link>https://jobs.co.uk/job/ai-evaluation-engineer-smartsourcing-ltd--cec58bbf-1cb2-4bba-aac9-a8ecf1033062</link>
      <guid>https://jobs.co.uk/job/ai-evaluation-engineer-smartsourcing-ltd--cec58bbf-1cb2-4bba-aac9-a8ecf1033062</guid>
      <pubDate>Tue, 15 Sep 2026 23:00:00 GMT</pubDate>
      <description>Location: London | Salary: &amp;pound;700 - &amp;pound;750/day | Type: Contract | AI Evaluation Engineer / AI Assurance Engineer / AI QA Engineer / AI Governance Engineer   4 Months | -750 per day (Inside IR35)  Location: London, Bristol or Manchester (2 days per week onsite)  Security Clearance: Active SC Clearance preferred / eligible and willing to undergo SC Clearance.  Alternative backgrounds may include-SDET, AI Evaluation Engineer, AI Assurance Engineer, Prompt Engineer, Harness Engineer or AI Quality Engineer.  This is a hands-on engineering role for someone with deep expertise in AI/ML evaluation, testing and quality assurance, rather than a cloud, infrastructure or platform engineer. You will play a key role in establishing robust evaluation frameworks for production-grade AI systems, helping assure the quality, safety and effectiveness of data-driven and AI-enabled products.   Responsibilities  Design, develop and deploy AI evaluation and risk management tooling. Build and maintain evaluation frameworks for AI, ML and LLM-based products. Create and develop code solutions, harnesses and testing capabilities from the ground up using Python. Define and execute evaluation strategies across model, agentic and application layers. Assess AI system quality, p...</description>
      <category>Contract</category>
    </item>
  </channel>
</rss>