1,435,541open jobs
84,045companies
218,048added this week
Browse all
Salary
≈ $13k – $31k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Middle · 3+ years exp

First seen by Alion on Oct 8, 2026.

Overview
Company
Impact
Profile match
BT Group is the largest telecommunications provider in the United Kingdom, with roots in the electric telegraph company founded in 1846. It runs consumer broadband and mobile through EE, enterprise networking and the Openreach access network. The company employs tens of thousands of engineers and operates one of the biggest research sites in Britain at Adastral Park.

The QA AI Evaluator will be responsible for ensuring the quality, reliability, governance, and business effectiveness of Agentic AI solutions. The role focuses on evaluating AI agents, prompts, LLM outputs, and end-to-end workflows to ensure accurate, consistent, and trusted outcomes.

Responsibilities:

  • Define AI evaluation frameworks, standards, KPIs, and success criteria.
  • Validate agent accuracy, reasoning, reliability, workflow quality, and business outcomes.
  • Assess prompts, LLM responses, hallucination risks, and AI-generated content quality.
  • Benchmark models, tools, prompts, and agent configurations to improve performance.
  • Ensure Responsible AI, security, privacy, governance, and compliance requirements are met.
  • Provide clear insights, risks, recommendations, and scorecards to leadership and stakeholders.
  • Define and govern AI quality standards, evaluation frameworks, KPIs, and success criteria.
  • Validate Agentic AI solutions for accuracy, reasoning, reliability, and business effectiveness.
  • Assess prompts, LLM outputs, and end-to-end agent workflows to ensure quality outcomes.
  • Benchmark models, prompts, tools, and agent configurations to optimize performance.
  • Measure and improve AI quality metrics, including accuracy, relevance, completeness, and hallucination rates.
  • Identify risks, quality gaps, and opportunities for continuous improvement.
  • Ensure Responsible AI, security, privacy, governance, and compliance requirements are met.
  • Support AI adoption, user acceptance, and organizational change initiatives.
  • Collaborate with product, engineering, architecture, QA, and business stakeholders.
  • Present insights, risks, performance metrics, and strategic recommendations to leadership.

Requirements:

  • Strong experience in AI testing, quality engineering, AI/LLM evaluation, or agentic AI assurance.
  • Deep understanding of LLMs, prompt engineering, RAG, AI agents, and multi-agent architectures.
  • Expertise in defining evaluation frameworks, quality metrics, KPIs, scorecards, and benchmarking approaches.
  • Experience validating AI accuracy, reasoning, reliability, hallucinations, and business outcomes.
  • Knowledge of Responsible AI, Security, Privacy, Governance, Risk, and Compliance practices.
  • Strong analytical and problem-solving skills with experience in root cause analysis and continuous improvement.
  • Familiarity with AI evaluation tools, test automation, and data-driven quality assessment techniques.
  • Ability to collaborate effectively with product, engineering, architecture, QA, and business stakeholders.
  • Excellent communication and stakeholder management skills, including presenting insights to leadership.
  • Experience driving AI adoption, change management, and business transformation initiatives.

Essential:

  • 3-4 years of experience with strong experience in AI testing, quality engineering, AI/LLM evaluation, or agentic AI assurance.
  • Good understanding of LLMs, prompt engineering, RAG, AI agents, and multi-agent workflows.
  • Ability to define evaluation frameworks, quality metrics, KPIs, scorecards, and success criteria.
  • Hands-on experience validating AI accuracy, reasoning, reliability, hallucination risk, relevance, completeness, and business outcomes.
  • Strong analytical, problem-solving, stakeholder management, and communication skills.

Desirable:

  • Minimum 3-4 years of experience with AI evaluation tools, automation frameworks, and data-driven quality assessment techniques.
  • Exposure to benchmarking models, prompts, tools, and agent configurations to improve performance.
  • Experience supporting AI adoption, user acceptance, change management, or wider business transformation initiatives.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,435,541 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Operations
Similar stack
Same company
Bengaluru
≈ $41k – $93k per year (Estimated) • In office • Full-Time • Amsterdam
Management
Microsoft Office
Apply
Warehouse Manager 22 min ago
≈ $57k – $132k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Abeokuta
Apply
In office • Full-Time • Abuja
Apply
In office • Internship • Port Harcourt
Management
Microsoft Office
Apply
≈ $72k – $143k per year (Estimated) • Equity • Hybrid • Full-Time • 3+ years exp • Master's Degree • San Mateo
Apply
$153k – $255k per year • Remote (likely United States) • Bachelor's Degree
Python
Databases
OpenSearch
AI/ML
Model Context Protocol
Function Calling
AI Agents
DeepEval
Langfuse
Promptfoo
Ragas
LLM
RAG
Hallucination
Human-in-the-Loop
LLM Evaluation
Multi-Agent Systems
Tool Use
DevOps
OpenShift
CI/CD
AWS
Management
Agile
QA
Playwright
Apply
Lead SRE 1 day ago
$153k – $255k per year • Remote (likely United States) • PhD
Python
Bash
Databases
OpenSearch
AI/ML
vLLM
AI Agents
Langfuse
Ollama
AgentOps
LLM
RAG
Hallucination
LLMOps
Multi-Agent Systems
DevOps
OpenShift
OpenTelemetry
CI/CD
Docker
Kubernetes
Self-Healing
Incident Management
Linux
Management
Agile
Apply
≈ $20k – $48k per year (Estimated) • In office • Full-Time • 15+ years exp • Master's Degree • Bengaluru
Python
Java
SQL
Scala
Databases
Snowflake
Databricks
Apache Iceberg
AI/ML
Copilot
Spark
Claude Code
Model Context Protocol
Embeddings
AI Agents
LLM
RAG
Context Engineering
DevOps
GCP
Azure
CI/CD
AWS
Incident Management
SLI/SLO/SLA
Analytics
Collibra
Alation
Apply
≈ $21k – $39k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Moscow
Databases
Apache Kafka
AI/ML
Model Context Protocol
AI Agents
RAG
DevOps
Rest API
Management
UML
Apply
≈ $119k – $233k per year (Estimated) • Remote (United States) • Full-Time • 8+ years exp • United States
Python
Go
Java
TypeScript
Databases
Snowflake
Databricks
AI/ML
AI Agents
AgentOps
RAG
LLMOps
Human-in-the-Loop
LLM Guardrails
Tool Use
DevOps
GCP
Azure
CI/CD
AWS
FinOps
Management
ServiceNow
Apply
≈ $44k – $119k per year (Estimated) • Hybrid • Dublin
Apply
≈ $44k – $119k per year (Estimated) • Hybrid • Dublin
Apply
≈ $44k – $119k per year (Estimated) • Hybrid • Dublin
Apply
In office • Manchester
Apply
≈ $38k – $97k per year (Estimated) • In office • Bristol
Apply
Team Lead 1 day ago
≈ $15k – $35k per year (Estimated) • In office • Full-Time • 7+ years exp • Bengaluru
Apply
≈ $29k – $71k per year (Estimated) • In office • Full-Time • 12+ years exp • Bengaluru • Gurgaon
Swift
Mobile
SwiftUI
Combine
SPM
GCD
CocoaPods
Clean Architecture
Management
Agile
Apply
DevOps Engineer 1 day ago
≈ $18k – $40k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru
DevOps
Helm
Azure DevOps
Azure
CI/CD
Docker
Kubernetes
Management
Agile
Apply
≈ $30k – $53k per year (Estimated) • Hybrid • Full-Time • 9+ years exp • Bengaluru
AI/ML
AI Agents
DevOps
Incident Management
Management
ServiceNow
ITIL
ITSM
Apply
≈ $34k – $70k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree • Bengaluru
AI/ML
AI Agents
Human-in-the-Loop
LLM Guardrails
Management
Agile
Apply
See all jobs
This is one of many
1,435,541 more open roles from verified company boards, updated every day.