368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$136k – $252k per year
Location
Remote/Hybrid (London, United Kingdom)
Seniority
Architect · 10+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Novartis is a Swiss multinational pharmaceutical company dedicated to the research, development, and manufacturing of innovative medicines. The organization focuses on addressing high-unmet patient needs across key therapeutic areas, including oncology, immunology, neuroscience, and cardiovascular-renal-metabolic diseases. Through its advanced technology platforms - such as gene and cell therapy, radioligand therapy, and xRNA - the firm aims to reimagine medicine to improve and extend lives worldwide.

Salary Range:

£100,240.00 - £186,160.00

Band

Level 6

Job Description Summary

Director: AI Systems Reliability, Testing & Performance

#LI-Hybrid

Location: London

Novartis is unable to offer relocation support for this role: please only apply if this location is accessible for you.

The Director, AI Systems Reliability, Testing & Performance is a senior technical AI leadership role within Data Science & AI, responsible for defining how AI systems across Novartis Development are evaluated, tested, validated, benchmarked, monitored, and continuously improved throughout their lifecycle.

This role owns the technical evidence required to determine whether AI systems are reliable, robust, secure, and fit-for-use across Novartis Development. The role defines common approaches for AI evaluation, benchmarking, testing, validation, monitoring, and production-readiness across agentic AI systems, predictive models, retrieval systems, digital twins, and other AI capabilities.

The Director provides technical leadership in AI evaluation science, reliability engineering, validation, adversarial testing, and performance assessment. Key areas of focus include model and agent evaluation, benchmarking, failure-mode analysis, drift detection, digital twin validation, AI red teaming, observability, traceability, and technical evidence generation supporting regulated and business-critical AI systems.

This is not a Governance, Product Management, PMO, or infrastructure operations role. Governance owns policies, risk frameworks, and approval processes. Product teams own roadmaps, adoption, and value realization. DDIT and engineering teams own platforms, infrastructure, and operational services.

Success means Development AI systems are supported by objective evidence demonstrating how they perform, where they fail, and whether they remain fit-for-use over time.

Job Description

Major Accountabilities

AI Evaluation & Benchmarking

  • Define evaluation methodologies for AI systems, models, agents, digital twins, and simulation environments across Development.
  • Establish benchmark suites, evaluation datasets, and testing harnesses used across AI initiatives.
  • Define objective measures for quality, reliability, robustness, grounding, agent effectiveness, and task success.
  • Ensure evaluation approaches remain scientifically rigorous, reproducible, and comparable across AI systems.
  • Build a common evidence framework for assessing AI capabilities, limitations, and fitness-for-use.

AI Testing & Validation

  • Establish approaches for hallucination testing, failure-mode analysis, robustness testing, and behavioral validation.
  • Define validation methodologies for agentic systems, digital twins, simulation environments, and human-in-the-loop workflows.
  • Develop production readiness criteria for AI systems operating in Development environments.
  • Support technical validation activities required for GxP-relevant and regulated AI systems.
  • Ensure AI systems are tested under realistic operating conditions and supported by objective validation evidence.

Reliability, Monitoring & Performance

  • Define how AI system reliability, degradation, drift, and operational performance are measured over time.
  • Establish monitoring requirements and performance assessment frameworks across Development AI systems.
  • Partner with engineering teams to ensure required evaluation, monitoring, tracing, and observability signals are available.
  • Define methods for identifying, diagnosing, and assessing AI system failures.
  • Promote evidence-based improvement of deployed AI capabilities.

AI Security & Adversarial Testing

  • Define approaches for AI red teaming, adversarial testing, and security evaluation.
  • Establish testing methodologies for prompt injection, jailbreaks, retrieval attacks, tool misuse, and agent manipulation scenarios.
  • Assess AI-specific vulnerabilities and resilience risks in collaboration with cybersecurity teams.
  • Ensure security testing forms part of AI system validation and production-readiness assessments.

Technical Evidence & AI Performance Intelligence

  • Establish portfolio-wide approaches for capturing technical evidence related to AI quality, reliability, robustness, and performance.
  • Define standard reporting for benchmark results, validation findings, monitoring signals, and reliability assessments.
  • Create transparency into AI system strengths, limitations, and failure modes.
  • Generate technical evidence supporting validation, audit, inspection, and governance activities.
  • Enable objective decisions on AI system readiness, reliability, and fitness-for-use.

Key Performance Indicators

· Percentage of AI systems meeting assurance standards.

· Coverage of evaluation and monitoring across AI portfolio.

· Reliability and availability of business-critical AI solutions.

· Audit and compliance readiness.

· Time to identify and resolve AI performance issues.

· Percentage of AI initiatives with measurable business outcomes.

· Leadership confidence in AI performance reporting.

Minimum Requirement: Work Experience

  • 10+ years in AI, machine learning, software engineering, quality, risk, validation, or technology governance.
  • Experience deploying or overseeing business-critical AI systems.
  • Experience establishing governance, controls, quality standards, or operational frameworks.
  • Experience defining evaluation approaches and performance metrics.
  • Experience managing operational risk in regulated environments.
  • Experience partnering with infrastructure and platform teams.

ewards

At Novartis, we’re committed to reimagining medicine together - and rewarding the people who make it happen.

The rewards of being part of our team go far beyond base pay and incentives. We also offer a variety of competitive benefits in kind to help you thrive personally and professionally, such as insurance plans, retirement plans, wellbeing resources and global recognition programs. In addition, we provide flexible and hybrid working options, where possible, and a minimum of 14 weeks paid parental leave.

Expected Annual Base Salary Range for role:

  • UK: 100,240.00 - 186,160.00 GBP Annual

The salary offered is determined based on gender-neutral objectives, such as relevant skills, competencies and experience in accordance with the Novartis pay setting policy and upon joining Novartis will be reviewed periodically.

In addition to your base salary, you may be eligible for a performance-based bonus depending on certain performance parameters. Further details will be provided during the application process.

Pay equity is a fundamental principle of our employment policy and reflects our commitment to create a diverse, equitable and inclusive environment that treats all employees with dignity and respect, as outlined in our Code of Ethics.

Read our brochure to learn more about our global total rewards offering: https://www.novartis.com/sites/novartis_com/files/novartis-life-handbook.pdf

Note: Benefits and compensation may vary by country and are subject to local legal requirements, including provisions of collective bargaining agreements where applicable. A full overview of your compensation package, including any relevant collective bargaining agreement details applicable to your role based on your employment location and Novartis employer entity, will be communicated separately to you during the application process.

Commitment to Diversity and Inclusion / EEO paragraph:

Novartis is committed to building an outstanding, inclusive work environment and diverse teams’ representative of the patients and communities we serve.

Why Novartis: Helping people with disease and their families takes more than innovative science. It takes a community of smart, passionate people like you. Collaborating, supporting and inspiring each other. Combining to achieve breakthroughs that change patients’ lives. Ready to create a brighter future together? https://www.novartis.com/about/strategy/people-and-culture

Skills Desired

Artificial Intelligence (AI), Business Value Creation, Change Management, Curious Mindset, Data Governance, Data Literacy, Data Quality, Data Science, Data Visualization, Deep Learning, Learning Agility, Machine Learning (ML), Machine Learning Algorithms, Mentorship, Stakeholder Engagement, Statistical Analysis, Time Series Analysis
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
London
$140k – $225k per year • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Seattle
TypeScript
Python
Python
FastAPI
Pydantic
Databases
pgvector
PostgreSQL
Redis
AI/ML
Fine-tuning
LangGraph
LLM
RAG
LangChain
AutoGen
CrewAI
Knowledge Distillation
NLP
Reranking
Human-in-the-Loop
LLM Evaluation
LLM Guardrails
OCR
Structured Outputs
AI Agents
Function Calling
Speech Recognition
DevOps
Docker
Terraform
AWS
Azure
Bicep
Robotics
Digital Twin
Marketing
Salesforce
Apply
In office • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$20k – $49k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
Python
Ruby
SQL
Databases
Amazon Neptune
Neo4j
AI/ML
Hallucination
LangChain
LangGraph
LLM
Model Context Protocol
Spark
AI Agents
LLM Guardrails
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
GitHub Actions
GitLab CI
Jenkins
Kubernetes
Rest API
Terraform
GitHub
GitLab
QA
Playwright
Postman
Selenium
Swagger
Apply
$26k – $66k per year (Estimated) • In office • Full-Time • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$26k – $66k per year (Estimated) • In office • Full-Time • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$25k – $56k per year (Estimated) • In office • Full-Time • Brazil
Analytics
Power BI
Marketing
Salesforce
Apply
$65k – $121k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Ljubljana
AI/ML
Computer Vision
Apply
$157k – $283k per year • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • Hanover
Apex
Apex
Salesforce Data Cloud
Databases
Snowflake
AI/ML
LLM
Agentforce
OpenAI
Marketing
Salesforce
Apply
$153k – $283k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • Hanover
Apply
$43k – $110k per year (Estimated) • In office • Contractor • 3+ years exp • Bachelor's Degree • Mexico
Analytics
QlikSense
Apply
$77k – $148k per year (Estimated) • Equity • In office • Master's Degree • London
JavaScript
Python
Scala
AI/ML
AI Agents
DevOps
GitHub
Apply
$87k – $159k per year (Estimated) • In office • Contractor • 5+ years exp • London • Stockholm
Design
Figma
Marketing
Zendesk
Apply
$105k – $204k per year (Estimated) • In office • Full-Time • London
Python
SQL
Databases
Snowflake
AI/ML
Dagster
dbt
Analytics
A/B Testing
Apply
$27k – $61k per year (Estimated) • In office • Full-Time • 5+ years exp • Pune • London
Analytics
Power BI
Tableau
Marketing
Salesforce
Apply
$34k – $85k per year (Estimated) • In office • Full-Time • Bachelor's Degree • London
AI/ML
AI Agents
Edge AI
QA
Appium
Cucumber
Cypress
Selenium
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.