435,331open jobs
15,193companies
61,867added this week
Browse all
Salary
$214k – $280k per year
Location
In office (London)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Apollo Research is an artificial intelligence safety organisation focused on deceptive model behaviour. Its evaluations test whether frontier models scheme or hide their intentions. The laboratory publishes interpretability research and advises policymakers on model risk.

THE OPPORTUNITY

Apollo Research works with most frontier AI companies (OpenAI, Anthropic, Google, Meta, Thinking Machines and others) to test their models before deployment and collaborate on fundamental scheming research. Our coding agent security product, Watcher, is deployed in production and monitors billions of agent tokens per month across engineering teams at agent-building scale-ups and enterprise.

Security exists at Apollo to safeguard the trust frontier labs place in us and to enable that research. Our own team uses AI agents extensively across its work, which makes Apollo both a target and a testbed. We’re hiring AI Security Researchers to join the Infra & Security Team.

In this role, you will identify, research and remediate both conventional threats and the novel risks introduced by AI agents that can affect Apollo and our mission. You will redefine and work on a new class of insider risk that didn’t exist before.

RESPONSIBILITIES

  • Hold responsibility for the security of Apollo’s internal surfaces. Red-team internal software, infrastructure, and AI agent access controls/monitoring. Build realistic attack trajectories.

  • Design solutions for novel or emerging threats in the AI security space, where no existing playbook applies. Set the standard for our security posture in these areas.

  • Track adversary tactics, techniques, and campaigns relevant to Apollo's threat landscape, and translate that intelligence into tuned, high-signal detections.

  • Own each finding through to a deployed fix. Build and roll out durable controls: checks, tests, defaults, detections. Socialise them, work with engineers to implement, and hold remediation to a high bar.

KEY REQUIREMENTS

    Must haves

  • 5+ years in security roles in a hands-on technical capacity (not purely GRC/compliance). You'd need to be able to think structurally about threat modelling and failure modes. You need to be able to read code, understand infrastructure, and evaluate technical controls.

  • Direct experience with offensive security. Threat modelling, red teaming, etc. Knowledge of application or cloud. Ideally you owned or significantly contributed to the security posture of an organisation or product that handles sensitive customer data.

  • Engineering mindset. You treat security as an engineering problem. You can translate your findings into fixes and controls, such as paved roads, custom detection rules, adversarial test suites, CI/CD integrations. You prioritize automation and systems-level thinking to scale security, and you are comfortable leveraging AI to accelerate development.

  • Startup pace. You are excited about a fast-moving environment, comfortable with ambiguity and changing priorities, and willing to grind when it matters.

  • Strong written communication. This role produces a lot of artifacts (threat models, reports, failure mode documentation) and they need to be clear and precise.

  • Nice to haves

  • Experience with AI/ML systems security or LLM security.

  • Detection engineering, SOC, or incident analysis experience.

  • Familiarity with insider threat programs or insider risk frameworks.

  • Explicitly not required

  • Formal AI safety research background. We need security practitioners who can learn the AI safety context, not AI safety researchers who need to learn security.

REPRESENTATIVE PROJECTS

  • Red-team Apollo's agent sandboxes used for evals: Test whether an agent can escape isolation, exfiltrate data, detect it's being evaluated, or otherwise undermine the validity of eval results. Your findings will harden the sandbox infrastructure the research team depends on to trust its own eval results.

  • Comprehensive coding agent threat model: Map every way a coding agent with internal access: credentials, code, network, execution ability, could attack Apollo, benchmarked against what a human insider with the same access could do.

BENEFITS

  • This role offers market competitive salary, equity, and competitive benefits.

  • Salary: San Francisco: $214,000 - $280,000; London: £144,000 - £189,000. We will be looking to meaningfully raise salaries soon.

  • Our engineers effectively have an unlimited token budget. If a better result costs more compute, use it.

  • Flexible work hours and schedule

  • Unlimited vacation

  • Unlimited sick leave

  • Up to 6 months of paid parental leave

  • Comprehensive health, dental and vision insurance

  • Retirement savings with competitive employer matching (e.g. 401(k) for US employees)

  • Lunch, dinner, and snacks are provided for all employees on workdays

  • Paid work trips, including staff retreats, business trips, and relevant conferences

  • A yearly $1,000 (USD) professional development budget

  • Relocation support and visa fees (if applicable)

LOGISTICS

  • Time Allocation: Full-time

  • Location: This is an in-person role working out of our London or San Francisco office. We offer flexible working hours and some wfh arrangements.

  • Visa sponsorship: We sponsor visas in both the UK and US. Sponsorship isn't guaranteed for every role or candidate, but if we make you an offer, we'll work with you to find the right visa route.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
435,331 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
London
Sr. Software Engineer 8 hours ago
$118k – $152k per year • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Scottsdale • San Francisco
SQL
C#
C#
ASP.NET Core
DevOps
CI/CD
Git
AWS
Docker
Kubernetes
Harbor
Apply
$218k – $365k per year • In office • Full-Time • 12+ years exp • Master's Degree • San Francisco • Palo Alto • Chicago • New York • Washington
AI/ML
AI Agents
LLM
Agentforce
Apply
$95k – $205k per year (Estimated) • In office • Full-Time • PhD • London
AI/ML
AI Agents
Agentforce
Apply
$135k – $180k per year • Remote • Full-Time • 4+ years exp • Bachelor's Degree • New York • Chicago • San Francisco • Atlanta
Apex
Apex
MuleSoft
AI/ML
Model Context Protocol
AI Agents
Agentforce
Multi-Agent Systems
DevOps
Rest API
Analytics
Informatica
Master Data Management
Apply
$60k – $113k per year (Estimated) • Remote/Hybrid • Full-Time • PhD • Mexico City
AI/ML
AI Agents
Agentforce
Design
Figma
Webflow
Apply
$227k – $296k per year • In office • Full-Time • 5+ years exp • London
AI/ML
LLM
OpenAI
Anthropic
DevOps
CI/CD
Apply
$204k – $385k per year • In office • Full-Time • 5+ years exp • London
AI/ML
AI Agents
LLM
OpenAI
Anthropic
Red Teaming
Cybersecurity
MITRE ATT&CK
Threat Modeling
Apply
$182k – $238k per year • In office • Full-Time • 2+ years exp • London
Python
AI/ML
LLM
Anthropic
Red Teaming
DevOps
Vector
Apply
$204k – $385k per year • In office • Full-Time • 2+ years exp • London
Python
SQL
AI/ML
Fine-tuning
AI Agents
LLM
Apply
Security Engineer 12 days ago
$214k – $280k per year • In office • Full-Time • 5+ years exp • London
AI/ML
AI Agents
LLM
OpenAI
Anthropic
Cybersecurity
Zero Trust
Threat Modeling
Apply
$105k – $175k per year (Estimated) • In office • Full-Time • London
Apply
$66k – $177k per year (Estimated) • In office • Full-Time • Bachelor's Degree • London
Marketing
Salesforce
Apply
Editor / Speechwriter 8 hours ago
$64k – $169k per year (Estimated) • In office • Full-Time • London
Apply
$73k – $122k per year (Estimated) • In office • Full-Time • 5+ years exp • London
Analytics
Power BI
Management
Outlook
Apply
$81k – $152k per year (Estimated) • Remote/Hybrid • Full-Time • London
Analytics
Power BI
Microsoft Excel
Apply
See all jobs
This is one of many
435,331 more open roles from verified company boards, updated every day.