368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$130k – $500k per year
Location
In office (San Francisco)
Employment
Full-Time
Overview
Company
Impact
Profile match
Mercor is an artificial intelligence infrastructure company headquartered in San Francisco, California, and founded in 2023. The company operates a platform that connects a global network of domain experts, including physicians, lawyers, and engineers, with AI labs to provide high-quality data for reinforcement learning from human feedback (RLHF) and model evaluation. It develops specialized benchmarks like APEX to measure model performance on economically valuable tasks and provides enterprises with tools to monetize their workflow data for AI training.

About Mercor

Mercor's mission is to organize human intelligence to power the AI economy. We're a leading AI data company, building the layer between human expertise and frontier models. Millions of domain experts on the platform are paid over $4 million per day to train frontier AI models. Mercor's APEX benchmark family measures AI's real-world impact on professional work. Mercor Enterprise brings this same infrastructure to Fortune 500 companies: helping companies capture how their best people actually work, translating that expertise directly back into agents.

Mercor is creating a new category of work where expertise powers AI advancement. Achieving this requires an ambitious, fast-paced and deeply committed team. You’ll work alongside researchers, operators, and AI companies at the forefront of shaping the systems that are redefining society. Mercor is a profitable Series C company valued at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices.

About the Role

We're looking for a strong engineer who can build agentic products that scale. You will work with:

  • Backend: Python, FastAPI, Django, Pydantic

  • Frontend: Next.js, React, TypeScript, Tailwind

  • Data: PostgreSQL, MySQL, Snowflake, DuckDB, Redis

  • Orchestration/Infra: Kubernetes, Temporal, Modal, Woz

  • Agents/LLM: LangGraph, LangChain, FastMCP, Harbor, NemoGym

  • Observability: Datadog, PostHog, LangSmith

At the end of the process, you’ll be team-matched to where you can have the most impact, on one of the following:

  • Automation - We build intelligent systems and agents that automate operational work at scale-handling talent management, decision-making insights, and knowledge access-so humans can focus on higher-level thinking.This is a newly formed, CEO-facing team focused on 0→1 product development, with a strong emphasis on business impact. The work is highly cross-functional, touching nearly every system across the company.

  • Studio - We own Mercor’s evaluation system & annotation platform for RL environments and tasks. We build harnesses, agents, verifiers, and the end-to-end infrastructure for producing frontier data. Our mission is to scale up high quality RL environments/tasks and expand their capabilities. We work closely with researchers at frontier AI labs to jointly shape the direction of next-generation models.

What You’ll Do

  • Own agentic features end-to-end - from scoping with researchers/ops partners through implementation, launch, and iteration on real customer feedback.

  • Design and ship LLM agents, harnesses, and verifiers - including the tools, prompts, and policies that make them reliable.

  • Build the Python/FastAPI services and Temporal/Modal pipelines that orchestrate agent runs, human-in-the-loop review and iterations.

  • Build state of the art RL environments that expand the capabilities of frontier agents, with realistic enterprise apps, simulated coworkers, and rich company data rooms that support tasks spanning hours to days.

  • Build tooling that turns agent trajectories into insight, from statistical analysis to automated failure mode detection.

  • Build and refine the full-stack surfaces and data infrastructure - craft Next.js/React interfaces where operators and experts work with agents, evolve data models to give agents the structured context and audit trails they need.

  • Define agent quality and drive continuous improvement - build evals, instrument traces, analyze failure modes, and iterate on prompts, tools, and guardrails while raising the bar for reliability, cost, latency, and UX.

  • Partner cross-functionally to shape agent autonomy - work with Product, Design, Research and Ops to draw the lines between autonomous action, propose-and-approve flows, and human-in-the-loop decisions.

Why Mercor

  • Impact: Your work powers how the world’s leading AI labs train and test their models.

  • Learning: Get early insights into frontier model capabilities months before the market.

  • Growth: Work on both infrastructure and research-adjacent projects with fast paths to ownership.

Benefits

  • Bi-annual performance bonus structure

  • Generous equity grant vested over 4 years

  • Up to $15k Relocation bonus

  • $10K housing bonus (if you live within 0.5 miles of our office)

  • $1.5K monthly stipend for meals

  • Free Equinox membership

  • $200 monthly laundry reimbursement

  • $200 monthly personal wellness reimbursement

  • Health, Dental, Vision insurance

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$23k – $57k per year (Estimated) • Remote • Full-Time • 6+ years exp
Python
Databases
OpenSearch
DevOps
AIOps
ArgoCD
AWS
Azure
CI/CD
Datadog
Docker
Dynatrace
GCP
GitHub Actions
Grafana
Incident Management
Jaeger
Jenkins
Kubernetes
New Relic
OpenTelemetry
Platform Engineering
Prometheus
SLI/SLO/SLA
Splunk
Terraform
GitHub
Apply
Cloud Engineer (AWS) 6 hours ago
In office • Full-Time • 5+ years exp • Bachelor's Degree • Dalian
Python
DevOps
AWS
CI/CD
CloudFormation
Docker
FinOps
GitHub Actions
Kubernetes
Terraform
GitHub
Apply
$20k – $50k per year (Estimated) • In office • Full-Time • Tomsk
C++
Go
Java
Kotlin
Databases
Apache Kafka
ClickHouse
ElasticSearch
DevOps
Jaeger
Kubernetes
Prometheus
SLI/SLO/SLA
Apply
$20k – $50k per year (Estimated) • In office • Full-Time • Perm
C++
Go
Java
Kotlin
Databases
Apache Kafka
ClickHouse
ElasticSearch
DevOps
Jaeger
Kubernetes
Prometheus
SLI/SLO/SLA
Apply
$19k – $47k per year (Estimated) • Remote • Full-Time • Perm
C#
C#
ASP.NET Core
Dapper
Entity Framework Core
Databases
Apache Kafka
ClickHouse
ElasticSearch
PostgreSQL
RabbitMQ
Redis
DevOps
CI/CD
Docker
Docker Compose
GitHub Actions
Kubernetes
TeamCity
GitHub
GitLab
Apply
$200k – $500k per year • Equity • In office • Full-Time • PhD • San Francisco
AI/ML
LLM
NLP
LLM Evaluation
Post-training
AI Agents
Apply
$130k – $500k per year • Equity • In office • Full-Time • New York
Python
TypeScript
JavaScript
Python
Django
FastAPI
Pydantic
Databases
DuckDB
MySQL
PostgreSQL
Redis
Snowflake
AI/ML
LangChain
LangGraph
LangSmith
LLM
AI Agents
Human-in-the-Loop
LLM Guardrails
Frontend
Next.js
React.js
Tailwind CSS
DevOps
Datadog
Kubernetes
Apply
$130k – $500k per year • Equity • In office • Full-Time • New York
Go
Python
Rust
AI/ML
Synthetic Data
Post-training
Apply
$130k – $500k per year • Equity • In office • Full-Time • 2+ years exp • New York
Apply
$130k – $500k per year • Equity • In office • Full-Time • New York
Python
Databases
PostgreSQL
DevOps
AWS
Apply
$180k – $210k per year • Equity • In office • Full-Time • San Francisco
Node JS
JavaScript
Databases
PostgreSQL
DevOps
PagerDuty
Web3
TRM Labs
Management
Slack
Apply
$252k – $335k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco
AI/ML
ChatGPT
Human-in-the-Loop
OpenAI
OpenAI Codex
DevOps
SLI/SLO/SLA
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
$160k – $283k per year • Equity • In office • 5+ years exp • San Francisco
AI/ML
AI Agents
Apply
$185k – $385k per year • Remote/Hybrid • Full-Time • 5+ years exp • San Francisco
JavaScript
Python
Databases
MySQL
PostgreSQL
AI/ML
OpenAI
Frontend
React.js
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.