368,910open jobs
9,449companies
47,822added this week
Browse all
Salary
$130k – $500k per year
Location
In office (San Francisco)
Seniority
Staff · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Mercor is an artificial intelligence infrastructure company headquartered in San Francisco, California, and founded in 2023. The company operates a platform that connects a global network of domain experts, including physicians, lawyers, and engineers, with AI labs to provide high-quality data for reinforcement learning from human feedback (RLHF) and model evaluation. It develops specialized benchmarks like APEX to measure model performance on economically valuable tasks and provides enterprises with tools to monetize their workflow data for AI training.

About Mercor

Mercor's mission is to organize human intelligence to power the AI economy. We're a leading AI data company, building the layer between human expertise and frontier models. Millions of domain experts on the platform are paid over $4 million per day to train frontier AI models. Mercor's APEX benchmark family measures AI's real-world impact on professional work. Mercor Enterprise brings this same infrastructure to Fortune 500 companies: helping companies capture how their best people actually work, translating that expertise directly back into agents.

Mercor is creating a new category of work where expertise powers AI advancement. Achieving this requires an ambitious, fast-paced and deeply committed team. You’ll work alongside researchers, operators, and AI companies at the forefront of shaping the systems that are redefining society. Mercor is a profitable Series C company valued at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices.

Location Preference

SF only.

About the Role

As a Member of Technical Staff on Talent Experience, you'll build the backend systems and agent infrastructure behind every journey in Mercor's talent network.

A growing share of Mercor's core work runs on agents: systems that source, screen, match, and support experts with far less human intervention than the workflows they replace. You'll design and ship the services those agents run on-orchestration, tool interfaces, evaluation and guardrail layers, and the data models that keep agent decisions auditable and reversible. This is production agent engineering at real volume, against a network of tens of thousands of experts and the demands of the world's leading AI labs.

This is a hands-on engineering role. You'll write code, own systems end to end, and work directly with product, operations, and AI researchers to decide what agents should do and where humans still belong in the loop.

What You'll Do

  • Design, build, and own the backend services powering Mercor's agentic workflows-orchestration, tool calling, retries and fallbacks, and human-in-the-loop review.

  • Ship agents into production and iterate on them against real usage: prompts, tools, guardrails, and cost and latency tradeoffs.

  • Build the evaluation infrastructure that tells us whether an agent is actually working-offline evals, online metrics, and regression detection.

  • Design core services, data models, and APIs that scale with the talent network and stay legible as agent behavior changes underneath them.

  • Make agent decisions observable, auditable, and reversible-so failures are caught early and understood quickly.

  • Partner with product, operations, and AI researchers to decide what should be an agent, what should stay deterministic, and where a human belongs in the loop.

  • Establish backend engineering standards and patterns, and raise the technical bar of the engineers around you.

What We're Looking For

  • 5+ years of professional backend engineering experience.

  • Strong fundamentals in distributed systems, service architecture, data modeling, and API design at scale.

  • Hands-on experience building with LLMs in production-agents, tool use, retrieval, or evaluation systems-or a clear track record of learning new systems fast and shipping them.

  • Comfort with ambiguity: agent systems fail in ways traditional services don't, and the right design is rarely obvious upfront.

  • A track record of owning systems end to end, from design through production operation.

  • Fluent in using modern AI development tools (e.g. Cursor, Claude Code, GitHub Copilot, or similar) to improve engineering productivity.

  • Strong opinions, loosely held: you drive alignment through technical judgment and clear writing, not authority.

  • Excellent communication skills-able to make complex tradeoffs legible to both engineers and leadership.

  • High ownership, pragmatism, and a bias toward shipping.

Why Mercor

  • Impact: Build the agent systems that Mercor's core business runs on, at a company scaling faster than almost any in its category.

  • Ownership: Real scope over the systems you build, with a direct line to engineering leadership and authority over technical direction.

  • Learning: Work alongside world-class engineers, product leaders, and AI researchers building at the frontier of applied AI.

  • Growth: Shape how Mercor builds with agents-the standards, the patterns, and the next generation of systems.

Benefits

  • Generous equity grant vested over 4 years

  • A $20K relocation bonus (if moving to the Bay Area)

  • A $10K housing bonus (if you live within 0.5 miles of our office)

  • A $1.5K monthly stipend for meals

  • Free Equinox membership

  • Health insurance

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,910 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
Senior Data Analyst 3 hours ago
$91k – $96k per year • Equity • Remote • Full-Time • 5+ years exp • Bachelor's Degree • Ontario
Python
SQL
Databases
Snowflake
AI/ML
Anomaly Detection
Claude
Copilot
Cursor
dbt
Edge AI
DevOps
AWS
Analytics
Tableau
Marketing
Salesforce
Apply
$44k – $115k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree • Mexico
Apex
JavaScript
Apex
Lightning Web Components
Salesforce CLI
AI/ML
AI Agents
ChatGPT
Claude
Cursor
Prompt Engineering
Agentforce
LLM Guardrails
Model Context Protocol
Frontend
Web Components
DevOps
CI/CD
Git
Rest API
GitHub
Cybersecurity
HIPAA
Marketing
Salesforce
Apply
$197k – $374k per year (Estimated) • In office • Full-Time • 8+ years exp • PhD • New York
AI/ML
AI Agents
Claude
LangChain
OpenAI
Vertex AI
Management
n8n
Zapier
Apply
$108k – $195k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • United States
Python
AI/ML
AI Agents
Amazon SageMaker
AWS Bedrock
AWS Bedrock AgentCore
LLM
DevOps
AWS
CI/CD
CloudFormation
Docker
IAM
Kubernetes
Platform Engineering
Terraform
Cybersecurity
FedRAMP
Apply
$113k – $189k per year • In office • Full-Time • Bachelor's Degree • Barcelona
AI/ML
Copilot
EU AI Act
Human-in-the-Loop
NIST AI RMF
DevOps
GitHub
Apply
$130k – $500k per year • Equity • In office • Full-Time • 3+ years exp • San Francisco
TypeScript
JavaScript
AI/ML
Claude
Claude Code
Copilot
Cursor
LLM
OpenAI Codex
Frontend
React.js
DevOps
GitHub
Apply
$200k – $500k per year • Equity • In office • Full-Time • PhD • San Francisco
AI/ML
LLM
NLP
LLM Evaluation
Post-training
AI Agents
Apply
$130k – $500k per year • Equity • In office • Full-Time • New York
Python
TypeScript
JavaScript
Python
Django
FastAPI
Pydantic
Databases
DuckDB
MySQL
PostgreSQL
Redis
Snowflake
AI/ML
LangChain
LangGraph
LangSmith
LLM
AI Agents
Human-in-the-Loop
LLM Guardrails
Frontend
Next.js
React.js
Tailwind CSS
DevOps
Datadog
Kubernetes
Apply
$130k – $500k per year • Equity • In office • Full-Time • New York
Go
Python
Rust
AI/ML
Synthetic Data
Post-training
Apply
$130k – $500k per year • Equity • In office • Full-Time • 2+ years exp • New York
Apply
$89k – $193k per year (Estimated) • In office • Full-Time • 3+ years exp • High School Diploma • San Francisco
Apply
$190k – $270k per year • Remote/Hybrid • Full-Time • 7+ years exp • San Francisco
SQL
AI/ML
LLM
Recommender Systems
Analytics
A/B Testing
Marketing
Salesforce
Apply
$170k – $220k per year • Equity 1–2.8% • In office • Full-Time • 3+ years exp • San Francisco
Python
SQL
Python
Django
AI/ML
AI Agents
Context Engineering
LLM
LLM Evaluation
RAG
Apply
$170k – $230k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • New York • San Francisco
Go
Java
Kotlin
Python
Scala
Databases
Apache Kafka
Databricks
Delta Lake
Snowflake
AI/ML
Dagster
dbt
Flink
Spark
Knowledge Graph
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
See all jobs
This is one of many
368,910 more open roles from verified company boards, updated every day.