368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$180k – $500k per year
Location
In office (San Francisco)
Employment
Full-Time
Overview
Company
Impact
Profile match
Mercor is an artificial intelligence infrastructure company headquartered in San Francisco, California, and founded in 2023. The company operates a platform that connects a global network of domain experts, including physicians, lawyers, and engineers, with AI labs to provide high-quality data for reinforcement learning from human feedback (RLHF) and model evaluation. It develops specialized benchmarks like APEX to measure model performance on economically valuable tasks and provides enterprises with tools to monetize their workflow data for AI training.

About Mercor

Mercor's mission is to organize human intelligence to power the AI economy. We're a leading AI data company, building the layer between human expertise and frontier models. Millions of domain experts on the platform are paid over $4 million per day to train frontier AI models. Mercor's APEX benchmark family measures AI's real-world impact on professional work. Mercor Enterprise brings this same infrastructure to Fortune 500 companies: helping companies capture how their best people actually work, translating that expertise directly back into agents.

Mercor is creating a new category of work where expertise powers AI advancement. Achieving this requires an ambitious, fast-paced and deeply committed team. You’ll work alongside researchers, operators, and AI companies at the forefront of shaping the systems that are redefining society. Mercor is a profitable Series C company valued at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices.

About the Role

You’ll work with large enterprises to capture their data and transform it into high-fidelity RL environments for capability evaluations and training datasets for frontier labs. We focus on pushing the frontier of world-building, verifier engineering, and more alongside our partners.

Your goal will be to automate the process of building evals for real work in the economy.

What You'll Do

  • Ship models for workflow extraction, classification, and grading.

  • Engineer autonomous task refinement processes which distill data taste into pipelines.

  • Deliver data to customers and deploy into real engagements.

  • Help define the future of agentic transformation for enterprises around the world.

  • Deeply learn about the intricacies of enterprises through building evaluations for all aspects of work.

  • Build end-to-end environments for labs & enterprises by platformizing sandbox app clones, load real data into the sandboxes, build prompts from real workflows, and write verifiers leveraging enterprise expertise & golden outputs.

  • Systematize the production of environments to scale throughput while maintaining high-quality worlds and verifiers.

What We're Looking For

  • Prior experience shipping environments - you’ve contributed to an OSS framework, built environments at previous companies, or worked on agentic evaluations.

  • Strong full-stack engineering skills - you’ll be responsible for everything from infrastructure to app code to analytics

  • Bias to action - this team is focused on shipping evals, not just philosophizing about them.

  • Curiosity - being biased towards understanding and digging deep into model behavior and actually looking at the data.

  • Sweat the details that make a simulation indistinguishable from the real thing and have systems-level thinking skills that allow you to scale up quality.

Nice to Have

  • Experience with Temporal, Modal, or similar orchestration/compute services

  • Experience with synthetic data generation for frontier models.Past work auditing and scrutinizing industry-standard evaluations

Benefits

  • Semi-annual performance bonus structure

  • Generous equity grant vested over 4 years

  • Up to $15k Relocation bonus

  • $10K housing bonus (if you live within 0.5 miles of our office)

  • $1.5K monthly stipend for meals

  • Free Equinox membership

  • $200 monthly laundry reimbursement

  • $200 monthly personal wellness reimbursement

  • Health, Dental, Vision insurance

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$26k – $66k per year (Estimated) • In office • Full-Time • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$26k – $66k per year (Estimated) • In office • Full-Time • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$150k – $250k per year • In office • Full-Time • 2+ years exp • San Francisco
Python
AI/ML
AI Agents
LLM
Reinforcement Learning
Synthetic Data
Post-training
DevOps
Docker
Apply
$200k – $500k per year • Equity • In office • Full-Time • PhD • San Francisco
AI/ML
LLM
NLP
LLM Evaluation
Post-training
AI Agents
Apply
$130k – $500k per year • Equity • In office • Full-Time • New York
Python
TypeScript
JavaScript
Python
Django
FastAPI
Pydantic
Databases
DuckDB
MySQL
PostgreSQL
Redis
Snowflake
AI/ML
LangChain
LangGraph
LangSmith
LLM
AI Agents
Human-in-the-Loop
LLM Guardrails
Frontend
Next.js
React.js
Tailwind CSS
DevOps
Datadog
Kubernetes
Apply
$130k – $500k per year • Equity • In office • Full-Time • New York
Go
Python
Rust
AI/ML
Synthetic Data
Post-training
Apply
$130k – $500k per year • Equity • In office • Full-Time • New York
Python
Databases
PostgreSQL
DevOps
AWS
Apply
$130k – $500k per year • Equity • In office • Full-Time • San Francisco
Go
Python
Databases
Apache Kafka
ElasticSearch
PostgreSQL
Redis
DevOps
Kubernetes
Terraform
Apply
$222k – $277k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Francisco
DevOps
CI/CD
Immutable Infrastructure
Apply
$70k – $196k per year • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
Databases
Databricks
Google BigQuery
SAP HANA
Snowflake
AI/ML
Knowledge Graph
DevOps
Azure
Apply
$70k – $196k per year • Remote/Hybrid • Full-Time • 5+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
DevOps
SLI/SLO/SLA
Apply
$70k – $206k per year • In office • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
AI/ML
AI Agents
Apply
$293k – $385k per year • In office • Full-Time • San Francisco
AI/ML
OpenAI
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.