498,476open jobs
16,622companies
73,829added this week
Browse all
Salary
$210k – $450k per year
Location
In office (San Francisco)
Employment
Full-Time
Overview
Company
Impact
Profile match

About AfterQuery

AfterQuery is an applied research lab curating data solutions for foundation model development. We serve every frontier AI lab with the mission of delivering the best data to power the best models. In doing so, we can make expertise that once took a lifetime to build available to anyone who needs it.

Our customers are the ones building the foundation models themselves and our work sits directly in the loop of how those systems improve. This is a rare opportunity to join a company at a defining moment in AI. We are YC's fastest unicorn, valued at $3.2 billion. We're based in San Francisco and backed by leading investors including Altos Ventures, BoxGroup, and Y Combinator and angels from Google DeepMind, OpenAI, Anthropic, Meta Superintelligence Labs, and Microsoft AI.

Why Apply

Massive Opportunity: We are YC's fastest unicorn valued at $3.2 billion and we're not slowing down.

Founding Impact: You will own and architect core infrastructure systems that power our platform from the ground up.

Equity & Growth: Competitive salary and meaningful equity. As we scale, you’ll have the opportunity to shape the engineering organization and lead major technical initiatives.

Strong Team: Our founding team has experience from Citadel Securities, Meta, Google, Silver Lake, and Morgan Stanley - work alongside world-class engineers and researchers.

Overview

AfterQuery is hiring Research Scientists to design and publish rigorous evaluations for frontier AI systems. The role spans agentic, coding, and safety evaluations, as well as expert-domain evaluations involving applied AI in healthcare, STEM, finance, and related fields. You will own evaluation development end to end and collaborate across disciplines to turn important capability gaps into rigorous public research.

Responsibilities

  • Lead the end-to-end design, validation, launch, and continuous improvement of frontier AI benchmarks.

  • Partner with researchers and domain experts to develop evaluations around meaningful model failures, gaps in existing coverage, and high-priority domains.

  • Analyze model capabilities and failure modes using rigorous experimental design and statistical methods.

  • Build reproducible evaluation systems, including harnesses, graders, and benchmark infrastructure.

  • Collaborate with researchers to post-train models and measure the resulting performance gains.

  • Communicate results through benchmark reports, technical articles, and research papers.

Required Qualifications

  • Strong record of publishing benchmarks or research papers.

  • Clear technical communication and strong scientific writing skills.

  • Commitment to experimental rigor, including baselines, ablations, statistical validity, and contamination controls.

  • Ability to take an ambiguous evaluation question from initial scoping through a reproducible public release.

  • Depth in agentic, coding, and safety evaluations or applied machine learning in an expert domain.

Preferred Qualifications

  • PhD in a related technical field.

  • Research publications at leading conferences or peer-reviewed journals.

  • Interest in multidisciplinary research and the creativity to combine methods and insights from AI, engineering, science, and other expert domains.

Company Benefits (For Eligible Employees):

  • Health Insurance: Medical, Vision, Dental

  • 401(k) with Employer Match

  • Daily Meals: Daily UberEats Stipend

  • Monthly Wellness Stipend

  • Commute Covered

We are an equal opportunity employer committed to providing a workplace free from discrimination and harassment. Employment decisions are made without regard to legally protected characteristics under applicable federal, state, or local law.

We comply with applicable pay transparency requirements and provide compensation ranges based on the position, qualifications, experience, and other relevant factors. Reasonable accommodations are available to qualified individuals with disabilities and for sincerely held religious beliefs, as required by law. This job description is intended to describe the general nature and level of work performed and is not an exhaustive list of all duties, responsibilities, qualifications, or working conditions associated with the position. We reserve the right to modify this job description as business needs change.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
498,476 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
Remote • Full-Time
Python
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
React.js
DevOps
GitHub
Apply
Remote • Full-Time
Python
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
React.js
DevOps
GitHub
Apply
Remote • Full-Time
Python
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
React.js
DevOps
GitHub
Apply
Remote • Full-Time
Python
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
React.js
DevOps
GitHub
Apply
Remote • Full-Time
Python
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
React.js
DevOps
GitHub
Apply
Corporate Counsel 10 days ago
$225k – $250k per year • In office • Full-Time • 3+ years exp • San Francisco
AI/ML
OpenAI
Anthropic
Apply
$250k – $300k per year • In office • Full-Time • 3+ years exp • San Francisco
AI/ML
OpenAI
Anthropic
Apply
Quality Control Lead 11 days ago
$150k – $275k per year • In office • Full-Time • 3+ years exp • San Francisco
AI/ML
Claude Code
RLHF
OpenAI
Anthropic
Apply
$150k – $225k per year • In office • Full-Time • 3+ years exp • Master's Degree • San Francisco
Python
MATLAB
AI/ML
Claude Code
OpenAI
Anthropic
Apply
Recruiter 12 days ago
$150k – $225k per year • In office • Full-Time • 7+ years exp • San Francisco
AI/ML
OpenAI
Anthropic
Apply
$106k – $130k per year • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • San Francisco • Chicago • Scottsdale
DevOps
GCP
Azure
CI/CD
AWS
Harbor
Platform Engineering
Incident Management
Apply
$186k – $345k per year (Estimated) • In office • Full-Time • 12+ years exp • Master's Degree • San Francisco
DevOps
Incident Management
Apply
$160k – $236k per year • Equity • Remote/Hybrid • Full-Time • 2+ years exp • Chicago • Austin • San Francisco
Apply
$63k – $122k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Chicago • San Francisco • Dallas • Boston • Philadelphia
Python
JavaScript
Java
TypeScript
SQL
C++
AI/ML
Prompt Engineering
Function Calling
AI Agents
LLM
Anthropic
Agentic Workflows
Tool Use
Frontend
Angular
React.js
DevOps
GCP
Azure
AWS
GitHub
Cybersecurity
Threat Modeling
Apply
In office • Full-Time • 12+ years exp • Associate's Degree • Atlanta • Milwaukee • Dallas • Columbus • Kirkland
Python
C++
AI/ML
Synthetic Data
Physical AI
Robotics
ROS
Gazebo
Isaac Sim
SLAM
Sensor Fusion
Motion Planning
Imitation Learning
Digital Twin
IoT
MQTT
OPC UA
Apply
See all jobs
This is one of many
498,476 more open roles from verified company boards, updated every day.