793,143open jobs
50,545companies
124,108added this week
Browse all
Salary
≈ $176k – $319k per year (Estimated)
Location
In office (San Francisco)
Seniority
Senior · 5+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 25, 2026. First seen by Alion on May 18, 2026. Maker Maker AI scores B on the Alion truth index.

Overview
Company
Impact
Profile match

ABOUT THE COMPANY

We're building autonomous research agents for recursive self-improvement (multi-agent systems that propose, run, and analyze machine learning experiments). We're a small team based in San Francisco, on-site

ABOUT THE ROLE

You'll be researching the agents at the core of our work: multi-agent systems that conduct automated machine learning research and discovery. You'll design how these agents plan, decompose problems, choose what to try next, evaluate their own outputs, and recover from mistakes.

This is a deeply open-ended research role. The benchmarks for agents that do real research don't exist yet, and inventing them is part of the job. You'll move between method design, careful experimentation, building evaluation frameworks, and shipping into production. Real autonomy, real ownership, and the corresponding responsibility for choosing well.

WHAT YOU'LL DO

  • Design methods that improve how our agents plan, decompose tasks, use tools, manage context, and recover from failures across long-horizon research workflows

  • Develop multi-agent coordination patterns: how multiple agents share context, divide labor, supervise each other, and combine their outputs

  • Build and maintain evaluation frameworks for agent capability on open-ended tasks (the kind where the right answer isn't pre-specified)

  • Run rigorous experiments to characterize what works, what doesn't, and why: controls, ablations, statistical significance

  • Co-design agent architectures with engineering teammates; ship the most promising methods into production

  • Read deeply across the agentic ML, planning, RL, and tool-use literature; bring useful work from outside in

  • Share findings internally so the rest of the team builds on them

  • Help shape research direction across the team: agentic research taste compounds when discussed openly

WHAT WE'RE LOOKING FOR

  • Strong track record of ML research with focus on agents, RL, LLMs, planning, tool use, or multi-agent systems

  • 5+ years of hands-on research experience in industry or academia

  • Comfort designing experiments and running them end-to-end at scale

  • Track record of building evaluation frameworks for capabilities that aren't easily benchmarked

  • Bias toward shipping research, not handing it off

  • Strong written communication: you can compress a result into a paragraph that changes what someone else does next

  • Comfort with ambiguity: open-ended problems without fixed benchmarks are the work, not a frustration

  • Published research at NeurIPS, ICML, ICLR, COLM, RLC, or comparable venues

NICE TO HAVE

  • PhD in ML, statistics, CS, or adjacent

  • Published research on agentic systems, tool use, long-horizon planning,

  • multi-agent coordination, or self-improvement methods

  • Open-source contributions in the agentic ML ecosystem (coding agents,

  • research assistants, autonomous workflows)

  • Experience with reasoning models, chain-of-thought / scratchpad methods,

  • or supervised fine-tuning for agentic behaviors

  • Background in evaluation methodology for capabilities that don't have

  • established benchmarks

THIS ROLE IS PROBABLY NOT FOR YOU IF

  • You want to focus on a single stable benchmark: our agents work on open-ended problems and the targets shift

  • You prefer to keep research paper-only; these agents need to actually work- You'd rather work alone than share research taste openly with a small team

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
793,143 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
San Francisco
$143k – $193k per year • Equity • In office • Full-Time • 4+ years exp • Master's Degree • Seattle
Python
Java
C++
AI/ML
AI Agents
Machine Learning
DevOps
AWS
Linux
Unix
Apply
$143k – $193k per year • Equity • In office • Full-Time • 4+ years exp • Master's Degree • Bellevue
Python
Java
C++
AI/ML
RLHF
Reinforcement Learning
Function Calling
AI Agents
LLM
RAG
DPO
SFT
PPO
GRPO
Post-training
Multi-Agent Systems
Tool Use
RLAIF
Reward Modeling
DevOps
Linux
Unix
Robotics
Reinforcement Learning
Apply
$143k – $193k per year • Equity • In office • Full-Time • 4+ years exp • Master's Degree • Bellevue
Python
Java
C++
AI/ML
RLHF
Reinforcement Learning
Function Calling
AI Agents
LLM
RAG
DPO
SFT
PPO
GRPO
Post-training
Multi-Agent Systems
Tool Use
RLAIF
Reward Modeling
DevOps
Linux
Unix
Robotics
Reinforcement Learning
Apply
$165k – $224k per year • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree • Cupertino
Python
Java
C#
C++
Perl
C++
PyTorch C++
AI/ML
vLLM
CUDA Toolkit
SGLang
TensorRT
PyTorch
LLM
CUDA
Triton
AWS Trainium
CUTLASS
Machine Learning
DevOps
CI/CD
AWS
GitHub
Management
Agile
Apply
$180k – $250k per year • Equity • In office • Full-Time • United States
Python
TypeScript
AI/ML
LLM
Apply
≈ $20k – $47k per year (Estimated) • In office • Full-Time • 10+ years exp • Chennai
AI/ML
Copilot
AI Agents
Multi-Agent Systems
DevOps
Azure
Apply
$163k – $414k per year • Hybrid • Full-Time • 12+ years exp • Associate's Degree • Boston • Milwaukee • Dallas • Columbus • Kirkland
SQL
Databases
Google BigQuery
Google Cloud Spanner
Google Bigtable
BigQuery
AI/ML
Vertex AI
AI Agents
Gemini
RAG
Machine Learning
DevOps
Terraform
GCP
GitHub Actions
CI/CD
Jenkins
Docker
Kubernetes
Google GKE
Cybersecurity
SIEM
Analytics
Looker
Management
Google Workspace
Agile
Apply
$133k – $302k per year • Hybrid • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
AI/ML
AI Agents
LLM Guardrails
Multi-Agent Systems
DevOps
Splunk
OpenTelemetry
Datadog
Dynatrace
PagerDuty
Self-Healing
AIOps
SLI/SLO/SLA
Management
ITIL
Apply
$87k – $253k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • New York • Milwaukee • Dallas • Columbus • Kirkland
AI/ML
Copilot
Claude
ChatGPT
AI Agents
Apply
≈ $27k – $80k per year (Estimated) • In office • Full-Time • 6+ years exp • Bengaluru
Python
SQL
AI/ML
LangChain
AI Agents
Semantic Kernel
LLM
RAG
OpenAI
DevOps
Rest API
Azure
Management
Microsoft Office
Apply
≈ $176k – $320k per year (Estimated) • In office • Full-Time • 6+ years exp • San Francisco
Python
AI/ML
JAX
PyTorch
Ray
Multi-Agent Systems
Machine Learning
DevOps
SLURM
Kubernetes
Apply
≈ $176k – $320k per year (Estimated) • In office • Full-Time • 6+ years exp • San Francisco
Python
AI/ML
JAX
PyTorch
Ray
Multi-Agent Systems
Machine Learning
DevOps
SLURM
Kubernetes
Apply
≈ $184k – $333k per year (Estimated) • In office • Full-Time • 5+ years exp • PhD • San Francisco
AI/ML
Quantization
Knowledge Distillation
PyTorch
Mixture of Experts
Post-training
Speculative Decoding
Multi-Agent Systems
Model Distillation
Machine Learning
Apply
≈ $189k – $343k per year (Estimated) • In office • Full-Time • 5+ years exp • PhD • San Francisco
AI/ML
RLHF
Reinforcement Learning
PyTorch
Synthetic Data
DPO
SFT
Post-training
Multi-Agent Systems
RLAIF
Reward Modeling
Machine Learning
Apply
RESEARCHER (GENERAL) 4 months ago
≈ $176k – $320k per year (Estimated) • In office • Full-Time • 5+ years exp • PhD • San Francisco
AI/ML
Function Calling
AI Agents
PyTorch
Multi-Agent Systems
Tool Use
Machine Learning
Apply
$125k – $175k per year • Equity 0.2–2% • In office • Full-Time • 3+ years exp • PhD • San Francisco
Python
AI/ML
Multimodal AI
Time Series Forecasting
Machine Learning
Robotics
Sensor Fusion
Apply
$180k – $250k per year • Remote (United States) • Full-Time • San Francisco
Python
TypeScript
Python
Hypothesis
AI/ML
AI Agents
Post-training
Machine Learning
Apply
≈ $134k – $359k per year (Estimated) • In office • Internship • San Francisco
Python
TypeScript
Python
Hypothesis
AI/ML
AI Agents
Post-training
Machine Learning
Apply
Founding Engineer 1 day ago
$110k – $180k per year • Equity 0.1–1% • In office • Full-Time • San Francisco
Python
JavaScript
Node JS
AI/ML
Vertex AI
OpenAI
Anthropic
Frontend
Next.js
React.js
DevOps
Azure
Kubernetes
Apply
$60k – $84k per year • In office • Internship • San Francisco
Python
JavaScript
TypeScript
AI/ML
Copilot
Cursor
Claude
Claude Code
Model Context Protocol
Vertex AI
AI Agents
LLM
OpenAI
Anthropic
LLM Guardrails
Tool Use
Frontend
Next.js
React.js
DevOps
Azure
AWS
Kubernetes
Apply
See all jobs
This is one of many
793,143 more open roles from verified company boards, updated every day.