746,581open jobs
44,818companies
107,195added this week
Browse all
Salary
$437k – $485k per year
Location
Hybrid (Seattle, United States)
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 24, 2026. First seen by Alion on Sep 24, 2026. OpenAI scores B on the Alion truth index.

Overview
Company
Impact
Profile match
OpenAI is an American artificial intelligence research and deployment company founded in 2015 with the mission of ensuring that artificial general intelligence benefits all of humanity. It develops the GPT family of large language models and turns them into consumer and developer products, including the ChatGPT assistant, the Sora video model, the Codex coding agent and a commercial API used by millions of developers. Structured as a public benefit corporation controlled by a non-profit foundation, the company is backed by Microsoft and SoftBank and operates from San Francisco.

About the Team

The Statsig team within OpenAI builds the experimentation, feature rollout, dynamic configuration, and analytics systems that help OpenAI ship products with speed, safety, and evidence. Our work sits on the critical path for how product, engineering, research, and go-to-market teams learn from real-world usage and make high-confidence decisions.

Statsig began as an independent company focused on helping builders move faster through trustworthy experimentation and feature management. After Statsig joined OpenAI, the team began the next chapter: bringing that deep product expertise, customer intuition, and mature platform infrastructure into OpenAI as the experimentation and rollout platform for every product we ship.

Today, we support teams across ChatGPT, Codex, model measurement, consumer experiences, business subscriptions, developer products, and the shared infrastructure that connects them. These teams rely on Statsig to safely introduce new capabilities, compare product and model behavior, measure impact, and roll changes forward or back with confidence.

We are at a defining moment in the platform journey. OpenAI has the data, product surface area, and pace of innovation to learn faster than almost any organization in the world, but that potential only becomes real if teams can experiment responsibly, measure clearly, and roll out changes safely.

We are evolving experimentation systems to help teams learn from product behavior and make better evidence-based decisions.

Based out of OpenAI’s Bellevue office, we are a close-knit team that values in-person collaboration, urgency, craft, and impact. We build for other builders, and the best version of this team is one where every OpenAI product team can move faster because the experimentation and rollout layer is dependable, fast, and easy to use.

About the Role

We are looking for a Machine Learning Engineer to lead the technical direction for ML-powered experimentation and insights capabilities. You will build production systems that learn from privacy-protected product and experimentation data to generate evidence-backed insights and support decision-making, and help teams decide which ideas are worth testing live.

This is an end-to-end, 0-to-1 role. You will work across ML modeling, retrieval and LLM systems, statistical methods, simulation, data and training pipelines, backend services, and user- and agent-facing product experiences. The hard part is not merely producing a plausible answer. It is making each insight and prediction traceable, calibrated, useful, and safe enough to influence real product decisions.

Live experiments remain the source of causal validation. You will design systems that make uncertainty explicit, backtest against historical outcomes, compare predictions with online results, learn from misses, and abstain when the evidence is weak. You will preserve clear review, permission, and approval boundaries as automation becomes more powerful.

You will collaborate closely with teams building ChatGPT, Codex, model measurement workflows, consumer products, Growth, business subscription experiences, developer products, and shared infrastructure. You will turn their most important learning and decision problems into general platform capabilities that can support the full company.

In This Role, You Will

  • Set and execute the technical roadmap for Generative Insights and Predictive Experimentation, from early prototypes through production adoption.

  • Build cross-experiment learning systems that retrieve and synthesize historical experiments, detect recurring effects and segment behavior, reanalyze prior results when data or methods improve, and generate hypotheses with clear evidence and provenance.

  • Develop predictive models and simulation workflows, including simulation-based evaluation approaches, to estimate likely impact, affected segments, regression risk, and uncertainty before a full live experiment.

  • Create high-quality datasets and feature or retrieval pipelines from exposures, events, metrics, experiment metadata, and replay data, with strong lineage, freshness, privacy, and data-quality controls.

  • Establish rigorous evaluation through offline benchmarks, backtests, calibration, drift monitoring, prediction-to-outcome comparisons, and explicit failure or abstention behavior.

  • Turn models into durable product, API, and agent workflows that move from an insight to experiment design, approval-gated action, and measured learning.

  • Partner deeply with data science and product teams on experiment design, causal inference, sequential decision-making, variance reduction, and the boundary between prediction and causal evidence.

  • Build reliable services and intuitive workflows so sophisticated ML capabilities are understandable and useful to teams making high-stakes product decisions.

  • Provide technical leadership across engineering, product, data science, and research partners, and raise the bar for production ML quality across the platform.

You Might Thrive In This Role If You

  • Have led ambiguous 0-to-1 production ML products where success was measured by better real-world decisions, not only offline model metrics.

  • Have strong hands-on experience across the ML lifecycle: dataset design, training or adaptation, evaluation, deployment, monitoring, and iteration.

  • Bring depth in one or more of LLM and retrieval systems, ranking or recommendation, forecasting or anomaly detection, causal ML or experiment analysis, or simulation. You do not need to have done all of them.

  • Have strong software engineering fundamentals and can build high-quality production systems in Python while working comfortably across data, backend, and platform boundaries.

  • Have a strong grounding in machine learning, statistics, computer science, or a related field through formal study or equivalent practical experience.

  • Understand experimentation and statistical reasoning, especially why predictive accuracy is not the same as causal validity.

  • Treat calibration, uncertainty, provenance, privacy, and human review as product requirements, not cleanup work.

  • Can translate ambiguous partner questions into a product and technical roadmap, and work well with product, data science, research, and infrastructure partners.

  • Enjoy building for internal power users and agents, and can make sophisticated ML capabilities feel clear and actionable.

  • Value in-person collaboration and want to help shape a growing Bellevue-based team.

Location and Workplace

This role is based in Bellevue, Washington. The team works in person and uses that time to move quickly, solve ambiguous problems together, and stay close to the product teams we support.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.

Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.

To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.

We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.

OpenAI Global Applicant Privacy Policy

At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
746,581 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Seattle
AI/ML Engineer II 1 hour ago
≈ $123k – $239k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Phoenix
Python
SQL
Python
FastAPI
pySpark
Databases
Snowflake
Google BigQuery
BigQuery
AI/ML
Spark
Airflow
XGBoost
Vertex AI
Scikit-learn
NLP
TensorFlow
Pandas
NumPy
PyTorch
NLTK
Machine Learning
DevOps
GCP
OpenShift
GitLab CI
Azure
CI/CD
AWS
Docker
Amazon EC2
Amazon S3
Robotics
Path Planning
Apply
≈ $171k – $298k per year (Estimated) • Remote (United States) • 5+ years exp • Bachelor's Degree • San Francisco
Python
Java
Scala
AI/ML
TensorFlow
PyTorch
Anomaly Detection
Machine Learning
Analytics
A/B Testing
Apply
$130k – $220k per year • Hybrid • Full-Time • 3+ years exp • PhD • Santa Clara
Python
C++
AI/ML
Machine Learning
Robotics
ROS
Path Planning
Motion Planning
Imitation Learning
Apply
$116k – $208k per year • Remote (United States) • Full-Time • 8+ years exp • Bachelor's Degree • United States
Databases
PostgreSQL
Redis
DynamoDB
Apache Kafka
OpenSearch
Amazon Aurora
AI/ML
LangGraph
AutoGen
LangChain
AI Agents
Semantic Kernel
LLM
LLM Guardrails
Multi-Agent Systems
DevOps
Rest API
gRPC
Terraform
Helm
GitHub Actions
CI/CD
GitOps
ArgoCD
Jenkins
AWS
Kubernetes
Platform Engineering
Chaos Engineering
Service Mesh
Amazon EKS
Progressive Delivery
SLI/SLO/SLA
IAM
Amazon ECS
Amazon Kinesis
API Gateway
Cybersecurity
ISO 27001
SOC 2
Zero Trust
SIEM
Chips/EDA
PoC Library
Apply
$89k – $141k per year • Hybrid • Full-Time • Somerville
Python
SQL
Python
Boto3
Databases
Databricks
AI/ML
Weights & Biases
Spark
MLFlow
Fine-tuning
PyTorch
Time Series Forecasting
Human-in-the-Loop
DevOps
AWS
Apply
In office • Full-Time
Python
AI/ML
Model Context Protocol
Multimodal AI
AI Agents
LLM
OpenAI
Anthropic
Apply
In office • Full-Time • 2+ years exp • Bachelor's Degree • Mexico
Python
SQL
C#
Cybersecurity
GDPR
Analytics
Power BI
Apply
≈ $41k – $109k per year (Estimated) • In office • 3+ years exp
Python
C++
AI/ML
Reinforcement Learning
Multimodal AI
DevOps
Git
Docker
Linux
Robotics
Isaac Sim
MuJoCo
Sensor Fusion
Perception
Sim-to-Real
Reinforcement Learning
Apply
≈ $48k – $121k per year (Estimated) • In office
Python
C++
AI/ML
Reinforcement Learning
Multimodal AI
World Models
Physical AI
DevOps
Git
Docker
Linux
Robotics
Imitation Learning
Reinforcement Learning
Apply
≈ $41k – $109k per year (Estimated) • In office • 3+ years exp
Python
C++
AI/ML
Reinforcement Learning
Multimodal AI
DevOps
Git
Docker
Linux
Robotics
Isaac Sim
MuJoCo
Sensor Fusion
Perception
Sim-to-Real
Reinforcement Learning
Apply
$381k – $555k per year • In office • Full-Time • Master's Degree • San Francisco
AI/ML
Fine-tuning
Knowledge Distillation
TensorFlow
PyTorch
OpenAI
SFT
Model Distillation
Machine Learning
Apply
$217k – $276k per year • Hybrid • Full-Time • San Francisco
Python
JavaScript
AI/ML
ChatGPT
AI Agents
LLM
OpenAI
OpenAI Codex
Apply
$221k – $278k per year • Hybrid • Full-Time • San Francisco
Python
JavaScript
AI/ML
ChatGPT
AI Agents
LLM
OpenAI
OpenAI Codex
Apply
$380k – $500k per year • In office • TS/SCI • Full-Time • 4+ years exp • Bachelor's Degree • San Francisco • Washington
AI/ML
RLHF
Reinforcement Learning
OpenAI
Post-training
Machine Learning
Apply
$380k – $500k per year • Hybrid • Full-Time • San Francisco
Verilog
SystemVerilog
AI/ML
Reinforcement Learning
Function Calling
OpenAI
Post-training
Tool Use
Chips/EDA
Formal Verification
Apply
$183k – $240k per year • Equity • Hybrid • Full-Time • 4+ years exp • New York • Seattle • San Francisco
Python
Go
JavaScript
TypeScript
AI/ML
AI Agents
LLM
DevOps
CI/CD
Cybersecurity
Threat Modeling
Apply
Remote (United States) • Confidential • Full-Time • Bachelor's Degree • Louisville • Las Vegas • Des Moines • Virginia Beach • Philadelphia
AI/ML
Copilot
Management
Microsoft Office
Apply
$70k per year • Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Seattle
Apply
≈ $72k – $126k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Seattle
Python
PowerShell
Bash
DevOps
VMWare
Windows Server
Hyper-V
Linux
Windows
TCP/IP
DNS
DHCP
Cybersecurity
Active Directory
Management
Jira
Google Workspace
ServiceNow
ITIL
Microsoft Office
Apply
$146k – $306k per year • Equity • In office • Seattle
Apply
See all jobs
This is one of many
746,581 more open roles from verified company boards, updated every day.