660,302open jobs
38,440companies
97,349added this week
Browse all
Salary
$203k – $461k per year (Estimated)
Location
In office (San Francisco)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Moonlake is a frontier AI lab building world models and simulation infrastructure for physical AI.

Introducing Moonlake, AI for creating world simulations.

About Moonlake

Moonlake is building the frontier of interactive world models: systems that generate, simulate, and reason over 3D environments for robotics, embodied AI, and interactive applications.

We develop the infrastructure that enables intelligent systems to learn, evaluate, and interact within realistic virtual environments before operating in the physical world.

Our work sits at the intersection of:

  • Robotics

  • Embodied AI

  • Interactive 3D Worlds

  • World Models

  • Simulation Infrastructure

  • Physical AI

Moonlake is building the next generation of AI infrastructure for interactive digital worlds. Our mission is to enable anyone to create, simulate, and interact with rich environments using natural language and multimodal inputs, turning simple ideas into worlds with structure, physics, and intelligent behavior.

Our team has raised $50M in seed funding from NVIDIA Ventures, Threshold Ventures, AIX Ventures, and notable angels including Naval Ravikant and Jeff Dean to build the foundational layer for the future of AI-powering everything from robotics training and simulation to digital twins and interactive environments.

We are looking for exceptional engineers to help build the simulation systems that will power the next generation of robotics and embodied intelligence.

The Role

We are hiring a Member of Technical Staff to lead reinforcement learning infrastructure and model post-training.

You will work closely with Qi and the research team to improve large vision-language and code-generating agents through fine-tuning, reinforcement learning, trajectory data, and scalable evaluation.

Moonlake already has deep expertise in 3D and world-building. This role adds the model-training experience needed to systematically improve agent performance and prepare the company for larger-scale RL across both digital and physical environments.

We are looking for a full-stack researcher and engineer who understands the complete training system and can make strong judgments about when training is necessary, which methods are likely to work, and what not to pursue.

What You’ll Do

  • Build RL and post-training pipelines for multimodal, vision-language, and code-generating agents

  • Develop infrastructure for supervised fine-tuning, preference optimization, reward modeling, and reinforcement learning

  • Create systems for collecting, filtering, replaying, and learning from agent trajectories

  • Design rewards, verifiers, and evaluations for long-horizon agent tasks

  • Improve agents’ ability to plan, write and execute code, use tools, recover from errors, and complete complex workflows

  • Scale distributed training and high-throughput rollout generation across multi-GPU environments

  • Improve training reliability, reproducibility, observability, and cost efficiency

  • Help define Moonlake’s long-term strategy for agent, robotics, and embodied-model training

What We’re Looking For

  • Real-world experience training large language, vision-language, multimodal, or code models

  • Strong experience in reinforcement learning, post-training, or large-scale fine-tuning

  • Experience building distributed training or high-throughput inference systems

  • Familiarity with supervised fine-tuning, preference optimization, reward modeling, and agentic RL

  • Experience with code-generation agents, long-horizon evaluation, or tool-using systems

  • Strong Python skills and experience with PyTorch, JAX, or similar frameworks

  • Ability to work across data, models, environments, rewards, evaluation, and infrastructure

  • Strong research judgment and a bias toward building reliable systems

Preferred Experience

  • Experience at a frontier AI lab or organization operating large-scale training systems

  • Experience with code-model post-training or autonomous coding agents

  • Experience with multimodal models, robotics, simulation, or embodied AI

  • Experience designing verifiable rewards or outcome-based training systems

  • Experience scaling RL workloads across large GPU clusters

What Success Looks Like

Within your first year, you will have:

  • Built Moonlake’s core post-training and RL infrastructure

  • Created scalable systems for learning from agent trajectories

  • Delivered measurable improvements in agent quality and task completion

  • Helped the team determine which problems require training and which do not

  • Established a reliable foundation for larger-scale agent and embodied-model training

Why This Role Matters

Moonlake’s agents must do more than generate content. They must understand complex requests, reason across vision and language, write and execute code, operate tools, build interactive worlds, and recover from mistakes.

This role will build the training systems that allow those agents to continuously improve.

We are committed to being an on-site, in-person team currently based in San Francisco.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
660,302 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
DevSecOps Engineer 5 hours ago
$79k – $160k per year • In office • Top Secret • 4+ years exp • Washington
Python
JavaScript
Java
PHP
SQL
PowerShell
Node JS
Bash
Java
Apache Tomcat
PHP
Drupal
AI/ML
Claude
Claude Code
Frontend
React.js
DevOps
GitHub Actions
CloudFormation
GitLab CI
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Nginx
Amazon EKS
GitLab
Amazon ECS
Apply
$120k – $218k per year (Estimated) • Remote • Plano
Python
JavaScript
PowerShell
Node JS
Bash
Node JS
Commander.js
Cybersecurity
Volatility
Autopsy
The Sleuth Kit
HIPAA
Magnet AXIOM
FTK
Apply
$76k per year • Remote/Hybrid • Full-Time • 1+ year exp • Bachelor's Degree • Charlotte • Richmond
Python
Java
SQL
SAS
Analytics
Tableau
Power BI
Microsoft Excel
Apply
$104k – $158k per year • In office • Full-Time • 10+ years exp • Charlotte • Jersey City • Plano
Python
Java
SQL
Scala
Python
pySpark
Databases
Db2
Oracle
Apache Iceberg
Apache Kafka
AI/ML
Hadoop
Spark
Flink
DevOps
GCP
Azure
CI/CD
AWS
Bitbucket
Management
Jira
ServiceNow
Agile
Apply
$72k per year • Remote/Hybrid • Full-Time • Bachelor's Degree • Charlotte • Richmond
Python
Java
SQL
Analytics
Tableau
Power BI
Microsoft Excel
Apply
Head of Design 1 month ago
$117k – $316k per year (Estimated) • In office • Full-Time • San Francisco
JavaScript
TypeScript
AI/ML
Multimodal AI
Edge AI
World Models
Physical AI
Embodied AI
Frontend
Tailwind CSS
Next.js
React.js
Robotics
Digital Twin
Apply
$156k – $348k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Multimodal AI
World Models
Embodied AI
Robotics
Digital Twin
Apply
$197k – $447k per year (Estimated) • In office • Full-Time • San Francisco
Python
AI/ML
Multimodal AI
Synthetic Data
World Models
Physical AI
Embodied AI
Robotics
Gazebo
Isaac Sim
MuJoCo
Sim-to-Real
Digital Twin
Apply
$194k – $440k per year (Estimated) • In office • Full-Time • Master's Degree • San Francisco
Python
C++
AI/ML
Multimodal AI
Synthetic Data
World Models
Physical AI
Embodied AI
Robotics
Drake
ROS2
Gazebo
PyBullet
Isaac Sim
MuJoCo
Sim-to-Real
Digital Twin
Apply
$154k – $310k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco
Python
JavaScript
TypeScript
AI/ML
Claude Code
Multimodal AI
AI Agents
LLM
OpenAI Codex
World Models
Embodied AI
Frontend
Tailwind CSS
Next.js
React.js
Radix UI
shadcn/ui
Robotics
Isaac Sim
MuJoCo
Digital Twin
Design
Blender
Apply
Business Risk Officer 8 hours ago
$163k – $204k per year • Remote • 7+ years exp • San Francisco
Apply
$128k – $192k per year • Equity • In office • Full-Time • 1+ year exp • San Francisco • Pleasanton
Apply
Investment Officer 9 hours ago
$129k – $223k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • San Francisco
Apply
$152k – $327k per year (Estimated) • In office • 13+ years exp • San Francisco
AI/ML
OpenAI
Anthropic
Management
Stripe
Apply
$152k – $279k per year (Estimated) • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
Python
JavaScript
Node JS
AI/ML
LLM
Frontend
React.js
DevOps
Terraform
AWS CDK
Azure
AWS
Docker
Kubernetes
Platform Engineering
Amazon EKS
AWS Fargate
AWS Lambda
Amazon EC2
HPC
Apply
See all jobs
This is one of many
660,302 more open roles from verified company boards, updated every day.