368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$200k – $300k per year
Location
In office (San Francisco)
Overview
Company
Impact
Profile match
World Labs is a spatial intelligence company, building frontier models that can perceive, generate, and interact with the 3D world.

About World Labs:

We build foundational world models that can perceive, generate, reason, and interact with the 3D world - unlocking AI's full potential through spatial intelligence by transforming seeing into doing, perceiving into reasoning, and imagining into creating. We believe spatial intelligence will unlock new forms of storytelling, creativity, design, simulation, and immersive experiences across both virtual and physical worlds. We bring together a world-class team, united by a shared curiosity, passion, and deep backgrounds in technology - from AI research to systems engineering to product design - creating a tight feedback loop between our cutting-edge research and products that empower our users.

Role Overview

We are looking for a Performance Engineer to make World Labs’ models train and serve as fast as the hardware allows.

Running large generative world models at scale is a novel systems problem. You will find the bottlenecks - in kernels, in the serving path, in the training loop, in how we use our GPUs - and eliminate them. Your ownership is technical and concrete: the throughput you unlock, the latency you cut, the utilization you win back, and the correctness you hold while doing it. You will work up and down the stack, from low-level tensor and kernel optimization to fleet-wide serving efficiency, in close partnership with the researchers whose models you are accelerating.

This is a hands-on, individual-contributor role. You will profile, design, build, and ship code directly.

What You Will Do:

  • Optimize inference and serving end to end - latency, throughput, batching, caching, and scheduling - to serve our models efficiently at production scale.
  • Write and tune GPU kernels (CUDA, Triton) for hot paths; drive kernel fusion, memory- and bandwidth-bound optimization, and low-precision (FP8/INT8) execution.
  • Optimize training throughput and GPU utilization: parallelism strategies, communication/compute overlap, mixed precision, and eliminating pipeline stalls.
  • Build performance models, profiling workflows, and observability that make throughput, latency, cost, utilization, and their tradeoffs legible across the stack.
  • Own numerical correctness across precision, kernel, and hardware changes - treating correctness as part of performance, not separate from it.
  • Partner with researchers to productionize models for serving and to make experiments run faster and more reliably.
  • Where needed, work on the distributed systems that training and inference run on - but the core of the job is squeezing the most out of every GPU.

Key Qualifications:

You should excel at the fundamentals below - we index on inference, serving, GPU optimization, and training performance. Distributed-systems breadth is welcome, but secondary.

  • Strong performance-engineering foundations: profiling, roofline analysis, latency/throughput optimization, and disciplined root-cause investigation.
  • Deep GPU programming and optimization experience (CUDA and/or Triton) - kernel-level tuning, memory hierarchy, and bandwidth optimization at scale.
  • Hands-on experience optimizing inference and serving for large models: batching, KV/prompt caching, quantization, and low-latency, high-throughput sampling.
  • Hands-on experience optimizing training performance: parallelism, distributed communication, mixed/low precision, and utilization.
  • Working knowledge of ML framework internals (PyTorch and/or JAX; torch.compile, XLA, or similar compiler paths).
  • Strong proficiency in Python, with the ability to drop into C++/CUDA (and Rust or Go) as the work demands.
  • High-ownership mindset - you measure yourself by throughput shipped and latency cut, not tickets closed.

Preferred Qualifications:

  • Experience at an AI lab or ML-native company, optimizing systems used directly by researchers and productionizing research code.
  • Low-precision and numerics depth: FP8/INT8 quantization, mixed-precision, and detecting numerical regressions across hardware platforms.
  • Distributed systems for large-scale training and inference - collective communication (NCCL), interconnects (NVLink), model and tensor parallelism, and fault tolerance. A strong plus, but not a substitute for the core skills above.
  • Experience serving generative, diffusion, video, or 3D/spatial models - not just text LLMs.
  • Multi-accelerator experience (GPU plus TPU or Trainium) and partnering with hardware vendors on accelerator capabilities.
  • Building performance-modeling and observability frameworks for GPU utilization and cost.

 

Who You Are:

  • Fearless Innovator: We need people who thrive on challenges and aren't afraid to tackle the impossible.
  • Resilient Builder: Impacting Large World Models isn't a sprint; it's a marathon with hurdles. We're looking for builders who can weather the storms of groundbreaking research and come out stronger.
  • Mission-Driven Mindset: Everything we do is in service of creating the best spatially intelligent AI systems, and using them to empower people.
  • Collaborative Spirit: We're building something bigger than any one person. We need team players who can harness the power of collective intelligence.

We're hiring the brightest minds from around the globe to bring diverse perspectives to our cutting-edge work. If you're ready to work on technology that will reshape how machines perceive and interact with the world, World Labs is your launchpad.

Join us, and let's make history together.

Equal Opportunity & Pay Transparency

Equal Employment Opportunity

World Labs is an equal opportunity employer. We do not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, genetic information, veteran status, or any other characteristic protected under applicable law. We welcome all qualified applicants and are committed to providing reasonable accommodations throughout the hiring process upon request.

California Pay Transparency

In accordance with California law, we disclose the following:

Pay Range

$200-$300k base salary (good-faith estimate for San Francisco Bay Area upon hire; actual offer based on experience, skills, and qualifications)

Total Compensation

Base salary plus equity awards

Salary History

We do not request or consider prior compensation in making offers

Compliance: Cal. Lab. Code §432.3 (pay scale disclosure & salary history ban); Cal. Lab. Code §1197.5 (Equal Pay Act); Cal. Gov. Code §12940 (FEHA); 42 U.S.C. §2000e (Title VII); 29 U.S.C. §621 (ADEA); 42 U.S.C. §12101 (ADA)

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$129k – $232k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Huntsville
Ada
C++
Fortran
MATLAB
Python
Cython
Cython
Meson
AI/ML
NumPy
DevOps
CI/CD
GitLab
Analytics
Matplotlib
Apply
$117k – $177k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • Chicago • Dallas
Apex
C#
Java
Python
Apex
Salesforce Data Cloud
AI/ML
Agentforce
AI Agents
Chain-of-Thought
Embeddings
Few-Shot Learning
Fine-tuning
Hallucination
LLM Guardrails
NLP
Prompt Engineering
RAG
RLHF
Semantic Search
Semantic Search
Frontend
GraphQL
DevOps
CI/CD
Git
Analytics
A/B Testing
Marketing
Salesforce
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Senior ML Engineer 1 hour ago
$149k – $224k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Francisco • Washington • Palo Alto
Python
Python
pySpark
Databases
Apache Kafka
AI/ML
AI Agents
Agentforce
Airflow
Anomaly Detection
Feature Store
Flink
Ray
Red Teaming
Spark
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
MITRE ATT&CK
Marketing
Salesforce
Apply
Software Developer 1 hour ago
$96k – $190k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Huntsville
Python
AI/ML
Copilot
Keras
PyTorch
Scikit-learn
TensorFlow
DevOps
Docker
GitHub
Kubernetes
Podman
Apply
$250k – $325k per year • Equity • In office • 3+ years exp • San Francisco
Python
AI/ML
Computer Vision
Diffusion Models
Fine-tuning
Knowledge Distillation
PyTorch
TensorFlow
Post-training
Pre-training
Apply
$250k – $350k per year • Equity • In office • 6+ years exp • San Francisco
C++
Python
AI/ML
Computer Vision
Robotics
Localization
Perception
Sensor Fusion
SLAM
Apply
$200k – $250k per year • Equity • In office • 8+ years exp • San Francisco
AI/ML
ChatGPT
Claude
Gemini
Management
Google Workspace
Slack
Apply
$250k – $325k per year • Equity • In office • 6+ years exp • San Francisco
TypeScript
JavaScript
Frontend
React.js
Three.JS
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Senior ML Engineer 1 hour ago
$149k – $224k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Francisco • Washington • Palo Alto
Python
Python
pySpark
Databases
Apache Kafka
AI/ML
AI Agents
Agentforce
Airflow
Anomaly Detection
Feature Store
Flink
Ray
Red Teaming
Spark
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
MITRE ATT&CK
Marketing
Salesforce
Apply
In office • Internship • 1+ year exp • Bachelor's Degree • San Francisco
Go
JavaScript
Ruby
Scala
Apply
$360k – $530k per year • In office • Full-Time • Bachelor's Degree • San Francisco
MATLAB
Python
MATLAB
Simulink
AI/ML
OpenAI
Robotics
Digital Twin
Apply
$222k – $277k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Francisco
DevOps
CI/CD
Immutable Infrastructure
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.