412,822open jobs
14,124companies
59,735added this week
Browse all
Salary
$196k – $402k per year (Estimated)
Location
In office (San Francisco)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Index : kitd this commitmain process supervisor for OpenBSDabout summary refs log tree commit diff log msgauthorcommitterrange KITD(8) System Manager's Manual KITD(8) NAME kitd — process supervisor SYNOPSIS kitd [-d] [-c cooloff] [-m maximum] [-n name] [-t restart] command ...

Our mission is general causal intelligence; AI that is capable of (1) predicting the future and (2) identifying the actions to alter it.

To achieve this breakthrough, we are building a Large Physics foundation Model (LPM) because physical systems, unlike text or images, are governed by verifiable cause and effect. We believe that scaling on physics will enable an understanding of causality required to predict and control physical systems, starting with weather.

Our founding team has built and deployed AI against the physical world in robotics, drug discovery, and particle physics at institutions like DeepMind, Waymo, Cruise, Insitro, Nabla Bio, and CERN.

We look for infrastructure engineers who are excited to tackle unsolved problems. Progress on an LPM is gated by how fast we can evaluate it: large-scale backtesting against decades of physical observations, ensemble generation, and rollout evaluation across model scales.

Responsibilities

Your mission is to make inference so fast and cheap that evaluation never gates research.

  • Build high-throughput inference systems for large-scale evaluation, backtesting, and scoring against historical physical observations

  • Design and implement techniques that improve latency, throughput, and efficiency for real-time inference

  • Optimize the inference stack to fully utilize hardware FLOPs, bandwidth, and memory

  • Extend orchestration frameworks (e.g. Kubernetes, Ray, Slurm) for distributed inference and large-batch evaluation sweeps

  • Establish standards for reliability, observability, and reproducibility across the inference stack, so every evaluation is trustworthy and repeatable

  • Collaborate with researchers to enable high-performance inference for novel architectures as they emerge

What we're looking for

We value a relentless approach to problem-solving, rapid execution, and the ability to quickly learn in unfamiliar domains.

  • Experience building or optimizing inference and serving systems for throughput and latency (e.g. TensorRT)

  • Understanding of distributed compute, GPU parallelism, and hardware-aware optimization

  • Deep familiarity with deep learning frameworks (e.g. PyTorch, JAX) and their underlying system architectures

  • Strong engineering skills: performant, maintainable code and the ability to debug complex codebases

  • Bonus: contributions to open-source inference or systems infrastructure (e.g. vLLM, SGLang, Triton)

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
412,822 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
In office • Bachelor's Degree
Python
C
C
FFmpeg
AI/ML
CUDA Toolkit
Triton Inference Server
Embeddings
Multimodal AI
Function Calling
Computer Vision
VLM
TensorRT
PyTorch
LLM
RAG
CUDA
Triton
NVIDIA NeMo
Structured Outputs
DevOps
AWS
Docker
Robotics
GStreamer
Apply
$182k – $374k per year (Estimated) • In office • Full-Time • San Francisco
Python
Rust
DevOps
Terraform
GCP
Pulumi
SLURM
Azure
AWS
Docker
Kubernetes
IAM
Cybersecurity
Okta
SOC 2
FedRAMP
Zero Trust
Threat Modeling
Microsoft Entra ID
Apply
$209k – $429k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Reinforcement Learning
Post-training
Robotics
Reinforcement Learning
Apply
$197k – $404k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Multimodal AI
Computer Vision
World Models
Robotics
Sensor Fusion
Apply
$202k – $415k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Interpretability
Apply
$184k – $377k per year (Estimated) • In office • Full-Time • San Francisco
DevOps
Terraform
GCP
Azure
AWS
Docker
Kubernetes
Apply
$60k – $100k per year • In office • Full-Time • 2+ years exp • San Francisco
Apply
$77k – $173k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco
Apply
$140k – $210k per year • Equity • In office • Full-Time • 4+ years exp • San Francisco
Python
Go
TypeScript
DevOps
GCP
CI/CD
Kubernetes
Management
Stripe
Apply
$200k – $250k per year • Equity 1–2% • In office • Full-Time • 3+ years exp • San Francisco
Apply
$160k – $250k per year • Equity 1–2% • In office • Full-Time • 3+ years exp • San Francisco
Python
AI/ML
PyTorch
DevOps
AWS
Apply
See all jobs
This is one of many
412,822 more open roles from verified company boards, updated every day.