698,236open jobs
41,342companies
99,387added this week
Browse all
Salary
$200k – $320k per year
Location
In office (San Francisco)
Employment
Full-Time
Overview
Company
Impact
Profile match
About Build AI. Build AI is the data hyperscaler for Physical AI. We're vertically integrated across hardware, manufacturing, logistics, collection, and model training to scale the physical labor dataset orders of magnitude more than anyone in the world.

About Build AI

Build AI is the data hyperscaler for Physical AI. We're vertically integrated across hardware, manufacturing, logistics, collection, and model training to scale the physical labor dataset orders of magnitude faster than anyone in the world.

Job Summary

Inference is about 90% of compute spend. Economics are heavily driven by inference optimization. We’re hiring someone to make inference cheaper, faster, and good enough that we can scale the data engine and the product without the GPU bill eating the company.

Key Responsibilities

  • Own inference performance: latency, throughput, and cost per unit of work (tokens, frames, or jobs)

  • Cut the 90% compute line: kernels, batching, quantization, compilation, serving, and hardware utilization

  • Profile pipelines (Nsight, PyTorch Profiler, or equivalent), find the real bottleneck, and ship the fix

  • Work with research and product so models that are accurate are also affordable to run at scale

  • Build the serving and eval path so experiments don’t hide the inference bill

  • Measure cost as a first-class metric, not an afterthought once quality is “done”

You may be a good fit if you have (Must-have qualifications)

  • Strong ML / systems engineer with real inference optimization experience (serving, compilers, CUDA/kernels, quantization, or similar)

  • Comfortable in Python and in C++ or Rust for performance-critical paths

  • You think in dollars and tokens/frames per second, not only in accuracy tables

  • Familiarity with PyTorch (or JAX) and with profiling tools

  • Comfortable in a small research team shipping under cost pressure

Strong candidates may also have experience with (Nice-to-have qualifications)

  • CUDA, kernels, compilers (TVM, MLIR, TensorRT), or quantization in production

  • You have owned GPU/accelerator cost as a first-class metric

  • Serving stacks for video or large models

  • Understanding of memory hierarchy, data movement, and low-precision compute

Benefits

  • Competitive pay

  • Medical, dental, and vision packages with generous premium coverage

  • $500 per month credit for waiving medical benefits

  • Housing subsidy of $2k per month for those living within walking distance of the office

  • Relocation support for those moving to San Francisco (Financial District) or Shenzhen (Nanshan)

  • Various wellness benefits covering fitness, mental health, and more

  • Daily lunch and dinner in our office

  • Unlimited compute budget subject to ROI justification

  • Unlimited Codex and Claude credits

  • Travel

How we're different

Build believes in the Bitter Lesson. By taking a general approach of learning from humans, our addressable market is all physical labor.

We are a fully in-person team in San Francisco (Financial District) and Shenzhen (Nanshan), and greatly value engineering skills. We do not have boundaries between engineering and research, and we expect all of our technical staff to contribute to both and work across disciplines as needed.

Build AI is an equal opportunity employer. We review every application. If you do not meet every bullet, still apply. Questions: [email protected]

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
698,236 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
In office • Internship • Bachelor's Degree • Taipei • Hsinchu
Python
C
C++
Perl
C
Valgrind
DevOps
GitHub Actions
CircleCI
CI/CD
Jenkins
Git
Docker
Kubernetes
Spinnaker
KVM
QEMU
Xen
GitHub
GitLab
Apply
$42k – $101k per year (Estimated) • In office • Full-Time • 6+ years exp • Master's Degree • Beijing • Shanghai • Shenzhen
Python
C++
DevOps
Linux
Apply
Engineer 11 hours ago
$80k – $200k per year • Equity 1–3% • In office • Full-Time • Brisbane
Python
TypeScript
AI/ML
AI Agents
Management
Zoom
Apply
In office • Internship • Bachelor's Degree
Python
JavaScript
SQL
PowerShell
AI/ML
Copilot
Claude
ChatGPT
DevOps
GitHub
Analytics
Power BI
Management
UiPath
Power Automate
Power Apps
SharePoint
Apply
In office • Internship • Bachelor's Degree
Python
JavaScript
SQL
PowerShell
AI/ML
Copilot
Claude
ChatGPT
DevOps
GitHub
Analytics
Power BI
Management
UiPath
Power Automate
Power Apps
SharePoint
Apply
Founding Paralegal 1 day ago
$140k – $180k per year • In office • Full-Time • 5+ years exp • San Francisco
AI/ML
Claude
Physical AI
Apply
$250k – $320k per year • In office • Full-Time • 6+ years exp • San Francisco
AI/ML
Claude
Physical AI
Apply
$130k – $180k per year • In office • Full-Time • 2+ years exp • San Francisco
AI/ML
Claude
Physical AI
Apply
$180k – $300k per year • In office • Full-Time • San Francisco • Shenzhen
AI/ML
Claude
Physical AI
Marketing
Meta Ads
Apply
Research Intern 22 days ago
$144k per year • In office • Internship • Bachelor's Degree • San Francisco
AI/ML
Claude
Multimodal AI
PyTorch
World Models
Physical AI
Apply
$155k – $256k per year • Equity • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • San Francisco
Python
Java
AI/ML
Claude Code
AI Agents
LLM
OpenAI Codex
LLM Guardrails
Machine Learning
DevOps
Splunk
Kubernetes
Apply
$150k – $220k per year • In office • Full-Time • San Francisco
AI/ML
Function Calling
LLM
RAG
Agentic Workflows
Tool Use
Apply
Senior AI Engineer 1 day ago
$150k – $220k per year • In office • Full-Time • San Francisco
Python
AI/ML
LLM
Apply
$190k – $225k per year • In office • Full-Time • 8+ years exp • San Francisco
SQL
Apply
Product Manager 1 day ago
$150k – $190k per year • In office • Full-Time • 4+ years exp • San Francisco
SQL
Analytics
A/B Testing
Apply
See all jobs
This is one of many
698,236 more open roles from verified company boards, updated every day.