1,469,448open jobs
87,948companies
231,463added this week
Browse all
Salary
$237k – $298k per year
Location
Hybrid (Boston, Foster City, United States)
Seniority
Senior · 8+ years exp
Visa
H-1B filings in 12 months: 373 · for this role: 169 · green card filings: 120
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 11, 2026. First seen by Alion on Sep 25, 2026. Zoox scores B on the Alion truth index.

Overview
Company
Impact
Profile match
Zoox is an American autonomous vehicle company founded in 2014 that designs a purpose-built robotaxi rather than retrofitting an existing car. Its vehicle has no steering wheel or driver position, seats four passengers facing each other, drives equally well in either direction and was developed alongside the sensors, compute and software that operate it, which required the company to certify the design under its own self-certification of federal motor vehicle safety standards. Amazon acquired the company in 2020 and it has since begun public robotaxi service in Las Vegas and San Francisco while testing in several other American cities.

Zoox made the world’s first purpose-built commercial robotaxi, combining advanced self-driving hardware and software. In order to ensure safe operation, low latency and consistent resource utilization are critical. The Planner Compute team is responsible for the performance of the largest single software component within the autonomy stack, the motion planner.

The Planner Compute team is looking for an expert in GPU performance. You will instrument, monitor, analyze and optimize GPU-based algorithms that are performance-critical for our solution. The scope for GPU usage ranges from machine learning inference to custom CUDA kernels written specifically for the motion planner.

In this role, you will:

  • Analyze performance metrics to identify GPU hotspots and optimizations

  • Contribute to the adaptation of our current planner to multiple GPU architectures with different resources

  • Optimize ML models for latency and GPU memory usage

  • Support engineers within Planner as a subject matter expert on CUDA & GPU performance

Qualifications

  • BS in computer science or related field and 3+ years of experience.

  • Strong knowledge of CUDA as applied to recent GPU microarchitectures (e.g., Ampere, Blackwell) and experience debugging/optimizing GPU kernels using tools like Nsight.

  • Strong knowledge of C++ and experience in large code bases, comfortable in Linux development environments.

  • Experience in development, debugging, and profiling of complex multiprocess systems (e.g., robotic systems, game engines).

Bonus Qualifications

    GPU kernel development experience

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,469,448 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Boston
$223k – $253k per year • In office • Full-Time • 5+ years exp • San Francisco
AI/ML
Transformers
Machine Learning
Apply
Applied AI Engineer 5 months ago
$200k – $300k per year • Remote (United States) • Full-Time • 5+ years exp
Python
Databases
Weaviate
Neo4j
Milvus
Pinecone
FAISS
Dgraph
ArangoDB
AI/ML
LangGraph
AutoGen
LangChain
Model Context Protocol
Function Calling
AI Agents
Arize Phoenix
DeepEval
Langfuse
LangSmith
Pydantic AI
Semantic Kernel
LLM
RAG
Hybrid Search
Structured Outputs
Agentic Workflows
Tool Use
DevOps
Rest API
Azure
Apply
≈ $184k – $378k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco
Python
JavaScript
TypeScript
Node JS
Python
Celery
Node JS
Prisma
Databases
PostgreSQL
Supabase
pgvector
AI/ML
Model Context Protocol
AI Agents
LLM
RAG
Frontend
Tailwind CSS
Next.js
React.js
DevOps
GitHub Actions
Vercel
QA
Playwright
Apply
Founding ML Engineer 4 months ago
≈ $136k – $262k per year (Estimated) • In office • Full-Time • 3+ years exp • Master's Degree • New York • San Francisco
Python
Rust
SQL
C++
C++
TensorFlow C++
PyTorch C++
Databases
Google Bigtable
AI/ML
Claude Code
TensorFlow
Keras
PyTorch
Time Series Forecasting
Machine Learning
DevOps
GCP
Azure
Docker
Kubernetes
Apply
≈ $190k – $390k per year (Estimated) • In office • Full-Time • San Francisco • New York
AI/ML
OpenAI
DevOps
AWS
Apply
Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Chengdu
C++
DevOps
KVM
Xen
Linux
Management
Agile
Apply
In office • Beijing
Python
Go
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
TensorFlow
PyTorch
NCCL
InfiniBand
DevOps
CI/CD
Linux
Apply
≈ $100k – $200k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Foster City
Java
C++
Apply
≈ $27k – $67k per year (Estimated) • Remote (China) • Full-Time • 10+ years exp • Bachelor's Degree • Chengdu
AI/ML
LLM
DevOps
Linux
Management
Agile
Apply
$11k – $18k per year • In office • Bachelor's Degree • Moscow
1C
DevOps
Linux
Windows
Apply
$276k – $343k per year • Hybrid • Full-Time • 8+ years exp • Foster City
AI/ML
Ray Serve
CUDA Toolkit
Quantization
JAX
Multimodal AI
Knowledge Distillation
TensorRT
PyTorch
Ray
CUDA
Triton
FSDP
Apache TVM
XLA
Model Distillation
Apply
$209k – $263k per year • Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Boston • Foster City
C++
AI/ML
CUDA Toolkit
CUDA
Machine Learning
DevOps
Linux
Apply
$230k – $300k per year • Hybrid • Full-Time • 10+ years exp • Foster City
Python
Databases
Weaviate
Milvus
Pinecone
OpenSearch
AI/ML
LangChain
Claude
LlamaIndex
Model Context Protocol
Vertex AI
AI Agents
AWS Bedrock
LLM
OpenAI
OpenAI Agents SDK
LLM Guardrails
Agentic Workflows
DevOps
AWS
Cloudflare
IAM
Cybersecurity
Okta
Microsoft Entra ID
Apply
≈ $206k – $455k per year (Estimated) • Hybrid • Full-Time • Foster City
Python
SQL
Scala
Python
pySpark
AI/ML
Spark
Fine-tuning
Multimodal AI
Function Calling
Computer Vision
NLP
VLM
LLM
RAG
Human-in-the-Loop
Agentic Workflows
Tool Use
Machine Learning
DevOps
Git
Apply
$177k – $270k per year • Hybrid • Full-Time • 2+ years exp • Foster City
AI/ML
DeepSpeed
Reinforcement Learning
JAX
PyTorch
Ray
Edge AI
DevOps
AWS
Apply
≈ $23k – $40k per year (Estimated) • In office • 1+ year exp • Boston
Apply
Line Cook | Lifted 6 hours ago
≈ $23k – $40k per year (Estimated) • In office • Boston
Apply
Pastry Line Cook 6 hours ago
≈ $23k – $40k per year (Estimated) • In office • Boston
Apply
Massage Therapist 6 hours ago
≈ $23k – $40k per year (Estimated) • In office • 1+ year exp • Boston
Apply
Spa Attendant 6 hours ago
≈ $23k – $40k per year (Estimated) • In office • Boston
Apply
See all jobs
This is one of many
1,469,448 more open roles from verified company boards, updated every day.