368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$220k – $320k per year
Location
In office (Redwood City)
Seniority
Senior · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Dyna Robotics is a robotics foundation model company headquartered in Redwood City, California, and founded in 2024. The company trains general purpose manipulation models that let a single robot arm learn commercial tasks such as folding, packing, and handling from demonstration rather than explicit programming. It targets repetitive physical work in laundries, restaurants, and light manufacturing, and publishes model progress through its DYNA series of releases.

Dyna Robotics builds general-purpose robots powered by a proprietary embodied AI foundation model with top-in-industry generalization and real-world performance. Already deployed with customers across multiple industries, our robots do commercial-grade work in the physical world. Our team comes from Google DeepMind, Meta, and Cruise, and we're backed by CRV, First Round, and other leading investors.

The Role

As a ML Training Infrastructure Engineer, you will architect and build the systems that turn our multi-cloud GPU fleet into a training engine our researchers love. Your charter is singular and broad: own training infrastructure end-to-end so that every GPU is busy, every run is reproducible, and every researcher's next experiment is one command away.

What You’ll Do

  • Scale Distributed Training: Architect and own the infrastructure for large-scale GPU clusters. You’ll implement sharding, activation checkpointing, and memory optimization (ZeRO, FSDP) to enable the training of massive multimodal models.

  • Optimize Researcher Ergonomics: Build a research codebase and job scheduling system (Kubernetes/SLURM) that prioritizes fast iteration, automated retries, and seamless failure recovery.

  • High-Performance Data Handling: Design high-throughput pipelines to ingest and transform terabytes of multimodal robot data (video, proprioception, 3D signals), ensuring dataloaders never starve the GPUs.

  • Production Inference: Build low-latency inference pipelines for real-time robot control. You’ll apply quantization, distillation, and model compilation (TensorRT, Triton) to move models from the lab to the physical world.

  • Deep Systems Profiling: Dive into the weeds of GPU utilization, I/O bottlenecks, and memory fragmentation to squeeze every bit of performance out of our expanding compute fleet.

What You’ll Bring

  • 7+ Years of Engineering: With a track record of leading technical projects in high-performance computing (HPC) or ML infrastructure.

  • ML Systems Mastery: Deep experience with PyTorch and distributed training frameworks (DeepSpeed, Accelerate). You understand the nuances of mixed precision and gradient accumulation.

  • Infrastructure Expertise: Hands-on experience managing cloud GPU environments (GCP/AWS) and container orchestration (Kubernetes).

  • Low-Level Intuition: A fundamental understanding of distributed systems, including race conditions, memory management, and NCCL/inter-node communication.

  • Ownership Mindset: You don't just "deploy" code; you design, build, and operate systems end-to-end to unblock fast-moving research.

Bonus Points For

  • Experience with Robotics Data Formats (MCAP, Protobuf) or multimodal models (VLAs).

  • Deep ML systems experience: custom kernels (Triton), compilers, or runtime optimization.

  • Experience as a founding or early-stage infrastructure hire.

At Dyna Robotics, we build technology for the real world, which requires a team as diverse as the environments our robots inhabit. We are an equal opportunity employer committed to technical rigor and mutual respect.

Don’t let a checklist stop you. Data shows that underrepresented groups often only apply if they meet 100% of the criteria. We value problem-solving and grit over keyword matching. If you’re passionate about the intersection of geometry and robotics, we want to hear from you-even if you don't check every box.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Redwood City
$93k – $126k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • United States
Node JS
SQL
TypeScript
JavaScript
Databases
MS SQL
Oracle
Mobile
JUnit
DevOps
AWS
Azure
CI/CD
GCP
Git
GitLab
GitLab CI
Jenkins
Management
Jira
QA
JMeter
Playwright
Postman
Rest-Assured
TestNG
Apply
$195k – $264k per year • In office • Full-Time • 15+ years exp • Master's Degree • United States
Python
AI/ML
Amazon SageMaker
Keras
Kubeflow
MLFlow
PyTorch
Scikit-learn
TensorFlow
Vertex AI
XGBoost
DevOps
AWS
Azure
CI/CD
CloudFormation
Docker
GCP
Kubernetes
Terraform
Cybersecurity
FedRAMP
NIST 800-53
Apply
$68k – $142k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree • Marseille
AI/ML
Knowledge Graph
DevOps
AWS
Azure
GCP
Management
ServiceNow
Apply
Cloud Scrum Master 8 hours ago
$31k – $64k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Hyderabad
DevOps
AWS
Azure
GCP
Management
Confluence
Apply
Data Scientist 9 hours ago
$17k – $47k per year (Estimated) • Remote/Hybrid • Full-Time • 1+ year exp • Bachelor's Degree • Pune
Python
SQL
AI/ML
MLFlow
PyTorch
RAG
Scikit-learn
TensorFlow
DevOps
AWS
Azure
GCP
Analytics
Matplotlib
Seaborn
Apply
$180k – $400k per year • In office • Full-Time • Redwood City
Apply
$167k – $309k per year (Estimated) • In office • Full-Time • 4+ years exp • Redwood City
AI/ML
Multimodal AI
Human-in-the-Loop
Time Series Forecasting
Robotics
Teleoperation
Apply
$160k – $200k per year • In office • Full-Time • Redwood City
Python
Databases
Amazon Aurora
AI/ML
Time Series Forecasting
DevOps
Datadog
Robotics
MuJoCo
Teleoperation
Marketing
Salesforce
Apply
$152k – $375k per year (Estimated) • In office • Full-Time • PhD • Redwood City
Python
Databases
Amazon Aurora
AI/ML
CUDA Toolkit
Diffusion Models
Gaussian Splatting
JAX
PyTorch
NeRF
Robotics
Imitation Learning
Isaac Lab
Isaac Sim
MuJoCo
Sim-to-Real
Design
Blender
Marketing
Salesforce
Apply
$81k – $229k per year (Estimated) • In office • Full-Time • 4+ years exp • Redwood City
Apply
$79k – $144k per year (Estimated) • Remote • 3+ years exp • Bachelor's Degree • Redwood City
AI/ML
Spark
DevOps
GitHub
Apply
$91k – $182k per year (Estimated) • In office • 3+ years exp • Redwood City
Python
DevOps
AWS
CI/CD
GCP
Terraform
Apply
$96k – $120k per year • In office • Internship • Bachelor's Degree • Redwood City
JavaScript
AI/ML
AI Agents
Apply
$96k – $120k per year • In office • Internship • Master's Degree • Redwood City
JavaScript
Python
AI/ML
AI Agents
Time Series Forecasting
DevOps
GitHub
Apply
$83k – $176k per year (Estimated) • In office • 12+ years exp • Bachelor's Degree • Redwood City
AI/ML
OpenAI Codex
Management
Confluence
Jira
Smartsheet
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.