1,438,441open jobs
85,144companies
219,560added this week
Browse all
Salary
≈ $149k – $297k per year (Estimated)
Location
In office (San Francisco)
Seniority
Middle · 4+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 11, 2026. First seen by Alion on Oct 9, 2026. Physical Intelligence scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Physical Intelligence is a San Francisco research company founded in 2024 that is building general-purpose foundation models for robots. Its flagship models are trained on data from many robot embodiments and tasks, aiming for a single policy that can be adapted to new hardware and chores. The company raised very large early rounds from investors including OpenAI, Jeff Bezos and Thrive Capital.

Who We Are

Physical Intelligence is bringing general-purpose AI into the physical world. We are a team of engineers, scientists, roboticists, and company builders developing foundation models and learning algorithms to power the robots of today and the physically-actuated devices of the future.

The Team

The Infrastructure team builds and operates the backbone of everything PI does: from training state-of-the-art VLA models, to orchestrating large-scale simulation, to reliably deploying intelligence across fleets of physical robots. The team works closely with researchers, robotics runtime, product, and platform engineers to ensure infrastructure scales from prototype to production-grade deployments.

In This Role You Will

  • Own and scale AI-native infrastructure: You will operate and evolve Kubernetes clusters and service deployment patterns, and help build a scalable microservice platform for internal systems such as evaluation services, operational tooling, and internal APIs with agent-use as the primary interface. This includes supporting safe rollouts, upgrades, and rollback strategies.

  • Drive security, observability and cost-aware infrastructure: You will treat security, logging, metrics, tracing, and alerting as first-class platform primitives, and build systems that surface reliability and performance issues early. You will also help improve cost visibility and enable cost-aware decision-making at the infrastructure level.

  • Harden platform foundations: You will own core infrastructure with multi-cloud considerations, designing authentication/authorization flows, networking architecture, quota and rate-limiting services, and cloud primitives that behave predictably. A major part of this work is reducing infra churn by standardizing patterns, abstractions, and interfaces.

  • Improve developer experience: You will build clear, documented interfaces for using platform infrastructure, reducing the gap between “I need infra” and “I can run my workload.” This includes supporting consistent local vs. remote development workflows and improving self-serve infrastructure usage.

  • Collaborate and lead through ownership: You will work closely with researchers and other engineers to understand requirements and constraints, translate fast-moving needs into reusable infrastructure, and own systems end-to-end, from design through operation.

What We Hope You'll Bring

  • Extremely strong first-principles thinking.

  • Familiarity with agent infrastructure, observability, security and sandboxing

  • Deep experience with cloud platforms (GCP, AWS) and distributed systems: compute orchestration, networking, autoscaling, service meshes, load balancing.

  • Ability to reason about system bottlenecks, performance tuning, and cost optimizations across compute, networking, and storage.

  • Comfort with Kubernetes, cluster-level reliability, and service-oriented architectures.

  • Solid intuition around scalability, performance, and failure modes.

  • Experience with infrastructure-as-code (e.g. Terraform), containerization, and modern platform engineering practices.

  • Familiarity with logging, metrics, tracing, incident response, SLOs, and debugging complex distributed systems.

  • Strong cross-functional communication and ownership mindset.

  • Experience (4-6 years) working in fast-moving or early-stage environments where ambiguity is normal with demonstrated growth trajectory.

Ultimately, we’re looking for someone who can spin up quickly on unfamiliar and ambiguous domains, has a strong sense of ownership, and cares deeply about our mission. Even if you don’t check all the boxes above, we strongly encourage you to apply if this sounds like you.

Bonus Points If You Have

  • Experience with large-scale ML training, evaluation, or simulation infrastructure.

  • Experience with secrets management systems (e.g., Doppler).

  • Background in observability, cost optimization, or internal platform tooling.

  • Exposure to robotics, simulation, or real-time systems.

Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,438,441 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
San Francisco
$195k – $262k per year • Remote (United States) • Palo Alto
Python
AI/ML
DeepSpeed
CUDA Toolkit
RLHF
Function Calling
TRL
Transformers
PyTorch
LLM
Ray
Synthetic Data
CUDA
Triton
DPO
SFT
PPO
GRPO
Post-training
Megatron-LM
FSDP
NCCL
InfiniBand
Agentic Workflows
Tool Use
RLAIF
Reward Modeling
Machine Learning
DevOps
SLURM
Kubernetes
Apply
$195k – $262k per year • Remote (United States) • Palo Alto • San Francisco
Python
AI/ML
Ray Serve
vLLM
Triton Inference Server
Quantization
Knowledge Distillation
Function Calling
AI Agents
VLM
SGLang
AWQ
GPTQ
TensorRT
TensorRT-LLM
PyTorch
LLM
Ray
Mixture of Experts
KServe
TPOT
Triton
Post-training
NCCL
InfiniBand
NVLink
Structured Outputs
Speculative Decoding
KV Cache
Tool Use
Model Distillation
Machine Learning
Apply
$180k – $225k per year • Remote (United States) • United States
AI/ML
vLLM
TensorRT
Apply
$195k – $262k per year • Remote (United States) • PhD • Palo Alto • San Francisco
Python
AI/ML
vLLM
CUDA Toolkit
RLHF
Quantization
Knowledge Distillation
VLM
SGLang
TensorRT
TensorRT-LLM
PyTorch
LLM
Mixture of Experts
Synthetic Data
CUDA
Triton
DPO
SFT
Post-training
Speculative Decoding
KV Cache
Model Distillation
RLAIF
Machine Learning
Apply
≈ $144k – $255k per year (Estimated) • Remote (United States) • 5+ years exp • United States
AI/ML
NCCL
NVLink
DevOps
Linux
Apply
≈ $61k – $151k per year (Estimated) • In office • Full-Time • Leeds
Python
Bash
DevOps
Terraform
GCP
Istio
OpenTelemetry
cert-manager
Dynatrace
Prometheus
CI/CD
GitOps
Kubernetes
Service Mesh
Google GKE
Linux
DNS
VPN
Cybersecurity
Aqua Security
Open Policy Agent
Least Privilege
OPA Gatekeeper
PKI
Apply
Solution Architect 1 day ago
≈ $81k – $165k per year (Estimated) • In office • Full-Time • Manchester
DevOps
GCP
Docker
Kubernetes
Management
Agile
Apply
≈ $75k – $186k per year (Estimated) • In office • Contractor • London
JavaScript
TypeScript
SQL
Node JS
Node JS
Nest.JS
TypeORM
Databases
PostgreSQL
DynamoDB
Amazon Aurora
Frontend
Next.js
React.js
DevOps
Terraform
CI/CD
AWS
GitHub
Amazon ECS
Apply
≈ $82k – $160k per year (Estimated) • In office • Full-Time • 7+ years exp • Dublin
SQL
DevOps
Terraform
Ansible
Azure DevOps
Jaeger
OpenTelemetry
Azure
JFrog Artifactory
Linux
Windows
SOAP
Cybersecurity
SonarQube
Apply
≈ $82k – $168k per year (Estimated) • In office • Full-Time • 6+ years exp • Dublin
Python
Python
FastAPI
Databases
PostgreSQL
Neo4j
pgvector
Pinecone
AI/ML
LangChain
Claude
LlamaIndex
Scikit-learn
Prompt Engineering
AI Agents
NLP
Llama
Mistral
Transformers
TensorFlow
Pandas
NumPy
PyTorch
Gemini
LLM
RAG
OpenAI
Anthropic
LLM Guardrails
DevOps
OpenShift
Azure DevOps
GitLab CI
Azure
CI/CD
ArgoCD
Jenkins
Kubernetes
Apply
≈ $160k – $353k per year (Estimated) • In office • Full-Time • San Francisco
Python
TypeScript
Databases
PostgreSQL
ClickHouse
AI/ML
Fine-tuning
LLM
Machine Learning
DevOps
GCP
WebSockets
Kubernetes
Apply
≈ $162k – $358k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
JAX
Multimodal AI
PyTorch
TPU
DevOps
GCP
SLURM
AWS
Kubernetes
Google GKE
Apply
≈ $163k – $361k per year (Estimated) • In office • Full-Time • San Francisco
Databases
ClickHouse
AI/ML
Multimodal AI
Flink
Ray
Machine Learning
Apply
Applied Researcher 9 months ago
≈ $160k – $353k per year (Estimated) • In office • Full-Time • PhD • San Francisco
Python
AI/ML
Fine-tuning
Apply
Research Scientist 2 years ago
≈ $160k – $355k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Machine Learning
Apply
≈ $256k – $461k per year (Estimated) • Equity • In office • 8+ years exp • San Francisco
Cybersecurity
HIPAA
Apply
≈ $227k – $410k per year (Estimated) • Equity • In office • 8+ years exp • San Francisco
Cybersecurity
HIPAA
Apply
≈ $256k – $461k per year (Estimated) • Equity • In office • 8+ years exp • San Francisco
Cybersecurity
HIPAA
Apply
$205k – $250k per year • Hybrid • Full-Time • 7+ years exp • Bachelor's Degree • San Francisco
Python
AI/ML
Red Teaming
DevOps
Terraform
Ansible
GitHub Actions
CI/CD
AWS
AWS Lambda
IAM
Amazon ECS
Linux
DNS
DHCP
Cybersecurity
pfSense
PKI
Apply
$132k – $170k per year • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Washington • San Francisco
Apply
See all jobs
This is one of many
1,438,441 more open roles from verified company boards, updated every day.