683,840open jobs
39,578companies
97,862added this week
Browse all
Salary
$150k – $250k per year
Location
In office (San Francisco)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
TypeSafe AI is an AI lab building machine-native intelligence infrastructure for automation, designed to make decisions within software. Try our first System One Model, Jev, in early access.

About TypeSafe

TypeSafe is an AI lab building intelligence beyond chat: AI designed to work inside software and power real-world automation. Our mission is to make intelligence composable-dependable enough for builders to invoke, inspect, constrain, and layer into production systems.

Today’s models are remarkably capable, but that intelligence is still hard to build on. We’re rethinking the AI stack from first principles so software can branch on meaning, judgment, and intent while code continues to handle exact computation.

We’re a small, fast-moving team from OpenAI, Google Brain, and Meta/FAIR. We care about technical rigor, real-world impact, and building systems people can trust. Join us to help turn today’s intelligence into the foundation for a new generation of software.

About the role

We're looking for an Infrastructure Engineer to build and operate the infrastructure behind TypeSafe AI's products at global scale. You'll own the systems that serve millions of users across regions - from provisioning Kubernetes clusters across multiple clouds to optimizing networking for low-latency AI inference.

This is a high-impact role on a small, fast-moving team. You'll work across the full infrastructure stack: cloud primitives, container orchestration, networking, observability, and the specialized infra that makes large-scale model inference efficient.

What you'll do

  • Design, deploy, and operate Kubernetes clusters across multiple regions and clouds

  • Build and maintain infrastructure for the platform that powers LLM inference workloads globally

  • Own networking, including VPCs, peering, load balancing, DNS, service mesh, CNI

  • Manage GPU infrastructure and autoscaling for ML workloads

  • Write and maintain infrastructure as code (Pulumi / Python)

  • Operate and improve observability: monitoring, alerting, tracing, logging

Requirements

  • Deep experience with Kubernetes in production at scale: networking, storage, scheduling, upgrades

  • Strong background in AWS

  • Hands-on experience with infrastructure as code (Pulumi, Terraform, or similar)

  • Solid understanding of Linux networking

  • Track record with high-traffic production ML systems

  • Programming fluency, Python preferred

Nice to have

  • Experience with large-scale LLM / ML inference infrastructure (GPU scheduling, model serving, vLLM, KubeRay, Kubernetes-native tooling)

  • Kubernetes networking depth with Cilium or other CNI plugins; service mesh (Istio, Envoy)

  • Multi-cloud infrastructure

  • Background in site reliability engineering including SLOs, incident response, capacity planning

Life at TypeSafe

We’re a small, flat, close-knit team working to make intelligence dependable enough to become part of everyday software. We work fully in person from our San Francisco office near Embarcadero station. We love what we do and care deeply about the work.

We strive for excellence and craftsmanship and won’t stop until we get there. When the team wins, we all win, and we enjoy collaborating and inspiring each other to grow-as a team and as individuals.

We value emotional honesty, kindness, and bringing your whole self to work. We build machines; we don’t try to be machines.

We want TypeSafe to be the place where you do the most impactful work of your career and help define our future as a company.

We provide

  • Base salary of $150k-250k plus equity, based on leveling

  • 100% covered health insurance

  • Daily lunch and dinner

  • Visa sponsorships

  • 401K plans

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
683,840 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$81k – $156k per year (Estimated) • Remote • Full-Time • 1+ year exp
Python
Go
DevOps
Terraform
OpenShift
Helm
OpenTelemetry
cert-manager
Kustomize
Prometheus
CI/CD
GitOps
ArgoCD
Kubernetes
Grafana
Platform Engineering
Configuration Management
SLI/SLO/SLA
Apply
$81k – $164k per year (Estimated) • Remote • Secret • Contractor • 5+ years exp • PhD
Python
Java
PowerShell
Java
Spring Boot
DevOps
Terraform
OpenShift
Azure DevOps
Azure
CI/CD
Git
Docker
Kubernetes
Bicep
Management
Power Automate
Power Apps
Apply
$94k – $175k per year (Estimated) • Remote • Full-Time • 1+ year exp
Python
Go
DevOps
Terraform
OpenShift
Helm
OpenTelemetry
cert-manager
Kustomize
Prometheus
CI/CD
GitOps
ArgoCD
Kubernetes
Grafana
Platform Engineering
Configuration Management
SLI/SLO/SLA
Apply
$72k – $151k per year (Estimated) • Remote • Secret • Full-Time • 3+ years exp
Python
Java
Kotlin
Java
Spring Boot
Quarkus
Apache Camel
Databases
ActiveMQ
Apache Kafka
DevOps
Rest API
OpenShift
Prometheus
CI/CD
Jenkins
Git
Kubernetes
Grafana
Tekton
Management
Agile
QA
Swagger
Apply
$120k – $135k per year • In office • 6+ years exp • Bachelor's Degree
DevOps
Azure
AWS
Apply
$150k – $250k per year • In office • Full-Time • 5+ years exp • San Francisco
Python
JavaScript
TypeScript
AI/ML
Claude Code
OpenAI
OpenAI Codex
Frontend
Tailwind CSS
Next.js
React.js
DevOps
PagerDuty
CI/CD
Kubernetes
Apply
$150k – $250k per year • In office • Full-Time • 2+ years exp • San Francisco
Python
JavaScript
TypeScript
AI/ML
Cursor
Claude Code
OpenAI
Frontend
Tailwind CSS
Next.js
React.js
DevOps
Kubernetes
Apply
$73k – $133k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • Los Angeles • San Jose • San Francisco • San Diego • Bellevue
Design
AutoCAD
Apply
$148k – $222k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Francisco
Apply
$70k – $111k per year • In office • Full-Time • 6+ years exp • Bachelor's Degree • Minneapolis • Seattle • San Francisco • Portland
Cybersecurity
Wireshark
Apply
$200k – $275k per year • In office • Full-Time • San Francisco
DevOps
Kubernetes
Apply
$79k – $188k per year (Estimated) • Remote • Full-Time • 3+ years exp • San Francisco
JavaScript
Frontend
React.js
Apply
See all jobs
This is one of many
683,840 more open roles from verified company boards, updated every day.