368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$165k – $330k per year
Location
Remote/Hybrid (San Francisco, United States)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Baseten is an AI infrastructure platform designed to help developers and machine learning teams deploy, serve, and scale open-source and custom AI models. The platform provides performant, low-latency inference infrastructure alongside developer tools like Truss, an open-source model packaging framework. Headquartered in San Francisco, California, Baseten enables companies to run state-of-the-art models in production seamlessly without managing underlying cloud infrastructure.

ABOUT BASETEN

Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.

THE ROLE

As the Engineering Manager for Baseten's Cloud Platform team, you will directly manage a team of cloud platform engineers responsible for building the systems and processes that keep our infrastructure scalable, reliable, and efficient - from automated deployments and monitoring to performance optimization and incident response.

You are a people-first leader with a strong cloud infrastructure background. You set a high bar for reliability and operational excellence, engage credibly in technical discussions and code reviews, and know how to build a culture of ownership and accountability. You'll spend most of your time close to the work: unblocking your team, shaping technical direction on day-to-day decisions, and developing your engineers. At Baseten, we work closely with our users to understand their struggles operationalizing ML - you'll keep your team connected to that mission and translate user learnings into better infrastructure.

RESPONSIBILITIES

  • Recruit, hire, and grow a high-performing team of cloud platform engineers; provide ongoing coaching, feedback, and career development through regular 1:1s.

  • Set clear performance expectations, hold a high bar, and create an environment where engineers do their best work.

  • Foster a culture of ownership, accountability, and continuous improvement.

  • Drive day-to-day technical decisions through design reviews, code reviews, and architectural discussions; translate the infrastructure roadmap into clear team priorities and milestones, and hold execution against them.

  • Establish standards and best practices for reliability, performance, and operational excellence; ensure the team owns projects end-to-end, from specification through production.

  • Exercise and encourage good judgment on tooling and architectural tradeoffs, with a bias against unnecessary complexity.

  • Partner with product and engineering teams to align infrastructure work with business priorities; represent your team's progress, capacity, and technical tradeoffs clearly to leadership.

  • Serve as the escalation point for major incidents; drive resolution with urgency, ensure the team learns systematically, and stay connected to users so those learnings feed back into infrastructure improvements.

REQUIREMENTS

  • Proven experience directly managing engineers in a cloud infrastructure, platform engineering, or SRE context (not managing managers).

  • Extensive hands-on experience with Kubernetes, with the ability to engage credibly in architectural discussions and code reviews.

  • Strong background building and maintaining scalable, production infrastructure.

  • Familiarity with infrastructure-as-code tools (e.g., Terraform, CloudFormation, Pulumi) and CI/CD tooling (e.g., GitHub Actions, GitLab CI, CircleCI, Jenkins).

  • Demonstrated ability to drive projects end-to-end - from specification to execution - and coach your team to do the same.

  • Strong written and verbal communication skills; able to influence across teams without direct authority.

  • Track record of recruiting and growing engineering talent.

  • Openness to learning about ML infrastructure; prior machine learning experience is not required.

NICE TO HAVE

  • Experience with OSS observability tooling (Prometheus, ELK stack, Grafana, OpenTelemetry).

  • Background in multi-cloud environments or GPU infrastructure.

  • Experience building and managing on-call rotations and incident response processes.

BENEFITS

  • Competitive compensation, including meaningful equity.

  • 100% coverage of medical, dental, and vision insurance for employee and dependents

  • Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)

  • Paid parental leave

  • Fertility and family-building stipend through Carrot

  • Company-facilitated 401(k)

  • Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.

Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.

At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.

We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$105k – $252k per year • Remote • Full-Time • 18+ years exp • Bachelor's Degree
Python
Java
Java
Gradle
DevOps
Ansible
AWS
CI/CD
CloudFormation
Configuration Management
Docker
GitHub Actions
GitLab CI
Helm
Jenkins
Kubernetes
Platform Engineering
Terraform
GitHub
GitLab
Cybersecurity
Sonatype Nexus IQ
Management
Confluence
Jira
Apply
$133k – $161k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Westminster
Bash
C++
Java
Python
DevOps
CI/CD
Git
Apply
$54k – $175k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
TypeScript
Frontend
React.js
Sass
DevOps
AWS
CI/CD
Docker
Kubernetes
Rest API
Apply
$35k – $113k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
TypeScript
Frontend
React.js
Sass
DevOps
AWS
CI/CD
Docker
Kubernetes
Rest API
Apply
$35k – $116k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
TypeScript
Frontend
React.js
Sass
DevOps
AWS
CI/CD
Docker
Kubernetes
Rest API
Apply
$165k – $330k per year • Remote/Hybrid • Full-Time • San Francisco • Toronto • New York • Montreal
AI/ML
Cursor
DeepSeek
LLM
SGLang
TensorRT
TensorRT-LLM
vLLM
Management
Notion
Apply
$190k – $240k per year • Remote/Hybrid • Full-Time • San Francisco • New York
JavaScript
TypeScript
AI/ML
Cursor
Frontend
Next.js
React Three Fiber
React.js
Three.JS
Design
Figma
Apply
Product Designer 5 days ago
$225k – $290k per year • Remote/Hybrid • Full-Time • San Francisco • New York
JavaScript
TypeScript
AI/ML
Cursor
Frontend
React.js
DevOps
Git
Apply
AI Engineer 12 days ago
$220k – $260k per year • Remote/Hybrid • Full-Time • 5+ years exp • San Francisco
Python
AI/ML
Claude
Claude Code
Cursor
Fine-tuning
LangChain
LLM
LoRA
Reinforcement Learning
Spark
Synthetic Data
PEFT
LLM Guardrails
OpenAI Codex
Post-training
SFT
AI Agents
Model Context Protocol
Apply
$170k – $220k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco • New York
Go
AI/ML
Cursor
LLM
Post-training
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
$160k – $283k per year • Equity • In office • 5+ years exp • San Francisco
AI/ML
AI Agents
Apply
$83k – $188k per year (Estimated) • In office • 2+ years exp • San Francisco
Python
AI/ML
AI Agents
LLM Guardrails
Model Context Protocol
DevOps
Terraform
Cybersecurity
Crowdstrike
GDPR
Least Privilege
Okta
SentinelOne
Management
Google Workspace
Slack
Apply
$171k – $273k per year • In office • Full-Time • 8+ years exp • PhD • San Francisco • Washington
AI/ML
A2A
Agentforce
AI Agents
Model Context Protocol
DevOps
AWS
GCP
Marketing
Salesforce
Apply
Security GRC Analyst 2 hours ago
$119k – $268k per year (Estimated) • Remote/Hybrid • 4+ years exp • Bachelor's Degree • San Francisco
AI/ML
Ignite
PyTorch
Cybersecurity
ISO 27001
NIST CSF
SOC 2
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.