Actively hiring
MLOps
AI Compute & Inference
As featured in
View 53 jobs
Follow
Overview
News
Technologies
Salaries
Products
People
Growth
Offices
Jobs
Financials
Funding and backers
Total raised
$1.7B
3 rounds on record
Latest round
$1.5B
Series F · Jun 2026 · 3 mo ago
Team
1001-5000
employees, estimated range
Hiring now
100
+28 posted in 30 days
Backers
14
AI Fund, South Park Commons
The money's age
Fresh money
3 months since the last round · and hiring
Momentum
56/100
hiring for its size, a new city, in the news, fresh money
Last sign of life
this month
a posting or a news item
Tractionpublic signals, read monthly
Funding rounds
| Round | Date | Amount | Investors | Source |
|---|---|---|---|---|
| Series F | Jun 2026 | $1.5B | — | TNW Artificial-intelligence |
| Series D | Sep 2025 | $150M | Bond lead | Fortune |
| Venture round | Feb 2025 | $75M | — | NBC New York |
Backed byfrom the funds' portfolios and accelerator catalogs
Similar startupsby what they do, their industry and backers

AI Cloud for All Models. Train, fine-tune, and serve across clouds: self-serve A100 and H100, B200 to GB300 as reserved capacity…

Makora inference endpoints use AI agents to optimize every layer of the stack — from GPU kernels and inference engines to algorithms and…

Deploy machine learning models that run in microseconds with low latency AI inference from myrtle.ai.

Modal is an AI infrastructure company headquartered in New York City and founded in 2021. It provides a serverless cloud platform featuring…

Lightning AI is an artificial intelligence technology company headquartered in New York City, New York, and established in 2019. The…

Hugging Face is an American company, founded in 2016 and headquartered in New York, that runs the largest open platform for sharing machine…
How Baseten hires
Truth index
B 73/100
Median time to fill
88 days
Ghost roles
0%
Stale roles
90%
Reposted roles
0%
Baseten scores B on the Alion truth index (A is best, E worst), measured over 19 open roles re-checked against its own hiring board: 0% look like ghost listings, 90% have gone stale and 0% are the same role posted again. Its roles close in a median of 88 days. It measures how much of what this employer publishes turns into a real, finite search - not the employer as a workplace. How the index is computed
Overview
Baseten is an American company founded in 2019 that runs machine learning models in production for companies that would rather not operate GPU infrastructure themselves. Its position is inference rather than training: it handles model packaging, autoscaling, cold start latency and multi-cloud capacity, which are the unglamorous problems that determine whether a model-powered product is fast and affordable enough to ship. Headquartered in San Francisco and backed at a multi-billion dollar valuation, it serves companies deploying open and custom models, and it competes with both the hyperscalers and the model providers' own hosted endpoints.
News
Modal, Fireworks, and Baseten gain cost advantage with Nvidia and AMD chips for Kimi K3
Modal, Fireworks AI, and Baseten leverage Nvidia GB300 and AMD MI350X chips to serve Kimi K3 at a fraction of the cost of direct Chinese API
Read more
Report
Baseten Acquires Blaxel To Build Integrated Agentic AI Infrastructure Platform
Baseten has acquired Blaxel, bringing together Baseten's AI model inference and training infrastructure with Blaxel's execution, storage and networking technology for autonomous agents. Financial terms of the transaction were not disclosed.
Read more
Report
Baseten acquires AI agent sandbox startup Blaxel
Baseten has acquired Blaxel, which builds sandboxes for AI agents. Terms were not disclosed. Blaxel keeps running, and Baseten will build on its tech.
Read more
Report
The efficient frontier of LLM inference
Inference techniques either move a deployment along the latency-throughput frontier or push the entire frontier out, creating more efficiency to allocate.
Read more
Report
Pro access
Upgrade to see all 45 mentions
Upgrade to a paid plan to read every media mention of this company - funding news, awards, product launches and press releases from all the outlets writing about it.
Every media mention and press release
Funding news, awards and product launches
Fresh coverage from every outlet writing about the company
Upgrade now
Cancel anytime. Secure checkout. Instant activation.
Tech stack
Technologies named in job postings at Baseten over the last 12 months, by layer, with the share of postings for each.
Technologies in use
147
Postings analysed
109
Stack modernity
79/100AI adopter
Languages
Python17%
Go8%
SQL7%
JavaScript4%
Frameworks & libraries
SGLang10%
TensorRT-LLM6%
PyTorch5%
OpenTelemetry4%
Databases
BigQuery4%
Google BigQuery4%
ClickHouse1%
PostgreSQL1%
Cloud & platforms
Kubernetes19%
Spark17%
vLLM10%
AWS7%
AI & ML
LLM13%
Claude6%
Whisper5%
DeepSeek2%
Sign in to see how your skills match this stack
Sign In
Full Baseten tech stack
147 technologies · by team · pay by technology · changes · similar stacks
Salary insights
Compensation ranges based on this company's open jobs
Median
$260k
Typical range
$60k – $400k
Open positions
95
67
Jobs
Junior
10.4%
Middle
19.4%
Senior
34.3%
Staff
28.4%
Architect
7.5%
| Level | Jobs | Median | Typical range |
|---|---|---|---|
| Junior | 7 | $230k | $110k – $400k |
| Middle | 13 | $220k | $190k – $360k |
| Senior | 23 | $235k | $160k – $360k |
| Staff | 19 | $275k | $205k – $380k |
| Architect | 5 | $180k | $170k – $330k |
Highest paying role
Forward Deployed Engineer (Training)
$400k
Normalized to USD per year
Products
Training and Fine-Tuning
Managed fine-tuning of open models on customer data. It captures workloads earlier in the lifecycle, so the resulting model is already running on the platform by the time it reaches actual production traffic.
baseten.co
Multi-Cloud Capacity
Access to GPU capacity across several cloud providers and regions. Availability is uneven and shifts constantly, so sourcing across clouds is a practical necessity rather than any kind of architectural preference.
baseten.co
Dedicated Deployments
Isolated GPU capacity reserved for a single customer's workload. Shared inference introduces latency variance that a customer-facing product frequently cannot tolerate at all, whatever the cost saving.
baseten.co
Truss
An open source format for packaging a model with its dependencies and serving configuration. Standardising packaging is what makes a model portable between environments instead of tied to whoever deployed it first.
baseten.co
Model Inference
Hosted serving of open and custom models with autoscaling and monitoring. Inference economics rather than model quality usually decide whether a feature ships, since a model too slow or costly to run is no use.
baseten.co
Growth
Hiring Momentum
77/100
Scaling fast
Open positions
100
27 opened / 9 closed in 30 days
Median time-to-fill
67 days
faster than 23% of the market
ATS activity
Every ~1 hours
Today
Hiring Dynamics
+100%
Hiring Focus
The percentage next to each role is its share of the company's job openings over the last 90 days; the arrow shows the shift versus the previous period.
Sales
17% ▲
Marketing
13% ▼
Operations
11% ▲
DevOps
10% ▼
AI/ML
7% ▼
Backend
7% ▼
Actively hiring for Sales (+7 in 30 days)
Activity Timeline
New location: Seattle
Sep 2026
New location: London
Sep 2026
Added Rest API to stack
Sep 2026
Added Fivetran to stack
Sep 2026
Added Linux to stack
Sep 2026
Added TCP/IP to stack
Sep 2026
Added Windows to stack
Sep 2026
Added SIEM to stack
Sep 2026
Added Machine Learning to stack
Sep 2026
Added Tokenomics to stack
Sep 2026
Added XLA to stack
Sep 2026
Added TPU to stack
Sep 2026
Offices
Where the company hires and what each office is for
Cities
6
Hiring now
6
Countries
3
Busiest office
San Francisco
Loading the map
Headquarters
Builds here
Office with open vacancies
Tap a pin for the office
San Francisco
United States
HQ
92 open roles 65 posted in 90 days 43% engineering 1000-5000 est. staff $261k Avg Salary
hires for Backend, Sales, Marketing
560 Davis St
New York
United States
Plant / site
58 open roles 33 posted in 90 days 45% engineering 500-1000 est. staff $277k Avg Salary
hires for Backend, Sales, Marketing
560 Davis St
Toronto
Canada
Engineering / R&D
24 open roles 8 posted in 90 days 92% engineering 200-500 est. staff $331k Avg Salary
hires for Backend, AI/ML, DevOps
Montreal
Canada
Engineering / R&D
24 open roles 8 posted in 90 days 92% engineering 200-500 est. staff $331k Avg Salary
hires for Backend, AI/ML, DevOps
Seattle
United States
Engineering / R&D
22 open roles 6 posted in 90 days 91% engineering 50-200 est. staff $331k Avg Salary
hires for Backend, AI/ML, Frontend
London
United Kingdom
Sales office
5 open roles 5 posted in 90 days 20% engineering 50-200 est. staff
hires for Sales, DevOps
Jobs
$150k – $180k per year • Hybrid • Full-Time • San Francisco • New York
Cursor
Baseten
Machine Learning
Management
Notion
Apply
≈ $128k – $238k per year (Estimated) • In office • Full-Time • 6+ years exp • London
Cursor
Baseten
Machine Learning
Apply
$150k – $235k per year • Hybrid • Full-Time • San Francisco
Cursor
Baseten
Machine Learning
Management
Notion
Apply
$185k – $260k per year • Hybrid • Full-Time • San Francisco
Cursor
Baseten
Machine Learning
Apply
$175k – $215k per year • Hybrid • Full-Time • San Francisco
Cursor
Baseten
Machine Learning
DevOps
Platform Engineering
HPC
Management
Notion
Apply
$175k – $190k per year • Hybrid • Full-Time • 3+ years exp • San Francisco
Databricks
Google BigQuery
BigQuery
AI/ML
Cursor
Claude
Claude Code
OpenAI Codex
OpenAI Agents SDK
Baseten
Mastra
Vercel AI SDK
Machine Learning
DevOps
Vercel
Management
Notion
n8n
Zapier
Apply
$185k – $260k per year • Hybrid • Full-Time • 5+ years exp • San Francisco
SQL
Databricks
AI/ML
Cursor
dbt
Time Series Forecasting
Baseten
Machine Learning
Apply
$200k – $240k per year • Hybrid • Full-Time • 5+ years exp • San Francisco
SQL
Databricks
Google BigQuery
BigQuery
AI/ML
Cursor
dbt
Baseten
Machine Learning
Analytics
Fivetran
Management
n8n
Zapier
Apply
$240k – $285k per year • Hybrid • Full-Time • 4+ years exp • San Francisco • Toronto • New York • Montreal • Seattle
Python
Go
Cursor
Spark
AI Agents
Baseten
Machine Learning
DevOps
Kubernetes
Management
Notion
Apply
$150k – $190k per year • Hybrid • Full-Time • 3+ years exp • San Francisco
SQL
Cursor
Baseten
Machine Learning
Management
Notion
Apply
$165k – $330k per year • Hybrid • Full-Time • San Francisco • Toronto • New York • Montreal • Seattle
Cursor
DeepSeek
vLLM
SGLang
TensorRT
TensorRT-LLM
Baseten
Machine Learning
Management
Notion
Apply
$225k – $290k per year • Hybrid • Full-Time • San Francisco • New York
JavaScript
TypeScript
Cursor
Baseten
Machine Learning
Frontend
React.js
DevOps
Git
Apply

AI Fund
South Park Commons
01 Advisors
BoxGroup
Conviction
Greylock
IVP
K5 Global
Sapphire Ventures
Scribble Ventures
Silicon Valley Investclub
Spark Capital