707,525open jobs
41,962companies
101,146added this week
Browse all
Salary
$170k – $220k per year
Location
Remote/Hybrid (San Francisco, New York, United States)
Seniority
Senior · 6+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Baseten is an American company founded in 2019 that runs machine learning models in production for companies that would rather not operate GPU infrastructure themselves. Its position is inference rather than training: it handles model packaging, autoscaling, cold start latency and multi-cloud capacity, which are the unglamorous problems that determine whether a model-powered product is fast and affordable enough to ship. Headquartered in San Francisco and backed at a multi-billion dollar valuation, it serves companies deploying open and custom models, and it competes with both the hyperscalers and the model providers' own hosted endpoints.

ABOUT BASETEN

Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products.

REQUIREMENTS:

Product Positioning & Technical Narrative

  • Own the positioning and messaging for Baseten’s dedicated inference and platform capabilities, including autoscaling, routing, failover, release safety, observability, and cost/performance capabilities.

  • Translate infrastructure-heavy product work into buyer narratives for ML engineering, platform engineering, security, compliance, procurement, and executive audiences.

  • Partner with Product and Engineering to understand the technical architecture, customer value, roadmap tradeoffs, and proof points behind each capability.

  • Define when a capability should be positioned as a platform differentiator, a dedicated inference requirement, a reliability story, a compliance story, or sales enablement.

  • Build messaging that is technically credible without being overly implementation-focused or generic.

Launch Strategy & GTM Execution

  • Build and execute launch plans for major dedicated inference and serving platform capabilities, from early internal enablement through external announcement.

  • Decide what deserves a full launch versus what should ship through docs, sales enablement, customer-specific materials, or targeted enterprise outreach.

  • Create launch assets including messaging briefs, landing pages, blog posts, sales decks, one-pagers, FAQs, demo storylines, competitive talk tracks, and customer-facing proof points.

  • Sequence launches and supporting assets based on customer demand, product readiness, sales need, and market opportunity.

  • Partner with Demand Gen, Sales, Solutions, and Customer Marketing to turn product capabilities into pipeline-generating campaigns and enterprise deal support.

Market, Customer & Competitive Intelligence

  • Build a deep understanding of how AI-native companies and enterprises run production inference today, where their current infrastructure breaks, and what they need as workloads scale.

  • Turn ambiguous customer and market signals into clear recommendations for Product, Sales, and Marketing.

  • Develop competitive messaging against inference providers, cloud platforms, and internal DIY alternatives

Cross-Functional Partnership

  • Partner with Engineering to ensure technical claims are accurate, defensible, and supported by real product behavior.

  • Work with Sales and Solutions teams to understand deal blockers, buyer objections, and enablement gaps.

  • Collaborate with Demand Gen and Customer Marketing to build campaigns, customer stories, and field programs

REQUIREMENTS:

  • 6+ years in product marketing, technical marketing, developer marketing, or a related GTM role, ideally at an infrastructure, developer tools, cloud, data, or AI company.

  • Strong technical fluency and the ability to quickly ramp on infrastructure concepts like autoscaling, routing, failover, capacity, observability, deployment models, and performance optimization.

  • Experience translating complex technical products into clear, differentiated positioning for both technical practitioners and enterprise decision-makers.

  • Track record building launch strategies and GTM plans for platform products, infrastructure products, or technical capabilities with long enterprise sales cycles.

  • Excellent storytelling skills, with the ability to make technical infrastructure feel concrete, credible, and commercially urgent.

NICE TO HAVE:

  • Experience with AI/ML infrastructure, inference, model serving, cloud infrastructure, Kubernetes, observability, DevOps, or distributed systems.

  • Experience marketing products to ML engineers, platform engineers, infrastructure leaders, security/compliance teams, or technical executives.

  • Experience building competitive positioning against cloud providers, inference providers, model API companies, or internal platforms.

BENEFITS

  • Competitive compensation, including meaningful equity

  • (U.S. only) 100% coverage of medical, dental, and vision insurance for employee and dependents

  • Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)

  • Paid parental leave

  • Fertility and family-building stipend through Carrot

  • (U.S. only) Company-facilitated 401(k)

  • Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.

Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.

At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.

We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
707,525 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
Junior QA engineer 12 hours ago
Remote • 1+ year exp • Bachelor's Degree • Tbilisi
PHP
SQL
PHP
Laravel
AI/ML
Cursor
ChatGPT
Claude Code
Mobile
TestFlight
Firebase App Distribution
DevOps
Rest API
Git
Docker
QA
Playwright
Postman
Apply
$15k – $42k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Makati
Python
SQL
Python
Flask
FastAPI
Databases
Snowflake
AI/ML
Copilot
Cursor
Claude
Claude Code
Prefect
AI Agents
Anthropic
Devin
GPT-4
DevOps
Rest API
Git
AWS
AWS Fargate
AWS Lambda
Amazon EC2
SLI/SLO/SLA
Amazon S3
Amazon ECS
Amazon EventBridge
AWS Step Functions
Analytics
Alteryx
Management
SharePoint
Apply
$33k – $90k per year (Estimated) • In office • Full-Time • Ho Chi Minh City
Python
Go
JavaScript
Java
Rust
Kotlin
TypeScript
SQL
Node JS
Scala
Python
FastAPI
Node JS
Nest.JS
Fastify
Databases
PostgreSQL
Snowflake
Databricks
ElasticSearch
Apache Kafka
Google BigQuery
Amazon Redshift
BigQuery
AI/ML
Copilot
Hadoop
Spark
Airflow
OpenCV
Dagster
Vertex AI
dbt
Embeddings
Scikit-learn
Kubeflow
TensorFlow
PyTorch
Anomaly Detection
TorchServe
Feature Store
Machine Learning
Frontend
GraphQL
Next.js
React.js
Storybook
DevOps
Rest API
gRPC
Terraform
GCP
Helm
GitHub Actions
Istio
Datadog
Kustomize
CI/CD
ArgoCD
AWS
Kubernetes
Cloudflare
Platform Engineering
Service Mesh
Argo Workflows
Google GKE
GitHub
Cybersecurity
Auth0
Analytics
ETL/ELT
Metabase
Looker
Design
Figma
Management
Slack
Miro
Confluence
Discord
Agile
Scrum
QA
Sentry
Apply
$34k – $73k per year (Estimated) • In office • 7+ years exp • Bachelor's Degree • Bengaluru
Python
JavaScript
Java
Node JS
AI/ML
Copilot
Claude
Model Context Protocol
Prompt Engineering
AI Agents
NLP
AWS Bedrock
LLM
Amazon SageMaker
AWS Bedrock AgentCore
DevOps
OpenShift
CloudFormation
CI/CD
Jenkins
AWS
Docker
Kubernetes
Platform Engineering
AWS Lambda
Amazon EC2
GitLab
Amazon S3
Apply
$39k – $103k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Karachi
Python
Go
PHP
Databases
MySQL
AI/ML
Cursor
Claude Code
DevOps
CI/CD
AWS
Management
Agile
Scrum
Apply
$170k – $210k per year • Remote/Hybrid • Full-Time • San Francisco • New York
AI/ML
Cursor
Baseten
Machine Learning
Management
Notion
Marketing
YouTube
LinkedIn
Apply
$122k – $291k per year (Estimated) • In office • Full-Time • 5+ years exp • London
AI/ML
Cursor
Baseten
Machine Learning
Management
Notion
Apply
Head of EMEA Sales 8 days ago
$136k – $308k per year (Estimated) • In office • Full-Time • London
AI/ML
Cursor
Baseten
Machine Learning
Management
Slack
Apply
$119k – $284k per year (Estimated) • In office • Full-Time • 8+ years exp • London
AI/ML
Cursor
Baseten
Machine Learning
Apply
$128k – $238k per year (Estimated) • In office • Full-Time • 6+ years exp • London
AI/ML
Cursor
Baseten
Machine Learning
Apply
$245k – $279k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco • McLean • Cambridge • San Jose • New York
Python
Go
Java
C#
C++
Scala
C++
PyTorch C++
AI/ML
CUDA Toolkit
AI Agents
PyTorch
LLM
CUDA
Hugging Face
LLM Guardrails
Agentic Workflows
Multi-Agent Systems
Machine Learning
DevOps
GCP
Azure
AWS
Apply
$245k – $279k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • McLean • San Francisco • Richmond • Chicago • New York
Python
Go
JavaScript
Rust
TypeScript
C#
Scala
AI/ML
AI Agents
Machine Learning
DevOps
GCP
Azure
AWS
HPC
Apply
Cloud Engineer 5 2 hours ago
$230k – $262k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • McLean • San Francisco • San Jose • Richmond • New York
Python
JavaScript
Java
Node JS
Scala
DevOps
Terraform
GCP
CloudFormation
Crossplane
Pulumi
Azure
CI/CD
AWS
Docker
Kubernetes
Apply
$197k – $225k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • San Jose • San Francisco • McLean • Cambridge • New York
Python
Go
Java
C#
C++
Scala
C++
PyTorch C++
AI/ML
CUDA Toolkit
Fine-tuning
AI Agents
PyTorch
LLM
CUDA
Hugging Face
TPU
LLM Guardrails
Agentic Workflows
Multi-Agent Systems
Machine Learning
DevOps
GCP
Azure
AWS
Apply
$219k – $250k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Jose • San Francisco • McLean • Cambridge • New York
AI/ML
Fine-tuning
RLHF
Quantization
NLP
Transfer Learning
VLM
PyTorch
LLM
Tokenization
Self-Supervised Learning
Hugging Face
SFT
Pre-training
Machine Learning
DevOps
AWS
Apply
See all jobs
This is one of many
707,525 more open roles from verified company boards, updated every day.