368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$170k – $300k per year
Location
In office (San Mateo)
Seniority
Staff · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Fireworks AI is an artificial intelligence infrastructure company headquartered in Redwood City, California, and founded in 2022. The company provides a high-performance inference and training platform that enables developers to deploy, fine-tune, and scale open-source generative models with optimized speed and cost. It operates globally as a cloud-based service provider, catering to technology firms and enterprises seeking to integrate specialized intelligence into their applications through a serverless or dedicated API.

About Us:

Fireworks is the platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and products. Founded by the team behind PyTorch and backed by AMD, Atreides, Benchmark Capital, Index Ventures, Lightspeed, NVIDIA, Sequoia Capital, and TCV, Fireworks powers production AI with hundreds of state-of-the-art open models across text, image, embedding, audio, and multimodal workloads. Today, Fireworks is a Series D company valued at $17.5 billion, bringing together an ambitious, collaborative team that's building the future of enterprise AI.

THE ROLE:

Inference is at the core of what Fireworks does. It’s how specialized intelligence reaches production, and the platform around it is what makes it fast to adopt, easy to trust, and economical at scale. We’re hiring a PM to work on our core inference product and general platform, spanning inference performance and control, accounts and fraud reduction.

As a PM working on our platform, you'll help set strategy, write specs, sit with customers running production traffic, and work across our inference, infrastructure, and go-to-market teams to deliver real customer impact. Example problems you may work on include figuring out whether we should let customers 5x their current rate limits, or cutting fraudulent free-tier usage without adding friction for legitimate users.

KEY RESPONSIBILITIES:

  • Own the roadmap, strategy, and success metrics for parts of the Fireworks inference and platform product, across API, UI, and CLI.

  • Work directly with AI-native startups and enterprises running production inference. Watch them work, unblock them when they stall, and convert repeated pain into productized capability.

  • Own the commercial and trust surfaces of the platform, including accounts, onboarding, quotas and rate limits, billing, and fraud and abuse prevention

  • Partner with product marketing, sales, and the field to launch platform capabilities that land, like pricing and packaging, docs, cookbooks, and enablement.

MINIMUM REQUIREMENTS:

  • 2 - 8+ years of product management experience building technical or developer-facing products (we are hiring at multiple levels for this role).

  • Strong technical background - CS/EE degree, production engineering experience, or equivalent depth earned on the job.

  • Familiarity with the inference lifecycle: model serving, latency and throughput tradeoffs, and how these connect to cost in production.

  • Demonstrated ownership of a product area end to end from strategy, spec, launch, to metrics.

  • Excellent written communication. You can write a spec, a launch post, and a customer-facing explanation of a tradeoff, and all three will be clear.

  • Comfort with ambiguity, and a bias toward shipping and learning over waiting for certainty.

  • Deep hunger and motivation. This isn't a 9-5 job and you'll be expected to step up, especially during periods of "wartime."

PREFERRED QUALIFICATIONS:

  • You've personally built on top of an LLM API and felt the latency, cost, and reliability tradeoffs firsthand.

  • Experience with self-serve or PLG products - signup and onboarding funnels, usage-based billing, quota and rate-limit design.

  • Experience with trust and safety, fraud, or abuse prevention at scale - payment fraud, free-tier abuse, or account takeover.

  • Understanding of GPU economics and how utilization, batching, and latency SLAs trade off against margin.

  • Experience with cloud or developer platforms with metered pricing, or with the account, identity, and org-management surfaces enterprises expect.

  • Early startup or founding experience.

Why Fireworks?

  • Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving.

  • Build What’s Next: Work with bleeding-edge technology that impacts how businesses and developers harness AI globally.

  • Ownership & Impact: Join a fast-growing, passionate team where your work directly shapes the future of AI-no bureaucracy, just results.

  • Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation.

Fireworks AI is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Mateo
$98k – $164k per year • In office • Full-Time • 5+ years exp • Master's Degree • United States
Python
AI/ML
LLM
Multimodal AI
PyTorch
TensorFlow
Knowledge Graph
OpenAI
Time Series Forecasting
DevOps
GitHub
Apply
$54k – $138k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Munich
Python
AI/ML
Computer Vision
Multimodal AI
PyTorch
Reinforcement Learning
TensorFlow
Robotics
Imitation Learning
Isaac Sim
MuJoCo
Perception
PyBullet
Reinforcement Learning
ROS
Sim-to-Real
Apply
$145k – $193k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Chicago • Washington • Denver • Jersey City • Boston
Python
AI/ML
AI Agents
Anomaly Detection
Embeddings
Fine-tuning
Function Calling
Hugging Face
LangChain
LLM
LLM Evaluation
LLM Guardrails
PyTorch
RAG
Red Teaming
Scikit-learn
Vertex AI
DevOps
Azure
GCP
Cybersecurity
Threat Modeling
Apply
$320k per year • In office • 10+ years exp • Bachelor's Degree • New York
AI/ML
Anthropic
Claude
Multimodal AI
Apply
Security GRC Analyst 4 hours ago
$119k – $268k per year (Estimated) • Remote/Hybrid • 4+ years exp • Bachelor's Degree • San Francisco
AI/ML
Ignite
PyTorch
Cybersecurity
ISO 27001
NIST CSF
SOC 2
Apply
$132k – $290k per year (Estimated) • In office • Full-Time • 3+ years exp • London
Python
AI/ML
Fine-tuning
Fireworks AI
Multimodal AI
PyTorch
Post-training
AI Agents
Apply
$77k – $214k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • London
Python
AI/ML
Fine-tuning
Fireworks AI
Multimodal AI
PyTorch
Quantization
AI Agents
Apply
GRC Analyst 10 days ago
$160k – $170k per year • In office • Full-Time • 3+ years exp • San Mateo
AI/ML
Fireworks AI
Multimodal AI
PyTorch
ISO 42001
DevOps
AWS
Azure
GCP
Cybersecurity
GDPR
HIPAA
ISO 27001
Least Privilege
NIST CSF
SOC 2
Management
ServiceNow
Apply
$180k – $210k per year • In office • Full-Time • 7+ years exp • San Mateo • New York
Python
AI/ML
Fireworks AI
Multimodal AI
PyTorch
Red Teaming
DevOps
AWS
Azure
GCP
Incident Management
PagerDuty
Splunk
Cybersecurity
Crowdstrike
MITRE ATT&CK
SentinelOne
Sumo Logic
Apply
$170k – $300k per year • In office • Full-Time • 8+ years exp • San Mateo
AI/ML
Fine-tuning
Fireworks AI
LoRA
Multimodal AI
PEFT
PyTorch
Transformers
Post-training
SFT
Apply
$235k per year • Remote/Hybrid • 2+ years exp • Master's Degree • San Mateo
Python
SQL
DevOps
Git
Analytics
A/B Testing
Apply
$216k – $388k per year (Estimated) • Equity • In office • Full-Time • San Mateo
AI/ML
AI Agents
Semantic Search
Apply
$192k – $346k per year (Estimated) • Equity • In office • Full-Time • 6+ years exp • San Mateo
Go
JavaScript
Node JS
Python
Ruby
AI/ML
AI Agents
Cybersecurity
FedRAMP
Apply
$150k – $320k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 8+ years exp • San Mateo
Databases
Apache Kafka
AI/ML
Flink
Feature Store
Apply
$222k – $416k per year (Estimated) • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • San Mateo
AI/ML
AI Agents
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.