428,801open jobs
14,629companies
63,028added this week
Browse all
Salary
$160k – $180k per year
Location
In office (San Mateo, New York)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Fireworks AI is an artificial intelligence infrastructure company headquartered in Redwood City, California, and founded in 2022. The company provides a high-performance inference and training platform that enables developers to deploy, fine-tune, and scale open-source generative models with optimized speed and cost. It operates globally as a cloud-based service provider, catering to technology firms and enterprises seeking to integrate specialized intelligence into their applications through a serverless or dedicated API.

About Us:

Fireworks is the platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and products. Founded by the team behind PyTorch and backed by AMD, Atreides, Benchmark Capital, Index Ventures, Lightspeed, NVIDIA, Sequoia Capital, and TCV, Fireworks powers production AI with hundreds of state-of-the-art open models across text, image, embedding, audio, and multimodal workloads. Today, Fireworks is a Series D company valued at $17.5 billion, bringing together an ambitious, collaborative team that's building the future of enterprise AI.

The Role

This role is built for new graduates who want to work on real AI systems, in production, from the start. As a MTS at Fireworks, you'll write and ship code that runs on one of the busiest inference platforms in the world - serving hundreds of models to developers and enterprises at massive scale.

We hire new grads into teams across Engineering, and we match you to a team based on your strengths and interests after we get to know you. Depending on where you land, your work could span inference and performance, model training and fine-tuning, distributed systems and cloud infrastructure, developer platform and APIs, or product and full-stack application engineering. One application puts you in front of all of them.

You'll join with a senior engineer as your mentor and a structured ramp-up: you'll start on well-scoped, real problems alongside your mentor, then take ownership of features and systems of your own. We move fast, our bar is high, and we'll invest in getting you there.

We welcome candidates who have completed a Bachelor's or Master's degree within the last 6 months, or expect to complete one by Summer 2027. Note that candidates graduating in December 2026 are highly preferred and prioritized.

What You'll Do

  • Ship production code: Design, build, and deploy features and services that customers depend on every day

  • Own problems end to end: Take a problem from ambiguous idea through design, implementation, testing, rollout, and operation

  • Make things faster and more reliable: Improve the performance, efficiency, and scalability of the systems you work on

  • Work with modern AI systems: Build with, on top of, or underneath large models - depending on your team, that could mean serving them, training them, evaluating them, or building products powered by them

  • Debug hard things: Dig into unfamiliar code and unfamiliar failures, and get to the root cause

  • Collaborate broadly: Work closely with engineers, researchers, product, and - on some teams - directly with customers

Minimum Qualifications

  • Bachelor's or Master's degree in Computer Science, Engineering, or a related technical field, completed within the last 6 months or before the end of 2026

  • Strong programming fundamentals and solid grounding in data structures, algorithms, and systems

  • Practical software engineering experience through internships, research, open-source contributions, or substantial personal projects

  • Proficiency in at least one of Python, C++, Go, Rust, TypeScript, or a comparable language

  • Fluency with AI coding harnesses and a habit of using them well

  • Clear written and verbal communication, and the judgment to ask for help early

  • Comfort with ambiguity, fast iteration, and high ownership

Preferred Qualifications

We don't expect any single candidate to have all of these.

  • Hands-on experience building with LLMs or other ML models - coursework, projects, and internships all count

  • Experience with GPU programming, CUDA, kernel optimization, or performance profiling

  • Experience with distributed systems, Kubernetes, or cloud infrastructure (AWS/GCP/Azure)

  • Experience with PyTorch, model training, fine-tuning (SFT, RLHF/RFT), or inference serving

  • Experience shipping a full-stack or agentic application end to end, not just a notebook

  • Familiarity with evaluation, RAG systems, or observability for ML systems

  • Open-source contributions, competitive programming, or research publications

Why Fireworks?

  • Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving.

  • Build What’s Next: Work with bleeding-edge technology that impacts how businesses and developers harness AI globally.

  • Ownership & Impact: Join a fast-growing, passionate team where your work directly shapes the future of AI-no bureaucracy, just results.

  • Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation.

Fireworks AI is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
428,801 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Mateo
Senior SQA Engineer 3 hours ago
$90k – $110k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Waltham
Python
JavaScript
C++
Frontend
React.js
DevOps
Rest API
CI/CD
Jenkins
Git
Buildkite
GitHub
Management
Jira
QA
TestRail
Playwright
Pytest
Apply
$124k – $222k per year (Estimated) • Remote • 5+ years exp
Python
PowerShell
C#
C++
AI/ML
Red Teaming
Cybersecurity
Burp Suite
Cobalt Strike
Fortify
Web3
Smart Contracts
Apply
$127k – $167k per year • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Gaithersburg
Python
JavaScript
TypeScript
Node JS
Python
SQLAlchemy
FastAPI
Databases
PostgreSQL
DynamoDB
Apache Kafka
AI/ML
Copilot
Cursor
Claude
Claude Code
Model Context Protocol
Langfuse
LiteLLM
AWS Bedrock
RAG
Frontend
Tailwind CSS
Next.js
React.js
Vite
Radix UI
shadcn/ui
pnpm
DevOps
GitHub Actions
OpenTelemetry
AWS CDK
AWS
Docker
Kubernetes
Amazon EKS
AWS Fargate
GitHub
Amazon S3
Amazon ECS
Apply
$71k – $163k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Singapore
Python
SPSS
Stata
AI/ML
Time Series Forecasting
Analytics
Tableau
Power BI
Apply
$130k – $170k per year • Equity • Remote/Hybrid • Full-Time • 6+ years exp • San Mateo
Python
Bash
DevOps
Terraform
GitHub Actions
Istio
Datadog
Prometheus
CI/CD
Jenkins
AWS
Docker
Kubernetes
Grafana
Service Mesh
GitHub
Apply
$200k – $230k per year • In office • Full-Time • PhD • San Mateo • New York
Python
Rust
C++
C++
PyTorch C++
AI/ML
vLLM
CUDA Toolkit
Multimodal AI
PyTorch
Ray
Fireworks AI
CUDA
Triton
NCCL
InfiniBand
ROCm
KV Cache
DevOps
SLURM
Kubernetes
HPC
Apply
$200k – $300k per year • In office • Full-Time • PhD • San Mateo • New York
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
JAX
Multimodal AI
TensorFlow
PyTorch
Fireworks AI
Apply
$178k – $355k per year (Estimated) • In office • Full-Time • 3+ years exp • London
Python
AI/ML
Fine-tuning
Multimodal AI
AI Agents
PyTorch
Fireworks AI
Post-training
Apply
$180k – $220k per year • Equity • In office • Full-Time • 6+ years exp • San Mateo
AI/ML
Multimodal AI
PyTorch
Fireworks AI
Apply
$180k – $220k per year • In office • Full-Time • 5+ years exp • San Mateo
AI/ML
Multimodal AI
PyTorch
Fireworks AI
Apply
$345k – $399k per year • Equity • Remote/Hybrid • 8+ years exp • San Mateo
Cybersecurity
GDPR
Apply
$190k – $265k per year • Equity • Remote/Hybrid • 10+ years exp • San Mateo
Python
Databases
Neo4j
AI/ML
Knowledge Graph
Analytics
Tableau
Looker
Apply
$102k – $127k per year • Remote/Hybrid • Full-Time • San Mateo
Management
Xero
Apply
$252k – $298k per year • Equity • Remote/Hybrid • 7+ years exp • San Mateo
Apply
Sales Engineer 9 hours ago
$90k – $110k per year • In office • Full-Time • 6+ years exp • Bachelor's Degree • San Mateo
Python
Management
Confluence
Marketing
HubSpot
Apply
See all jobs
This is one of many
428,801 more open roles from verified company boards, updated every day.