368,910open jobs
9,449companies
47,822added this week
Browse all
Salary
$250k – $400k per year
Location
In office (San Mateo, New York)
Employment
Full-Time
Overview
Company
Impact
Profile match
Fireworks AI is an artificial intelligence infrastructure company headquartered in Redwood City, California, and founded in 2022. The company provides a high-performance inference and training platform that enables developers to deploy, fine-tune, and scale open-source generative models with optimized speed and cost. It operates globally as a cloud-based service provider, catering to technology firms and enterprises seeking to integrate specialized intelligence into their applications through a serverless or dedicated API.

About Us:

Fireworks is the platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and products. Founded by the team behind PyTorch and backed by AMD, Atreides, Benchmark Capital, Index Ventures, Lightspeed, NVIDIA, Sequoia Capital, and TCV, Fireworks powers production AI with hundreds of state-of-the-art open models across text, image, embedding, audio, and multimodal workloads. Today, Fireworks is a Series D company valued at $17.5 billion, bringing together an ambitious, collaborative team that's building the future of enterprise AI.

About the Role

We are looking for a Research Engineer to join our team, operating at the critical intersection of model research and training infrastructure.

In this role, your time will be split between tackling open-ended research problems-such as designing novel architectures and improving algorithmic efficiency - and building the distributed training systems required to make those research breakthroughs a reality. You won't just be handed a paper to implement; you will be expected to reproduce state-of-the-art results from the literature, identify their limitations, and build the infrastructure needed to push beyond them.

The most significant advances in deep learning require massive scale. We need engineers who are as comfortable reasoning about gradient descent and loss landscapes as they are about distributed systems, GPU cluster utilization, and data pipelines.

What You'll Do

  • Conduct Open-Ended Research: Explore new model architectures, training objectives, and optimization techniques. Formulate hypotheses, design experiments, and iterate quickly based on empirical results.

  • Reproduce and Extend State-of-the-Art: Implement and reproduce results from recent machine learning papers. Identify bottlenecks, propose improvements, and scale these methods to larger datasets and models.

  • Build and Scale Training Infrastructure: Design, implement, and maintain high-performance, distributed machine learning systems. Optimize training loops, data loaders, and communication overhead across large GPU clusters.

  • Bridge Science and Engineering: Translate abstract mathematical concepts and research ideas into robust, bug-free, and efficient code.

  • Collaborate Cross-Functionally: Work closely with Research Scientists to unblock their experiments by providing tooling, optimizing code, and co-designing experiments that are hardware-aware.

We Expect You To Have:

  • Strong programming skills (Python, C++, or Rust) and a commitment to writing clean, maintainable code.

  • Deep practical knowledge of machine learning frameworks (PyTorch, JAX, or TensorFlow).

  • Experience working with large distributed systems and parallel computing (e.g., CUDA, NCCL, MPI).

  • A strong foundation in linear algebra, calculus, probability, and statistics.

  • A proven track record of implementing complex deep learning algorithms from scratch.

Nice to Have:

  • A Master’s or PhD in Computer Science, Machine Learning, Physics, Mathematics, or a related field (or equivalent industry experience).

  • Experience with low-level GPU programming (CUDA/Triton) or hardware co-design.

  • Familiarity with the challenges of training Large Language Models (LLMs)

  • Familiarity with the challenges of inference, and OSS inference engines such as SGLang and vLLM

Why Fireworks?

  • Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving.

  • Build What’s Next: Work with bleeding-edge technology that impacts how businesses and developers harness AI globally.

  • Ownership & Impact: Join a fast-growing, passionate team where your work directly shapes the future of AI-no bureaucracy, just results.

  • Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation.

Fireworks AI is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,910 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Mateo
$117k – $258k per year (Estimated) • Equity • Remote • Full-Time • United States
C++
Java
Python
Cybersecurity
Crowdstrike
Apply
$98k – $164k per year • In office • Full-Time • 5+ years exp • Master's Degree • United States
Python
AI/ML
LLM
Multimodal AI
PyTorch
TensorFlow
Knowledge Graph
OpenAI
Time Series Forecasting
DevOps
GitHub
Apply
$48k – $112k per year (Estimated) • In office • Full-Time • Master's Degree • Toronto
C++
MATLAB
Python
Apply
$115k – $204k per year (Estimated) • In office • Full-Time • United States
Bash
C++
Python
C++
CMake
DevOps
GitHub
GitHub Actions
Apply
$164k – $307k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp
C++
Assembly
Assembly
WinDbg
DevOps
VMWare
Cybersecurity
Zero Trust
Zscaler
Apply
$132k – $290k per year (Estimated) • In office • Full-Time • 3+ years exp • London
Python
AI/ML
Fine-tuning
Fireworks AI
Multimodal AI
PyTorch
Post-training
AI Agents
Apply
$77k – $214k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • London
Python
AI/ML
Fine-tuning
Fireworks AI
Multimodal AI
PyTorch
Quantization
AI Agents
Apply
GRC Analyst 10 days ago
$160k – $170k per year • In office • Full-Time • 3+ years exp • San Mateo
AI/ML
Fireworks AI
Multimodal AI
PyTorch
ISO 42001
DevOps
AWS
Azure
GCP
Cybersecurity
GDPR
HIPAA
ISO 27001
Least Privilege
NIST CSF
SOC 2
Management
ServiceNow
Apply
$180k – $210k per year • In office • Full-Time • 7+ years exp • San Mateo • New York
Python
AI/ML
Fireworks AI
Multimodal AI
PyTorch
Red Teaming
DevOps
AWS
Azure
GCP
Incident Management
PagerDuty
Splunk
Cybersecurity
Crowdstrike
MITRE ATT&CK
SentinelOne
Sumo Logic
Apply
$170k – $300k per year • In office • Full-Time • 8+ years exp • San Mateo
AI/ML
Fireworks AI
LLM
Multimodal AI
PyTorch
Apply
$151k – $272k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • San Mateo
AI/ML
Time Series Forecasting
Apply
$235k per year • Remote/Hybrid • 2+ years exp • Master's Degree • San Mateo
Python
SQL
DevOps
Git
Analytics
A/B Testing
Apply
$216k – $388k per year (Estimated) • Equity • In office • Full-Time • San Mateo
AI/ML
AI Agents
Semantic Search
Apply
$192k – $346k per year (Estimated) • Equity • In office • Full-Time • 6+ years exp • San Mateo
Go
JavaScript
Node JS
Python
Ruby
AI/ML
AI Agents
Cybersecurity
FedRAMP
Apply
$280k – $420k per year • In office • Full-Time • San Mateo
Go
Python
AI/ML
Hugging Face
PyTorch
Ray
Reinforcement Learning
Transformers
DevOps
Grafana
Helm
Kubernetes
OpenTelemetry
Platform Engineering
Prometheus
SLURM
Terraform
Robotics
Reinforcement Learning
Apply
See all jobs
This is one of many
368,910 more open roles from verified company boards, updated every day.