368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$200k – $260k per year
Location
In office (San Mateo, New York)
Seniority
Staff · 10+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Fireworks AI is an artificial intelligence infrastructure company headquartered in Redwood City, California, and founded in 2022. The company provides a high-performance inference and training platform that enables developers to deploy, fine-tune, and scale open-source generative models with optimized speed and cost. It operates globally as a cloud-based service provider, catering to technology firms and enterprises seeking to integrate specialized intelligence into their applications through a serverless or dedicated API.

About Us:

Fireworks is the platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and products. Founded by the team behind PyTorch and backed by AMD, Atreides, Benchmark Capital, Index Ventures, Lightspeed, NVIDIA, Sequoia Capital, and TCV, Fireworks powers production AI with hundreds of state-of-the-art open models across text, image, embedding, audio, and multimodal workloads. Today, Fireworks is a Series D company valued at $17.5 billion, bringing together an ambitious, collaborative team that's building the future of enterprise AI.

In the last few months alone we launched Fireworks Training, partnered with Microsoft Azure Foundry, and published research straight from our production systems. A few examples of what that looks like in practice:

  • Frontier RL is cheaper than the mega-cluster narrative suggests: we ran cross-region rollouts using 98% sparse weight deltas and published what we learned. (blog)

  • Open source agents with frontier advisors: matching frontier performance through training and harness engineering. (blog)

  • The fine-tuning bottleneck is not the algorithm: integration friction and iteration speed are what actually stall teams; we documented the patterns across dozens of customer engagements. (blog)

The Role:

AI Field Engineers at Fireworks are the technical tip of the spear. You embed with our most ambitious customers and technology partners to turn complex AI problems into production systems, fast. The role sits at the intersection of engineering, product, and customer delivery. You are hands-on-keyboard building POCs, MVPs, and production integrations, while also holding your own in executive-level conversations about architecture, strategy, and business outcomes.

You spend most of your time building. You ship code, run benchmarks, debug production issues, and architect deployments. But you also lead discovery conversations, align stakeholders, and translate customer pain points into product improvements that compress the feedback loop from field to roadmap. This is a role for engineers who are comfortable on-site with customers, building the relationships and trust that happen in person, not just over a call.

The Segment

As a Field Engineer in the AI Native segment you will work with the most innovative AI-native companies building at the frontier, where GenAI is the core product, not a feature, and where Fireworks is the platform they depend on to ship and scale it. These engagements move fast with fewer stakeholders, so you will spend more time in the code and iterate alongside their engineering teams, while still holding executive-level conversations on architecture and strategy. You will embed deeply with a small set of high-velocity accounts where the quality of your engineering is the relationship.

What You'll Work On

Technical Delivery and Deployment

  • Build end-to-end POCs and MVPs alongside customer engineering teams, working inside their codebases, infrastructure, and constraints.

  • For customers whose core product is built on GenAI, architect the inference foundations that capability depends on, and size deployments so they can scale in their market without infrastructure becoming the bottleneck.

  • Run load tests and establish latency, throughput, and cost baselines against realistic customer traffic profiles, and tune deployments to hit those targets

  • Deploy and validate new model families on inference frameworks (vLLM, SGLang), determining optimal shapes, quantization configs, and serving patterns across workloads.

Model Strategy and Fine-Tuning

  • Guide customers on model selection, fine-tuning strategy (SFT, DPO, RFT), and evaluation methodology.

  • Build and run fine-tuning pipelines directly with customers, navigating trade-offs between model families, compute cost, and quality targets.

  • Design and implement evaluation frameworks that measure production-quality metrics, not just benchmark scores.

Customer Engagement and Stakeholder Management

  • Many of our customers exist because of GenAI. Help them bake frontier model capabilities into their core offering and turn that into a durable competitive edge.

  • Lead structured discovery conversations to unpack customer pain points, constraints, and success criteria before proposing solutions.

  • Own the technical relationship from first engagement through production deployment. Embed with their engineering team as a peer, your credibility comes from what you build alongside them.

  • Spend time on-site with customers. Build trust and momentum in person, embedding with their teams where the work happens.

Product Feedback and Platform Improvement

  • Identify recurring customer pain points and translate them into concrete product proposals, working directly with engineering and product to ship fixes and features.

  • Codify repeatable deployment patterns and contribute them back to internal tooling, documentation, and the platform itself.

  • Feed customer signals (deployment patterns, failure modes, feature gaps) back into the product roadmap with specificity and urgency.

What We're Looking For:

Minimum Qualifications

  • 5+ years in a hands-on, customer-facing technical role: Forward Deployed Engineer, Applied AI Engineer, Solutions Architect, ML Engineer with field exposure, or technical founder.

  • Demonstrated ability to build production software with customers, not just advise on it. You have shipped code running in someone else's production environment.

  • Strong Python skills. Comfortable reading, writing, and debugging production code. Familiarity with Kubernetes and infrastructure engineering.

  • Working knowledge of the LLM stack: inference trade-offs, model serving, fine-tuning workflows (SFT at minimum; DPO/RFT a strong plus).

  • Experience with cloud infrastructure (AWS, Azure, GCP) and deploying models on GPU infrastructure.

  • Exceptional communication: able to run a sharp discovery call, present to a VP, and debug a latency issue with an ML engineer in the same afternoon.

  • Experience building or integrating agentic systems, tool-use chains, or AI-native developer toolchains.

Preferred Qualifications

  • 10+ years in technical field or engineering roles.

  • Experience with inference serving frameworks (vLLM, SGLang, TensorRT-LLM) and tuning deployments for real workloads.

  • Prior experience at a company with a forward-deployed or embedded engineering model (Palantir, Scale AI, Anthropic, OpenAI, BCG X, McKinsey Quantum Black, AI Native startups with FDE motions).

  • Prior experience as a technical founder or early engineer at an AI-native company is a strong signal.

  • Track record taking GenAI POCs from prototype to production-scale deployments.

  • Experience with hyperscaler AI platforms (Azure AI Foundry, AWS Bedrock/SageMaker, GCP Vertex).

Why Fireworks?

  • Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving.

  • Build What’s Next: Work with bleeding-edge technology that impacts how businesses and developers harness AI globally.

  • Ownership & Impact: Join a fast-growing, passionate team where your work directly shapes the future of AI-no bureaucracy, just results.

  • Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation.

Fireworks AI is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Mateo
$220k – $325k per year • Remote/Hybrid • Full-Time • 15+ years exp • New York
DevOps
AWS
Azure
CI/CD
GCP
Kubernetes
Platform Engineering
Cybersecurity
Threat Modeling
Apply
Remote/Hybrid • 7+ years exp • Bachelor's Degree
C#
JavaScript
Python
SQL
TypeScript
C#
ASP.NET Core
Blazor
Dapper
Databases
Oracle
Frontend
Angular
Bootstrap
JQuery
Vue.js
DevOps
AWS
Azure
Azure DevOps
CI/CD
Git
Apply
$26k – $66k per year (Estimated) • In office • Full-Time • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
In office • Contractor • Dublin
DevOps
Azure
Cybersecurity
Okta
Management
Google Workspace
Jira
Slack
Apply
$21k – $57k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Noida
C#
SQL
TypeScript
JavaScript
C#
.NET
Frontend
Angular
DevOps
Azure
Azure DevOps
CI/CD
Git
TeamCity
Apply
$132k – $290k per year (Estimated) • In office • Full-Time • 3+ years exp • London
Python
AI/ML
Fine-tuning
Fireworks AI
Multimodal AI
PyTorch
Post-training
AI Agents
Apply
$77k – $214k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • London
Python
AI/ML
Fine-tuning
Fireworks AI
Multimodal AI
PyTorch
Quantization
AI Agents
Apply
GRC Analyst 10 days ago
$160k – $170k per year • In office • Full-Time • 3+ years exp • San Mateo
AI/ML
Fireworks AI
Multimodal AI
PyTorch
ISO 42001
DevOps
AWS
Azure
GCP
Cybersecurity
GDPR
HIPAA
ISO 27001
Least Privilege
NIST CSF
SOC 2
Management
ServiceNow
Apply
$180k – $210k per year • In office • Full-Time • 7+ years exp • San Mateo • New York
Python
AI/ML
Fireworks AI
Multimodal AI
PyTorch
Red Teaming
DevOps
AWS
Azure
GCP
Incident Management
PagerDuty
Splunk
Cybersecurity
Crowdstrike
MITRE ATT&CK
SentinelOne
Sumo Logic
Apply
$170k – $300k per year • In office • Full-Time • 8+ years exp • San Mateo
AI/ML
Fireworks AI
LLM
Multimodal AI
PyTorch
Apply
$216k – $388k per year (Estimated) • Equity • In office • Full-Time • San Mateo
AI/ML
AI Agents
Semantic Search
Apply
$192k – $346k per year (Estimated) • Equity • In office • Full-Time • 6+ years exp • San Mateo
Go
JavaScript
Node JS
Python
Ruby
AI/ML
AI Agents
Cybersecurity
FedRAMP
Apply
$150k – $320k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 8+ years exp • San Mateo
Databases
Apache Kafka
AI/ML
Flink
Feature Store
Apply
$222k – $416k per year (Estimated) • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • San Mateo
AI/ML
AI Agents
Apply
$245k per year • Remote/Hybrid • 4+ years exp • Bachelor's Degree • San Mateo
JavaScript
TypeScript
AI/ML
LLM
Recommender Systems
Frontend
Babel
GraphQL
React.js
Webpack
DevOps
CI/CD
Docker
QA
Jest
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.