368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$77k – $214k per year (Estimated)
Location
Remote/Hybrid (London, United Kingdom)
Seniority
Middle · 3+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Fireworks AI is an artificial intelligence infrastructure company headquartered in Redwood City, California, and founded in 2022. The company provides a high-performance inference and training platform that enables developers to deploy, fine-tune, and scale open-source generative models with optimized speed and cost. It operates globally as a cloud-based service provider, catering to technology firms and enterprises seeking to integrate specialized intelligence into their applications through a serverless or dedicated API.

About Us:

Fireworks is the platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and products. Founded by the team behind PyTorch and backed by AMD, Atreides, Benchmark Capital, Index Ventures, Lightspeed, NVIDIA, Sequoia Capital, and TCV, Fireworks powers production AI with hundreds of state-of-the-art open models across text, image, embedding, audio, and multimodal workloads. Today, Fireworks is a Series D company valued at $17.5 billion, bringing together an ambitious, collaborative team that's building the future of enterprise AI.

The Role:

As an AI Forward Deployed Engineer, you are the technical owner of a customer deployment. You embed with a small number of accounts, write code in their environment, and unblock whatever stands between them and production on Fireworks: capacity, model selection, integration, latency, cost. You scope high-stakes deals before signature and then deliver what you scoped. This role suits engineers with a founder CTO mentality who genuinely enjoy customer work.

What we value most is first principles thinking: reasoning about unfamiliar problems from the ground up rather than from playbooks. You will be evaluated on this directly in the interview process, and strength here can outweigh gaps elsewhere in your background.

Key Responsibilities:

  • Deployment Ownership: Own the technical outcome of customer deployments end to end, from architecture through production.

  • Embedded Pilots: Design, scope, and run proof-of-value pilots inside customer environments, with clear success criteria.

  • Pre-Signature Scoping: Act as the technical owner on complex deals (dedicated deployments, BYOC, compliance-sensitive architectures) and carry them through delivery.

  • Unblocking: Diagnose and resolve capacity, performance, integration, and model selection issues, pulling in specialists where needed.

  • Model and Platform Fit: Guide customers to the right models, deployment tiers, and optimization paths (e.g. quantization, speculative decoding, fine-tuning handoff to AMLEs).

  • Product Signal: Feed structured, prioritized signal from the field back to product and engineering.

Minimum Requirements:

  • First principles and critical thinking: the ability to break down novel problems and reason to a solution without a reference implementation. This is the primary interview signal and can outweigh other qualifications.

  • Bachelor's degree in Computer Science, Engineering, or a related technical field.

  • 3+ years of experience as a software engineer, with depth in backend systems and infrastructure.

  • Strong coding skills in Python and at least one systems language; comfortable shipping in unfamiliar codebases and customer environments.

  • Fluent in AI-assisted and agentic engineering: uses coding agents and AI tooling as a core part of how you build, and knows when to trust the output and when not to.

  • Working knowledge of modern generative AI: inference, serving, and the tradeoffs across models and deployment options.

  • Proven ability to own ambiguous technical problems under time pressure and drive them to resolution.

  • Genuine appetite for customer-facing work and strong communication with senior technical stakeholders.

Preferred Qualifications:

  • Founder or founding engineer experience, especially as a technical co-founder or CTO.

  • Experience with GPU infrastructure, distributed serving, or performance engineering.

  • Prior forward deployed, solutions architecture, or professional services experience at a technical product company.

  • Experience in a startup or fast-paced environment.

Why Fireworks?

  • Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving.

  • Build What’s Next: Work with bleeding-edge technology that impacts how businesses and developers harness AI globally.

  • Ownership & Impact: Join a fast-growing, passionate team where your work directly shapes the future of AI-no bureaucracy, just results.

  • Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation.

Fireworks AI is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
London
$98k – $164k per year • In office • Full-Time • 5+ years exp • Master's Degree • United States
Python
AI/ML
LLM
Multimodal AI
PyTorch
TensorFlow
Knowledge Graph
OpenAI
Time Series Forecasting
DevOps
GitHub
Apply
$78k – $150k per year (Estimated) • Equity • In office • Full-Time • 2+ years exp • Bachelor's Degree • United States
MATLAB
Python
DevOps
Error Budget
Apply
$48k – $112k per year (Estimated) • In office • Full-Time • Master's Degree • Toronto
C++
MATLAB
Python
Apply
RTL Design Engineer 2 hours ago
$87k – $168k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Austin • Phoenix • Hillsboro
Assembly
Perl
Python
SystemVerilog
Verilog
VHDL
Chips/EDA
Synopsys Design Compiler
Apply
In office • Internship • Bachelor's Degree • Boise
MATLAB
Python
AI/ML
Claude
Apply
$132k – $290k per year (Estimated) • In office • Full-Time • 3+ years exp • London
Python
AI/ML
Fine-tuning
Fireworks AI
Multimodal AI
PyTorch
Post-training
AI Agents
Apply
GRC Analyst 10 days ago
$160k – $170k per year • In office • Full-Time • 3+ years exp • San Mateo
AI/ML
Fireworks AI
Multimodal AI
PyTorch
ISO 42001
DevOps
AWS
Azure
GCP
Cybersecurity
GDPR
HIPAA
ISO 27001
Least Privilege
NIST CSF
SOC 2
Management
ServiceNow
Apply
$180k – $210k per year • In office • Full-Time • 7+ years exp • San Mateo • New York
Python
AI/ML
Fireworks AI
Multimodal AI
PyTorch
Red Teaming
DevOps
AWS
Azure
GCP
Incident Management
PagerDuty
Splunk
Cybersecurity
Crowdstrike
MITRE ATT&CK
SentinelOne
Sumo Logic
Apply
$170k – $300k per year • In office • Full-Time • 8+ years exp • San Mateo
AI/ML
Fireworks AI
LLM
Multimodal AI
PyTorch
Apply
$170k – $300k per year • In office • Full-Time • 8+ years exp • San Mateo
AI/ML
Fine-tuning
Fireworks AI
LoRA
Multimodal AI
PEFT
PyTorch
Transformers
Post-training
SFT
Apply
$72k – $136k per year (Estimated) • Equity • Remote/Hybrid • 5+ years exp • London
Python
SQL
Databases
Snowflake
AI/ML
AI Agents
DevOps
Kibana
Marketing
Salesforce
Apply
$77k – $148k per year (Estimated) • Equity • In office • Master's Degree • London
JavaScript
Python
Scala
AI/ML
AI Agents
DevOps
GitHub
Apply
$87k – $159k per year (Estimated) • In office • Contractor • 5+ years exp • London • Stockholm
Design
Figma
Marketing
Zendesk
Apply
$105k – $204k per year (Estimated) • In office • Full-Time • London
Python
SQL
Databases
Snowflake
AI/ML
Dagster
dbt
Analytics
A/B Testing
Apply
$27k – $61k per year (Estimated) • In office • Full-Time • 5+ years exp • Pune • London
Analytics
Power BI
Tableau
Marketing
Salesforce
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.