368,746open jobs
9,444companies
47,506added this week
Browse all
Salary
$180k – $300k per year
Location
In office (San Francisco)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match

Bfl

Black Forest Labs is building visual intelligence: models that understand, reason, and act in the world. Use FLUX models via our API.

About Black Forest Labs

We're the team behind Latent Diffusion, Stable Diffusion, and FLUX-foundational technologies that changed how the world creates images and video. We’re creating the generative models that power how people make images and video-tools used by millions of creators, developers, and businesses worldwide. Our FLUX models are among the most advanced in the world, and we're just getting started.

Headquartered in Freiburg, Germany with a growing presence in San Francisco, we're scaling fast while staying true to what makes us different: research excellence, open science, and building technology that expands human creativity.

Why This Role

Our research team moves fast. Models improve weekly. New capabilities emerge constantly.

What slows us down is not model quality-it’s productionization.

Without this role:

  • Research checkpoints sit longer before becoming usable APIs

  • Inference is slower than it needs to be

  • APIs struggle under load

  • Demos don’t reflect the true potential of our models

This role removes the bottleneck between frontier research and production reality. Once hired, researchers ship faster, demos launch faster, and customers experience models at their best.

What You’ll Work On

You will own the bridge between research breakthroughs and production systems.

  • Turn research checkpoints into production-ready inference services

  • Design and maintain high-performance APIs serving millions of requests

  • Optimize inference latency and throughput across GPU infrastructure

  • Build scalable serving architectures that handle unpredictable traffic

  • Improve reliability, monitoring, and observability across model-serving systems

  • Prototype and ship demos that showcase new capabilities in days, not weeks

  • Collaborate closely with researchers to move from idea to live endpoint rapidly

Tools & Context - Model Serving & API Infrastructure

  • Python, FastAPI, async systems

  • GPU infrastructure, CUDA, inference optimization

  • Docker and Kubernetes

  • Redis, Postgres, distributed task queues

  • Cloud platforms (AWS, GCP, or Azure)

  • Observability stacks (metrics, logging, tracing)

This role spans backend systems, GPU performance, and production ML serving.

What We’re Looking For

You’ve built and operated systems at meaningful scale. You understand the difference between a research prototype and a production system. You are comfortable navigating ambiguity, making tradeoffs, and improving systems under real-world constraints.

You demonstrate:

  • Strong judgment around performance, reliability, and cost tradeoffs

  • Experience scaling APIs or ML systems under load

  • Comfort working in fast-moving, research-adjacent environments

  • Ownership from system design through debugging and deployment

Role-specific experience we value:

  • Building and operating ML inference services in production

  • Designing scalable API architectures with async processing

  • Optimizing GPU workloads (batching, quantization, compilation, CUDA)

  • Managing distributed systems and task queues under variable load

  • Implementing monitoring and observability for production ML systems

  • Debugging performance bottlenecks across model, infrastructure, and network layers

Bonus experience includes:

  • Real-time or low-latency inference systems

  • TensorRT, reduced precision, layer fusion, or model compilation techniques

  • Frontend demo tooling (Streamlit, Gradio, React)

  • CI/CD and automated testing for ML systems

  • Security best practices for API and model serving

How We Work Together

We’re a distributed team with real offices that people actually use. Depending on your role, you’ll either join us in Freiburg or SF at least 2 days a week (or one full week every other week), or work remotely with a monthly in-person week to stay connected. We’ll cover reasonable travel costs to make this possible. We think in-person time matters, and we’ve structured things to make it accessible to all. We’ll discuss what this will look like for the role during our interview process.

Everything we do is grounded in four values:

  • Obsessed. We are a frontier research lab. The science has to be right, the understanding deep, the product beautiful.

  • Low Ego. The work speaks. The best idea wins, no matter who said it. Credit is shared. Nobody is above any task.

  • Bold. We take the ambitious bet. We ship, we do not wait for conditions to be perfect.

  • Kind. People over politics. We treat each other with genuine warmth. Agency without empathy creates chaos.

Base Annual Salary: $180,000-$300,000 USD

We're based in Europe and value depth over noise, collaboration over hero culture, and honest technical conversations over hype. Our models have been downloaded hundreds of millions of times, but we're still a ~50-person team learning what's possible at the edge of generative AI.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$23k – $58k per year (Estimated) • In office • Full-Time • 10+ years exp • Pune
PowerShell
Python
SQL
Databases
MS SQL
Oracle
PostgreSQL
DevOps
Ansible
AWS
Azure
Chef
CI/CD
Configuration Management
GCP
Grafana
Helm
Kubernetes
OpenShift
Prometheus
Apply
$17k – $49k per year (Estimated) • Remote/Hybrid • Kaliningrad
Node JS
Python
JavaScript
Node JS
BullMQ
Fastify
Python
FastAPI
Flask
Databases
Apache Kafka
Chroma
Pinecone
PostgreSQL
Qdrant
RabbitMQ
Redis
AI/ML
Claude
Copilot
Cursor
LangChain
LlamaIndex
LLM
Model Context Protocol
Prompt Engineering
RAG
Anthropic
OpenAI Codex
Structured Outputs
Function Calling
Frontend
Next.js
React.js
DevOps
AWS
Azure
CI/CD
Docker
GCP
Git
Gitflow
Rest API
WebSockets
GitHub
Management
Jira
Apply
In office • Full-Time • Pune
Java
DevOps
Azure
Azure DevOps
CI/CD
Git
Jenkins
Marketing
Salesforce
QA
BrowserStack
Playwright
Apply
$12k – $30k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Mumbai
Java
AI/ML
AI Agents
Edge AI
DevOps
AWS
Azure
Apply
$39k – $124k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Israel
C#
Java
DevOps
Azure
Azure DevOps
CI/CD
Git
QA
Appium
TestNG
Apply
Product Engineer 14 days ago
$163k – $221k per year • In office • Full-Time • Freiburg im Breisgau
TypeScript
JavaScript
AI/ML
Stable Diffusion
Frontend
Next.js
React.js
Analytics
A/B Testing
Design
Figma
Apply
$151k – $279k per year • In office • Full-Time • San Francisco • Freiburg im Breisgau
AI/ML
CUDA
CUDA Toolkit
LLM
Multimodal AI
PyTorch
Quantization
Stable Diffusion
Triton
FSDP
NCCL
Apply
$140k – $209k per year • In office • Full-Time • Freiburg im Breisgau
AI/ML
Diffusion Models
Fine-tuning
Knowledge Distillation
Reinforcement Learning
Stable Diffusion
Robotics
LeRobot
Reinforcement Learning
Apply
$151k – $395k per year • In office • Full-Time • Freiburg im Breisgau
AI/ML
Fine-tuning
LLM
LoRA
Multimodal AI
Stable Diffusion
Tokenization
VLM
PEFT
SFT
Apply
$151k – $395k per year • In office • Full-Time • Freiburg im Breisgau
AI/ML
Fine-tuning
Knowledge Distillation
PyTorch
RLHF
Stable Diffusion
DPO
Post-training
SFT
Apply
$170k – $220k per year • Equity 1–2.8% • In office • Full-Time • 3+ years exp • San Francisco
Python
SQL
Python
Django
AI/ML
AI Agents
Context Engineering
LLM
LLM Evaluation
RAG
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Senior ML Engineer 2 hours ago
$149k – $224k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Francisco • Washington • Palo Alto
Python
Python
pySpark
Databases
Apache Kafka
AI/ML
AI Agents
Agentforce
Airflow
Anomaly Detection
Feature Store
Flink
Ray
Red Teaming
Spark
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
MITRE ATT&CK
Marketing
Salesforce
Apply
In office • Internship • 1+ year exp • Bachelor's Degree • San Francisco
Go
JavaScript
Ruby
Scala
Apply
$360k – $530k per year • In office • Full-Time • Bachelor's Degree • San Francisco
MATLAB
Python
MATLAB
Simulink
AI/ML
OpenAI
Robotics
Digital Twin
Apply
See all jobs
This is one of many
368,746 more open roles from verified company boards, updated every day.