1,433,153open jobs
84,640companies
216,921added this week
Browse all
Salary
$170k – $230k per year
Location
Remote (United States)
Seniority
Senior · 5+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 10, 2026. First seen by Alion on Oct 9, 2026. Fal scores A on the Alion truth index.

Overview
Company
Impact
Profile match

Fal

Fal is a generative media AI infrastructure platform headquartered in San Francisco, California. Founded in 2021 by former Coinbase and Amazon engineers Burkay Gur and Gorkem Yurtseven, the enterprise is backed by prominent venture investors including Andreessen Horowitz (a16z), Bessemer Venture Partners, and Salesforce Ventures.

fal is the generative media ecosystem powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from idea to production, and do it at scale without compromise. For developers and enterprises, fal is the foundation that makes generative media not just possible, but practical: a unified platform where high-performance inference, orchestration, and observability come together to unlock new categories of AI-native products.

As generative media reshapes industries across a market projected to grow by hundreds of billions over the next decade, fal is becoming the ecosystem that ambitious teams build on.

About this role:

Help fal's ML team move faster by building the automation, infrastructure, and developer tooling that makes developing, testing, and deploying generative AI models seamless.

You'll own and improve the CI/CD systems supporting our rapidly growing collection of ML models and inference pipelines. Your focus will be on eliminating manual work, accelerating development cycles, and building reliable systems that allow ML engineers to ship new models and optimizations with confidence.

This is a high-impact engineering role where you'll work closely with our Applied ML and ML Performance teams. You'll build everything from automated model validation and performance benchmarking to AI-powered development workflows that help engineers iterate faster.

The ideal candidate thinks beyond traditional CI/CD and sees automation as a force multiplier for the entire engineering organization.

What you’ll do:

  • Own ML CI/CD infrastructure Design, build, and maintain automated testing, validation, and deployment pipelines for our ML models and inference services.

  • Accelerate development cycles.Dramatically reduce CI execution times through intelligent parallelization, caching, test selection, and efficient use of compute resources.

  • Build automated model validation.Develop systems that test model outputs, detect quality regressions, and validate changes across different models, GPU architectures, and configurations.

  • Automate performance benchmarking. Build continuous performance testing that detects regressions in inference latency, throughput, GPU utilization, and cost.

  • Build automated pricing and deployment checks. Ensure model pricing, billing configurations, API schemas, and deployments are validated automatically before reaching production.

  • Extend our agentic engineering workflows. Develop AI-powered automation and agentic coding systems that automatically diagnose CI failures, identify regressions, propose fixes, and streamline engineering workflows.

  • Improve deployment reliability. Build automated safeguards, deployment verification, rollback mechanisms, and monitoring to ensure new model releases are reliable.

  • Eliminate engineering toil. Identify repetitive tasks across the ML team and build tools and systems that automate them, allowing engineers to focus on developing new models and improving performance.

Qualifications/Nice-to-haves:

  • 5+ years of software engineering background with proficiency in Python and experience building production infrastructure and developer tooling.

  • Experience designing and operating CI/CD systems using GitHub Actions or comparable technologies.

  • Deep understanding of automated testing, build systems, dependency management, caching, and parallel execution.

  • Experience working with containerized workloads, Docker, and cloud infrastructure.

  • Ability to design reliable distributed systems and debug complex infrastructure failures.

  • Strong understanding of observability, including logs, metrics, tracing, and automated alerting.

  • A passion for developer productivity and a demonstrated ability to eliminate manual processes through automation.

  • Comfortable working independently, identifying high-impact problems, and building end-to-end solutions.

  • Experience with ML infrastructure, PyTorch, GPU workloads, or model-serving systems.

  • Familiarity with NVIDIA GPU architectures and multi-GPU environments.

  • Experience building automated inference benchmarks or ML quality evaluation frameworks.

  • Experience with agentic coding tools such as Codex or Claude Code, or building custom AI engineering agents.

  • Experience optimizing CI/CD pipelines at scale, including distributed test execution and ephemeral compute environments.

  • Experience developing internal developer platforms or infrastructure-as-code tooling.

Tech Stack:

  • Python, PyTorch, Docker, GitHub Actions

  • NVIDIA GPUs and distributed GPU infrastructure

  • Model serving and inference pipelines

  • Observability and performance monitoring tools

  • AI coding agents and automated engineering workflows

  • Access to fal's massive GPU cluster for testing and development

What we offer at fal:

  • Interesting and challenging work

  • A lot of learning and growth opportunities

  • Health, dental, and vision insurance (US)

  • Regular team events and offsites

U.S. EQUAL EMPLOYMENT OPPORTUNITY INFORMATION:

fal provides equal employment opportunities to applicants and employees without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability, or any other classification protected by applicable law.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,433,153 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
In your city
$370k – $574k per year • Equity • Hybrid • Full-Time • 1+ year exp • Master's Degree • San Jose
Python
SQL
AI/ML
Computer Vision
TensorFlow
PyTorch
LLM
Machine Learning
Apply
$200k – $275k per year • Remote (United States) • Full-Time • United States
AI/ML
LangGraph
LangChain
vLLM
Fine-tuning
Prompt Engineering
AI Agents
TensorRT
LLM
RAG
Triton
Knowledge Graph
LLM Guardrails
Machine Learning
Apply
$170k – $258k per year • Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Sunnyvale • Washington • Austin • San Francisco
C++
AI/ML
CUDA Toolkit
CUDA
CUTLASS
DevOps
HPC
Apply
$170k – $241k per year • Remote (United States) • Full-Time • 5+ years exp • Bachelor's Degree • Sunnyvale
Python
AI/ML
TensorFlow
PyTorch
FSDP
Machine Learning
DevOps
GCP
Azure
AWS
Apply
$159k – $231k per year • Remote (United States) • Full-Time • Master's Degree • Sunnyvale • Washington
Python
SQL
AI/ML
Spark
Fine-tuning
Reinforcement Learning
Multimodal AI
Pandas
NumPy
PyTorch
Self-Supervised Learning
Synthetic Data
SFT
Pre-training
Embodied AI
Machine Learning
Mobile
AVFoundation
Robotics
Sim-to-Real
Imitation Learning
Reinforcement Learning
Apply
≈ $114k – $237k per year (Estimated) • Remote (United States) • Bachelor's Degree • Kansas City
JavaScript
TypeScript
SQL
PowerShell
C#
C#
ASP.NET Core
AI/ML
Copilot
Claude
Frontend
Vue.js
Angular
React.js
DevOps
Rest API
Azure
CI/CD
AWS
VPN
Cybersecurity
Microsoft Entra ID
Management
Agile
Apply
≈ $124k – $241k per year (Estimated) • Remote (United States) • Bachelor's Degree • Kansas City
JavaScript
TypeScript
SQL
PowerShell
C#
C#
ASP.NET Core
AI/ML
Copilot
Claude
Frontend
Vue.js
Angular
React.js
DevOps
Rest API
Azure
CI/CD
AWS
VPN
Cybersecurity
Microsoft Entra ID
Management
Agile
Apply
≈ $59k – $144k per year (Estimated) • In office • 20+ years exp • Bengaluru
Python
JavaScript
Java
TypeScript
Frontend
Angular
DevOps
CI/CD
Platform Engineering
Progressive Delivery
Apply
≈ $59k – $96k per year (Estimated) • Hybrid • Full-Time • Warsaw
Java
Java
Maven
Spring Boot
Hibernate
Hazelcast
Databases
Oracle
RabbitMQ
DevOps
Rest API
GitLab CI
CI/CD
Jenkins
Git
GitLab
SOAP
Cybersecurity
Sonatype Nexus IQ
Management
Jira
Apply
≈ $56k – $88k per year (Estimated) • Remote (Poland) • Full-Time • 4+ years exp • Warsaw
JavaScript
TypeScript
SQL
C#
C#
.NET
Entity Framework Core
Dapper
Databases
PostgreSQL
MS SQL
Apache Kafka
Azure Cosmos DB
Frontend
Angular
DevOps
Rest API
Azure DevOps
Azure
CI/CD
Git
Docker
Kubernetes
Azure AKS
Management
Agile
Scrum
Apply
$150k – $200k per year • In office • Full-Time • 3+ years exp • San Francisco
Python
DevOps
CI/CD
Apply
$180k – $230k per year • In office • Full-Time • 3+ years exp • San Francisco
Python
AI/ML
Fine-tuning
Computer Vision
Diffusers
PyTorch
Hugging Face
Post-training
Machine Learning
Apply
$180k – $250k per year • In office • Full-Time • San Francisco
Python
AI/ML
Machine Learning
DevOps
Kubernetes
Apply
$200k – $250k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco
Apply
Hybrid • Full-Time • 5+ years exp
Python
AI/ML
Diffusion Models
DevOps
Kubernetes
Incident Management
Apply
See all jobs
This is one of many
1,433,153 more open roles from verified company boards, updated every day.