368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$160k – $200k per year
Location
In office (San Francisco)
Seniority
Staff · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match

Fal

Fal is a generative media AI infrastructure platform headquartered in San Francisco, California. Founded in 2021 by former Coinbase and Amazon engineers Burkay Gur and Gorkem Yurtseven, the enterprise is backed by prominent venture investors including Andreessen Horowitz (a16z), Bessemer Venture Partners, and Salesforce Ventures.

fal is the generative media ecosystem powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from idea to production, and do it at scale without compromise. For developers and enterprises, fal is the foundation that makes generative media not just possible, but practical: a unified platform where high-performance inference, orchestration, and observability come together to unlock new categories of AI-native products.

As generative media reshapes industries across a market projected to grow by hundreds of billions over the next decade, fal is becoming the ecosystem that ambitious teams build on.

You will own the operational and commercial performance of the hyperscalers, neoclouds, and GPU infrastructure providers that power fal’s compute fleet. You will translate contract commitments into measurable operating standards, resolve performance issues, recover credits and remedies, and ensure our capacity is ready when customers need it.

This is a hands-on role for someone who understands both infrastructure operations and commercial vendor management. You will build the systems, reporting, and operating cadence for a vendor portfolio that is growing quickly in scale and complexity.

What you’ll own

  • Own post-signature relationships with GPU infrastructure providers, including hyperscalers, neoclouds, colocation partners, and other strategic compute vendors.

  • Translate contractual commitments into measurable SLAs, scorecards, escalation procedures, and internal operating requirements.

  • Track availability, provisioning timelines, capacity delivery, support responsiveness, and other vendor-performance measures.

  • Lead escalations when vendors miss commitments, coordinating with Infrastructure, Capacity Planning, Finance, Legal, and executive stakeholders.

  • Pursue service credits, remedies, and corrective-action plans, and ensure issues remain open until they are fully resolved.

  • Run vendor governance, including regular operating reviews, executive business reviews, performance reporting, renewal readiness, and contract-compliance tracking.

  • Partner with Capacity Planning and Strategic Sourcing to align committed capacity with demand forecasts and identify supply risks early.

  • Reconcile invoices against contracted pricing, delivered capacity, usage, credits, and other commercial terms.

  • Maintain accurate records of commitments, renewals, open claims, invoices, and vendor obligations.

  • Build the vendor-management processes, dashboards, and operating cadence fal needs as its infrastructure footprint scales.

What you bring

  • 5+ years in infrastructure vendor management, cloud partnerships, strategic sourcing, procurement operations, or a related function.

  • Direct experience managing commercial relationships with hyperscalers, neoclouds, GPU providers, data-center operators, or other large-scale infrastructure vendors.

  • Experience operationalizing complex infrastructure agreements involving committed capacity, SLAs, service credits, provisioning requirements, renewals, and commercial remedies.

  • A track record of resolving vendor-performance issues and holding strategic partners accountable without damaging long-term relationships.

  • Strong analytical and operational skills, including performance reporting, invoice reconciliation, credit calculations, and contract-compliance tracking.

  • Enough technical fluency to work effectively with infrastructure engineering and capacity-planning teams and understand the operational impact of vendor failures.

  • Experience running QBRs, vendor scorecards, executive escalations, and corrective-action plans.

  • Comfort building processes and systems from scratch in a fast-moving environment.

  • Clear, direct communication and excellent attention to detail.

Bonus points

  • Experience managing GPU capacity across multiple providers.

  • Experience allocating constrained capacity during shortages or periods of rapid demand growth.

  • Familiarity with GPU availability, provisioning, utilization, networking, and cloud infrastructure economics.

  • Experience recovering material service credits or negotiating remedies following major performance failures.

  • Experience establishing a vendor-management function at a hyperscaler, neocloud, AI infrastructure company, or high-growth technology company.

What success looks like

Within your first six months:

  • Every strategic infrastructure vendor has clear ownership, performance measures, and a regular review cadence.

  • Contractual commitments, renewals, credits, and open escalations are tracked in one reliable system.

  • Invoice and capacity discrepancies are identified and resolved consistently.

  • Vendor risks are surfaced before they affect customers or constrain growth.

  • fal has a repeatable process for holding infrastructure partners accountable.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$180k – $250k per year • Remote • Full-Time • 5+ years exp
Python
AI/ML
InfiniBand
NCCL
NVLink
DevOps
Ansible
Cilium
containerd
etcd
Kubernetes
KubeVirt
KVM
OpenStack
QEMU
SLURM
Cybersecurity
Calico
Tcpdump
Wireshark
Apply
$180k – $250k per year • Remote • Full-Time • 5+ years exp
Python
DevOps
Ansible
AWS
Azure
eBPF
GCP
Kubernetes
Cybersecurity
Tcpdump
Wireshark
Apply
$220k – $290k per year • In office • Full-Time • 10+ years exp • High School Diploma • San Francisco
Python
SQL
AI/ML
Diffusion Models
Multimodal AI
DevOps
AWS
Azure
GCP
VMWare
Apply
$200k – $250k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco
Apply
$180k – $250k per year • In office • Full-Time • 5+ years exp • San Francisco
Python
Rust
SQL
TypeScript
JavaScript
Databases
PostgreSQL
Frontend
Next.js
React.js
Apply
$180k – $210k per year • Equity • In office • Full-Time • San Francisco
Node JS
JavaScript
Databases
PostgreSQL
DevOps
PagerDuty
Web3
TRM Labs
Management
Slack
Apply
$252k – $335k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco
AI/ML
ChatGPT
Human-in-the-Loop
OpenAI
OpenAI Codex
DevOps
SLI/SLO/SLA
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
$160k – $283k per year • Equity • In office • 5+ years exp • San Francisco
AI/ML
AI Agents
Apply
$185k – $385k per year • Remote/Hybrid • Full-Time • 5+ years exp • San Francisco
JavaScript
Python
Databases
MySQL
PostgreSQL
AI/ML
OpenAI
Frontend
React.js
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.