368,958open jobs
9,466companies
48,032added this week
Browse all
Salary
$338k – $438k per year
Location
Remote/Hybrid (Bellevue, San Francisco, United States)
Seniority
Principal · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Lambda is a specialized AI infrastructure provider that offers high-performance GPU cloud compute, clusters, and hardware tailored for deep learning and machine learning workloads. The company enables AI developers and research teams to train, fine-tune, and deploy large language models efficiently through scalable cloud instances and dedicated on-premise GPU servers.

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.

If you'd like to build the world's best AI cloud, join us.

*Note: This position requires presence in our Bellevue or San Francisco office location 4 days per week; Lambda's designated work from home day is currently Tuesday.

About the Role

Lambda's hardware is the product. Every training run, fine-tune, and inference workload our customers launch runs on hardware someone at Lambda decided to buy, configure, and ship as a product. This role owns that decision. You'll manage the GPU (graphics processing unit) fleet lifecycle as a product: new NVIDIA platform introductions such as the B200 and H200 class and the generations that follow, node and cluster configurations, InfiniBand fabric options, and the roadmap for what hardware Lambda offers, when, and at what configuration.

You'll sit at the layer where Lambda meets silicon. Externally, you'll work directly with NVIDIA and with ODM (original design manufacturer) partners on roadmap alignment. Internally, you'll work with our data center, supply chain, and infrastructure engineering teams to turn silicon roadmaps into sellable, reliable products: On-Demand GPU Instances, 1-Click Clusters, and our largest multi-node reserved deployments. You'll translate customer demand signals and benchmark data into fleet investment recommendations that shape where Lambda puts its capital.

Great product managers at Lambda are defined by three things: insight, influence, and execution. Insight means you look at the data, determine what it means for customers and business, and then figure out what to do about it. But, a great idea doesn’t mean anything in a vacuum. That is where influence comes in. Influence means you take that idea and get others to want to buy into it; you win over engineers, designers, executives, and partners without relying on authority. But a great idea that everyone is excited about doesn’t matter unless it is delivered to customers. Execution means you work with the right people to get the idea launched, then measure and iterate. We hire product managers who learn new domains fast and reason rigorously from evidence. Deep background in GPU hardware, systems, or the semiconductor ecosystem are also desired.

We value diverse backgrounds, experiences, and skills, and we are excited to hear from candidates who can bring unique perspectives to our team. If you do not exactly meet this description but believe you may be a good fit, please still apply and help us understand your readiness for this role.

What You'll Do

  • Own the Hardware Roadmap: Define what GPU platforms Lambda offers, when, and at what configuration, from today's B200 and H200 class systems through next-generation NVIDIA platforms.

  • Drive New Platform Introductions: Lead new NVIDIA platform introductions end to end, from early roadmap alignment through general availability as On-Demand GPU Instances and 1-Click Clusters.

  • Translate Demand Into Investment: Turn customer demand signals and benchmark data into fleet investment recommendations, and defend those recommendations with executives and finance.

  • Define Node and Cluster Configurations: Specify node configurations and InfiniBand fabric options for clusters from 64 to 1,024+ GPUs, working with infrastructure engineering to keep NCCL (NVIDIA Collective Communications Library) and MPI (Message Passing Interface) pre-configured and performant out of the box. This role owns defining this performance standard.

  • Align Partner Roadmaps: Work directly with NVIDIA and ODM partners so Lambda's product plans and our partners' silicon and systems roadmaps land together, not months apart.

  • Turn Silicon Into Product: Partner with data center, supply chain, and infrastructure engineering teams to convert silicon roadmaps into reliable, sellable SKUs (stock keeping units) with clear positioning and launch plans.

  • Win Adoption: Bring engineers, designers, executives, and customers along with your roadmap through clear writing, honest data, and direct conversation.

  • Ship, Measure, Iterate: Launch new hardware products, define the metrics that tell you whether they are working, and iterate on configuration and positioning based on what the data says.

You

  • Have 7+ years of product management experience on technical infrastructure, hardware, systems, or platform products; senior candidates should bring 10+ years and ownership of a multi-team or multi-product roadmap

  • Turn data and customer signal into a clear decision about what to build next, and can walk through examples where you saw the what and the so what, and determined the now what

  • Have a track record of getting engineers, executives, and external partners to adopt a plan because you made them want to, not because you outranked them

  • Have shipped products that required coordinating hardware, software, and operations teams against hard external deadlines

  • Are fluent in technical conversations with hardware and systems engineers, and comfortable discussing GPU architectures, interconnects, memory, and data center constraints

  • Have made hardware, capacity, or fleet investment recommendations that committed significant capital under uncertainty, and can walk through how the decision played out

  • Write documents that lead with the conclusion and the evidence, not the background

  • Able to define iterative plans that move an organization from the current state towards the desired outcome

Nice to Have

  • Have direct experience in GPU or accelerator hardware, or systems and HPC (high-performance computing) product management

  • Have worked inside the semiconductor or original equipment manufacturer (OEM)/ODM ecosystem, or directly with NVIDIA or another silicon vendor on roadmap alignment

  • Have owned cloud infrastructure capacity planning or fleet economics

  • Have hands-on familiarity with distributed training infrastructure, including InfiniBand fabrics and NCCL

  • Have run benchmark programs and used the results to drive buy or configure decisions

Salary Range Information

The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.

About Lambda

  • Founded in 2012, with 500+ employees, and growing fast

  • Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove

  • We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG

  • Our values are publicly available: https://lambda.ai/careers

  • We offer generous cash & equity compensation

  • Health, dental, and vision coverage for you and your dependents

  • Wellness and commuter stipends for select roles

  • 401k Plan with 2% company match (USA employees)

  • Flexible paid time off plan that we all actually use

Equal Opportunity Employer

Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,958 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bellevue
$42k – $104k per year (Estimated) • In office • Internship • 5+ years exp
Bash
PowerShell
Python
DevOps
Amazon EC2
Amazon EKS
Ansible
AWS
AWS Lambda
Azure
CentOS Stream
Chef
CI/CD
CloudFormation
Configuration Management
Datadog
FinOps
GCP
Git
GitHub Actions
GitLab CI
Grafana
Hyper-V
Istio
Jenkins
Kubernetes
KVM
Linkerd
Platform Engineering
Prometheus
Puppet
Service Mesh
Splunk
Terraform
Ubuntu
VMWare
Windows Server
Amazon CloudWatch
Amazon ECS
Amazon S3
API Gateway
AWS Step Functions
GitHub
GitLab
IAM
Cybersecurity
GDPR
ISO 27001
SOC 2
Apply
Data Engineer 1 day ago
In office • Full-Time • Singapore
Node JS
Python
SQL
JavaScript
Python
Beautiful Soup
Databases
Apache Kafka
MySQL
PostgreSQL
RabbitMQ
SQLite
AI/ML
Hadoop
Spark
DevOps
AWS
AWS Lambda
Azure
CI/CD
GCP
Amazon ECS
Amazon EventBridge
Amazon S3
Analytics
ETL/ELT
QA
Selenium
Apply
up to $48k per year (net) • Remote/Hybrid • Moscow
Node JS
Python
TypeScript
JavaScript
Node JS
Nest.JS
Python
Django
FastAPI
Databases
Apache Kafka
pgvector
Pinecone
PostgreSQL
Qdrant
RabbitMQ
Redis
AI/ML
Chain-of-Thought
Claude
Claude Code
Copilot
Cursor
LangChain
LlamaIndex
LLM
Prompt Engineering
RAG
Anthropic
Function Calling
OpenAI
Structured Outputs
Frontend
GraphQL
Next.js
React.js
Redux
Redux Toolkit
Zustand
Mobile
State Management
DevOps
AWS
CI/CD
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Kubernetes
Prometheus
Yandex Cloud
GitHub
GitLab
Apply
$24k – $64k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Chennai
Python
Scala
SQL
TypeScript
JavaScript
Java
Python
pySpark
Java
Spring Boot
Databases
Apache Kafka
Databricks
AI/ML
AI Agents
Copilot
Google ADK
LLM
NLP
Prompt Engineering
Spark
Devin
Model Context Protocol
Frontend
Angular
React.js
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
OpenShift
GitHub
Analytics
ETL/ELT
Apply
$12k – $35k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Bengaluru
Databases
MySQL
Oracle
PostgreSQL
AI/ML
AI Agents
DevOps
AIOps
Amazon EC2
AWS
CI/CD
CloudFormation
GitHub Actions
Kubernetes
Terraform
Amazon CloudWatch
Amazon S3
GitHub
IAM
Apply
$89k – $119k per year • Equity • In office • Full-Time • Atlanta
DevOps
AWS Lambda
AWS
Marketing
Zendesk
Apply
$380k – $445k per year • Equity • Remote/Hybrid • Full-Time • 1+ year exp • Master's Degree • San Francisco
Go
Python
SQL
Databases
DynamoDB
Presto
DevOps
AWS Lambda
AWS
Amazon S3
Apply
$297k – $440k per year • Equity • Remote/Hybrid • Full-Time • 3+ years exp • San Francisco • San Jose • Bellevue
AI/ML
InfiniBand
DevOps
AWS Lambda
Incident Management
AWS
HPC
Apply
$278k – $325k per year • Equity • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco • San Jose
Databases
Google BigQuery
DevOps
AWS Lambda
AWS
Cybersecurity
Okta
Management
Google Workspace
Jira
Marketing
Salesforce
Apply
$137k – $183k per year • Equity • Remote/Hybrid • Full-Time • 5+ years exp • Elk Grove Village
AI/ML
InfiniBand
DevOps
AWS Lambda
AWS
Apply
$161k – $310k per year (Estimated) • In office • Bachelor's Degree • Bellevue
AI/ML
Edge AI
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
Apply
$236k – $339k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Bellevue
Java
Python
Databases
Snowflake
AI/ML
AI Agents
Feature Store
Apply
$215k – $260k per year • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco • Sunnyvale • Bellevue
C++
Go
DevOps
AWS
Azure
Cilium
eBPF
GCP
KVM
VMWare
Apply
$160k – $210k per year • Remote/Hybrid • Full-Time • 6+ years exp • Denver • Bellevue
C#
C++
Go
Python
SQL
Databases
Azure Cosmos DB
Azure SQL Database
DynamoDB
MySQL
AI/ML
Edge AI
DevOps
AWS
Azure
Azure AKS
CI/CD
GCP
Google GKE
Kubernetes
SLI/SLO/SLA
Analytics
Power BI
Management
UiPath
Apply
$159k – $285k per year (Estimated) • In office • Full-Time • 10+ years exp • Bellevue
AI/ML
AI Agents
Management
Smartsheet
Apply
See all jobs
This is one of many
368,958 more open roles from verified company boards, updated every day.