372,254open jobs
9,642companies
49,764added this week
Browse all
Salary
$266k – $395k per year
Location
Remote/Hybrid (San Francisco, San Jose, Bellevue, United States)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Lambda is a specialized AI infrastructure provider that offers high-performance GPU cloud compute, clusters, and hardware tailored for deep learning and machine learning workloads. The company enables AI developers and research teams to train, fine-tune, and deploy large language models efficiently through scalable cloud instances and dedicated on-premise GPU servers.

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.

If you'd like to build the world's best AI cloud, join us.

*Note: This position requires presence in our San Francisco, San Jose, or Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.

The Compute Software Engineering role plays a key part in the foundation of both the Public and Private Cloud offerings by enabling Bare-Metal and Virtual Machine provisioning. In this role you will: Develop and integrate workflows for compute instance lifecycle, system health, host and cluster validation, and contribute to and evolve the internal platform tooling, and workflows. As part of the Compute Engineering team, this position focuses on improving the bare-metal and VM developer and customer experiences, streamlining the software delivery lifecycle, building, testing, and shipping software efficiently and reliably.

The position works closely with various engineering teams across the company and contributes directly to Lambda’s ability to scale engineering productivity, maintain high delivery velocity, and support the company’s continued growth by making the right engineering workflows easier, faster, and more reliable.

What You’ll Do

We are seeking a Software Engineer with good experience in backbone infrastructure that build, scale and optimize GPU-first cloud infrastructure that enables High Performance Computing and demanding AI workloads. In this role you will be responsible for:

  • Design, develop and maintain software for GPU/CPU compute infrastructure with focus on performance, scalability, and reliability.

  • Implement and develop services for baremetal and VM instancing.

  • Develop distributed systems for managing and orchestrating compute resources across various SKU’s.

  • Troubleshoot and debug complex issues in a production and development environment.

  • On-call and incident ownership

  • Collaboration across multiple teams and drive ambiguity in requirements or solutions on RFC’s.

You

  • 5+ years of experience working with Go (Golang) or Python in production environments.

  • 5+ years of experience with bare metal & virtualization hardware management and configuration.

  • Are comfortable working in Linux environments and debugging issues at the OS, hardware, and networking layers.

  • Can independently troubleshoot complex systems and communicate effectively across software, infrastructure, and vendor teams.

Nice to Have

  • Familiarity with GPU Infrastructure or high-performance computing environments.

  • Experience with Slurm or Kubernetes-based cluster management.

  • Experience with core public cloud internals (Virtualization, KVM, QEMU, Security and Fleet health)

  • Experience with durable execution platforms like Temporal.

Salary Range Information

The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.

About Lambda

  • Founded in 2012, with 500+ employees, and growing fast

  • Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove

  • We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG

  • Our values are publicly available: https://lambda.ai/careers

  • We offer generous cash & equity compensation

  • Health, dental, and vision coverage for you and your dependents

  • Wellness and commuter stipends for select roles

  • 401k Plan with 2% company match (USA employees)

  • Flexible paid time off plan that we all actually use

Equal Opportunity Employer

Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
372,254 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
In office • Full-Time • PhD • Guadalajara
Python
SQL
Databases
Amazon Redshift
Snowflake
Trino
AI/ML
dbt
DevOps
AWS
AWS Lambda
CI/CD
Git
GitHub Actions
Terraform
GitHub
Apply
$54k – $128k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Mexico City
C#
JavaScript
Python
TypeScript
C#
ASP.NET Core
Python
FastAPI
Databases
PostgreSQL
AI/ML
AI Agents
Anthropic
Embeddings
Function Calling
LLM
LLM Guardrails
Model Context Protocol
OpenAI
Semantic Search
Semantic Search
Frontend
Angular
React.js
DevOps
AWS
Azure
CI/CD
GitHub
GitHub Actions
Vector
Vercel
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Robot SWE 1 day ago
$100k – $200k per year • Equity 0.2–1.2% • In office • Full-Time • Bachelor's Degree • Los Angeles
Bash
C++
Python
AI/ML
Computer Vision
Human-in-the-Loop
DevOps
Docker
Git
Ubuntu
Robotics
Autonomous Navigation
Motion Planning
Obstacle Avoidance
ROS2
Sensor Fusion
SLAM
Teleoperation
Apply
$117k – $177k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • Chicago • Dallas
Apex
C#
Java
Python
Apex
Salesforce Data Cloud
AI/ML
Agentforce
AI Agents
Chain-of-Thought
Embeddings
Few-Shot Learning
Fine-tuning
Hallucination
LLM Guardrails
NLP
Prompt Engineering
RAG
RLHF
Semantic Search
Semantic Search
Frontend
GraphQL
DevOps
CI/CD
Git
Analytics
A/B Testing
Marketing
Salesforce
Apply
$89k – $119k per year • Equity • In office • Full-Time • Atlanta
DevOps
AWS Lambda
AWS
Marketing
Zendesk
Apply
$380k – $445k per year • Equity • Remote/Hybrid • Full-Time • 1+ year exp • Master's Degree • San Francisco
Go
Python
SQL
Databases
DynamoDB
Presto
DevOps
AWS Lambda
AWS
Amazon S3
Apply
$297k – $440k per year • Equity • Remote/Hybrid • Full-Time • 3+ years exp • San Francisco • San Jose • Bellevue
AI/ML
InfiniBand
DevOps
AWS Lambda
Incident Management
AWS
HPC
Apply
$278k – $325k per year • Equity • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco • San Jose
Databases
Google BigQuery
DevOps
AWS Lambda
AWS
Cybersecurity
Okta
Management
Google Workspace
Jira
Marketing
Salesforce
Apply
$137k – $183k per year • Equity • Remote/Hybrid • Full-Time • 5+ years exp • Elk Grove Village
AI/ML
InfiniBand
DevOps
AWS Lambda
AWS
Apply
$94k – $266k per year • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Boston • Milwaukee • Dallas • Columbus • Kirkland
Python
AI/ML
Claude
Claude Code
CoreWeave
CUDA
CUDA Toolkit
NCCL
DevOps
Ansible
Configuration Management
HPC
Incident Management
Kubernetes
Rest API
SLURM
Terraform
Apply
$94k – $266k per year • Remote/Hybrid • Full-Time • 5+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
Databases
Databricks
Google BigQuery
SAP HANA
Snowflake
AI/ML
Knowledge Graph
DevOps
Azure
Apply
$133k – $366k per year • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
JavaScript
Python
TypeScript
AI/ML
Anthropic
Edge AI
Gemini
OpenAI
Management
ServiceNow
Marketing
Salesforce
Apply
$218k – $365k per year • In office • Full-Time • 15+ years exp • PhD • Washington • New York • San Francisco
AI/ML
Agentforce
AI Agents
DevOps
AppDynamics
AWS
Azure
CI/CD
Docker
GitLab
Grafana
Jenkins
Kubernetes
Spinnaker
Terraform
Cybersecurity
Crowdstrike
Prisma Cloud
Qualys Cloud Platform
Threat Modeling
Wiz
Marketing
Salesforce
Apply
$149k – $224k per year • In office • Full-Time • 3+ years exp • PhD • San Francisco • Washington
Apex
Apex
MuleSoft
AI/ML
AI Agents
Claude
Claude Code
Copilot
Agentforce
Frontend
GraphQL
DevOps
AWS
CI/CD
Docker
gRPC
Kong
Kubernetes
Platform Engineering
Terraform
Marketing
Salesforce
Apply
See all jobs
This is one of many
372,254 more open roles from verified company boards, updated every day.