376,406open jobs
9,791companies
48,125added this week
Browse all
Salary
$105k – $140k per year
Location
In office (San Jose)
Employment
Full-Time
Overview
Company
Impact
Profile match
Lambda is a specialized AI infrastructure provider that offers high-performance GPU cloud compute, clusters, and hardware tailored for deep learning and machine learning workloads. The company enables AI developers and research teams to train, fine-tune, and deploy large language models efficiently through scalable cloud instances and dedicated on-premise GPU servers.

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.

If you'd like to build the world's best AI cloud, join us.

*Note: This position requires presence in our San Jose Data Centers 5 days; shift work

What You'll Do

  • Ensure new server, storage and network infrastructure is properly racked, labeled, cabled, and configured

  • Troubleshoot hardware and software issues in some of the world’s most advanced systems

  • Document data center layout and network topology in DCIM software

  • Work with supply chain & manufacturing teams to ensure timely deployment of systems and project plans for large-scale deployments

  • Manage a parts depot inventory and track equipment through the delivery-store-stage-deploy-handoff process in each of our data centers

  • Work closely with HW Support team to ensure data center infrastructure-related support tickets are resolved

  • Work with RMA team to ensure faulty parts are returned and replacements are ordered

  • Follow installation standards and documentation for placement, labeling, and cabling to drive consistency and discoverability across all data centers

You

  • Are familiar with critical infrastructure systems supporting data centers, such as power distribution, air flow management, environmental monitoring, capacity planning, DCIM software, structured cabling, and cable management

  • Are Fluent in English and Spanish

  • Are someone who pays attention to detail and has the ability to follow instructions

  • Are action-oriented and have a strong willingness to learn

  • Are willing to travel for bring up of new data center locations

Nice to Have

  • Experience with troubleshooting server hardware

  • Experience with/or knowledge of network topology

  • Familiarity with ticketing systems like JIRA and Zendesk

  • Experience with Linux administration

  • Experience with working in large-scale distributed data center environments

  • Experience with Supermicro & Nvidia hardware

Salary Range Information

This is a salaried non-exempt role, eligible for overtime. The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.

About Lambda

  • Founded in 2012, with 500+ employees, and growing fast

  • Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove

  • We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG

  • Our values are publicly available: https://lambda.ai/careers

  • We offer generous cash & equity compensation

  • Health, dental, and vision coverage for you and your dependents

  • Wellness and commuter stipends for select roles

  • 401k Plan with 2% company match (USA employees)

  • Flexible paid time off plan that we all actually use

Equal Opportunity Employer

Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
376,406 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Jose
$115k – $224k per year (Estimated) • Remote/Hybrid • Full-Time • 7+ years exp • Bachelor's Degree • Wayne • Charlotte • Dallas
Python
DevOps
Amazon ECS
AWS
AWS Lambda
CI/CD
Docker
Kubernetes
OpenTelemetry
Platform Engineering
SLI/SLO/SLA
Apply
In office • 8+ years exp • Bachelor's Degree
Java
Node JS
TypeScript
JavaScript
Java
Spring Boot
Databases
Amazon DocumentDB
DynamoDB
AI/ML
Copilot
LLM
Prompt Engineering
DevOps
Amazon S3
AWS
AWS Fargate
AWS Lambda
CI/CD
GitHub
IAM
Apply
$89k – $190k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Toronto
Scala
Databases
Apache Kafka
Databricks
PostgreSQL
Redis
DevOps
Amazon Kinesis
AWS
Incident Management
Kubernetes
Terraform
Apply
$83k – $177k per year (Estimated) • In office • Full-Time • 10+ years exp • Wellington
PowerShell
Python
SQL
DevOps
Ansible
AWS
Azure
CI/CD
Platform Engineering
Terraform
Windows Server
Apply
$73k – $190k per year (Estimated) • In office • Full-Time • Wellington
PowerShell
SQL
Databases
Cassandra
Databricks
MS SQL
DevOps
Ansible
AWS
Azure
Azure DevOps
CI/CD
GitHub
Jenkins
Kubernetes
SaltStack
Terraform
Windows Server
Apply
$231k – $342k per year • Equity • Remote/Hybrid • Full-Time • 5+ years exp • San Francisco • San Jose • Bellevue
SQL
AI/ML
CUDA
CUDA Toolkit
Fine-tuning
InfiniBand
LLM
NVIDIA NeMo
NVIDIA NIM
RAG
DevOps
AWS Lambda
HPC
Kubernetes
SLI/SLO/SLA
SLURM
AWS
Apply
$180k – $240k per year • Equity • Remote/Hybrid • Full-Time • 7+ years exp • San Jose • San Francisco
AI/ML
CUDA
CUDA Toolkit
Fine-tuning
InfiniBand
NCCL
PyTorch
DevOps
AWS Lambda
HPC
Kubernetes
SLURM
AWS
Apply
$89k – $119k per year • Equity • In office • Full-Time • Atlanta
DevOps
AWS Lambda
AWS
Marketing
Zendesk
Apply
$380k – $445k per year • Equity • Remote/Hybrid • Full-Time • 1+ year exp • Master's Degree • San Francisco
Go
Python
SQL
Databases
DynamoDB
Presto
DevOps
AWS Lambda
AWS
Amazon S3
Apply
$297k – $440k per year • Equity • Remote/Hybrid • Full-Time • 3+ years exp • San Francisco • San Jose • Bellevue
AI/ML
InfiniBand
DevOps
AWS Lambda
Incident Management
AWS
HPC
Apply
$122k – $195k per year • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree • San Jose
Apply
$91k – $183k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Longmont • San Jose
Python
DevOps
CI/CD
Apply
$180k – $225k per year • In office • Full-Time • 15+ years exp • Bachelor's Degree • San Jose
Apply
$216k – $270k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Jose
AI/ML
AI Agents
Apply
$145k – $180k per year • In office • Full-Time • 5+ years exp • San Jose
Verilog
Apply
See all jobs
This is one of many
376,406 more open roles from verified company boards, updated every day.