697,753open jobs
40,891companies
106,255added this week
Browse all
Salary
$200k – $322k per year
Location
In office (Santa Clara, United States)
Seniority
Principal · 12+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA's AI Factory Infrastructure team develops global reference designs, de-risks technologies needed for the next generation of compute and network products and builds AI infrastructure at scale to validate solutions at scale. As a TPM on the team, you’ll be responsible for leading teams that are highly multi-functional, including hardware, software and facility infrastructure teams, to deliver solutions and infrastructure. This role offers a truly unique blend of developing future solutions and products with building infrastructure at scale.

What You Will Be Doing:

  • Collaborate with product owners and technical leads to identify and collect requirements for next-generation AI Factories.

  • Build, supervise, and complete long-term programs including schedules, resourcing, and checkpoints.

  • Work with data center and hardware teams to find creative solutions to hard problems, and co-develop solutions and mitigation strategies.

  • Lead planning with key internal partners on capacity demands with engineering roadmaps and data center expansions.

  • Own end-to-end delivery of 100MW+ AI factory data center deployments, from construction readiness through commissioning, turn-up, and handoff to operations.

  • Coordinate general contractors, colocation providers, utilities, and OEMs to align electrical/mechanical scope, long-lead equipment, and site logistics for large-scale AI cluster deployments.

  • Drive integrated readiness reviews and acceptance criteria across power, liquid cooling, networking, and platform/hardware teams to ensure performance and reliability targets are met for AI factory applications.

  • Develop program plans for government grants and initiatives.

  • Translate program requirements into Basis Of Design documents

  • Bring together team members and foster a collaborative approach to delivery while holding team members accountable to action items and timelines.

What We Need To See:

  • Outstanding long-term planning and execution skills to carry our data center lifecycle planning, including large-scale AI factory buildouts and expansions.

  • Experience managing end-to-end data center deployments for high-density AI infrastructure, including commissioning, readiness reviews, turn-up, and operational handoff.

  • Demonstrated ability to coordinate across colocation providers, general contractors, utilities, and OEMs to deliver complex electrical and mechanical scope (e.g., at 100MW+ campus scale).

  • Strong technical and program leadership across power delivery, liquid cooling, networking, and compute/platform teams to define acceptance criteria and ensure performance and reliability targets are met.

  • 12+ years of experience providing program and project management leadership for data center projects covering construction of mechanical, electrical, and plumbing with large-scale server, storage, and network deployments.

  • BS or MS degree in Engineering (or equivalent experience).

Ways To Stand Out From the Crowd:

  • In-depth knowledge of infrastructure (hardware and software) data center facilities infrastructure (electrical and mechanical) technologies.

  • Familiarity with NVIDIA’s AI compute technology stack and ability to translate platform requirements into data center infrastructure designs (power delivery, liquid cooling, space, and network topology) at scale.

  • Experience with colocation data center environments

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 200,000 USD - 322,000 USD for Level 5, and 240,000 USD - 379,500 USD for Level 6.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 24, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
697,753 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$37k – $86k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru
Python
DevOps
GCP
AWS
Docker
Kubernetes
Linux
Windows
Apply
$39k – $84k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru
AI/ML
CUDA Toolkit
ONNX
OpenCL
TensorFlow
PyTorch
CUDA
Apply
$39k – $83k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Bengaluru
Python
Go
Rust
DevOps
SLURM
CI/CD
ArgoCD
Jenkins
Kubernetes
Incident Management
Linux
DNS
DHCP
Apply
$41k – $101k per year (Estimated) • In office • Full-Time • 2+ years exp • Shanghai • Beijing
C
C++
C
Pthreads
C++
TensorFlow C++
LLVM
AI/ML
CUDA Toolkit
Speech Recognition
TensorRT
OpenMP
TensorFlow
CUDA
cuDNN
MLIR
Apache TVM
CUTLASS
DevOps
CI/CD
Apply
$34k – $84k per year (Estimated) • In office • Full-Time • 5+ years exp • Shenzhen • Shanghai
C++
DevOps
Linux
Apply
$148k – $325k per year (Estimated) • In office • Master's Degree • Santa Clara
Python
AI/ML
LangChain
Reinforcement Learning
AI Agents
PyTorch
RAG
Ray
Hugging Face
Multi-Agent Systems
Robotics
Digital Twin
Apply
$167k – $291k per year • Equity • In office • Full-Time • 8+ years exp • Santa Clara
Python
JavaScript
TypeScript
AI/ML
Machine Learning
Frontend
Vue.js
GraphQL
Angular
React.js
Mobile
JUnit
DevOps
CI/CD
Jenkins
Unix
Management
ServiceNow
Agile
Scrum
QA
TestNG
Selenium
Playwright
Apply
$191k – $334k per year • Equity • In office • Full-Time • 12+ years exp • Bachelor's Degree • Santa Clara
Python
Java
Databases
ElasticSearch
OpenSearch
AI/ML
Embeddings
AI Agents
RAG
Semantic Search
Semantic Search
Recommender Systems
Machine Learning
Mobile
Algolia
DevOps
Platform Engineering
Management
ServiceNow
Apply
$100k – $127k per year • In office • 10+ years exp • High School Diploma • Santa Clara
Apply
up to $160k per year • In office • 10+ years exp • Bachelor's Degree • Santa Clara
DevOps
AWS
Apply
See all jobs
This is one of many
697,753 more open roles from verified company boards, updated every day.