1,350,029open jobs
78,789companies
207,011added this week
Browse all
Salary
$272k – $431k per year
Location
In office (United States, Santa Clara)
Seniority
Principal · 15+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 8, 2026. First seen by Alion on Aug 11, 2026. NVIDIA scores A on the Alion truth index.

Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA is seeking a principal-level software engineer to build the next generation of our Kubernetes platform. Our teams build foundational capabilities for self-service GPU infrastructure, managed Kubernetes control planes, management of cluster operations, and automation for large-scale AI and improved computational environments.

In this role, you will work at the intersection of platform architecture and production-grade Kubernetes lifecycle systems. You will lead technical strategy and execution for systems that make clusters easier to provision, upgrade, operate, and scale across cloud and on-premises environments.

What you will be doing:

  • Lead the architecture and development of core Kubernetes platform capabilities, including cluster management, control plane services, fleet lifecycle, and day-2 operations.

  • Design and build highly reliable distributed systems and APIs for provisioning, managing, upgrading, and remediating Kubernetes clusters at scale.

  • Define technical requirements, validation criteria, production-readiness practices, and the direction for declarative workflows and automation across the Kubernetes stack.

  • Collaborate across engineering teams to create cohesive platform experiences spanning management APIs, lifecycle orchestration, runtime integration, and fleet consistency.

  • Lead the diagnosis and resolution of complex platform issues spanning infrastructure, runtime, networking, hardware, and operations, improving the scalability, resilience, and operability of systems supporting large-scale AI deployments.

  • Influence engineering standards, architectural decisions, and long-term platform strategy.

  • Mentor senior engineers and raise the bar for design quality, execution, and engineering rigor across the organization.

What we need to see:

  • BS or MS degree in Computer Science, Computer Engineering, or a related field, or equivalent experience.

  • 15+ years of relevant software engineering experience, including experience building and operating large-scale production systems.

  • Deep expertise in Kubernetes internals, APIs, controllers or operators, and cluster lifecycle management.

  • A strong background in distributed systems design, reliability, scalability, and failure recovery.

  • Proven experience building platform software, infrastructure control planes, or foundations for managed services.

  • Strong programming skills in one or more systems or cloud-native languages, such as Go, Python, Rust, or C++.

  • Experience designing clear APIs and abstractions for platform consumers and engineering teams.

  • Demonstrated ability to provide technical leadership across team boundaries and drive ambiguous, cross-functional initiatives to completion.

  • Excellent communication and collaboration skills, backed by a sustained record of significant technical contributions and recognized expertise influencing department-level architecture and high-priority company initiatives.

Ways to stand out from the crowd:

  • Experience building Kubernetes platforms or managed Kubernetes services.

  • Expertise in fleet management, cluster upgrades, node lifecycle, remediation, or day-2 operations.

  • Experience with declarative infrastructure, Kubernetes controllers, GitOps, or policy-driven platform automation.

  • Familiarity with both public-cloud and bare-metal infrastructure environments.

  • Experience supporting AI, GPU, HPC, or other large-scale accelerated computing platforms.

With competitive salaries and a generous benefits package, NVIDIA is widely considered to be one of the technology industry's most desirable employers. We have some of the most forward-thinking and versatile people in the world working with us, and our engineering teams are growing fast in some of the most impactful fields of our generation: GPU, Cloud, Cloud Development or Distributed Cloud. If you're a creative Principal Software Engineer who enjoys autonomy and shares our passion for technology, we want to hear from you.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 15, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,350,029 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
Santa Clara
≈ $143k – $290k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Seattle
Management
Google Maps
Apply
≈ $156k – $316k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Sunnyvale
DevOps
GCP
Apply
$186k – $265k per year • Hybrid • Full-Time • 12+ years exp • Santa Clara
SQL
Databases
MySQL
PostgreSQL
Cassandra
DynamoDB
RabbitMQ
Apache Kafka
Amazon Aurora
DevOps
Jaeger
Istio
OpenTelemetry
Linkerd
Prometheus
GitOps
ArgoCD
AWS
Docker
Kubernetes
Chaos Engineering
Service Mesh
Amazon EKS
Progressive Delivery
Incident Management
SLI/SLO/SLA
Cybersecurity
Zscaler
Zero Trust
SBOM
SLSA
Management
Agile
Apply
≈ $137k – $277k per year (Estimated) • Hybrid • TS/SCI • Full-Time • 10+ years exp • Bachelor's Degree • Herndon
Python
JavaScript
TypeScript
Python
FastAPI
Django
Databases
PostgreSQL
AI/ML
AI Agents
Knowledge Graph
Machine Learning
Frontend
D3.js
React.js
DevOps
CI/CD
Git
Docker
Kubernetes
Linux
Analytics
Plotly
Management
Agile
Apply
≈ $230k – $419k per year (Estimated) • In office • 8+ years exp • New York
Python
C#
C#
.NET
Databases
Apache Kafka
Redpanda
DevOps
Terraform
OpenTelemetry
Azure
AWS
Kubernetes
Amazon EKS
Amazon S3
Cybersecurity
GDPR
Apply
$104k – $166k per year • In office • Secret • 12+ years exp • PhD • United States
Python
JavaScript
Rust
Cybersecurity
Zero Trust
Apply
≈ $105k – $199k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Munich
Python
C++
Bash
DevOps
RTOS
CI/CD
Docker
Gerrit
Linux
Apply
$168k – $203k per year • Equity • In office • 8+ years exp • Bachelor's Degree • Santa Cruz
Python
C++
Management
Agile
Apply
$20k per year • In office • 2+ years exp • Tomsk
Python
PowerShell
Bash
DevOps
Linux
Windows
TCP/IP
VPN
Cybersecurity
Nessus
ISO 27001
Ghidra
OpenVAS
VirusTotal
IBM QRadar
SIEM
DLP
Kaspersky
Apply
$176k – $282k per year • In office • TS/SCI • 12+ years exp • Associate's Degree • Fort Meade
Python
SQL
AI/ML
Anomaly Detection
Machine Learning
DevOps
Azure
AWS
Apply
$184k – $288k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Santa Clara
Python
Go
JavaScript
TypeScript
SQL
Node JS
Databases
Neo4j
AI/ML
Model Context Protocol
Prompt Engineering
AI Agents
LLM
RAG
GraphRAG
LLM Evaluation
LLM Guardrails
Multi-Agent Systems
Frontend
GraphQL
Next.js
React.js
DevOps
Terraform
Ansible
GitLab CI
CI/CD
Jenkins
Apply
≈ $190k – $319k per year (Estimated) • Remote (United States) • Full-Time • 8+ years exp • Bachelor's Degree • United States
Rust
DevOps
Kubernetes
Linux
Unix
Apply
$184k – $288k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Santa Clara • Austin • Durham
Python
Go
JavaScript
Rust
Node JS
Rust
PyO3
AI/ML
AI Agents
LLM
RAG
LLM Guardrails
DevOps
OpenTelemetry
Apply
$152k – $242k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Seattle • Santa Clara
DevOps
Terraform
Ansible
CI/CD
Kubernetes
GitLab
Linux
Unix
Apply
$124k – $196k per year • In office • Full-Time • Master's Degree • Hillsboro
AI/ML
CUDA Toolkit
CUDA
Apply
$221k – $387k per year • In office • Full-Time • 15+ years exp • Santa Clara
Python
Java
AI/ML
AI Agents
Human-in-the-Loop
DevOps
AIOps
Management
ITSM
Apply
$167k – $291k per year • Remote (United States) • Full-Time • 8+ years exp • Santa Clara
Python
JavaScript
TypeScript
Frontend
Vue.js
Angular
React.js
DevOps
Kubernetes
Management
ServiceNow
Apply
≈ $178k – $324k per year (Estimated) • In office • Full-Time • Santa Clara
Python
Python
Asyncio
AI/ML
AI Agents
Machine Learning
DevOps
gRPC
OpenTelemetry
Apply
$191k – $334k per year • Remote (United States) • Full-Time • 10+ years exp • Santa Clara
Python
JavaScript
C#
Databases
MySQL
PostgreSQL
Neo4j
AI/ML
Cursor
Windsurf
Recommender Systems
Mobile
JUnit
DevOps
Rest API
Terraform
Puppet
Ansible
CI/CD
Jenkins
GitLab
IAM
Unix
Cybersecurity
FedRAMP
Management
ServiceNow
ITSM
QA
TestNG
Selenium
Playwright
Apply
$119k – $131k per year • In office • Full-Time • High School Diploma • Santa Clara
AI/ML
Transformers
Apply
See all jobs
This is one of many
1,350,029 more open roles from verified company boards, updated every day.