1,283,717open jobs
74,382companies
213,626added this week
Browse all
Salary
$272k – $431k per year
Location
In office (Santa Clara)
Seniority
Principal · 15+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 6, 2026. First seen by Alion on Aug 5, 2026. NVIDIA scores A on the Alion truth index.

Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's (Infrastructure, Planning and Process) Cloud Infrastructure Team. IPP is a global organization within NVIDIA. This group works with various other groups within NVIDIA such as Graphics Processors, Mobile Processors, Deep Learning, Artificial Intelligence and Autonomous Vehicles to cater to their infrastructure needs. These cloud services provide almost half a million automated jobs per day on thousands of servers helping with the efficiency of thousands of NVIDIA's software engineers worldwide. The cloud hosts various machines and devices with operating systems like Windows, Linux, and Android. It supports hardware platforms including NVIDIA GPUs and Tegra Processors. It delivers unified CI/CD solutions and cloud-based software development. Are you passionate about distributed infrastructure and looking for sophisticated, critical issues, ready to build the next generation of cloud services, design creative solutions, mine through data to uncover real problems and fix them?

What you'll be doing:

  • Serve as an SRE Architect part of GPU Private Cloud team used by thousands of NVIDIANs globally for interactive development, centralized CI/CD, and QA testing.

  • Evaluating, identifying and developing software solutions to optimize critical software development workflows across various organizations within NVIDIA.

  • Architecting, implementing, and supporting end-to-end CI/CD system using open-source and NVIDIA proprietary software.

  • Customer (NVIDIA Internal development teams) onboarding to Private cloud infrastructure with a good discovery of the use case and available solutions within the cloud.

  • Identify performance bottlenecks and optimize the speed and cost efficiency of AI development and testing systems.

  • Leading software development projects and technically direct a team of brilliant engineers and guide them to provide efficient and impactful solutions.

  • Looking for problems within software systems and resolving the issues

  • Craft and implement critical metrics using various analytics methods and dashboards.

What we need to see:

  • BS or MS in Electrical Engineering, Computer Science, or relevant field (or equivalent experience).

  • 15+ years of systems software development including at least 1 year dedicated to developing/exploring AI.

  • Experience of maintaining cloud infrastructure and highly available production environment.

  • Strong programming and software development skills in JAVA, Python, Shell-script along with good understanding of distributed systems and REST APIs.

  • Experience in working with SQL/NoSQL database systems such as MySQL, Cassandra, MongoDB or Elasticsearch.

  • Excellent knowledge and working experience with Docker containers and Virtual Machines.

  • Good background of Cloud technologies like: OpenStack, Docker, Kubernetes, Chef/Puppet, Hadoop/Ceph/SwiftStack, LXC, Git, Perforce, JFrog, Kafka.

  • Ability to work across organizational boundaries effectively to improve alignment and productivity between teams in a multi-national, multi-time-zone corporate environment.

Ways to stand out from the crowd:

  • Depth in AI, Machine Learning and Deep Learning algorithms and techniques.

  • Strong collaborative and interpersonal skills, with a consistent record of guiding and influencing others in dynamic environments.

  • Experience developing large-scale software systems using modular architecture under real-time performance requirements.

  • Background in designing high-performance, scalable software systems with a strong focus on hardware cost optimization.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 9, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,283,717 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Santa Clara
≈ $88k – $182k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Newport Beach
DevOps
Azure DevOps
Azure
Linux
Windows
Management
Jira
ServiceNow
Agile
Scrum
ITSM
Apply
$105k – $162k per year • In office • Secret • Full-Time • 10+ years exp • Bachelor's Degree • Berkeley • Seattle
DevOps
GCP
OpenShift
Helm
Azure DevOps
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Platform Engineering
Configuration Management
Azure AKS
GitLab
Apply
$105k – $162k per year • In office • Secret • Full-Time • 10+ years exp • Bachelor's Degree • Berkeley • Seattle
PowerShell
DevOps
Terraform
Ansible
Azure
Bicep
DNS
Apply
$147k – $224k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • Moorpark
DevOps
Azure
CI/CD
AWS
Docker
GitHub
GitLab
Amazon ECS
Linux
Windows
Management
SharePoint
Apply
≈ $103k – $210k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Pittsburgh
Python
JavaScript
Java
TypeScript
Bash
Java
Maven
Spring Boot
Apache Tomcat
AI/ML
Edge AI
Frontend
Angular
npm
DevOps
Splunk
Terraform
Ansible
Helm
Datadog
GitLab CI
CI/CD
GitOps
Windows Server
Git
AWS
Docker
Kubernetes
Nginx
Apache HTTP Server
Amazon EKS
Amazon EC2
Incident Management
Amazon CloudWatch
Windows
Apply
≈ $15k – $40k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Izhevsk
Python
JavaScript
TypeScript
SQL
Databases
PostgreSQL
AI/ML
Embeddings
TensorFlow
PyTorch
LLM
RAG
Frontend
React.js
DevOps
Rest API
Git
Docker
Analytics
ETL/ELT
Apply
Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Gurgaon
Python
PHP
Databases
MySQL
Redis
DevOps
AWS
Docker
AWS Lambda
Amazon EC2
Management
Agile
Scrum
Apply
≈ $19k – $37k per year (Estimated) • Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • Gurgaon
Python
SQL
Databases
PostgreSQL
AI/ML
Machine Learning
DevOps
CI/CD
Jenkins
Git
AWS
Amazon S3
Analytics
ETL/ELT
AWS Glue
Management
Agile
Scrum
Apply
≈ $26k – $55k per year (Estimated) • Remote (likely EAEU) • 3+ years exp • Moscow
Python
JavaScript
C#
C#
.NET
Databases
MySQL
PostgreSQL
AI/ML
Claude Code
AI Agents
DevOps
Docker
Apply
≈ $15k – $26k per year (Estimated) • In office • 1+ year exp • Bachelor's Degree • Yekaterinburg
Python
SQL
Python
pySpark
Databases
DuckDB
AI/ML
Spark
NumPy
Analytics
Matplotlib
QlikSense
Plotly
Management
Telegram
Apply
$184k – $288k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Santa Clara
AI/ML
LLM
Mixture of Experts
NVLink
DevOps
HPC
Apply
$184k – $288k per year • In office • Full-Time • 6+ years exp • Bachelor's Degree • Santa Clara • Redmond
Python
Rust
AI/ML
Fine-tuning
Prompt Engineering
Multimodal AI
AI Agents
NLP
VLM
LLM
DevOps
CI/CD
Docker
Kubernetes
Apply
$184k – $288k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Santa Clara
Python
AI/ML
vLLM
CUDA Toolkit
AI Agents
SGLang
TensorRT
TensorRT-LLM
PyTorch
CUDA
NCCL
DevOps
Terraform
GCP
Azure
CI/CD
GitOps
ArgoCD
AWS
Kubernetes
Self-Healing
Linux
Apply
$208k – $334k per year • In office • Full-Time • 12+ years exp • Master's Degree • Santa Clara
Python
Ruby
DevOps
HPC
DNS
DHCP
VPN
BGP
MPLS
Apply
$92k – $155k per year • In office • Full-Time • Bachelor's Degree • Santa Clara
DevOps
Linux
Windows
Apply
$152k – $242k per year • In office • Full-Time • 5+ years exp • Master's Degree • Santa Clara
Python
C
C++
C
Pthreads
MPI
AI/ML
CUDA Toolkit
AI Agents
OpenMP
CUDA
DevOps
CI/CD
HPC
Apply
Lead Product Manager 6 hours ago
$213k – $266k per year • Hybrid • 7+ years exp • Plano • Santa Clara • Boise
Cybersecurity
Zero Trust
Apply
≈ $99k – $191k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Los Angeles • San Francisco • Pasadena • San Mateo • Santa Clara
Apply
$110k – $155k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Chicago • Pasadena • Fresno • Santa Rosa • San Francisco
Analytics
Microsoft Excel
Management
Outlook
Apply
$116k – $192k per year • Equity • Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara
Databases
Snowflake
AI/ML
Copilot
Claude
Analytics
Power BI
Microsoft Excel
Management
ServiceNow
Outlook
Apply
See all jobs
This is one of many
1,283,717 more open roles from verified company boards, updated every day.