686,202open jobs
39,765companies
97,316added this week
Browse all
Salary
$184k – $288k per year
Location
In office (Santa Clara)
Seniority
Senior · 6+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

As a Senior DevOps Engineer, you will help lead the evolution of infrastructure operations within our Networking Software group. Building on a strong Linux systems administration foundation, you will build, automate, and operate scalable platforms that support networking software development and testing. This role offers an outstanding opportunity to work with elite technology and collaborate with ambitious engineers across global sites. If you are passionate about automation, reliability, technical leadership, and continuous improvement, this is the perfect opportunity for you!

What you'll be doing:

  • Build, provision, configure, and maintain scalable Linux infrastructure for networking feature creation and validation, including physical servers, network switches, virtualization platforms, containers, and remote-management interfaces.

  • Develop automation for infrastructure provisioning, configuration management, software deployment, upgrades, and day-to-day operations using infrastructure-as-code and configuration-management practices.

  • Build reusable tools and self-service capabilities that simplify infrastructure operations, improve engineering efficiency, and reduce repetitive manual work.

  • Diagnose and resolve complex issues spanning hardware, firmware, operating systems, virtualization, containers, storage, network communications, and application environments.

  • Implement monitoring, observability, capacity management, and reliability practices to improve infrastructure performance, availability, and operational readiness.

  • Partner with engineering, IT, facilities, security, and network teams to define technical standards, maintain documentation and runbooks, and establish scalable operational processes.

  • Provide technical leadership, guide infrastructure initiatives, and mentor team members in automation, troubleshooting, and operational guidelines.

What we need to see:

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or equivalent experience.

  • 6+ years of experience in systems engineering, DevOps, site reliability engineering, or infrastructure operations, including significant hands-on experience coordinating production or engineering Linux environments.

  • Experience working in the semiconductor industry or a hardware-focused engineering environment, with deep hands-on expertise in bare-metal Linux systems and server components-including CPUs, GPUs, memory, PCIe devices, NICs, storage, BIOS/UEFI, BMC/IPMI/Redfish, power, and cooling.

  • Proficiency in generative AI tools and skill in applying them effectively to automation, troubleshooting, documentation, operational analysis, and engineering efficiency.

  • Skilled at diagnosing complex issues across hardware, firmware, and operating-system layers.

  • Strong data-center networking knowledge, including TCP/IP, DNS, DHCP, VLANs, routing, switching, firewalls, and network troubleshooting tools.

  • Strong analytical, problem-solving, written communication, and cross-departmental collaboration skills, with the ability to guide technical initiatives to completion.

Ways to stand out from the crowd:

  • Experience managing Linux KVM/QEMU virtualization, Kubernetes clusters, multi-user engineering lab environments, NFS or distributed storage systems, and automated OS or cluster provisioning platforms.

  • Experience supporting fast-growing engineering labs, large-scale data-center environments, or globally distributed infrastructure.

  • Familiarity with observability platforms, including metrics, logging, tracing, alerting, incident management, and service-level objectives.

  • Experience crafting self-service infrastructure platforms and reusable automation that improves developer efficiency and theaAbility to establish clear, reliable, and scalable engineering and operational practices from evolving requirements.

  • Proven technical leadership, team leadership, or managerial experience, including mentoring engineers, prioritizing work, coordinating cross-functional initiatives, and driving projects from planning through completion.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 21, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
686,202 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$46k – $99k per year (Estimated) • Remote/Hybrid • Full-Time • Paris
DevOps
GCP
Helm
FluxCD
GitOps
Kubernetes
Apply
$34k – $91k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Vélizy-Villacoublay
DevOps
Terraform
GCP
Kubernetes
Google GKE
Cybersecurity
ISO 27001
Zero Trust
Apply
$45k – $97k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Amersfoort
Python
AI/ML
LLM
DevOps
Rest API
Terraform
Ansible
GCP
OpenShift
Helm
GitLab CI
CI/CD
AWS
Kubernetes
Platform Engineering
AIOps
Incident Management
SLI/SLO/SLA
GitLab
IAM
Cybersecurity
Zero Trust
Apply
$54k – $160k per year (Estimated) • In office • Full-Time • 1+ year exp • Bachelor's Degree • Yokneam • Tel Aviv
Python
Perl
AI/ML
Copilot
Claude
InfiniBand
NVLink
DevOps
VMWare
Git
Gerrit
KVM
Apply
$56k – $165k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Israel
Python
DevOps
GitLab CI
CI/CD
Jenkins
KVM
QA
Pytest
Apply
$100k – $167k per year • In office • Full-Time • Bachelor's Degree • Santa Clara
Python
Apply
$108k – $178k per year • In office • Full-Time • Bachelor's Degree • Santa Clara • Seattle
Python
C++
AI/ML
Flash Attention
AI Agents
LLM
MLIR
Apache TVM
Apply
$76k – $188k per year • In office • Internship • PhD • Santa Clara • Westford • Seattle
C++
AI/ML
CUDA Toolkit
CUDA
Apply
$76k – $188k per year • In office • Internship • PhD • Santa Clara
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
CUDA Toolkit
Reinforcement Learning
Multimodal AI
Diffusion Models
TensorFlow
PyTorch
CUDA
World Models
Embodied AI
Robotics
MuJoCo
Model Predictive Control
Imitation Learning
Reinforcement Learning
Apply
$53k – $137k per year (Estimated) • In office • Full-Time • 5+ years exp • Bengaluru
C++
AI/ML
InfiniBand
NVLink
DevOps
HPC
Apply
$94k – $184k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Chandler • Hillsboro • Folsom • Santa Clara
Analytics
Power BI
Microsoft Excel
Apply
$118k – $230k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara • Hillsboro
Apply
$58k – $173k per year (Estimated) • In office • Internship • Bachelor's Degree • Chicago • Tampa • New York • Portland • Irvine
AI/ML
LLM
Management
Outlook
Apply
$180k – $270k per year • In office • 3+ years exp • Bachelor's Degree • Santa Clara
Python
SQL
AI/ML
Dagster
Scikit-learn
PyTorch
Feature Store
DevOps
GCP
Azure
AWS
Apply
$180k – $270k per year • In office • 6+ years exp • Bachelor's Degree • Santa Clara
Python
Go
Java
C++
DevOps
gRPC
Prometheus
Docker
Platform Engineering
Apply
See all jobs
This is one of many
686,202 more open roles from verified company boards, updated every day.