368,941open jobs
9,452companies
47,951added this week
Browse all
Salary
$106k – $266k per year (Estimated)
Location
In office (Yokneam)
Seniority
Senior · 12+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s an outstanding legacy of innovation that’s fueled by phenomenal technology - and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

We are seeking a Senior Site Reliability Engineer - Storage, you will own the reliability, performance, and scalability of our global NAS, SAN, and Object Storage platforms that power critical internal and external services. You will combine deep storage expertise with strong automation and SRE practices to design, build, and operate highly available storage systems at scale.

What You Will Be Doing:

  • Lead design, deployment, and operations of production NAS, SAN, and Object Storage platforms, ensuring reliability, performance, and security.

  • Capture requirements from partner teams, architect storage solutions, and drive end-to-end implementation for new and existing services.

  • Develop, maintain, and improve automation for provisioning, configuration, monitoring, incident response, and lifecycle management of storage infrastructure.

  • Participate in on-call and incident response, lead troubleshooting of complex storage and performance issues, and drive root cause analysis and preventive actions.

  • Define and track SLOs/SLIs and error budgets for storage services, using observability and analytics to continuously improve reliability and efficiency.

  • Build and maintain runbooks, standard operating procedures, and comprehensive documentation for storage services and automation.

  • Analyze capacity and usage trends, perform forecasting, and recommend scaling or optimization strategies to support business growth.

  • Collaborate closely with SRE, infrastructure, networking, and application teams in a follow-the-sun model to deliver consistent, high-quality service.

  • Mentor junior engineers, share best practices, and help drive adoption of SRE principles across the team.

What We Need to See:

  • 12+ years of experience in Site Reliability, DevOps, or Infrastructure Engineering, with significant focus on storage systems.

  • Bachelor’s degree in Computer Science, Computer Engineering, or a related technical field or equivalent practical experience.

  • Strong hands-on experience with design, deployment, and operations of enterprise-grade NAS, SAN, and/or Object Storage platforms.

  • Solid understanding of SRE concepts (SLOs/SLIs, error budgets, incident management, observability, postmortems).

  • Proficiency with Infrastructure as Code and configuration management tools (e.g., Terraform, Ansible, Puppet, SaltStack) and source control systems.

  • Experience building and operating highly available, scalable infrastructure, including automation for provisioning, monitoring, and remediation.

  • Experience with container and virtualization platforms (e.g., Docker, Kubernetes, hypervisors) and modern CI/CD and version control tools.

  • Strong scripting or programming skills (e.g., Python, Go, Shell) to build tools, automate workflows, and integrate systems.

  • Excellent communication and collaboration skills, with the ability to work effectively across distributed and cross-functional teams.

Ways to Stand Out from the Crowd:

  • Experience with storage for high-performance computing, AI/ML workloads, or large-scale data analytics.

  • Proven ability to debug complex, distributed systems and storage performance issues.

  • History of driving reliability improvements through data-driven analysis and automation.

  • Experience leading technical initiatives, mentoring engineers, or acting as a technical lead on critical projects.

With competitive salaries and a generous benefits package, NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most brilliant and talented people in the world working for us. If you're creative and motivated, we want to hear from you!

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,941 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Yokneam
$105k – $252k per year • Remote • Full-Time • 18+ years exp • Bachelor's Degree
Python
Java
Java
Gradle
DevOps
Ansible
AWS
CI/CD
CloudFormation
Configuration Management
Docker
GitHub Actions
GitLab CI
Helm
Jenkins
Kubernetes
Platform Engineering
Terraform
GitHub
GitLab
Cybersecurity
Sonatype Nexus IQ
Management
Confluence
Jira
Apply
$54k – $175k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
TypeScript
Frontend
React.js
Sass
DevOps
AWS
CI/CD
Docker
Kubernetes
Rest API
Apply
$35k – $113k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
TypeScript
Frontend
React.js
Sass
DevOps
AWS
CI/CD
Docker
Kubernetes
Rest API
Apply
$133k – $161k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Westminster
C++
Python
DevOps
CI/CD
SpaceTech
NASA cFS
Apply
In office • Part-Time • 2+ years exp • Bachelor's Degree • Ness Ziona
C#
Java
AI/ML
Copilot
Cursor
DevOps
CI/CD
GitHub
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara
Perl
Python
Apply
$98k – $252k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree • Switzerland
Assembly
C++
Fortran
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenMP
DevOps
HPC
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Hsinchu
Perl
Python
Apply
In office • Full-Time • 5+ years exp • Hsinchu • Taipei
C++
Python
AI/ML
InfiniBand
Apply
$156k – $348k per year (Estimated) • Remote • Full-Time • 10+ years exp • Bachelor's Degree • United Kingdom
AI/ML
CUDA
CUDA Toolkit
AI Agents
NVIDIA NeMo
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Yokneam • Tel Aviv
Verilog
VHDL
AI/ML
ChatGPT
Apply
$103k – $260k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Yokneam
Go
Python
Databases
MS SQL
MySQL
Oracle
DevOps
Blue-Green Deployment
CI/CD
Kubernetes
Vector
Apply
$93k – $274k per year (Estimated) • In office • Full-Time • 1+ year exp • Bachelor's Degree • Tel Aviv • Yokneam
C++
AI/ML
Edge AI
NCCL
DevOps
Kubernetes
Apply
$104k – $294k per year (Estimated) • In office • Full-Time • 5+ years exp • Yokneam
DevOps
CI/CD
Apply
$165k – $339k per year (Estimated) • In office • Full-Time • 12+ years exp • Bachelor's Degree • Yokneam • Tel Aviv
Apply
See all jobs
This is one of many
368,941 more open roles from verified company boards, updated every day.