405,710open jobs
14,091companies
78,515added this week
Browse all
Salary
$140k – $224k per year
Location
In office (Santa Clara)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA is the world leader in GPU Computing. We are passionate about markets include gaming, automotive, vision, HPC, datacenters and networking in addition to our traditional OEM business. NVIDIA is also well positioned as the ‘AI Computing Company’, and NVIDIA GPUs are the brains powering Deep Learning software frameworks, analytics, data centers, and driving autonomous vehicles. We have some of the most experienced and dedicated people in the world working for us. If you are dedicated, forward-thinking, and hard-working technical people across countries sounds exciting, this job is for you. NVIDIA is looking for an outstanding individual who thrives in a diverse work environment, has outstanding interpersonal skills and possesses a strong sense of engagement and continuous process improvement. This candidate must have enterprise server integration, strong Linux experience, reliability testing with various telemetries, scale out cluster, test plan development, track record in developing AI tools and NLP, DevOps, CI/CD experience to join our platform SWQA team.

What you’ll be doing:

  • Responsible for the development and execution of NVIDIA HGX/DGX/MGX platform test plan on servers, OS, FW and CUDA SW stack from design doc.

  • Installing and testing various systems OS, server firmware and SW stack.

  • Drive support for root cause analysis on reliability and validation test failures to identify root cause(s) and achieve mitigation.

  • Build, develop/debug server and OS level automation front-end and back-end framework and tests

  • Review partner and supplier test results and prescribe additional reliability testing on components, servers, and packaging as needed.

  • Work in an agile software development team with very high production quality standards.

  • Manage bug lifecycle and collaborate with inter-groups to drive for solutions.

What we need to see:

  • Bachelor’s Degree (or equivalent experience) in a STEM (Science, Technology, Engineering, Math or Physics) field

  • 5+ years proven experience; or master’s degree.

  • Proven years of OS and server level automation, CI/CD process and DevOps experience using Python, SHELL, Ansible, Jenkins, C/C++, Java, JavaScript

  • Strong server and Linux(Ubuntu, RedHat, CentOS, SuSE, Fedora and etc…) troubleshooting and debugging experience in a bare-metal and KVM/VMWare/Hyper-V environment.

  • Good knowledge and hands-on experience in model testing, AI tools/frameworks (TensorFlow, Pytorch, Cursor and etc…), NLP and LLM benchmarking

  • Experience in using AI development tools for test plans creation, test cases development and test cases automation

  • Strong experience in FW, BMC/OpenBMC, Network protocol, internal/external enterprise storage devices, PCIe buses and devices, IO sub-devices, CPU and memory, ACPI, UEFI spec, Redfish - huge plus

  • Proven years of experience in GitHub/Gitlab/Gerrit, PXE, SLURM, Stack/Kubernetes/Docker) - huge plus

Ways to stand out from the crowd:

  • AI related tools, LLM and NLP.

  • Experience working with NVIDIA GPU hardware is a strong plus.

  • Good to have solid understanding of virtualization in Linux (KVM, Docker orchestrated with Kubernetes)

  • Background in parallel programming ideally CUDA/OpenCL is a plus

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 140,000 USD - 224,250 USD for Level 3, and 168,000 USD - 270,250 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 4, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
405,710 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$34k – $80k per year (Estimated) • In office • Full-Time • Bengaluru
PowerShell
Python
TypeScript
Databases
DynamoDB
DevOps
Amazon CloudWatch
Amazon EC2
Amazon S3
Ansible
AWS
AWS CDK
AWS Lambda
AWS Step Functions
CI/CD
CloudFormation
Configuration Management
GitHub
GitHub Actions
GitLab
Grafana
IAM
Prometheus
Splunk
Terraform
Apply
$109k – $218k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Malvern • Charlotte
JavaScript
TypeScript
Frontend
React.js
DevOps
CI/CD
Git
Rest API
Apply
$94k – $266k per year • In office • Full-Time • 12+ years exp • Associate's Degree • Columbus • Tampa • Dallas • Atlanta • Houston
AI/ML
AI Agents
Embeddings
LLM
Prompt Engineering
RAG
Frontend
GraphQL
DevOps
CI/CD
Docker
gRPC
Helm
Kubernetes
Platform Engineering
Vector
Apply
$47k – $121k per year (Estimated) • Remote • Full-Time • São Paulo • Vitoria-Gasteiz • Belo Horizonte • Lima • Porto Alegre
Python
SQL
Python
pySpark
Databases
BigQuery
Google BigQuery
MySQL
AI/ML
Airflow
Spark
DevOps
AWS
Azure
CI/CD
GCP
Analytics
ETL/ELT
Apply
Data Engineer 2 3 days ago
$72k – $93k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • Toronto
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Airflow
Spark
DevOps
AWS
Azure
CI/CD
GCP
Analytics
ETL/ELT
Apply
$44k – $104k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree • Bengaluru • Pune
Go
Java
Python
SQL
Node JS
JavaScript
Node JS
Commander.js
Databases
MySQL
PostgreSQL
AI/ML
Anomaly Detection
LLM
DevOps
AWS
AWS CDK
Azure
CI/CD
CloudFormation
Docker
GCP
Grafana
Incident Management
Kubernetes
OpenTelemetry
Platform Engineering
Prometheus
Self-Healing
Terraform
Apply
In office • Full-Time • 4+ years exp • Master's Degree • Shanghai
C++
Python
AI/ML
AI Agents
Copilot
Cursor
LLM
Reinforcement Learning
Robotics
Isaac Lab
Isaac Sim
Perception
Reinforcement Learning
ROS
Sim-to-Real
Teleoperation
Apply
$152k – $242k per year • In office • Full-Time • 4+ years exp • PhD • Santa Clara
C++
Go
Perl
Python
AI/ML
Structured Outputs
DevOps
CI/CD
Apply
$58k – $101k per year • Remote • Full-Time • Bachelor's Degree • Italy
AI/ML
InfiniBand
DevOps
HPC
Apply
$118k – $284k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Yokneam
C++
Python
AI/ML
AI Agents
Apply
$221k – $387k per year • Equity • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
DevOps
Incident Management
Management
ServiceNow
Apply
$56k – $94k per year • In office • Contractor • Santa Clara
Apply
$64k – $74k per year • In office • Contractor • Santa Clara
Apply
$56k per year • In office • Contractor • Santa Clara
Apply
$84k – $94k per year • In office • Contractor • Santa Clara
Apply
See all jobs
This is one of many
405,710 more open roles from verified company boards, updated every day.