394,870open jobs
13,845companies
76,904added this week
Browse all
Salary
$168k – $270k per year
Location
In office (Santa Clara)
Seniority
Senior · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

We are seeking a highly skilled and hard-working Senior Test Developer / test engineer to join our multifaceted Enterprise Software QA team. This role offers an outstanding opportunity to leave your mark on the design, construction, optimization and testing of large-scale infrastructure for various foundational NVIDIA unified cloud services and data center offerings. If you are a dedicated engineer with strong expertise in cloud infrastructure and distributed systems and want to apply your skills with AI tools, this role could fit you perfectly. You will thrive in an exciting, innovative environment.

What you'll be doing:

  • Work with development teams on test plans for all layers of SW stack for cloud infrastructure, execution, reviews, failure analysis and assessing overall quality and risk. Work with customer PMs on software issues including technical feedback from OEMs and CSPs. Develop key benchmarks to track execution and deploy process improvements to improve efficiency

  • Leverage AI skills to expedite the test scope, test plan, execution and automation workflows.

  • Lead NVIDIA Cloud and Data Center bring up activities which will involve validation, reporting, working with engineering to debug issues, providing design input at times, adding coverage in different areas.

  • Design, develop and maintain CI/CD pipelines for continuous testing in cloud environments when needed.

  • Perform performance, scalability, and reliability testing of cloud services.

  • Implement and maintain test environments in cloud platforms such as AWS, Azure, or Google Cloud.

  • Supervise the infrastructure to alert on significant events, ensuring the highest level of system performance and reliability.

  • Work with various different partner teams to ensure availability of clusters to test on and take the lead in resolve all issues.

  • Working with teams to ensure quality of the cloud products getting delivered focusing on critical areas like security, storage, workloads, performance on latest SW and FW components.

What we need to see:

  • A Master's or Ph.D. in Computer Science or a related field, or equivalent experience.

  • Experience with AI development tools used in creating test cases, automating test cases, code coverage, triaging.

  • 8+ years of hands-on experience in cluster management and related tools, including Docker Containers, Slurm, Kubernetes, and Ansible.

  • 2+ years strong experience with cloud infrastructure platforms like AWS, Azure, Google, OCI Cloud.

  • Hands-on experience with network, storage, security, cluster configuration and debugging, cloud infrastructure management tools like terraform, ansible.

  • Expertise in administering, operating, and configuring Kubernetes.

  • Experience in CI/CD tools such as Gitlab and Jenkins and the GitOps model.

  • Proficiency in various monitoring tools :Prometheus, Grafana, Cloudwatch, and Thanos.

  • Proficiency in debugging issues involving networks, DHCP, DNS, HTTP, Linux, and containers.

Ways to Stand Out from the Crowd:

  • Familiarity with "Base Command Manager" for managing and monitoring high performance computing.

  • Experience in writing automation for web application using tools like selenium, playwright.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 270,250 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 4, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
394,870 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
In office • 8+ years exp • PhD
Python
SQL
Databases
Databricks
Snowflake
AI/ML
Dagster
Feature Store
Google AI Studio
MLFlow
Spark
DevOps
AWS
Azure
Azure AKS
Azure DevOps
CI/CD
Docker
GCP
GitHub
GitHub Actions
Kubernetes
Platform Engineering
Apply
$14k – $38k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Hyderabad
JavaScript
Python
SQL
TypeScript
Frontend
npm
React.js
DevOps
AWS
Azure
CI/CD
Docker
Git
Grafana
Kubernetes
Prometheus
Rest API
QA
Cypress
Jest
Apply
$24k – $59k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Bengaluru
DevOps
Azure
Azure DevOps
CI/CD
GitHub
IAM
Management
Confluence
Jira
Apply
In office • 6+ years exp • Master's Degree
C++
Python
SQL
C++
PyTorch C++
TensorFlow C++
AI/ML
AI Agents
Google AI Studio
PyTorch
TensorFlow
DevOps
AWS
Azure
GCP
Apply
In office
Databases
Databricks
Snowflake
AI/ML
AI Agents
DevOps
AWS
Azure
Apply
$272k – $431k per year • In office • Full-Time • 15+ years exp • PhD • Santa Clara • New York
AI/ML
AI Agents
Fine-tuning
Function Calling
LLM
Multimodal AI
NVIDIA NeMo
Post-training
Pre-training
Reinforcement Learning
Structured Outputs
Synthetic Data
TGI
vLLM
DevOps
CI/CD
Git
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • PhD • Santa Clara
C++
Python
Apply
$144k – $230k per year • In office • Full-Time • PhD • Santa Clara
Apex
Apex
MuleSoft
Apply
In office • Internship • Master's Degree • Beijing • Shanghai • Shenzhen
C++
AI/ML
CUDA
CUDA Toolkit
Speech Recognition
Apply
In office • Full-Time • 5+ years exp • Master's Degree • Shanghai
C++
Python
AI/ML
AI Agents
Copilot
Cursor
LLM
Reinforcement Learning
Robotics
Isaac Lab
Isaac Sim
Perception
Reinforcement Learning
ROS
Sim-to-Real
Teleoperation
Apply
$171k – $324k per year (Estimated) • Equity • In office • 15+ years exp • Master's Degree • Santa Clara
AI/ML
AI Agents
DevOps
AWS
Azure
GCP
Cybersecurity
Zero Trust
Marketing
Instagram
LinkedIn
Apply
$100k – $137k per year • Equity • In office • Full-Time • 3+ years exp • Master's Degree • Santa Clara
Chips/EDA
Cadence Allegro
OrCAD
Apply
$114k – $228k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Santa Clara
Marketing
X (Twitter)
Apply
$272k – $431k per year • In office • Full-Time • 15+ years exp • PhD • Santa Clara • New York
AI/ML
AI Agents
Fine-tuning
Function Calling
LLM
Multimodal AI
NVIDIA NeMo
Post-training
Pre-training
Reinforcement Learning
Structured Outputs
Synthetic Data
TGI
vLLM
DevOps
CI/CD
Git
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • PhD • Santa Clara
C++
Python
Apply
See all jobs
This is one of many
394,870 more open roles from verified company boards, updated every day.