368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$140k – $224k per year
Location
In office (Santa Clara)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA is the world leader in GPU Computing. We are passionate about markets include gaming, automotive, vision, HPC, datacenters and networking in addition to our traditional OEM business. NVIDIA is also well positioned as the ‘AI Computing Company’, and NVIDIA GPUs are the brains powering Deep Learning software frameworks, analytics, data centers, and driving autonomous vehicles. We have some of the most experienced and dedicated people in the world working for us. If you are dedicated, forward-thinking, and hard-working technical people across countries sounds exciting, this job is for you. NVIDIA is looking for an outstanding individual who thrives in a diverse work environment, has outstanding interpersonal skills and possesses a strong sense of engagement and continuous process improvement. This candidate must have enterprise server integration, strong Linux experience, reliability testing with various telemetries, scale out cluster, test plan development, track record in developing AI tools and NLP, DevOps, CI/CD experience to join our platform SWQA team.

What you’ll be doing:

  • Responsible for the development and execution of NVIDIA HGX/DGX/MGX platform test plan on servers, OS, FW and CUDA SW stack from design doc.

  • Installing and testing various systems OS, server firmware and SW stack.

  • Drive support for root cause analysis on reliability and validation test failures to identify root cause(s) and achieve mitigation.

  • Build, develop/debug server and OS level automation front-end and back-end framework and tests

  • Review partner and supplier test results and prescribe additional reliability testing on components, servers, and packaging as needed.

  • Work in an agile software development team with very high production quality standards.

  • Manage bug lifecycle and collaborate with inter-groups to drive for solutions.

What we need to see:

  • Bachelor’s Degree (or equivalent experience) in a STEM (Science, Technology, Engineering, Math or Physics) field

  • 5+ years proven experience; or master’s degree.

  • Proven years of OS and server level automation, CI/CD process and DevOps experience using Python, SHELL, Ansible, Jenkins, C/C++, Java, JavaScript

  • Strong server and Linux(Ubuntu, RedHat, CentOS, SuSE, Fedora and etc…) troubleshooting and debugging experience in a bare-metal and KVM/VMWare/Hyper-V environment.

  • Good knowledge and hands-on experience in model testing, AI tools/frameworks (TensorFlow, Pytorch, Cursor and etc…), NLP and LLM benchmarking

  • Experience in using AI development tools for test plans creation, test cases development and test cases automation

  • Strong background in FW, BMC/OpenBMC, Network protocol, internal/external enterprise storage devices, PCIe buses and devices, IO sub-devices, CPU and memory, ACPI, UEFI spec, Redfish - huge plus

  • Proven years of experience in GitHub/Gitlab/Gerrit, PXE, SLURM, Stack/Kubernetes/Docker) - huge plus

Ways to stand out from the crowd:

  • AI related tools, LLM and NLP.

  • Experience working with NVIDIA GPU hardware is a strong plus.

  • Good to have solid understanding of virtualization in Linux (KVM, Docker orchestrated with Kubernetes)

  • Background in parallel programming ideally CUDA/OpenCL is a plus

    Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 140,000 USD - 224,250 USD for Level 3, and 168,000 USD - 270,250 USD for Level 4.

    You will also be eligible for equity and benefits.

    Applications for this job will be accepted at least until August 24, 2026.

    This posting is for an existing vacancy.

    NVIDIA uses AI tools in its recruiting processes.

    NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
    Free account
    Stop reading job ads. Get the ones that fit.
    One free account turns this page into a shortlist built around your stack, your level and your pay.
    Match on every job. Stack, seniority, pay and location, scored against your profile.
    368,634 open roles. Read straight off company career pages, refreshed every day.
    Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
    3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
    Create a free account
    Free forever. No card. Under a minute.

    Your match

    How well do you fit this role?
    Two answers are enough for a real match. No account needed.
    Check my fit
    Answers stay in this browser until you create an account.

    Recommended for you based on this role

    Similar stack
    Same company
    Santa Clara
    $17k – $43k per year (Estimated) • Remote • Moscow
    C++
    Java
    Python
    Java
    Maven
    DevOps
    Ansible
    CI/CD
    Docker
    Git
    Graylog
    HAProxy
    Jenkins
    Nginx
    Prometheus
    Zabbix
    Apply
    $40k – $86k per year (Estimated) • Remote/Hybrid • Full-Time • Élancourt
    Bash
    PowerShell
    Python
    SQL
    Databases
    MySQL
    PostgreSQL
    Redis
    DevOps
    Amazon CloudWatch
    Ansible
    AWS
    Azure
    CentOS Stream
    CI/CD
    Debian
    Docker
    GCP
    GitLab
    GitLab CI
    Grafana
    Hyper-V
    IAM
    Jenkins
    Kubernetes
    Nagios
    Prometheus
    Proxmox VE
    Terraform
    Ubuntu
    VMWare
    Windows Server
    Zabbix
    Apply
    $143k – $173k per year • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Wiesbaden
    Python
    DevOps
    Ansible
    Terraform
    Cybersecurity
    Defense in Depth
    Wireshark
    Zero Trust
    Apply
    $93k – $126k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • United States
    Node JS
    SQL
    TypeScript
    JavaScript
    Databases
    MS SQL
    Oracle
    Mobile
    JUnit
    DevOps
    AWS
    Azure
    CI/CD
    GCP
    Git
    GitLab
    GitLab CI
    Jenkins
    Management
    Jira
    QA
    JMeter
    Playwright
    Postman
    Rest-Assured
    TestNG
    Apply
    $195k – $264k per year • In office • Full-Time • 15+ years exp • Master's Degree • United States
    Python
    AI/ML
    Amazon SageMaker
    Keras
    Kubeflow
    MLFlow
    PyTorch
    Scikit-learn
    TensorFlow
    Vertex AI
    XGBoost
    DevOps
    AWS
    Azure
    CI/CD
    CloudFormation
    Docker
    GCP
    Kubernetes
    Terraform
    Cybersecurity
    FedRAMP
    NIST 800-53
    Apply
    $136k – $213k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara
    Perl
    Python
    Apply
    $98k – $252k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree • Switzerland
    Assembly
    C++
    Fortran
    C
    C
    MPI
    AI/ML
    CUDA
    CUDA Toolkit
    OpenMP
    DevOps
    HPC
    Apply
    In office • Full-Time • 5+ years exp • Bachelor's Degree • Hsinchu
    Perl
    Python
    Apply
    In office • Full-Time • 5+ years exp • Hsinchu • Taipei
    C++
    Python
    AI/ML
    InfiniBand
    Apply
    $156k – $348k per year (Estimated) • Remote • Full-Time • 10+ years exp • Bachelor's Degree • United Kingdom
    AI/ML
    CUDA
    CUDA Toolkit
    AI Agents
    NVIDIA NeMo
    Apply
    $72k – $99k per year • Equity • In office • Full-Time • Santa Clara
    Apply
    $166k – $290k per year • Equity • In office • Full-Time • 8+ years exp • Santa Clara
    Management
    ServiceNow
    Apply
    $133k – $272k per year (Estimated) • In office • Santa Clara
    Go
    Python
    AI/ML
    Edge AI
    LLM
    RAG
    Apply
    $80k – $110k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
    MATLAB
    Python
    Apply
    $142k – $256k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara • Toronto
    C++
    Go
    IoT
    MQTT
    OPC UA
    Apply
    See all jobs
    This is one of many
    368,634 more open roles from verified company boards, updated every day.