740,080open jobs
44,473companies
105,813added this week
Browse all
Salary
$152k – $242k per year
Location
Remote/Hybrid (Santa Clara, Westford, Austin, Durham, United States)
Seniority
Senior · 5+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 24, 2026. First seen by Alion on Aug 29, 2026. NVIDIA scores A on the Alion truth index.

Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

As a member of the Hardware Infrastructure EDA Compute team, you will optimize, scale, and support workload scheduling systems that directly impact design velocity and infrastructure efficiency. Success in this role requires both operational precision along with developing and supporting forward-looking resource management solutions that address evolving compute demands. Beyond day-to-day operations, the role drives improvements in observability, service reliability, and automation, ensuring the EDA compute environment remains resilient, measurable, and aligned with long-term engineering demands.

What you'll be doing:

  • Manage, scale, and optimize job scheduling systems (LSF, Slurm, etc.) in a large-scale, multi-site environment supporting EDA and other compute-intensive workloads

  • Analyze scheduler and infrastructure performance data to identify systemic bottlenecks and drive measurable improvements in utilization, throughput, and turnaround time

  • Lead problem solving across scheduler, OS, and workload layers, ensuring timely resolution of service-impacting issues

  • Identify recurring operational challenges and implement targeted automation or process improvements to reduce manual effort and prevent repeat incidents

  • Help define and track reliable metrics and SLOs for service performance and reliability, partnering with customers to ensure expectations are realistic and measurable

  • Contribute to operational standards, documentation, and best practices to improve consistency across sites

  • Partner directly with customer teams to clarify requirements, translate technical tradeoffs, and drive issues to closure

What we need to see:

  • Bachelor’s degree in Computer Science or related field, or equivalent experience

  • Minimum 5+ years of experience operating and supporting large-scale Linux-based compute infrastructure

  • Strong hands-on experience supporting and tuning job scheduling systems (LSF, Slurm, etc.) in HPC or silicon design environments

  • Proficiency in Linux systems administration (CentOS/RHEL)

  • Strong problem solving skills and the ability to independently analyze complex system behavior under load

  • Clear and effective communication skills, including the ability to articulate technical tradeoffs and reliability metrics to engineering stakeholders

Ways to stand out from the crowd:

  • Experience implementing reliability engineering practices within HPC scheduling environments

  • Deep knowledge of job scheduling systems (LSF, Slurm, etc.) configuration tuning, scheduler internals, and advanced troubleshooting techniques

  • Experience building or enhancing observability systems, including metrics collection, monitoring pipelines, alerting strategies, and performance dashboards

  • Background with container technologies such as Docker, Singularity, or Podman in HPC environments

  • Experience influencing adoption of new infrastructure standards across multiple teams or sites

NVIDIA offers highly competitive salaries and a comprehensive benefits package. We have some of the most forward-thinking and hardworking people in the world on our team and our collaborative talent continues to drive NVIDIA's growth. We are seeking creative and independent engineers with real passion for technology!

#LI-Hybrid

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 23, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
740,080 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Industrial Engineering
Similar stack
Same company
Santa Clara
≈ $123k – $238k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Clay
Design
AutoCAD
Apply
$48k – $52k per year • In office • Full-Time • 12+ years exp • High School Diploma • Dunkirk
Management
Microsoft Office
Apply
≈ $66k – $133k per year (Estimated) • Remote (United States) • Full-Time • 4+ years exp • Bachelor's Degree • United States
Management
Microsoft Office
Apply
≈ $75k – $145k per year (Estimated) • In office • Full-Time • 3+ years exp • Associate's Degree • The Dalles
Apply
$84k – $126k per year • Equity • In office • Full-Time • 2+ years exp • High School Diploma • Brooklyn Center
Apply
≈ $98k – $185k per year (Estimated) • In office • Full-Time • Australia
JavaScript
Java
Kotlin
Node JS
Databases
Redis
Apache Kafka
AI/ML
Copilot
AI Agents
RAG
Frontend
React.js
DevOps
Rest API
Azure
CI/CD
AWS
Docker
Kubernetes
Management
Agile
Apply
$38k – $43k per year • In office • Bachelor's Degree • Naples
Python
JavaScript
Java
TypeScript
Java
Spring Boot
Databases
MongoDB
Apache Kafka
AI/ML
Spark
Frontend
Angular
React.js
DevOps
Rest API
Docker
Kubernetes
SOAP
Analytics
Informatica
Apply
≈ $151k – $301k per year (Estimated) • Equity • Remote (Poland) • Full-Time • 1+ year exp • Bachelor's Degree
AI/ML
Model Context Protocol
Vertex AI
AI Agents
AWS Bedrock
LLM
OpenAI
LLMOps
LLM Guardrails
DevOps
Terraform
GCP
Azure DevOps
GitHub Actions
CloudFormation
GitLab CI
Azure
CI/CD
Jenkins
AWS
Docker
Kubernetes
Platform Engineering
Bicep
FinOps
GitHub
GitLab
Web3
Abstract
Apply
$117k – $177k per year • In office • Full-Time • 3+ years exp • PhD • Washington
Python
Go
Java
SQL
AI/ML
Copilot
Cursor
Claude Code
Prompt Engineering
AI Agents
OpenAI Codex
Agentforce
DevOps
Rest API
Terraform
CloudFormation
CI/CD
Git
AWS
Docker
Kubernetes
IAM
Linux
Windows
Cybersecurity
OWASP Top 10
Zero Trust
PKI
Management
Agile
Apply
≈ $131k – $238k per year (Estimated) • Remote/Hybrid • Top Secret • 5+ years exp • Bachelor's Degree • Huntsville
Python
Java
SQL
C#
C++
C++
PyTorch C++
Databases
Databricks
RabbitMQ
Apache Kafka
AI/ML
OpenCV
MLFlow
Computer Vision
AI Agents
PyTorch
ClearML
Amazon SageMaker
Machine Learning
DevOps
CI/CD
AWS
Docker
Kubernetes
AWS Lambda
GitLab
Amazon S3
Apply
$127k – $190k per year • In office • Full-Time • PhD • Santa Clara
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • PhD • Santa Clara
Python
AI/ML
LLM
Apply
$136k – $219k per year • In office • Full-Time • 3+ years exp • PhD • Santa Clara
Python
Perl
Chips/EDA
Cadence Virtuoso
Apply
$132k – $207k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Austin
Python
AI/ML
InfiniBand
DevOps
HPC
Linux
Apply
$88k – $150k per year • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Santa Clara
Python
Apply
$127k – $190k per year • In office • Full-Time • PhD • Santa Clara
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • PhD • Santa Clara
Python
AI/ML
LLM
Apply
$182k – $273k per year • In office • Santa Clara
Management
Asana
Jira
Marketing
Marketo
Apply
$136k – $213k per year • In office • Full-Time • Bachelor's Degree • Santa Clara
Apply
≈ $53k – $111k per year (Estimated) • In office • 2+ years exp • High School Diploma • Santa Clara
Apply
See all jobs
This is one of many
740,080 more open roles from verified company boards, updated every day.