411,854open jobs
14,107companies
70,575added this week
Browse all
Salary
$191k – $373k per year (Estimated)
Location
In office (Santa Clara, United States)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

We are now looking for a Senior Machine Learning Applications and Compiler Engineer!

NVIDIA is seeking engineers to develop algorithms and optimizations for our LPX inference and compiler stack. You will work at the intersection of large-scale systems, compilers, and deep learning, crafting how neural network workloads map onto future NVIDIA platforms. This is your chance to be part of something outstandingly innovative!

What you’ll be doing:

  • Build, develop, and maintain high-performance runtime and compiler components, focusing on end-to-end inference optimization.

  • Define and implement mappings of large-scale inference workloads onto NVIDIA’s systems.

  • Extend and integrate with NVIDIA’s SW ecosystem, contributing to libraries, tooling, and interfaces that enable seamless deployment of models across platforms.

  • Benchmark, profile, and monitor key performance and efficiency metrics to ensure the compiler generates efficient mappings of neural network graphs to our inference hardware.

  • Collaborate closely with hardware architects and design teams to feedback software observations, influence future architectures, and codesign features that unlock new performance and efficiency points.

  • Prototype and evaluate new compilation and runtime techniques, including graph transformations, scheduling strategies, and memory/layout optimizations tailored to spatial processors.

  • Publish and present technical work on novel compilation approaches for inference and related spatial accelerators at top tier ML, compiler, and computer architecture venues.

What we need to see:

  • MS or PhD in Computer Science, Electrical/Computer Engineering, or related field, or equivalent experience, with 5 years of relevant experience.

  • Strong software engineering background with proficiency in systems level programming (e.g., C/C++ and/or Rust) and solid CS fundamentals in data structures, algorithms, and concurrency.

  • Hands on experience with compiler or runtime development, including IR design, optimization passes, or code generation.

  • Experience with LLVM and/or MLIR, including building custom passes, dialects, or integrations.

  • Familiarity with deep learning frameworks such as TensorFlow and PyTorch, and experience working with portable graph formats such as ONNX.

  • Solid understanding of parallel and heterogeneous compute architectures, such as GPUs, spatial accelerators, or other domain specific processors.

  • Strong analytical and debugging skills, with experience using profiling, tracing, and benchmarking tools to drive performance improvements.

  • Excellent communication and collaboration skills, with the ability to work across hardware, systems, and software teams.

  • Ideal candidates will have direct experience with MLIR based compilers or other multilevel IR stacks, especially in the context of graph based deep learning workloads.

Ways to stand out from the crowd:

  • Prior work on spatial or dataflow architectures, including static scheduling, pipeline parallelism, or tensor parallelism at scale.

  • Contributions to opensource ML frameworks, compilers, or runtime systems, particularly in areas related to performance or scalability.

  • Demonstrated research impact, such as publications or presentations at conferences like PLDI, CGO, ASPLOS, ISCA, MICRO, MLSys, NeurIPS, or similar.

  • Experience with large-scale AI distributed inference or training systems, including performance modeling and capacity planning for multi rack deployments.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 17, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
411,854 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$91k – $135k per year • In office • Bachelor's Degree • Menlo Park
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
OpenCV
Computer Vision
TensorFlow
PyTorch
DevOps
Docker
Game Dev
Unreal Engine
OpenXR
Robotics
ROS
Gazebo
Isaac Sim
Localization
Sensor Fusion
Teleoperation
Digital Twin
Design
SolidWorks
Fusion 360
Apply
$114k – $282k per year (Estimated) • In office • Full-Time • 5+ years exp • Master's Degree • Tel Aviv
Python
AI/ML
LangChain
AI Agents
PyTorch
LLM
RAG
Hugging Face
Agentforce
Agentic Workflows
Cybersecurity
OWASP Top 10
Apply
$72k – $78k per year • Remote/Hybrid • 2+ years exp • Bachelor's Degree • San Diego
Python
JavaScript
SQL
C++
Apply
$73k – $161k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Singapore
Python
C++
AI/ML
Copilot
DevOps
RTOS
CI/CD
Git
GitHub
Apply
$71k – $151k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Bristol
Python
C++
Cybersecurity
Metasploit
Cyber Kill Chain
Apply
$13k – $22k per year (Estimated) • In office • Internship • Master's Degree • Shanghai
Python
Perl
AI/ML
LangChain
LlamaIndex
Function Calling
AI Agents
LLM
OpenAI
Agentic Workflows
Tool Use
DevOps
CI/CD
Git
Apply
$14k – $23k per year (Estimated) • In office • Internship • PhD • Shanghai
Python
C
C++
C
MPI
AI/ML
CUDA Toolkit
OpenCL
CUDA
DevOps
HPC
Apply
$13k – $22k per year (Estimated) • In office • Internship • Master's Degree • Shanghai
Python
C#
C++
AI/ML
Model Context Protocol
AI Agents
Apply
In office • Internship • Master's Degree • Shanghai
Python
Perl
AI/ML
AI Agents
Apply
$13k – $21k per year (Estimated) • In office • Internship • PhD • Shanghai
AI/ML
AI Agents
RAG
Apply
$154k – $313k per year • Equity • In office • 15+ years exp • Master's Degree • Santa Clara
AI/ML
AI Agents
DevOps
GCP
Azure
AWS
Cybersecurity
Zero Trust
Apply
$124k – $155k per year • Remote/Hybrid • 5+ years exp • Bachelor's Degree • Santa Clara
Python
DevOps
VMWare
OpenStack
Cybersecurity
Wireshark
Management
Jira
QA
TestRail
Apply
$208k – $260k per year • Remote/Hybrid • Santa Clara
AI/ML
AI Agents
Apply
$184k – $230k per year • Remote/Hybrid • 10+ years exp • Bachelor's Degree • Santa Clara
Verilog
Apply
$140k – $175k per year • Remote/Hybrid • 8+ years exp • Bachelor's Degree • Santa Clara
DevOps
VMWare
Azure
AWS
Apply
See all jobs
This is one of many
411,854 more open roles from verified company boards, updated every day.