399,808open jobs
13,937companies
77,696added this week
Browse all
Salary
$184k – $288k per year
Location
In office (Santa Clara, Austin, Hillsboro, Durham, United States, Redmond)
Seniority
Senior · 6+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology-and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our tightly coupled CPU, GPU and DPU technology acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world!

NVIDIA is searching for a highly motivated, technical engineer to join the Tegra system-on-chip (SoC) software organization. You will work on key aspects of our ARM SW ecosystem and system software architecture. With a targeted charter to enable best-in-class datacenter-scale performance and efficiency for our next generation of datacenter products, including CPUs and CPU+GPU Superchips.

What you will be doing:

  • Design, develop, test, and optimize software for our next-generation SoCs. In both pre-silicon and post-silicon phases of execution.

  • Review architectural performance bottlenecks for various system wide work loads. Identify HW/SW policies to drive performance and performance/watt leadership.

  • Using strong communication skills, build and drive architecture, analysis documents and communications to internal and/or external audiences about our technology.

  • Competitive analysis comparing uArchitecture & workload performance metrics on NVIDIA's ARM SoCs against emerging processors from other silicon vendors.

  • Influence and drive full-stack adoption of performance optimizations and best practices across NVIDIA SW products & OSS SDKs

What we need to see:

  • BS or MS degree in Computer Engineering, Computer Science, or related degree (or equivalent experience).

  • 6+ years of relevant computer architecture or SW development experience.

  • Proven leadership skills and strong ownership on past projects.

  • Hands on technical experience and demonstrated excellence in an environment with complex software and hardware designs.

  • Strong understanding of multicore hardware, operating systems design, concurrency, virtual memory, caching, interrupts, device drivers and real-time programming.

  • Strong stills in performance analysis, data analysis and performance optimization.

  • Strong use of linux perf tool, application/library performance optimizations.

Ways to stand out from the crowd:

  • Deep expertise in ARM architecture and SW ecosystem.

  • Proficient in analyzing, debugging and tuning performance of complex system software stacks.

  • Experience with CPU server system workloads and performance analysis.

  • Familiarity with CUDA programming and/or GPUs.

  • Experience with HPC or large-scale computing environments.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until June 30, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
399,808 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$130k – $180k per year • Remote • 10+ years exp • Bachelor's Degree
C++
C
C++
LLVM
PyTorch C++
TensorFlow C++
C
MPI
AI/ML
CUDA
CUDA Toolkit
CUTLASS
DeepSpeed
JAX
MLIR
NCCL
PyTorch
ROCm
TensorFlow
TensorRT
Triton
vLLM
DevOps
AWS
Azure
GCP
HPC
Apply
$130k – $180k per year • Remote • 10+ years exp • Bachelor's Degree
C++
Python
C++
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
CUTLASS
DeepSpeed
KV Cache
LLM
NCCL
PyTorch
Quantization
Ray
ROCm
Speculative Decoding
TensorBoard
TensorRT
TensorRT-LLM
Triton
Triton Inference Server
vLLM
DevOps
AWS
Azure
FinOps
GCP
HPC
Apply
$100k – $150k per year • Remote • 6+ years exp • Bachelor's Degree
C++
C++
LLVM
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
CUTLASS
JAX
MLIR
NCCL
PyTorch
TensorRT
Triton
vLLM
DevOps
HPC
Apply
$80k – $107k per year • Remote • 7+ years exp • Bachelor's Degree
C++
C++
LLVM
AI/ML
CUDA
CUDA Toolkit
CUTLASS
MLIR
NCCL
TensorRT
Triton
vLLM
Apply
$85k – $110k per year • Remote • 6+ years exp • Bachelor's Degree
C++
C++
LLVM
AI/ML
CUDA
CUDA Toolkit
CUTLASS
MLIR
NCCL
TensorRT
Triton
vLLM
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
Agentic Workflows
AI Agents
Function Calling
LangChain
LlamaIndex
LLM
OpenAI
Tool Use
DevOps
CI/CD
Git
Apply
In office • Internship • PhD • Shanghai
C++
Python
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenCL
DevOps
HPC
Apply
In office • Internship • Master's Degree • Shanghai
C#
C++
Python
AI/ML
AI Agents
Model Context Protocol
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
AI Agents
Apply
In office • Internship • PhD • Shanghai
AI/ML
AI Agents
RAG
Apply
$171k – $324k per year (Estimated) • Equity • In office • 15+ years exp • Master's Degree • Santa Clara
AI/ML
AI Agents
DevOps
AWS
Azure
GCP
Cybersecurity
Zero Trust
Marketing
Instagram
LinkedIn
Apply
$100k – $137k per year • Equity • In office • Full-Time • 3+ years exp • Master's Degree • Santa Clara
Chips/EDA
Cadence Allegro
OrCAD
Apply
$114k – $228k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Santa Clara
Marketing
X (Twitter)
Apply
$272k – $431k per year • In office • Full-Time • 15+ years exp • PhD • Santa Clara • New York
AI/ML
AI Agents
Fine-tuning
Function Calling
LLM
Multimodal AI
NVIDIA NeMo
Post-training
Pre-training
Reinforcement Learning
Structured Outputs
Synthetic Data
TGI
vLLM
DevOps
CI/CD
Git
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • PhD • Santa Clara
C++
Python
Apply
See all jobs
This is one of many
399,808 more open roles from verified company boards, updated every day.