397,589open jobs
13,889companies
77,448added this week
Browse all
Salary
$152k – $242k per year
Location
In office (Santa Clara)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

At NVIDIA, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

As a Developer Technology Engineer, you will be at the forefront of innovation, working with leading industry partners and pioneering open-source projects to enable professional agentic AI workflows at the edge powered by NVIDIAs RTX and DGX platforms. This role offers an outstanding opportunity to collaborate with world-class talent and make a significant contribution to the evolving landscape of enterprise and consumer agentic AI.

What you'll be doing:

  • Work closely across internal engineering and product teams as well as external app developers and enterprise ISVs on solving local end-to-end agentic AI GPU deployment challenges on NVIDIA RTX & DGX.

  • Apply powerful profiling and debugging tools for analyzing most demanding accelerated end-to-end agentic AI workflows to detect insufficient system utilization resulting in suboptimal runtime performance.

  • Conduct hands-on trainings, develop sample code and host presentations to give good guidance on efficient end-to-end agentic AI deployment targeting optimal runtime performance.

  • Improve LLM & GenAI user experience by working on feature and performance enhancements of OSS software, including but not limited to projects like GGML, Llama.cpp, Ollama, vLLM, ONNX Runtime.

  • Collaborate with GPU driver and architecture teams as well as NVIDIA research to influence next generation GPU features by providing real-world workflows and giving feedback on partner and customer needs.

  • Providing technical leadership and mentorship to junior engineers, encouraging an inclusive and high-performing team environment.

What we need to see:

  • A proven track record 5+ years of professional experience in local GPU deployment, profiling and optimization.

  • A Bachelor's or Master's degree or equivalent experience in Computer Science, Engineering, or a related field.

  • Strong proficiency in C/C++, Python, software design, programming techniques..

  • Familiarity with and development experience on Windows and Linux.

  • Experience with CUDA and NVIDIA's Nsight GPU profiling and debugging suite.

  • Some travel is required for conferences and for on-site visits with external partners.

  • Strong problem-solving skills and the ability to work both independently and collaboratively in a fast-paced environment.

  • Excellent interpersonal and communication skills and a passion for keeping track with the latest advancements in AI technology.

Ways to stand out from the crowd:

  • Experience with GPU-accelerated AI inference driven by NVIDIA APIs and SDKs, specifically TensorRT-RTX, cuDNN, NVIDIA Model Optimizer.

  • Expertise with professional agentic AI use cases, i.e., digital content creation and productivity workflows.

  • Experience working with open-source LLM and GenAI software.

  • Detailed knowledge of the latest generation GPU architectures.

  • Experience with AI deployment on NPUs and ARM architectures.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative, autonomous and love a challenge, we want to hear from you.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 26, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
397,589 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$130k – $180k per year • Remote • 10+ years exp • Bachelor's Degree
C++
C
C++
LLVM
PyTorch C++
TensorFlow C++
C
MPI
AI/ML
CUDA
CUDA Toolkit
CUTLASS
DeepSpeed
JAX
MLIR
NCCL
PyTorch
ROCm
TensorFlow
TensorRT
Triton
vLLM
DevOps
AWS
Azure
GCP
HPC
Apply
$80k – $107k per year • Remote • 7+ years exp • Bachelor's Degree
C++
C++
LLVM
AI/ML
CUDA
CUDA Toolkit
CUTLASS
MLIR
NCCL
TensorRT
Triton
vLLM
Apply
In office • Full-Time • Xuzhou
C++
Java
Apply
$120k – $200k per year • Equity 0.5–1.5% • Remote • Full-Time • San Francisco
AI/ML
AI Agents
Claude
Claude Code
OpenAI Codex
Marketing
LinkedIn
YouTube
Apply
Algorithm Engineer 2 hours ago
Remote/Hybrid • Full-Time • 3+ years exp • Netanya
C++
MATLAB
Python
AI/ML
Computer Vision
Quantization
DevOps
CI/CD
HPC
Analytics
A/B Testing
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
Agentic Workflows
AI Agents
Function Calling
LangChain
LlamaIndex
LLM
OpenAI
Tool Use
DevOps
CI/CD
Git
Apply
In office • Internship • PhD • Shanghai
C++
Python
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenCL
DevOps
HPC
Apply
In office • Internship • Master's Degree • Shanghai
C#
C++
Python
AI/ML
AI Agents
Model Context Protocol
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
AI Agents
Apply
In office • Internship • PhD • Shanghai
AI/ML
AI Agents
RAG
Apply
$171k – $324k per year (Estimated) • Equity • In office • 15+ years exp • Master's Degree • Santa Clara
AI/ML
AI Agents
DevOps
AWS
Azure
GCP
Cybersecurity
Zero Trust
Marketing
Instagram
LinkedIn
Apply
$100k – $137k per year • Equity • In office • Full-Time • 3+ years exp • Master's Degree • Santa Clara
Chips/EDA
Cadence Allegro
OrCAD
Apply
$114k – $228k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Santa Clara
Marketing
X (Twitter)
Apply
$272k – $431k per year • In office • Full-Time • 15+ years exp • PhD • Santa Clara • New York
AI/ML
AI Agents
Fine-tuning
Function Calling
LLM
Multimodal AI
NVIDIA NeMo
Post-training
Pre-training
Reinforcement Learning
Structured Outputs
Synthetic Data
TGI
vLLM
DevOps
CI/CD
Git
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • PhD • Santa Clara
C++
Python
Apply
See all jobs
This is one of many
397,589 more open roles from verified company boards, updated every day.