368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$224k – $357k per year
Location
Remote/Hybrid (Santa Clara, Redmond, United States)
Seniority
Staff · 10+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

Artificial intelligence is moving from passive assistance to agents that can reason, use tools, and complete work on a person's device. NVIDIA is building the software foundation that helps developers and partners deliver these experiences privately, efficiently, and responsibly across GeForce RTX, NVIDIA RTX PRO, RTX Spark, DGX Spark, and DGX Station systems.

We are looking for an Engineering Manager to build and lead the team developing local and hybrid AI agent technology. You will set technical direction, grow engineers, and deliver production software spanning agent runtimes, local inference, Windows integration, developer tools, and security controls. You will turn lessons from Hermes, OpenClaw, Perplexity, and domain-specific agents into platform capabilities that work across NVIDIA systems.

What you'll be doing:

  • Lead the architecture and implementation of local AI agent runtimes on Windows and NVIDIA RTX systems, with clear interfaces across models, tools, GPU software, operating-system services, and applications.

  • Guide hands-on development through prototypes, design and code reviews, performance analysis, and debugging; help the team turn research ideas into reliable C++ and Python software.

  • Build core agent capabilities including MCP and tool routing, persistent memory, skills, scheduling, multi-agent coordination, computer use, and policy-aware local or hybrid model routing.

  • Own CI/CD, evaluation, compatibility, and regression systems across Windows releases, drivers, GPUs, models, and agent frameworks; add telemetry and diagnostics that make failures reproducible.

  • Partner with AI research, GPU and driver, Windows platform, product security, Microsoft, OEM, ISV, and open-source teams to land durable interfaces and production integrations.

  • Hire and develop engineers, set a high technical bar, allocate capacity, and create an environment where the team can make sound decisions and deliver maintainable software.

What we need to see:

  • Bachelor's degree in Computer Science, Computer Engineering, or a related field, or equivalent experience in practice.

  • 10+ years of software engineering experience, including 3+ years managing engineers or leading a comparable technical organization.

  • Experience hiring and developing engineers, setting goals, providing useful feedback, and building an inclusive team environment.

  • A record of delivering complex software platforms or systems from architecture through production operation.

  • Technical fluency in AI agents, model inference, systems software, and local or hybrid deployment, with the judgment to guide architecture and execution.

  • Ability to align engineering, research, product, security, and external partners around clear decisions and measurable outcomes.

  • Proficiency with C++ or Python.

Ways to stand out from the crowd:

  • Hands-on experience with Windows internals, process isolation, virtualization, containers, or sandboxing.

  • Familiarity with OpenClaw, Hermes, LangChain or similar agent frameworks; MCP; memory, skills, multi-agent, or computer-use patterns.

  • Experience with TensorRT-LLM, Ollama, llama.cpp, vLLM, PyTorch, ONNX Runtime, Windows ML, model optimization, or quantization.

  • Experience with CUDA, GPU performance, local models on consumer hardware, or privacy-aware inference routing.

  • Experience contributing to open-source projects or working with Microsoft, OEMs, ISVs, and developer communities.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD for Level 3, and 272,000 USD - 431,250 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 28, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$19k – $28k per year (net) • Remote • Full-Time • Moscow
C#
C++
C++
CMake
DevOps
CI/CD
Git
Management
Jira
Slack
Apply
$123k – $251k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Dallas • Denver • Birmingham
Java
SQL
Java
Gradle
Hibernate
Maven
Spring Boot
Spring Framework
Databases
Apache Kafka
MySQL
Redis
DevOps
CI/CD
Dynatrace
Jenkins
Kubernetes
OpenShift
Cybersecurity
SonarQube
Apply
$230k – $260k per year • Equity • Remote • Internship • Bachelor's Degree
Python
DevOps
Amazon EKS
AWS
Azure
CI/CD
GCP
Helm
Kubernetes
Terraform
Cybersecurity
FedRAMP
Orca Security
Apply
$67k – $160k per year (Estimated) • In office • Full-Time • France
C++
C++
PyTorch C++
TensorFlow C++
AI/ML
Computer Vision
Multimodal AI
OpenCV
PyTorch
TensorFlow
Edge AI
Apply
$19k – $51k per year (Estimated) • Remote/Hybrid • Full-Time • Moscow
C++
C
C++
STL
C
Valgrind
DevOps
RTOS
IoT
FreeRTOS
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara
Perl
Python
Apply
$98k – $252k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree • Switzerland
Assembly
C++
Fortran
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenMP
DevOps
HPC
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Hsinchu
Perl
Python
Apply
In office • Full-Time • 5+ years exp • Hsinchu • Taipei
C++
Python
AI/ML
InfiniBand
Apply
$156k – $348k per year (Estimated) • Remote • Full-Time • 10+ years exp • Bachelor's Degree • United Kingdom
AI/ML
CUDA
CUDA Toolkit
AI Agents
NVIDIA NeMo
Apply
$72k – $99k per year • Equity • In office • Full-Time • Santa Clara
Apply
$166k – $290k per year • Equity • In office • Full-Time • 8+ years exp • Santa Clara
Management
ServiceNow
Apply
$133k – $272k per year (Estimated) • In office • Santa Clara
Go
Python
AI/ML
Edge AI
LLM
RAG
Apply
$80k – $110k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
MATLAB
Python
Apply
$142k – $256k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara • Toronto
C++
Go
IoT
MQTT
OPC UA
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.