394,870open jobs
13,845companies
76,904added this week
Browse all
Salary
$125k – $278k per year (Estimated)
Location
Remote (Germany)
Seniority
Senior · 4+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

JAX is the framework of choice for the most ambitious AI research, and the datacenters it runs on are evolving fast. NVIDIA supercomputers are becoming more heterogeneous combining GPUs, CPUs, and LPUs, and JAX needs to evolve to perform well across all of them. NVIDIA seeks a Senior Engineering Manager to define and drive NVIDIA's JAX strategy, coordinating multiple teams and workstreams to ensure JAX delivers peak performance across the platform, from single-accelerator workloads to hundreds of thousands of devices on the largest supercomputers ever built.

A critical part of this role is ensuring JAX is ready on next-generation architectures, including Rubin Ultra, LPUs, and the CPUs that drive them, from day one. Doing that well requires keeping up with where AI is heading. This means supporting emerging needs across training, post-training, inference, and robotics, bridging new hardware capabilities with AI trends. Come join us to build the team and technology behind the next AI breakthroughs.

What you will be doing:

  • Drive the engineering contribution strategy across the JAX ecosystem stack - including JAX core, XLA, Shardy, MPMD, and application-level codebases like MaxText - prioritizing where NVIDIA's contributions create maximum impact on performance, adoption, and ecosystem influence

  • Promote teamwork across organizational boundaries to shape NVIDIA's AI platform across compiler (XLA, TileIR, CUDA), runtime, networking, and hardware teams

  • Build deep partnerships with key open-source projects, especially Google's JAX team - supporting our partners while representing NVIDIA's interests in shaping the project's direction

  • Design processes and define clear success criteria to keep teams aligned to outcomes

  • Build staffing plans, review team capacity, and make strategic hiring decisions to stay ahead of current and future needs

  • Lead, grow, and mentor a high-performing engineering organization - developing technical leaders and creating space for your team to experiment, prototype, and build innovative systems that improve products and processes

What we need to see:

  • MS or PhD in Computer Science, Computer Engineering, or a related field (or equivalent experience)

  • 10+ overall years of software engineering experience, with 4+ years in engineering management or technical leadership

  • Strong technical background in system software, compilers, high-performance computing, or AI frameworks

  • Proven experience leading domain-level strategy and projects that span organizational boundaries, building consensus, influencing without direct authority, and tailoring communication to the audience from engineers to executives

  • Track record of owning outcomes across multiple projects and teams, balancing trade-offs, and driving progress with incomplete information

  • Experience identifying development cycle gaps and defining workflows across teams and organizations to improve efficiency

  • Experience leading teams that contribute to large-scale projects with external collaborators

Ways to stand out from the crowd:

  • Hands-on experience with JAX and its systems like XLA, HLO, PJRT, Shardy, or JAX's tracing and transformation system

  • Experience leading teams working on compiler or runtime infrastructure for deep learning frameworks (JAX, PyTorch, TensorFlow or equivalent)

  • Familiarity with NVIDIA's AI platform components: GPU programming, CUDA, cuDNN, NCCL, TensorRT, Transformer Engine and heterogeneous compute targets (GPU, CPU, LPU)

  • Experience driving performance optimization efforts for large-scale distributed AI training, post-training (RLHF, alignment), inference, or robotics workloads (multi-node, multi-GPU)

  • Experience driving framework or software bring-up on new hardware architectures, coordinating across silicon, driver, compiler, and framework teams

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
394,870 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Munich
$100k – $150k per year • Remote • 6+ years exp • Bachelor's Degree
C++
C++
LLVM
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
CUTLASS
JAX
MLIR
NCCL
PyTorch
TensorRT
Triton
vLLM
DevOps
HPC
Apply
$100k – $175k per year • Remote • 6+ years exp • Bachelor's Degree
C++
C++
LLVM
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
CUTLASS
JAX
MLIR
NCCL
PyTorch
TensorRT
Triton
vLLM
DevOps
HPC
Apply
$85k – $110k per year • Remote • 6+ years exp • Bachelor's Degree
C++
C++
LLVM
AI/ML
CUDA
CUDA Toolkit
CUTLASS
MLIR
NCCL
TensorRT
Triton
vLLM
Apply
$130k – $150k per year • Remote • 6+ years exp • Bachelor's Degree
C++
C++
LLVM
AI/ML
CUDA
CUDA Toolkit
CUTLASS
MLIR
NCCL
TensorRT
Triton
vLLM
DevOps
HPC
Apply
$100k – $150k per year • Remote • 6+ years exp • Bachelor's Degree
C++
Go
Python
C++
PyTorch C++
AI/ML
DeepSpeed
FSDP
InfiniBand
JAX
Megatron-LM
NCCL
PyTorch
Ray
DevOps
CI/CD
FinOps
HPC
Kubernetes
SLURM
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
Agentic Workflows
AI Agents
Function Calling
LangChain
LlamaIndex
LLM
OpenAI
Tool Use
DevOps
CI/CD
Git
Apply
In office • Internship • PhD • Shanghai
C++
Python
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenCL
DevOps
HPC
Apply
In office • Internship • Master's Degree • Shanghai
C#
C++
Python
AI/ML
AI Agents
Model Context Protocol
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
AI Agents
Apply
In office • Internship • PhD • Shanghai
AI/ML
AI Agents
RAG
Apply
$81k – $105k per year • In office • Full-Time • 7+ years exp • Munich
TypeScript
JavaScript
Databases
PostgreSQL
AI/ML
LLM
Frontend
React.js
tRPC
DevOps
CI/CD
Docker
Kubernetes
Apply
$70k – $93k per year • In office • Full-Time • 3+ years exp • Munich
Python
TypeScript
JavaScript
Databases
PostgreSQL
AI/ML
LLM
Frontend
React.js
tRPC
DevOps
CI/CD
Docker
Kubernetes
Apply
$38k – $126k per year (Estimated) • In office • Full-Time • Munich
AI/ML
AI Agents
ChatGPT
Claude
Cursor
Design
Figma
Apply
$70k – $141k per year (Estimated) • Equity • In office • Full-Time • Berlin • Munich • Hamburg • Frankfurt am Main • Dortmund
DevOps
Amazon EKS
Ansible
AWS
CI/CD
CloudFormation
FluxCD
GitOps
Grafana
K3s
Kubernetes
Loki
OpenTelemetry
Prometheus
SOPS
Apply
$53k – $120k per year (Estimated) • Remote/Hybrid • Full-Time • Essen • Dresden • Stuttgart • Hamburg • Jena
AI/ML
AI Agents
LLM
RAG
Apply
See all jobs
This is one of many
394,870 more open roles from verified company boards, updated every day.