368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$184k – $288k per year
Location
In office (Santa Clara)
Seniority
Senior · 6+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology-and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Join our group and discover how you can develop a lasting impact on the world.

NVIDIA BioNeMo is building the computational foundation for the next generation of biological discovery. We are looking for a Senior Software Engineer to join the cuEquivariance team - an NVIDIA library that accelerates geometric neural networks on NVIDIA GPUs, enabling researchers in molecular biology, materials science, and physics to train and deploy equivariant models at scale. This team builds and ships the production GPU kernels and software interfaces that power equivariant deep learning throughout the scientific field. The work spans CUDA kernel engineering, Python library development involving both PyTorch and JAX, and direct collaboration with research teams and external framework developers. If you want to work where GPU computing meets graph-based deep learning, this is the role for you. Your work will run in production pipelines across the scientific community.

What You Will Be Doing:

  • Build, implement, and optimize CUDA kernels for equivariant neural network primitives - tensor products, segmented polynomials, and triangle-based operations - targeting peak performance across NVIDIA GPU generations.

  • Be responsible for the end-to-end delivery of GPU-accelerated geometric ML primitives: from implementation to validated, production-quality software that external frameworks depend on.

  • Build and maintain the interfaces for PyTorch and JAX that expose cuEquivariance primitives to application developers and researchers.

  • Drive CI/CD infrastructure for multi-GPU kernel builds, automated correctness testing, and performance regression tracking.

  • Collaborate with Applied Science and research teams to evaluate new equivariant architectures and translate prototypes into production kernels.

  • Engage directly with third-party framework developers and partners to align on interfaces and ensure delivered software integrates cleanly into production pipelines.

What We Need to See:

  • 6+ years of software engineering experience with a strong background in CUDA and GPU programming.

  • Deep proficiency in C++ and Python; experience building and shipping production libraries used by external developers.

  • Good foundation in GPU computing: memory hierarchy, warp-level execution, occupancy, and performance profiling methodology.

  • Experience building or chipping in to production scientific software libraries, ML frameworks, or developer-facing GPU APIs.

  • Familiarity with concepts in geometric machine learning - equivariance, group representations, irreducible representations, or tensor products - sufficient to work efficiently in the domain.

  • BS/MS in Computer Science, Physics, Applied Mathematics, or a related field, or equivalent experience.

Ways to Stand Out from the Crowd:

  • You have chipped in to or deeply used a major neural network framework that respects equivariance: e3nn, MACE, NequIP, SE(3)-Transformers, or similar.

  • Hands-on experience with Triton kernel development or other GPU kernel authoring tools alongside CUDA.

  • Experience with mixed-precision or tensor-core-aware algorithm design for scientific or ML workloads.

  • PhD or equivalent experience in computational chemistry, biophysics, physics, or computer science with a focus on geometric deep learning or HPC.

  • Contributions to open-source geometric ML or GPU computing projects.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 21, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$140k – $253k per year (Estimated) • In office • 5+ years exp • Long Beach
C++
Rust
C++
Protobuf
AI/ML
Human-in-the-Loop
DevOps
CI/CD
Vector
Apply
$220k – $325k per year • Remote/Hybrid • Full-Time • 15+ years exp • New York
DevOps
AWS
Azure
CI/CD
GCP
Kubernetes
Platform Engineering
Cybersecurity
Threat Modeling
Apply
$100k – $500k per year • In office • Full-Time • 1+ year exp • Fort Collins
C++
Python
SystemVerilog
Cython
C++
CMake
Cython
PyBind11
AI/ML
ChatGPT
Claude
Copilot
Edge AI
Chips/EDA
Synopsys ZeBu
Apply
$21k – $57k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Noida
C#
SQL
TypeScript
JavaScript
C#
.NET
Frontend
Angular
DevOps
Azure
Azure DevOps
CI/CD
Git
TeamCity
Apply
$16k – $42k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Noida
C#
SQL
TypeScript
JavaScript
C#
.NET
Frontend
Angular
DevOps
Azure
Azure DevOps
CI/CD
Git
TeamCity
Apply
In office • Full-Time • 3+ years exp • Master's Degree • Shanghai • Shenzhen
C++
C
C++
TBB
C
MPI
Pthreads
AI/ML
CUDA
CUDA Toolkit
cuDF
OpenMP
RAPIDS
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara
Perl
Python
Apply
$98k – $252k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree • Switzerland
Assembly
C++
Fortran
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenMP
DevOps
HPC
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Hsinchu
Perl
Python
Apply
In office • Full-Time • 5+ years exp • Hsinchu • Taipei
C++
Python
AI/ML
InfiniBand
Apply
$100k – $137k per year • Equity • In office • Full-Time • 2+ years exp • Bachelor's Degree • Santa Clara
Apply
$72k – $99k per year • Equity • In office • Full-Time • Santa Clara
Apply
$166k – $290k per year • Equity • In office • Full-Time • 8+ years exp • Santa Clara
Management
ServiceNow
Apply
$133k – $272k per year (Estimated) • In office • Santa Clara
Go
Python
AI/ML
Edge AI
LLM
RAG
Apply
$80k – $110k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
MATLAB
Python
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.