686,079open jobs
39,763companies
97,004added this week
Browse all
Salary
$76k – $188k per year
Location
In office (Santa Clara, Westford, Seattle)
Seniority
Intern
Employment
Internship
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

The NVIDIA Architecture Research Group is seeking an intern to help define the future of GPU and data-center computing. Our research spans GPU microarchitecture, memory systems and memory models, large-scale data-center hardware, emerging accelerators, and the architecture implications of new workloads in AI, scientific computing, robotics, and other rapidly evolving domains. Are you excited about exploring foundational research questions, developing new architectural ideas, and helping translate those ideas into future NVIDIA systems?

In this position, you will apply your knowledge of computer architecture, parallel computing, compilers, runtime systems, and workloads to explore new hardware and hardware-software co-design opportunities. You will work with researchers and product architects to formulate research questions, develop and evaluate architecture concepts, and build the models, simulators, prototypes, and experimental infrastructure needed to test them. Projects may range from mechanisms within a GPU to memory consistency and programmability, rack- and data-center-scale architectures, and specialized hardware for emerging applications.

You should have a strong foundation in computer architecture and parallel systems, an ability to work across hardware and software boundaries, and experience with some combination of architecture modeling, simulation, workload analysis, compilers, runtime systems, or CPU/GPU programming. You should also be comfortable building robust research prototypes and communicating the insights produced by your work. NVIDIA pioneered programmable GPUs and CUDA and continues to shape the future of accelerated computing. This position offers an opportunity to conduct ambitious research and have a direct impact on future products.

What you’ll be doing:

  • Investigate new architecture concepts for future GPUs, memory systems, accelerators, and data-center-scale computing platforms.

  • Study emerging workloads and identify the architectural bottlenecks and opportunities they create.

  • Explore hardware-software co-design across architecture, compilers, runtime systems, programming models, and applications.

  • Develop models, simulators, prototypes, and experimental tools to evaluate new architecture ideas.

  • Research new approaches to memory hierarchy, coherence, consistency, data movement, and system-level programmability.

  • Collaborate with NVIDIA researchers, GPU architects, software teams, and product groups to develop and evaluate promising concepts.

  • Clearly communicate research findings through presentations, technical reports, and potentially research publications.

  • Help transfer successful research ideas, methodologies, and tools into NVIDIA product teams.

What we need to see:

  • Pursuing PhD Degree in relevant discipline(s) (CS, CE, EE, Physics, Math).

  • Relevant industrial and University experience. Relevant industries include hardware, software, and algorithm development in PC or workstation graphics, digital video or image processing, video game or console, cell phones or consumer electronics, rendering software, and computing.

  • Strong programming ability in C/C++, and scripting languages.

  • Experience as a CUDA programmer.

  • Background with building computer system simulators.

  • Experience building efficient low-level software tools such as runtime systems, binary translators, or compilers.

  • Strong background in computer architecture and parallel computer architectures.

Ways to stand out from the crowd:

  • Prior research experience and/or research publications at ISCA/MICRO/ASPLOS/HPCA/MLSYS

  • Versatile in using generative AI coding tools

  • Versatile in using GPU profiling tools and running DL models on GPUs

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. Are you creative and independently driven? Do you love a challenge? Come join our Architecture Research Group and help us invent the future of Computing.

Our internship hourly rates are a standard pay based on the position, your location, year in school, degree, and experience. The hourly rate for our interns is 38 USD - 94 USD.

You will also be eligible for Intern benefits.

Applications for this job will be accepted at least until September 21, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
686,079 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$76k – $188k per year • In office • Internship • PhD • Santa Clara
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
CUDA Toolkit
Reinforcement Learning
Multimodal AI
Diffusion Models
TensorFlow
PyTorch
CUDA
World Models
Embodied AI
Robotics
MuJoCo
Model Predictive Control
Imitation Learning
Reinforcement Learning
Apply
$37k – $96k per year (Estimated) • In office • Full-Time • 4+ years exp • Master's Degree • Beijing
C++
C++
PyTorch C++
AI/ML
CUDA Toolkit
TensorRT
PyTorch
LLM
CUDA
cuDNN
Apply
$17k – $32k per year (Estimated) • In office • Internship • Bachelor's Degree • Beijing • Shanghai • Shenzhen
Python
C++
AI/ML
CUDA Toolkit
Reinforcement Learning
CUDA
SFT
Post-training
Pre-training
World Models
Physical AI
Robotics
Reinforcement Learning
Apply
$135k – $305k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Singapore
Python
Bash
AI/ML
CUDA Toolkit
AI Agents
CUDA
InfiniBand
DevOps
Terraform
Ansible
Red Hat
Loki
Prometheus
CI/CD
Kubernetes
Ubuntu
Grafana
HPC
Apply
In office • Internship • Master's Degree
Python
Java
C++
DevOps
CI/CD
Git
Management
Agile
Apply
$100k – $167k per year • In office • Full-Time • Bachelor's Degree • Santa Clara
Python
Apply
$108k – $178k per year • In office • Full-Time • Bachelor's Degree • Santa Clara • Seattle
Python
C++
AI/ML
Flash Attention
AI Agents
LLM
MLIR
Apache TVM
Apply
$76k – $188k per year • In office • Internship • PhD • Santa Clara
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
CUDA Toolkit
Reinforcement Learning
Multimodal AI
Diffusion Models
TensorFlow
PyTorch
CUDA
World Models
Embodied AI
Robotics
MuJoCo
Model Predictive Control
Imitation Learning
Reinforcement Learning
Apply
$53k – $137k per year (Estimated) • In office • Full-Time • 5+ years exp • Bengaluru
C++
AI/ML
InfiniBand
NVLink
DevOps
HPC
Apply
$54k – $160k per year (Estimated) • In office • Full-Time • 1+ year exp • Bachelor's Degree • Yokneam • Tel Aviv
Python
Perl
AI/ML
Copilot
Claude
InfiniBand
NVLink
DevOps
VMWare
Git
Gerrit
KVM
Apply
$94k – $184k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Chandler • Hillsboro • Folsom • Santa Clara
Analytics
Power BI
Microsoft Excel
Apply
$118k – $230k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara • Hillsboro
Apply
$58k – $173k per year (Estimated) • In office • Internship • Bachelor's Degree • Chicago • Tampa • New York • Portland • Irvine
AI/ML
LLM
Management
Outlook
Apply
$180k – $270k per year • In office • 3+ years exp • Bachelor's Degree • Santa Clara
Python
SQL
AI/ML
Dagster
Scikit-learn
PyTorch
Feature Store
DevOps
GCP
Azure
AWS
Apply
$180k – $270k per year • In office • 6+ years exp • Bachelor's Degree • Santa Clara
Python
Go
Java
C++
DevOps
gRPC
Prometheus
Docker
Platform Engineering
Apply
See all jobs
This is one of many
686,079 more open roles from verified company boards, updated every day.