1,460,829open jobs
87,370companies
229,922added this week
Browse all
Salary
$76k – $188k per year
Location
In office (Santa Clara, United States)
Seniority
Intern
Employment
Internship

Confirmed on the employer's own hiring board on Oct 11, 2026. First seen by Alion on Oct 5, 2026. NVIDIA scores A on the Alion truth index.

Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA is searching for an outstanding PhD intern working on efficient deep learning to join the Deep Learning Efficiency Research (DLER) team. We are passionate about research that pushes boundaries but also has impact in the real world. The team has two core focuses: (1) efficient diffusion language models and multimodal generative models, and (2) efficient agentic AI with hybrid inference orchestration across cloud and edge. We are also excited about post-training model optimization (pruning, quantization, NAS), efficient architecture design, adaptive/dynamic inference, and resource-efficient training and finetuning.

You will work within an amazing and collaborative research team that consistently publishes at the top venues in computer vision and machine learning. Our existing expertise includes computer vision, deep learning, generative models, diffusion LLMs, multimodal models, and hybrid cloud-edge agentic systems. Your contributions have the chance to create real impact on our products.

What you'll be doing:

  • Research, design, and implement novel methods for efficient deep learning in one or both of the team’s focus areas:
    • Diffusion LLMs and multimodal models - sampling efficiency, adaptive unmasking, self-speculation / parallel decoding, training and distillation pipelines, and multimodal generation.
    • Efficient agentic AI - hybrid inference orchestration across cloud and edge, routing and scheduling policies, on-device vs. cloud expert delegation, and resource-aware agent loops.
  • Publish original research.
  • Collaborate with other team members and teams.
  • Work with product groups to transfer technology.
  • Collaborate with external researchers.

What we need to see:

  • Pursuing a Ph.D. in Computer Science/Engineering, Electrical Engineering, etc.
  • Excellent knowledge of theory and practice of machine learning and deep learning.
  • Experience with large language models, diffusion language models, multimodal / vision-language models, or agentic systems is required.
  • Hands-on experience with large-scale model training including data preparation and model parallelization (tensor and pipeline) is required.
  • Outstanding research track record with at least one top-tier conference (ICML, ICLR, NeurIPS, CVPR, ICCV, etc.).
  • Excellent communication skills.

Ways to stand out from the crowd:

  • Parallel programming (e.g., CUDA).
  • Interest or experience in hybrid cloud-edge inference, orchestration, or adaptive routing.
  • Background in pruning, quantization, NAS, or efficient backbones.

NVIDIA is widely considered to be one of the technology world’s most desirable employers with competitive salaries and a generous benefits package, we have some of the most forward-thinking and hardworking people in the world working for us. And, due to unprecedented growth, our best-in-class engineering teams are rapidly growing. If you're a creative and autonomous engineer with a real passion for computer architecture and technology, we want to hear from you!

Our internship hourly rates are a standard pay based on the position, your location, year in school, degree, and experience. The hourly rate for our interns is 38 USD - 94 USD.

You will also be eligible for Intern benefits.

Applications for this job will be accepted at least until October 9, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,460,829 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Early Careers
Similar stack
Same company
Santa Clara
DNA Examiner 6 days ago
$65k – $105k per year • Equity • In office • TS/SCI • Full-Time • 1+ year exp • Bachelor's Degree • Forest Park
Apply
≈ $32k – $54k per year (Estimated) • In office • Internship • Bachelor's Degree • Melbourne
Analytics
Power BI
Microsoft Excel
Apply
≈ $35k – $59k per year (Estimated) • In office • Internship • Melbourne
Python
C++
Apply
≈ $33k – $55k per year (Estimated) • In office • Internship • Bachelor's Degree • Grain Valley
Management
SharePoint
Microsoft Office
Apply
≈ $31k – $51k per year (Estimated) • In office • Internship • Melbourne
Management
Microsoft Office
Apply
≈ $17k – $39k per year (Estimated) • In office • 8+ years exp • Santiago
AI/ML
Multimodal AI
Apply
$166k – $238k per year • Hybrid • Full-Time • 5+ years exp • Menlo Park • Bellevue
Databases
Snowflake
AI/ML
Fine-tuning
AI Agents
LLM
RAG
Streamlit
DevOps
Prometheus
Platform Engineering
Cortex
Apply
AI/ML Engineer 11 hours ago
Hybrid • 3+ years exp • Bachelor's Degree • Budapest
Python
Databases
Databricks
AI/ML
Spark
Model Context Protocol
MLFlow
Fine-tuning
Prompt Engineering
AI Agents
NLP
Kubeflow
LLM
RAG
Semantic Search
Semantic Search
Knowledge Graph
Multi-Agent Systems
DevOps
Terraform
GCP
Azure DevOps
GitHub Actions
Azure
CI/CD
Analytics
Azure Data Factory
Apply
Hybrid • 3+ years exp • Bachelor's Degree • Szeged
Python
Databases
Databricks
AI/ML
Spark
Model Context Protocol
MLFlow
Fine-tuning
Prompt Engineering
AI Agents
NLP
Kubeflow
LLM
RAG
Semantic Search
Semantic Search
Knowledge Graph
Multi-Agent Systems
DevOps
Terraform
GCP
Azure DevOps
GitHub Actions
Azure
CI/CD
Analytics
Azure Data Factory
Apply
$94k – $137k per year • Hybrid • Full-Time • 3+ years exp • Master's Degree • Boston
Python
AI/ML
Multimodal AI
Machine Learning
Apply
$40k – $142k per year • In office • Internship • Master's Degree • Santa Clara
Python
AI/ML
Multimodal AI
Diffusion Models
Computer Vision
AI Agents
Self-Supervised Learning
Synthetic Data
World Models
Physical AI
Embodied AI
Machine Learning
Apply
$76k – $188k per year • In office • Internship • PhD • Santa Clara
Python
C++
C++
PyTorch C++
AI/ML
CUDA Toolkit
Reinforcement Learning
Computer Vision
PyTorch
CUDA
Machine Learning
Apply
$76k – $188k per year • In office • Internship • PhD • Santa Clara • Westford • Durham
Python
Rust
C++
C++
PyTorch C++
AI/ML
Multimodal AI
Computer Vision
PyTorch
Apply
$76k – $188k per year • In office • Internship • 5+ years exp • PhD • Santa Clara
Python
Rust
C++
AI/ML
Computer Vision
AI Agents
Machine Learning
DevOps
HPC
Apply
$76k – $188k per year • In office • Internship • 1+ year exp • PhD • Santa Clara
Python
AI/ML
CUDA Toolkit
JAX
TensorFlow
PyTorch
CUDA
Machine Learning
Apply
$40k – $142k per year • In office • Internship • Master's Degree • Santa Clara
Python
AI/ML
Multimodal AI
Diffusion Models
Computer Vision
AI Agents
Self-Supervised Learning
Synthetic Data
World Models
Physical AI
Embodied AI
Machine Learning
Apply
$48k per year • In office • Full-Time • High School Diploma • Santa Clara
Management
Agile
Apply
$102k – $210k per year • Equity • In office • 10+ years exp • Bachelor's Degree • Santa Clara
DevOps
GCP
Azure
AWS
Analytics
Microsoft Excel
Apply
$140k – $200k per year • In office • Full-Time • 8+ years exp • Santa Clara
MATLAB
DevOps
Red Hat
VMWare
SLURM
Azure
CI/CD
Jenkins
AWS
Ubuntu
KVM
Xen
HPC
Linux
Windows
Cybersecurity
Active Directory
LDAP
Apply
$213k – $288k per year • Equity • In office • Full-Time • 7+ years exp • Santa Clara
AI/ML
vLLM
Reinforcement Learning
Computer Vision
AI Agents
TensorFlow
PyTorch
Amazon SageMaker
SFT
Post-training
Megatron-LM
FSDP
TPU
Machine Learning
Apply
See all jobs
This is one of many
1,460,829 more open roles from verified company boards, updated every day.