709,913open jobs
42,197companies
100,209added this week
Browse all
Location
In office (Beijing, Shanghai)
Seniority
Intern
Employment
Internship
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

At present, NVIDIA is exploring the immense possibilities of AI to define the forthcoming era of computing. Our division develops compiler technologies within the CUDA software stack that support the transformation of high-level parallel programs into efficient performance on NVIDIA GPUs. This internship grants a hands-on experience working closely with compilers, GPU architecture, and parallel programming languages. The role is notable for its mix of systems perspective, software knowledge, and significant exposure to authentic compiler issues across advanced GPU platforms.

What you’ll be doing:

  • Develop programs in PTX, CUDA, C/C++, or low-level GPU languages to validate compiler and evaluate generated code across key features and architectures.

  • Build software components, libraries, and technical solutions that advance compiler development and open up more innovative ways to explore and exercise compiler behavior.

  • Collaborate with compiler and hardware engineers on challenging architectural and compiler interactions, helping uncover deeper insights into how advanced GPU software features behave across releases.

  • Contribute to expansive engineering projects that improve the robustness, scalability, and future-readiness of the CUDA software stack for new GPU platforms.

What we need to see:

  • Pursuing an MS or PhD in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.

  • Strong programming skills in Python, C++, or similar languages used for systems or software development.

  • Solid understanding of data structures, algorithms, and software debugging fundamentals.

  • Familiarity with compiler concepts, computer architecture, operating systems, or parallel programming.

  • Ability to read technical details carefully, reason about complex behaviors, and communicate findings clearly in a collaborative engineering environment.

Ways to stand out from the crowd:

  • Exposure to compiler internals, code generation, static analysis, LLVM, or MLIR.

  • Background in GPU programming, CUDA, PTX, or parallel computing models.

  • Hands-on project work involving low-level systems, program analysis, or software aimed at optimizing software operation.

  • Passion for compilers, GPU frameworks, or software tools intended for developer use.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
709,913 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Beijing
$17k – $33k per year (Estimated) • In office • Internship • Master's Degree • Shanghai • Beijing
Python
C++
C++
PyTorch C++
LLVM
AI/ML
CUDA Toolkit
TensorRT
TensorRT-LLM
PyTorch
LLM
Mixture of Experts
CUDA
Edge AI
MLIR
Apply
$17k – $33k per year (Estimated) • In office • Internship • Bachelor's Degree • Shanghai • Beijing
Python
C++
C++
LLVM
AI/ML
CUDA Toolkit
TensorRT
CUDA
cuDNN
Edge AI
MLIR
Apply
$26k – $75k per year (Estimated) • In office • Full-Time • 2+ years exp • Master's Degree • Shanghai • Beijing • Shenzhen
Python
C++
AI/ML
CUDA Toolkit
CUDA
Machine Learning
DevOps
HPC
Apply
$17k – $33k per year (Estimated) • In office • Internship • Master's Degree • Shanghai • Beijing
C
C
Pthreads
AI/ML
CUDA Toolkit
OpenMP
TensorFlow
Mixture of Experts
CUDA
MLIR
Apache TVM
CUTLASS
DevOps
CI/CD
Apply
$60k – $122k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
Python
Verilog
C++
SystemVerilog
SystemC
AI/ML
LangChain
AI Agents
LLM
Anomaly Detection
NVIDIA NIM
OpenAI
NVIDIA NeMo
DevOps
HPC
Apply
$26k – $75k per year (Estimated) • In office • Full-Time • 2+ years exp • Master's Degree • Shanghai • Beijing • Shenzhen
Python
C++
AI/ML
CUDA Toolkit
CUDA
Machine Learning
DevOps
HPC
Apply
$60k – $129k per year (Estimated) • In office • Full-Time • 5+ years exp • Master's Degree • Taipei
Apply
$60k – $129k per year (Estimated) • In office • Full-Time • 5+ years exp • Hsinchu
Apply
$60k – $129k per year (Estimated) • In office • Full-Time • 5+ years exp • Hsinchu
Apply
$17k – $33k per year (Estimated) • In office • Internship • Master's Degree • Shanghai
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
ChatGPT
TensorRT
TensorRT-LLM
TensorFlow
PyTorch
LLM
Machine Learning
Apply
SAP FI Consultant 1 day ago
In office • Full-Time • Shanghai • Guangzhou • Shenzhen • Dalian • Chengdu
Apply
In office • Full-Time • Guangzhou • Shenzhen • Shanghai • Dalian • Chengdu
Apply
$21k – $47k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Beijing
AI/ML
Machine Learning
Apply
$46k – $92k per year (Estimated) • In office • Full-Time • 10+ years exp • Master's Degree • Shanghai • Beijing
Verilog
VHDL
Apply
$14k – $37k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Beijing
Apply
See all jobs
This is one of many
709,913 more open roles from verified company boards, updated every day.