407,354open jobs
14,154companies
73,129added this week
Browse all
Location
In office (Shanghai)
Seniority
Intern
Employment
Internship
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

We are now looking for an Accelerated Computing Architect intern. NVIDIA is developing software and hardware system architectures for accelerated high performance computing, scientific computing, machine learning, artificial intelligence, datacenter, and automotive computing. This position offers you the opportunity to make a meaningful impact in a fast-moving, technology focused company.

What you'll be doing:

  • Performing in-depth analysis and optimization to ensure the best possible performance on current and/or next-generation NVIDIA GPUs.

  • Understanding and analyzing the interplay of hardware and software architectures on core algorithms, programming models, and applications.

  • Actively collaborating with the hardware design, software engineering, product, and research teams to guide the direction of accelerated computing.

  • Diving into accelerated computing applications to facilitate software-hardware co-design.

  • Writing up and presenting your work by writing white papers, conference publications, official blog posts, patent applications, etc. as appropriate.

What we need to see:

  • Pursuing B.Sc., M.Sc., or Ph.D. in relevant discipline (CS, EE, CE).

  • A passion for performance analysis and optimization.

  • Hands-on experience with the massively parallel GPU programming model, e.g. CUDA or OpenCL. Familiarity with APIs for multi-node communication, like MPI or OpenSHMEM/NVSHMEM, is a plus.

  • Solid background in GPU and computer systems architecture.

  • Strong knowledge of C and C++ with a solid understanding of software design, programming techniques, and algorithms. Familiarity with Python is a plus.

  • Good communication and organization skills, with a logical approach to problem solving, good time management, and task prioritization skills.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most brilliant and talented people in the world working for us. Are you creative and autonomous? Do you love the challenge of pushing an architecture to its limits? If so, we want to hear from you.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
407,354 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Shanghai
$73k – $175k per year (Estimated) • In office • Full-Time • 8+ years exp • Master's Degree • São Paulo
AI/ML
CUDA
CUDA Toolkit
NLP
Synthetic Data
DevOps
Docker
HPC
Kubernetes
Apply
$60k – $300k per year • Equity 0.1–0.5% • Remote • Full-Time • 6+ years exp • San Francisco
C++
Elixir
JavaScript
Kotlin
Node JS
Python
Rust
Swift
TypeScript
AI/ML
CUDA
CUDA Toolkit
Edge AI
Embeddings
Semantic Search
Semantic Search
Frontend
npm
WebAssembly
DevOps
Vector
Apply
Remote • Freelance
AI/ML
CUDA
CUDA Toolkit
Triton
Apply
In office • Internship • Master's Degree • Beijing • Shanghai • Shenzhen
C++
AI/ML
CUDA
CUDA Toolkit
Speech Recognition
Apply
$191k – $255k per year • Equity • In office • Full-Time • 10+ years exp • Bachelor's Degree
DevOps
AWS Lambda
HPC
AWS
Management
Smartsheet
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
Agentic Workflows
AI Agents
Function Calling
LangChain
LlamaIndex
LLM
OpenAI
Tool Use
DevOps
CI/CD
Git
Apply
In office • Internship • Master's Degree • Shanghai
C#
C++
Python
AI/ML
AI Agents
Model Context Protocol
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
AI Agents
Apply
In office • Internship • PhD • Shanghai
AI/ML
AI Agents
RAG
Apply
In office • Full-Time • 10+ years exp • Bachelor's Degree • Tel Aviv • Yokneam
Apply
In office • PhD • Shanghai
Design
SolidWorks
Apply
Engineer (Ph.D.) 1 day ago
In office • PhD • Shanghai
Apply
In office • PhD • Shanghai
Apply
In office • PhD • Shanghai
Apply
In office • PhD • Shanghai
Apply
See all jobs
This is one of many
407,354 more open roles from verified company boards, updated every day.