397,589open jobs
13,889companies
77,448added this week
Browse all
Salary
$187k – $410k per year (Estimated)
Location
In office (Tel Aviv, Israel, Yokneam)
Seniority
Architect · 12+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

We are now looking for a Senior GPU Architect, Deep Learning! The NVIDIA GPU Architecture group is looking for a world class hardware architect to help define and drive future GPU architectures for deep learning and accelerated computing. In this role, you will work at the earliest stages of product definition, with primary focus on defining architectural features for future GPUs, while also contributing to relevant microarchitectural direction and tradeoffs. You will help drive new capabilities from concept through modeling, design collaboration, validation, and silicon readiness.

A key part of NVIDIA’s strength is our ability to translate workload insight into differentiated hardware. We are constantly looking for ways to improve our GPU architecture for training and inference workloads while advancing performance, efficiency, scalability, and programmability. In this position, you will focus primarily on defining new architectural features, while also shaping the supporting microarchitectural direction, evaluating design alternatives, and partnering closely with ASIC design, verification, performance, and software teams to bring new capabilities into future products.

What you'll be doing:

  • Define and architect new GPU hardware features for future deep learning and parallel processing workloads.

  • Drive microarchitectural exploration across key areas such as compute pipelines, memory hierarchy, data movement, synchronization, and performance efficiency.

  • Analyze workload behavior and translate bottlenecks into clear architectural requirements and hardware feature proposals.

  • Evaluate performance, power, area, complexity, and programmability tradeoffs for new architectural directions.

  • Develop and use functional and performance models to study new features and refine the architecture before implementation.

  • Work closely with RTL, design, verification, compiler, and software teams to ensure successful execution from architecture definition to productization.

  • Create clear architecture specifications, validation plans, and success criteria for the features you define.

  • Be ready to learn, dig deep, and work across the full stack when required - from workloads and models to RTL and silicon.

What we need to see:

  • BS, MS, or PhD in Computer Science, Electrical Engineering, Computer Engineering, or equivalent experience.

  • 12+ years of relevant industry experience in GPU architecture, computer architecture, or other parallel processing architectures.

  • Strong background in hardware architecture and microarchitecture.

  • Experience defining and evaluating architectural features with solid understanding of performance, power, and area tradeoffs.

  • Strong programming and scripting skills in C, C++, and Python.

  • Experience with architectural modeling, simulation, or performance analysis.

  • Background in parallel computing, memory systems, high performance computing, or deep learning acceleration.

  • Strong communication skills and the ability to drive technical work across distributed, interdisciplinary teams.

Ways to stand out from the crowd:

  • Deep understanding of modern GPU architecture and the interaction between hardware and AI workloads.

  • Experience with memory subsystem architecture, interconnects, coherence, scheduling, or execution pipelines.

  • Experience with pre-silicon performance studies, workload characterization, and architectural correlation.

  • Familiarity with training and inference behavior for large-scale deep learning models.

  • Experience with silicon bring-up, debug, or post-silicon analysis.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative, autonomous, and love a challenge, consider joining our GPU Architecture team and help us build the next generation of AI computing platforms.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
397,589 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Tel Aviv
Algorithm Engineer 1 hour ago
Remote/Hybrid • Full-Time • 3+ years exp • Netanya
C++
MATLAB
Python
AI/ML
Computer Vision
Quantization
DevOps
CI/CD
HPC
Analytics
A/B Testing
Apply
$130k – $180k per year • Remote • 10+ years exp • Bachelor's Degree
C++
C
C++
LLVM
PyTorch C++
TensorFlow C++
C
MPI
AI/ML
CUDA
CUDA Toolkit
CUTLASS
DeepSpeed
JAX
MLIR
NCCL
PyTorch
ROCm
TensorFlow
TensorRT
Triton
vLLM
DevOps
AWS
Azure
GCP
HPC
Apply
In office • Full-Time • Xuzhou
C++
Java
Apply
$145k – $160k per year • In office • 7+ years exp • Bachelor's Degree
Java
MATLAB
Python
SQL
AI/ML
Hadoop
Spark
Analytics
Tableau
Apply
In office • 2+ years exp • PhD • Princeton
MATLAB
Python
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
Agentic Workflows
AI Agents
Function Calling
LangChain
LlamaIndex
LLM
OpenAI
Tool Use
DevOps
CI/CD
Git
Apply
In office • Internship • PhD • Shanghai
C++
Python
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenCL
DevOps
HPC
Apply
In office • Internship • Master's Degree • Shanghai
C#
C++
Python
AI/ML
AI Agents
Model Context Protocol
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
AI Agents
Apply
In office • Internship • PhD • Shanghai
AI/ML
AI Agents
RAG
Apply
Remote/Hybrid • Full-Time • 1+ year exp • Tel Aviv
Marketing
Salesforce
Apply
$66k – $221k per year (Estimated) • In office • Full-Time • 7+ years exp • Tel Aviv
Node JS
TypeScript
JavaScript
Frontend
GraphQL
React.js
Mobile
React Native
Apply
In office • Full-Time • Tel Aviv
Apply
DevOps Engineer 1 day ago
$58k – $172k per year (Estimated) • In office • Full-Time • 4+ years exp • Tel Aviv
Kotlin
Node JS
Python
JavaScript
Databases
Amazon Aurora
Apache Kafka
RabbitMQ
Redis
Kafka
AI/ML
AWS Bedrock
Copilot
Vertex AI
OpenAI
DevOps
Amazon EC2
Amazon EKS
Ansible
Argo Workflows
ArgoCD
AWS
CI/CD
Cloudflare
Configuration Management
Docker
GCP
GitHub Actions
Grafana
Jenkins
Kubernetes
Nginx
Platform Engineering
Prometheus
Terraform
Terragrunt
GitHub
IAM
Cybersecurity
HashiCorp Vault
Cryptography
Vault
Apply
$86k – $244k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Tel Aviv
C++
Go
AI/ML
AI Agents
Apply
See all jobs
This is one of many
397,589 more open roles from verified company boards, updated every day.