652,130open jobs
37,899companies
91,174added this week
Browse all
Salary
$63k – $142k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Architect · 3+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA has continuously reinvented itself over two decades. Our invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI - the next era of computing. NVIDIA is a “learning machine” that constantly evolves by adapting to new opportunities which are hard to seek, which only we can pursue, and which matter to the world. This is our life’s work: to amplify human inventiveness and intelligence.

NVIDIA is now driving innovation at the intersection of visual processing, high performance computing and artificial intelligence. We are looking for passionate, highly motivated, creative engineers to be part of our HW architecture team. As part of this team, you would be working on projects that will help make our next generation visual computing, automotive, GPU, HPC systems better. You will get to work on high performance CPU and Memory sub-systems, Next-Gen GPUs , NOC based Interconnect Fabric etc. Make the choice to join us today.

What you'll be doing:

  • System level performance analysis/ bottleneck analysis of complex, high performance GPUs and System-on-Chips (SoCs).

  • Work on hardware models of different levels of abstraction, including performance models, RTL test benches ,emulators and silicon to analyze performance and find performance bottlenecks in the system.

  • Understand key performance use-cases of the product. Develop workloads and test suits targeting graphics, machine learning, automotive, video, compute vision applications running on these products.

  • Work closely with the architecture and design teams to explore architecture trade-offs related to system performance, area, and power consumption.

  • Develop required infrastructure including performance models, testbench components, performance analysis and visualization tools.

  • Drive methodologies for improving turnaround time, finding representative data-sets and enabling performance analysis early in the product development cycle.

What we need to see:

  • BE/BTech, or MS/MTech in relevant area, PhD is a plus, or equivalent experience.

  • 3+ years of experience with exposure to performance analysis and complex system on chip and/or GPU architectures.

  • Strong understanding of System-on-Chip (SoC) architecture, graphics pipeline, memory subsystem architecture and Network-on-Chip (NoC)/Interconnect architecture.

  • Expert hands on competence in programming (C/C++) and scripting (Perl/Python). Exposure to Verilog/System Verilog, SystemC/TLM is a strong plus.

  • Strong debugging and analysis (including data and statistical analysis) skills, including use for RTL dumps to debug failures.

  • Hands on experience developing performance simulators, cycle accurate/approximate models for pre-silicon performance analysis is a strong plus.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative, autonomous and love a challenge, we want to hear from you. Come, join our Deep Learning Automotive team and help build the real-time, cost-effective computing platform driving our success in this exciting and quickly growing field. We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
652,130 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
$36k – $83k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Master's Degree • Pune
Python
C++
DevOps
GCP
AWS
Docker
Kubernetes
HPC
Apply
$48k – $120k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree
JavaScript
Node JS
Prolog
AI/ML
vLLM
CUDA Toolkit
Quantization
SGLang
TensorRT
TensorRT-LLM
Kubeflow
CUDA
NCCL
InfiniBand
Frontend
Bootstrap
DevOps
Cilium
Loki
Prometheus
SLURM
GitOps
Kubernetes
Grafana
kubeadm
KVM
QEMU
KubeVirt
HPC
Apply
$74k – $187k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree
JavaScript
Node JS
Prolog
AI/ML
vLLM
CUDA Toolkit
Quantization
SGLang
TensorRT
TensorRT-LLM
Kubeflow
CUDA
NCCL
InfiniBand
Frontend
Bootstrap
DevOps
Cilium
Loki
Prometheus
SLURM
GitOps
Kubernetes
Grafana
kubeadm
KVM
QEMU
KubeVirt
HPC
Apply
$85k – $213k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree
JavaScript
Node JS
Prolog
AI/ML
vLLM
CUDA Toolkit
Quantization
SGLang
TensorRT
TensorRT-LLM
Kubeflow
CUDA
NCCL
InfiniBand
Frontend
Bootstrap
DevOps
Cilium
Loki
Prometheus
SLURM
GitOps
Kubernetes
Grafana
kubeadm
KVM
QEMU
KubeVirt
HPC
Apply
$62k – $155k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree
JavaScript
Node JS
Prolog
AI/ML
vLLM
CUDA Toolkit
Quantization
SGLang
TensorRT
TensorRT-LLM
Kubeflow
CUDA
NCCL
InfiniBand
Frontend
Bootstrap
DevOps
Cilium
Loki
Prometheus
SLURM
GitOps
Kubernetes
Grafana
kubeadm
KVM
QEMU
KubeVirt
HPC
Apply
$85k – $236k per year (Estimated) • In office • Full-Time • 2+ years exp • PhD • Israel
Python
DevOps
CI/CD
Apply
In office • Full-Time • 3+ years exp • PhD • Yokneam
Apply
$26k – $51k per year (Estimated) • In office • Internship • Hsinchu
Python
Perl
Apply
$45k – $94k per year (Estimated) • Remote/Hybrid • Full-Time • 7+ years exp • Bachelor's Degree • Beijing
Python
AI/ML
Qwen
vLLM
Fine-tuning
Quantization
Knowledge Distillation
SGLang
Ollama
PyTorch
MiniMax
Hugging Face
Model Distillation
Apply
$35k – $87k per year (Estimated) • In office • Internship • Shanghai
Verilog
SystemVerilog
Apply
$20k – $46k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
C#
C++
Databases
Oracle
Analytics
Power BI
Apply
$27k – $62k per year (Estimated) • Remote/Hybrid • Full-Time • Pune • Bengaluru
Java
SQL
Java
Spring Boot
Databases
PostgreSQL
Redis
DevOps
CI/CD
AWS
AWS Lambda
Amazon S3
Amazon ECS
API Gateway
Management
Agile
Apply
Revenue Manager 1 day ago
In office • 8+ years exp • Bachelor's Degree • Bengaluru
Databases
ElasticSearch
Apply
Workplace Host 1 day ago
$13k – $32k per year (Estimated) • In office • Full-Time • 2+ years exp • Bengaluru
Apply
$15k – $35k per year (Estimated) • In office • Full-Time • Bengaluru
Apply
See all jobs
This is one of many
652,130 more open roles from verified company boards, updated every day.