709,913open jobs
42,197companies
100,209added this week
Browse all
Salary
$60k – $122k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Architect · 3+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA has continuously reinvented itself. Our invention of the GPU sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. Today, research in artificial intelligence is booming worldwide, which calls for highly scalable and massively parallel computation horsepower that NVIDIA GPUs excel.

NVIDIA is a “learning machine” that constantly evolves by adapting to new opportunities that are hard to solve, that only we can address, and that matter to the world. This is our life’s work , to amplify human creativity and intelligence. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join our diverse team and see how you can make a lasting impact on the world. As part of this team, you would be working on projects that will help make our next generation visual computing, automotive, GPU, HPC systems better. You will get to work on high performance CPU and Memory sub-systems, Next-Gen GPUs , NOC based Interconnect Fabric etc. Make the choice to join us today. Our work spans the full pre-silicon and post-silicon lifecycle: performance models, RTL simulation, emulation platforms, and silicon bringup. We catch bugs early, influence architecture and design decisions, and help ensure that the final product delivers the performance that our customers depend on. If you are passionate about building the hardware that underpins the future of AI and computing, this is where you belong.

What You'll Be Doing:

  • Partner with System Architecture, Usecase Modeling, Unit Architecture/PV, and Design/Verification teams to define comprehensive performance test plans that reflect real product usecases.

  • Drive full-chip SoC performance verification across all real engines and subsystems, running realistic concurrent workloads that represent how the product will be used by customers in the field.

  • Execute performance test plans across the full verification stack: cycle-accurate performance models, RTL simulation, emulation platforms, and silicon, identifying bottlenecks and regressions at each stage.

  • Debug performance failures through waveform analysis, signal-level queries, trace analysis, and system-level profiling to root-cause issues in the memory subsystem, fabric, or individual engines.

  • Influence architecture and microarchitecture decisions by surfacing performance data and trade-off analysis to the design and architecture teams early in the product development cycle.

  • Develop and maintain performance workloads, test suites, and infrastructure - including testbench components, performance simulators, analysis scripts, and automated regression flows.

  • Leverage AI-assisted tools and automation to accelerate repetitive analysis tasks, improve coverage, and free up engineering time for higher-level problem solving. This includes using LLM-based assistants for querying results and specs, AI-driven triage of regressions, and intelligent tooling that improves over time.

  • Drive methodology improvements to reduce verification turnaround time, improve coverage of representative workloads, and enable earlier performance insight in the product development cycle.

What We Need to See:

  • B.E./B.Tech or M.S./M.Tech (or equivalent experience) in Electrical Engineering, Computer Science, or a related field.

  • 3+ years of relevant experience in SoC or system-level architecture, performance verification, or hardware validation.

  • Strong understanding of SoC architecture including GPU and CPU pipelines, memory subsystem design (caches, DRAM controllers, coherency), Network-on-Chip (NoC)/fabric architecture, and high-speed IO interfaces.

  • Hands-on experience with RTL simulation and debug, including waveform-based debug and signal-level querying to isolate performance failures.

  • Solid programming skills in Python and C/C++; scripting proficiency in Bash/Python for automation and analysis. Exposure to Verilog/SystemVerilog or SystemC/TLM is a strong plus.

  • Strong debugging, data analysis, and statistical analysis skills - ability to synthesize large volumes of performance data into actionable insights.

  • Experience with or exposure to pre-silicon performance analysis methodologies, including performance models, cycle-approximate simulators, or emulation platforms.

  • Excellent communication skills and the ability to work effectively in a large, globally distributed engineering organization.

Ways to Stand Out from the Crowd:

  • Deep experience with RTL-level performance debug - particularly the ability to formulate precise signal queries and interpret waveforms to root-cause complex system-level interactions.

  • Background in system-level performance analysis for GPU, AI accelerators, or high-bandwidth memory subsystems, with knowledge of bottleneck identification across multiple concurrent engines.

  • Demonstrated use of AI and LLM-based tools (e.g., NVIDIA NIM/NeMo, OpenAI APIs, LangChain, or similar agentic frameworks) to measurably improve your own engineering productivity - whether for automated analysis, natural-language querying of data, intelligent triage, or workflow automation.

  • Experience building or deploying ML/AI-assisted tooling in an engineering or EDA context (e.g., regression analysis, anomaly detection, test generation, coverage closure).

  • Expertise in data analysis and visualization - ability to build dashboards and tooling that surface performance trends clearly to both engineering and architecture stakeholders.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
709,913 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
$26k – $75k per year (Estimated) • In office • Full-Time • 2+ years exp • Master's Degree • Shanghai • Beijing • Shenzhen
Python
C++
AI/ML
CUDA Toolkit
CUDA
Machine Learning
DevOps
HPC
Apply
In office • Internship • Master's Degree • Beijing • Shanghai
Python
C++
C++
LLVM
AI/ML
CUDA Toolkit
CUDA
MLIR
Apply
$17k – $33k per year (Estimated) • In office • Internship • Master's Degree • Shanghai • Beijing
Python
C++
C++
PyTorch C++
LLVM
AI/ML
CUDA Toolkit
TensorRT
TensorRT-LLM
PyTorch
LLM
Mixture of Experts
CUDA
Edge AI
MLIR
Apply
$17k – $33k per year (Estimated) • In office • Internship • Bachelor's Degree • Shanghai • Beijing
Python
C++
C++
LLVM
AI/ML
CUDA Toolkit
TensorRT
CUDA
cuDNN
Edge AI
MLIR
Apply
$28k – $65k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Noida
Python
C++
MATLAB
Analytics
Microsoft Excel
Management
Agile
Scrum
Apply
$26k – $75k per year (Estimated) • In office • Full-Time • 2+ years exp • Master's Degree • Shanghai • Beijing • Shenzhen
Python
C++
AI/ML
CUDA Toolkit
CUDA
Machine Learning
DevOps
HPC
Apply
$60k – $129k per year (Estimated) • In office • Full-Time • 5+ years exp • Master's Degree • Taipei
Apply
$60k – $129k per year (Estimated) • In office • Full-Time • 5+ years exp • Hsinchu
Apply
$60k – $129k per year (Estimated) • In office • Full-Time • 5+ years exp • Hsinchu
Apply
In office • Internship • Master's Degree • Beijing • Shanghai
Python
C++
C++
LLVM
AI/ML
CUDA Toolkit
CUDA
MLIR
Apply
$98k – $164k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Bengaluru
Apply
In office • Full-Time • Bachelor's Degree • Kochi • Bengaluru
Apex
Apex
MuleSoft
Analytics
Master Data Management
Apply
Technology Lead 1 day ago
$28k – $66k per year (Estimated) • In office • Full-Time • 6+ years exp • Bengaluru • Hyderabad
Python
JavaScript
TypeScript
SQL
Python
Flask
FastAPI
Databases
PostgreSQL
pgvector
OpenSearch
AI/ML
AWS Bedrock
RAG
Semantic Search
OpenAI
Semantic Search
Frontend
React.js
DevOps
Rest API
Terraform
CI/CD
Git
AWS
AWS Fargate
Amazon EC2
Amazon S3
Cybersecurity
ISO 27001
NIST 800-53
OWASP
Apply
$36k – $75k per year (Estimated) • Remote/Hybrid • Full-Time • 13+ years exp • Bengaluru
Python
JavaScript
PowerShell
AI/ML
AI Agents
DevOps
Azure
CI/CD
AWS
Docker
Kubernetes
Configuration Management
Unix
Cybersecurity
Microsoft Entra ID
LDAP
SIEM
Management
Jira
Apply
Senior UX Designer 1 day ago
$28k – $66k per year (Estimated) • Equity • In office • 5+ years exp • Bachelor's Degree • Bengaluru
C#
C#
.NET
Design
Sketch
InVision
Management
Scrum
Apply
See all jobs
This is one of many
709,913 more open roles from verified company boards, updated every day.