520,299open jobs
17,886companies
71,691added this week
Browse all
Salary
$124k – $196k per year
Location
In office (Santa Clara)
Seniority
Architect · 2+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA’s Executive Briefing Center (EBC) Solutions Architect (SA) team is looking for a highly hands-on Solutions Architect with exemplary communication skills. The role involves developing, demonstrating (in the NVIDIA EBC), and packaging agentic AI systems. Partnering with account SAs you will co-develop proof of concepts (POC) and "uplift" their presentation quality to match the NVIDIA branding and messaging used with Executive meetings.

This is a builder’s and presenter's role! You will spend time architecting and writing code. You will develop multi-agent systems, retrieval pipelines, and optimized inference stacks on NVIDIA’s full-stack accelerated computing platform. We want a creative, diligent, and curious engineer energized by agentic AI and ready to make significant change. If that’s you, join us!

What you’ll be doing:

  • Architect, build, and ship end-to-end Agentic AI applications for a variety of use cases-spanning multi-agent coordination, long-horizon reasoning, planning, and tool use.

  • Act as Technical Advisor alongside fellow Subject Matter Experts (SME) in Executive Briefings.

  • Creating and presenting demos that are used at Trade Shows or Customer Meetings.

  • Partner with NVIDIA engineering, product, and sales teams to secure build wins, translate customer feedback into actionable product and roadmap insights, and scale global expertise through technical collateral, workshops, and developer communities.

What we need to see:

  • BS/MS/PhD in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, AI/ML, or a related field (or equivalent experience)

  • 2+ years as an ML/Software Engineer or Solutions Architect writing production-level code in Python and/or C/C++ in Linux environments.

  • Validated experience building sophisticated agentic and multi-agent AI systems using orchestration frameworks such as LangGraph, LlamaIndex, CrewAI, LangChain, OpenAI Agents SDK -including tool-using and routing agents. Solid understanding of MCP and A2A is vital.

  • Strong background in PyTorch and distributed GPU (post-)training. Able to quickly prototype and build scalable GPU-accelerated architectures. Applies test-time compute, reinforcement learning, inference optimization, and post-training. Deploys workloads at scale on public cloud (AWS, GCP, Azure, OCI) or on-premise.

  • Strong grasp of the way C-Suite and Industry Leaders think paired with excellent communication and presentation skills. Able to explain sophisticated ideas to both technical and non-technical groups. Leads projects from start to finish in a fast-paced, multitasking setting.

Ways to stand out from the crowd:

  • Practical experience working directly with the NVIDIA agentic AI software stack-NVIDIA NIM, NeMo Framework, NeMo Retriever, NeMo Agent Toolkit, Dynamo, Triton Inference Server, TensorRT-LLM, and AI Blueprints.

  • Expertise building LLM evaluation harnesses, benchmarking systems, observability platforms, and safety guardrails, plus fine-tuning and optimizing reasoning-focused LLMs and SLMs through timely engineering and quantization.

  • Experience developing production-grade deployment patterns using Kubernetes/OpenShift, CI/CD automation, and secure cloud-native infrastructure, with familiarity with modern agent architectures and emerging communication protocols such as MCP (Model Context Protocol) or Google A2A.

  • Proven experience handling NVIDIA GPU architectures, CUDA-X libraries (cuBLAS, cuDNN, RAPIDS), and HPC technologies (NCCL, InfiniBand, MPI, NVLink), along with familiarity in large-scale data processing and distributed/parallel computing frameworks (e.g., Spark, Dask).

  • A strong public profile (blogs, GitHub, conference talks) that demonstrates your expertise and passion for agentic AI.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 124,000 USD - 195,500 USD for Level 2, and 152,000 USD - 241,500 USD for Level 3.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 14, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
520,299 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
Engineering Intern 2 hours ago
$38k – $58k per year (Estimated) • In office • Internship • Bachelor's Degree • Princeton
Python
C#
C++
DevOps
Git
Apply
$184k – $288k per year • In office • Full-Time • 6+ years exp • Master's Degree • Santa Clara
Python
C++
C++
PyTorch C++
AI/ML
vLLM
CUDA Toolkit
Multimodal AI
SGLang
PyTorch
LLM
CUDA
Triton
NCCL
CUTLASS
Management
Agile
Apply
$75k – $114k per year • In office • Top Secret • Full-Time • 3+ years exp • Bachelor's Degree • Albuquerque • Dayton • Simi Valley • Herndon • Arlington
Python
PowerShell
Cybersecurity
Microsoft Sentinel
VirusTotal
MITRE ATT&CK
Cyber Kill Chain
KEV
Tanium
Microsoft Entra ID
Apply
$91k – $139k per year • In office • Top Secret • Full-Time • 3+ years exp • Bachelor's Degree • Albuquerque • Dayton • Simi Valley • Herndon • Arlington
Python
SQL
PowerShell
DevOps
Rest API
Cybersecurity
Microsoft Defender
CIS Benchmarks
CVSS
KEV
Tanium
Apply
$48k – $116k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Mexico City
Python
Apply
$24k – $58k per year (Estimated) • In office • Full-Time • 5+ years exp • Pune
Python
AI/ML
CUDA Toolkit
Multimodal AI
Function Calling
Computer Vision
AI Agents
TensorRT
Semantic Search
CUDA
Semantic Search
Tool Use
Physical AI
DevOps
Helm
GitHub Actions
GitLab CI
CI/CD
Jenkins
Docker
Kubernetes
Apply
$184k – $288k per year • In office • Full-Time • 6+ years exp • Master's Degree • Santa Clara
Python
C++
C++
PyTorch C++
AI/ML
vLLM
CUDA Toolkit
Multimodal AI
SGLang
PyTorch
LLM
CUDA
Triton
NCCL
CUTLASS
Management
Agile
Apply
$192k – $305k per year • In office • Full-Time • Master's Degree • Redmond • Santa Clara
AI/ML
CUDA Toolkit
CUDA
Apply
$124k – $196k per year • Remote/Hybrid • Full-Time • 5+ years exp • PhD • Santa Clara
Apply
$136k – $219k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara
DevOps
HPC
Chips/EDA
Cadence Allegro
Apply
$46k – $71k per year (Estimated) • In office • Internship • PhD • Santa Clara • Irvine • Austin • Morrisville • Chandler
Python
AI/ML
Copilot
LangGraph
AutoGen
LangChain
Claude
ChatGPT
LlamaIndex
Model Context Protocol
Fine-tuning
Reinforcement Learning
Prompt Engineering
Multimodal AI
Diffusion Models
AI Agents
TensorFlow
PyTorch
CrewAI
LLM
RAG
Hugging Face
A2A
Agentic Workflows
Multi-Agent Systems
Tool Use
DevOps
Git
Management
n8n
Apply
$61k – $100k per year • In office • Bachelor's Degree • Santa Clara
Apply
$167k – $291k per year • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree • Santa Clara
Python
SQL
Bash
Databases
Azure Cosmos DB
AI/ML
AI Agents
Anomaly Detection
LLM Guardrails
DevOps
Terraform
GCP
CloudFormation
Azure
CI/CD
GitOps
AWS
Kubernetes
Platform Engineering
Service Mesh
Self-Healing
Amazon EKS
Google GKE
Azure AKS
AWS Lambda
Amazon EC2
AIOps
Incident Management
SLI/SLO/SLA
Management
ServiceNow
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara
Apply
$168k – $270k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Santa Clara
Python
SQL
AI/ML
Spark
DevOps
Terraform
Azure
AWS
Docker
Kubernetes
Analytics
Tableau
Power BI
ETL/ELT
SAP BusinessObjects
Management
Outlook
Apply
See all jobs
This is one of many
520,299 more open roles from verified company boards, updated every day.