411,234open jobs
14,368companies
73,679added this week
Browse all
Salary
$184k – $288k per year
Location
In office (Santa Clara, United States)
Seniority
Architect · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA is seeking outstanding AI Solutions Architects to assist and support customers that are building solutions with our newest AI and accelerated computing technologies. At NVIDIA, our solutions architects work across product, engineering, sales, developer relations, business development, and partner teams to help customers design, deploy and optimize AI infrastructure.

This role will focus on helping ISVs adopt NVIDIA accelerated infrastructure for training, fine-tuning, inference, retrieval, and agentic AI workloads. This role is an excellent opportunity to work in an interdisciplinary team at NVIDIA! You will serve as a technical advisor for accelerated systems architecture, GPU and networking systems, cluster design, architectures, orchestration, validation, and production deployment for AI data centers.

What You Will Be Doing:

  • Partner with ISVs on discovery, architecture reviews, technical deep dives, POCs, benchmarks, demos, and production deployment guidance

  • Advise on the design, build-out, and optimization of accelerated AI infrastructure, including large-scale clusters

  • Support infrastructure design across compute, networking, storage, containers, observability, security, power, and data center operations

  • Drive adoption of systems monitoring, telemetry, and management tools to improve cluster utilization, reliability, performance and workload insight

  • Build repeatable reference architectures, deployment guides, sizing guidance, benchmark reports, technical playbooks, demos and whitepapers

  • Travel up to 20% customer meetings may be required

What We Need To See:

  • BS, MS, or PhD in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, other Engineering or related fields (or equivalent experience)

  • 8+ years of hands-on experience in AI infrastructure, accelerated computing, distributed systems, cloud infrastructure, high-performance computing, or machine learning platforms

  • Strong experience designing, deploying, and operating accelerated computing infrastructure at scale

  • In-depth knowledge of AI cluster orchestration, scheduling, automation and CI/CD deployment pipelines

  • Understanding of data center networking technologies such as InfiniBand, Ethernet, RDMA, network configuration or performance tuning

  • Familiarity with infrastructure requirements for AI workloads, including distributed training, inference serving, model deployment, storage performance, and cluster reliability

  • Excellent presentation, communication, problem-solving, documentation, and collaboration skills

Ways To Stand Out From The Crowd:

  • Experience architecting AI factories, large GPU clusters, multi-node training environments, production inference platforms

  • Experience deploying LLM training, fine-tuning, RAG, and inference workflows on large-scale AI infrastructure

  • Experience evaluating cluster performance using benchmarks such as MLPerf, HPL, or workload-specific performance tests

  • Applications and systems-level knowledge of OpenMPI, NCCL, distributed training frameworks, and GPU communication patterns

  • Experience delivering technical training, workshops, whitepapers, blogs, or mentoring engineers, researchers, and customers on AI/HPC infrastructure

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 20, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
411,234 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$28k – $72k per year (Estimated) • In office • Full-Time • 15+ years exp • Master's Degree • Hyderabad
Python
Databases
Chroma
Pinecone
Qdrant
Weaviate
AI/ML
AI Agents
AutoGen
Claude
Copilot
Copilot Studio
CrewAI
DeepSeek
Fine-tuning
Gemini
GraphRAG
Hugging Face
Knowledge Graph
LangChain
Llama
LlamaIndex
LLM
LoRA
Mistral
MLFlow
Multi-Agent Systems
Multimodal AI
NLP
PEFT
Prompt Engineering
PyTorch
RAG
Reinforcement Learning
RLHF
Semantic Kernel
Semantic Search
Semantic Search
SFT
Synthetic Data
TensorFlow
Transformers
Vertex AI
DevOps
AWS
Azure
CI/CD
Docker
Kubernetes
Platform Engineering
Vector
Robotics
Digital Twin
Apply
$17k – $39k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Hyderabad
Python
SQL
Python
pySpark
Databases
Databricks
Snowflake
AI/ML
AI Agents
Spark
DevOps
CI/CD
Analytics
ETL/ELT
Apply
Developer - DevOps 9 min ago
$15k – $44k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Bengaluru
Python
DevOps
Ansible
CI/CD
Docker
Git
GitHub
GitLab
Grafana
Helm
Jenkins
JFrog Artifactory
Kubernetes
OpenShift
Prometheus
Splunk
Cybersecurity
SonarQube
Apply
$15k – $43k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
Bash
Java
Python
Java
Maven
Spring Boot
DevOps
Ansible
Chef
CI/CD
Configuration Management
Jenkins
Prometheus
Puppet
Red Hat
Self-Healing
Splunk
Zabbix
Cybersecurity
SonarQube
Apply
Solution Architect 9 min ago
$27k – $63k per year (Estimated) • In office • Full-Time • 10+ years exp • Master's Degree • Bengaluru
JavaScript
SQL
TypeScript
Java
Java
Gradle
Hibernate
Maven
Spring Framework
Databases
MS SQL
PostgreSQL
Frontend
React.js
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
Jenkins
Kubernetes
OpenShift
Prometheus
Splunk
Management
Confluence
QA
Postman
Selenium
TestNG
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
Agentic Workflows
AI Agents
Function Calling
LangChain
LlamaIndex
LLM
OpenAI
Tool Use
DevOps
CI/CD
Git
Apply
In office • Internship • PhD • Shanghai
C++
Python
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenCL
DevOps
HPC
Apply
In office • Internship • Master's Degree • Shanghai
C#
C++
Python
AI/ML
AI Agents
Model Context Protocol
Apply
In office • Internship • Master's Degree • Shanghai
Perl
Python
AI/ML
AI Agents
Apply
In office • Internship • PhD • Shanghai
AI/ML
AI Agents
RAG
Apply
$221k – $387k per year • Equity • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
DevOps
Incident Management
Management
ServiceNow
Apply
$56k – $94k per year • In office • Contractor • Santa Clara
Apply
$64k – $74k per year • In office • Contractor • Santa Clara
Apply
$56k per year • In office • Contractor • Santa Clara
Apply
$84k – $94k per year • In office • Contractor • Santa Clara
Apply
See all jobs
This is one of many
411,234 more open roles from verified company boards, updated every day.