664,807open jobs
38,858companies
99,669added this week
Browse all
Salary
$184k – $288k per year
Location
In office (Santa Clara, United States)
Seniority
Architect · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by extraordinary technology-and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing.

NVIDIA is searching for an AI/ML Solutions Architect focusing on Hyperscale customers and Cloud Service Providers. Your primary responsibilities will be to lead software customer technical engagement for AI training, inference and infrastructure being deployed at vast scale. You will work across multiple organizations within NVIDIA as well as at the customer to ensure successful and trouble-free deployments. If you would you like to partner with a large company to build automation and management to create a robust large scale artificial intelligence infrastructure and are interested in the optimization and characterization of customer specific AI models and pipelines - you should apply!

What you’ll be doing:

  • As a key technical member of a focused account team, you will serve as the main point of contact for NVIDIA products, enabling internet giants and cloud providers to have an innovative AI/ML software infrastructure.

  • Work directly with best-in-class engineering teams to secure design wins, address challenges, bring solutions to production, and support them throughout their lifecycle.

  • Become a trusted advisor to your customer by understanding their environment, constraints, and long-term strategy. Translate these insights into product requirements and innovative solutions.

  • Help your customer enhance the value of NVIDIA technology, and provide feedback to NVIDIA for future product improvements.

  • Facilitate the resolution of customer issues, offering timely and proactive communications to mitigate risks.

  • Lead workshops, demos, and proof-of-concepts to showcase NVIDIA’s AI/ML capabilities.

  • Guide customers on standard processes for scalable AI model deployment and inference optimization.

What we need to see:

  • Minimum of a BS/MS in Computer Science, Electrical Engineering, or equivalent experience.

  • 8+ years of engineering experience with a proven track record in AI/ML-focused projects or enterprise-grade solutions.

  • Proven understanding of Linux, including solving, optimization, and customization for AI/ML workloads.

  • Strong understanding of data science and machine learning infrastructure-software and hardware.

  • Professional-level communication skills, including the ability to tailor messages for varying technical audiences and maintain composure in high-pressure situations.

  • Excellent follow-up and interpersonal skills, with a true passion for problem-solving.

  • Proficient in Python, with the ability to develop scripts and build custom tools. Experience with parallel programming or GPU acceleration (e.g., CUDA) is helpful.

  • Shown eagerness to learn and apply new technologies.

Ways to stand out from the crowd:

  • Experience with Chatbots, RAG pipelines, vector databases, and distributed training or inference workloads.

  • Experience or background in HPC (High Performance Computing) environments for AI or ML applications.

  • Familiarity with multi-node GPU clusters and performance tuning for large-scale AI workloads.

  • Experience developing in cloud and/or virtualized environments, containerized solutions, with knowledge of Docker, Kubernetes

  • Background with common deep learning frameworks such as PyTorch or JAX.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 18, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
664,807 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$28k – $84k per year (Estimated) • In office • Bachelor's Degree • Minato
Python
SQL
Apply
Network Engineer 10 hours ago
$121k – $257k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • San Francisco
Python
TypeScript
AI/ML
Cursor
Devin
Scale AI
DevOps
Terraform
Ansible
GCP
Azure
CI/CD
Git
AWS
Cloudflare
AIOps
Cybersecurity
Zero Trust
Least Privilege
Apply
$81k – $184k per year (Estimated) • Remote • Full-Time
Python
Dart
Python
FastAPI
AI/ML
Copilot
Claude
LLM
OpenAI Codex
Mobile
Flutter
DevOps
gRPC
GitHub Actions
CircleCI
Datadog
CI/CD
QA
Playwright
Pytest
Sentry
Apply
$155k – $312k per year (Estimated) • In office • Full-Time • San Francisco
Python
JavaScript
TypeScript
AI/ML
Reinforcement Learning
Frontend
React.js
Apply
$39k – $79k per year (Estimated) • In office • Internship • Bachelor's Degree • Austin
Python
Chips/EDA
Ansys HFSS
Apply
$131k – $237k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Yokneam
AI/ML
InfiniBand
Apply
$31k – $81k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Shanghai
AI/ML
Reinforcement Learning
TensorFlow
PyTorch
DevOps
GitHub
Robotics
Isaac Gym
Isaac Sim
MuJoCo
Isaac Lab
Motion Planning
Imitation Learning
Reinforcement Learning
Apply
$272k – $431k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • Westford • Austin • Durham
DevOps
Platform Engineering
eBPF
IAM
HPC
Cybersecurity
SOC 2
Zero Trust
Apply
$152k – $230k per year • In office • Full-Time • 8+ years exp • PhD • Santa Clara
AI/ML
AI Agents
Apply
$144k – $342k per year (Estimated) • In office • Full-Time • 10+ years exp • Master's Degree • Tel Aviv
Apply
$107k – $202k per year (Estimated) • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara
AI/ML
Copilot
ChatGPT
Apply
$156k – $299k per year (Estimated) • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree • Santa Clara
AI/ML
Copilot
ChatGPT
Apply
$106k – $232k per year (Estimated) • Equity • In office • Full-Time • Bachelor's Degree • Santa Clara • Irvine
AI/ML
Copilot
ChatGPT
Apply
$184k – $288k per year • In office • Full-Time • PhD • Santa Clara
Apply
$168k – $270k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Santa Clara
Python
AI/ML
vLLM
CUDA Toolkit
SGLang
TensorRT
TensorRT-LLM
PyTorch
CUDA
NCCL
DevOps
Splunk
Terraform
Puppet
Ansible
GCP
OpenTelemetry
Chef
Prometheus
Azure
AWS
Kubernetes
Grafana
Argo Workflows
KubeVirt
Incident Management
AWS Step Functions
Apply
See all jobs
This is one of many
664,807 more open roles from verified company boards, updated every day.