368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$184k – $288k per year
Location
Remote (United States)
Seniority
Architect · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

This a Full Remote job, the offer is available from: California (USA)

We are looking for a Senior Solutions Architect to help leading Enterprise ISVs design, build, optimize, and deploy production-grade Agentic AI systems on NVIDIA’s accelerated computing platform. In this role, you will partner with strategic software companies to translate frontier AI capabilities into reliable enterprise products, spanning multi-agent orchestration, RAG, tool use, model customization, and production deployment.

We work at the intersection of partner engineering, NVIDIA’s AI platform, and product development. You will lead sophisticated technical projects from initial exploration through architecture, prototyping, performance optimization, production rollout, and field enablement. Together, we will help partners adopt NVIDIA technologies and GPU-accelerated infrastructure to bring next-generation AI products to market.

What you'll be doing:

  • Lead technical delivery for strategic Agentic AI partner engagements from discovery and architecture through PoC, production readiness, rollout, and scale.

  • Design and build enterprise-grade agentic systems, including multi-agent workflows, tool-using agents, RAG-integrated applications, planning, memory, evaluation, and guardrail patterns.

  • Lead deep architecture reviews with partner engineering teams, driving tradeoffs across model quality, latency, efficiency, cost, retrieval quality, reliability, safety, security, and observability.

  • Build hands-on PoCs, benchmarks, reference architectures, and reusable blueprints that help Enterprise ISVs and NVIDIA field teams move from exploration to production.

  • Guide partners on model customization and post-training workflows, including supervised fine-tuning, reinforcement learning methods, human or AI feedback, direct preference optimization, PEFT/LoRA, synthetic data generation, evaluation, and regression analysis.

  • Work with agent harnesses and execution environments such as OpenShell, OpenAI Agents SDK, LangGraph, LlamaIndex, LangChain, CrewAI, Semantic Kernel, or similar frameworks.

  • Translate partner deployment findings into actionable feedback for NVIDIA Product and Engineering, so we can improve our platforms, tools, and field guidance.

What we need to see:

  • BS/MS/PhD in Computer Science, Electrical Engineering, AI/ML, or equivalent experience.

  • 8+ years of engineering, solutions architecture, applied ML, or technical deployment experience.

  • Consistent track record leading complex AI, ML, distributed systems, or enterprise software deployments from prototype to production.

  • Hands-on experience building LLM, generative AI, RAG, or agentic AI applications in production or production-like environments.

  • Strong programming and debugging skills in Python and Linux environments, with experience in PyTorch, TensorFlow, or similar deep learning frameworks.

  • Deep understanding of agentic AI architectures, including tool use, orchestration, memory, retrieval, planning, evaluation, guardrails, and failure handling.

  • Experience with model customization or post-training techniques such as SFT, RL/RLHF/RLAIF, DPO or relevant equivalent experience, reward modeling, LoRA/PEFT, quantization-aware optimization, or model evaluation.

  • Ability to lead ambiguous partner engagements, influence senior engineering collaborators, and communicate clearly with technical and executive audiences.

Ways to stand out from the crowd:

  • Hands-on experience with NVIDIA AI software such as NIM, NeMo Framework, NeMo Retriever, NeMo Guardrails, NeMo Agent Toolkit, Dynamo, Nemotron, Triton, TensorRT-LLM, or NIM Operator.

  • Experience building post-training pipelines for reasoning, tool use, domain adaptation, enterprise task performance, or agent behavior improvement.

  • Experience with agent harnesses, sandboxed execution, policy enforcement, OpenShell-like environments, or secure enterprise agent runtime build.

  • Recognized expertise in RAG, model customization, agent orchestration, enterprise AI security, or GPU-accelerated AI infrastructure, with field-facing technical presence through workshops, architecture reviews, talks, whitepapers, or developer enablement.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 20, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$105k – $175k per year • In office • Full-Time • Raleigh
Python
AI/ML
AI Agents
Amazon SageMaker
AutoGen
AWS Bedrock
CrewAI
Edge AI
Function Calling
LangChain
LangGraph
Model Context Protocol
NLP
Prompt Engineering
PyTorch
RAG
TensorFlow
DevOps
Amazon EC2
Amazon ECS
Amazon EKS
Amazon S3
AWS
AWS Lambda
Azure
GCP
Git
Kubernetes
Apply
$145k – $193k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Chicago • Washington • Denver • Jersey City • Boston
Python
AI/ML
AI Agents
Anomaly Detection
Embeddings
Fine-tuning
Function Calling
Hugging Face
LangChain
LLM
LLM Evaluation
LLM Guardrails
PyTorch
RAG
Red Teaming
Scikit-learn
Vertex AI
DevOps
Azure
GCP
Cybersecurity
Threat Modeling
Apply
$135k – $147k per year • In office • Full-Time • 7+ years exp • PhD • New York
Python
Databases
Databricks
Delta Lake
AI/ML
A2A
AI Agents
Computer Vision
Gemini
Google ADK
LangChain
LangGraph
LLM
Model Context Protocol
Prompt Engineering
RAG
Vertex AI
DevOps
GCP
Apply
$141k – $306k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Belfast
Python
TypeScript
AI/ML
A2A
AI Agents
Anthropic
Claude
Claude Code
LangChain
LangGraph
LLM
Model Context Protocol
OpenAI
OpenAI Agents SDK
OpenAI Codex
DevOps
AWS
OpenTelemetry
Apply
$87k – $195k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Belfast
Python
TypeScript
AI/ML
A2A
AI Agents
Anthropic
Claude
Claude Code
LangChain
LangGraph
LLM
Model Context Protocol
OpenAI
OpenAI Agents SDK
OpenAI Codex
DevOps
AWS
OpenTelemetry
Apply
In office • Full-Time • 8+ years exp • Bachelor's Degree • Pune
DevOps
Platform Engineering
SRE
Apply
In office • Full-Time • 3+ years exp • Master's Degree • Shanghai • Shenzhen
C++
C
C++
TBB
C
MPI
Pthreads
AI/ML
CUDA
CUDA Toolkit
cuDF
OpenMP
RAPIDS
Apply
$136k – $219k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • Santa Clara • Durham
C++
SystemVerilog
Verilog
Chips/EDA
Synopsys Verdi
UVM
Apply
In office • Full-Time • 8+ years exp • Bachelor's Degree • Tel Aviv
SystemVerilog
Verilog
Chips/EDA
UVM
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Yokneam • Tel Aviv
Verilog
VHDL
AI/ML
ChatGPT
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.