368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$224k – $357k per year
Location
In office (Santa Clara, United States)
Seniority
Architect · 12+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology-and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

We’re looking for a Senior Architect to help shape the next generation of Agentic AI platforms for NVIDIA Marketing. In this role, you’ll combine technical leadership with hands-on engineering. You’ll translate marketing opportunities-including personalization, customer journeys, content intelligence, recommendations, campaign operations, and field enablement-into reliable AI agents and reusable platform capabilities. You’ll join a team with strong foundations in recommendation systems, retrieval-augmented generation (RAG), embedding-based retrieval, model evaluation, conversational AI, model adaptation, and production infrastructure. Your role will be to provide senior technical architecture, business translation, and platform leadership, helping the team evolve from high-value AI applications into a durable agentic AI ecosystem for Marketing. Our team works at the intersection of applied AI, agentic systems, marketing technology, enterprise data, and production platforms. You’ll have the opportunity to lead complex initiatives from early discovery through architecture, prototyping, evaluation, deployment, observability, governance, and scale.

What You’ll Be Doing:

  • Lead the architecture and delivery of Agentic AI solutions that support NVIDIA Marketing, including personalization, content discovery, campaign intelligence, recommendations, internal copilots, workflow automation, and customer-facing AI experiences.

  • Translate business and marketing needs into practical agent architectures, including goals, tools, retrieval, memory, planning, human-in-the-loop workflows, evaluation criteria, and production operating models.

  • Advance NVIDIA Marketing’s agentic AI platform by building on capabilities such as agent-tool gateways, multi-agent orchestration, conversational data assistants, recommendation APIs, embedding pipelines, contextual retrieval, model adapters, catalog intelligence, and production AI infrastructure.

  • Shape the platform roadmap across capabilities such as agent registries, tool catalogs, permission models, memory and state services, evaluation frameworks, observability, reusable agent patterns, and lifecycle management.

  • Partner with engineers to design reliable interfaces for agent invocation, tool execution, response formats, memory, state management, and integrations with marketing platforms, analytics systems, chat experiences, and content repositories.

  • Establish production practices for agentic systems across evaluation, regression testing, observability, latency, cost, reliability, safety, access controls, auditability, fallback behavior, and incident response.

  • Guide technical tradeoffs across model quality, retrieval precision, inference cost, throughput, latency, personalization, data freshness, privacy, security, and business impact.

  • Build prototypes, reference architectures, technical blueprints, and reusable components that help move promising ideas into scalable production systems.

What We Need to See:

  • A BS, MS, or PhD in Computer Science, AI/ML, Electrical Engineering, Data Science, a related technical field, or equivalent experience.

  • 12+ years of experience in one or more areas such as software engineering, AI/ML engineering, solutions architecture, applied AI, data platforms, or large-scale production systems.

  • Experience building and deploying applications involving LLMs, generative AI, RAG, recommendation systems, conversational AI, or agentic AI.

  • Strong programming skills in Python, along with experience working with APIs, Linux environments, distributed systems, containers, cloud-native infrastructure, and production debugging.

  • Understanding of agentic AI system design, including tool use, orchestration, planning, memory, retrieval, evaluation, guardrails, human approval, and failure handling.

  • Experience designing integrations between AI agents and enterprise tools using approaches such as MCP, function calling, API gateways, or related interoperability patterns.

  • Experience developing conversational AI experiences grounded in structured or semi-structured data. This may include text-to-SQL, intent classification, multi-turn dialogue, retrieval, or connections to live data sources.

  • Experience with production AI or software infrastructure, such as model serving, Kubernetes, Docker, CI/CD, observability, monitoring, health checks, performance testing, or cost optimization.

  • Demonstrated ownership of technical solutions across architecture, development, deployment, integration, and ongoing operations.

  • Ability to navigate ambiguous business problems, translate them into technical approaches, and communicate effectively with both technical and non-technical audiences.

Ways to Stand Out from the crowd::

  • Production multi-agent systems, agent runtimes, b frameworks, or tool-using agents.

  • Agent frameworks such as NVIDIA NeMo Agent Toolkit, LangGraph, LlamaIndex, LangChain, CrewAI, Semantic Kernel, OpenAI Agents SDK, Google ADK, or similar technologies.

  • NVIDIA AI software, including NIM, NeMo, NeMo Retriever, NeMo Guardrails, NeMo Agent Toolkit, Nemotron, Triton, NVIDIA AI Enterprise, or GPU-enabled Kubernetes environments.

  • MCP or agent-tool interoperability, including authenticated tool routing, server registries, enterprise tool catalogs, or policy-aware agent gateways.

  • Agent evaluation and observability, including traces, tool-call monitoring, offline and online evaluations, regression testing, quality dashboards, or business outcome measurement.

NVIDIA is widely considered to be one of the technology world’s most desirable employers! We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 28, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$25k – $42k per year • Equity 0–0.2% • Remote • Full-Time • 3+ years exp
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
$100k – $210k per year • Equity 0–0.5% • Remote • Full-Time • 3+ years exp • San Francisco
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
$100k – $200k per year • Equity 0.5–5% • In office • Full-Time • 1+ year exp • New York
Python
TypeScript
JavaScript
Python
FastAPI
Databases
DynamoDB
PostgreSQL
AI/ML
Claude
LLM
OpenAI
AI Agents
Frontend
Next.js
Tailwind CSS
React.js
DevOps
AWS
Docker
Vercel
GitHub
Management
Slack
Apply
$19k – $28k per year (net) • Remote • Full-Time • Moscow
C#
C++
C++
CMake
DevOps
CI/CD
Git
Management
Jira
Slack
Apply
$120k – $180k per year • Equity 0.2–1.5% • In office • Full-Time • 1+ year exp • San Francisco
Python
AI/ML
AI Agents
ChatGPT
DevOps
Error Budget
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara
Perl
Python
Apply
$98k – $252k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree • Switzerland
Assembly
C++
Fortran
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenMP
DevOps
HPC
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Hsinchu
Perl
Python
Apply
In office • Full-Time • 5+ years exp • Hsinchu • Taipei
C++
Python
AI/ML
InfiniBand
Apply
$156k – $348k per year (Estimated) • Remote • Full-Time • 10+ years exp • Bachelor's Degree • United Kingdom
AI/ML
CUDA
CUDA Toolkit
AI Agents
NVIDIA NeMo
Apply
$72k – $99k per year • Equity • In office • Full-Time • Santa Clara
Apply
$166k – $290k per year • Equity • In office • Full-Time • 8+ years exp • Santa Clara
Management
ServiceNow
Apply
$133k – $272k per year (Estimated) • In office • Santa Clara
Go
Python
AI/ML
Edge AI
LLM
RAG
Apply
$80k – $110k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
MATLAB
Python
Apply
$142k – $256k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara • Toronto
C++
Go
IoT
MQTT
OPC UA
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.