429,307open jobs
14,518companies
64,629added this week
Browse all
Salary
$152k – $242k per year
Location
In office (Santa Clara, Seattle)
Seniority
Senior · 2+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA Cloud Functions team is looking for a motivated, product-minded AI/ML Engineer with domain expertise in AI platform engineering at scale. Our team builds and operates a serverless deployment platform for enabling AI applications. Our product enables and scales AI inferencing workloads using globally distributed orchestration of workloads on GPU-backed cloud-agnostic Kubernetes clusters. You will be working with a team of passionate and skilled engineers that are continuously innovating at the speed of light to provide the best product possible, for both external customers and internal NVIDIA teams. We are looking for someone to join us at the forefront of defining cloud engineering paradigms for AI at scale.

What You'll be Doing:

  • Becoming a trusted subject matter expert by understanding user challenges and constraints. Translate this into product requirements and solutions, accelerating delivery of AI models and inference hosted on the NVCF platform.

  • Leading implementation of key features. Conducting user-acceptance testing, load testing and performance evaluations. Emphasis on customer experience, performance optimization and platform reliability.

  • Mentoring and embedding with other engineering teams building products on top of our platform on best practices for AI/ML workloads at scale with excellent performance, including ML reliability engineering at scale.

  • Producing reference architectures to guide customer use cases, applying the newest NVCF product features with the latest AI technologies.

  • Shepherding customer issues to resolution and providing timely warning of issues and risks.

  • Evaluating new and innovative technologies and tooling as the AI-at-scale landscape evolves to ensure we have a competitive product and forward-looking roadmap.

What We Need to See:

  • Masters, PhD, or equivalent experience in Computer Science, Artificial Intelligence, Applied Math, or related field

  • At least 2 years work experience with Python, Rust, Golang, Linux or Bash.

  • Experience in Deep Learning and Machine Learning; expertise in using AI/DL frameworks and inferencing software such as SGLang, vLLM, TensorRT-LLM, or Dynamo.

  • Knowledge of CPU and GPU architecture.

  • Excellent interpersonal skills including ability to explain sophisticated technical topics to non-experts.

  • Experience in the design, implementation, and release of AI/ML products to market. A flexible technologist familiar with all aspects of the software development lifecycle.

Ways to stand out from the crowd:

  • Demonstrate a strong desire to share knowledge with clients, partners and co-workers, able to show this through previous work.

  • Demonstrate expertise through projects or Open Source contributions in HPC, Data Analytics, Machine Learning, Deep Learning, Cloud Native Projects, Kubernetes, Slurm, or enabling GPU workloads.

  • Show a willingness and ability to dig into unfamiliar territories to tackle complex problems through examples in previous work.

  • Prior experience in building distributed systems.

NVIDIA offers highly competitive salaries and a comprehensive benefits package. NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most brilliant and talented people in the world working for us. Are you a creative engineer with a drive for advancing the state of AI and bringing it to the cloud? If you love to tackle problems and advocate for continuous, innovative improvement, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 13, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
429,307 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$61k – $171k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Yokneam
Python
Bash
Perl
AI/ML
CUDA Toolkit
CUDA
Apply
$103k – $254k per year (Estimated) • Remote • Full-Time • 8+ years exp • Australia
Python
JavaScript
Java
C++
Node JS
Bash
Databases
InfluxDB
DevOps
Terraform
Puppet
Ansible
Chef
Prometheus
CI/CD
Git
Kubernetes
Grafana
Configuration Management
OpenStack
Amazon S3
HPC
Apply
$146k – $264k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree • Zurich
Python
Rust
C++
AI/ML
LLM
NCCL
InfiniBand
NVLink
Edge AI
DevOps
Ansible
Kubernetes
Apply
$91k – $156k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • London
Python
SQL
Databases
Snowflake
Databricks
Microsoft Fabric
AI/ML
Pandas
NumPy
Analytics
Power BI
ETL/ELT
Apply
$184k – $288k per year • Remote • Full-Time • 6+ years exp • Bachelor's Degree • Austin
Python
DevOps
Puppet
Ansible
Chef
SLURM
CI/CD
GitOps
Configuration Management
Apply
$184k – $288k per year • Remote • Full-Time • 6+ years exp • Bachelor's Degree • Austin
Python
DevOps
Puppet
Ansible
Chef
SLURM
CI/CD
GitOps
Configuration Management
Apply
$200k – $322k per year • In office • Full-Time • 12+ years exp • Master's Degree • Santa Clara
Python
Java
C++
AI/ML
Claude
OpenAI Codex
DevOps
Prometheus
CI/CD
Apply
$196k – $311k per year • In office • Full-Time • 12+ years exp • Master's Degree • Santa Clara
Python
Ruby
AI/ML
AI Agents
DevOps
GCP
Azure
AWS
IAM
Cybersecurity
MITRE ATT&CK
NIST CSF
Zero Trust
Threat Modeling
Apply
$152k – $242k per year • Remote • Full-Time • Bachelor's Degree • Redmond • Santa Clara
AI/ML
CUDA Toolkit
CUDA
Apply
$152k – $242k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • United States
AI/ML
CUDA Toolkit
CUDA
MLIR
DevOps
HPC
Apply
AI/ML Director 4 hours ago
$198k – $369k per year (Estimated) • Equity • In office • Full-Time • Santa Clara
MATLAB
Apply
$90k – $124k per year • Equity • In office • Full-Time • Santa Clara
Apply
$100k – $137k per year • Equity • In office • Full-Time • Santa Clara
Apply
$100k – $110k per year • In office • Internship • Master's Degree • Santa Clara
Python
C
MATLAB
C
Embedded C
MATLAB
Simulink
Apply
$40k – $98k per year (Estimated) • Equity • In office • Full-Time • Santa Clara
AI/ML
Copilot
ChatGPT
Management
Outlook
Apply
See all jobs
This is one of many
429,307 more open roles from verified company boards, updated every day.