1,293,330open jobs
75,165companies
205,432added this week
Browse all
Salary
$200k – $322k per year
Location
Remote (United States)
Seniority
Senior · 12+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 7, 2026. First seen by Alion on Oct 5, 2026. NVIDIA scores A on the Alion truth index.

Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA’s DGX Cloud organization seeks a Senior Data Platform Engineer to contribute to building the shared data infrastructure that underpins decision-making within DGX Cloud. The DGX Cloud Data Platform transforms infrastructure telemetry and operational data into trustworthy data products for engineering, operations, finance, security, and product teams. These tools aid in monitoring fleet condition, capacity management, utilization tracking, cost oversight, reliability, governance, and the sustained expansion of large GPU fleets across cloud service providers and NVIDIA Cloud Partners. We would love for you to apply today!

What you'll be doing:

  • Define and guide the technical vision for a key DGX Cloud Data Platform domain covering various services, pipelines, data products, and consumer teams. Take responsibility for its architecture, interfaces, and growth, while foreseeing needs related to scale, reliability, performance, security, compatibility, and cost.
  • Lead the technical delivery of complex, cross-team initiatives. Transform unclear requirements into well-defined architectures and interfaces, coordinate implementation methods among contributors, personally write essential code, overcome technical obstacles, and guide integrations securely into production.
  • Architect, implement, and evolve batch and streaming systems that ingest, transform, reconcile, and serve fleet, capacity, utilization, cost, scheduling, and operational telemetry at scale across multiple environments and consumers.
  • Build shared platform capabilities-including libraries, workflow and orchestration abstractions, deployment tooling, and implementation standards-that are adopted across teams and measurably improve delivery speed, reliability, operational effort, and cost.
  • Serve as the technical lead for high-impact production investigations spanning pipelines, applications, query engines, distributed processing, storage, networks, and cloud services. Coordinate across owners, establish root cause, drive durable resolution, and ensure preventive improvements are implemented.
  • Establish and drive adoption of engineering standards for automated testing, data quality, reconciliation, lineage, service-level objectives, observability, secure service identities, least-privilege access, release readiness, and auditable deployments.
  • Establish robust data models, semantics, ownership boundaries, and serving interfaces across teams. Provide tables, APIs, automation, dashboards, and internal applications that ensure trusted DGX Cloud data is widely accessible without sacrificing accuracy or maintainability.
  • Provide technical leadership through architecture and build reviews, hands-on mentorship of senior engineers, and evidence-based resolution of difficult tradeoffs. Raise engineering quality across teams through reusable patterns, clear decisions, and sustained follow-through.

What we need to see:

  • 12+ years of relevant industry experience with a Bachelor’s or equivalent experience, and a Master’s degree or equivalent experience in Computer Science, Engineering, or a related field.
  • A sustained record of personally crafting, implementing, and operating production software, data platforms, databases, or distributed systems. This includes end-to-end technical ownership of a multi-system platform domain or a complex cross-team engineering initiative.
  • Experience includes deep hands-on work with distributed processing, analytical or relational databases, production ETL, change-data capture, streaming or event processing, or backend and cloud systems handling large data volumes.
  • Strong software engineering fundamentals and production proficiency in a backend or systems language, with deep experience using data-processing and platform libraries or frameworks to build reliable systems. Experience crafting reusable abstractions, reviewing substantial changes, and debugging critical code paths. Equivalent depth across different technology stacks is welcome.
  • Strong SQL and data-modeling skills, including practical depth in query execution, incremental processing, schema evolution, consistency, analytical consumption, idempotency, replay, late-arriving data, partial failure, and cross-system correctness.
  • Demonstrated skill in diagnosing failures across various systems by analyzing logs, metrics, traces, query plans, profiles, and controlled experiments, followed by applying and confirming long-lasting solutions.
  • Strong architectural judgment in assessing tradeoffs among reliability, performance, cost, security, compatibility, and maintainability, including experience guiding major migrations or architectural changes across teams without interrupting production service.
  • Experience establishing production safeguards and engineering practices that multiple teams adopt, including automated testing, CI/CD, monitoring, alerting, rollback, incident response, and secure deployment.

Ways to stand out from the crowd:

  • Deep experience with distributed data processing and lakehouse architectures, including optimization, reliability, and operation at production scale. Equivalent experience with large-scale database or data-processing platforms is welcome.
  • Experience crafting and operating distributed streaming or event-driven systems, including partitioning, consumer behavior, flow control, replay, delivery guarantees, and schema evolution.
  • Experience leading the scaling, migration, or performance improvement of relational, distributed, time-series, object-storage, or data systems specialized in managing searchable content.
  • Background operating cloud infrastructure, container orchestration, workload schedulers, compute or GPU clusters, and fleet-scale telemetry.
  • Experience defining and owning the production adoption of agentic systems or workflow automation. You should focus on evaluation, permissions, observability, failure recovery, and measurable improvements in engineering efficiency or operational outcomes.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 200,000 USD - 322,000 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until October 9, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,293,330 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
United States
Data Engineer 5 days ago
≈ $127k – $253k per year (Estimated) • Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • United States
Python
SQL
PowerShell
C#
C#
.NET
Databases
Cassandra
MS SQL
Azure SQL Database
AI/ML
Hadoop
Spark
DevOps
AWS
GitLab
Analytics
Tableau
Power BI
ETL/ELT
SSIS
SSAS
Management
Agile
Scrum
Kanban
Apply
$153k – $255k per year • Remote (United States) • 8+ years exp • Bachelor's Degree
Python
SQL
Python
Flask
FastAPI
AI/ML
MLFlow
Scikit-learn
AI Agents
Kubeflow
TensorFlow
Pandas
NumPy
PyTorch
Anomaly Detection
Time Series Forecasting
Agentic Workflows
Machine Learning
DevOps
OpenShift
AWS
Management
Confluence
Jira
Agile
Apply
≈ $21k – $43k per year (Estimated) • Remote (India) • 7+ years exp • India
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
Microsoft Fabric
AI/ML
Spark
Model Context Protocol
AI Agents
LLM
Analytics
Dimensional Modeling
Apply
≈ $59k – $86k per year (Estimated) • Remote (Poland) • Full-Time • Warsaw
Python
SQL
Python
pySpark
Databases
Snowflake
Databricks
Apache Kafka
AI/ML
Spark
DevOps
Terraform
GCP
Azure
CI/CD
AWS
Bicep
Apply
Cloud Data Architect 3 hours ago
≈ $55k – $132k per year (Estimated) • Remote (United Kingdom, Greece) • Athens
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
MS SQL
Microsoft Fabric
AI/ML
Spark
Machine Learning
DevOps
Azure DevOps
Azure
CI/CD
Git
FinOps
Analytics
Power BI
ETL/ELT
Azure Data Factory
Dimensional Modeling
Master Data Management
Management
Agile
Apply
≈ $162k – $274k per year (Estimated) • Remote (United States) • Full-Time • 15+ years exp • PhD
Python
SQL
Python
Flask
FastAPI
pySpark
Databases
Databricks
Apache Kafka
Google BigQuery
Amazon Redshift
BigQuery
AI/ML
LangGraph
LangChain
Claude
Spark
Airflow
LlamaIndex
Model Context Protocol
MLFlow
Vertex AI
dbt
Fine-tuning
Embeddings
Scikit-learn
Prompt Engineering
Function Calling
Computer Vision
AI Agents
NLP
LangSmith
Ragas
Llama
Kubeflow
Transformers
TensorFlow
PyTorch
Gemini
LLM
RAG
Hallucination
Reranking
Semantic Search
Hybrid Search
Hugging Face
Google AI Studio
Semantic Search
LLM Guardrails
Agentic Workflows
Multi-Agent Systems
Vertex AI Agent Builder
Machine Learning
DevOps
Rest API
Terraform
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Google Cloud Run
AWS Lambda
FinOps
GitHub
Amazon S3
IAM
Amazon CloudWatch
Amazon Kinesis
AWS Step Functions
Cybersecurity
Least Privilege
Analytics
ETL/ELT
AWS Glue
Dimensional Modeling
Master Data Management
Management
Agile
Apply
≈ $22k – $49k per year (Estimated) • Hybrid • Full-Time • 7+ years exp • Madurai
Python
JavaScript
SQL
C#
Python
Flask
FastAPI
Django
C#
.NET
Databases
PostgreSQL
Redis
Snowflake
Databricks
pgvector
Pinecone
RabbitMQ
Apache Kafka
Amazon Redshift
AI/ML
LangGraph
AutoGen
LangChain
Model Context Protocol
MLFlow
Prompt Engineering
Computer Vision
AI Agents
Semantic Kernel
AWS Bedrock
CrewAI
LLM
RAG
OpenAI
LLM Guardrails
Frontend
GraphQL
React.js
Mobile
State Management
DevOps
Rest API
Terraform
WebSockets
Azure
CI/CD
Git
AWS
Docker
Kubernetes
AIOps
Management
Agile
QA
Playwright
Apply
≈ $8k – $24k per year (Estimated) • In office • Full-Time • 1+ year exp • Bachelor's Degree • Pune
Python
SQL
AI/ML
Copilot
Prompt Engineering
AI Agents
Copilot Studio
DevOps
Azure DevOps
Azure
Git
AWS
GitHub
Analytics
Power BI
Management
Jira
Power Automate
Power Apps
SharePoint
Agile
Apply
In office • Top Secret • 10+ years exp
Java
C++
Java
Spring Boot
DevOps
Rest API
CI/CD
Jenkins
AWS
Docker
Kubernetes
Linux
Management
Agile
Apply
≈ $9.5k – $21k per year (Estimated) • In office • 2+ years exp • Saint Petersburg
AI/ML
AI Agents
GigaChat
DevOps
Rest API
Management
Confluence
Jira
Apply
$168k – $270k per year • Remote (United States) • Full-Time • 8+ years exp • Bachelor's Degree • United States
SQL
AI/ML
AI Agents
Time Series Forecasting
DevOps
CI/CD
Cybersecurity
Least Privilege
Analytics
ETL/ELT
Apply
$184k – $288k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
Python
Go
JavaScript
Node JS
Scala
Databases
PostgreSQL
Databricks
Apache Iceberg
Delta Lake
Apache Kafka
AI/ML
dbt
DevOps
Rest API
GCP
Azure
AWS
Docker
Kubernetes
Self-Healing
Linux
Analytics
ETL/ELT
Apply
$168k – $265k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
Python
JavaScript
TypeScript
SQL
Node JS
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Spark
Prompt Engineering
AI Agents
LLM
RAG
Tool Use
Frontend
Angular
DevOps
CI/CD
Jenkins
Git
Self-Healing
Analytics
ETL/ELT
Informatica
Master Data Management
Management
Jira
Apply
$140k – $224k per year • Remote (United States) • Full-Time • 5+ years exp • Bachelor's Degree • United States
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
ElasticSearch
Apache Kafka
OpenSearch
AI/ML
Spark
AI Agents
LLM
Time Series Forecasting
DevOps
GCP
SLURM
Azure
CI/CD
AWS
Kubernetes
Cybersecurity
Least Privilege
Analytics
ETL/ELT
Apply
$168k – $265k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Santa Clara
Python
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Copilot
Claude
Spark
ChatGPT
AI Agents
Gemini
LLM
RAG
Knowledge Graph
DevOps
Self-Healing
Analytics
ETL/ELT
Informatica
Master Data Management
Apply
$191k – $248k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • New York • Pittsburgh • Boston
Marketing
Salesforce
Apply
≈ $116k – $231k per year (Estimated) • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree • Houston
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
Apache Kafka
Azure SQL Database
AI/ML
Spark
MLFlow
Scikit-learn
MLlib
TensorFlow
Keras
PyTorch
Ray
Machine Learning
DevOps
Rest API
Azure DevOps
Azure
CI/CD
Git
Kubernetes
Azure AKS
Analytics
ETL/ELT
Azure Data Factory
Apply
≈ $52k – $91k per year (Estimated) • Equity • In office • Full-Time • 3+ years exp • High School Diploma • United States
Apply
≈ $39k – $89k per year (Estimated) • Hybrid • Full-Time • United States
Apply
≈ $73k – $141k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Charlotte
Analytics
Microsoft Excel
Management
Microsoft Office
Apply
See all jobs
This is one of many
1,293,330 more open roles from verified company boards, updated every day.