1,329,188open jobs
78,176companies
213,498added this week
Browse all
Salary
$168k – $270k per year
Location
Remote (United States)
Seniority
Senior · 8+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 7, 2026. First seen by Alion on Oct 5, 2026. NVIDIA scores A on the Alion truth index.

Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

NVIDIA’s DGX Cloud organization seeks a Senior Data Platform Engineer to contribute to building the shared data infrastructure that underpins decision-making within DGX Cloud. The DGX Cloud Data Platform transforms infrastructure telemetry and operational data into trustworthy data products for engineering, operations, finance, security, and product teams. These tools aid in monitoring fleet condition, capacity management, utilization tracking, cost oversight, reliability, governance, and the sustained expansion of large GPU fleets across cloud service providers and NVIDIA Cloud Partners. We would love for you to apply today!

What you'll be doing:

  • Define and guide the technical vision for a key DGX Cloud Data Platform domain covering various services, pipelines, data products, and consumer teams. Take responsibility for its architecture, interfaces, and growth, while foreseeing needs related to scale, reliability, performance, security, compatibility, and cost.
  • Lead the technical delivery of complex, cross-team initiatives. Transform unclear requirements into well-defined architectures and interfaces, coordinate implementation methods among contributors, personally write essential code, overcome technical obstacles, and guide integrations securely into production.
  • Architect, implement, and evolve batch and streaming systems that ingest, transform, reconcile, and serve fleet, capacity, utilization, cost, scheduling, and operational telemetry at scale across multiple environments and consumers.
  • Build shared platform capabilities-including libraries, workflow and orchestration abstractions, deployment tooling, and implementation standards-that are adopted across teams and measurably improve delivery speed, reliability, operational effort, and cost.
  • Serve as the technical lead for high-impact production investigations spanning pipelines, applications, query engines, distributed processing, storage, networks, and cloud services. Coordinate across owners, establish root cause, drive durable resolution, and ensure preventive improvements are implemented.
  • Establish and drive adoption of engineering standards for automated testing, data quality, reconciliation, lineage, service-level objectives, observability, secure service identities, least-privilege access, release readiness, and auditable deployments.
  • Establish robust data models, semantics, ownership boundaries, and serving interfaces across teams. Provide tables, APIs, automation, dashboards, and internal applications that ensure trusted DGX Cloud data is widely accessible without sacrificing accuracy or maintainability.
  • Provide technical leadership through architecture and build reviews, hands-on mentorship of senior engineers, and evidence-based resolution of difficult tradeoffs. Raise engineering quality across teams through reusable patterns, clear decisions, and sustained follow-through.

What we need to see:

  • 8+ years of relevant industry experience with a Bachelor’s or equivalent experience, and a Master’s degree or equivalent experience in Computer Science, Engineering, or a related field.
  • A sustained record of personally crafting, implementing, and operating production software, data platforms, databases, or distributed systems. This includes end-to-end technical ownership of a multi-system platform domain or a complex cross-team engineering initiative.
  • Experience includes deep hands-on work with distributed processing, analytical or relational databases, production ETL, change-data capture, streaming or event processing, or backend and cloud systems handling large data volumes.
  • Strong software engineering fundamentals and production proficiency in a backend or systems language, with deep experience using data-processing and platform libraries or frameworks to build reliable systems. Experience crafting reusable abstractions, reviewing substantial changes, and debugging critical code paths. Equivalent depth across different technology stacks is welcome.
  • Strong SQL and data-modeling skills, including practical depth in query execution, incremental processing, schema evolution, consistency, analytical consumption, idempotency, replay, late-arriving data, partial failure, and cross-system correctness.
  • Demonstrated skill in diagnosing failures across various systems by analyzing logs, metrics, traces, query plans, profiles, and controlled experiments, followed by applying and confirming long-lasting solutions.
  • Strong architectural judgment in assessing tradeoffs among reliability, performance, cost, security, compatibility, and maintainability, including experience guiding major migrations or architectural changes across teams without interrupting production service.
  • Experience establishing production safeguards and engineering practices that multiple teams adopt, including automated testing, CI/CD, monitoring, alerting, rollback, incident response, and secure deployment.

Ways to stand out from the crowd:

  • Deep experience with distributed data processing and lakehouse architectures, including optimization, reliability, and operation at production scale. Equivalent experience with large-scale database or data-processing platforms is welcome.
  • Experience crafting and operating distributed streaming or event-driven systems, including partitioning, consumer behavior, flow control, replay, delivery guarantees, and schema evolution.
  • Experience leading the scaling, migration, or performance improvement of relational, distributed, time-series, object-storage, or data systems specialized in managing searchable content.
  • Background operating cloud infrastructure, container orchestration, workload schedulers, compute or GPU clusters, and fleet-scale telemetry.
  • Experience defining and owning the production adoption of agentic systems or workflow automation. You should focus on evaluation, permissions, observability, failure recovery, and measurable improvements in engineering efficiency or operational outcomes.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 270,250 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until October 9, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,329,188 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
United States
≈ $72k – $168k per year (Estimated) • Remote (likely Brazil)
Python
SQL
Databases
Databricks
Google BigQuery
BigQuery
DevOps
GCP
Git
Analytics
ETL/ELT
Apply
$96k – $138k per year • Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • Boston
Python
SQL
Databases
KDB+
DevOps
Rest API
Azure
CI/CD
Git
Docker
Linux
Unix
Apply
Big Data Engineer 1 hour ago
$60k – $135k per year • In office • 5+ years exp • Saint Charles
AI/ML
Hadoop
Apply
≈ $115k – $288k per year (Estimated) • Equity • Remote (North America) • 6+ years exp
Python
SQL
AI/ML
Spark
LLM
LLM Evaluation
Apply
$125k – $150k per year • Hybrid • 6+ years exp • Bachelor's Degree • Alexandria
Python
SQL
SPSS
AI/ML
Machine Learning
Analytics
Microsoft Excel
Management
Agile
Apply
$85k – $141k per year • In office • Public Trust • Full-Time • 7+ years exp • Bachelor's Degree • Tysons
Python
Java
Bash
Java
Apache Tomcat
Databases
PostgreSQL
AI/ML
AI Agents
Gemini
DevOps
Splunk
Terraform
Ansible
GCP
OpenShift
Helm
Datadog
Knative
Prometheus
GitLab CI
CI/CD
GitOps
Windows Server
ArgoCD
Jenkins
Git
AWS
Kubernetes
Nginx
Grafana
Tekton
Bitbucket
Amazon EKS
Google GKE
Google Cloud Run
AIOps
GitLab
Linux
DNS
Apache HTTP Server
Apply
≈ $81k – $162k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Chicago
Python
JavaScript
Java
SQL
PowerShell
Node JS
Groovy
Perl
SAS
Java
Spring Boot
Gradle
Databases
PostgreSQL
Oracle
Frontend
React.js
DevOps
Rest API
Azure
CI/CD
Jenkins
Git
AWS
Kubernetes
SOAP
Analytics
Tableau
Power BI
ETL/ELT
Management
Agile
Scrum
QA
Postman
Apply
≈ $81k – $162k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • Austin • San Antonio
Python
JavaScript
TypeScript
Node JS
Databases
MySQL
PostgreSQL
Frontend
React.js
DevOps
CI/CD
AWS
Docker
AWS Lambda
Amazon S3
IAM
Amazon CloudWatch
API Gateway
Management
Agile
Apply
Software Engineer 2 days ago
≈ $90k – $166k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • San Antonio
JavaScript
TypeScript
Node JS
Databases
PostgreSQL
DynamoDB
Frontend
React.js
DevOps
Terraform
AWS CDK
CloudFormation
CI/CD
Jenkins
Git
AWS
Kubernetes
Amazon EKS
AWS Lambda
GitHub
Amazon S3
Management
Agile
Apply
≈ $29k – $67k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Costa Rica
AI/ML
AI Agents
Analytics
Power BI
Marketing
Salesforce
Apply
$200k – $322k per year • Remote (United States) • Full-Time • 12+ years exp • Bachelor's Degree • United States
SQL
AI/ML
AI Agents
Time Series Forecasting
DevOps
CI/CD
Cybersecurity
Least Privilege
Analytics
ETL/ELT
Apply
$184k – $288k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
Python
Go
JavaScript
Node JS
Scala
Databases
PostgreSQL
Databricks
Apache Iceberg
Delta Lake
Apache Kafka
AI/ML
dbt
DevOps
Rest API
GCP
Azure
AWS
Docker
Kubernetes
Self-Healing
Linux
Analytics
ETL/ELT
Apply
$168k – $265k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
Python
JavaScript
TypeScript
SQL
Node JS
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Spark
Prompt Engineering
AI Agents
LLM
RAG
Tool Use
Frontend
Angular
DevOps
CI/CD
Jenkins
Git
Self-Healing
Analytics
ETL/ELT
Informatica
Master Data Management
Management
Jira
Apply
$140k – $224k per year • Remote (United States) • Full-Time • 5+ years exp • Bachelor's Degree • United States
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
ElasticSearch
Apache Kafka
OpenSearch
AI/ML
Spark
AI Agents
LLM
Time Series Forecasting
DevOps
GCP
SLURM
Azure
CI/CD
AWS
Kubernetes
Cybersecurity
Least Privilege
Analytics
ETL/ELT
Apply
$168k – $265k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Santa Clara
Python
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Copilot
Claude
Spark
ChatGPT
AI Agents
Gemini
LLM
RAG
Knowledge Graph
DevOps
Self-Healing
Analytics
ETL/ELT
Informatica
Master Data Management
Apply
$38k per year • In office • Full-Time • United States
Marketing
X (Twitter)
Instagram
Apply
$226k – $438k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • United States
AI/ML
Fine-tuning
AI Agents
LLMOps
Machine Learning
DevOps
Kubernetes
Web3
Tokenomics
Apply
≈ $37k – $81k per year (Estimated) • In office • United States
Cybersecurity
HIPAA
Apply
Machine Operator 3 hours ago
≈ $36k – $62k per year (Estimated) • In office • Full-Time • 1+ year exp • High School Diploma • United States
Apply
≈ $54k – $105k per year (Estimated) • In office • Full-Time • 1+ year exp • High School Diploma • United States
DevOps
Windows
Apply
See all jobs
This is one of many
1,329,188 more open roles from verified company boards, updated every day.