1,061,370open jobs
62,137companies
177,439added this week
Browse all
Salary
≈ $91k – $181k per year (Estimated)
Location
Remote (Spain)
Seniority
Senior · 8+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 1, 2026. First seen by Alion on Oct 1, 2026. NVIDIA scores A on the Alion truth index.

Overview
Company
Impact
Profile match
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.

The DGX Cloud organization at NVIDIA brings together cutting-edge hardware and software innovation to deliver industry-leading accelerated computing for the world's most adventurous AI workloads. We're a team of innovative engineers dedicated to solving some of the world's biggest challenges, constantly driving advancements, and impacting millions of lives worldwide!

We are looking for an outstanding Senior Systems Software Engineer with deep experience in distributed systems, open-source technologies such as Kubernetes and containers, and a strong background in systems performance and scalability. The ideal candidate brings broad, end-to-end experience across the stack - from GPU operator and device plugins to distributed inference serving and cloud platforms - along with the technical depth to investigate and address exciting, real-world problems at scale. In this pivotal role, you will take on the challenge of scaling AI infrastructure while optimizing total cost of ownership, driving down cost per token to unlock the next generation of AI innovation and AI factories!

What you'll be doing:

  • Drive end-to-end performance and scale characterization for the NVIDIA DGX Cloud software stack, from Kubernetes control and data planes through NVIDIA components such as GPU Operator, Network Operator, DCGM, NIM, and distributed inference serving, following issues from orchestration down to the metal.

  • Collaborate with AI researchers, developers and customers to develop innovative, automated tests that simulate real user workloads using custom-built and leading open-source tools and frameworks.

  • Deep dive into performance and scale issues in complex distributed systems, including interactions between Kubernetes and the NVIDIA software stack, to identify and resolve root causes.

  • Design and develop monitoring, reporting and analysis tools for performance and scale testing across software, GPU and CPU resources.

  • Triage, debug and root cause issues related to operating Kubernetes clusters at ultra-large scale, ensuring reliability and efficiency.

  • Build and maintain a high-velocity framework that enables continuous, always-on performance and scale testing via a modern CI/CD pipeline.

  • Document research, methodologies and results clearly and concisely, and present findings at internal and external venues, including community conferences such as KubeCon and GTC.

  • Engage efficiently with upstream communities - including Kubernetes, CNCF and NVIDIA open-source projects - to validate performance and scalability of AI workloads early and help shape design and development decisions.

What we need to see:

  • 8+ years of experience Computer Architecture, Networking, Storage systems, Accelerators and Bachelors/Masters in Engineering (preferably, Electrical Engineering, Computer Engineering, or Computer Science) or equivalent experience

  • Expertise in Kubernetes and familiarity with related CNCF projects

  • Background in working with large scale parallel and distributed accelerator-based systems

  • Expertise optimizing performance and AI workloads on large scale systems

  • Experience with performance modeling and benchmarking at scale

  • Proficiency in Golang/Python

  • Background with the NVIDIA software ecosystem in both training and inference domains

  • Expertise with at least one of public CSP infrastructure (GCP, AWS, Azure, OCI for example)

Ways to stand out from the crowd:

  • Strong operational experience with any one of the Kubernetes distributions

  • Prior experience scaling Kubernetes clusters to ultra-large node and object counts

  • Demonstrated history of working in the open-source community

  • Excellent communication and interpersonal abilities

  • PhD in relevant areas

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. For Poland: The base salary range is 292,500 PLN - 507,000 PLN for Level 4, and 375,000 PLN - 650,000 PLN for Level 5.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,061,370 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
Spain
≈ $95k – $222k per year (Estimated) • Remote (Switzerland) • Full-Time
Go
Java
DevOps
Splunk
OpenShift
CI/CD
ArgoCD
Kubernetes
Tekton
Apply
≈ $133k – $222k per year (Estimated) • Remote (United States) • Full-Time • 4+ years exp • Bachelor's Degree • United States
Python
JavaScript
SQL
C#
Bash
C#
.NET
Databases
PostgreSQL
RabbitMQ
MS SQL
Frontend
React.js
DevOps
Rest API
CI/CD
Git
AWS
Ubuntu
Gitflow
Linux
SOAP
Management
Stripe
Agile
Scrum
Apply
$90k – $160k per year • Remote (Argentina) • Full-Time • 4+ years exp • Buenos Aires
JavaScript
TypeScript
Node JS
Node JS
TypeORM
Databases
PostgreSQL
Frontend
Redux
React.js
React Query
Redux Toolkit
Analytics
Microsoft Excel
Apply
≈ $42k – $116k per year (Estimated) • Remote (Argentina) • Full-Time • 8+ years exp • Buenos Aires
Java
SQL
Java
Spring Boot
Databases
MySQL
PostgreSQL
Apache Kafka
DevOps
Rest API
ArgoCD
AWS
Docker
Kubernetes
GitLab
QA
Swagger
Apply
≈ $18k – $50k per year (Estimated) • Remote (Philippines) • Makati
JavaScript
TypeScript
Ruby
C#
Ruby
Ruby on Rails
C#
.NET
Databases
MySQL
ElasticSearch
AI/ML
Claude Code
Frontend
Angular
DevOps
CI/CD
Trunk-Based Development
Management
Agile
Scrum
Apply
≈ $130k – $248k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • United States
Python
SQL
Databases
PostgreSQL
PostGIS
AI/ML
Polars
Fine-tuning
Scikit-learn
SciPy
Multimodal AI
TensorFlow
NumPy
PyTorch
Statsmodels
Time Series Forecasting
Feature Store
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Cybersecurity
NIST 800-171
Robotics
Digital Twin
Analytics
Tableau
Power BI
Plotly
Dimensional Modeling
SpaceTech
GDAL
PDAL
Apply
In office • Bachelor's Degree • Bengaluru
Python
SQL
Bash
Databases
Google BigQuery
BigQuery
DevOps
Terraform
GCP
Azure
CI/CD
AWS
Kubernetes
Platform Engineering
Google GKE
FinOps
IAM
Linux
Apply
Senior QA Engineer 8 hours ago
≈ $71k – $189k per year (Estimated) • In office • Full-Time • Melbourne
DevOps
CI/CD
Shift-Left
Cybersecurity
Shift-Left Security
Management
Agile
Apply
≈ $87k – $167k per year (Estimated) • Hybrid • 7+ years exp • Bachelor's Degree • Charlotte
Python
AI/ML
Machine Learning
Design
AutoCAD
Apply
$71k per year • Hybrid • Bachelor's Degree • Greenwood Village
Python
JavaScript
Design
AutoCAD
Apply
$152k – $242k per year • Remote (United States) • Full-Time • 5+ years exp • Bachelor's Degree • United States
Go
Rust
C++
Databases
ClickHouse
AI/ML
CUDA Toolkit
CUDA
Frontend
GraphQL
DevOps
Datadog
Kubernetes
Grafana
HPC
Apply
$184k – $288k per year • Remote (United States) • Full-Time • 6+ years exp • Bachelor's Degree • United States
DevOps
Helm
ArgoCD
Kubernetes
Apply
≈ $141k – $275k per year (Estimated) • Remote (Canada) • Full-Time • 10+ years exp • Bachelor's Degree • Toronto
Python
AI/ML
AI Agents
NVIDIA NeMo
DevOps
SLURM
Docker
Kubernetes
Apply
≈ $86k – $133k per year (Estimated) • Remote (Poland) • Full-Time • 6+ years exp • Bachelor's Degree • Poland • Czech Republic • Romania • Hungary
Rust
C++
AI/ML
CUDA Toolkit
CUDA
DevOps
Linux
TCP/IP
Apply
$152k – $242k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Seattle • Santa Clara
DevOps
Terraform
Ansible
CI/CD
Kubernetes
GitLab
Linux
Unix
Apply
≈ $52k – $149k per year (Estimated) • In office • Spain
Databases
ElasticSearch
DevOps
GCP
Cilium
Kibana
Crossplane
Envoy
Azure
AWS
Kubernetes
DNS
Apply
Category Manager 11 hours ago
≈ $37k – $77k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Bengaluru • Istanbul
Apply
$73k per year • Remote (Spain) • Full-Time • Spain
Apply
Dev/Ops Developer 1 day ago
≈ $25k – $59k per year (Estimated) • Remote (Spain) • Full-Time • 3+ years exp • Spain
JavaScript
SQL
Frontend
JQuery
DevOps
Incident Management
Management
Agile
Scrum
ITIL
Apply
Remote (Spain) • Full-Time • 2+ years exp • Spain
Python
SQL
Python
pySpark
Databases
Snowflake
AI/ML
Spark
dbt
DevOps
AWS CDK
CloudFormation
CI/CD
Jenkins
AWS
Amazon S3
AWS Step Functions
Management
ITIL
Apply
See all jobs
This is one of many
1,061,370 more open roles from verified company boards, updated every day.