368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$180k – $339k per year
Location
In office (Santa Clara)
Seniority
Senior · 10+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Jobs for Humanity is a global movement and recruitment platform dedicated to connecting underrepresented talent with inclusive employers committed to fair-chance and diversity hiring. The platform operates tailored job boards and career programs for specific groups, including neurodivergent individuals, refugees, blind or visually impaired job seekers, single parents, Black leaders, returning citizens (formerly incarcerated individuals), and LGBTQIA+ professionals. By providing job coaching, automated applicant tracking system (ATS) integrations, and diversity, equity, and inclusion (DEI) training for employers, the organization aims to reduce systemic employment barriers and foster equal workplace opportunities worldwide.
Jobs for Humanity is collaborating with Upwardly Global and with Nvidia to build an inclusive and just employment ecosystem. We support individuals coming from all walks of life.

Company Name: Nvidia

Senior System Software Engineer - Scientific Computing PaaS

locations: US, CA, Santa Clara; US, Remote

time type: Full time

posted on: Posted Today

job requisition id: JR1979896

We are seeking a Sr System Software Engineer to help us build out our scientific computing platform on Nvidia DGX Cloud. We are building a cloud-based accelerated scientific computing platform as a service on the Nvidia DGX cloud. This DGX scientific computing cloud platform enables Physics-based Numerical Simulation Solvers, AI-based Training, Inference, and Visualization workflow for physical science and engineering problems. Those applications include Weather prediction, Climate modeling, Industrial design, and Digital twins simulation in various domains e.g. Aerospace, Automotive, Sports, Renewable energy, Bio-medical, and many more.

Are you passionate about solving rewarding problems at scale? Do you enjoy crafting robust, critical services for compute and data-intensive workloads? If so, you may be a phenomenal fit for our team!

What you'll be doing:

Design, Build, Deploy, and Operate Cloud-native microservices and APIs for scientific computing workload on DGX cloud.

Design services and take ownership of underlying cloud infrastructure for physics-informed and data-driven scientific workflows.

Design novel algorithms and actively engage with operations to increase overall system performance, it spans across the stack e.g. deep understanding of application code e.g. DL Framework, Numerical Solvers, Microservices, APIs, and Heterogeneous accelerated computing with CPUs and GPUs.

Design, Build, Deploy, and Operate scalable I/O infrastructure for checkpointing, data loading, pre & post-processing of data.

Optimize compute, storage, and network architecture specific to physics & simulation-driven applications.

What we need to see:

- BS/MS degree in Computer Science or related areas or equivalent experience.

- 10+ years experience working on building and operating distributed compute and data-intensive platform as a service on cloud

- Proven skill in a compiled language (Go, Rust, C++ or otherwise).

- Strong foundational knowledge in Cloud Computing e.g. "The Datacenter is a Computer" architecture, cloud security architecture, virtualization - CPU, Memory and IO, Resource pooling and elasticity.

- Proven skills in Distributed Systems & Parallel Processing e.g. System model of distributed computation e.g. topology abstraction, logical time. Synchronization and deadlock detection in distributed systems, Fault Tolerance and Failure Detection, Consensus and Agreement protocols, Parallel algorithms, shared memory and distributed memory architecture, message passing (MPI, NCCL), Cluster scalability and performance.

- Hands-on Debugging skills with Process, Threads, Deadlock and Synchronization, Scheduling, IPC, Memory management, File system, and I/O structure.

- Strong Evidence of Algorithmic Thinking & System Design skills e.g. Recursion, Graph, Tree, Stack, and Queue, Large scale loosely coupled distributed system design and operational experience.

- Be self-motivated, have strong interpersonal skills, and be able to work independently with multiple teams with minimal direction.

Ways to stand out from the crowd:

- Have built, deployed, and operated AI platforms on HPC clusters. Have built, deployed, and operated cloud-native system including distributed storage, scheduling, and orchestration among compute, storage, and network.

- Configuring and troubleshooting hardware, operating systems, kernels, compilers for maximum performance.

- Hands-on debugging skills to optimize performance of compute, networking, and I/O framework. Extensively worked on third-party source code for debugging and customization.

NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most forward-thinking and hardworking people on the planet working for us. If you're creative and autonomous, we want to hear from you!

The base salary range is 180,000 USD - 339,250 USD. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. You will also be eligible for equity and benefits.

NVIDIA accepts applications on an ongoing basis. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status, or any other characteristic protected by law.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$30k – $64k per year (Estimated) • Remote/Hybrid • Bachelor's Degree • Moscow
C++
Rust
SQL
Databases
PostgreSQL
Redis
DevOps
CI/CD
Git
Grafana
Kibana
Apply
$117k – $234k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Sunnyvale
C#
C++
Python
AI/ML
Computer Vision
OpenCV
Apply
$28k – $67k per year (Estimated) • In office • Full-Time • 7+ years exp • Bengaluru
C++
Python
DevOps
RTOS
Apply
$53k – $108k per year • In office • Internship • Bachelor's Degree • Colorado Springs
C++
JavaScript
Perl
Python
Apply
$53k – $108k per year • In office • Internship • Bachelor's Degree • Rome
C++
JavaScript
Perl
Python
Apply
$65k – $163k per year (Estimated) • Remote/Hybrid • Contractor • Glasgow
Java
Python
DevOps
CI/CD
Grafana
Helm
Kubernetes
Prometheus
Terraform
Apply
Python Developer 4 days ago
$75k – $169k per year (Estimated) • Remote/Hybrid • Contractor • Glasgow
Python
SQL
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
OpenShift
Rest API
Apply
$160k – $190k per year • Remote/Hybrid • Full-Time • 5+ years exp • New York
Python
SQL
DevOps
CI/CD
Apply
$160k – $190k per year • Remote/Hybrid • Full-Time • New York
Management
Confluence
Apply
UX/UI Designer 11 days ago
up to $48k per year • In office • Full-Time • Riyadh
Design
Figma
ProtoPie
Apply
$100k – $137k per year • Equity • In office • Full-Time • 2+ years exp • Bachelor's Degree • Santa Clara
Apply
$221k – $387k per year • Equity • In office • Full-Time • Bachelor's Degree • Santa Clara
AI/ML
AI Agents
Context Engineering
Embeddings
LLM Guardrails
Management
ServiceNow
Apply
Field CTO 3 hours ago
$346k – $485k per year • In office • Full-Time • 12+ years exp • Santa Clara
DevOps
AWS
Azure
GCP
Apply
$72k – $99k per year • Equity • In office • Full-Time • Santa Clara
Apply
$166k – $290k per year • Equity • In office • Full-Time • 8+ years exp • Santa Clara
Management
ServiceNow
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.