1,059,896open jobs
62,082companies
176,030added this week
Browse all
Salary
$165k – $200k per year
Location
In office (San Jose)
Seniority
Staff · 12+ years exp

Confirmed on the employer's own hiring board on Oct 1, 2026. First seen by Alion on Sep 30, 2026.

Overview
Company
Impact
Profile match
Supermicro designs and builds servers, storage and rack-scale systems. Its building block architecture allows rapid configuration for artificial intelligence and cloud workloads. The company assembles much of its hardware in Silicon Valley.

Job Req ID: 29832

About Supermicro:

Supermicro® is a Top Tier provider of advanced server, storage, and networking solutions for Data Center, Cloud Computing, Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon Valley Top 50 technology firms. Our unprecedented global expansion has provided us with the opportunity to offer a large number of new positions to the technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.

Job Summary:

As Staff Data Center Solutions Engineer you will be the primary technical point of contact from the Data Center side for customers planning, deploying, and validating large-scale GPU clusters - at customer sites, in Supermicro-owned facilities, and in partner colocation environments. This is a fulfillment role serving customers directly, day one as an individual contributor, with a clear path to leading a small team as the group scales.

This role suits a Staff Data Center Solutions Engineer who wants to grow into broad GPUaaS solution engineering while advising customers on Supermicro's DCBBS liquid-cooled (DLC) and air-cooled product lines - CDUs, Sidecars, cooling towers, and chillers. You don't need to be a professional software developer, but you should understand how software is built well enough to guide internal tooling work and speak credibly to technical customers and internal engineering teams alike. Strong presentation and communication skills are essential - you'll regularly be the person customers trust in the room.

Essential Duties and Responsibilities:

Drive DataCenter Systems Planning & Deployment: Work directly with customers to plan, deploy, and validate GPU cluster deployments, coordinating infrastructure build-out, rack power-up, network fabric bring-up, cooling commissioning, and system validation.

Customer Advisory on Cooling & Power Products: Advise customers on Supermicro's DCBBS DLC and air-cooled product lines - CDUs, Sidecars, cooling towers, and chillers - helping them select and configure the right solution for their deployment.

BMS & Infrastructure Monitoring Integration: Connect and configure Building Management System (BMS) monitoring integrations alongside datacenter infrastructure monitoring tools (power, cooling, environmental) to ensure clusters are deployed into healthy, well-instrumented environments.

AI/NeoCloud & GPUaaS Expertise: Apply strong working knowledge of NeoCloud infrastructure and GPUaaS platforms, including Agentic AI tooling, to guide deployment decisions, automate diagnostics, and troubleshoot issues unique to AI infrastructure.

Internal Tooling Direction: Partner with engineering to scope and direct internal tooling and automation using Agentic Coding tools, applying software development understanding without requiring hands-on coding as a core function.

Cross-Team Technical Leadership: Diagnose complex, multi-disciplinary issues spanning power, cooling, networking, and compute, and lead (often dotted-line) teams through resolution during critical deployment windows.

Customer-Facing Delivery: Serve as the primary technical point of contact for customers throughout planning, deployment, and validation, managing expectations and communicating status/issues clearly.

Partner Datacenter Coordination: Work with partner colocation facility teams to align on power, cooling, and space readiness ahead of and during deployments.

AI Factory Service Strategy: Contribute to strategies and improvements for Supermicro's end-to-end AI Factory deployment services, drawing on direct customer fulfillment experience to identify gaps and opportunities.

Documentation & Runbooks: Develop and maintain installation runbooks, deployment checklists, and troubleshooting playbooks to support repeatable, high-quality deployments as the team scales.

Qualifications:

  • Bachelor's degree in Electrical Engineering, Computer Engineering, Mechanical Engineering, Computer Science, or a related field-or equivalent experience-with 12+ years of experience in datacenter infrastructure or GPU cluster deployment.
  • Strong understanding of AI/NeoCloud infrastructure architectures and GPUaaS platforms.
  • Strong understanding of how software is built; able to direct and evaluate internal tooling work without necessarily writing production code.
  • Practical experience using Agentic AI tools and workflows for diagnostics or automation.
  • Hands-on experience planning, deploying, and validating GPU compute clusters (NVIDIA/CUDA preferred).
  • Working knowledge of BMS integration and datacenter infrastructure monitoring tools (e.g., Hyperview, Grafana, PagerDuty, or similar).
  • Solid understanding of datacenter cooling and power systems - liquid cooling (DLC), CDUs, Sidecars, cooling towers, chillers, and three-phase power distribution.
  • Excellent presentation and communication skills; comfortable and credible presenting technical recommendations directly to customers.
  • Demonstrated ability to lead cross-functional technical teams without formal authority (dotted-line leadership).

Preferred Qualifications:

  • Background in Electrical or Computer Engineering.
  • Prior experience deploying GPU clusters in colocation or multi-tenant partner datacenter environments.
  • Familiarity with high-density rack architectures (OCP, MGX, or similar).
  • Understanding of network fabrics used in GPU clusters (InfiniBand, RoCEv2).
  • Experience with ServiceNow CSM or similar customer-facing IT service platforms.
  • Experience building or directing agentic AI/coding workflows for infrastructure diagnostics or operations automation.
  • Interest or experience in contributing to service strategy or process improvement initiatives.

Salary Range

$165,000 - $200,000

The salary offered will depend on several factors, including your location, level, education, training, specific skills, years of experience, and comparison to other employees already in this role. In addition to a comprehensive benefits package, candidates may be eligible for other forms of compensation, such as participation in bonus and equity award programs.

EEO Statement

Supermicro is an Equal Opportunity Employer and embraces diversity in our employee population. It is the policy of Supermicro to provide equal opportunity to all qualified applicants and employees without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, protected veteran status or special disabled veteran, marital status, pregnancy, genetic information, or any other legally protected status.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,059,896 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
San Jose
≈ $108k – $224k per year (Estimated) • In office • Secret • Full-Time • 5+ years exp • White Sands
Python
PowerShell
Bash
DevOps
Red Hat
VMWare
Linux
Windows
Cybersecurity
Active Directory
Apply
$96k – $182k per year • In office • Secret • Full-Time • 5+ years exp • El Segundo
DevOps
Splunk
VMWare
Hyper-V
Windows
DNS
DHCP
Cybersecurity
Nessus
Active Directory
Apply
$157k – $215k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Louisville • Centennial
Python
C++
Bash
DevOps
Terraform
Ansible
GCP
Prometheus
Azure
CI/CD
GitOps
AWS
Docker
Kubernetes
Grafana
GitHub
GitLab
Linux
Windows
Management
Agile
Scrum
Kanban
Apply
≈ $103k – $201k per year (Estimated) • In office • 15+ years exp • Bachelor's Degree • Phoenix
Apply
$124k – $170k per year • In office • Secret • 12+ years exp • Bachelor's Degree • Bedford
Design
SolidWorks
Management
SharePoint
Microsoft Office
Apply
ABAP AI Developer - US 11 hours ago
≈ $152k – $287k per year (Estimated) • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • New York
JavaScript
Java
Kotlin
ABAP
AI/ML
AI Agents
LLM
Frontend
React.js
Vite
Apply
In office • Full-Time • 3+ years exp • Johannesburg
Python
Bash
Databases
PostgreSQL
DevOps
Terraform
Ansible
GCP
Prometheus
CI/CD
AWS
Docker
Ubuntu
Grafana
Linux
DNS
Apply
≈ $48k – $131k per year (Estimated) • In office • Full-Time • Cork
AI/ML
AI Agents
Apply
≈ $162k – $267k per year (Estimated) • Remote (United States) • 5+ years exp
SQL
Databases
MS SQL
AI/ML
AI Agents
Agentforce
DevOps
AWS
GitHub
Analytics
Pentaho
Management
Confluence
Jira
Agile
Apply
Senior DevOps Engineer 11 hours ago
≈ $41k – $107k per year (Estimated) • In office • Full-Time • Gerakas
Python
Go
Java
PHP
Databases
PostgreSQL
Redis
Cassandra
Couchbase
ElasticSearch
DevOps
Terraform
Ansible
Helm
Loki
VMWare
Prometheus
GitLab CI
Azure
CI/CD
Jenkins
AWS
Docker
Kubernetes
Grafana
GitLab
Linux
Management
Agile
Apply
$150k – $165k per year • In office • 8+ years exp • Master's Degree • San Jose
Python
AI/ML
Hadoop
DevOps
HPC
Linux
Apply
$90k – $115k per year • Equity • In office • 3+ years exp • Bachelor's Degree • San Jose
Python
AI/ML
Hadoop
DevOps
HPC
Apply
Sr. System Engineer 3 days ago
$137k – $156k per year • Equity • In office • 8+ years exp • Bachelor's Degree • San Jose
Python
AI/ML
Hadoop
Fine-tuning
TensorFlow
PyTorch
LLM
NCCL
DevOps
Azure
AWS
Docker
Kubernetes
OpenStack
HPC
TCP/IP
DNS
DHCP
Apply
$75k – $100k per year • Equity • In office • 2+ years exp • Bachelor's Degree • San Jose
AI/ML
Hadoop
DevOps
HPC
Apply
$150k – $185k per year • Equity • In office • 8+ years exp • Bachelor's Degree • San Jose
AI/ML
Hadoop
DevOps
HPC
Windows
Apply
$150k – $165k per year • In office • 8+ years exp • Master's Degree • San Jose
Python
AI/ML
Hadoop
DevOps
HPC
Linux
Apply
$138k – $226k per year • Equity • Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Milpitas • San Jose • San Francisco
AI/ML
LLM
RAG
LLM Evaluation
Agentic Workflows
Management
Jira
Agile
Apply
$243k – $364k per year • Equity • Remote (United States) • Full-Time • 15+ years exp • Bachelor's Degree • San Jose • Austin • Seattle
Management
Agile
Apply
$126k – $203k per year • Equity • Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • San Jose
Python
Perl
Chips/EDA
Synopsys Design Compiler
Cadence Innovus
Synopsys PrimeTime
Apply
$44k – $185k per year • Equity • Hybrid • Full-Time • Bachelor's Degree • San Jose
Python
Chips/EDA
Cadence Allegro
Apply
See all jobs
This is one of many
1,059,896 more open roles from verified company boards, updated every day.