368,941open jobs
9,452companies
47,951added this week
Browse all
Salary
$202k – $241k per year
Location
In office (Glendale)
Employment
Full-Time
Overview
Company
Impact
Profile match
Fluidstack is an AI cloud platform that designs, builds, and operates high-performance GPU clusters for frontier AI laboratories, enterprises, and governments. The company provides enterprise-grade bare-metal compute infrastructure - scaled across tens of thousands of state-of-the-art NVIDIA GPUs - specifically optimized for training large language models (LLMs) and running high-throughput inference.

About Fluidstack

We exist to make humanity more free. For most of human history, you farmed or you starved. Technology gave people more time for the things they wanted to do, instead of things they had to do. Powerful AI will be the biggest lever for human choice we've ever built - but only if models are aligned with what humanity actually wants. There are groups building AI who don't share these goals. Whoever deploys frontier compute infrastructure fastest will decide whether AI expands human freedom or shrinks it.

We're singularly focused on delivering 10 to 100s of GWs of compute faster than anyone else, rethinking every layer of the stack. We acquire power, design and build data centers, and operate them - with teams spanning hardware and software. Speed and scale are our key differentiators. Come be a part of building civilization-scale infrastructure for AI.

We hire people who care deeply about this problem space. If that is you, please apply!

How We Operate

  • Extreme ownership. Full autonomy. Own things end to end often taking on scope outside your core role without being asked to get things done.

  • Velocity. We drive everything forward as fast as possible.

  • First principles. Challenge every assumption. Zero analogy thinking, no egos, the best idea wins.

  • Love of the game. The frontier of AI is the most interesting problem of our time. We put in long hours at high intensity to push the frontier forward.

The Fluidstack Labs Team

Examples of key problems the team is working on

  • Qualify the hardware the frontier runs on before it runs anywhere else. First samples of next-generation accelerators, switches, storage, and liquid cooling land here, and leave as production-ready platforms with runbooks the whole fleet inherits.

  • Compress silicon-to-production to weeks. The lab closes the gap between vendor sample and customer-ready gigawatt infrastructure, and every week cut here pulls the entire 10 GW deployment curve forward.

  • Run the lab like a production site. Provisioning, telemetry, demand management, and liquid cooling mirror production architecture exactly, so a qualification pass in the lab is a deployment guarantee in the field.

  • Prove the power envelope nobody else will touch. Dynamic demand management lets AI compute deploy beyond nominal electrical capacity, and the lab validates the full detection-to-shutdown response chain that makes it safe.

Role Scope

  • Bring up first-sample accelerator platforms (NVIDIA, AMD, custom accelerators) end to end: rack integration, liquid cooling commissioning, firmware baseline establishment, network connectivity, and software stack validation at rack densities up to and beyond 120 kW.

  • Validate network platforms across Broadcom Tomahawk, Broadcom Jericho, and NVIDIA Spectrum silicon, driving Keysight Ixia traffic generation for RFC 2544/2889 benchmarking, line-rate stress, and protocol correctness.

  • Verify the optical layer with EXFO test equipment, covering BER characterization and power budget analysis across the link inventory a qualification depends on.

  • Qualify CDUs and liquid cooling across nVent, CoolIT, and Vertiv platforms: commissioning procedures, BMS telemetry integration, leak detection validation, coolant chemistry verification, and the operational runbooks that fall out of each pass.

  • Exercise the full power oversubscription response chain: graceful and forced shutdown paths, power-cap and p-state levers via BMC, ATS transfer scenarios, breaker-trip detection, rPDU commissioning with outlet-level telemetry, and repeatable power-virus stress harnesses against accelerator hardware.

  • Co-develop qualification matrices with hardware partners, evaluate converged local-NVMe storage platforms (Weka, Hammerspace, VAST Data), and turn results into runbooks production teams inherit.

What We're Looking For

The below is a starting point. We always make space for exceptional people, so if you don't fit this role exactly,tell us where you would.

  • You've personally brought up servers, accelerators, or switches from first power-on: racking, cabling, firmware, first boot, and everything that goes wrong in between.

  • You script your test harnesses and automation rather than clicking through runs, so your results are repeatable and your throughput compounds.

  • You're deep in Linux and manage hardware at the BMC level: IPMI, Redfish, sensor telemetry, and power and boot control from the command line.

  • You document methodically, producing test reports and runbooks that someone else can execute without you in the room.

  • You've worked hands-on with liquid-cooled, high-power hardware and treat its safety procedures as part of the craft.

  • You debug across hardware, firmware, and software layers yourself, without waiting for someone else to isolate the fault.

  • Bonus: RDMA/RoCEv2 fabrics. Optics testing depth. Kubernetes-based provisioning stacks. Power and electrical instrumentation experience.

We are committed to pay equity and transparency.

Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

You will receive a confirmation email once your application has successfully been accepted. If there is an error with your submission and you did not receive a confirmation email, please email [email protected] with your resume/CV, the role you've applied for, and the date you submitted your application-- someone from our recruiting team will be in touch.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,941 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Glendale
Java Developer 1 day ago
$31k – $54k per year (Estimated) • Remote • 5+ years exp • Tula
C#
JavaScript
TypeScript
Java
C#
.NET
Java
Spring Boot
Databases
Apache Kafka
ElasticSearch
PostgreSQL
RabbitMQ
Frontend
Angular
Bootstrap
React.js
DevOps
Docker
Git
Jenkins
Kubernetes
Rest API
GitLab
QA
Swagger
Apply
$27k – $58k per year (Estimated) • Remote/Hybrid • 3+ years exp • Bachelor's Degree • Moscow
C#
JavaScript
SQL
TypeScript
C#
ASP.NET Core
Databases
Apache Kafka
MS SQL
Frontend
Angular
React.js
DevOps
Docker
Kubernetes
Apply
$19k – $53k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Pune
Bash
JavaScript
Python
TypeScript
Frontend
Angular
React.js
DevOps
AWS
Azure
Datadog
Docker
GCP
Grafana
Kubernetes
Prometheus
Splunk
IAM
Cybersecurity
Keycloak
Apply
$77k – $194k per year (Estimated) • In office • Contractor • 5+ years exp • Singapore
Java
SQL
DevOps
CI/CD
Docker
Grafana
Incident Management
Kubernetes
OpenShift
SRE
Apply
$131k – $281k per year (Estimated) • In office • Contractor • 15+ years exp • Singapore
Java
SQL
Java
Maven
Databases
MS SQL
Oracle
DevOps
AWS
CI/CD
Incident Management
Kubernetes
OpenShift
OpenStack
Apply
$173k – $250k per year • In office • Full-Time • San Francisco • New York • Seattle • Austin
TypeScript
Databases
PostgreSQL
AI/ML
AI Agents
Claude
Claude Code
Cursor
LLM
LLM Guardrails
Model Context Protocol
Apply
$224k – $279k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • San Francisco • New York • Seattle • Austin
Python
JavaScript
Databases
PostgreSQL
Redis
AI/ML
Time Series Forecasting
Frontend
Bootstrap
DevOps
Ansible
CI/CD
Docker
Grafana
Incident Management
OpenTelemetry
Platform Engineering
Prometheus
Terraform
Robotics
Digital Twin
Apply
$258k – $300k per year • Remote • Full-Time
Python
SQL
Apply
$173k – $279k per year • In office • Full-Time • San Francisco • New York • Seattle • Austin
Python
AI/ML
Time Series Forecasting
IoT
OPC UA
Apply
$224k – $264k per year • In office • Full-Time • Austin • New York • San Francisco • Seattle
Python
SQL
Apply
$152k – $178k per year • In office • Full-Time • Glendale
Bash
Python
Apply
$120k – $122k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Glendale
Python
Databases
Amazon Redshift
Game Dev
Houdini
Design
Adobe After Effects
Adobe Photoshop
Blender
Maya
Apply
$117k – $240k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Glendale
C#
JavaScript
SQL
Frontend
Bootstrap
Apply
$115k – $229k per year (Estimated) • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • Boca Raton • Glendale
ABAP
ABAP
SAP Fiori
Apply
$98k – $131k per year • In office • Full-Time • Bachelor's Degree • Glendale
Java
Python
Scala
SQL
Databases
Apache Kafka
Databricks
Presto
Snowflake
AI/ML
Flink
Spark
DevOps
AWS
Docker
Kubernetes
Terraform
Amazon Kinesis
Apply
See all jobs
This is one of many
368,941 more open roles from verified company boards, updated every day.