368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$245k – $295k per year
Location
In office (San Francisco, Sunnyvale)
Seniority
Senior · 10+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Crusoe (formerly Crusoe Energy Systems) is an energy-first AI infrastructure and cloud computing company headquartered in Denver, Colorado. The company specializes in building and operating high-performance AI data centers powered by stranded, wasted, or underutilized energy sources - such as flared natural gas from oil fields and surplus renewable power.

Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack - from electrons to tokens - to power the world's most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.

We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that - with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.

We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.

If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.

About the Role

We are seeking a Senior Manager, Infrastructure Platform Engineering to lead a team building core systems that turn large-scale compute infrastructure into reliable, secure, and efficiently allocatable capacity. The team owns foundational services spanning resource pooling and allocation, capacity and utilization intelligence, fleet and system lifecycle management, and platform security and trust.

This is a hands-on management role for a leader who has come up through infrastructure and systems software engineering, understands the realities of operating compute at scale across cloud and on-premise environments, and is energized by building the control and platform systems that other engineering teams depend on. You'll lead a growing team of infrastructure software engineers, set technical direction across the platform, and partner closely with adjacent infrastructure, production engineering, and security teams to keep the substrate reliable, well-utilized, and easy to build on.

While this is an infrastructure-focused role rather than a traditional product role, the systems this team builds are essential to the experience our customers have on the platform. Reliable capacity, healthy systems, and a trustworthy substrate are what make a seamless, dependable customer experience possible - so the team's work directly underpins the business, even as its immediate users are the internal engineering teams building and operating workloads on top of it.

What You'll Be Working On

  • Leading the team responsible for the platform services that abstract underlying infrastructure into reliable, allocatable capacity, and for the systems that track and reconcile state across a large fleet

  • Setting the technical roadmap across capacity and utilization intelligence, resource lifecycle and state management, and platform security and trust frameworks

  • Driving the design of secure, well-instrumented platform systems - from Kubernetes-based orchestration and automation to lower-level system and hardware integration

  • Hiring, mentoring, and growing a team of infrastructure software engineers; building a high-performing organization from a strong foundation

  • Partnering with infrastructure, production engineering, and security teams to align platform capabilities with operational reliability, capacity, and trust requirements

  • Improving platform efficiency and availability - characterizing bottlenecks, reducing stranded resources, and shortening operational and recovery cycles

  • Establishing engineering standards for infrastructure software development: code quality, testing, deployment safety, and on-call practices for systems that span the platform

  • Translating a vertically integrated infrastructure stack into reliable platform primitives that engineering teams can build on

  • Staying technically hands-on - reviewing designs, contributing to architecture decisions, and being credible to the engineers you lead

What You'll Bring to the Team

  • 10+ years of experience in infrastructure or systems software development, with at least 3+ years in an engineering leadership role

  • Deep expertise in large-scale infrastructure platforms - building services that pool, allocate, and reconcile compute resources at scale

  • Strong background with Kubernetes and cloud platforms (GCP, AWS, or Azure) - orchestration, automation, and operating distributed systems in production

  • Experience with distributed state management and control systems - modeling resource and system lifecycle, reconciling desired vs. actual state, and handling failure gracefully across a large fleet

  • Experience with efficiency, capacity, or performance engineering - characterizing system behavior, identifying bottlenecks, and driving measurable improvements in utilization or availability

  • A player-coach approach to management: hands-on enough to make technical calls, structured enough to grow a team and ship through them

  • Track record of hiring strong infrastructure engineers and helping them grow into more senior roles

  • Comfortable operating in a fast-moving environment where the path isn't fully paved - willing to drive ambiguity to clarity

Bonus Points

  • Experience operating Kubernetes on bare-metal infrastructure as well as on managed cloud services (GKE, EKS, AKS)

  • Familiarity with the operational challenges of GPU clusters, AI training, and inference workloads

  • Working knowledge of platform security and trust concepts - secure boot, measured boot, TPMs, and hardware attestation

  • Experience with capacity forecasting, demand modeling, or allocation optimization at scale

  • Hands-on background with telemetry and observability platforms at scale (Prometheus, OpenTelemetry, Grafana)

  • Prior experience building infrastructure platforms at hyperscalers or cloud providers where internal engineers are the primary customer

  • Familiarity with hardware-software co-design - understanding how platform choices affect physical infrastructure utilization

Benefits

  • Competitive compensation and equity packages

  • Restricted Stock Units

  • Paid time off, paid holidays & leave of absence programs

  • Comprehensive health, dental & vision insurance

  • Employer contributions to HSA account

  • Paid parental leave

  • Paid life insurance, short-term and long-term disability

  • Professional development & tuition reimbursement

  • Mental health & wellness support

  • Commuter benefits (parking & transit)

  • Cell phone stipend

  • 401(k) Retirement plan with company match up to 4% of salary

  • Volunteer time off

  • Global travel insurance & emergency assistance

  • Daily meals allowance

  • Additional perks & programs specific to location

Compensation Range

Compensation will be paid in the range of up to $245,000 - $295,000 + Bonus. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant's knowledge, education, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$220k – $325k per year • Remote/Hybrid • Full-Time • 15+ years exp • New York
DevOps
AWS
Azure
CI/CD
GCP
Kubernetes
Platform Engineering
Cybersecurity
Threat Modeling
Apply
Remote/Hybrid • 7+ years exp • Bachelor's Degree
C#
JavaScript
Python
SQL
TypeScript
C#
ASP.NET Core
Blazor
Dapper
Databases
Oracle
Frontend
Angular
Bootstrap
JQuery
Vue.js
DevOps
AWS
Azure
Azure DevOps
CI/CD
Git
Apply
In office • Contractor • Dublin
DevOps
Azure
Cybersecurity
Okta
Management
Google Workspace
Jira
Slack
Apply
$21k – $57k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Noida
C#
SQL
TypeScript
JavaScript
C#
.NET
Frontend
Angular
DevOps
Azure
Azure DevOps
CI/CD
Git
TeamCity
Apply
$16k – $42k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Noida
C#
SQL
TypeScript
JavaScript
C#
.NET
Frontend
Angular
DevOps
Azure
Azure DevOps
CI/CD
Git
TeamCity
Apply
$200k – $240k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • San Francisco • Sunnyvale
Apply
$140k – $160k per year • Equity • In office • Full-Time • Bachelor's Degree • San Francisco
C++
Java
Rust
DevOps
Ansible
Chef
Puppet
Terraform
Apply
$215k – $260k per year • Equity • In office • Full-Time • 8+ years exp • San Francisco
Go
Python
DevOps
AWS
CI/CD
GCP
Kubernetes
Terraform
Cybersecurity
Wiz
Zero Trust
Apply
$225k – $255k per year • Equity • In office • Full-Time • 15+ years exp • Bachelor's Degree • San Francisco
SQL
Apex
Apex
MuleSoft
AI/ML
LLM
RAG
AI Agents
Model Context Protocol
DevOps
AWS
Azure
GCP
Self-Healing
IAM
Cybersecurity
SOC 2
QA
Postman
Apply
$285k – $335k per year • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree • Sunnyvale • San Francisco
AI/ML
InfiniBand
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Senior ML Engineer 29 min ago
$149k – $224k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Francisco • Washington • Palo Alto
Python
Python
pySpark
Databases
Apache Kafka
AI/ML
AI Agents
Agentforce
Airflow
Anomaly Detection
Feature Store
Flink
Ray
Red Teaming
Spark
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
MITRE ATT&CK
Marketing
Salesforce
Apply
In office • Internship • 1+ year exp • Bachelor's Degree • San Francisco
Go
JavaScript
Ruby
Scala
Apply
$360k – $530k per year • In office • Full-Time • Bachelor's Degree • San Francisco
MATLAB
Python
MATLAB
Simulink
AI/ML
OpenAI
Robotics
Digital Twin
Apply
$222k – $277k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Francisco
DevOps
CI/CD
Immutable Infrastructure
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.