368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$215k – $260k per year
Location
In office (San Francisco)
Seniority
Staff · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Crusoe (formerly Crusoe Energy Systems) is an energy-first AI infrastructure and cloud computing company headquartered in Denver, Colorado. The company specializes in building and operating high-performance AI data centers powered by stranded, wasted, or underutilized energy sources - such as flared natural gas from oil fields and surplus renewable power.

Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack - from electrons to tokens - to power the world's most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.

We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that - with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.

We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.

If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.

About the Role:

We are seeking a Staff Software Engineer to lead distributed systems design and development within Crusoe Cloud's Cloud Monitoring Service. This team builds the observability backbone of Crusoe Cloud: the telemetry systems that collect, process, store, and serve the metrics and logs our customers depend on to run AI workloads at scale. You will own the design and evolution of high-throughput, multi-tenant data systems, from collection at the edge through ingestion, storage, and query. This is a full-time position.

What You'll Be Working On:

  • Distributed Systems Ownership: Own the architecture and evolution of large-scale telemetry pipelines, including high-volume ingestion, stream processing, time-series and log storage, and low-latency query paths. Design systems that stay correct, available, and cost-efficient as data volume grows 10x.

  • Scalable, Multi-Tenant Design: Design services that are highly scalable, durable, and fair across tenants. Solve the hard problems in this space: hot shards, high-cardinality data, noisy neighbors, backpressure, retention and compaction at scale, and graceful degradation under load.

  • Reliability and Operational Excellence: Build for operability from day one. Improve pipeline reliability and data freshness, reduce on-call burden through better system design rather than more process, and participate in a customer-facing on-call rotation, leading by example.

  • Technical Leadership: Set the technical direction for the team's distributed systems work. Drive design reviews, identify one-way door decisions early, and raise the bar on how the team scopes, builds, and operates systems.

  • Cross-Team Collaboration: Work with product, compute, networking, and platform teams to make sure observability decisions are made with full context. Represent the team's technical position in cross-org conversations.

  • Mentorship: Coach senior and mid-level engineers through design work, code review, and incident response. Build patterns and frameworks that make the team better without requiring your direct involvement.

What You'll Bring to the Team:

  • Distributed Systems Depth: Deep, hands-on experience designing and operating distributed systems at scale. You have solved real problems in sharding, replication, consistency, load balancing, and concurrency, not just studied them.

  • Observability Data Experience: Experience building or operating large-scale data infrastructure such as time-series databases, log aggregation, streaming pipelines, or distributed tracing backends. Familiarity with technologies like Prometheus, VictoriaMetrics, Loki, OpenTelemetry, Kafka, Vector, or similar.

  • Technical Proficiency: Strong programming fundamentals in Go or another modern compiled language (Go strongly preferred). Comfort with Kubernetes, microservices, and CI/CD as the operating environment for everything you build.

  • Design Judgment at Staff Level: You proactively scope ambiguous problems, surface non-functional requirements without being prompted, and reason about tradeoffs in terms of customer and business outcomes, not just technical elegance.

  • Operational Mindset: On-call experience on a customer-facing service. You treat incidents as design feedback and systematically eliminate the conditions that cause them.

  • Cross-Functional Influence: A track record of driving technical outcomes that span team boundaries, and the communication skills to bring product, support, and adjacent engineering teams along.

  • Mentorship: You make the engineers around you better through design guidance, thoughtful review, and honest, constructive feedback.

  • Professional Experience: 8+ years of software development experience, with sustained ownership of production distributed systems.

Benefits:

  • Competitive compensation and equity packages

  • Restricted Stock Units

  • Paid time off, paid holidays & leave of absence programs

  • Comprehensive health, dental & vision insurance

  • Employer contributions to HSA account

  • Paid parental leave

  • Paid life insurance, short-term and long-term disability

  • Professional development & tuition reimbursement

  • Mental health & wellness support

  • Commuter benefits (parking & transit)

  • Cell phone stipend

  • 401(k) Retirement plan with company match up to 4% of salary

  • Volunteer time off

  • Global travel insurance & emergency assistance

  • Daily meals allowance

  • Additional perks & programs specific to location

Compensation Range

Compensation will be paid in the range of up to $215,000 -$260,000 + Bonus. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant's knowledge, education, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$19k – $47k per year (Estimated) • Remote • Full-Time • Perm
C#
C#
ASP.NET Core
Dapper
Entity Framework Core
Databases
Apache Kafka
ClickHouse
ElasticSearch
PostgreSQL
RabbitMQ
Redis
DevOps
CI/CD
Docker
Docker Compose
GitHub Actions
Kubernetes
TeamCity
GitHub
GitLab
Apply
$19k – $47k per year (Estimated) • Remote • Full-Time • Tomsk
C#
C#
ASP.NET Core
Dapper
Entity Framework Core
Databases
Apache Kafka
ClickHouse
ElasticSearch
PostgreSQL
RabbitMQ
Redis
DevOps
CI/CD
Docker
Docker Compose
GitHub Actions
Kubernetes
TeamCity
GitHub
GitLab
Apply
Cloud Engineer (AWS) 5 hours ago
In office • Full-Time • 5+ years exp • Bachelor's Degree • Dalian
Python
DevOps
AWS
CI/CD
CloudFormation
Docker
FinOps
GitHub Actions
Kubernetes
Terraform
GitHub
Apply
$20k – $50k per year (Estimated) • In office • Full-Time • Tomsk
C++
Go
Java
Kotlin
Databases
Apache Kafka
ClickHouse
ElasticSearch
DevOps
Jaeger
Kubernetes
Prometheus
SLI/SLO/SLA
Apply
$20k – $50k per year (Estimated) • In office • Full-Time • Perm
C++
Go
Java
Kotlin
Databases
Apache Kafka
ClickHouse
ElasticSearch
DevOps
Jaeger
Kubernetes
Prometheus
SLI/SLO/SLA
Apply
$200k – $240k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • San Francisco • Sunnyvale
Apply
$140k – $160k per year • Equity • In office • Full-Time • Bachelor's Degree • San Francisco
C++
Java
Rust
DevOps
Ansible
Chef
Puppet
Terraform
Apply
$215k – $260k per year • Equity • In office • Full-Time • 8+ years exp • San Francisco
Go
Python
DevOps
AWS
CI/CD
GCP
Kubernetes
Terraform
Cybersecurity
Wiz
Zero Trust
Apply
$225k – $255k per year • Equity • In office • Full-Time • 15+ years exp • Bachelor's Degree • San Francisco
SQL
Apex
Apex
MuleSoft
AI/ML
LLM
RAG
AI Agents
Model Context Protocol
DevOps
AWS
Azure
GCP
Self-Healing
IAM
Cybersecurity
SOC 2
QA
Postman
Apply
$285k – $335k per year • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree • Sunnyvale • San Francisco
AI/ML
InfiniBand
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
$160k – $283k per year • Equity • In office • 5+ years exp • San Francisco
AI/ML
AI Agents
Apply
$83k – $188k per year (Estimated) • In office • 2+ years exp • San Francisco
Python
AI/ML
AI Agents
LLM Guardrails
Model Context Protocol
DevOps
Terraform
Cybersecurity
Crowdstrike
GDPR
Least Privilege
Okta
SentinelOne
Management
Google Workspace
Slack
Apply
$171k – $273k per year • In office • Full-Time • 8+ years exp • PhD • San Francisco • Washington
AI/ML
A2A
Agentforce
AI Agents
Model Context Protocol
DevOps
AWS
GCP
Marketing
Salesforce
Apply
Security GRC Analyst 2 hours ago
$119k – $268k per year (Estimated) • Remote/Hybrid • 4+ years exp • Bachelor's Degree • San Francisco
AI/ML
Ignite
PyTorch
Cybersecurity
ISO 27001
NIST CSF
SOC 2
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.