1,227,840open jobs
69,546companies
213,293added this week
Browse all
Salary
$125k – $160k per year
Location
In office (Hawthorne)
Seniority
Junior · 2+ years exp

Confirmed on the employer's own hiring board on Oct 5, 2026. First seen by Alion on Sep 23, 2026. SpaceX scores A on the Alion truth index.

Overview
Company
Impact
Profile match
SpaceX is an American aerospace manufacturer and launch services provider founded by Elon Musk in 2002 with the goal of making life multiplanetary. It developed the reusable Falcon 9, now by far the most flown orbital rocket in the world, the Dragon capsule that carries NASA astronauts to the International Space Station, and Starship, the fully reusable super-heavy vehicle intended for lunar and Mars missions. The company also operates Starlink, a satellite constellation of thousands of spacecraft providing broadband to consumers, aircraft, ships and militaries, which has become its largest source of revenue.

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.

SITE RELIABILITY ENGINEER (HIGH PERFORMANCE COMPUTING)

SpaceX HPC is a shared compute platform used across the company - vehicle and structures simulation, machine learning, AI inference, and more. We support every program at SpaceX to design and operate the worlds most advanced rockets and satellites. This role exists to put a real Site Reliability Engineer operating model on these capabilities and accelerating the world class engineering at SpaceX: toil reduction, automation, observability, and a sustainable incident process.

We are looking for a Site Reliability Engineer who wants to own everything from Linux machines and our Infrastructure as Code, storage, and user facing applications - the whole ecosystem as a product, not as a ticket queue. You do not need a prior HPC title. You do need production instincts - you have operated real infrastructure, you write code to delete toil, and you care about whether users can actually get work done, not just whether nodes ping. You’ll work alongside HPC systems engineers who design and commission clusters to help make them more reliable and provide world class services for world class engineers.

Aerospace experience is not required. We value engineers who treat teammates with fairness and respect, who are self-critical, and who will take ownership of hard production problems.

RESPONSIBILITIES:

  • Participate in the team's on-call rotation; practice sustainable incident response and blameless postmortems
  • Manage node lifecycle with infrastructure as code: OS images, firmware, configuration management, kernel and driver stack
  • Build observability for both HPC administrators and end users - cluster, node, and storage health for operators, and job/workflow-level signal for the people running work on the platform
  • Reduce toil with automation; split time between operating production systems and writing the software that makes that work smaller
  • Sustainably manage resources, including compute and storage
  • Lead capacity planning with users across the company: understand what they will need next, and turn that into a concrete picture of tomorrow's compute and storage
  • Collaborate with HPC systems engineers and with engineers across all disciplines across the company on operable, maintainable infrastructure

BASIC QUALIFICATIONS:

  • Bachelor's degree in computer science, engineering, math, or a scientific discipline; OR 2+ years of professional experience operating production infrastructure in lieu of a degree
  • 2+ years of experience with Linux operating systems in production
  • 2+ years of experience operating production infrastructure (servers, services, or networks), including monitoring, debugging, and repairing what you own

PREFERRED SKILLS AND EXPERIENCE:

  • 2+ years of professional experience in SRE, DevOps, or production infrastructure engineering
  • Experience with monitoring and alerting (Prometheus, Grafana, Nagios, or similar)
  • Experience deploying and maintaining configuration management or infrastructure as code (Ansible, Puppet, Terraform, or similar)
  • Experience writing scripts/code (eg. Python or similar languages) to automate common tasks
  • Experience with containers (Docker, Podman, Singularity/Apptainer)
  • Experience with Kubernetes administration for on-premise deployment
  • Experience with distributed or high-performance storage (VAST or similar), including capacity, performance, and lifecycle management
  • Familiarity with HPC clusters, schedulers (Slurm, PBS, LSF), or GPU compute - not required; we will teach this
  • Familiarity with scientific computing (CFD, FEA) and/or ML training workloads (PyTorch, TensorFlow, CUDA) and/or AI inference workloads
  • Good understanding of version control, testing, continuous integration, build, deployment and monitoring
  • Ability to communicate clearly with users, peers, and vendors in both incident and design settings
  • Comfortable working with mission-critical and sensitive systems, with a sense of urgency appropriate to the responsibilities
  • Eligibility for access to classified material up to TS/SCI with polygraph

ADDITIONAL REQUIREMENTS:

  • Position is based in Hawthorne, CA and is primarily on-site
  • Must be able to participate in an on-call rotation
  • Must be willing to work extended hours and weekends as needed for incidents, cluster bring-up, and time-critical failures

COMPENSATION AND BENEFITS:

Pay Range:

Level 1: $125,000.00 - $160,000.00

Level 2: $145,000.00 - $195,000.00

Your actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience.

Base salary is just one part of your total rewards package at SpaceX. You may also be eligible for long-term incentives, in the form of company stock or long-term cash awards, as well as potential discretionary bonuses and the ability to purchase additional stock at a discount through an Employee Stock Purchase Plan. You will also receive access to comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, and various other discounts and perks. You may also accrue 3 weeks of paid vacation and will be eligible for 10 or more paid holidays per year. Employees accrue paid sick leave pursuant to Company policy which satisfies or exceeds the accrual, carryover, and use requirements of the law.

ITAR REQUIREMENTS:

  • To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here.  

SpaceX is an Equal Opportunity Employer; employment with SpaceX is governed on the basis of merit, competence and qualifications and will not be influenced in any manner by race, color, religion, gender, national origin/ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability or any other legally protected status.

Applicants wishing to view a copy of SpaceX’s Affirmative Action Plan for veterans and individuals with disabilities, or applicants requiring reasonable accommodation to the application/interview process should reach out to  [email protected]. 

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,227,840 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Hawthorne
≈ $77k – $165k per year (Estimated) • In office • 1+ year exp • Bachelor's Degree • Brownsville
SQL
Apply
$100k – $130k per year • Equity • In office • 1+ year exp • Bachelor's Degree • Hawthorne
Python
AI/ML
Ray
Machine Learning
Apply
$50k – $70k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Medford
DevOps
Azure
Windows Server
Windows
Cybersecurity
Active Directory
Management
Microsoft Office
Apply
≈ $83k – $169k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Charleston
Design
SolidWorks
Apply
≈ $117k – $214k per year (Estimated) • Remote (United States) • Full-Time
Python
Bash
AI/ML
AI Agents
Machine Learning
DevOps
Terraform
CloudFormation
Prometheus
CI/CD
Jenkins
AWS
Docker
Kubernetes
Grafana
Bitbucket
AWS Lambda
Amazon EC2
Amazon S3
IAM
Amazon ECS
Amazon CloudWatch
Linux
Cybersecurity
SOC 2
Management
Agile
Apply
Presales Engineer 1 day ago
≈ $66k – $170k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Cairo
AI/ML
Machine Learning
Management
Agile
Scrum
Apply
$28k – $36k per year • In office • Bachelor's Degree • Mississauga
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Spark
MLFlow
Scikit-learn
NumPy
Time Series Forecasting
Machine Learning
DevOps
Git
GitHub
Apply
≈ $138k – $273k per year (Estimated) • Hybrid • TS/SCI • Full-Time • Boston • Washington
Python
SQL
Python
SQLAlchemy
Django
Alembic
Databases
PostgreSQL
AI/ML
AI Agents
Machine Learning
DevOps
GCP
Azure
AWS
Docker
Apply
$145k – $185k per year • Remote (location not specified, CT hours) • 5+ years exp
Python
JavaScript
PHP
TypeScript
SQL
Ruby
C#
AI/ML
AI Agents
Frontend
Angular
DevOps
Rest API
Azure
Git
Linux
TCP/IP
DNS
SOAP
Cybersecurity
Active Directory
Apply
$115k – $150k per year • Remote (location not specified, CT hours) • 7+ years exp
Python
JavaScript
PHP
TypeScript
SQL
Ruby
C#
AI/ML
AI Agents
Frontend
Angular
DevOps
Rest API
Azure
Git
Linux
TCP/IP
DNS
SOAP
Cybersecurity
Active Directory
Apply
$100k – $130k per year • Equity • In office • 2+ years exp • Bachelor's Degree • Hawthorne
Apply
$145k – $195k per year • Equity • In office • TS/SCI • 2+ years exp • Bachelor's Degree • Hawthorne
Python
Bash
AI/ML
Machine Learning
DevOps
Kubernetes
SRE
Linux
Apply
$125k – $160k per year • Equity • In office • TS/SCI • 1+ year exp • Bachelor's Degree • Hawthorne
Python
Go
C++
Bash
DevOps
Terraform
Ansible
CI/CD
Kubernetes
SRE
Linux
TCP/IP
Apply
$165k – $240k per year • Equity • In office • 5+ years exp • Bachelor's Degree • Hawthorne
Python
PowerShell
DevOps
Ansible
VMWare
Configuration Management
KVM
Hyper-V
Apply
$105k – $140k per year • Equity • In office • 1+ year exp • Bachelor's Degree • Hawthorne
Python
AI/ML
Human-in-the-Loop
Apply
$100k – $130k per year • Equity • In office • 2+ years exp • Bachelor's Degree • Hawthorne
Apply
$100k – $130k per year • Equity • In office • 2+ years exp • Bachelor's Degree • Hawthorne
Apply
≈ $77k – $166k per year (Estimated) • In office • 2+ years exp • Bachelor's Degree • Hawthorne
Apply
$77k – $134k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • Hawthorne
Apply
≈ $50k – $132k per year (Estimated) • In office • Contractor • Hawthorne
Apply
See all jobs
This is one of many
1,227,840 more open roles from verified company boards, updated every day.