368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$198k – $395k per year (Estimated)
Location
In office (Memphis)
Seniority
Middle · 3+ years exp
Overview
Company
Impact
Profile match

xAI

xAI is an American artificial intelligence company founded by Elon Musk in 2023 with the stated goal of building models that help humans understand the universe. It develops the Grok family of large language models, distributes them through a consumer assistant, a developer API and deep integration with the X social platform, and adds image and video generation through Grok Imagine. The company runs its own Colossus supercomputer clusters in Memphis, Tennessee, is headquartered in Palo Alto, California, and merged with X Corp in 2025 to combine model development with a large consumer distribution channel.

SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.

ABOUT THE ROLE:

The Data Center Engineering team builds the internal systems and platforms that keep our data centers running at the scale and reliability required for frontier AI training and inference. We partner closely with datacenter operations, research, and infrastructure teams to deliver high-leverage tools that turn raw operational data into clear insight and action.

RESPONSIBILITIES:

  • Build and operate the software stacks that make site operations scalable, auditable, and fast - including repair trackers, vendor turnback workflows, operational dashboards, and their integrations. Your users are the technicians, managers, NOC operators, and leadership who run the fleet.
  • Design, build, and operate multi-service production systems (UI, APIs, data pipelines, auth, and on-call) for systems such as SRT-style repair/maintenance trackers and vendor turnover/turnback state machines.
  • Own correctness of operational state: node state accuracy, queue ownership, and audit trails - the data that decides what work happens on the floor.
  • Build and maintain integrations with ticketing, inventory/rack systems, telemetry stores, and vendor portals.
  • Keep the tools themselves reliable: uptime, data integrity, reconciliation, access control, and safe deploys.
  • Embed with SiteOps and NOC users; measure workflow adoption, not just feature delivery.

BASIC QUALIFICATIONS: 

  • Bachelor’s degree in Computer Science, Engineering, or related fields.
  • 3+ years building and operating production software (backend and/or full-stack).
  • Strong fundamental knowledge of computer science - data structures, algorithms, operating systems and networking.
  • Strong proficiency in at least one programming language e.g. Rust, Python, JavaScript, Java, C++, etc.
  • Experience designing and developing RESTful APIs for mission-critical applications.
  • Experience working with at least one database like Postgres, MongoDB, MySQL, DynamoDB, etc.
  • Experience collaborating with cross-functional teams.
  • Experience working in cloud platforms like GCP, Microsoft Azure, AWS, OCI or similar.
  • Experience using observability tools and dashboards like New Relic, Splunk, Grafana, etc.
  • Experience writing unit tests and integration tests.

PREFERRED SKILLS AND EXPERIENCE:

  • Full-stack or backend + data experience, especially with workflow and state-machine systems.
  • Experience automating deployments using CI/CD tools like Azure Devops, ArgoCD, Jenkins, GitHub Actions, etc.
  • Experience building high-correctness operational UIs where a wrong value can dispatch a human to the wrong rack.
  • On-call discipline and a track record of treating internal platforms with production rigor.
  • Experience designing and shipping multi-service systems (APIs, data stores, and at least one of: UI, pipelines, or auth).
  • Experience delivering high-quality internal tools in rapidly changing environments (as a tech lead, former founder, etc.).
  • Proven ownership of correctness-sensitive systems - state machines, workflows, or operational data where inaccurate state has real-world impact.
  • Experience integrating with external systems via APIs (e.g. ticketing, inventory, telemetry, or vendor portals).
  • MS in Computer Science or related field.
  • Experience collaborating closely with operations, NOC, and infrastructure teams.
  • Experience working with performance load testing tools like BlazeMeter, k6, etc.

ADDITIONAL REQUIREMENTS:

  • The role is fully onsite in Memphis, TN or Southhaven, MS. Candidates are expected to be located near the area or open to relocation.

SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Memphis
$22k – $58k per year (Estimated) • In office • Full-Time • Pune
AI/ML
Claude
Claude Code
Copilot
AI Agents
DevOps
ArgoCD
AWS
Azure
Backstage
CI/CD
Docker
GCP
GitHub Actions
Jenkins
Kubernetes
Platform Engineering
GitHub
Apply
$29k – $70k per year (Estimated) • Remote/Hybrid • Full-Time • 11+ years exp • Bachelor's Degree • Hyderabad
Python
SQL
Databases
Databricks
Snowflake
DevOps
AWS
Azure
CI/CD
Apply
$34k – $75k per year (Estimated) • Remote/Hybrid • Full-Time • 13+ years exp • Bachelor's Degree • Hyderabad
Java
Node JS
TypeScript
JavaScript
Databases
Cassandra
DynamoDB
Memcached
MySQL
PostgreSQL
Redis
Frontend
GraphQL
DevOps
AWS
CI/CD
Apply
$21k – $55k per year (Estimated) • In office • Full-Time • 3+ years exp • Navi Mumbai
C#
Java
C#
.NET
Java
Spring Boot
Databases
PostgreSQL
Redis
AI/ML
Copilot
DevOps
AWS
Azure
CI/CD
Amazon S3
GitHub
Apply
$33k – $78k per year (Estimated) • Equity • Remote • Full-Time • 8+ years exp • Bachelor's Degree • India
Apex
JavaScript
Python
TypeScript
Apex
Copado
Lightning Web Components
AI/ML
AutoGen
CrewAI
Fine-tuning
Hallucination
LangChain
LangGraph
LlamaIndex
LLM
RAG
Semantic Kernel
Semantic Search
Synthetic Data
Vertex AI
Agentforce
AWS Bedrock AgentCore
Semantic Search
AI Agents
Model Context Protocol
DevOps
AWS
CI/CD
GitHub Actions
Jenkins
Vector
GitHub
Cybersecurity
Crowdstrike
Management
Slack
Marketing
Salesforce
Apply
$117k – $239k per year (Estimated) • In office • 3+ years exp • Memphis
Apply
$440k per year • In office • Palo Alto
C++
AI/ML
CUDA Toolkit
CUDA
Frontend
Sass
Apply
$167k – $358k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Memphis
Python
SQL
AI/ML
BERT
InfiniBand
DevOps
HPC
Apply
$440k per year • In office • Palo Alto
C++
Python
Rust
AI/ML
LLM
Reinforcement Learning
Apply
$600k per year • In office • Palo Alto
AI/ML
Reinforcement Learning
RLHF
DPO
Post-training
Apply
$117k – $239k per year (Estimated) • In office • 3+ years exp • Memphis
Apply
$167k – $358k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Memphis
Python
SQL
AI/ML
BERT
InfiniBand
DevOps
HPC
Apply
$204k – $396k per year (Estimated) • In office • 5+ years exp • High School Diploma • Memphis
Bash
Management
Jira
Apply
$150k – $316k per year (Estimated) • In office • 4+ years exp • High School Diploma • Memphis
Bash
Management
Jira
Apply
$150k – $316k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Memphis
MATLAB
MATLAB
Simulink
Design
AutoCAD
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.