368,746open jobs
9,444companies
47,506added this week
Browse all
Salary
$172k – $365k per year (Estimated)
Location
In office (Sunnyvale)
Seniority
Staff · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Cerebras Systems is an American computer hardware company founded in 2016 and headquartered in Sunnyvale, California that builds accelerators for artificial intelligence at wafer scale. Instead of assembling clusters from many small chips, it manufactures a single processor the size of an entire silicon wafer, the Wafer Scale Engine, which removes most of the communication overhead in large model training and inference. The company sells CS-series systems to research laboratories and enterprises, operates its own inference cloud known for very high token throughput, and has built large supercomputers with partners including the Gulf technology group G42.

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.

Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.

About The Role

We are seeking an experienced IT SRE Team Lead to build and run the reliability function for Cerebras' internal technology estate.

The IT SRE Team Lead will be responsible for the availability, performance, and operational quality of the systems Cerebras employees rely on every day, including identity, endpoint management, collaboration, SaaS, and internal networking. The right candidate will bring a software engineering mindset to IT operations, treating corporate infrastructure as code, with measurable SLOs, automated remediation, and a ruthless focus on eliminating toil.

You will build and lead a small, high-leverage team of engineers who build tooling, write automation, and respond when things break. You will partner closely with the security, networking, and infrastructure teams to make sure the internal environment stays fast, stable, and secure as the company scales.

Responsibilities

  • Define and own the reliability strategy for internal IT systems, including SLOs, error budgets, and operational health reporting.

  • Build and lead a team of IT SRE engineers focused on automation, observability, and incident response for corporate systems.

  • Design and implement automation to eliminate manual IT work across provisioning, access management, patching, and lifecycle operations.

  • Instrument internal services and SaaS integrations with monitoring, alerting, and on-call workflows.

  • Run incident response for IT outages, including root cause analysis and durable remediation.

  • Drive infrastructure-as-code and GitOps practices across IT-owned systems.

  • Partner with security and networking teams on identity, access, and network reliability.

Skills And Qualifications

  • Minimum 8 years of experience in SRE, DevOps, or IT engineering roles, with at least 2 years in a leadership capacity.

  • Direct hands on experience with AI coding tools, building and deploying AI agents for triage and bug fixes.

  • Strong software engineering background with hands-on experience in Python, Go, or similar, and comfort writing production-grade automation.

  • Deep experience with identity platforms (Okta, Entra), endpoint management (Jamf, Intune), and SaaS integration patterns.

  • Hands-on experience with infrastructure-as-code tools (Terraform) and CI/CD pipelines applied to IT systems.

  • Proven track record of running on-call rotations, defining SLOs, and driving operational maturity in a fast-moving environment.

  • Experience supporting highly technical engineering populations where uptime and speed both matter.

  • Strong organizational skills with the ability to multitask and prioritize.

  • Detail-oriented with the ability to anticipate the needs of customers and internal stakeholders.

  • Proactive, adaptable, and able to thrive in a rapidly changing environment.

  • Excellent verbal and written communication skills.

Why Join Cerebras

People who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:

  • Build a breakthrough AI platform beyond the constraints of the GPU.

  • Publish and open source their cutting-edge AI research.

  • Work on one of the fastest AI supercomputers in the world.

  • Enjoy job stability with startup vitality.

  • Our simple, non-corporate work culture that respects individual beliefs.

Find out more about what it's like to work at Cerebras here!

Apply today and become part of the forefront of groundbreaking advancements in AI!

Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.

This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Sunnyvale
$80k – $147k per year (Estimated) • Remote/Hybrid • Full-Time • Sydney
DevOps
CI/CD
Apply
$108k – $242k per year (Estimated) • Remote • Full-Time • 5+ years exp
C#
SQL
C#
.NET
Databases
MS SQL
RabbitMQ
AI/ML
ChatGPT
Claude
Claude Code
Copilot
Cursor
LLM
RAG
Model Context Protocol
DevOps
CI/CD
Git
GitHub
Apply
$40k – $107k per year (Estimated) • Remote • 5+ years exp • Tbilisi
C#
SQL
C#
.NET
Databases
MS SQL
RabbitMQ
AI/ML
ChatGPT
Claude
Claude Code
Copilot
Cursor
LLM
RAG
Model Context Protocol
DevOps
CI/CD
Git
GitHub
Apply
$93k – $210k per year (Estimated) • In office • Contractor • 8+ years exp • Bachelor's Degree • Singapore
DevOps
CI/CD
Apply
$77k – $194k per year (Estimated) • In office • Contractor • 5+ years exp • Singapore
Java
SQL
DevOps
CI/CD
Docker
Grafana
Incident Management
Kubernetes
OpenShift
SRE
Apply
$175k – $275k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Sunnyvale
AI/ML
Cerebras
AI Agents
Edge AI
OpenAI
Apply
$290k – $502k per year (Estimated) • In office • Full-Time • 2+ years exp • Master's Degree • Sunnyvale
JavaScript
Node JS
Python
Python
Flask
Databases
ElasticSearch
PostgreSQL
Redis
AI/ML
Cerebras
Ray
Edge AI
OpenAI
AI Agents
DevOps
Amazon EC2
Amazon EKS
Ansible
AWS
AWS CDK
AWS Fargate
AWS Lambda
CI/CD
CloudFormation
Docker
Git
Grafana
Helm
Jenkins
Kibana
Kubernetes
Logstash
Prometheus
Terraform
Amazon CloudWatch
Amazon ECS
Management
Jira
Apply
Simulation Engineer 5 days ago
$171k – $425k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Sunnyvale • Toronto
C++
Python
AI/ML
Cerebras
AI Agents
Edge AI
OpenAI
DevOps
HPC
Apply
$150k – $220k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Sunnyvale
AI/ML
Cerebras
AI Agents
Edge AI
OpenAI
Apply
$226k – $458k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree
AI/ML
Cerebras
AI Agents
Edge AI
OpenAI
DevOps
HPC
Apply
$127k – $185k per year • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • Sunnyvale
AI/ML
AI Agents
LLM
Apply
Sr. SDET 4 hours ago
$109k – $163k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Sunnyvale
Java
Python
SQL
Java
Spring Framework
Mobile
JUnit
DevOps
CI/CD
Kubernetes
QA
Cypress
JMeter
Postman
Apply
$168k – $356k per year • In office • Full-Time • 8+ years exp • Master's Degree • Sunnyvale
Python
AI/ML
AI Agents
ChatGPT
Gemini
LLM
NLP
RAG
Recommender Systems
Semantic Search
Semantic Search
Vertex AI
Apply
$170k – $205k per year • Equity • In office • Full-Time • 3+ years exp • Sunnyvale
C++
Go
Java
Rust
C++
Protobuf
Databases
Apache Kafka
NATS
PostgreSQL
AI/ML
Human-in-the-Loop
DevOps
API Gateway
CI/CD
Docker
gRPC
Kubernetes
Terraform
Apply
$215k – $260k per year • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco • Sunnyvale • Bellevue
C++
Go
DevOps
AWS
Azure
Cilium
eBPF
GCP
KVM
VMWare
Apply
See all jobs
This is one of many
368,746 more open roles from verified company boards, updated every day.