489,618open jobs
16,287companies
72,295added this week
Browse all
Salary
$153k – $366k per year (Estimated)
Location
Remote/Hybrid (Sunnyvale, United States, Toronto, Canada)
Seniority
Middle · 3+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Cerebras Systems is an American computer hardware company founded in 2016 and headquartered in Sunnyvale, California that builds accelerators for artificial intelligence at wafer scale. Instead of assembling clusters from many small chips, it manufactures a single processor the size of an entire silicon wafer, the Wafer Scale Engine, which removes most of the communication overhead in large model training and inference. The company sells CS-series systems to research laboratories and enterprises, operates its own inference cloud known for very high token throughput, and has built large supercomputers with partners including the Gulf technology group G42.

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.

Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.

About The Role

The Host and Network IO Team develops the full IO path implementation between a distributed system of server nodes, through the cluster, down to the custom RoCE network stack implemented in Cerebras' system, and over the proprietary IOs onto the WSE. As a software developer on the team, you will interface between AI application-level IO teams, cluster architecture teams, and FPGA/ASIC teams to develop solutions that optimize bandwidth and latency while minimizing congestion, pauses, pause spreading, unfairness, etc. Strong skills in socket programming will enable you to deploy robust management operations, while deftness in RDMA Verbs will enable you to optimize CPU resources and shape network traffic to deliver real world impact on AI performance metrics, as well as developing tools for gaining insight and visibility into network behavior. Meticulous analysis and rigour are key tenants of this role, harnessing that together with a deeply-understood mental model of the server, NIC, protocol, switch, and custom hardware behavior will enable you to lead network debug, optimize traffic patterns, and prescribe architectural changes.

Responsibilities

  • Develop x86 & ARM software to expose next-generation hardware IO capabilities for AI/HPC application teams

  • Govern a generic IO API with multiple internal users.

  • Develop control and configuration subsystems directly interacting with Cerebras hardware

  • Drive network performance debug of large AI clusters

  • Gather and analyze network statistics and packet traces to root cause and alleviate bottlenecks and sub-optimalities.

  • Develop tools/telemetry for increasing visibility into the network and IO datapath.

  • Optimize cpu/mem utilization leveraging kernel bypass and zero-copy techniques

  • Integrate leading edge networking technologies and protocols

  • Lead cross-functional technical projects spanning multiple teams and integrating diverse software and hardware components to deliver an improved network IO solution.

  • Foster clear and effective communication across teams and stakeholders.

Skills & Qualifications

  • Master's/PhD in Computer Science or Electrical Engineering + 1 year industry experience, OR 3+ years industry experience.

  • Experience in large software environments.

  • Embedded systems, HW/SW co-design, and some driver development.

  • Network protocol familiarity (TCP, RoCE) and network debug tools such as Wireshark, or willingness to learn

  • Some network switch environment familiarity or willingness to learn (Arista, Juniper, etc.).

  • Detail-oriented but keen to learn the bigger picture and step out of comfort zone to embrace the unknown.

Why Join Cerebras

People who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:

  • Build a breakthrough AI platform beyond the constraints of the GPU.

  • Publish and open source their cutting-edge AI research.

  • Work on one of the fastest AI supercomputers in the world.

  • Enjoy job stability with startup vitality.

  • Our simple, non-corporate work culture that respects individual beliefs.

Find out more about what it's like to work at Cerebras here!

Apply today and become part of the forefront of groundbreaking advancements in AI!

Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.

This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
489,618 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Sunnyvale
$48k – $72k per year • Remote • Internship • Bachelor's Degree • Atlanta
Python
JavaScript
TypeScript
SQL
PowerShell
C#
Node JS
C#
.NET
AI/ML
Prompt Engineering
AI Agents
LLM
LLM Guardrails
DevOps
Rest API
Analytics
Tableau
Power BI
Management
UiPath
Apply
$145k – $261k per year • Equity • In office • Full-Time • 6+ years exp • Seattle
Python
TypeScript
AI/ML
AI Agents
LLM
Human-in-the-Loop
Structured Outputs
Context Engineering
DevOps
CI/CD
Cloudflare
Fastly
Akamai
Cybersecurity
ModSecurity
CVSS
EPSS
KEV
AWS WAF
Apply
$93k – $189k per year • Remote • Full-Time • 10+ years exp • Bachelor's Degree • Columbus • Austin • Chicago • Dallas • Detroit
Databases
Google BigQuery
BigQuery
AI/ML
Vertex AI
AI Agents
Gemini
Google ADK
LLM Guardrails
DevOps
GCP
CI/CD
Platform Engineering
Google Cloud Run
Vector
Incident Management
Apply
Solution Architect 1 hour ago
$66k – $151k per year (Estimated) • Remote/Hybrid • Full-Time • Prague
AI/ML
Model Context Protocol
AI Agents
RAG
Tokenization
Agentic Workflows
DevOps
Azure
AWS
Cybersecurity
PCI DSS
Management
Agile
Apply
$68k – $157k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Toronto • Waterloo
Python
TypeScript
Python
FastAPI
Databases
Databricks
Delta Lake
Azure Cosmos DB
AI/ML
Spark
MLFlow
Prompt Engineering
RAG
OpenAI
DevOps
Rest API
Azure DevOps
Azure
CI/CD
Jenkins
Git
Docker
Kubernetes
Analytics
Power BI
Informatica
Management
ServiceNow
Apply
$188k – $352k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Sunnyvale • Toronto
Python
C++
C++
PyTorch C++
AI/ML
Quantization
AI Agents
PyTorch
LLM
Ray
Cerebras
OpenAI
NCCL
InfiniBand
ROCm
Edge AI
KV Cache
DevOps
Prometheus
SLURM
CI/CD
Kubernetes
Grafana
Chaos Engineering
Apply
$175k – $370k per year (Estimated) • In office • Full-Time • 7+ years exp • Sunnyvale
Python
AI/ML
AI Agents
Cerebras
OpenAI
Edge AI
DevOps
Terraform
Ansible
GitOps
AWS
Cloudflare
HPC
Apply
$176k – $373k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Toronto
Python
AI/ML
Model Context Protocol
AI Agents
Cerebras
OpenAI
Edge AI
DevOps
gRPC
etcd
Prometheus
Kubernetes
Grafana
eBPF
HPC
Apply
$124k – $300k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Sunnyvale
Python
Python
Asyncio
AI/ML
AI Agents
Cerebras
OpenAI
Edge AI
DevOps
Kubernetes
QA
Pytest
Apply
$162k – $355k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Sunnyvale • Toronto
Python
AI/ML
AI Agents
Cerebras
OpenAI
Edge AI
DevOps
Terraform
Helm
CI/CD
ArgoCD
AWS
Kubernetes
SRE
Platform Engineering
Apply
$146k – $267k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Sunnyvale
Kotlin
Kotlin
Kotlin Coroutines
Databases
Apache Kafka
Mobile
Jetpack Compose
MVVM
Room
Clean Architecture
ViewModel
DevOps
Azure
CI/CD
Apply
Data Scientist 2 1 hour ago
$102k – $205k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Sunnyvale
Python
SQL
Databases
Snowflake
AI/ML
Claude
Claude Code
AI Agents
LLM
RAG
Agentic Workflows
DevOps
CI/CD
Git
Apply
$129k – $273k per year (Estimated) • In office • Full-Time • 8+ years exp • Sunnyvale
Design
SolidWorks
Apply
Category Manager 2 hours ago
$103k – $210k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Master's Degree • Sunnyvale
Analytics
Microsoft Excel
Management
Agile
Kanban
Apply
$104k – $155k per year • In office • Full-Time • 9+ years exp • Bachelor's Degree • Sunnyvale
Design
AutoCAD
Apply
See all jobs
This is one of many
489,618 more open roles from verified company boards, updated every day.