368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$200k – $420k per year
Location
In office (Palo Alto)
Seniority
Staff · 5+ years exp
Overview
Company
Impact
Profile match
River AI is an artificial intelligence company headquartered in Palo Alto, California, and founded in 2026 by xAI co-founder Igor Babuschkin. The company is building what it calls an open AI stack, spanning personal AI models, tooling for developers to train and serve their own models, and a custom system on chip with an onboard machine learning accelerator. It raised 1.1 billion dollars led by General Catalyst with strategic investment from NVIDIA and AMD Ventures.

At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, custom training infrastructure, next-generation UIs, and frontier deep learning research.

Who we are

We are scientists, engineers, and builders from the industry's top tech companies and AI labs. We bring a proven track record of scaling consumer systems for hundreds of millions of users and architecting the pre-training infrastructure behind today's frontier models.

About the Role

We are seeking exceptional hardware performance engineers to architect, model, and correlate high-performance custom silicon. You will develop high-fidelity simulators to predict how our AI accelerator architecture and SoC system will handle real-world AI models. You will take ownership of performance models, ISA and kernel optimization, and FPGA emulators for software development. You will be collaborating both up and down the stack with compiler, IR, and software teams, as well as with RTL design engineers.

What You’ll Do

  • Simulator Development: Design and implement high-performance and functional models of complex hardware using C++ and/or SystemC.
  • Micro-architectural Exploration: Conduct "what-if" studies to evaluate architectural changes (e.g., cache sizes, branch predictors, pipeline depths, scatter/gather, matmul shaping) and their impact on IPC, MFU, TTFT, and total execution time.
  • Workload Characterization: Analyze and profile AI kernels and software stacks to generate representative traces that stress-test the hardware models.
  • HW/SW Co-Design: Collaborate with compiler and kernel teams to optimize software mapping to hardware, ensuring the architecture supports emerging algorithmic breakthroughs efficiently.
  • Performance Correlation: Validate the performance model against RTL and pre-silicon emulators to ensure the model’s accuracy remains within strict tolerance levels.
  • Bottleneck Analysis: Identify and quantify system-level bottlenecks, ranging from instruction-level parallelism (ILP) limits to bandwidth throttling to utilization.

Skills and Qualifications

Minimum Qualifications:

  • Bachelor’s degree in Electrical Engineering or Computer Engineering or Computer Science, and 5+ years practical industry experience working with advanced process nodes (7nm or below).
  • Expert proficiency in C/C++ or event-driven simulation environments like SystemC
  • Hands-on experience with how compilers transform code (LLVM/GCC/XLA) and how high-performance kernels (CUDA/Triton) interact with the underlying ISA.
  • Expert knowledge in Computer Architecture of at least one style of chip, including SoCs, CPUs, GPUs, or AI accelerators
  • Experience with profiling hardware with performance counters, hardware profilers, and trace analysis tools to dissect application behavior.
  • A highly collaborative mindset to push boundaries and co-design effectively with other engineers.

Preferred Qualifications: (We encourage you to apply even if you don't meet all of these)

  • Hands-on experience in pre-silicon RTL/emulator and/or post-silicon performance validation
  • Knowledge or experience of QEMU models for pre-silicon software development
  • Proficiency in scripting for data post-processing, visualization of simulation results, and automation of massive regression suites.

Logistics & Benefits

  • Location: This role is based in Austin, Texas or Palo Alto, California.
  • Compensation: Depending on background, skills, and experience, the expected annual salary range for this position is $200,000 - $420,000 USD.
  • Visa Sponsorship: We sponsor visas. We can't guarantee success for every candidate or role, but if you're the right fit, we're committed to working through the visa process.
  • Benefits: River AI offers generous health, dental, and vision benefits, unlimited PTO, and relocation support as needed.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Palo Alto
$48k – $90k per year (Estimated) • Remote/Hybrid • Full-Time • Nice
C++
Apply
$175k – $250k per year • Remote/Hybrid • Full-Time • 4+ years exp • Master's Degree • New York
C++
Python
SQL
C++
PyTorch C++
TensorFlow C++
AI/ML
PyTorch
Scikit-learn
TensorFlow
Apply
$124k – $208k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • Austin • San Jose
C++
Python
SystemVerilog
Chips/EDA
Formal Verification
Apply
$71k – $112k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Austria
C#
C++
Python
Visual Basic
Apply
$150k per year • In office • Full-Time • New York
C++
Python
Apply
$200k – $420k per year • In office • Bachelor's Degree • Palo Alto
Python
Rust
TypeScript
JavaScript
AI/ML
LLM
Pre-training
Structured Outputs
Function Calling
Frontend
React.js
Mobile
React Native
DevOps
CI/CD
Kubernetes
Terraform
Apply
$200k – $420k per year • In office • 5+ years exp • Bachelor's Degree • Palo Alto
C++
AI/ML
CUDA Toolkit
Pre-training
DevOps
Vector
Apply
$200k – $420k per year • In office • 5+ years exp • Bachelor's Degree • Palo Alto
C++
C++
LLVM
PyTorch C++
AI/ML
CUDA Toolkit
PyTorch
Pre-training
Apply
$131k – $333k per year (Estimated) • In office • Palo Alto
AI/ML
Pre-training
Apply
$200k – $420k per year • In office • Bachelor's Degree • Palo Alto
AI/ML
Diffusion Models
Fine-tuning
JAX
Multimodal AI
PEFT
PyTorch
Reinforcement Learning
RLHF
Transformers
Edge AI
Pre-training
Apply
Chief of Staff 8 hours ago
$120k – $150k per year • Equity 0.4–0.7% • In office • Full-Time • 3+ years exp • Palo Alto
AI/ML
AI Agents
Apply
$140k – $310k per year (Estimated) • Remote/Hybrid • Bachelor's Degree • Palo Alto
Databases
Apache Kafka
NATS
DevOps
AWS
Azure
CI/CD
Docker
GCP
Grafana
gRPC
Kubernetes
OpenTelemetry
Platform Engineering
Prometheus
Robotics
EtherCAT
IoT
MQTT
OPC UA
Apply
$139k – $294k per year (Estimated) • In office • Palo Alto
Python
AI/ML
Fine-tuning
Hybrid Search
LLM
Prompt Engineering
RAG
Human-in-the-Loop
Knowledge Graph
AI Agents
Function Calling
DevOps
AWS
Apply
$137k – $292k per year (Estimated) • In office • Palo Alto
Python
AI/ML
Hybrid Search
LLM
Prompt Engineering
RAG
Human-in-the-Loop
AI Agents
Function Calling
DevOps
AWS
Apply
$150k – $271k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Palo Alto
C++
Java
Python
Rust
AI/ML
LLM
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.