928,299open jobs
56,686companies
155,823added this week
Browse all
Salary
≈ $111k – $301k per year (Estimated)
Location
Hybrid (San Francisco, United States)
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 28, 2026. First seen by Alion on Sep 28, 2026. Lumai scores C on the Alion truth index.

Overview
Company
Impact
Profile match
Lumai delivers optical compute for the inference era, enabling up to 50x AI performance and 90% less power. Supports billion-parameter LLMs today.

The Opportunity

Lumai is moving from breakthrough optical-compute research to production AI systems. We need an engineer to build the performance and power modelling capability that guides architecture and compiler decisions before silicon exists, helping us turn a working system into a competitive product.

The Role

You will own the performance and power modelling framework that predicts how Lumai’s photonic AI hardware will perform against real inference workloads. Your models will give architecture and compiler teams a shared source of truth for latency, throughput and power across hardware configurations.

The hard part is balancing speed, fidelity and usability: you will model concurrency, event-driven behaviour and compute-block latency at enough detail to support decisions, while keeping simulations fast enough to compare configurations in minutes. You will characterise operator-level workloads against hardware, RTL or emulation results so the model stays grounded in evidence.

By month six, you will have established a reliable modelling flow for priority workloads and device configurations, demonstrated a clear path for closing gaps between model and hardware, and made the simulator a tool that architecture and compiler teams use to make decisions. You will report to the Head of Architecture or VP Engineering and may build and lead a small team around this capability.

What You'll Do

  • Design, build and maintain a Python-based framework that models hardware and software latency, hardware concurrency and event-driven relationships.

  • Model compute engines, interconnect and synchronisation fabric, memory and bridge interfaces, and accelerator blocks across device configurations.

  • Characterise convolution, depthwise convolution, pooling, attention and other operator workloads against hardware, RTL or emulation results.

  • Build compiler-agnostic interfaces that ingest compiled models and MLIR from multiple internal compiler flows.

  • Analyse simulation traces with architecture and compiler teams to improve layer-to-hardware mappings and resolve performance discrepancies.

  • Run industry benchmarks and representative models, including UL Procyon, Geekbench AI, Stable Diffusion and contemporary LLMs, to produce decision-ready KPIs.

  • Develop block-level power models and validate worst-case power scenarios with micro-architecture, verification and physical-design teams.

  • Own the CI/CD pipelines and simulation flows that keep results reliable and reproducible as usage scales.

  • Translate modelling findings into architecture and compiler decisions, while setting technical direction and developing engineers or interns as the team grows.

What We're Looking For

Must-have

  • Proven experience building or substantially extending a hardware/software performance-modelling or simulation framework for AI/ML accelerators, ideally in Python.

  • Strong understanding of hardware concurrency, event-driven modelling, and AI inference accelerator micro-architecture, including compute, memory hierarchy, interconnect and dataflow.

  • Hands-on experience characterising AI operators and correlating model predictions with real hardware or hardware-accurate reference data.

  • Familiarity with compiler internals and intermediate representations such as MLIR, including tooling that consumes compiler output to drive simulation.

  • Strong software-engineering and technical-leadership practice, including framework design, CI/CD ownership and direction or review of work by junior engineers or interns.

Strong preference for

  • Direct experience with NPU, GPU or other edge or client AI inference accelerator architectures.

  • Experience working across multiple compiler toolchains and reconciling architecture-level and production compiler behaviour.

  • Exposure to power modelling, RTL, emulation, FPGA prototyping, photonic compute or another non-traditional compute architecture.

About Lumai

Lumai is the optical compute company building the next generation of AI infrastructure. Spun out of optics research at the University of Oxford in 2021, we compute with light instead of electrons. Our 3D optical technology carries out the matrix multiplications at the heart of AI inside beams of light travelling through free space, which lets it go beyond the limits of both silicon GPUs and integrated photonics.

In April 2026 we launched Iris Nova, the world's first optical computing system to run billion-parameter large language models in real time, using up to 90% less energy than conventional GPU-based systems. Iris Nova, the first server in the family, is now available for evaluation by hyperscalers, neoclouds, enterprises and research institutions. Aura and Tetra will follow.

Our work won the Falling Walls Award for Science Breakthrough of the Year 2025 and 'Best Overall Technology' at the OCP Future Technologies Symposium. We are headquartered in Oxford.

Why Lumai

You'll work on a new kind of computer. Optical computing for AI has been promised for decades. We have a working system running real models, and the hard part left is taking it to volume.

Your work ships. We are moving from first product to volume production, so what you build this year goes into the servers our customers run.

You'll work across disciplines. Optical engineers, machine learning researchers, and hardware and software engineers solve problems together. You will learn things that don't appear on your job description.

The work matters beyond Lumai. AI's appetite for energy is one of the defining constraints of the next decade. Our mission is sustainable intelligence at global scale: AI that is faster, cheaper to run and far less power-hungry.

You'll join early. You'll have a real say in how we build the product, the team and the way we work.

Equal Opportunity

Lumai is an equal opportunity employer. We make hiring decisions based on skills, experience and potential, and we welcome applications from people of all backgrounds. If you need an adjustment at any stage of the hiring process, let us know and we will do our best to support you.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
928,299 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
San Francisco
AI Enablement Lead 1 hour ago
≈ $106k – $212k per year (Estimated) • Remote (United Kingdom) • Full-Time
Databases
Snowflake
AI/ML
Copilot
Claude
ChatGPT
Model Context Protocol
LLM
LLM Guardrails
Agentic Workflows
Machine Learning
Management
Agile
Apply
Applied AI Researcher 3 hours ago
≈ $102k – $278k per year (Estimated) • In office • Full-Time • London
Python
AI/ML
Fine-tuning
Reinforcement Learning
Function Calling
LLM
SFT
Post-training
Computer Use
Tool Use
Apply
$90k – $150k per year • Remote (United Kingdom, Singapore) • Full-Time • London • San Francisco
Apply
$90k – $150k per year • In office • Full-Time • London
AI/ML
OpenAI
Anthropic
Apply
$90k – $150k per year • In office • Full-Time • London
Cybersecurity
PKI
Apply
MLOps Engineer 1 day ago
≈ $55k – $101k per year (Estimated) • Hybrid • Warsaw
Python
AI/ML
MLFlow
Kubeflow
Machine Learning
DevOps
Terraform
GCP
Helm
Azure
CI/CD
GitOps
Git
AWS
Docker
Kubernetes
Apply
≈ $14k – $32k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Databases
Databricks
AI/ML
Spark
DevOps
Azure
CI/CD
Analytics
Azure Data Factory
Management
Agile
Apply
$187k – $229k per year • In office • Full-Time • 8+ years exp • Chicago
Python
JavaScript
TypeScript
SQL
C#
Node JS
Databases
MySQL
PostgreSQL
ElasticSearch
Azure Cosmos DB
AI/ML
Prompt Engineering
Function Calling
AI Agents
LLM
Structured Outputs
Context Engineering
Tool Use
Frontend
Vue.js
Angular
React.js
DevOps
Azure
CI/CD
Octopus Deploy
Apply
In office • Contractor • Bachelor's Degree • Bengaluru
Python
SQL
Python
pySpark
Databases
MySQL
PostgreSQL
Oracle
AI/ML
Spark
DevOps
Rest API
GitHub Actions
CI/CD
Jenkins
Git
AWS
AWS Lambda
GitLab
Amazon S3
IAM
Amazon CloudWatch
Amazon EventBridge
AWS Step Functions
SOAP
Cybersecurity
Least Privilege
Analytics
ETL/ELT
AWS Glue
Apply
≈ $87k – $171k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Portland
SQL
Databases
Snowflake
Databricks
Azure SQL Database
DevOps
CI/CD
Git
Analytics
Power BI
ETL/ELT
SSIS
Dimensional Modeling
Apply
≈ $89k – $175k per year (Estimated) • Hybrid • Full-Time • Oxford
AI/ML
Machine Learning
DevOps
CI/CD
Apply
≈ $59k – $135k per year (Estimated) • In office • Full-Time • Spain
AI/ML
Machine Learning
Chips/EDA
Cadence Virtuoso
Apply
≈ $79k – $189k per year (Estimated) • Hybrid • Full-Time • Oxford
Python
AI/ML
vLLM
CUDA Toolkit
TensorRT
TensorRT-LLM
PyTorch
CUDA
Triton
KV Cache
Machine Learning
DevOps
Docker
Linux
Apply
≈ $82k – $202k per year (Estimated) • Hybrid • Full-Time • Oxford
Python
AI/ML
Machine Learning
DevOps
HPC
Apply
≈ $163k – $311k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Machine Learning
Apply
$150k – $300k per year • Equity • In office • Full-Time • 6+ years exp • San Francisco
AI/ML
Vapi
Voice Agents
Management
WhatsApp
Apply
$150k – $300k per year • Equity • In office • Full-Time • 6+ years exp • San Francisco
AI/ML
Model Context Protocol
Vapi
Voice Agents
Apply
$150k – $300k per year • Equity • In office • Full-Time • 6+ years exp • San Francisco
AI/ML
Multimodal AI
Vapi
Voice Agents
Apply
$150k – $300k per year • Equity • In office • Full-Time • 1+ year exp • San Francisco
AI/ML
Vapi
Voice Agents
Apply
$150k – $300k per year • Equity • In office • Full-Time • 3+ years exp • San Francisco
AI/ML
Vapi
Voice Agents
Cybersecurity
Threat Modeling
Apply
See all jobs
This is one of many
928,299 more open roles from verified company boards, updated every day.