824,589open jobs
53,146companies
135,057added this week
Browse all
Salary
≈ $82k – $195k per year (Estimated)
Location
Hybrid (Oxford, United Kingdom)
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 26, 2026. First seen by Alion on Jun 17, 2026.

Overview
Company
Impact
Profile match
Lumai delivers optical compute for the inference era, enabling up to 50x AI performance and 90% less power. Supports billion-parameter LLMs today.

The Opportunity

Lumai is redefining how the world computes. We are an ambitious, venture-backed UK startup pioneering a breakthrough AI accelerator for data centers which uses 3D optical compute. Our radical technology uses light to perform computation at orders of magnitude faster speeds and at far greater scales than ever before, all whilst consuming far less energy than traditional approaches.

Lumai is unlocking performance and efficiency gains that could transform the economics of AI and compute infrastructure and reshape how intelligence scales globally.

If you are passionate about bringing groundbreaking technology to market, and want to be part of a team pushing the boundaries of what is physically possible, Lumai is where you can make it happen.

About Lumai

Founded in 2022, Lumai is a University of Oxford spinout using optical processing to accelerate large language models (LLMs) and other transformer-based AI systems. The team combines expertise in optical computing, machine learning, and physics.

Lumai has already secured over $15 million in investment from leading deep-tech investors like Constructor Capital, IP Group, PhotonVentures and government grants, and is scaling rapidly to deploy the fastest optical compute currently available globally.

The Role

We are bringing the world's first optical AI compute platform to market. As we move from development into field deployment, we are looking for a Software Inference Deployment Engineer to own the software-side integration and customer support of Lumai Iris servers in third-party data centre environments.

You will begin by working alongside our software and engineering teams - helping integrate the Iris software stack, supporting model onboarding through the toolchain, and getting hands-on with the disaggregated prefill/decode runtime. This is intentional: the best way to develop deep expertise in a novel platform is to build with it. As deployments go live, you will take ownership in the field - supporting customer integration into their inference stacks, troubleshooting software issues, and acting as a primary technical contact for customer ML and infrastructure engineering teams.

This is an opportunity to work at the cutting edge of efficient AI inference - deploying a genuinely novel compute platform into production for the first time, and playing a central role in how it reaches the world.

What You'll Do

  • Work alongside Lumai's software and engineering teams to integrate, test, and harden the Iris software stack ahead of deployment

  • Support model onboarding through the Iris toolchain - loading, conversion, and framework integration

  • Develop hands-on familiarity with the disaggregated prefill/decode runtime, including how Iris servers operate alongside decode processors

  • Support customer integration of Lumai Iris into their own frameworks

  • Own software-side troubleshooting in the field, acting as the first line of response post-deployment

  • Train and enable customer ML and infrastructure engineering teams on the Iris software platform

  • Feed field findings, integration issues, and customer feedback back into product and engineering

What We're Looking For

Must-Have

  • Hands-on software engineering experience in AI infrastructure, inference serving, accelerator integration, or comparable deep-tech hardware-software environments

  • Strong Python skills and familiarity with major ML frameworks (PyTorch in particular)

  • Practical experience with model deployment workflows - loading, format conversion, quantisation, or framework integration

  • Comfortable working with inference serving stacks (for example vLLM, TensorRT-LLM, or similar)

  • Familiarity with Linux, containerisation (Docker), and cluster environments

  • Comfortable in a customer-facing role, able to communicate clearly with ML and infrastructure engineering teams

  • Comfortable working in a fast-moving, early-stage environment where the product and the deployment approach are both still being developed

Strong Preference For

  • Experience integrating accelerator hardware (GPUs, FPGAs, ASICs, NPUs, or novel architectures) into customer inference workflows

  • Familiarity with the NVIDIA inference stack - CUDA, TensorRT, Triton

  • Exposure to disaggregated inference architectures, prefill/decode separation, or KV cache management

Compensation & Benefits

  • Highly Competitive Salary: We are not saying our salary is a blank check, but let's just say it won't be a source of your stress

  • Share Option Scheme: We are all in this together! We believe in shared success while we build the Lumai of tomorrow

  • Pension Scheme: Plan for retirement with AVIVA

  • Private Health Insurance: We firmly believe that you come first, and a happy you is a healthy you! Look after yourself and your loved ones with AXA

  • Cycle to Work: Spread the cost of a bike, a bike and accessories or just accessories and save on tax

  • L&D Allowance: Stay at the forefront of your field with a £500 annual development budget

  • Subsidised On-site Lunches: Enjoy on-site healthy meals at half the price, as Lumai covers 50% of the cost

  • Holidays: Enjoy some deserved "me time" with 25 days paid holiday (plus bank holidays) per year

  • Socials: Be part of an inclusive community enjoying occasional all-company off-sites, lunches and socials

Interview Process

Our process is four stages. An initial conversation with our HR team to understand what you want from the role and what we want from it. Two technical sessions with our Product and Leadership team. Finally, an HR-team session covering scope, terms, and any final questions. We aim to move fast on candidates we are excited about; expect roughly three to four weeks end to end.

Lumai is an equal opportunity employer. We make hiring decisions on merit, scope-fit, and the strength of the working relationship we expect to build with each hire. Applications welcome from candidates of any background. If you are not sure whether you are a fit, send a note anyway.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
824,589 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
Oxford
≈ $87k – $174k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • Cambridge
Python
TypeScript
Python
FastAPI
Pydantic
Databases
DuckDB
AI/ML
Model Context Protocol
LLM
Knowledge Graph
DevOps
Rest API
Terraform
AWS CDK
AWS
AWS Fargate
AWS Lambda
Amazon S3
IAM
Amazon ECS
Cybersecurity
Least Privilege
Apply
≈ $82k – $153k per year (Estimated) • In office • Full-Time • Manchester
Python
JavaScript
TypeScript
Databases
PostgreSQL
ElasticSearch
OpenSearch
Frontend
Next.js
React.js
DevOps
Rest API
Terraform
Ansible
OpenShift
Azure
AWS
Docker
Kubernetes
Management
Agile
Scrum
QA
Cypress
Swagger
Jest
Mocha
Apply
≈ $82k – $195k per year (Estimated) • Equity • In office • Full-Time • London
Python
TypeScript
Databases
PostgreSQL
AI/ML
Machine Learning
DevOps
AWS
Docker
Kubernetes
Apply
≈ $81k – $192k per year (Estimated) • Equity • In office • Full-Time • London
Python
JavaScript
TypeScript
Frontend
React.js
Apply
Software Engineer II 25 days ago
≈ $65k – $191k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Dundee
JavaScript
Java
TypeScript
Java
Spring Framework
Maven
Spring Boot
Hibernate
Gradle
AI/ML
Cursor
Frontend
Angular
DevOps
Rest API
CI/CD
Jenkins
Docker
SOAP
Management
Agile
Apply
In office • 7+ years exp • Bachelor's Degree • Bengaluru
Python
AI/ML
Scikit-learn
SciPy
Multimodal AI
Function Calling
AI Agents
NLP
TensorFlow
Pandas
NumPy
Keras
PyTorch
LLM
Agentic Workflows
Tool Use
Machine Learning
DevOps
GCP
Azure
AWS
Robotics
Digital Twin
Management
Agile
Apply
≈ $109k – $233k per year (Estimated) • Remote (United States) • Master's Degree • Phoenix
Python
SQL
Databases
Amazon Redshift
AI/ML
Amazon SageMaker
Machine Learning
DevOps
AWS
AWS Lambda
Apply
Hybrid • 3+ years exp • Bachelor's Degree • Makati
Python
JavaScript
PHP
TypeScript
Node JS
DevOps
Linux
Windows
Cybersecurity
PCI DSS
Management
Microsoft Office
Apply
$150k – $220k per year • Hybrid • 5+ years exp • Fremont
Python
SQL
Python
Django
Databases
PostgreSQL
ElasticSearch
OpenSearch
DevOps
GCP
Podman
Docker Swarm
AWS
Docker
Kubernetes
Amazon EKS
Amazon ECS
Linux
Apply
Python Architect 1 day ago
In office
Python
JavaScript
TypeScript
Python
Django
Celery
Django REST Framework
Databases
PostgreSQL
Redis
Frontend
RxJS
Angular
Angular Material
DevOps
CI/CD
Git
AWS
AWS Fargate
QA
Pytest
Apply
≈ $91k – $177k per year (Estimated) • Hybrid • Full-Time • Oxford
AI/ML
Machine Learning
DevOps
CI/CD
Apply
≈ $59k – $128k per year (Estimated) • In office • Full-Time • Spain
AI/ML
Machine Learning
Chips/EDA
Cadence Virtuoso
Apply
≈ $82k – $186k per year (Estimated) • Hybrid • Full-Time • Oxford
Python
AI/ML
Machine Learning
DevOps
HPC
Apply
≈ $163k – $313k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Machine Learning
Apply
≈ $89k – $171k per year (Estimated) • Hybrid • Full-Time • Oxford
AI/ML
Machine Learning
Apply
≈ $84k – $168k per year (Estimated) • Hybrid • 5+ years exp • Oxford
Python
JavaScript
TypeScript
C++
Databases
PostgreSQL
AI/ML
Machine Learning
Frontend
React.js
DevOps
Helm
CI/CD
Docker
Kubernetes
Management
Agile
Apply
Software Engineer 3 days ago
≈ $51k – $118k per year (Estimated) • Hybrid • 2+ years exp • Oxford
Python
JavaScript
TypeScript
C++
Databases
PostgreSQL
AI/ML
Machine Learning
Frontend
React.js
DevOps
Helm
CI/CD
Docker
Kubernetes
Management
Agile
Apply
Leasing Consultant 1 day ago
In office • PhD • Oxford
Management
Agile
Apply
≈ $29k – $60k per year (Estimated) • In office • 3+ years exp • High School Diploma • Oxford
Management
Agile
Apply
≈ $41k – $99k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Oxford
Design
SolidWorks
Apply
See all jobs
This is one of many
824,589 more open roles from verified company boards, updated every day.