368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$188k – $412k per year (Estimated)
Location
Remote/Hybrid (London, United Kingdom)
Seniority
Staff
Overview
Company
Impact
Profile match
Wayve is a British autonomous driving technology company that develops end-to-end "Embodied AI" foundation models for self-driving vehicles and advanced driver assistance systems (ADAS). Founded in 2017 by University of Cambridge researchers, Wayve pioneered what it calls the "AV2.0" approach to autonomous vehicles. Unlike traditional self-driving stacks that rely on rigid hand-coded rules, high-definition (HD) 3D maps, and specialized sensors, Wayve uses a unified deep learning model.

About us

Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and safety of automated driving systems.

Our vision is to create autonomy that propels the world forward. Our intelligent, mapless, and hardware-agnostic AI products are designed for automakers, accelerating the transition from assisted to automated driving.

In our fast-paced environment big problems ignite us-we embrace uncertainty, leaning into complex challenges to unlock groundbreaking solutions. We aim high and stay humble in our pursuit of excellence, constantly learning and evolving as we pave the way for a smarter, safer future.

At Wayve, your contributions matter. We value diversity, embrace new perspectives, and foster an inclusive work environment; we back each other to deliver impact.

Make Wayve the experience that defines your career!

The role

As a Staff ML Performance Engineer, you’ll play a key role in high-impact projects, optimising ML inference for edge accelerators and GPUs. The focus of this team is to run large transformer-based models efficiently on low-cost, low-power edge devices to enable Wayve’s first driving product.

You’ll help set the technical direction for turning these models into production systems that run reliably on in-vehicle compute. This is a hands-on role working across ML systems, compilers, runtimes, kernels, and embedded deployment, contributing to several early-stage, high-impact projects at Wayve.

Key responsibilities:

  • Profile and pinpoint bottlenecks across the full inference stack (model graph, compiler/runtime, kernel execution, memory movement) and deliver measurable improvements.
  • Implement and validate optimisations in compilers, runtimes, and/or kernels (e.g. operator fusion, scheduling, quantisation-aware performance, custom kernels).
  • Build robust benchmarking and regression testing to ensure performance improvements hold across models, devices, and software releases.
  • Optimise for multiple targets (e.g. NVIDIA Orin/Thor, Qualcomm) and work with teams to support these in a maintainable way
  • Collaborate with model developers to influence architecture and training/deployment decisions that affect on-device performance.
  • Contribute to technical roadmaps and tooling and help raise the standard of performance engineering across the team

About you

Essential

  • Proven experience improving performance in production systems with tight constraints (latency, memory, bandwidth, power/thermal, or cost).
  • Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly.
  • Comfort operating at multiple levels of abstraction - from high-level model behaviour down to low-level kernel/runtime execution.
  • Strong software engineering fundamentals (debugging, profiling, testing, and maintainable code).
  • Clear communicator and collaborative teammate; able to align multiple stakeholders on performance trade-offs and priorities.

Desirable

  • Exposure to embedded or edge deployment of ML models, including benchmarking on real devices and handling system-level constraints.
  • Experience with NVIDIA and/or Qualcomm SoCs and performance tooling.
  • Python and C++ proficiency.
  • Experience mentoring others and/or driving technical direction in a small, fast-moving team.

This is a full-time role based in our office in London. At Wayve we want the best of all worlds so we operate a hybrid working policy that combines time together in our offices and workshops to fuel innovation, culture, relationships and learning, and time spent working from home.

Wayve is committed to creating an inclusive interview experience. If you require any accommodations or adjustments to participate fully in our interview process, please let us know.

We understand that everyone has a unique set of skills and experiences and that not everyone will meet all of the requirements listed above. If you’re passionate about self-driving cars and think you have what it takes to make a positive impact on the world, we encourage you to apply.

At Wayve we're committed to creating a diverse, fair and respectful culture that is inclusive of everyone based on their unique skills and perspectives, and regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, veteran status, pregnancy or related condition (including breastfeeding) or any other basis as protected by applicable law.

For more information visit Careers at Wayve.

To learn more about what drives us, visit Values at Wayve

For US candidates only, please visit E-Verify Notice and Participation and Right to Work

DISCLAIMER: We will not ask about marriage or pregnancy, care responsibilities or disabilities in any of our job adverts or interviews. However, we do look to capture information about care responsibilities, and disabilities among other diversity information as part of an optional DEI Monitoring form to help us identify areas of improvement in our hiring process and ensure that the process is inclusive and non-discriminatory.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
London
NLP / LLM Engineer 10 hours ago
$18k – $24k per year (net) • In office • Full-Time • 3+ years exp • Tashkent
Python
Python
FastAPI
Databases
ElasticSearch
Milvus
Pinecone
Qdrant
Weaviate
AI/ML
AI Agents
ChatGPT
DeepEval
Embeddings
Gemini
Hybrid Search
LangChain
Langfuse
LangGraph
LlamaIndex
LLM
LoRA
NLP
PEFT
Prompt Engineering
PyTorch
QLoRA
RAG
Reranking
Semantic Search
Synthetic Data
Tokenization
Triton
vLLM
Transformers
Anthropic
DPO
GraphRAG
Hugging Face
OCR
OpenAI
Semantic Search
SFT
Structured Outputs
Function Calling
TGI
DevOps
CI/CD
Docker
Git
GitHub
Analytics
A/B Testing
Apply
In office • Full-Time • 3+ years exp • Bachelor's Degree • Petaling Jaya
Python
SQL
AI/ML
Claude
LangChain
LangGraph
LightGBM
LLM
PyTorch
Qwen
RAG
Scikit-learn
Spark
Synthetic Data
TensorFlow
Transformers
XGBoost
AI Agents
Apply
$26k – $62k per year (Estimated) • In office • 5+ years exp • Moscow
Python
SQL
Python
pySpark
AI/ML
Pandas
PyTorch
Transformers
Spark
Apply
$20k – $51k per year (Estimated) • In office • Full-Time • 2+ years exp • Almaty
Python
SQL
AI/ML
NLP
NumPy
CatBoost
LLM
MLFlow
PyTorch
DevOps
Docker
Git
Apply
$28k – $78k per year (Estimated) • In office • Full-Time • 3+ years exp • Hyderabad
Python
AI/ML
LLM
PyTorch
Context Engineering
Post-training
SFT
AI Agents
Function Calling
Management
ServiceNow
Apply
$336k – $370k per year • Equity • Remote/Hybrid • 5+ years exp • Sunnyvale
C++
Python
C++
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
Ignite
Multimodal AI
PyTorch
VLM
AI Agents
Apply
$336k – $370k per year • Equity • Remote/Hybrid • Sunnyvale
C++
Python
C++
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
Ignite
Multimodal AI
PyTorch
Reinforcement Learning
Mixture of Experts
Robotics
Motion Planning
Reinforcement Learning
Apply
$110k – $306k per year (Estimated) • Remote/Hybrid • 4+ years exp • London
Python
AI/ML
Ignite
Knowledge Distillation
Multimodal AI
PyTorch
Ray
Spark
Synthetic Data
Flyte
DevOps
AWS
Azure
GCP
Apply
$110k – $307k per year (Estimated) • Remote/Hybrid • 4+ years exp • London
Python
AI/ML
Ignite
Multimodal AI
PyTorch
Reinforcement Learning
Robotics
Reinforcement Learning
Sensor Fusion
Sim-to-Real
Apply
$181k – $398k per year (Estimated) • Remote/Hybrid • London
AI/ML
Ignite
PyTorch
TensorRT
DevOps
CI/CD
GitHub Actions
Grafana
GitHub
Apply
$77k – $148k per year (Estimated) • Equity • In office • Master's Degree • London
JavaScript
Python
Scala
AI/ML
AI Agents
DevOps
GitHub
Apply
$87k – $159k per year (Estimated) • In office • Contractor • 5+ years exp • London • Stockholm
Design
Figma
Marketing
Zendesk
Apply
$105k – $204k per year (Estimated) • In office • Full-Time • London
Python
SQL
Databases
Snowflake
AI/ML
Dagster
dbt
Analytics
A/B Testing
Apply
$27k – $61k per year (Estimated) • In office • Full-Time • 5+ years exp • Pune • London
Analytics
Power BI
Tableau
Marketing
Salesforce
Apply
$34k – $85k per year (Estimated) • In office • Full-Time • Bachelor's Degree • London
AI/ML
AI Agents
Edge AI
QA
Appium
Cucumber
Cypress
Selenium
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.