1,431,920open jobs
84,427companies
218,560added this week
Browse all
Salary
≈ $130k – $324k per year (Estimated)
Location
Hybrid (London, United Kingdom)
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 10, 2026. First seen by Alion on Oct 8, 2026.

Overview
Company
Impact
Profile match
Luma AI is an artificial intelligence company that specializes in building advanced 3D visual technology and generative AI tools. Founded in 2021, the company is best known for its realistic text-to-3D models, NeRF (Neural Radiance Fields) capture app, and its high-quality video generation model, Dream Machine. By enabling creators and developers to easily capture, generate, and edit photorealistic 3D assets and video, it aims to make next-generation digital media creation accessible to everyone.

You'll own how Luma's models get served - integrating new architectures into the inference engine, scaling deployments across thousands of machines, and keeping expensive GPU fleets busy while meeting internal SLOs.

This is large-scale inference systems work: scheduling, fleet management, deployment pipelines, and reliability across clusters and hardware providers. It fits a strong systems engineer comfortable with model serving and Kubernetes at scale. If you want pure modeling rather than the systems that run models, this is firmly the systems side.

What You'll Own

  • Ship new model architectures by integrating them into the inference engine.

  • Collaborate across research, engineering, and infrastructure to optimize model efficiency and deployments.

  • Build internal tooling to measure, profile, and track the lifetime of inference jobs and workflows.

  • Automate, test, and maintain inference services for maximum uptime and reliability.

  • Manage and optimize inference workloads across clusters and hardware providers, and scale deployments across thousands of machines.

  • Build scheduling systems that use expensive GPU resources optimally while meeting SLOs, and maintain CI/CD for model checkpoints and SDKs.

First 90 Days

One way the first 90 could unfold.

  • Days 1-30 - Immerse & Diagnose: Learn the inference stack, the fleets, and where reliability or utilization break.

  • Days 30-60 - Ship & Validate: Integrate a model or ship tooling/scheduling that improves uptime or GPU utilization.

  • Days 60-90 - Scale & Systemize: Harden deployment pipelines and scheduling across clusters and providers.

What You Bring

  • Strong Python and system-architecture skills.

  • Experience deploying models with PyTorch, Hugging Face, vLLM, SGLang, TensorRT-LLM, or similar.

  • Experience with queues, scheduling, traffic control, and fleet management at scale.

  • Experience with Linux, Docker, and Kubernetes, and with orchestration, deployment, and scheduling.

  • Familiarity with Redis and S3-compatible storage.

Nice to Have

  • Modern networking stacks including RDMA (RoCE, InfiniBand, NVLink).

  • High-performance large-scale ML systems (100+ GPUs).

  • CUDA, and FFmpeg or multimedia processing.

About Luma: Luma's mission is to build unified general intelligence that can generate, understand, and operate in the physical world. We believe multimodality is critical for intelligence - the next step beyond language models comes from vision. Luma is an equal opportunity employer.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,431,920 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
London
≈ $76k – $188k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • Belfast
Java
Rust
C++
DevOps
Linux
Apply
Software Engineer 3 hours ago
≈ $83k – $163k per year (Estimated) • Hybrid • Barcelona
Go
Databases
MySQL
PostgreSQL
RabbitMQ
Apache Kafka
DevOps
Kong
CI/CD
AWS
SLI/SLO/SLA
Management
Agile
Apply
Lead Software Engineer 4 months ago
$132k – $172k per year • Hybrid • Full-Time • London
Python
JavaScript
TypeScript
AI/ML
LLM
LLM Evaluation
Machine Learning
Frontend
React.js
DevOps
Azure
Apply
$132k – $152k per year • Hybrid • Full-Time • London
Python
JavaScript
TypeScript
AI/ML
LLM
LLM Evaluation
Machine Learning
Frontend
React.js
DevOps
Terraform
Azure
Docker
Bicep
Apply
full stack Developer 6 hours ago
≈ $75k – $186k per year (Estimated) • In office • Contractor • London
JavaScript
TypeScript
SQL
Node JS
Node JS
Nest.JS
TypeORM
Databases
PostgreSQL
DynamoDB
Amazon Aurora
Frontend
Next.js
React.js
DevOps
Terraform
CI/CD
AWS
GitHub
Amazon ECS
Apply
≈ $55k – $100k per year (Estimated) • Hybrid • Full-Time • Warsaw
Python
PowerShell
Bash
DevOps
Terraform
Ansible
GCP
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Platform Engineering
Windows
Management
ITIL
Apply
≈ $125k – $263k per year (Estimated) • Remote (location not specified) • Full-Time • 6+ years exp
Python
Python
FastAPI
Pydantic
Databases
Weaviate
Chroma
Pinecone
AI/ML
LangGraph
LangChain
Embeddings
LangSmith
huggingface_hub
LLM
RAG
TruLens
OpenAI
Anthropic
Tool Use
DevOps
Azure
Apply
Senior Data Scientist 7 hours ago
≈ $63k – $103k per year (Estimated) • In office • Full-Time • 5+ years exp • Warsaw
Python
AI/ML
Scikit-learn
TensorFlow
PyTorch
Machine Learning
DevOps
GCP
Azure
AWS
Apply
$150k – $250k per year • In office • 8+ years exp • Bachelor's Degree • El Monte
JavaScript
TypeScript
COBOL
COBOL
IBM MQ
Databases
Redis
Apache Kafka
Frontend
React.js
DevOps
gRPC
Azure
CI/CD
AWS
Kubernetes
Platform Engineering
Azure AKS
Management
Agile
Apply
Data Analyst 7 hours ago
$75k – $115k per year • In office • 3+ years exp • Bachelor's Degree • Pasadena
Python
SQL
SAS
Analytics
Tableau
Power BI
A/B Testing
Microsoft Excel
Apply
$170k – $360k per year • Hybrid • Full-Time • Redwood City
Python
AI/ML
Spark
Multimodal AI
Ray
Machine Learning
Apply
$235k – $353k per year • Hybrid • Full-Time • 8+ years exp • Redwood City
AI/ML
AI Agents
Apply
≈ $128k – $294k per year (Estimated) • Hybrid • Full-Time • London
AI/ML
CUDA Toolkit
Multimodal AI
ONNX
TensorRT
PyTorch
CUDA
Triton
XLA
Apply
≈ $128k – $294k per year (Estimated) • Hybrid • Full-Time • London
C
C
MPI
AI/ML
vLLM
RLHF
Reinforcement Learning
Multimodal AI
Function Calling
AI Agents
SGLang
TRL
Transformers
PyTorch
LLM
Ray
PPO
GRPO
Post-training
FSDP
NCCL
LLM Evaluation
Computer Use
Tool Use
DevOps
Kubernetes
Apply
Remote (location not specified) • Full-Time
C
C
MPI
AI/ML
vLLM
RLHF
Reinforcement Learning
Multimodal AI
Function Calling
AI Agents
SGLang
TRL
Transformers
PyTorch
LLM
Ray
PPO
GRPO
Post-training
FSDP
NCCL
LLM Evaluation
Computer Use
Tool Use
DevOps
Kubernetes
Apply
≈ $34k – $76k per year (Estimated) • In office • Full-Time • London
Apply
≈ $34k – $76k per year (Estimated) • In office • Full-Time • London
Apply
≈ $51k – $148k per year (Estimated) • In office • 5+ years exp • London • Derry • Belfast • Birmingham
Management
Outlook
Agile
Apply
≈ $88k – $179k per year (Estimated) • In office • Full-Time • London • Derry • Belfast • Birmingham
Apply
≈ $109k – $201k per year (Estimated) • In office • Full-Time • London
Apply
See all jobs
This is one of many
1,431,920 more open roles from verified company boards, updated every day.