700,010open jobs
41,477companies
98,955added this week
Browse all
Salary
$103k – $207k per year (Estimated)
Location
Remote (United Kingdom)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Build with generative AI on a unified inference API. Image generation, video generation, audio, 3D, and large language models. 400K+ models, managed infrastructure, usage-based pricing.

Join Runware as a Senior Machine Learning Engineer and be at the forefront of developing innovative AI solutions across various media modalities including text, image, video, 3D, and audio. We're building a powerful AI media creation platform designed to revolutionize how content is generated.

As a Senior Machine Learning Engineer, you’ll take the lead on critical projects, guiding the end-to-end lifecycle from research and experimentation to production deployment and performance monitoring. Your work will help shape the capabilities of our platform and enhance the experiences of users who rely on our cutting-edge AI technologies.

What You'll Be Doing

    • Integrate open-source and third-party models into our inference platform
    • Lead fine-tuning initiatives (LoRA, adapters, PEFT, domain adaptation)
    • Optimise inference workloads for latency, batching, memory efficiency, and throughput
    • Benchmark model quality vs cost vs performance across modalities
    • Improve inference startup times and stability under high load
    • Build evaluation frameworks and internal tooling for model validation
    • Work closely with Infrastructure and Backend teams on scalable serving systems
    • Monitor production performance and drive continuous optimisation
    • Mentor engineers and help raise the ML engineering bar across the team

Requirements

What We’re Looking For

    • Deep, hands-on expertise in Python and PyTorch, comfortable working at a level well beyond standard framework usage, including the quirks and edge cases that come with pushing Python for ML workloads
    • Demonstrated experience building ML applications and models yourself, not just running or deploying existing ones (e.g. serving a pre-built model via vLLM doesn't qualify on its own)
    • Hands-on experience writing GPU kernels in CUDA/C++, or in Triton
    • Low-level experience with model internals e.g. working directly with diffusion model architectures (diffusers or equivalent), not just calling high-level APIs
    • Experience with the PyTorch compiler (torch.compile) or comparable low-level PyTorch tooling
    • Practical experience fine-tuning large models as a repeatable service (LoRA, PEFT, adapters), built for speed and reuse across many models, not a single one-off training run
    • Real, verifiable project history e.g. GitHub repos, demonstrating what you have personally wrote
    • High ownership and comfort operating in a fast-paced startup environment

Nice to have

    • Experience optimizing inference workloads in GPU environments
    • Experience with vLLM or custom inference servers
    • Experience with diffusion models, LLMs, or multimodal architectures more broadly
    • Experience with Kubernetes, Docker, or containerised ML workloads
    • Experience building internal ML tooling or developer-facing APIs
    • Experience working in high-throughput distributed systems
    • Background in AI media generation (image, video, audio)

Benefits

We’re a remote-first team that comes together in person twice a year to plan, collaborate, and celebrate wins. Day to day we keep a few core hours for teamwork, but outside of that you set the schedule that helps you do your best work.

Our environment is fast-moving and ambitious. Big pushes are part of building category-defining products, but we balance that with flexible working, generous time off, and regular retreats so the team can stay sharp and motivated.

    • Generous paid time off - vacation, sick days, public holidays
    • Meaningful stock options - share in the upside you create
    • Remote-first setup - work from home anywhere we can employ you
    • Flexible hours - own your schedule outside core collaboration blocks
    • Family leave - paid maternity, paternity, and caregiver time
    • Company retreats - twice-yearly gatherings in inspiring locations
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
700,010 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$39k – $83k per year (Estimated) • In office • 5+ years exp • Moscow
Python
AI/ML
Weights & Biases
OpenCV
LoRA
MLFlow
YOLO
Fine-tuning
Prompt Engineering
Computer Vision
ONNX
TensorRT
PEFT
OpenVINO
Transformers
PyTorch
LLM
TensorBoard
CLIP
Vision Transformer
CNN
LLaVA
World Models
DevOps
Kubernetes
Apply
DevOps-инженер 5 hours ago
$19k – $37k per year (Estimated) • In office • 3+ years exp • Moscow
Python
JavaScript
Bash
Python
Flask
Django
Databases
PostgreSQL
Redis
MinIO
ElasticSearch
OpenSearch
Frontend
Vue.js
DevOps
Terraform
Ansible
Docker Compose
Zabbix
Loki
VMWare
Envoy
GitLab CI
CI/CD
Jenkins
Docker
Nginx
Grafana
Mimir
GitLab
Amazon S3
Linux
Cybersecurity
HashiCorp Vault
Apply
$10k – $25k per year (Estimated) • In office • 1+ year exp • Almaty
Python
SQL
Databases
PostgreSQL
Oracle
AI/ML
Pandas
NumPy
Machine Learning
Analytics
Matplotlib
Microsoft Excel
Apply
$185k – $339k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 8+ years exp • New York
Python
TypeScript
AI/ML
Cursor
Claude Code
LLM
Agentic Workflows
DevOps
AWS
Platform Engineering
Apply
SDET Engineer(Python) 5 hours ago
$19k – $46k per year (Estimated) • In office • Full-Time • 7+ years exp • Gurgaon
Python
Java
DevOps
Linux
Management
Agile
QA
Selenium
Cucumber
Pytest
Apply
Equity • Remote • Full-Time
AI/ML
Stable Diffusion
Fine-tuning
ComfyUI
Kling
Apply
$10k – $29k per year (Estimated) • In office • Full-Time • Bucharest
Apply
$99k – $199k per year (Estimated) • Equity • Remote • Full-Time
PHP
PHP
Symfony
Doctrine
DevOps
Rest API
WebSockets
Management
Stripe
Apply
$137k – $226k per year (Estimated) • Equity • Remote • Full-Time
Go
PHP
Apply
Engineering Manager 1 month ago
$96k – $195k per year (Estimated) • Equity • Remote • Full-Time • Master's Degree • London
Python
Go
PHP
Rust
DevOps
CI/CD
Apply
See all jobs
This is one of many
700,010 more open roles from verified company boards, updated every day.