700,010open jobs
41,477companies
98,955added this week
Browse all
Salary
$94k – $175k per year (Estimated)
Location
Remote (United Kingdom)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Build with generative AI on a unified inference API. Image generation, video generation, audio, 3D, and large language models. 400K+ models, managed infrastructure, usage-based pricing.

Runware is building the API layer for the next generation of AI products. Our platform gives teams fast, reliable access to real-time inference across thousands of models through a single flexible API. We help customers build and scale media generation products with better performance, lower cost, and less operational complexity.

Behind this is an infrastructure platform built for speed, reliability, and GPU scale. New models launch constantly. Customer traffic can grow quickly. Performance matters at every layer.

We are looking for a Staff/Senior DevOps Engineer to help build, operate, and scale the infrastructure behind Runware’s global AI inference platform. You’ll play a critical role in making our systems faster, more resilient, easier to operate, and ready for the next stage of growth.

About the role

Runware’s infrastructure is the engine behind some of the fastest-growing AI products in the world. As a Staff/Senior DevOps Engineer, you’ll help design, build, and operate the systems that power real-time AI inference across large-scale GPU fleets and a global production platform.

This is not a traditional DevOps role. You’ll be working at the intersection of bare-metal infrastructure, GPUs, networking, automation, observability, and high-performance distributed systems. Your work will directly shape how quickly we can launch new models, scale customer traffic, recover from failures, and deliver low-latency AI experiences to millions of users.

You’ll turn complex, hardware-driven infrastructure into reliable, automated, developer-friendly platforms. From provisioning and orchestration to deployment pipelines, monitoring, incident response, and capacity scaling, you’ll help remove friction so engineering teams can move faster without compromising reliability.

You’ll build the foundations that let Runware scale with confidence: infrastructure that is fast, resilient, observable, secure, and built for the demands of real-time AI.

What you’ll do

  • Build and scale the infrastructure that powers real-time AI inference across GPU fleets, bare-metal servers, serverless and containerised production systems
  • Help evolve Runware’s platform toward more elastic, on-demand infrastructure that can scale quickly with customer traffic and model demand
  • Make Runware faster, more reliable and more resilient by improving the critical paths behind our request entrypoints, inference services, queues, storage, load balancers and networking layer
  • Automate the hard parts of infrastructure operations, from provisioning and configuration through to CI/CD, deployment safety, progressive rollouts and rapid rollback
  • Build the observability backbone for a high-performance AI platform, with the signals needed to spot issues early, understand capacity and fix problems before customers feel them
  • Play a leading role in production operations, incident response, debugging and post-incident improvements, helping us turn operational challenges into a stronger platform
  • Strengthen the security and compliance foundations of our infrastructure through patching, secrets management, access controls, hardening, auditability, documentation and repeatable operational processes

Requirements

  • Strong experience as a DevOps Engineer, SRE, Infrastructure Engineer, Platform Engineer or similar, with a track record of running production systems at scale
  • Deep Linux knowledge and confidence debugging real production issues across networking, storage, performance, services and system behaviour
  • Hands-on experience building automation, Infrastructure-as-Code, CI/CD pipelines and deployment workflows that make infrastructure safer and easier to operate
  • Experience operating high-availability, low-latency or high-throughput platforms where reliability and performance directly affect customers
  • Strong networking fundamentals across TCP/IP, DNS, load balancing, routing, firewalls, proxies, TLS and HTTP
  • A calm and pragmatic approach under pressure, with strong communication, good judgement and a bias toward automation over manual toil

Bonus

  • Experience operating GPU infrastructure for AI/ML inference, including NVIDIA drivers, CUDA, container runtimes, GPU monitoring, capacity planning and workload isolation
  • Familiarity with inference serving and optimisation frameworks such as vLLM, TensorRT, Triton or similar

Benefits

We’re a remote-first collective, meeting in person twice a year to plan, brainstorm, celebrate wins, and enjoy some face-to-face time. We have core hours for cooperative working and calls, but outside of that your calendar is yours. Work the hours that let you perform at your peak while also building a healthy life.

Our release cycles are fast and intense, but they’re followed by real downtime. After big pushes we expect the team to unplug, recharge, and come back ready & stronger than ever for the next leap.

  • Generous paid time off - vacation, sick days, public holidays
  • Meaningful stock options - share in the upside you create
  • Remote-first setup - work from home anywhere we can employ you
  • Flexible hours - own your schedule outside core collaboration blocks
  • Family leave - paid maternity, paternity, and caregiver time
  • Company retreats - twice-yearly gatherings in inspiring locations
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
700,010 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$75k – $182k per year (Estimated) • Remote • Full-Time • 7+ years exp • Bachelor's Degree • United States
JavaScript
AI/ML
Model Context Protocol
Frontend
Next.js
React.js
DevOps
Rest API
CI/CD
SOAP
Apply
$63k – $166k per year (Estimated) • In office • Full-Time • 7+ years exp • Luxembourg City
JavaScript
Java
SQL
Java
Maven
Databases
Db2
Oracle
ElasticSearch
DevOps
Rest API
Azure DevOps
Kibana
Azure
CI/CD
Jenkins
Git
Docker
Kubernetes
Linux
Windows
Management
Agile
Scrum
Apply
$187k – $308k per year • Equity • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • United States
AI/ML
Machine Learning
DevOps
CI/CD
Cybersecurity
Zero Trust
Apply
$46k – $94k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Kraków
Python
JavaScript
Java
Node JS
Databases
MySQL
PostgreSQL
Apache Kafka
AI/ML
Copilot
Cursor
DevOps
CI/CD
Git
Apply
$85k – $135k per year (Estimated) • Remote/Hybrid • Full-Time • 7+ years exp • Bachelor's Degree • Kraków
Python
Go
Java
SQL
C++
AI/ML
Agentic Workflows
DevOps
gRPC
Splunk
GCP
Azure
CI/CD
AWS
Kubernetes
Apply
$103k – $207k per year (Estimated) • Equity • Remote • Full-Time • Bachelor's Degree
Python
C++
C++
PyTorch C++
AI/ML
LoRA
vLLM
CUDA Toolkit
Fine-tuning
Multimodal AI
Diffusion Models
PEFT
Transformers
PyTorch
CUDA
Triton
Edge AI
Machine Learning
DevOps
Docker
Kubernetes
GitHub
Apply
$10k – $29k per year (Estimated) • In office • Full-Time • Bucharest
Apply
$99k – $199k per year (Estimated) • Equity • Remote • Full-Time
PHP
PHP
Symfony
Doctrine
DevOps
Rest API
WebSockets
Management
Stripe
Apply
$137k – $226k per year (Estimated) • Equity • Remote • Full-Time
Go
PHP
Apply
Engineering Manager 1 month ago
$96k – $195k per year (Estimated) • Equity • Remote • Full-Time • Master's Degree • London
Python
Go
PHP
Rust
DevOps
CI/CD
Apply
See all jobs
This is one of many
700,010 more open roles from verified company boards, updated every day.