368,910open jobs
9,449companies
47,822added this week
Browse all
Salary
$86k – $179k per year (Estimated)
Location
Remote/Hybrid (United Kingdom)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
AbcdefghijklmnopqrstuvwxyzABCDEFGHIJKLMNOPQRSTUVWXYZ ABOUT PHOTO APP What you write here is totally up to you, there is no right or wrong way to complete your 'About Us' page. We advise making your language friendly and approachable, while being informative and professional.

About Radiant

Radiant is redefining how AI infrastructure is built.

We design and operate AI-native cloud platforms engineered for sovereignty, performance, and scale. Our infrastructure powers GPU-native workloads, multi-tenant control planes, and high-performance AI systems designed for the most demanding environments.

We are not building a generic cloud. We are building purpose-built AI infrastructure - from powered land, to compute, to software .

As we scale our platform and expand our engineering organisation, we are looking for leaders who can build strong teams, uphold high standards, and deliver reliably at pace.

Role Responsibilities

  • Deploy and Manage Kubernetes Clusters, deployed at scale to support AI centric workloads, across both our bare metal clusters and via trusted partner infrastructure

  • Develop Kubernetes Manifests and Operators: Facilitate application deployments and maintain Kubernetes-native services for networking, storage, security, identity and infrastructure management

  • Optimize Linux system configuration including kernel, driver, filesystem and services to support workloads running via our orchestration layer

  • Build and maintain automation scripts and infrastructure as code to support platform lifecycle, as well as simplifying troubleshooting for Incident resolution and provision of tooling for our support organisation

  • Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement.

  • Maintain and enhance Radiant’s observability stack: Prometheus, Grafana, and custom monitoring integrations

  • Operate and support services in 24x7 production environments, including on-call rotation

  • Contribute to Incident postmortem analyses, root cause analysis, document learnings, and automate remediations

  • Mentor junior engineers and act as an Operational requirements consultant to other departments

  • Communicate technical decisions clearly to non-technical stakeholders and customers

  • Uphold a culture of: do, document, automate

  • Willingness to cross train with Platform Engineering/Platform SRE to fully support both our infrastructure and platform stacks.

  • Willingness to cross train with HPC Engineering, supported by NVIDIA to enhance our HPC supportability offering

Requirements

  • 5+ Years Proven experience in globally scaled, performance-intensive environments operating to a 24/7 support model in an SRE or equivalent role

  • 3+ years experience in both running, deploying and optimising orchestration platforms with a strong emphasis on Kubernetes

  • Expert-level Linux administration, especially Ubuntu distributions

  • Proficiency in system tuning, disk I/O optimization, and hardware-level performance tweaks

  • Strong networking fundamentals: TCP/IP, DNS, DHCP, VLANs, routing, switching

  • Strong experience with API interrogation

  • Strong experience with infrastructure scripting and automation (Bash, Python, Ansible)

  • Deep understanding of observability principles and tools (Prometheus, Grafana preferred)

  • Strong grasp of ITSM and service operation best practices

  • Excellent communication and mentorship skills

  • Comfortable interfacing with internal stakeholders and external customers

  • Bonus: Knowledge of running AI workloads via orchestration platforms

Bonus Requirements

  • Bachelor or Masters Level degree in Computer Science, Engineering or related field, or equivalent experience.

  • LPIC Certifications

  • ITIL Foundation level qualification or equivalent experience

  • Certified Kubernetes Administrator (CKA)

Qualities we look for:

  • You approach problems with a systems mindset - balancing practical execution with long-term scalability

  • You elevate the team, setting high standards for technical quality and engineering excellence.

  • You hold yourself and others accountable - giving direct feedback and expecting the same

  • You take initiative, owning challenges end-to-end and proactively driving solutions.

  • You invest in others, mentoring to build both capability and confidence.

Why should you join us?

What sets us apart is our blend of modern technology, competitive benefits, and an open, welcoming work culture that enables our people to thrive.

Here are just some of the great things you can expect from us:

  • 25 days of annual leave

  • A culture that emphasises results over hierarchy, process & ego: we place great emphasis on the quality, ingenuity and creativity of work.

  • Open communication, regular feedback: we value smooth collaboration, direct and actionable feedback, and believe that leading with empathy and a growth mindset makes us better together.

  • Learning Time: we all have dedicated learning time to focus on new skills, projects or interests that lay outside of your day-to-day job.

  • Health & Wellbeing: we want everyone to feel healthy and happy, so we offer private medical insurance via Bupa.

  • Cycle to Work Scheme: we're committed to building a sustainable business, so we encourage cycling to work.

  • Gympass subscription to a variety of gyms and wellbeing apps

  • Participation in the company shares program

  • Enhanced parental pay & leave

Diversity, Equality, Inclusion and Belonging

We are an equal opportunity employer and we strive to reduce unconscious bias throughout our hiring process. All applicants will be considered for employment without attention to ethnicity, religion, sexual orientation, gender identity, family or parental status, national origin, veteran, neurodiversity status or disability status. To ensure our recruitment processes provide an equal opportunity for all applicants to succeed, we encourage you to let us know if there are any adjustments that we can make.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,910 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
DevOps Engineer 8 hours ago
$49k – $88k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Warsaw
Bash
Python
DevOps
AWS
Azure
Azure AKS
CI/CD
Docker
GCP
Git
GitHub
GitHub Actions
GitOps
Grafana
Kubernetes
Loki
Prometheus
Terraform
Thanos
Apply
$141k – $265k per year (Estimated) • Equity • Remote • Contractor • 8+ years exp
Python
Ruby
DevOps
AWS
Blue-Green Deployment
CI/CD
GCP
GitOps
Grafana
Incident Management
Kubernetes
OpenTelemetry
Prometheus
Pulumi
Terraform
Cybersecurity
FedRAMP
NIST 800-53
SentinelOne
Apply
$173k – $260k per year • In office • Full-Time • PhD • San Francisco
JavaScript
Node JS
Python
Python
Celery
Django
Flask
Databases
RabbitMQ
Redis
AI/ML
Agentforce
AI Agents
DevOps
Akamai
AWS
CI/CD
Cloudflare
CloudFormation
Helm
Jenkins
Kubernetes
Spinnaker
Terraform
Marketing
Salesforce
Apply
$117k – $177k per year • In office • Full-Time • 4+ years exp • PhD • Washington
Java
Python
Rust
Databases
Apache Kafka
HBase
Trino
AI/ML
Agentforce
AI Agents
Flink
Spark
DevOps
AWS
GCP
Kubernetes
Marketing
Salesforce
Apply
$191k – $297k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • Seattle
AI/ML
AI Agents
DevOps
AWS
Azure
CI/CD
GCP
IAM
Platform Engineering
Cybersecurity
Least Privilege
Microsoft Entra ID
PCI DSS
Threat Modeling
Zero Trust
Apply
$80k – $191k per year (Estimated) • In office • Full-Time • London
DevOps
Incident Management
Apply
Engineering Manager 25 days ago
$126k – $227k per year (Estimated) • Remote/Hybrid • Full-Time • London
DevOps
CI/CD
Kubernetes
Apply
Cluster Architect 28 days ago
$91k – $215k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • London
C++
AI/ML
CUDA
CUDA Toolkit
InfiniBand
DevOps
Docker
Docker Swarm
Kubernetes
SLURM
HPC
Apply
DevOps Engineer 2 months ago
$69k – $171k per year (Estimated) • Remote/Hybrid • Full-Time • London
Go
Python
DevOps
Ansible
Blue-Green Deployment
CI/CD
Git
GitHub Actions
Helm
Kubernetes
Progressive Delivery
Terraform
GitHub
Apply
Senior SDET 2 months ago
$59k – $146k per year (Estimated) • Remote/Hybrid • Full-Time • London
Go
Databases
PostgreSQL
AI/ML
KServe
LLM
vLLM
DevOps
CI/CD
gRPC
Helm
Kubernetes
Apply
See all jobs
This is one of many
368,910 more open roles from verified company boards, updated every day.