899,872open jobs
55,626companies
151,463added this week
Browse all
Salary
$180k – $250k per year
Location
Remote (likely United Kingdom)
Seniority
Principal · 10+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 28, 2026. First seen by Alion on Sep 16, 2026.

Overview
Company
Impact
Profile match
Built for enterprise-scale integration, MetaRouter enables real-time customer data collection, privacy-first identity resolution, and seamless activation across AI, retail media networks, and digital ecosystems.

About The Role

As a Principal Site Reliability Engineer, you set the reliability strategy for the platform. You will define how we build, deploy, observe, and operate a distributed system that runs both in our own cloud and inside customer-controlled environments - and you will hold the organization to that standard.

We run dedicated, isolated environments per customer, which makes repeatability and automation the central engineering problem rather than an afterthought. Depth of judgment about reliability engineering matters far more here than experience with any particular cloud, orchestrator, or observability vendor.

This is an individual contributor role with organization-level influence.

Core Responsibilities

  • Own the reliability architecture of the platform: deployment topology, failure domains, capacity strategy, and the automation that makes environments reproducible.

  • Define service level objectives with product and engineering leadership, and drive the work needed to meet them.

  • Set the standard for observability - dashboards, logs, metrics, tracing, and alerting - so that issues are detected before customers report them.

  • Lead major incidents, run blameless postmortems, and make sure the corrective work actually lands.

  • Contribute to design and architecture across infrastructure and applications, with automation, performance, reliability, and security as first-class concerns.

  • Drive infrastructure lifecycle at scale: provisioning, upgrades, and decommissioning across many isolated environments.

  • Ensure infrastructure and applications meet or exceed enterprise compliance requirements, and design identity and access controls across platforms and services.

  • Partner with enterprise customers on custom infrastructure requirements, translating their constraints into repeatable patterns rather than one-off work.

  • Raise the bar through code and design review, and mentor SREs and product engineers on reliability practice.

  • Improve and maintain infrastructure and process documentation.

  • Participate in and help evolve the on-call rotation, including how the team balances operational load against project work.

Qualifications and Experience

  • 10+ years in infrastructure, SRE, or platform engineering, including deep experience operating large-scale distributed systems in production.

  • Expertise designing, analyzing, and troubleshooting distributed systems, with a track record of reliability decisions that held up under growth.

  • Deep experience with at least one major public cloud provider, and the ability to reason across providers rather than within one.

  • Strong command of container orchestration: cluster operation, workload scheduling, networking, and the failure modes that come with them.

  • Fluency with infrastructure as code, configuration management, and CI/CD pipeline design.

  • Strong scripting and automation ability, and comfort reading and debugging application code in the languages your services are written in.

  • Experience defining observability strategy - instrumentation, query languages, dashboards, and alert design that minimizes noise.

  • Demonstrated ability to influence without authority and align multiple teams behind a technical direction.

  • Experience operating under enterprise security and compliance frameworks.

  • Solid understanding of Unix/Linux operating systems and networking fundamentals.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
899,872 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
In your city
$70k – $140k per year • Remote (United States) • Full-Time • 7+ years exp • Bachelor's Degree • Columbus • Atlanta • Austin • Dallas • Cleveland
DevOps
Platform Engineering
TCP/IP
DNS
DHCP
Management
Microsoft Teams
Apply
$144k – $245k per year • Remote (United States) • Public Trust • Full-Time • 15+ years exp • Bachelor's Degree • Reston
DevOps
GCP
Azure
CI/CD
AWS
Platform Engineering
GitHub
DNS
VPN
Cybersecurity
CIS Benchmarks
NIST 800-53
FedRAMP
Zero Trust
Analytics
Power BI
Management
Slack
Google Workspace
ServiceNow
Power Automate
Power Apps
SharePoint
Agile
ITIL
Apply
$116k – $191k per year • Remote (United States) • Full-Time • 10+ years exp • Bachelor's Degree • Georgia
DevOps
GCP
Azure
AWS
Apply
$16k per year • Remote (Russia) • Full-Time • Rostov-on-Don
1C
DevOps
VMWare
Prometheus
Grafana
Proxmox VE
Linux
Windows
VPN
VLAN
Cybersecurity
КриптоПро
Apply
≈ $89k – $170k per year (Estimated) • Remote (United Kingdom) • Full-Time
C#
C#
.NET
Databases
PostgreSQL
DevOps
Terraform
CI/CD
AWS
Kubernetes
Amazon EKS
Amazon EC2
Amazon S3
IAM
Management
Agile
Apply
≈ $30k – $74k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Moscow
Python
C++
C++
CMake
AI/ML
OpenCV
DevOps
Git
Linux
Unix
Apply
≈ $97k – $226k per year (Estimated) • Equity • Remote (Poland) • Full-Time • 7+ years exp • PhD
Python
Java
Rust
TypeScript
C++
AI/ML
Anthropic
DevOps
GCP
OpenTelemetry
Azure
CI/CD
GitOps
AWS
Kubernetes
Platform Engineering
Management
Agile
Apply
≈ $80k – $219k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Petah Tikva
Python
Go
JavaScript
Kotlin
TypeScript
C++
Scala
Databases
MySQL
PostgreSQL
Frontend
React.js
DevOps
GCP
CI/CD
AWS
Management
Agile
Apply
≈ $14k – $35k per year (Estimated) • Remote (EAEU) • Moscow
DevOps
Linux
Wi-Fi
Game Dev
Meta Quest SDK
Chips/EDA
PoC Library
IoT
LoRaWAN
Zigbee
Apply
≈ $82k – $170k per year (Estimated) • Hybrid • 8+ years exp • Bachelor's Degree • Marietta
DevOps
Linux
Cybersecurity
Microsoft Sentinel
Google SecOps
SIEM
Apply
$150k – $220k per year • Remote (United States) • Full-Time • 6+ years exp
DevOps
CI/CD
Platform Engineering
Configuration Management
Linux
Unix
Analytics
ETL/ELT
Management
Agile
Apply
$140k – $170k per year • Remote (likely United States) • Full-Time • 10+ years exp
JavaScript
Lua
Apply
$180k – $240k per year • Remote (United States) • Full-Time • 4+ years exp
Apply
$160k – $210k per year • Remote (likely United Kingdom) • Full-Time • 5+ years exp
Go
JavaScript
SQL
Databases
PostgreSQL
Apache Kafka
Frontend
GraphQL
React.js
Material UI
DevOps
GCP
AWS
Docker
Analytics
ETL/ELT
Management
Agile
Apply
$290k – $300k per year • Remote (likely United States) • Full-Time • 7+ years exp
Apply
See all jobs
This is one of many
899,872 more open roles from verified company boards, updated every day.