1,431,920open jobs
84,427companies
218,560added this week
Browse all
Salary
≈ $85k – $195k per year (Estimated)
Location
Remote (location not specified)
Seniority
Senior · 5+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 10, 2026. First seen by Alion on Oct 9, 2026.

Overview
Company
Impact
Profile match
Madiff is an IT consulting and staff augmentation company that provides remote agile teams, nearshore development teams and individual IT contractors to enterprise clients, with most of its hiring run from Warsaw. It works across retail, banking and insurance, telecom and media, automotive, energy, transportation and life sciences, recruits consultants in Poland and Portugal and sells its services into the UK, Ireland, France, Switzerland, Northern Europe and the US. Its openings cover .NET and full stack developers, data engineers and data scientists, AI architects, system and business analysts, QA testers, project managers and IT sales staff.

This is a remote position.

We are looking for a Senior Site Reliability Engineer to support advanced AI platforms responsible for production-grade applications and pipelines. The role focuses on building and maintaining reliability, scalability, and operational excellence across multiple AI-driven systems.

The engineer will work on a central operational layer for monitoring and managing AI workloads, improving system stability, and reducing incidents. This is a hands-on role requiring direct involvement in diagnosing production issues, implementing fixes, and optimising monitoring, alerting, and CI/CD processes.

The position requires close collaboration with engineering teams to improve release quality, standardise telemetry, and ensure stable and predictable system behaviour in a distributed cloud environment.

Responsibilities

  • Build and maintain central monitoring and alerting layer for AI applications and pipelines
  • Define and implement SLIs, alerts, and operational dashboards
  • Manage incidents including triage, coordination, root cause analysis, and prevention
  • Standardise telemetry across systems including latency, throughput, and failures
  • Optimise CI CD pipelines and introduce quality gates for reliability
  • Work closely with engineering teams to reduce recurring issues and improve stability

Requirements

  • Minimum5+ years of experience in SRE, Platform, or Production Engineering
  • Strong hands on experience withKubernetes and production environments
  • Experience withAzure and Azure DevOps
  • Experience with monitoring tools such asDatadog
  • Strong understanding ofincident management and root cause analysis
  • Ability to build practical monitoring and alerting systems

Nice to have

  • Experience withAI or LLM pipelines
  • Experience building monitoring platforms across multiple systems
  • Experience withGrafana
  • Experience working in large scale or distributed environments

Expectations

  • Strong ownership mindset and accountability for system stability
  • Proactive approach to identifying risks and improvements
  • Hands on engineer actively working with systems, not only coordinating
  • Comfortable working in dynamic and evolving environments

Benefits

  • Solid, competitive salary
  • Work in a multinational environment on international projects
  • Comprehensive healthcare
  • Long-term B2B contract with a stable project pipeline
  • Work model: fully remote
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,431,920 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
In your city
≈ $83k – $158k per year (Estimated) • Remote (United States) • Full-Time • 3+ years exp • Associate's Degree • New York
SQL
DevOps
Linux
Windows
Apply
≈ $35k – $89k per year (Estimated) • Remote (Hungary) • Hungary
Python
Go
AI/ML
LLM
DevOps
Terraform
Puppet
Ansible
Loki
VMWare
Prometheus
CI/CD
GitOps
AWS
Kubernetes
Grafana
Cybersecurity
FortiGate
Apply
≈ $117k – $246k per year (Estimated) • Remote (United States) • Full-Time
Python
SQL
Python
pySpark
Databases
Databricks
MS SQL
Apache Kafka
Azure Cosmos DB
Azure SQL Database
AI/ML
Spark
DevOps
Azure
Analytics
Azure Data Factory
Apply
$160k – $175k per year • Remote (United States) • Full-Time
DevOps
Terraform
GCP
Crossplane
Pulumi
Azure
AWS
Kubernetes
Apply
≈ $68k – $140k per year (Estimated) • Remote (Germany) • Full-Time • 3+ years exp
Python
JavaScript
C#
DevOps
Terraform
Ansible
Helm
CI/CD
GitOps
Kubernetes
Apply
≈ $88k – $173k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Alpharetta
TypeScript
SQL
C#
C#
.NET
Databases
Redis
Snowflake
Apache Kafka
Azure SQL Database
AI/ML
Copilot
AI Agents
OpenAI
DevOps
Splunk
Azure
CI/CD
Kubernetes
Azure AKS
GitLab
API Gateway
Analytics
SSIS
Management
Confluence
Agile
QA
Postman
SoapUI
Apply
$98k – $134k per year • Remote (United States) • Full-Time • 4+ years exp • Bachelor's Degree • United States
Python
AI/ML
Copilot
Claude
Model Context Protocol
AI Agents
DevOps
Terraform
Ansible
GCP
Azure DevOps
GitHub Actions
GitLab CI
Azure
CI/CD
AWS
VPN
BGP
Cybersecurity
Zscaler
HIPAA
Zero Trust
Management
Agile
ITSM
Apply
$118k – $162k per year • Remote (United States) • Full-Time • 5+ years exp • Bachelor's Degree • United States
DevOps
Azure DevOps
Azure
CI/CD
Platform Engineering
Cybersecurity
HIPAA
Management
Agile
Scrum
Apply
SCRUM MASTER 1 day ago
≈ $74k – $157k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Montreal
DevOps
Azure DevOps
Azure
Management
Agile
Scrum
Kanban
Apply
≈ $107k – $187k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Wilmington
Python
Java
SQL
PowerShell
Perl
Java
Apache Tomcat
Databases
MySQL
MariaDB
DevOps
Splunk
Red Hat
Dynatrace
Kubernetes
Linux
Windows
Unix
Management
ITIL
Apply
≈ $90k – $204k per year (Estimated) • Remote (location not specified) • Full-Time • 5+ years exp
Rust
C++
DevOps
ZooKeeper
etcd
Apply
Remote (location not specified) • Full-Time
Python
Databases
Databricks
DevOps
Terraform
Azure
Apply
≈ $58k – $158k per year (Estimated) • Remote (location not specified) • Full-Time • 3+ years exp
DevOps
Terraform
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Cybersecurity
GDPR
Apply
DevOps Engineer 1 day ago
Remote (location not specified) • Full-Time • 2+ years exp
PowerShell
DevOps
Terraform
Ansible
GCP
Helm
Azure DevOps
Azure
CI/CD
AWS
Kubernetes
TeamCity
Octopus Deploy
Linux
Windows
DNS
Cybersecurity
Qualys Cloud Platform
Veracode
Apply
≈ $102k – $224k per year (Estimated) • Remote (location not specified) • Full-Time
DevOps
Terraform
Azure DevOps
Azure
CI/CD
Git
Kubernetes
Bicep
Linux
Apply
See all jobs
This is one of many
1,431,920 more open roles from verified company boards, updated every day.