1,409,356open jobs
81,787companies
210,025added this week
Browse all
Salary
≈ $63k – $106k per year (Estimated)
Location
Hybrid (Kraków, Poland)
Seniority
Senior · 6+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 9, 2026. First seen by Alion on Aug 13, 2026.

Overview
Company
Impact
Profile match

IG

Bereik meer omdat wij u meer bieden. U ziet de mogelijkheden, wij bieden de tools. Trade CFD's, Knock-Outs en opties op ons krachtige platform. Geniet van ons innovatieve productaanbod en prijswinnende service.

Job Title

Senior Platform SRE

Job Description

So, who are we?

IG is a FTSE 100 fintech operatingacross five continents, serving over 1.3m customers and handling billions of dollars in transactions - built on scale, trust, and proof. We didn'tpivot to innovation; it'show we'vealways operated. What that means for the people who work here is real: genuinely complex problems to solve, the technology and resources to tackle them properly, and the kind of scope that'srare in established businesses.

The bar is high - bring a curious and forward-thinking mindset and we'llgive you the platform to define what comes next. Join us at IG - the future gets built here.

Your team

The Platform SRE team is the engine of IG’s reliability programme. We sit within Infrastructure & Operations, working across IG’s hybrid estate of on-premisesHashiCorpNomad and AWS.

We are not a reactive ops team. We build the platform, standards, and tooling that make reliability the default for every engineering team at IG. Through the SRE Guild, we connect with Domain SREs and Reliability Champions across the organisation, setting the bar and lifting it together.

Your role in the Team's Success

You will be a hands-on technical contributor at the heart of the Platform SRE team, owning pieces of the reliability platform that hundreds of engineers depend on. You will work at the intersection of software engineering, observability, and systems reliability, turning reliability from a reactive concern into a proactive engineering discipline.

You will partner with Platform Engineering, product teams, and Reliability Champions to define what good looks like in production and then make it the default. You will contribute to the SRE Guild, mentor engineers across the organisation, and when things go wrong, you will be on the call helping to mitigate, understand, and prevent a repeat.

What you'll do

Build and own the reliability platform

  • Implement comprehensive monitoring and observability usingOpenTelemetryand distributed tracing.Maintain SLO, error budgetsand burn-rate tracking

  • Establish andmaintain24/7 operational readiness including automated deployments, blue/green releases, and zero-downtime patching strategies

  • Engineer self-healing capabilities: auto-remediation, error-budget-gated rollback, and automated traffic rerouting

  • Design and run chaos experiments across the AWS estate, turning severe-but-plausible failure scenarios into engineering improvements

  • Build automation tools and CI/CD pipelines that embed reliability practices, whileapplyingsoftware engineering discipline including version control, code reviews, and testing.

  • Contribute to the SRE AI agent, IG’s agentic tooling for incident investigation and reliability review, built on AWS frontier models

  • Mentor junior SREs and Reliability Champions on reliability patterns and production engineering discipline

Set and uphold standards

  • Author and evolve the SRE standards that underpin the Guild: SLOmethodology, error budget policy, observability instrumentation guide, andProduction Readiness Review(PRR)checklist

  • Mentor developers on reliability patterns including circuit breakers, retry logic, and fault tolerance

  • Work with development teams and Reliability Champions to design SLOs on customer journeys rather than per-service.

  • Assistand guide teamsin system design, capacity planning,architecturalreviewsandclosingobservability gaps.

Own incident response and learning

  • Facilitate blameless post-incident reviews (PIRs) within five working days using contributing-factormethodology

  • Maintain the Lessons Register, track remediation actions to closure, and surface patterns across incidents quarterly

What you'll need for this role:

6+ years of experience in all the below mentioned areas.

Essential Technical Skills

  • Observability and instrumentation: hands-onOpenTelemetryexperience (spans, metrics, traces, context propagation) and production use of Honeycomb, Datadog, Dynatrace, or Grafana; able to instrument Java or Python services directly.

  • SLOs and error budgets: proventrack recorddesigning customer-meaningful SLIs, setting error budgets, configuring multi-window burn-rate alerts, and working with development teams on reliability measurement

  • CI/CD and release engineering: experience building pipelines with safety mechanisms: blue/green and canary releases, automated rollback, and DORA metrics integration

  • Container orchestration: Kubernetes (EKS, AKS, or GKE)required;HashiCorpNomad is a strong advantage on IG’s hybrid estate; solid understanding of cloud networking andIaC(Terraform preferred)

  • Software engineering: production-quality coding in Java and/or Python; comfortable contributing to application codebases to implement reliability patterns, not just configuring infrastructure around them

  • Distributed systems: strong understanding of how large-scale systems fail and how to make them fail safely; circuit breakers, bulkheads, idempotency, graceful degradation, and load-shedding; high-throughput, low-latency environments preferred

  • Incident management: on-call experience on production systems, blameless PIR facilitation, contributing-factor analysis, and driving action items to closure; PagerDuty and ServiceNow familiarity helpful

  • Chaos engineering: experience designing and executing hypothesis-driven experiments with blast-radius controls and gap-to-impact-tolerance analysis; AWS FIS, Gremlin, or equivalent

  • Community and standards: at ease in a guild or community-of-practice model; comfortable writing RFCs, presenting at engineering forums, and building standards that others will adopt

Experience Requirements

  • Track recordin high-throughput, production environments (financial services, trading platforms, or similar mission-critical systems preferred)

  • Demonstrated ability to improve system reliability and performance at scale

  • Experience working collaboratively with development teams to implement observability and reliability improvements

  • Strong troubleshooting skills in distributed systems environments

Core Competencies

  • Systems thinking approach to problem-solving

  • Excellent communication skills for cross-functional collaboration and technical enablement

  • Ability to balance hands-on development work with operational responsibilities

  • Strong bias toward automation andeliminatingmanual toil

How we work

We try to take a thoughtful approach to our ways of working as a company. We follow a hybrid working model with 3 days in the office -- which we think balances the need to collaborate effectively and connect with each other. When it comes to how we deliver, there are 5 things we want everyone to do to drive high performance, better learning and career satisfaction:

  • Lead and Inspire: Drives trust, alignment, and enthusiasm
  • Think Big: Focus on the problems that most impact commercial outcomes
  • Champion the client: Understand and prioritise client's needs
  • Deliver at pace: Push for fast, sustainable growth;
  • Raise the bar: Take ownership, be accountable and share feedback

We believe that diversity is vital to success, it fuels creativity, drives innovation and sets us up for global success. We're committed to building teams with a variety of perspectives and skills to help us realise our vision and strategy, that's why we encourage applications from people with diverse backgrounds and experiences to join us on this journey. Learn more about our D&I approach here.

The Perks

Your growth fuels our success! Thrive with tailored development programs, mentoring opportunities with leaders, and clear career progression. Expand your network through committees, sports and social clubs. Enjoy extra time off for volunteering and community work.

Learn more about the Perks here!

Join us for this exciting journey. Apply now!

Number of openings

1
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,409,356 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Kraków
≈ $61k – $103k per year (Estimated) • In office • Full-Time • 6+ years exp • Warsaw
Databases
PostgreSQL
Redis
AI/ML
LLM
LLM Guardrails
DevOps
Terraform
GitHub Actions
Azure
CI/CD
AWS
Management
Agile
Scrum
Apply
IT Architect 3 months ago
up to $92k per year (net) • In office • Warsaw
Python
Java
Databases
Apache Kafka
Google BigQuery
BigQuery
DevOps
Terraform
GCP
GitHub
Apply
DevOps 3 hours ago
≈ $36k – $84k per year (Estimated) • In office • Full-Time • 4+ years exp • Warsaw
Python
SQL
Bash
Groovy
DevOps
Terraform
Ansible
GCP
Prometheus
Azure
CI/CD
Jenkins
AWS
Docker
Kubernetes
Grafana
Linux
Unix
Apply
≈ $62k – $105k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Warsaw
Python
PowerShell
Bash
Databases
MS SQL
DevOps
Terraform
Ansible
GCP
Helm
Azure DevOps
Azure
CI/CD
AWS
Kubernetes
TeamCity
Octopus Deploy
Windows
DNS
Cybersecurity
Qualys Cloud Platform
Veracode
Apply
Senior Linux Engineer 3 hours ago
≈ $61k – $103k per year (Estimated) • In office • Full-Time • Warsaw
Python
Java
Bash
Java
Spring Boot
Apache Tomcat
DevOps
Ansible
Red Hat
Linux
Apache HTTP Server
Cybersecurity
PKI
Management
Confluence
Apply
Systems Engineer 1 day ago
≈ $64k – $107k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Kraków
DevOps
GCP
VMWare
Azure
Windows Server
AWS
IAM
Cybersecurity
Okta
Microsoft Entra ID
Apply
DevOps Engineer 1 day ago
$35k – $44k per year (gross) • Hybrid • Confidential • 3+ years exp • Kraków
Python
PowerShell
DevOps
Splunk
Terraform
Ansible
Azure DevOps
VMWare
Azure
Jenkins
Kubernetes
Self-Healing
Linux
Windows
Management
ServiceNow
Power Automate
Power Apps
ITIL
Apply
$46k – $55k per year • Hybrid • Full-Time • Kraków
Python
JavaScript
TypeScript
AI/ML
Copilot
Cursor
Claude Code
Model Context Protocol
Gemini
LLM
RAG
OpenAI
Anthropic
OpenAI Codex
Chips/EDA
PoC Library
Management
n8n
Zapier
Apply
≈ $41k – $89k per year (Estimated) • Hybrid • 5+ years exp • Bachelor's Degree • Kraków
Analytics
Master Data Management
Management
Agile
Apply
$82k – $102k per year • Hybrid • Full-Time • Kraków
JavaScript
Java
Kotlin
TypeScript
Java
Spring Boot
Kotlin
Mockito
Databases
MySQL
PostgreSQL
Apache Kafka
Frontend
Redux
Webpack
GraphQL
Babel
React.js
ESLint
Mobile
JUnit
State Management
DevOps
GCP
WebSockets
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Management
Agile
Scrum
QA
Jest
Apply
See all jobs
This is one of many
1,409,356 more open roles from verified company boards, updated every day.