{"id":2120224,"url":"https://alion.io/job/ig-senior-platform-sre","title":"Senior Platform SRE","company":{"id":38407,"name":"IG","domain":"ig.com","url":"https://alion.io/company/ig","size_band":"5000+","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Kraków, Poland"],"countries":["PL"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":63000,"max_usd":106000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":94},"experience_years_min":6,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AI Agents","optional":false},{"name":"Amazon EKS","optional":false},{"name":"AWS","optional":false},{"name":"Azure AKS","optional":false},{"name":"Chaos Engineering","optional":false},{"name":"CI/CD","optional":false},{"name":"Datadog","optional":false},{"name":"Dynatrace","optional":false},{"name":"Error Budget","optional":false},{"name":"Google GKE","optional":false},{"name":"Grafana","optional":false},{"name":"Honeycomb","optional":false},{"name":"Incident Management","optional":false},{"name":"Java","optional":false},{"name":"Kubernetes","optional":false},{"name":"PagerDuty","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Python","optional":false},{"name":"Self-Healing","optional":false},{"name":"ServiceNow","optional":false},{"name":"SLI/SLO/SLA","optional":false},{"name":"Terraform","optional":false},{"name":"Azure","optional":true},{"name":"GCP","optional":true}],"status":"live","first_seen_at":"2026-08-13T00:00:00Z","employer_posted_date":"2026-08-13","last_verified_at":"2026-10-09T23:27:48Z","board_verified":true,"closed_at":null,"days_open":58,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":58},"description":"Job Title\nSenior Platform SREJob Description\nSo, who are we?\nIG is a FTSE 100 fintech operatingacross five continents, serving over 1.3m customers and handling billions of dollars in transactions - built on scale, trust, and proof. We didn'tpivot to innovation; it'show we'vealways operated. What that means for the people who work here is real: genuinely complex problems to solve, the technology and resources to tackle them properly, and the kind of scope that'srare in established businesses.\nThe bar is high - bring a curious and forward-thinking mindset and we'llgive you the platform to define what comes next. Join us at IG - the future gets built here.\nYour team\nThe Platform SRE team is the engine of IG’s reliability programme. We sit within Infrastructure & Operations, working across IG’s hybrid estate of on-premisesHashiCorpNomad and AWS.\nWe are not a reactive ops team. We build the platform, standards, and tooling that make reliability the default for every engineering team at IG. Through the SRE Guild, we connect with Domain SREs and Reliability Champions across the organisation, setting the bar and lifting it together.\nYour role in the Team's Success\nYou will be a hands-on technical contributor at the heart of the Platform SRE team, owning pieces of the reliability platform that hundreds of engineers depend on. You will work at the intersection of software engineering, observability, and systems reliability, turning reliability from a reactive concern into a proactive engineering discipline.\nYou will partner with Platform Engineering, product teams, and Reliability Champions to define what good looks like in production and then make it the default. You will contribute to the SRE Guild, mentor engineers across the organisation, and when things go wrong, you will be on the call helping to mitigate, understand, and prevent a repeat.\nWhat you'll do\nBuild and own the reliability platform\nImplement comprehensive monitoring and observability usingOpenTelemetryand distributed tracing.Maintain SLO, error budgetsand burn-rate tracking\n\nEstablish andmaintain24/7 operational readiness including automated deployments, blue/green releases, and zero-downtime patching strategies\n\nEngineer self-healing capabilities: auto-remediation, error-budget-gated rollback, and automated traffic rerouting\n\nDesign and run chaos experiments across the AWS estate, turning severe-but-plausible failure scenarios into engineering improvements\n\nBuild automation tools and CI/CD pipelines that embed reliability practices, whileapplyingsoftware engineering discipline including version control, code reviews, and testing.\n\nContribute to the SRE AI agent, IG’s agentic tooling for incident investigation and reliability review, built on AWS frontier models\n\nMentor junior SREs and Reliability Champions on reliability patterns and production engineering discipline\n\nSet and uphold standards\nAuthor and evolve the SRE standards that underpin the Guild: SLOmethodology, error budget policy, observability instrumentation guide, andProduction Readiness Review(PRR)checklist\n\nMentor developers on reliability patterns including circuit breakers, retry logic, and fault tolerance\n\nWork with development teams and Reliability Champions to design SLOs on customer journeys rather than per-service.\n\nAssistand guide teamsin system design, capacity planning,architecturalreviewsandclosingobservability gaps.\n\nOwn incident response and learning\nFacilitate blameless post-incident reviews (PIRs) within five working days using contributing-factormethodology\n\nMaintain the Lessons Register, track remediation actions to closure, and surface patterns across incidents quarterly\n\nWhat you'll need for this role:\n6+ years of experience in all the below mentioned areas.\nEssential Technical Skills\nObservability and instrumentation: hands-onOpenTelemetryexperience (spans, metrics, traces, context propagation) and production use of Honeycomb, Datadog, Dynatrace, or Grafana; able to instrument Java or Python services directly.\n\nSLOs and error budgets: proventrack recorddesigning customer-meaningful SLIs, setting error budgets, configuring multi-window burn-rate alerts, and working with development teams on reliability measurement\n\nCI/CD and release engineering: experience building pipelines with safety mechanisms: blue/green and canary releases, automated rollback, and DORA metrics integration\n\nContainer orchestration: Kubernetes (EKS, AKS, or GKE)required;HashiCorpNomad is a strong advantage on IG’s hybrid estate; solid understanding of cloud networking andIaC(Terraform preferred)\n\nSoftware engineering: production-quality coding in Java and/or Python; comfortable contributing to application codebases to implement reliability patterns, not just configuring infrastructure around them\n\nDistributed systems: strong understanding of how large-scale systems fail and how to make them fail safely; circuit breakers, bulkheads, idempotency, graceful degradation, and load-shedding; high-throughput, low-latency environments preferred\n\nIncident management: on-call experience on production systems, blameless PIR facilitation, contributing-factor analysis, and driving action items to closure; PagerDuty and ServiceNow familiarity helpful\n\nChaos engineering: experience designing and executing hypothesis-driven experiments with blast-radius controls and gap-to-impact-tolerance analysis; AWS FIS, Gremlin, or equivalent\n\nCommunity and standards: at ease in a guild or community-of-practice model; comfortable writing RFCs, presenting at engineering forums, and building standards that others will adopt\n\nExperience Requirements\nTrack recordin high-throughput, production environments (financial services, trading platforms, or similar mission-critical systems preferred)\n\nDemonstrated ability to improve system reliability and performance at scale\n\nExperience working collaboratively with development teams to implement observability and reliability improvements\n\nStrong troubleshooting skills in distributed systems environments\n\nCore Competencies\nSystems thinking approach to problem-solving\n\nExcellent communication skills for cross-functional collaboration and technical enablement\n\nAbility to balance hands-on development work with operational responsibilities\n\nStrong bias toward automation andeliminatingmanual toil\n\nHow we work\nWe try to take a thoughtful approach to our ways of working as a company. We follow a hybrid working model with 3 days in the office -- which we think balances the need to collaborate effectively and connect with each other. When it comes to how we deliver, there are 5 things we want everyone to do to drive high performance, better learning and career satisfaction:\nLead and Inspire: Drives trust, alignment, and enthusiasm\nThink Big: Focus on the problems that most impact commercial outcomes\nChampion the client: Understand and prioritise client's needs\nDeliver at pace: Push for fast, sustainable growth;\nRaise the bar: Take ownership, be accountable and share feedback\nWe believe that diversity is vital to success, it fuels creativity, drives innovation and sets us up for global success. We're committed to building teams with a variety of perspectives and skills to help us realise our vision and strategy, that's why we encourage applications from people with diverse backgrounds and experiences to join us on this journey. Learn more about our D&I approach here.\nThe Perks\nYour growth fuels our success! Thrive with tailored development programs, mentoring opportunities with leaders, and clear career progression. Expand your network through committees, sports and social clubs. Enjoy extra time off for volunteering and community work.\nLearn more about the Perks here!\nJoin us for this exciting journey. Apply now!\nNumber of openings\n1","description_format":"text","description_chars":7790,"description_truncated":false,"requirements":{"experience_years_min":6,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Hybrid work"],"hiring_locations":[{"name":"Poland","iso":"PL","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Online Brokerage & Trading Platforms","Forex & CFD Brokers"],"lifecycle":[{"event":"open","at":"2026-10-08T23:55:15Z"}],"visa":[],"liveness":{"score":19,"band":"cold","label":"Long shot","p_open":1,"p_active":0.557,"p_room":0.35,"age_days":57,"expected_fill_days":22,"reasons":["conf:2","win:tail"],"computed_at":"2026-10-09T06:01:00Z"},"pay":null,"html_url":"https://alion.io/job/ig-senior-platform-sre","json_url":"https://alion.io/job/ig-senior-platform-sre.json","meta":{"generated_at":"2026-10-10T01:11:37Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":2042,"day_limit":5000,"remaining_today":2958,"minute_limit":60,"resets_at":"2026-10-11T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":38407},"rest":"https://alion.io/mcp/rest/get_company?id=38407"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fig-senior-platform-sre"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fig-senior-platform-sre"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fig-senior-platform-sre"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/ig-senior-platform-sre\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fig-senior-platform-sre"}]}