369,078open jobs
9,456companies
47,986added this week
Browse all
Location
In office (Hyderabad)
Seniority
Principal · 10+ years exp
Overview
Company
Impact
Profile match
Entain is a Computer Software company that engages in providing online and retail gaming in regulated markets. It owns brands like Bwin, Ladbrokes, Gala Coral, Partypoker, FoxyBingo, Sportingbet, etc.

We are seeking a Principal Site Reliability Engineer (SRE) to set the technical direction for the reliability, performance, and scalability of our hybrid infrastructure. As the most senior technical authority in the SRE function, you will define architecture and engineering standards across cloud and on-premises environments, drive the adoption of modern SRE practices, and ensure operational excellence. This is a hands-on, deeply technical role (~90%) with significant cross-organizational influence and technical mentorship responsibilities (~10%). You will operate as a force multiplier, raising the technical bar for the entire engineering organization rather than managing a team directly.

The core responsibilities for the job include the following:

Technical Responsibilities (~90%)

  • Architect scalable, secure, and cost-efficient cloud and hybrid solutions, and establish the patterns and reference implementations other teams build on.
  • Define the observability strategy across the organization: monitoring, logging, alerting, SLIs/SLOs, and error budgets using Datadog, Elastic, OpenTelemetry, and similar tooling.
  • Set standards for high availability, disaster recovery, and backup across hybrid environments, and validate them through resilience and failure testing.
  • Partner with development, security, and platform teams to shape deployment pipelines (CI/CD) and GitOps workflows at scale.
  • Establish and maintain organization-wide technical documentation, runbooks, architectural decision records, and operational standards.
  • Identify systemic reliability bottlenecks, lead complex root-cause analysis, and drive down operational load through automation and platform improvements.
  • Troubleshoot and resolve the most complex, high-impact production issues, and lead major incident response.
  • Provide deep 3rd-line technical support and participate in the on-call rotation as an escalation point.

Technical Leadership & Influence (~10%)

  • Act as a technical mentor and role model for SREs and engineers across teams, raising the overall engineering bar through coaching and knowledge-sharing.
  • Lead structured up-skilling and enablement across Windows, Linux, AWS, and Kubernetes.
  • Influence and align Product, Engineering, Security, Platform, and Operations teams around reliability goals and technical direction.
  • Shape the SRE roadmap together with SRE Leadership and SRE Guild and drive prioritization of key reliability and platform initiatives.
  • Communicate technical strategy and trade-offs clearly to both engineering teams and senior stakeholders, and drive accountability for reliability outcomes.
  • Champion a culture of operational excellence, ownership, and continuous improvement across the organization.

Requirements:

  • 10+ years of experience in SRE, DevOps, platform engineering, or cloud engineering, with a track record of operating at a senior or principal level.
  • Deep, hands-on expertise with on-prem and large-scale AWS environments.
  • Expert-level experience running Kubernetes in production (EKS, EKS Anywhere, and EKS Hybrid preferred), including cluster design, scaling, and hardening.
  • Proven background leading migrations of IIS/. NET and Linux-based applications to cloud and Kubernetes.
  • Expert-level Infrastructure-as-Code skills, particularly Terraform, including module design and standards for large teams.
  • Deep knowledge of GitOps (ArgoCD), automated deployments, and configuration management.
  • Exceptional troubleshooting skills across Windows, Linux, networking, and distributed systems.
  • Strong experience with observability platforms (Datadog, Prometheus, OpenTelemetry) and defining SLI/SLO/error-budget practices.
  • Strong understanding of security best practices in cloud and hybrid environments.
  • Experience operating in 24/7 high-availability, mission-critical environments.
  • Excellent communication and cross-team collaboration skills, with the ability to influence technical decisions without direct authority.
  • Demonstrated success mentoring senior engineers and driving org-wide technical standards.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
369,078 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Hyderabad
$220k – $325k per year • Remote/Hybrid • Full-Time • 15+ years exp • New York
DevOps
AWS
Azure
CI/CD
GCP
Kubernetes
Platform Engineering
Cybersecurity
Threat Modeling
Apply
Remote/Hybrid • 7+ years exp • Bachelor's Degree
C#
JavaScript
Python
SQL
TypeScript
C#
ASP.NET Core
Blazor
Dapper
Databases
Oracle
Frontend
Angular
Bootstrap
JQuery
Vue.js
DevOps
AWS
Azure
Azure DevOps
CI/CD
Git
Apply
$26k – $66k per year (Estimated) • In office • Full-Time • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$140k – $253k per year (Estimated) • In office • 5+ years exp • Long Beach
C++
Rust
C++
Protobuf
AI/ML
Human-in-the-Loop
DevOps
CI/CD
Vector
Apply
$21k – $57k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Noida
C#
SQL
TypeScript
JavaScript
C#
.NET
Frontend
Angular
DevOps
Azure
Azure DevOps
CI/CD
Git
TeamCity
Apply
Remote/Hybrid • Full-Time • Hyderabad
Bash
Groovy
Java
Python
SQL
JavaScript
Java
Maven
Databases
PostgreSQL
AI/ML
Replicate
Frontend
npm
DevOps
AWS
Azure
Bamboo
CI/CD
Docker
GCP
Jenkins
Kubernetes
Cybersecurity
SBOM
Snyk
Sonatype Nexus IQ
Apply
Backend Engineer 1 day ago
In office • 2+ years exp • Hyderabad
Java
Python
Ruby
SQL
Ruby
Ruby on Rails
DevOps
Docker
Git
Kubernetes
Rest API
Apply
In office • 7+ years exp • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
Sr. Data Engineer 1 day ago
In office • 5+ years exp • Hyderabad
Python
SQL
Python
pySpark
Databases
Apache Kafka
Databricks
Delta Lake
MS SQL
Snowflake
Microsoft Fabric
AI/ML
Claude
Copilot
dbt
Spark
DevOps
Azure
CI/CD
Git
SLI/SLO/SLA
GitHub
Game Dev
Unity
Analytics
ETL/ELT
Apply
In office • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
See all jobs
This is one of many
369,078 more open roles from verified company boards, updated every day.