{"id":1250548,"url":"https://alion.io/job/searce-sr-engineer-cloud-reliability-engineer","title":"Sr Engineer - Cloud Reliability Engineer","company":{"id":180441,"name":"Searce","domain":"searce.com","url":"https://alion.io/company/searce","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":null,"truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Coimbatore, India"],"countries":["IN"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":30000,"max_usd":65000,"period":"year","method":"role_seniority_country_cell","sample_n":8},"experience_years_min":3,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Active Directory","optional":false},{"name":"Amazon EKS","optional":false},{"name":"AWS","optional":false},{"name":"Azure","optional":false},{"name":"Azure AKS","optional":false},{"name":"CI/CD","optional":false},{"name":"Commander.js","optional":false},{"name":"Crossplane","optional":false},{"name":"DNS","optional":false},{"name":"Docker","optional":false},{"name":"GCP","optional":false},{"name":"GitOps","optional":false},{"name":"Google GKE","optional":false},{"name":"Helm","optional":false},{"name":"Incident Management","optional":false},{"name":"Istio","optional":false},{"name":"Kubernetes","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Terraform","optional":false},{"name":"Windows","optional":false},{"name":"Zero Trust","optional":false},{"name":"JavaScript","optional":true},{"name":"Node JS","optional":true}],"status":"live","first_seen_at":"2026-09-24T04:41:55Z","employer_posted_date":null,"last_verified_at":"2026-09-24T04:41:55Z","board_verified":false,"closed_at":null,"days_open":2,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":2},"description":"Senior Cloud Site Reliability Engineer (CSRE) – Azure\n\nAbout Searce:\nSearce is an AI-native, engineering-led modern technology consultancy that empowers\nclients to futurify their businesses by delivering real, intelligent business outcomes. As a\ntrusted partner for over 3,000 clients globally, Searce specializes in cloud modernization,\ndata engineering, applied AI, and robust cloud platform security. Driven by a \"HAPPIER\"\ncultural mindset and our proprietary evlos problem-solving framework, we eliminate\nbureaucratic fluff to build working prototypes fast and scale enterprise production\nenvironments intelligently. We don't just fix systems; we leverage multi-cloud technologies\nto transform client operations into distinct competitive advantages.\nPosition Overview:\nWe are looking for a high-caliber Senior or Lead Cloud Site Reliability Engineer (CSRE) to\narchitect, secure, and stabilize next-generation hybrid and multi-cloud environments.\nOperating at the intersection of infrastructure design, security compliance, and production\noperations, you will serve as the technical Subject Matter Expert (SME) across GCP, Azure,\nand AWS.\nWhether optimizing a microservice mesh on GKE, tuning autoscaling on AKS, or driving a\nmassive disaster recovery drill across AWS regions, your focus will be absolute reliability. For\nthe Lead path, you will couple this deep engineering toolkit with stakeholder management\nand mentorship to drive an elite operational culture.\n\nExperience & Level Expectation:\nYears of Experience: 3 to 10 years of intensive, hands-on production operations\nexperience in a dedicated DevOps, Cloud Platform Engineering, or SRE role.\nAssociate level (3-5 Years): Expected to show flawless execution of IaC, advanced\ntriaging of infrastructure failures, and ownership of the CI/CD and deployment\nlifecycles.\nIntermediate level (5-10 Years): Expected to take architectural ownership, serve as\nprimary Incident Commander for complex outages, design cross-cloud governance\nframeworks, and act as a reliable bridge between technical teams and client\nleadership.\n\nKey Responsibilities & Role Expectations:\nMulti-Cloud Platforms & Orchestration: Design, configure, and maintain\nproduction-grade Kubernetes clusters across major platforms (AKS).\nManage advanced network routing, service meshes (e.g., Istio), and multi-tenant\nisolation.\nInfrastructure as Code (IaC) & GitOps: Build declarative, enterprise-grade, reusable\ninfrastructure components using Terraform or Crossplane. Standardize automated\nenvironment provisioning to eliminate configuration drift across multi-branch\nenvironments.\nIncident Management & Reliability (SRE): Own and optimize the production on-call\nrotation. Lead rapid mitigation strategies for Sev-1/Sev-2 system outages, reducing\nMean Time to Recovery (MTTR) through centralized log and metric correlation.\nRoot Cause Analysis (RCA): Facilitate rigorous, blameless post-incident reviews to\nidentify core architectural vulnerabilities and establish long-term fixes preventing\nrecurrence.\nLifecycle, Patching & Upgrades: Plan and execute zero-downtime cluster upgrades,\noperating system patching strategies (Linux/Windows), database lifecycle updates,\nand multi-region Disaster Recovery (DR) failover drills.\nCore Core Operations & Legacy Integration: Manage enterprise-level hybrid\nnetworking architecture (VPCs, Firewalls, Load Balancers, DNS routing, and DHCP\nconfigurations) while effectively connecting cloud native services to legacy\ninfrastructures like Active Directory.\nSecurity & Governance: Embed Zero Trust policies, secure secrets management\n(Secrets Manager/Key Vault), and continuous vulnerability patching into the\nautomated SDLC pipeline.\n\nRequired Technical Skills:\n- Microsoft Azure: Azure Virtual Machines, Virtual Networks, Azure Active Directory, Azure Update Management.\n- Containers & Orchestration\nProduction-level management of GKE, AKS, and EKS.\nAdvanced mastery of Docker, Helm, Kubernetes StatefulSets, Pod Disruption\nSkills\nWindows Azure, AKS, DevOps, Microsoft Windows Azure","description_format":"text","description_chars":4046,"description_truncated":false,"requirements":{"experience_years_min":3,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Artificial Intelligence","Science & Engineering","Engineering Services"],"lifecycle":[{"event":"open","at":"2026-09-25T18:00:00Z"}],"liveness":{"score":86,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.86,"p_room":1,"age_days":2,"expected_fill_days":24,"reasons":["seen:2","win:early"],"computed_at":"2026-09-26T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/searce-sr-engineer-cloud-reliability-engineer","json_url":"https://alion.io/job/searce-sr-engineer-cloud-reliability-engineer.json","meta":{"generated_at":"2026-09-27T02:57:44Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":2681,"day_limit":5000,"remaining_today":2319,"minute_limit":60,"resets_at":"2026-09-28T00:00:00Z"}}}