{"id":1172825,"url":"https://alion.io/job/amgen-site-reliability-engineer-data-platforms","title":"Site Reliability Engineer – Data Platforms","company":{"id":4775,"name":"Amgen","domain":"amgen.com","url":"https://alion.io/company/amgen","size_band":"5000+","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":{"grade":"B","score":80,"open_postings":337,"ghost_share":0,"stale_share":0.816,"repost_share":0.003,"time_to_fill_p50_days":24,"computed_at":"2026-09-25T05:45:01Z"}},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Hyderabad, India"],"countries":["IN"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":27000,"max_usd":58000,"period":"year","method":"role_seniority_country_cell","sample_n":8},"experience_years_min":5,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Amazon CloudWatch","optional":false},{"name":"Amazon EC2","optional":false},{"name":"Amazon EKS","optional":false},{"name":"Amazon S3","optional":false},{"name":"AWS","optional":false},{"name":"AWS Lambda","optional":false},{"name":"IAM","optional":false},{"name":"ITSM","optional":false},{"name":"SRE","optional":false},{"name":"Agile","optional":true},{"name":"Amazon ECS","optional":true},{"name":"CI/CD","optional":true},{"name":"Docker","optional":true},{"name":"ITIL","optional":true},{"name":"Jira","optional":true},{"name":"Kubernetes","optional":true},{"name":"Platform Engineering","optional":true}],"status":"live","first_seen_at":"2026-09-24T08:55:07Z","employer_posted_date":"2026-09-24","last_verified_at":"2026-09-25T22:27:26Z","board_verified":true,"closed_at":null,"days_open":1,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":1},"description":"Career Category\nEngineeringJob Description\nJoin Amgen’s Mission of Serving Patients\nAt Amgen, if you feel like you’re part of something bigger, it’s because you are. Our shared mission-to serve patients living with serious illnesses-drives all that we do.\nSince 1980, we’ve helped pioneer the world of biotech in our fight against the world’s toughest diseases. With our focus on four therapeutic areas -Oncology, Inflammation, General Medicine, and Rare Disease- we reach millions of patients each year. As a member of the Amgen team, you’ll help make a lasting impact on the lives of patients as we research, manufacture, and deliver innovative medicines to help people live longer, fuller happier lives.\nOur award-winning culture is collaborative, innovative, and science based. If you have a passion for challenges and the opportunities that lay within them, you’ll thrive as part of the Amgen team. Join us and transform the lives of patients while transforming your career.\nAbout Amgen\nAmgen harnesses the best of biology and technology to fight the world’s toughest diseases, and make people’s lives easier, fuller and longer. We discover, develop, manufacture and deliver innovative medicines to help millions of patients. Amgen helped establish the biotechnology industry more than 40 years ago and remains on the cutting-edge of innovation, using technology and human genetic data to push beyond what’s known today.\nAbout the Role\nEDSE Data Platforms is seeking a Site Reliability Engineer who is equally comfortable improving production systems in AWS and establishing how services should be operated.\nOur SRE team manages AWS platform operations and the infrastructure supporting applications across EDSE. As our service portfolio grows, we need an experienced individual contributor who can create clear, practical operating arrangements across SRE, application teams, Security, business stakeholders, and other service partners.\nThis is a deliberately balanced role, with responsibilities split approximately:\n50% service governance, operational readiness, resilience, and process leadership\n50% hands-on AWS site reliability engineering\nThe balance will be measured over a typical quarter and may vary during major incidents, application handovers, continuity exercises, audits, and delivery periods.\nThis is not a policy-only, PMO, or ITSM process-administration role. You will remain hands-on with AWS infrastructure, automation, observability, incident response, and reliability improvement while creating lightweight, enforceable operational controls.\nWhat you will do\nRoles & Responsibilities:\nService governance and operating model - approximately 50%\nEstablish and operate a consistent application handover process, with readiness criteria covering ownership, architecture, dependencies, monitoring, access, security findings, runbooks, recovery, known risks, knowledge transfer, hypercare, and formal acceptance.\nIdentify incomplete or unsafe handovers and recommend deferring acceptance. Ensure exceptions are documented, time-bound, and approved by the appropriate business, service, security, or risk owner.\nMaintain a service catalogue and operating profile for every supported application, including criticality, owners, recurring tasks, dependencies, support hours, access requirements, backup coverage, recovery objectives, and escalation paths.\nDefine clear responsibilities, decision rights, support boundaries, and segregation of duties across SRE, application engineering, Product, Security, business owners, vendors, and other stakeholders.\nAgree service indicators and internal objectives, proposed SLAs, support hours, severity definitions, response and restoration expectations, maintenance windows, recovery expectations, and escalation paths.\nLead business-continuity and disaster-recovery planning, including business-impact analysis, dependency mapping, recovery procedures, technical recovery tests, and remediation tracking.\nEstablish repeatable security-review and audit-readiness processes covering technical evidence, control self-assessments, access reviews, vulnerability remediation, and audit findings while preserving the independence of Security and Audit.\nDefine standards and ownership for essential operational knowledge, including service profiles, architecture and dependency records, runbooks, incident playbooks, recovery procedures, access models, and known-risk registers.\nLead regular service and reliability reviews, turning operational performance, incidents, recovery readiness, security actions, and risks into prioritised improvements with accountable owners.\nHands-on AWS site reliability engineering - approximately 50%\nOperate and improve secure, reliable, scalable, and cost-effective application infrastructure on AWS.\nBuild and maintain infrastructure as code, delivery pipelines, and automation that removes repetitive operational effort.\nImprove observability through meaningful metrics, logs, traces, dashboards, alerts, and service-health reporting.\nParticipate in on-call and incident response, including diagnosis, mitigation, recovery, communication, escalation, and blameless post-incident reviews.\nImprove availability, performance, capacity, resilience, backup and recovery, patching, infrastructure lifecycle management, and security posture.\nPartner with application teams on architecture, production readiness, deployment safety, application resilience, and operational quality.\nApply SRE practices such as SLIs, SLOs, error budgets, automation, and learning from failure to reduce repeat incidents, alert fatigue, manual toil, and recovery time.\nWhat we expect of you\nBasic Qualifications and Experience:\nMaster's or Bachelor’s degree in Computer Science, Engineering, Information Systems, or related field\n5-8 years of relevant experience\nMust-Have Skills:\nStrong understanding of AWS infrastructure, networking, identity and access management, security, monitoring, backup, and resilience patterns with depth in the services most relevant to data platforms: IAM, VPC, PrivateLink, S3, EC2, KMS, CloudWatch, CloudTrail, Secrets Manager, and STS. Working knowledge of EKS, Lambda, Glue, EMR, RDS will be good to have.\nExperience in operating business-critical production services on AWS.\nStrong hands-on experience in Site Reliability Engineering, Platform Engineering, Cloud Operations, or Infrastructure Engineering.\nPractical experience with infrastructure as code, CI/CD, operational automation, monitoring, observability, and incident response.\nDemonstrated experience establishing or improving service handovers, support models, operational ownership, service levels, business continuity, disaster recovery, or security-control processes.\nExperience defining and testing recovery objectives, recovery plans, and technical recovery procedures.\nUnderstanding of security controls, audit evidence, access governance, vulnerability management, risk acceptance, and segregation of duties.\nAbility to convert unclear cross-team responsibilities into practical, agreed, and measurable operating arrangements.\nStrong facilitation, negotiation, documentation, and stakeholder-management skills.\nAbility to influence technical, business, security, and leadership stakeholders without relying on direct reporting authority.\nA pragmatic approach to governance that improves accountability and reliability without introducing unnecessary bureaucracy.\nStrong analytical and problem-solving ability, with a record of turning complex or ambiguous platform challenges into scalable and reusable solutions.\nExperience working in Agile delivery environments and using planning and delivery tools such as Jira or Jira Align.\nGood-to-Have Skills:\nExperience supporting a portfolio of applications with different criticality and support requirements.\nExperience in a regulated, security-sensitive, or audit-intensive environment.\nFamiliarity with frameworks such as AWS Well-Architected, ITIL, and business-continuity standards.\nExperience with containers and orchestration platforms such as Docker, Amazon ECS, Amazon EKS, or Kubernetes.\nRelevant AWS, SRE, IT service-management, security, risk, or business-continuity certifications, including:AWS Certified Solutions Architect\nAWS Certified Security - Specialty\nSAFe Agilist or another SAFe certification\n\nFunctional Skills:\nExcellent written and verbal communication, with the ability to explain complex platform concepts, architecture decisions, risks, and trade-offs in clear, business-relevant language.\nStrong influencing and consensus-building skills, including the ability to establish standards and drive adoption across teams without relying solely on formal authority.\nA platform-product mindset focused on reusable capabilities, paved roads, developer experience, measurable outcomes, and long-term platform health.\nStrong systems-thinking and structured problem-solving skills, with the ability to diagnose issues across application, platform, cloud, governance, security, and operating-model boundaries.\nHigh degree of ownership, initiative, and follow-through, with the ability to move ambiguous topics from exploration through decision, implementation, adoption, and continuous improvement.\nCollaborative and globally minded, with experience working effectively across architecture, governance, cybersecurity, operations, engineering, product, and business teams.\nStrong planning, estimation, prioritization, and execution skills, with the ability to manage multiple initiatives while maintaining high standards for security, reliability, quality, and reusability.\nAbility to balance innovation with enterprise risk, distinguishing between experimentation, limited preview adoption, and production-ready capabilities.\nStrong coaching and enablement skills, with the ability to create clear documentation, facilitate technical workshops, mentor engineers, and build an active platform community.\nA growth mindset and commitment to continuous learning, modern engineering practices, constructive challenge, and responsible adoption of AI.\nWhat you can expect of us\nAs we work to develop treatments that take care of others, we also work to care for your professional and personal growth and well-being. From our competitive benefits to our collaborative culture, we’ll support your journey every step of the way.\nIn addition to the base salary, Amgen offers competitive and comprehensive Total Rewards Plans that are aligned with local industry standards.\nApply now\nEQUAL OPPORTUNITY STATEMENT\nAmgen is an Equal Opportunity employer and will consider all qualified applicants for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability status, or any other basis protected by applicable law.\nWe will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.\nReady to Apply for the Job?\nWe highly recommend utilizing Workday's robust Career Profile feature to complete the application process. A link to update your profile is available when you click Apply. You can then complete your Workday profile in minutes with the “Upload My Experience” functionality to upload an updated copy of your resume or you can simply edit the individual sections of your Career Profile.\nPlease note that you should be in your current position for at least 18 months before applying to internal positions. Staff must notify their current manager if invited for an interview. In addition, Staff are ineligible to apply for open positions if (a) their performance is currently being managed on a performance improvement plan (PIP) or other locally utilized formal coaching document or (b) their most recent performance rating was not a “Partially Meets Expectations” o...","description_format":"text","description_chars":12162,"description_truncated":true,"requirements":{"experience_years_min":5,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"bachelor","optional":false},"security_clearance":false,"languages":[]},"benefits":["Continuous learning"],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Prescription Drugs","Biologics & Biosimilars","Immunology & Inflammation Therapeutics","Cardiometabolic Therapeutics"],"lifecycle":[{"event":"open","at":"2026-09-24T08:55:07Z"}],"liveness":{"score":63,"band":"ok","label":"Likely open","p_open":1,"p_active":0.632,"p_room":1,"age_days":0,"expected_fill_days":24,"reasons":["conf:5","stale_co","velocity","win:early","comp:brand"],"computed_at":"2026-09-25T05:45:01Z"},"pay":null,"html_url":"https://alion.io/job/amgen-site-reliability-engineer-data-platforms","json_url":"https://alion.io/job/amgen-site-reliability-engineer-data-platforms.json","meta":{"generated_at":"2026-09-26T00:51:47Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":792,"day_limit":5000,"remaining_today":4208,"minute_limit":60,"resets_at":"2026-09-27T00:00:00Z"}}}