{"id":1459795,"url":"https://alion.io/job/abbott-laboratories-site-reliability-engineer-ii","title":"Site Reliability Engineer II","company":{"id":2710,"name":"Abbott Laboratories","domain":"abbott.com","url":"https://alion.io/company/abbott","size_band":"5000+","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":{"grade":"A","score":100,"open_postings":99,"ghost_share":0.01,"stale_share":0,"repost_share":0.03,"time_to_fill_p50_days":15,"computed_at":"2026-10-01T05:45:00Z"}},"role":"DevOps","role_family":"DevOps","seniority":"middle","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Sylmar, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":81500,"max":141300,"currency":"USD","period":"year","gross":null,"usd_annual":141300},"salary_estimate":null,"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":".NET","optional":false},{"name":"Azure","optional":false},{"name":"Azure AKS","optional":false},{"name":"Azure DevOps","optional":false},{"name":"Bash","optional":false},{"name":"DNS","optional":false},{"name":"Incident Management","optional":false},{"name":"Kubernetes","optional":false},{"name":"Linux","optional":false},{"name":"PowerShell","optional":false},{"name":"Python","optional":false},{"name":"TCP/IP","optional":false},{"name":"C#","optional":true},{"name":"CI/CD","optional":true},{"name":"Datadog","optional":true},{"name":"Docker","optional":true},{"name":"FinOps","optional":true},{"name":"Grafana","optional":true},{"name":"HIPAA","optional":true},{"name":"Prometheus","optional":true}],"status":"live","first_seen_at":"2026-09-28T00:00:00Z","employer_posted_date":"2026-09-28","last_verified_at":"2026-10-01T14:12:13Z","board_verified":true,"closed_at":null,"days_open":3,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":3},"description":"Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices, nutritionals and branded generic medicines. Our 122,000 colleagues serve people in more than 160 countries.JOB DESCRIPTION:\nPosition Title: Site Reliability Engineer II\nTeam: CRM DevOps Employment Type: Full-Time\nAbout the Role\nThis Site Reliability Engineer II position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.\nWe are seeking a highly skilled and mission-driven Site Reliability Engineer (SRE) to join our DevOps team. In this critical role, you will be responsible for ensuring the reliability, scalability, performance, and operational excellence of Merlin.net - a remote monitoring platform designed to help doctors, cardiologists, and care teams automatically collect and review data from patients with implanted cardiac devices.\nThis isn't just about keeping servers up; it's about building and maintaining the resilient backbone for systems where failure is not an option, and where our success directly impacts patient care around the world. You will embed within our DevOps team, acting as a bridge between development and operations.\nWhat You'll Do\nDesign, implement, and maintain highly available, fault-tolerant, and resilient systems that meet demanding uptime and safety requirements.\nIdentify and eliminate performance bottlenecks in software and infrastructure, ensuring low-latency, high-throughput, and real-time responsiveness for customer-facing services.\nDefine, monitor, and uphold Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets.\nDevelop and implement comprehensive monitoring, logging, tracing, and alerting solutions to provide deep insights into system health and behavior at scale.\nAutomate away manual operational tasks, from provisioning and deployment to testing and recovery.\nDevelop and implement strategies for scaling our services and infrastructure to meet evolving business demands, including distributed systems and cloud deployments in Azure.\nWork closely with software engineering, security, quality, and compliance teams to integrate SRE best practices into our operational processes and infrastructure, ensuring the integrity, availability, and confidentiality of our systems that carry sensitive patient health data.\nCreate clear, concise, and comprehensive documentation, runbooks, and playbooks for operational procedures. Lead blameless postmortem processes following incidents and drive systematic follow-through on action items.\nWork with a multi-disciplinary team on challenging problems in a fast-paced environment, contributing across architecture reviews, incident response, capacity planning, and reliability roadmap planning.\nRequired Qualifications\nBachelor's in Computer Science, Software Engineering, Systems Engineering, or a related technical discipline; equivalent professional experience will be considered.\nExcellent communication skills with the demonstrated ability to work effectively in cross-functional teams, translate technical complexity for non-technical stakeholders, and collaborate with development, quality, security, marketing, and regulatory teams.\nStrong analytical, problem-solving, and debugging skills with a methodical and structured approach to diagnosing complex, distributed system issues under pressure - including production incidents with patient safety implications.\n Proficiency in at least one systems or automation programming language (e.g., Python, Go, Bash, PowerShell) for building tooling, automation, and operational systems.\n Demonstrated expertise with Microsoft Azure - including Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, and related managed services.\nSolid Linux & networking fundamentals - DNS, TCP/IP, HTTP/S, TLS, load balancing, and networking in cloud environments.\n Incident management experience - including on-call rotation participation, structured incident response, root cause analysis (RCA), and systematic prevention of recurrence.\nPreferred Qualifications\nExperience working in a regulated healthcare, medical device, or life sciences environment, with familiarity in compliance frameworks such as HIPAA.\n Relevant professional certifications such as: Microsoft Certified Azure DevOps Engineer Expert, Azure Solutions Architect Expert, Certified Kubernetes Administrator (CKA), or Certified Kubernetes Security Specialist (CKS).\nBackground in cost optimization for cloud-native architectures, including FinOps practices for Azure environments.\nDeep understanding of distributed systems - including load balancing, service meshes, microservices, message queues, and fault-tolerant design patterns.\nContainer orchestration expertise - hands-on production experience with Kubernetes and Docker at scale, including deployment strategies, resource management, and cluster operations.\nExperience designing and operating CI/CD pipelines for continuous delivery of software in production environments, including safe deployment strategies such as blue/green, canary, and feature flag-gated rollouts.\nObservability platform experience with tools such as Prometheus, Grafana, the ELK/EFK stack, Datadog, Azure Monitor, or similar enterprise-grade monitoring and tracing platforms.\nExperience contributing to, or driving disaster recovery (DR) design, business continuity planning, and tabletop exercises.\nThe base pay for this position is\n$81,500.00 - $141,300.00In specific locations, the pay range may vary from the range posted.\nJOB FAMILY:\nProduct DevelopmentDIVISION:\nCRM Cardiac Rhythm ManagementLOCATION:\nUnited States > Sylmar : 15900 Valley View CourtADDITIONAL LOCATIONS:\nWORK SHIFT:\nStandardTRAVEL:\nNoMEDICAL SURVEILLANCE:\nNoSIGNIFICANT WORK ACTIVITIES:\nContinuous sitting for prolonged periods (more than 2 consecutive hours in an 8 hour day), Keyboard use (greater or equal to 50% of the workday)Abbott is an Equal Opportunity Employer of Minorities/Women/Individuals with Disabilities/Protected Veterans.EEO is the Law link - English: http://webstorage.abbott.com/common/External/EEO_English.pdfEEO is the Law link - Espanol: http://webstorage.abbott.com/common/External/EEO_Spanish.pdf","description_format":"text","description_chars":6380,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"bachelor","optional":false},"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"United States","iso":"US","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Health Care","Prescription Drugs","Cardiovascular Devices","Neurotech & Neuromodulation"],"lifecycle":[{"event":"open","at":"2026-09-29T11:44:14Z"}],"liveness":{"score":87,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.875,"p_room":1,"age_days":3,"expected_fill_days":15,"reasons":["conf:2","velocity","win:early","comp:brand"],"computed_at":"2026-10-01T05:45:00Z"},"pay":{"stated_usd_annual":141300,"is_top_pay":true},"html_url":"https://alion.io/job/abbott-laboratories-site-reliability-engineer-ii","json_url":"https://alion.io/job/abbott-laboratories-site-reliability-engineer-ii.json","meta":{"generated_at":"2026-10-01T20:28:33Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":2946,"day_limit":5000,"remaining_today":2054,"minute_limit":60,"resets_at":"2026-10-02T00:00:00Z"}}}