406,377open jobs
14,133companies
78,486added this week
Browse all
Salary
$40k – $100k per year (Estimated)
Location
In office (Sant Cugat del Vallès, Madrid, Poland)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Roche is a Swiss healthcare group founded in Basel in 1896 and is unusual in operating two large divisions of comparable importance: prescription medicines and in vitro diagnostics. The pharmaceutical business is built on oncology, neurology, immunology, ophthalmology and haemophilia, with products such as Ocrevus, Hemlibra, Perjeta and Vabysmo, while the diagnostics division supplies laboratory analysers, molecular tests and sequencing systems to hospitals worldwide. The group owns the American biotechnology company Genentech outright and the Japanese firm Chugai in majority, is controlled by a family shareholder pool, and spends among the largest research budgets in the industry.

At Roche you can show up as yourself, embraced for the unique qualities you bring. Our culture encourages personal expression, open dialogue, and genuine connections, where you are valued, accepted and respected for who you are, allowing you to thrive both personally and professionally. This is how we aim to prevent, stop and cure diseases and ensure everyone has access to healthcare today and for generations to come. Join Roche, where every voice matters.

The Position

The Position

We are building a global Site Reliability Engineering (SRE) team to support critical commercial and internal platforms and applications. As an SRE, you will help design, build, and scale reliable distributed systems that power healthcare innovation worldwide.

This role is focused on reliability, scalability, automation and operational excellence. You will influence system design, define reliability standards and reduce operational toil through engineering solutions.

This role includes participation in a structured on-call rotation.

Who We Are

At Roche, we are passionate about transforming patients’ lives, and we are bold in both decision and action - we believe that good business means a better world. That is why we come to work every single day. We commit ourselves to scientific rigor, unassailable ethics and access to medical innovations for all. We do this today to build a better tomorrow.

Roche is strongly committed to a diverse and inclusive workplace. We strive to build teams that represent a range of backgrounds, perspectives and skills. Embracing diversity enables us to create a great place to work and to innovate for patients.

Step into the Future of IT with Roche!

As a seasoned Site Reliability Engineer (SRE) at Roche, you will leverage your deep software engineering expertise to propel our products to new heights of robustness, scalability and reliability. This isn't just a role-it's an invitation to shape the backbone of technological innovations forward.

Your Mission

Design and maintain cutting-edge tools, scripts and frameworks that automate repetitive tasks, streamline software deployment and manage expansive systems with unparalleled efficiency. Partner closely with forward-thinking development teams to architect and implement high-performance solutions that elevate system efficiency, optimize resource utilization and enhance deployment processes for superior uptime and user satisfaction.

Your Impact

Lead the charge in incident management and response. Detect system anomalies, troubleshoot swiftly and conduct thorough root cause analyses to prevent recurring issues.

Champion continuous improvement by refining monitoring and alerting mechanisms, conducting insightful post-incident reviews and embedding best practices in software lifecycle management. Your strategic foresight and meticulous planning will ensure our systems are not only reliable but also superlatively performant.

By joining our elite team, you will play a pivotal role in delivering seamless experiences to our end-users, exceeding business and customer demands, and solidifying Roche's reputation as a leader in IT innovation.

Your Core Responsibilities

Reliability Engineering & Architecture

  • Define and implement SLIs, SLOs, and error budgets with product and engineering teams

  • Conduct reliability reviews for new and existing services

  • Design scalable, fault-tolerant architectures in AWS and Azure environments

  • Lead capacity planning, performance and cost optimization initiatives

  • Improve system resilience through automation and self-healing patterns

  • Drive organizational observability maturity (metrics, logs, traces, alert quality)

Incident Management & Continuous Improvement

  • Perform complex root cause analysis and drive rapid mitigation

  • Participate in blameless postmortems and follow-through

  • Improve MTTR, reduce incident frequency, and elevate production standards

  • Collaborate seamlessly with engineering teams to enable timely and effective resolutions

  • Handle requests and incidents, create and maintain runbooks

  • Participation in a structured 24*7 on-call rotation

Automation & Platform Engineering

  • Reduce operational toil through tooling and automation (Python or similar)

  • Improve CI/CD reliability and deployment safety mechanisms

  • Build and maintain infrastructure-as-code (Terraform or equivalent)

  • Enhance Kubernetes platform reliability (EKS, AKS, or similar)

Cross-Functional Leadership

  • Partner with business, engineering, security, and cloud teams to embed reliability early in the software development life cycle

  • Mentor mid-level engineers and help shape SRE best practices

  • Championing a culture of ownership, accountability, and continuous improvement

Who You Are:

  • Minimum bachelor’s degree in computer science, Engineering, or a related field, or equivalent professional experience.

  • Experience in either site reliability engineering, software engineering or related fields with production on-call experience.

  • Solid experience with AWS and/or Azure, including setting up, monitoring, and maintaining cloud resources (incl. Kubernetes, EKS, AKS, GKE, etc knowledge).

  • Proficiency with observability tools

  • Hands-on experience with incident management tools

  • Proficiency in scripting languages for automation purposes

  • Demonstrated proficiency in troubleshooting, especially in cloud and distributed system environments

  • Excellent communication, teamwork and documentation skills, with a proactive and self-motivated approach to improving system reliability and operational efficiencies.

  • We value and encourage candidates from diverse backgrounds and experiences, believing that diverse perspectives drive innovation and success.

  • Excelling in both spoken and written English communication.

Where pay transparency applies, details are provided based on the primary posting location. For this role, the primary location is Sant Cugat del Vallès. If you are interested in additional locations where the role may be available, we will provide the relevant compensation details later in the hiring process.

Who we are

A healthier future drives us to innovate. Together, more than 100’000 employees across the globe are dedicated to advance science, ensuring everyone has access to healthcare today and for generations to come. Our efforts result in more than 26 million people treated with our medicines and over 30 billion tests conducted using our Diagnostics products. We empower each other to explore new possibilities, foster creativity, and keep our ambitions high, so we can deliver life-changing healthcare solutions that make a global impact.

Let’s build a healthier future, together.

Roche is an Equal Opportunity Employer.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
406,377 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Sant Cugat del Vallès
$117k – $146k per year • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree • Princeton
JavaScript
TypeScript
PHP
PHP
Drupal
WordPress
Frontend
GraphQL
Next.js
React.js
DevOps
AWS
Azure
GCP
Apply
$121k – $225k per year • Remote • Full-Time • Chicago
Python
SQL
Apex
Apex
MuleSoft
Databases
Db2
MongoDB
Neo4j
Oracle
AI/ML
AI Agents
Fine-tuning
Knowledge Graph
LangChain
Langfuse
LangGraph
Prompt Engineering
RAG
Multi-Agent Systems
DevOps
AWS
Azure
Rest API
Analytics
Tableau
Apply
Software Developer 3 days ago
$88k – $129k per year • Equity • In office • Full-Time • Toronto • Columbia
JavaScript
Python
TypeScript
DevOps
AWS
CI/CD
Apply
$70k – $105k per year • In office • Full-Time • 1+ year exp • Bachelor's Degree • Wichita
Bash
PowerShell
DevOps
Azure
Cybersecurity
Microsoft Entra ID
Management
Jira
ServiceNow
Apply
$160k – $250k per year • In office • Full-Time • Bachelor's Degree • Dallas
C++
DevOps
CI/CD
RTOS
Apply
$27k – $80k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Master's Degree • Madrid
Bash
PowerShell
Python
AI/ML
AI Agents
Model Context Protocol
DevOps
Ansible
CI/CD
Git
Jenkins
VMWare
GitHub
Apply
$23k – $61k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Hyderabad
Python
Python
FastAPI
Databases
DynamoDB
AI/ML
AI Agents
AWS Bedrock
AWS Strands Agents
Claude
LangChain
LangGraph
LLM
LLM Guardrails
Prompt Engineering
RAG
DevOps
Amazon ECS
Amazon EventBridge
Amazon S3
AWS
AWS Fargate
AWS Lambda
CI/CD
Docker
GCP
Git
Rest API
Terraform
Vector
Apply
$71k – $164k per year (Estimated) • In office • Full-Time • 8+ years exp • Switzerland
AI/ML
AI Agents
Knowledge Graph
LLM
DevOps
Platform Engineering
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Petaling Jaya
Apply
$35k – $70k per year (Estimated) • In office • Full-Time • 2+ years exp • Warsaw
Java
SQL
Kotlin
Java
Spring Boot
Spring Data JPA
Spring MVC
Spring Security
Kotlin
Mockito
Databases
PostgreSQL
AI/ML
AI Agents
Copilot
Cursor
Prompt Engineering
Mobile
JUnit
DevOps
AWS
Git
Apply
In office • Full-Time • 2+ years exp • Bachelor's Degree • Sant Cugat del Vallès
Analytics
Tableau
Apply
Remote/Hybrid • Full-Time • PhD • Sant Cugat del Vallès
Apply
Remote/Hybrid • Full-Time • PhD • Sant Cugat del Vallès
Apply
Remote/Hybrid • Full-Time • PhD • Sant Cugat del Vallès
Apply
Remote/Hybrid • Full-Time • PhD • Sant Cugat del Vallès
Apply
See all jobs
This is one of many
406,377 more open roles from verified company boards, updated every day.