368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$38k – $96k per year (Estimated)
Location
In office (Madrid)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Roche Poland is a major pharmaceutical and diagnostics organization operating as part of the global F. Hoffmann-La Roche group. The company focuses on developing and delivering innovative medicines in oncology, neurology, immunology, and rare diseases, alongside advanced diagnostic systems for healthcare institutions. Through its local research and development capabilities - including a prominent global IT and data analytics center based in Poland - it plays a vital role in advancing personalized healthcare and clinical trials across the region.

At Roche you can show up as yourself, embraced for the unique qualities you bring. Our culture encourages personal expression, open dialogue, and genuine connections, where you are valued, accepted and respected for who you are, allowing you to thrive both personally and professionally. This is how we aim to prevent, stop and cure diseases and ensure everyone has access to healthcare today and for generations to come. Join Roche, where every voice matters.

The Position

Job description

As an Infrastructure Provisioning and Management Engineer within the Accelerated Compute Engineering (ACE) team, you will be responsible for overseeing and advancing our core infrastructure management and provisioning tech stack. This role has a strong focus on driving configuration-as-code, infrastructure-as-code (IaC), and modern automated provisioning best practices across our high-performance compute (HPC) and industry-leading AI Factory.

You will own the lifecycle, deployment, and optimization of bare-metal and virtualized compute environments that power Roche's advanced computing initiatives. By treating infrastructure strictly as code and eliminating manual configurations, you will ensure our advanced clusters are highly reproducible, securely patched, and rapidly scalable to meet the evolving demands of computational science and large-scale AI workloads.

Description of the area

Hosting and Infrastructure (HI) provides mission-critical on-premise infrastructure, cloud hosting, connectivity, and technology products that enable all functions at every Roche site to develop, innovate, connect, and deliver compliant digital products across the Roche Enterprise.

The Value Streams - Accelerated Compute Engineering (ACE) Team is focused on driving both customer success and platform success by acting as a center of excellence and delivery for the High Performance Compute and AI Infrastructure supporting AI and HPC use cases across Roche. This team facilitates seamless onboarding and adoption for business vertical customers needing accelerated compute-helping those infrastructure consumers with needs optimized for high availability, seamless data transfer, flexibility, speed, and the rapidly changing needs of AI-helping achieve rapid time-to-value.

Job Responsibilities

Automated Provisioning & Cluster Orchestration

  • Design, deploy, and manage large-scale automated provisioning systems for multi-node HPC and AI Factory environments.

  • Own and maintain the infrastructure management and provisioning tech stack underpinning the orchestration, monitoring, and provisioning of complex GPU and CPU workloads.

  • Streamline bare-metal provisioning and node imaging pipelines to ensure minimal downtime and rapid expansion capabilities.

Infrastructure-as-Code (IaC) & Configuration Governance

  • Enforce a strict configuration-as-code and infrastructure-as-code mindset, replacing manual interventions with repeatable automation scripts.

  • Author, review, and maintain complex Ansible playbooks and roles for configuration management, patch deployment, and compliance drift remediation.

  • Establish robust CI/CD pipelines using GitLab to test, validate, and deploy infrastructure changes safely across development, staging, and production clusters.

Operating System Engineering & Lifecycle Management

  • In partnership with Enterprise OS teams, standardize and manage operating system builds, with dual proficiency across HPC and AI Factory platforms.

  • Utilize solutions such as Red Hat Image Builder and NVIDIA Base Command Manager to create optimized, compliant, and secure custom golden images tailored for AI and high-performance computing workloads.

  • Manage OS lifecycles, including kernel tuning, automated package updates, and vulnerability management, ensuring alignment with global security standards.

Platform Reliability & Collaboration

  • Implement proactive monitoring and alerting for infrastructure provisioning health, node availability, and configuration drifts.

  • Address and help resolve complex, systemic infrastructure failures, contributing to post-mortem analyses to continuously improve platform resilience.

Qualifications

Education / Experience

  • Bachelor’s or an advanced degree in Computer Science, Computer Engineering, or a similar technical discipline.

  • 5+ years of experience in systems engineering, DevOps, or platform infrastructure roles, with a proven track record of managing enterprise Linux environments at scale.

  • Deep, practical knowledge of operating system internals for both RHEL and Ubuntu OS.

Technical & Business Skills:

  • Automation & Orchestration: Advanced capability with Ansible on the command line and experience building scalable infrastructure pipelines using GitLab CI/CD.

  • Provisioning Tooling: Experience using NVIDIA Base Command Manager (Bright Cluster Manager) and Red Hat Image Builder (or related tools like Kickstart/Satellite).

  • Modern Engineering Mindset: Strong adherence to git-based workflows, code-review methodologies, and infrastructure-as-code principles.

  • Troubleshooting Depth: Ability to isolate complex, multi-layered faults bridging hardware, kernel configurations, and automation scripts.

Leadership & Mindset:

  • Lean & Agile Mindset: Passionate about continuous improvement, eliminating technical debt, and automating repetitive tasks to achieve scale.

  • Collaboration & Communication: Strong collaborative skills with an enterprise mindset, capable of working fluidly across team boundaries to drive platform success.

  • Intellectual Curiosity: Highly self-motivated to explore and adopt emerging technologies in the fast-evolving landscape of HPC and AI infrastructure engineering

Who we are

A healthier future drives us to innovate. Together, more than 100’000 employees across the globe are dedicated to advance science, ensuring everyone has access to healthcare today and for generations to come. Our efforts result in more than 26 million people treated with our medicines and over 30 billion tests conducted using our Diagnostics products. We empower each other to explore new possibilities, foster creativity, and keep our ambitions high, so we can deliver life-changing healthcare solutions that make a global impact.

Let’s build a healthier future, together.

Roche is an Equal Opportunity Employer.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Madrid
AI Engineer 2 hours ago
$120k – $130k per year • In office • Full-Time • Texas
Python
SQL
Databases
Apache Kafka
Snowflake
AI/ML
Amazon SageMaker
Hadoop
DevOps
CI/CD
Git
GitLab
Analytics
ETL/ELT
Apply
$128k – $173k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • United States
JavaScript
Node JS
TypeScript
Frontend
React.js
DevOps
Amazon EC2
AWS
AWS CDK
AWS Lambda
CI/CD
GitHub Actions
Jenkins
Terraform
Amazon CloudWatch
Amazon S3
API Gateway
GitHub
GitLab
IAM
Apply
$35k per year • In office • Internship • Bachelor's Degree • Freiburg im Breisgau
C++
Java
Node JS
Python
C#
JavaScript
C#
.NET
AI/ML
AI Agents
DevOps
CI/CD
GitLab
Apply
$99k – $135k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Washington
JavaScript
TypeScript
DevOps
Azure
Azure DevOps
CI/CD
Management
Power Apps
Power Automate
QA
Playwright
Apply
$58k – $136k per year (Estimated) • Remote/Hybrid • Full-Time • Guadalajara
Bash
Node JS
Python
JavaScript
Python
FastAPI
Databases
RabbitMQ
Frontend
Next.js
React.js
DevOps
ArgoCD
AWS
CI/CD
Datadog
Docker
Git
GitHub
GitHub Actions
GitOps
Grafana
Helm
Jenkins
Karpenter
KEDA
Kubernetes
OpenTelemetry
Prometheus
Terraform
Apply
In office • Full-Time • San José
Apply
$26k – $51k per year (Estimated) • In office • Full-Time • 8+ years exp • Hyderabad
Apex
Apex
MuleSoft
AI/ML
AI Agents
Model Context Protocol
DevOps
Azure
CI/CD
Apply
$35k – $87k per year (Estimated) • In office • Freelance • 6+ years exp • Bachelor's Degree • Vienna
JavaScript
Kotlin
SQL
Swift
Mobile
Espresso
JUnit
Realm
DevOps
AWS
CI/CD
GitHub
GitHub Actions
Cybersecurity
GDPR
HIPAA
Management
Confluence
QA
Appium
BrowserStack
Charles Proxy
JMeter
Postman
TestNG
TestRail
XCUITest
Apply
$84k – $212k per year (Estimated) • In office • Full-Time • 6+ years exp • Master's Degree • Basel
Python
DevOps
Docker
Git
GitLab
SLURM
Management
Jira
Apply
IT Technical Analyst 12 hours ago
$24k – $54k per year (Estimated) • In office • Full-Time • Madrid
AI/ML
Claude
Management
ServiceNow
Apply
$63k – $137k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Madrid • Barcelona
AI/ML
AI Agents
Apply
$36k – $93k per year (Estimated) • Equity • In office • Contractor • 3+ years exp • Madrid
C#
C#
.NET
Cybersecurity
GDPR
HIPAA
ISO 27001
PCI DSS
SOC 2
Apply
$62k – $142k per year (Estimated) • In office • Full-Time • 6+ years exp • PhD • Madrid
Apex
JavaScript
Python
TypeScript
Databases
Databricks
Google BigQuery
Snowflake
AI/ML
Agentforce
AI Agents
Claude
Cursor
LangChain
LlamaIndex
LLM
Prompt Engineering
Marketing
Salesforce
Apply
$43k – $110k per year (Estimated) • In office • Madrid
Python
SQL
Databases
Google BigQuery
DevOps
GCP
Google Cloud Run
Apply
Analytics Engineer 11 hours ago
$57k – $154k per year (Estimated) • Remote/Hybrid • Full-Time • London • Lisbon • Copenhagen • Madrid
SQL
Databases
Google BigQuery
AI/ML
dbt
DevOps
GCP
Git
Marketing
Amplitude
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.