Location
In office
Seniority
Staff · 8+ years exp
Overview
Company
Impact
Profile match
Wij bieden in de Benelux een scala aan verzekeringsproducten voor particulieren, families en bedrijven. Wij hebben onder andere verzekeringsproducten op het gebied van Ongevallen & Welzijn, Financial Lines, Transport, Brand- en Bedrijfsschade, Aansprakelijkheid, Reizen en Overlijdensrisico.
Key Objective:
- Lead the Site Reliability Engineering function to define and drive the organisation’s reliability engineering strategy - bridging software development and operations through engineering discipline, not manual process.
- Own the end-to-end reliability posture of production systems: define SLO/SLI frameworks, govern error budgets, and enforce production-readiness standards to protect business continuity.
- Build, mentor, and scale a high-performing SRE team that prioritises engineering over toil - automating manual work, embedding reliability into the SDLC, and driving down mean time to recovery through systematic improvement.
- Champion observability-led engineering through full-stack Dynatrace adoption, AIOps integration, and data-driven reliability decision-making at every layer of the stack.
- Serve as the primary reliability engineering partner to development and platform leadership, shaping architecture decisions, release policies, and automation strategy.
Key Responsibilities:
- Define and drive the SRE strategy and multi-year roadmap aligned to business priorities.
- Lead and develop the SRE team, including hiring, onboarding, performance management, career development, and succession planning.
- Own incident management, including severity classification, escalation, response SLAs, and leadership of major incidents.
- Champion blameless postmortems, root cause analysis, and implementation of systemic fixes.
- Establish and govern SLOs, SLIs, and error budgets, ensuring reliability targets are aligned to business needs.
- Drive resilience engineering, including chaos engineering, GameDays, production readiness reviews, and failure mode analysis.
- Reduce toil through automation, improved runbooks, and continuous operational improvement.
- Own observability and alerting standards, including monitoring strategy, dashboards, and alert quality.
- Partner with engineering, architecture, product, and leadership teams to embed reliability into design and delivery.
- Represent the SRE function in senior forums and provide reporting on reliability, risk, and operational performance.
Qualifications:
- Degree in Computer Science, Software Engineering, IT, or a related technical field.
- 8+ years’ experience in software engineering, platform reliability, or SRE, including 3+ years in people leadership.
- Strong hands-on coding ability in Python, Go, or similar, with experience building automation and self-healing solutions.
- Proven experience leading enterprise-scale SRE or platform reliability functions.
- Experience defining and operating RTO/RPO and SLI/SLO frameworks.
- Strong background in observability, production readiness, error budgets, and chaos engineering.
- Experience leading on-call models, incident response, and executive stakeholder engagement.
- Solid understanding of SDLC, Agile, and DevOps delivery models.
- ITIL Foundation is desirable.
Managerial & Soft Skills:
- Proven people leader with experience building and coaching high-performing teams.
- Strategic thinker who can turn business priorities into reliability roadmaps.
- Strong communicator who can explain technical risk in business terms.
- Calm and decisive during major incidents.
- Influential partner across engineering, product, and leadership teams.
- Strong advocate for developer experience and sustainable on-call practices.
- Data-driven and able to balance reliability, speed, and cost.
- Champions psychological safety, continuous learning, and operational excellence.
Technical Skills:
- Expert in Dynatrace, with experience in observability, monitoring, SLOs, tracing, and log management.
- Proficient in Grafana, Prometheus, Splunk, ELK, Azure Monitor, and Log Analytics.
- Strong knowledge of OpenTelemetry and telemetry pipeline design.
- Experience with ServiceNow, CI/CD tools, Kubernetes, Docker, Terraform, and Bicep.
- Familiar with Java, .NET, databases, APIs, Kafka, and cloud platforms, especially Azure.
- Experience with AIOps, AI-assisted triage, and automation tooling.
- Able to support reliability engineering through scripting, auto-remediation, and operational automation.
Desired:
- Experience in insurance or financial services.
- Dynatrace, Azure, ITIL 4, or Google Cloud/SRE-related certifications.
- Experience with chaos engineering, AIOps, MLOps, and FinOps.
- Strong analytical skills and experience working with large operational datasets.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
Similar stack
Same company
In your city
Founding DevOps Engineer - Jordan
1 day ago
$25k – $42k per year • Equity 0–0.2% • Remote • Full-Time • 3+ years exp
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
Founding DevOps Engineer - California
1 day ago
$100k – $210k per year • Equity 0–0.5% • Remote • Full-Time • 3+ years exp • San Francisco
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
Security Researcher
1 day ago
$120k – $220k per year • Equity 0.2–0.8% • Remote/Hybrid • Full-Time • 6+ years exp • San Francisco
C++
Go
JavaScript
Python
TypeScript
Apply
Senior Golang Developer
2 days ago
≈ $17k – $47k per year (Estimated) • Remote/Hybrid • Internship • 5+ years exp • Minsk
Go
JavaScript
Python
AI/ML
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
Prompt Engineering
RAG
Anthropic
Function Calling
OpenAI
Frontend
React.js
DevOps
CI/CD
Docker
Git
Terraform
GitHub
Apply
Backend Engineer
2 days ago
≈ $61k – $206k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Tel Aviv
C++
Go
JavaScript
Node JS
Rust
TypeScript
DevOps
Akamai
AWS
Cloudflare
Fastly
Kubernetes
Nginx
Platform Engineering
Apply
Manager, CX/UX and Visual Design, Renters
4 days ago
$174k – $200k per year • In office • 5+ years exp • Bachelor's Degree • Jersey City
JavaScript
Design
Figma
Apply
Senior Data Analyst
4 days ago
Equity • Remote/Hybrid • 7+ years exp • Bachelor's Degree
PowerShell
Python
SQL
Databases
Azure SQL Database
MS SQL
PostgreSQL
Snowflake
AI/ML
dbt
DevOps
Azure
Git
Platform Engineering
SLI/SLO/SLA
Cybersecurity
GDPR
Analytics
Power BI
ETL/ELT
Apply
In office • 7+ years exp • Bachelor's Degree
JavaScript
SQL
TypeScript
DevOps
CI/CD
QA
Playwright
Selenium
Apply
Site Support Analyst
5 days ago
≈ $57k – $116k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Philadelphia
AI/ML
Model Context Protocol
Apply
System Support
5 days ago
Remote • 3+ years exp • Bachelor's Degree • Bangkok
SQL
Databases
MS SQL
DevOps
Windows Server
Apply
This is one of many
368,746 more open roles from verified company boards, updated every day.

