374,200open jobs
9,699companies
47,772added this week
Browse all
Salary
$73k – $146k per year (Estimated)
Location
Remote/Hybrid (Berlin, Germany)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
KNIME develops a widely used open source analytics platform where data science workflows are built visually as node graphs. Founded in 2008 in Zurich with roots at the University of Konstanz, it supports everything from data blending to deep learning. Its free desktop tool is paired with a commercial server for production deployment.

Mission

The Site Reliability Engineer at KNIME ensures that our next-generation cloud platform is built, operated, and scaled with reliability, security, and cost-efficiency as first-class engineering outcomes. Through active on-call ownership and hands-on engagement with engineering teams, this role sets and drives the operational standards that keep KNIME SaaS stable, observable, and production-ready at scale.

Role Overview

We are currently designing, building and launching our next generation of products and services in the cloud. We are looking for someone eager to be a part of this innovative process. This includes adapting our industry leading data science and analytics platform into a managed platform capable of serving thousands of users. The ability to handle Infrastructure as Code development is crucial as we strive to match our quality and functionality with innovative solutions that can address growth, cost and durability concerns. You will work in a cloud platform team interfacing with multiple development teams to drive KNIME products into production ready environments.

Responsibilities

  • Using code to automate the deployment and operations of large scale SaaS systems. Experience with Kubernetes operators is a plus.

  • Building out infrastructure as code using tools such as Helm, Terraform, Amazon CloudFormation and Azure ARM.

  • Participating in on-call rotations, incident triage and mitigation, troubleshooting issues in live environments, providing root cause analysis and issue resolution.

  • Setting standards for product deployments including reliability, scalability, traceability and monitoring. Communicate with product and development teams to help drive adoption.

  • Instrument deployed systems for performance, reliability and cost effectiveness.

  • Embeds with product and engineering to lead planning, own dependency risk, and drive consistent adoption of reliability and operational standards across teams.

Requirements

  • You hold one or more current certifications on a cloud platform such as AWS, Kubernetes, Linux or similar technologies

  • Strong cloud experience with at least one among AWS and Azure cloud providers. The ideal candidate would master both. You possess in-depth knowledge of VPC, IAM, EKS, ECR, EC2, S3, RDS, CloudWatch and their counterparts in the Azure environments.

  • Have experience deploying software systems to a Kubernetes environment. Have a working knowledge of Kubernetes concepts and the ability to craft deployment solutions using common Kubernetes patterns.

  • Scripting knowledge in Python, Shell are required, additional programming experience in Go, Java are a plus.

  • Systems level knowledge and experience with Linux. Expertise in networking, including security, routing, load balancers, and firewalls.

  • Knowledge of best practices around service telemetry, including metrics aggregation, distributed logging, and tracing in large, distributed systems

  • Working knowledge of OAuth/OIDC identity providers such as Keycloak

  • Working knowledge of relational databases such as Postgres

  • Ability to work independently and within a team environment. This includes clear and concise communication across an organization that is geographically and culturally dispersed

What Success Looks Like

  • Platform stability at scale: The multi-tenant SaaS platform maintains its availability SLO across all tenant tiers as the infrastructure grows from its current state toward a globally distributed commercial release.

  • Operational excellence embedded in engineering: Every team shipping to the platform follows a shared production readiness standard, reducing escaped defects and repeated incidents.

  • Automated, scalable operations: Tenant onboarding, deployments, and incident remediation are pipeline-driven, eliminating manual toil and enabling the team to scale without growing headcount at the same rate.

  • Commercial readiness: The platform meets enterprise security and compliance gates, supports consumption-based metering, and can sustain the onboarding velocity required for a commercial launch without operations becoming the bottleneck.

What we offer

  • Purpose-driven impact:The opportunity to be a driving force and cloud advocate as we design and build our next generation of service offerings

  • Craft & collaboration:Work alongside experienced engineers in a systems-first culture that values simplicity, maintainability, and clean design.

  • Learning:Continuous growth through hands-on challenges, peer exchange, and exposure to cutting-edge AI and data analytics topics.

  • Flexibility, health and wellbeing:Hybrid working, flexible hours, subsidised sports or yoga courses, physiotherapy, and flu shots at select locations.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
374,200 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Berlin
$100k – $195k per year (Estimated) • In office • Full-Time • United States
DevOps
Amazon CloudWatch
Amazon EC2
Amazon ECS
Amazon EKS
Amazon S3
AWS
AWS Lambda
FinOps
IAM
Incident Management
Kubernetes
Apply
$25k – $65k per year (Estimated) • In office • Full-Time • 10+ years exp • India
Java
SQL
Java
Hibernate
Spring Boot
Spring Framework
Databases
Apache Kafka
AI/ML
Airflow
DevOps
AWS
Azure
Azure DevOps
CI/CD
Docker
GCP
GitHub
GitHub Actions
Incident Management
Jenkins
Kubernetes
Rest API
Apply
$94k – $237k per year (Estimated) • In office • Full-Time • 3+ years exp • Singapore
Python
AI/ML
Computer Vision
Hallucination
Knowledge Distillation
LangChain
LlamaIndex
LLM
Multimodal AI
PyTorch
Quantization
RAG
TensorRT
TensorRT-LLM
Triton
Triton Inference Server
vLLM
LLM Evaluation
LLM Guardrails
AI Agents
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
Apply
$113k – $279k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Quebec
Java
Python
SQL
Databases
PostgreSQL
AI/ML
AI Agents
LLM
RAG
LLMOps
Red Teaming
DevOps
Azure
Azure AKS
CI/CD
Vector
Kubernetes
Apply
$40k – $114k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Mexico City
Python
Databases
Apache Kafka
AI/ML
AI Agents
LLM
OCR
DevOps
AWS
AWS Step Functions
CI/CD
KEDA
Kubernetes
Apply
$75k – $150k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Berlin
Go
Java
Java
Quarkus
DevOps
AWS
Azure
Helm
Kubernetes
Rest API
Cybersecurity
Keycloak
Apply
$79k – $176k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Berlin
Cybersecurity
ISO 27001
SBOM
SOC 2
Apply
$31k – $81k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Essen • Dresden • Stuttgart • Hamburg • Jena
Java
Apply
In office • Full-Time • Berlin
Apply
In office • Full-Time • Essen • Dresden • Stuttgart • Hamburg • Munich
Apply
$79k – $172k per year (Estimated) • Remote/Hybrid • Full-Time • Hamburg • Dresden • Stuttgart • Mannheim • Jena
AI/ML
LLMOps
DevOps
AWS
Azure
Bicep
CI/CD
CloudFormation
Docker
GCP
GitOps
Kubernetes
OpenShift
OpenTofu
Platform Engineering
Pulumi
Terraform
Apply
$76k – $179k per year (Estimated) • Remote/Hybrid • Full-Time • Hamburg • Dresden • Stuttgart • Mannheim • Jena
AI/ML
LLMOps
DevOps
AWS
Azure
Bicep
CloudFormation
GCP
Kubernetes
OpenShift
OpenTofu
Platform Engineering
Pulumi
Terraform
Apply
See all jobs
This is one of many
374,200 more open roles from verified company boards, updated every day.