1,443,753open jobs
85,355companies
221,329added this week
Browse all
Salary
$55k – $70k per year
Location
Remote (Poland)
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 11, 2026. First seen by Alion on Aug 18, 2026.

Overview
Company
Impact
Profile match

XTB

XTB is a Polish online brokerage headquartered in Warsaw that offers retail and institutional clients trading and investing in forex, CFDs, stocks, ETFs, indices and commodities through its own trading platform and mobile app. Founded in 2002 and listed on the Warsaw Stock Exchange since 2016, it serves more than 2.9 million customers worldwide, holds licences in the UK, Cyprus and other jurisdictions, and has offices across Europe, Latin America, the Middle East and Southeast Asia. It hires data, Java and site reliability engineers, product managers, AML and KYC specialists, risk analysts, multilingual customer support staff and business development and partnership managers.

XTB is a global company from the financial industry, focusing on online trading of financial instruments. We are the largest FinTech in Poland and a leader in Central and Eastern Europe, and the range of our operations covers several countries, including Asia and South America. At XTB, we focus on the development of our employees, giving them opportunities to gain knowledge and skills in various fields, as well as offering a number of training and development programs. If you are looking for challenges and want to gain valuable experience in an international business environment, XTB is the right place for you.

We are a certified Great Place to Work company.

We are looking for a Site Reliability Engineer to define and drive the reliability of XTB systems at the scale of millions of clients. In this role, you will strengthen SRE practices and shape the resilience of our entire technology stack through high-impact observability, ensuring our systems remain robust and scalable.

Responsibilities

  • Observability Platform Engineering: Develop a standardized observability ecosystem. Implement a conscious telemetry model focusing on structured events, distributed tracing, and intelligent sampling strategies - that provides deep, actionable insights into system behavior.
  • Reliability Enablement: Act as a strategic partner to product engineering teams, providing the platform, standards, and data they need to own service reliability. Use error budgets and alerting as the primary language for balancing feature velocity with stability.
  • Proactive Resilience & Protection: Enhance detection capabilities to identify issues before they impact the customer. Leverage early-warning systems and AI/ML for automated anomaly detection and intelligent data analysis to continuously verify and strengthen system resilience.
  • Operations & Tooling: Build internal automation and tooling that streamlines SRE workflows, automates routine operational tasks, and enhances efficiency across the technology stack.
  • Incident Management & On-Call Rotation: Participate in an on-call rotation to provide incident management, ensuring rapid incident resolution, effective communication, and post-incident analysis to drive continuous improvement.

Requirements

  • Professional Background: Professional experience in SRE, Infrastructure, or DevOps roles managing high-scale, distributed environments.
  • Technical Engineering: Advanced programming skills in Python, with a strong focus on building scalable automation, internal tooling, and robust scripts.
  • Cloud & Orchestration: Hands-on expertise in managing production-grade Kubernetes environments, configuration management tools like Ansible, and designing resilient infrastructure architectures within Azure Kubernetes Service and on-prem environments.
  • Observability Engineering: Proficiency in building standardized telemetry ecosystems. You have mastered self-hosted opensource tools for observability data collection, storage and visualization, like Prometheus, Grafana, ELK Stack, Tempo, Thanos, Jaeger and similar.
  • Operational & Soft Skills: Ability to drive incident management, conduct thorough post-incident analysis, and foster a culture of reliability and shared ownership.

Nice to have

  • Experience with commercial APM platforms (e.g., Datadog, Splunk, New Relic) and chaos engineering tooling.
  • Experience with cloud cost management and FinOps principles.
  • Experience defining and tracking SRE metrics (SLI/SLOs) and managing error budgets to drive reliability.
  • Experience with AI/ML techniques for SRE tasks, such as AIOps, automated anomaly detection, log analysis, and optimizing reliability workflows.
  • Experience in building and managing strategies to proactively manage technical debt and align team output with organizational goals.

What we offer

  • Real influence on the development of the company and the product.
  • Work in an experienced team that is happy to share its knowledge.
  • A clear vision of development thanks to regular feedback and clear career paths.
  • Regular team-building meetings.

Benefits

  • A training budget for courses and conferences that interest you.
  • An extra day off on your birthday.
  • An extra day off for parents.
  • Equipment tailored to your needs.
  • Private medical care and group insurance.
  • Access to an e-learning platform for learning English and a benefits platform.
  • Access to a wellbeing platform and the opportunity to take advantage of workshops and private therapy sessions.
  • Remote work, from the office in Warsaw or from a coworking space in your city.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,443,753 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Warsaw
Remote (Europe, United States)
C++
AI/ML
vLLM
CUDA Toolkit
Multimodal AI
SGLang
TensorRT
TGI
CUDA
Triton
ROCm
Apache TVM
CUTLASS
DevOps
Linux
Apply
≈ $113k – $209k per year (Estimated) • Equity • Remote (United States) • United States
AI/ML
NCCL
DevOps
Platform Engineering
Linux
Apply
Remote (Europe) • Amsterdam
Python
DevOps
CI/CD
eBPF
Linux
Apply
≈ $107k – $243k per year (Estimated) • Remote (Europe) • 6+ years exp
Python
Bash
DevOps
Terraform
GCP
Azure
AWS
Kubernetes
IAM
Linux
VPN
Cybersecurity
Threat Modeling
Apply
≈ $79k – $201k per year (Estimated) • Remote (Europe) • Amsterdam • Berlin • London • Prague
Python
AI/ML
vLLM
Multimodal AI
Ray
Triton
DevOps
Terraform
Prometheus
Kubernetes
Grafana
Self-Healing
Apply
DevOps инженер 19 days ago
≈ $17k – $39k per year (Estimated) • Hybrid • Moscow
Python
AI/ML
Machine Learning
DevOps
Ansible
OpenShift
Helm
Debian
Prometheus
GitLab CI
CI/CD
Docker
Kubernetes
Ubuntu
Grafana
CentOS Stream
Linux
Apply
≈ $13k – $37k per year (Estimated) • In office • Full-Time • Ho Chi Minh City
Python
Go
JavaScript
Java
Rust
Kotlin
TypeScript
SQL
Node JS
Scala
Python
FastAPI
Node JS
Nest.JS
Fastify
Databases
PostgreSQL
Snowflake
Databricks
ElasticSearch
Apache Kafka
Google BigQuery
Amazon Redshift
BigQuery
AI/ML
Copilot
Hadoop
Spark
Airflow
OpenCV
Dagster
Vertex AI
dbt
Embeddings
Scikit-learn
Kubeflow
TensorFlow
PyTorch
Anomaly Detection
TorchServe
Feature Store
Machine Learning
Frontend
GraphQL
Next.js
React.js
Storybook
DevOps
Rest API
gRPC
Terraform
GCP
Helm
GitHub Actions
Istio
Datadog
Kustomize
CI/CD
ArgoCD
AWS
Kubernetes
Cloudflare
Platform Engineering
Service Mesh
Argo Workflows
Google GKE
GitHub
Cybersecurity
Auth0
Analytics
ETL/ELT
Metabase
Looker
Design
Figma
Management
Slack
Miro
Confluence
Discord
Agile
Scrum
QA
Sentry
Apply
≈ $21k – $47k per year (Estimated) • In office • 2+ years exp • Moscow
Python
Python
FastAPI
Asyncio
Databases
PostgreSQL
AI/ML
NumPy
DevOps
Helm
GitLab CI
Docker
Kubernetes
QA
Pytest
Apply
≈ $18k – $37k per year (Estimated) • In office • Bachelor's Degree • Moscow
Python
1C
AI/ML
Claude
ChatGPT
YandexGPT
GigaChat
Management
n8n
Power Automate
Microsoft Office
Apply
$78k – $130k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Berlin
Python
Management
Agile
Scrum
Apply
≈ $62k – $104k per year (Estimated) • In office • Full-Time • Warsaw
Python
Bash
DevOps
Ansible
Red Hat
Configuration Management
Linux
Cybersecurity
PKI
Management
Confluence
Apply
≈ $55k – $100k per year (Estimated) • Hybrid • Full-Time • Warsaw
Python
PowerShell
Bash
DevOps
Terraform
Ansible
GCP
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Platform Engineering
Windows
Management
ITIL
Apply
≈ $48k – $107k per year (Estimated) • Hybrid • 3+ years exp • Warsaw
Analytics
Microsoft Excel
Apply
≈ $38k – $73k per year (Estimated) • In office • Full-Time • Warsaw
Java
Java
Maven
Spring Boot
Hibernate
Hazelcast
Databases
Oracle
RabbitMQ
DevOps
Rest API
Jenkins
Git
GitLab
SOAP
Cybersecurity
Fortify
Sonatype Nexus IQ
Management
Jira
Apply
≈ $48k – $85k per year (Estimated) • In office • Full-Time • 3+ years exp • Warsaw
SQL
Databases
Oracle
DevOps
Linux
Unix
Apply
See all jobs
This is one of many
1,443,753 more open roles from verified company boards, updated every day.