368,634open jobs
9,437companies
50,578added this week
Browse all
Location
Remote/Hybrid (Baku, Azerbaijan)
Seniority
Middle · 3+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Xsolla is a global video game commerce company that provides specialized financial and operational tools tailored for the gaming industry. The firm helps game developers and publishers fund, launch, market, and monetize their titles across PC, mobile, web, and cloud platforms. By operating as a merchant of record with support for over 1,000 local payment methods, it enables direct-to-consumer sales and seamless cross-border transactions for gaming studios worldwide.

ABOUT YOU

We are looking for a Site Reliability Engineer who is pragmatic, product-minded, and equally comfortable writing code and running production systems to join our Infrastructure department's SRE team. The best candidate will be someone who thrives in a fast-paced, highly collaborative, and exceptionally dynamic setting and is excited to own the application-level infrastructure and reliability of a high-traffic commerce domain end to end - from deploy pipelines and Kubernetes manifests to SLOs, capacity planning, and production readiness.

Strong Kubernetes, observability, and software engineering skills are essential, along with experience in operating production services in a cloud environment (GCP/GKE or comparable) and partnering closely with product development teams. The ability to hold a dual perspective - understanding both how developers ship features and what infrastructure needs to stay reliable - and to bring the reliability lens into design decisions early will be key to your success in this role.

If you're passionate about making complex distributed systems boringly reliable and love building the commerce and monetization backbone that lets game developers around the world get paid, we would love to hear from you!

ABOUT US

Xsolla is a global commerce company with robust tools and services to help developers solve the inherent challenges of the video game industry. From indie to AAA, companies partner with Xsolla to help them fund, distribute, market, and monetize their games. Grounded in the belief in the future of video games, Xsolla is resolute in the mission to bring opportunities together, and continually make new resources available to creators. Headquartered and incorporated in Los Angeles, California, Xsolla operates as the merchant of record and has helped over 1,500+ game developers to reach more players and grow their businesses around the world. With more paths to profits and ways to win, developers have all the things needed to enjoy the game.

For more information, visit xsolla.com.

Responsibilities

  • Own the application-level infrastructure: Helm charts, Terraform configurations, Kubernetes deployments, runtime configuration, and service-level networking and integrations
  • Own the domain's observability: design and implement SLOs/SLIs, monitors, alerts, and dashboards for critical services on Datadog and OpenTelemetry-based tooling
  • Help to set up and evolve CI/CD pipelines for domain services (GitLab CI, GitHub Actions), including deploy and rollback automation
  • Perform capacity planning and performance tuning ahead of expected load - product launches, sales events, and regional rollouts - including load testing and performance regression investigation
  • Run Production Readiness Reviews for new services and major changes; define and enforce what "production-ready" means for the domain
  • Support domain incident response: assist with deep investigation of complex incidents, contribute to post-mortems, drive follow-up reliability improvements, and maintain runbooks
  • Build domain-specific automation that reduces operational toil: runbook automation, deploy helpers, recurring operational scripts
  • Maintain and drive a forward-looking reliability roadmap for the domain together with product engineering leads
  • Participate in product team planning, refinements, and architecture reviews, bringing the reliability perspective before design decisions become expensive to change
  • Co-author company-wide SLO/SLI, capacity, and operational standards together with the broader SRE team; contribute improvements directly to shared SRE-operated subsystems
  • Participate in the SRE duty rotation, supporting developers across the company

Qualifications & Skills

    • 3+ years of proven SRE, DevOps, or platform engineering experience: on-call or incident response duty, SLO/monitoring ownership, deploy pipeline and infrastructure work for production services
    • Software development background: you have built and shipped backend services, not only operated them - comfortable reading application code during an investigation and writing production-quality automation in at least one language (e.g., Go, PHP)
    • Hands-on Kubernetes experience: Helm, manifests, deploy strategies, debugging application-level performance and networking issues (GKE or another managed Kubernetes)
    • Solid observability practice: building monitors, dashboards, and SLOs/SLIs on a modern platform (Datadog preferred; Prometheus/Grafana experience also relevant), familiarity with OpenTelemetry
    • Infrastructure as Code exposure (Terraform/Terragrunt) for collaboration with platform teams
    • GCP experience (IAM, networking, managed services)
    • Experience building and maintaining CI/CD pipelines (GitLab CI and/or GitHub Actions)
    • Programming/scripting proficiency sufficient to build automation and tooling (e.g., Python, Go, or Bash)
    • Practical experience with incident response, post-mortems, and driving reliability improvements from incidents
    • Strong collaboration and communication skills - this role works embedded with product development teams daily
    • Experience in payments, fintech, e-commerce, or gaming - high-traffic transactional systems
    • Nice to Have:

    • Kubernetes certifications
    • Google Cloud Platform certifications
    • HashiCorp certifications
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Baku
DevOps Engineer 4 hours ago
$49k – $88k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Warsaw
Bash
Python
DevOps
AWS
Azure
Azure AKS
CI/CD
Docker
GCP
Git
GitHub
GitHub Actions
GitOps
Grafana
Kubernetes
Loki
Prometheus
Terraform
Thanos
Apply
$91k – $182k per year (Estimated) • In office • 3+ years exp • Redwood City
Python
DevOps
AWS
CI/CD
GCP
Terraform
Apply
$171k – $273k per year • In office • Full-Time • 8+ years exp • PhD • San Francisco • Washington
AI/ML
A2A
Agentforce
AI Agents
Model Context Protocol
DevOps
AWS
GCP
Marketing
Salesforce
Apply
IT Administrator 3 hours ago
$83k – $113k per year • In office • Full-Time
Node JS
Python
JavaScript
DevOps
AWS
GCP
GitHub
Cybersecurity
Okta
SOC 2
Management
Confluence
Jira
Slack
Apply
$65k – $174k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Erlangen
SystemVerilog
VHDL
DevOps
CI/CD
Chips/EDA
Xilinx Vivado
Apply
$14k – $26k per year (Estimated) • In office • Contractor • 4+ years exp • Bachelor's Degree • Vladivostok
Node JS
TypeScript
JavaScript
Frontend
esbuild
GraphQL
React.js
styled-components
Webpack
Zod
DevOps
CI/CD
GitLab CI
GitLab
QA
Cypress
Jest
Playwright
Vitest
Apply
$15k – $37k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Perm
Bash
Python
SQL
Databases
MySQL
PostgreSQL
DevOps
AWS
Datadog
GCP
Grafana
Prometheus
Puppet
Terraform
Zabbix
Apply
$14k – $35k per year (Estimated) • In office • Perm
Go
PHP
SQL
DevOps
CI/CD
Apply
$12k – $27k per year (Estimated) • In office • 3+ years exp • Perm
Go
PHP
Python
DevOps
CI/CD
Datadog
GCP
GitHub Actions
GitLab CI
Google GKE
Grafana
Helm
Kubernetes
OpenTelemetry
Prometheus
SLI/SLO/SLA
Terraform
Terragrunt
GitHub
GitLab
IAM
Apply
IT Support Engineer 4 days ago
$31k – $67k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Berlin
Cybersecurity
Okta
Management
Google Workspace
Apply
Remote • 2+ years exp • Baku
Go
SQL
TypeScript
JavaScript
Frontend
React.js
Apply
$24k – $36k per year (gross) • Remote • Full-Time • Baku
TypeScript
Databases
PostgreSQL
Supabase
Apply
Remote • 7+ years exp • Baku
DevOps
AWS
AWS Fargate
CI/CD
CircleCI
CloudFormation
Datadog
GitHub Actions
Grafana
Kubernetes
New Relic
OpenTelemetry
Platform Engineering
Terraform
Amazon ECS
GitHub
IAM
Cybersecurity
Snyk
SonarQube
Apply
Remote • Internship • Bachelor's Degree • Baku
PHP
Python
SQL
Databases
Apache Kafka
RabbitMQ
DevOps
CI/CD
Git
Helm
Jenkins
Kubernetes
Terraform
WebSockets
Web3
DeFi
Smart Contracts
Apply
Tech Lead 5 days ago
In office • Full-Time • 5+ years exp • Baku
Go
JavaScript
PHP
Frontend
Next.js
React.js
DevOps
CI/CD
Git
Incident Management
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.