514,690open jobs
17,495companies
76,717added this week
Browse all
Salary
$21k – $52k per year (Estimated)
Location
In office (Shenzhen)
Seniority
Senior
Overview
Company
Impact
Profile match
kodypay is on a mission to make in-person payment acceptance easy. today, paying in person presents common problems for businesses, such as high costs, long queues, and limited choice of payment methods. kodypay fully integrates the payment ecosys...

Job Summary

Kody is seeking a Senior Site Reliability Engineer (8+ years of experience) to drive the reliability, availability, scalability, and operational excellence of our global payment platform. Based in Shenzhen, you will take end-to-end ownership of production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating across Europe, Asia, and North America.

Key Responsibilities

  • Incident Management & On-Call: Participate in a follow-the-sun production on-call rotation as a senior incident responder. Lead incident management during SEV1/SEV2 events to optimize MTTR and operational effectiveness.
  • Production Operations: Diagnose, triage, mitigate, and coordinate the resolution of complex production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure.
  • SLO & Reliability Engineering: Define, implement, and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes across distributed services.
  • Continuous Optimization: Drive systemic reliability improvements through infrastructure automation, observability enhancement, capacity planning, performance tuning, and post-incident root-cause analysis (RCA).
  • Security & Compliance: Partner with global engineering teams to strengthen architectural resilience, security posture, and operational maturity in PCI-DSS-regulated payment environments.
  • Technical Leadership: Mentor junior engineers, eliminate operational toil through automation, and influence engineering teams to adopt resilience-by-design practices.

Requirements

Qualifications & Requirements

  • Experience: 8+ years of hands-on experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting high-availability, mission-critical production systems.
  • Core Technical Stack: Strong expertise in AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms (e.g., Datadog, Prometheus, Grafana).
  • Distributed Systems Mastery: Deep understanding of distributed systems architecture, high availability, disaster recovery, capacity planning, and microservices orchestration.
  • Domain Expertise: Proven track record operating in payment, banking, fintech, or other highly regulated environments with strict PCI-DSS, security, and uptime standards.
  • SRE Methodology: Deep knowledge of core SRE principles, including SLO/SLI design, error budget management, alert governance, and toil reduction.
  • Location & Communication: Based in Hong Kong or Shenzhen. Excellent command of English (written and spoken) to lead cross-functional incident responses and collaborate seamlessly with global teams.

Leadership & Operational Excellence

  • Ownership: Demonstrates strong end-to-end accountability for service reliability and customer impact under high pressure.
  • Structured Problem Solving: Applies a systematic and data-driven approach to troubleshooting, telemetry analysis, and incident resolution in complex distributed environments.
  • Crisis Management: Proven ability to command cross-functional incident response efforts, align stakeholders, and maintain clear communication during critical outages.
  • Engineering Culture: Champions a blameless post-incident culture, operational readiness, continuous learning, and technical mentorship.

Benefits

- Competitive package

- A dynamic and innovative team

- Collaborative, inclusive working environment where your contributions are recognized

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
514,690 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Shenzhen
In office
DevOps
Incident Management
Management
Jira
Agile
Apply
$27k – $55k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru
SQL
Databases
Databricks
Delta Lake
Amazon Aurora
DevOps
AWS
Docker
Kubernetes
Amazon EKS
Amazon ECS
AWS Step Functions
Analytics
ETL/ELT
Apply
$149k – $224k per year • Remote/Hybrid • Full-Time • 6+ years exp • PhD • San Francisco
Python
Java
SQL
Databases
ElasticSearch
Apache Kafka
Trino
AI/ML
AI Agents
Agentforce
DevOps
Terraform
GCP
Azure
AWS
Docker
Kubernetes
Amazon S3
Apply
$151k – $312k per year (Estimated) • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Columbus • Tampa • Dallas • Atlanta
Python
Java
AI/ML
Claude
Vertex AI
AI Agents
RAG
OpenAI
Anthropic
Context Engineering
LLM Evaluation
Agentic Workflows
DevOps
Terraform
Helm
CI/CD
Docker
Kubernetes
Apply
$56k – $164k per year (Estimated) • Remote/Hybrid • Internship • Singapore
JavaScript
Java
TypeScript
C#
Java
Spring Boot
C#
.NET
Databases
Apache Kafka
Frontend
Angular
DevOps
GitLab CI
CI/CD
Kubernetes
Apply
$37k – $91k per year (Estimated) • In office • Hong Kong
Databases
PostgreSQL
Redis
Apache Kafka
DevOps
Terraform
Datadog
Prometheus
AWS
Kubernetes
Grafana
Platform Engineering
Amazon EKS
Incident Management
Error Budget
SLI/SLO/SLA
Cybersecurity
PCI DSS
Apply
$46k – $116k per year (Estimated) • In office • 3+ years exp • London
Apply
$56k – $116k per year (Estimated) • In office • Hong Kong
Apply
HR Specialist- UK 21 day ago
$46k – $116k per year (Estimated) • In office • 3+ years exp • London
Apply
$107k – $257k per year (Estimated) • Equity • Remote • 5+ years exp • Singapore
Apply
Sales Manager 1 day ago
$34k – $87k per year (Estimated) • In office • 5+ years exp • Shenzhen
Apply
Account Executive 1 day ago
$18k – $48k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Shenzhen
Apply
$16k – $44k per year (Estimated) • Remote • Full-Time • 4+ years exp • High School Diploma • Shenzhen
Apply
Packaging Engineer 2 days ago
$14k – $31k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Shenzhen
Design
Adobe Photoshop
SolidWorks
Apply
Product Designer 2 days ago
$14k – $39k per year (Estimated) • In office • Full-Time • 3+ years exp • Master's Degree • Shenzhen
Apply
See all jobs
This is one of many
514,690 more open roles from verified company boards, updated every day.