Salary
≈ $135k – $262k per year (Estimated)
Location
In office (San Francisco)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
kodypay is on a mission to make in-person payment acceptance easy. today, paying in person presents common problems for businesses, such as high costs, long queues, and limited choice of payment methods. kodypay fully integrates the payment ecosys...
Senior Site Reliability Engineer (Payments Infrastructure)
Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America.
Responsibilities
- Participate in a follow-the-sun production on-call rotation as a primary incident responder.
- Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure.
- Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes.
- Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews.
- Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments.
- Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR.
Requirements
- 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems.
- Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms.
- Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization.
- Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements.
- Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence.
Leadership & Operational Excellence
- Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact.
- Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure.
- Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments.
- Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements.
- Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil.
- Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events.
- Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability.
- Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design.
Benefits
- Competitive packages aligned with California market standards
- Lead a dynamic and innovative team in a very rapidly growing company
- Collaborative, inclusive environment where your contributions are recognized and valued
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
Similar stack
Same company
San Francisco
Engineering Manger - Dev Platform
8 hours ago
≈ $73k – $183k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Zug
Python
Python
FastAPI
Databases
PostgreSQL
Redis
AI/ML
AI Agents
AWS Bedrock
AWS Bedrock AgentCore
LLM
DevOps
Amazon EKS
AWS
AWS CDK
CI/CD
Datadog
Kubernetes
OpenTelemetry
Platform Engineering
Apply
$136k – $253k per year • Equity • Remote/Hybrid • Full-Time • 10+ years exp • Frisco • New York • Toronto • Ann Arbor
Python
SQL
Java
Java
Flyway
AI/ML
AWS Bedrock
Claude
LLM
Anthropic
AWS Bedrock AgentCore
LLM Guardrails
LLMOps
DevOps
Amazon EKS
AWS
CI/CD
Datadog
Docker
Kubernetes
Apply
Software Engineer (SQL & Java)
8 hours ago
$32k – $42k per year • Remote/Hybrid • Full-Time • 4+ years exp • Cyprus
Java
SQL
Java
Spring Boot
Databases
Apache Kafka
Oracle
DevOps
AWS
Docker
Jenkins
Kubernetes
Rest API
Apply
≈ $24k – $63k per year (Estimated) • In office • Full-Time • 6+ years exp • Master's Degree • India
Crystal
Groovy
JavaScript
Perl
Python
Ruby
SQL
TypeScript
Java
Java
Apache Tomcat
Gradle
Hibernate
Maven
Spring Boot
Spring MVC
Databases
Apache Kafka
Db2
Oracle
PostgreSQL
RabbitMQ
AI/ML
Fine-tuning
Frontend
Angular
JQuery
DevOps
Apache HTTP Server
AWS
Azure
CI/CD
Docker
GCP
Jenkins
Kubernetes
Rest API
Cybersecurity
Checkmarx
SonarQube
Apply
SO MLOps Integration Engineer, GMET
8 hours ago
≈ $64k – $189k per year (Estimated) • Remote/Hybrid • Full-Time • 1+ year exp • Bachelor's Degree • Singapore
Python
SQL
Databases
Apache Kafka
AI/ML
Amazon SageMaker
Kubeflow
MLFlow
Spark
Vertex AI
DevOps
AWS
Azure
Azure DevOps
CI/CD
Docker
GCP
GitLab
GitLab CI
Jenkins
Kubernetes
Apply
Senior Product Manager- Payment
12 days ago
≈ $106k – $218k per year (Estimated) • Remote • Singapore
Apply
Senior Product Manager- payment
12 days ago
≈ $111k – $175k per year (Estimated) • In office • London
Apply
Apply
Apply
Apply
⚡️ Founding Full Stack Software Engineer ⚡️
2 hours ago
$170k – $220k per year • Equity 1–2.8% • In office • Full-Time • 3+ years exp • San Francisco
Python
SQL
Python
Django
AI/ML
AI Agents
Context Engineering
LLM
LLM Evaluation
RAG
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Senior ML Engineer
3 hours ago
$149k – $224k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Francisco • Washington • Palo Alto
Python
Python
pySpark
Databases
Apache Kafka
AI/ML
AI Agents
Agentforce
Airflow
Anomaly Detection
Feature Store
Flink
Ray
Red Teaming
Spark
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
MITRE ATT&CK
Marketing
Salesforce
Apply
Software Engineer, Intern (Summer or Winter)
3 hours ago
In office • Internship • 1+ year exp • Bachelor's Degree • San Francisco
Go
JavaScript
Ruby
Scala
Apply
Data Center Infrastructure Architect
3 hours ago
$360k – $530k per year • In office • Full-Time • Bachelor's Degree • San Francisco
MATLAB
Python
MATLAB
Simulink
AI/ML
OpenAI
Robotics
Digital Twin
Apply
This is one of many
368,746 more open roles from verified company boards, updated every day.

