368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$25k – $64k per year (Estimated)
Location
Remote/Hybrid (Bengaluru, India)
Seniority
Senior · 10+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Nium's Purpose: We exist to solve the payment challenges that unlock the full potential of the global economy. Nium's Mission: We want to build THE global payments infrastructure for on-demand money movement. Nium's Vision: Our payments in...

The Role

Nium is looking for a Senior Manager, Site Reliability Engineering to lead the teams responsible for the availability, performance, scalability, and operational excellence of our global payments platform. This is a hands-on leadership role: you will build and grow a team of SREs, define the reliability roadmap, and partner closely with product engineering, security, and infrastructure teams to ensure Nium's systems meet the always-on expectations of a regulated financial platform operating across 100+ markets.

You will own incident management, observability, capacity planning, and production readiness practices, while championing a culture of blameless postmortems, proactive risk reduction, and engineering-driven automation. This role sits at the intersection of engineering leadership and operational rigor, and reports into Nium's engineering leadership.

Responsibilities

    • Lead, mentor, and grow a team of SREs and reliability engineers across multiple time zones, setting clear goals, career paths, and performance expectations.

    • Own Nium's reliability strategy - defining and driving SLIs/SLOs/error budgets across critical payment, card issuance, and compliance services.

    • Drive incident management end-to-end: on-call structure, escalation paths, major incident response, and blameless postmortems that produce durable fixes, not just tickets.

    • Partner with product engineering leaders to embed reliability, scalability, and operational readiness into the software development lifecycle from design through launch.

    • Build and scale observability (metrics, logging, tracing, alerting) so that issues are detected and diagnosed before they impact customers or partner banks.

    • Lead capacity planning and performance engineering for systems processing high-volume, real-time financial transactions across a global, multi-region infrastructure.

    • Champion automation and self-healing systems to reduce toil, eliminate manual runbooks, and improve mean-time-to-detect and mean-time-to-resolve.

    • Own disaster recovery, business continuity, and chaos engineering practices, running regular game days to validate resilience assumptions.

    • Collaborate with Security and Compliance teams to ensure infrastructure practices meet regulatory and audit requirements (PCI-DSS, SOC 2, ISO 27001, and regional financial regulations).

    • Manage the reliability budget: tooling investments, cloud cost/performance trade-offs, and staffing plans, in partnership with finance and engineering leadership.

    • Represent SRE in executive reviews, translating technical risk and system health into business-relevant reporting for leadership and the board.

    • Establish and continuously refine production readiness reviews, runbooks, and operational standards across all engineering teams.

Requirements

    • 10+ years of experience in software engineering, infrastructure, or site reliability engineering, with 4+ years in a people-management or technical leadership role leading SRE/DevOps/Infrastructure teams.

    • Proven track record operating and scaling production systems for a high-availability, transaction-heavy platform - fintech, payments, banking, or e-commerce experience strongly preferred.

    • Deep hands-on expertise with cloud infrastructure (AWS) Kubernetes, container orchestration, and infrastructure-as-code (Terraform, CloudFormation, or similar).

    • Strong background in observability stacks (Prometheus, Grafana, Datadog, ELK/OpenSearch, or equivalent) and building alerting that reduces noise while catching real issues.

    • Demonstrated experience defining and operationalizing SLOs/error budgets, and using them to drive engineering prioritization.

    • Solid understanding of distributed systems, databases, caching, messaging queues, and API-driven microservice architectures at scale.

    • Experience leading major incident response for critical, customer-facing systems, including postmortem processes that drive real change.

    • Familiarity with security and compliance frameworks relevant to financial services (PCI-DSS, SOC 2, ISO 27001) and how they shape infrastructure and access practices.

    • Excellent communication skills - able to translate technical reliability concepts for engineering peers, product leaders, and executives alike.

    • A pragmatic, metrics-driven approach to engineering decisions: you measure before you optimize, and you test assumptions in production rather than in theory.

    • Experience with scripting/automation (Python, Go, or Bash) and CI/CD pipelines (Jenkins, ArgoCD, GitLab CI, or similar).

    Nice to Have

    • Experience operating systems subject to real-time payment rails, card networks, or banking core integrations.

    • Prior experience running SRE or infrastructure functions through hypergrowth or rapid international expansion.

    • Exposure to FinOps practices and cloud cost optimization at scale.

    • Experience building or scaling an SRE function from the ground up within an existing engineering organization.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
$129k – $232k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Austin
C++
DevOps
CI/CD
Git
RTOS
Apply
$133k – $161k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Westminster
C++
Python
DevOps
CI/CD
SpaceTech
NASA cFS
Apply
$133k – $161k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Westminster
C++
C
C
U-Boot
DevOps
CI/CD
Apply
$133k – $161k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Westminster
C++
Python
DevOps
CI/CD
SpaceTech
NASA cFS
Apply
In office • Part-Time • 2+ years exp • Bachelor's Degree • Ness Ziona
C#
Java
AI/ML
Copilot
Cursor
DevOps
CI/CD
GitHub
Apply
$33k – $67k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bengaluru
SQL
AI/ML
ChatGPT
Claude
Copilot
Apply
Staff Product Manager 11 days ago
$106k – $199k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • London
Apply
$22k – $54k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bengaluru
Python
SQL
Databases
Amazon Redshift
Apache Kafka
Snowflake
AI/ML
Airflow
Prefect
DevOps
AWS
Rest API
SLI/SLO/SLA
Amazon S3
Analytics
ETL/ELT
Apply
Data Engineer II 14 days ago
$20k – $49k per year (Estimated) • In office • Full-Time • 2+ years exp • Mumbai
Python
SQL
Databases
Amazon Redshift
Apache Kafka
MySQL
AI/ML
dbt
DevOps
AWS
AWS Lambda
CI/CD
Amazon S3
Apply
Sr Data Scientist 26 days ago
$30k – $59k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Bengaluru
Apply
$16k – $34k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Mumbai • Bengaluru
JavaScript
PowerShell
SQL
C#
C#
.NET
Databases
Azure SQL Database
MS SQL
DevOps
Azure
Rest API
Cybersecurity
Microsoft Entra ID
QA
Postman
Swagger
Apply
$41k – $89k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Bengaluru
C#
TypeScript
JavaScript
C#
.NET
Databases
Apache Kafka
AI/ML
Copilot
LLM
OpenAI
Frontend
Angular
GraphQL
DevOps
Azure
Azure AKS
Azure DevOps
CI/CD
Docker
GitHub
GitHub Actions
Grafana
Kubernetes
Prometheus
Rest API
Apply
$38k – $83k per year (Estimated) • In office • Full-Time • 12+ years exp • Bachelor's Degree • Bengaluru
Databases
Oracle
DevOps
AWS
Platform Engineering
Apply
Data Architect 2 hours ago
$38k – $91k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru • Pune
Node JS
Python
SQL
JavaScript
Databases
Databricks
MongoDB
Redis
Apply
$28k – $71k per year (Estimated) • In office • Full-Time • 5+ years exp • Bengaluru
DevOps
CI/CD
Platform Engineering
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.