845,403open jobs
53,832companies
143,151added this week
Browse all
Salary
≈ $131k – $282k per year (Estimated)
Location
In office (Noida)
Seniority
Architect · 15+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 27, 2026. First seen by Alion on Sep 1, 2026. Netspend scores B on the Alion truth index.

Overview
Company
Impact
Profile match

About the Company:

Netspend Corporation is a global, vertically-integrated financial services and technology company dedicated to the delivery of innovative financial empowerment solutions to consumers worldwide. Netspend's financial products and services span prepaid, debit, cross-border payments, and loyalty solutions for consumers and enterprise partners.

Netspend provides prepaid and debit account solutions that connect customers with secure, convenient access to global payment networks so they can manage their money and make everyday purchases. With a nationwide U.S. retail network, customers can purchase and reload Netspend products at 130,000 reload points and over 100,000 distributing locations.

Since our founding in 1999 by industry pioneers, Netspend products have processed billions of dollars in transaction volume and served millions of customers worldwide. The company is headquartered in Austin, Texas with employees worldwide.

Role Overview

We are seeking a visionary Director - Observability, Response & Reliability (ORR) with 15+ years of overall technical experience, including 5+ years in engineering leadership, to serve as our primary authority on system resilience, full-stack observability, and enterprise incident management across Netspend's global financial ecosystem.

In this high-visibility role, you will lead our India-based ORR organization, driving the strategy that transforms how Netspend monitors, predicts, responds to, and resolves critical operational events. You will establish a world-class Observability and AIOps practice, institutionalize Site Reliability Engineering (SRE) principles across all product teams, and safeguard systems handling ACH processing, core payment rails, millions of daily card transactions, and bank-sensitive data.

Key Responsibilities

1. Observability Strategy & Telemetry Architecture

  • Unified Telemetry Vision: Design and execute the enterprise Observability roadmap, establishing full-stack, end-to-end visibility across microservices, legacy monoliths, cloud infrastructure, and data pipelines (Kafka, Cassandra).

  • MELT Standardizer: Define unified logging, metrics, traces, and synthetics standards (OpenTelemetry, Prometheus, Splunk, Dynatrace) across all engineering groups to enable granular transaction-level tracing for payment flows.

  • AIOps & Self-Healing Systems: Pioneer the adoption of AIOps, Machine Learning, and anomaly detection to shift the engineering culture from reactive alert firefighting to proactive noise reduction, predictive fault detection, and automated self-healing workflows.

  • Financial Data Observability: Build custom monitoring, alert thresholds, and real-time dashboards for critical FinTech protocols, payment rails, ACH file watching, and bank connectivity channels (Axway, SFTP, AS2).

2. Reliability Engineering, Resilience & SLO Management

  • SRE Practice Ownership: Institutionalize SRE paradigms (SLIs, SLOs, Error Budgets, Reliability Reviews) across all software product squads, embedding reliability directly into the software development lifecycle (SDLC).

  • Chaos Engineering & Testing: Establish proactive resilience practices, including failure injection, Chaos Engineering, and regular multi-AZ/multi-region Disaster Recovery (DR) simulations for critical financial services.

  • Cost-Effective Reliability (FinOps): Partner with cloud and finance leadership to balance extreme uptime demands with cost efficiency across Splunk, Dynatrace, AWS telemetry storage, and log ingestion limits.

3. Enterprise Incident Response & Operational Governance

  • Major Incident Command (P0/P1): Oversee the high-severity Incident Management framework, ensuring 24/7 incident readiness, rapid mean time to detect (MTTD), and swift mean time to resolve/recover (MTTR).

  • Blameless Post-Mortems & RCA: Champion a culture of psychological safety through rigorous, blameless post-mortems and Root Cause Analyses (RCAs) to drive systemic platform fixes and prevent recurring outages.

  • Executive Reliability Reporting: Establish transparent reporting frameworks to translate uptime metrics, availability SLAs, error budget consumption, and platform risks directly to C-suite leadership (CTO, CIO).

4. Strategic Cloud Transformation & Infrastructure Evolution

  • Strangler Pattern Execution: Provide reliability oversight for moving critical functions off legacy platforms (Xymon, Puppet, SVN) to cloud-native AWS architectures without risking live financial traffic or data loss.

  • CI/CD & IaC Standards: Enforce strict Infrastructure-as-Code (Terraform/Ansible) and GitLab CI/CD pipeline reliability, integrating automated canary deployments, security scans, and instant rollback mechanisms.

  • Security & Compliance Oversight: Partner with security teams to ensure all observability logging, tracing, and automation frameworks comply strictly with PCI-DSS, SOC 2, mTLS, and banking audit standards.

5. Organizational Leadership & Vendor Ownership

  • India ORR Center of Excellence: Build, mentor, and scale a high-performing team of Site Reliability Engineers, Observability Specialists, and Incident Response Leads in India.

  • Talent Development: Define technical career paths, performance benchmarks, and continuous learning opportunities in observability and SRE for mid-level and senior engineers.

  • Vendor Governance: Manage strategic relationships, licensing, contractual negotiations, and technical roadmaps with key observability and cloud partners (Splunk, Dynatrace, AWS, GitLab).

Required Qualifications & Experience

Education & Overall Experience

  • Bachelor’s or Master’s Degree in Computer Science, Software Engineering, Information Technology, or a related quantitative field.

  • 15+ years of total experience in SRE, Systems/Platform Engineering, Infrastructure Architecture, and Enterprise Observability.

Core Observability & Reliability Expertise

  • Observability Mastery: Deep expertise architecting enterprise telemetry solutions using tools such as Splunk, Dynatrace, Datadog, Prometheus, Grafana, and OpenTelemetry.

  • AIOps & Automation: Proven track record implementing AIOps tools, automated incident remediation, and AI/ML-based anomaly detection engines.

  • SRE Leadership: Expert knowledge of SRE best practices, error budget policy enforcement, SLO modeling, and incident response frameworks.

Technical & Systems Foundation

  • Cloud & Modern Infrastructure: Strong hands-on architectural understanding of AWS services (EC2, ALB/NLB, Direct Connect, IAM, KMS, Transfer Family).

  • Legacy-to-Cloud Modernization: Practical experience executing "strangler fig" migration strategies off legacy tooling (Puppet, SVN, Xymon) to modern GitOps/IaC (GitLab CI/CD, Terraform, Ansible).

  • Data & Middleware: Experience monitoring and troubleshooting distributed middleware and database technologies (Apache Kafka, Cassandra) under heavy throughput.

  • FinTech & Secure Protocols: Understanding of secure file transfer (SFTP, AS2), mTLS, cryptographic key management (HSMs, Virtucrypt), and high-availability payment processing environments.

Preferred Certifications

  • Observability: Splunk Certified Architect, Dynatrace Master / Professional, or equivalent.

  • Cloud & DevOps: AWS Certified Solutions Architect - Professional

  • SRE & Methodologies: Certified Site Reliability Engineer (SRE), ITIL v4 (Incident/Problem Management focus).

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
845,403 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Noida
≈ $142k – $307k per year (Estimated) • Hybrid • 5+ years exp • High School Diploma • Oakland
DevOps
GCP
Azure
CI/CD
GitOps
AWS
Platform Engineering
Shift-Left
Cybersecurity
Shift-Left Security
Management
Linear
Notion
ITSM
Apply
≈ $140k – $303k per year (Estimated) • In office • Bachelor's Degree • New York
DevOps
Kubernetes
Akamai
FinOps
Cybersecurity
Zero Trust
Apply
$200k – $270k per year • In office • 12+ years exp • New York
Databases
Microsoft Fabric
AI/ML
Copilot
AI Agents
Semantic Kernel
OpenAI
Copilot Studio
DevOps
Azure
Management
Agile
Apply
≈ $119k – $256k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Pittsburgh
Python
Java
Databases
Oracle
AI/ML
Edge AI
Mobile
JUnit
DevOps
Dynatrace
CI/CD
Git
Docker
Kubernetes
Management
Agile
Apply
≈ $111k – $240k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Pittsburgh
PowerShell
DevOps
Windows
Cybersecurity
Active Directory
Apply
≈ $63k – $150k per year (Estimated) • In office • Berlin
DevOps
Terraform
Ansible
Red Hat
VMWare
Linux
Management
Confluence
Jira
Apply
$98k – $182k per year • In office • 8+ years exp • Bachelor's Degree • Chicago
Python
JavaScript
TypeScript
SQL
Node JS
Python
pySpark
Node JS
Nest.JS
Express
Fastify
Databases
PostgreSQL
Redis
Snowflake
Databricks
DynamoDB
RabbitMQ
Apache Kafka
Amazon Aurora
AI/ML
Cursor
Spark
Claude Code
AI Agents
RAG
Frontend
Next.js
React.js
DevOps
Rest API
Terraform
Ansible
GCP
Azure
CI/CD
AWS
Apply
≈ $26k – $58k per year (Estimated) • In office • 10+ years exp • Bengaluru
Python
JavaScript
Frontend
npm
DevOps
Terraform
Ansible
GCP
Azure
AWS
Configuration Management
DNS
DHCP
BGP
OSPF
Cybersecurity
ISO 27001
PCI DSS
Apply
In office • 3+ years exp • Bengaluru
Python
DevOps
Terraform
Ansible
Azure
Windows
Cybersecurity
Okta
Auth0
Active Directory
Apply
DevOps-инженер 2 hours ago
≈ $16k – $36k per year (Estimated) • Remote (likely EAEU) • Moscow
JavaScript
Kotlin
Bash
Databases
PostgreSQL
ClickHouse
Apache Kafka
DevOps
Terraform
GCP
Helm
FluxCD
Prometheus
GitLab CI
Azure
CI/CD
GitOps
Git
AWS
Docker
Kubernetes
Grafana
Proxmox VE
Linux
Apply
≈ $136k – $265k per year (Estimated) • Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Noida
Java
SQL
Java
Spring Framework
Databases
Oracle
ActiveMQ
Apache Kafka
DevOps
Splunk
Dynatrace
Unix
Apply
≈ $85k – $170k per year (Estimated) • In office • Contractor • 5+ years exp • Bachelor's Degree • Austin
Marketing
LinkedIn
Apply
≈ $41k – $96k per year (Estimated) • In office • Full-Time • Austin
Apply
≈ $152k – $285k per year (Estimated) • In office • Full-Time • 12+ years exp • Noida
Java
SQL
Java
Spring Framework
Databases
Oracle
ActiveMQ
Apache Kafka
DevOps
Splunk
Dynatrace
Unix
Apply
≈ $28k – $60k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Noida
Java
SQL
Java
Spring Boot
Databases
PostgreSQL
Oracle
Apache Kafka
DevOps
AWS
Bitbucket
Apply
In office • 15+ years exp • Noida
Python
Verilog
SystemVerilog
VHDL
Chips/EDA
UVM
Apply
≈ $28k – $58k per year (Estimated) • In office • Full-Time • Noida
Python
Python
pySpark
Databases
Snowflake
Databricks
Delta Lake
Apache Kafka
AI/ML
Spark
Airflow
DevOps
Azure
AWS
Platform Engineering
Amazon Kinesis
Apply
≈ $28k – $61k per year (Estimated) • In office • Full-Time • 3+ years exp • Noida
Python
JavaScript
Node JS
Python
Flask
FastAPI
Django
AI/ML
Copilot
Claude Code
Frontend
React.js
DevOps
CI/CD
Apply
In office • Full-Time • 3+ years exp • Noida
Python
JavaScript
Node JS
Python
Flask
FastAPI
Django
AI/ML
Copilot
Claude Code
Frontend
React.js
DevOps
CI/CD
Apply
≈ $28k – $60k per year (Estimated) • In office • Full-Time • 3+ years exp • Noida
Python
C#
Python
Flask
FastAPI
Django
C#
.NET
AI/ML
Copilot
Claude Code
DevOps
CI/CD
Apply
See all jobs
This is one of many
845,403 more open roles from verified company boards, updated every day.