698,699open jobs
41,391companies
99,310added this week
Browse all
Location
In office (Mumbai)
Seniority
Middle · 3+ years exp
Overview
Company
Impact
Profile match
JPMorganChase is the largest bank in the United States by assets and one of the most systemically important financial institutions in the world, with a lineage running back through more than a thousand predecessor firms to the 1799 founding of the Bank of the Manhattan Company. It combines a dominant investment bank and markets business with Chase, the largest retail banking franchise in America, plus commercial banking and asset and wealth management. Headquartered in New York, the group is unusual among banks for the scale of its technology spending, running one of the largest engineering organisations of any financial institution and deploying its own internal AI platform across the firm.

Join us at the center of a rapidly growing technology field, where your work helps modernize complex, mission-critical systems. We’ll value your ideas, support your growth, and empower you to make reliability improvements that matter. Guidelines.docx Raw Posting.docx

Job summary

As a Site Reliability Engineer III at JPMorgan Chase within the Corporate Technology, you solve broad business problems with simple, straightforward solutions while improving the availability, reliability, and scalability of your application or platform. You use code and cloud infrastructure to configure, maintain, monitor, and optimize applications and their associated infrastructure, and you contribute meaningfully by sharing end-to-end operational knowledge across the team. We work collaboratively, communicate clearly during incidents, and focus on iterative improvements that reduce toil and improve outcomes.

Job responsibilities

  • Design appropriate-level reliability designs, guide and assist others, and build consensus with peers while supporting adoption of site reliability engineering best practices within your team.
  • Collaborate with software engineers and partner teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery (CI/CD) pipelines.
  • Implement infrastructure, configuration, and network as code for the applications and platforms in your remit, and iteratively improve solutions by decomposing problems into smaller, actionable changes.
  • Operate and optimize applications and their associated infrastructure by configuring, maintaining, and monitoring services to meet availability, reliability, and scalability expectations.
  • Resolve complex problems with technical experts, key stakeholders, and team members by using service level indicators (SLIs) and service level objectives (SLOs) to proactively address issues before they impact customers.
  • Use enterprise-authorized AI capabilities to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
  • Identify patterns in operational signals that indicate reliability risk or recurring toil, prioritize reuse-first improvements tied to SLO outcomes, and recognize roadblocks while exploring new technologies where appropriate.

Required qualifications, capabilities, and skills

  • Formal training or certification on site reliability engineering concepts and 3+ years applied experience (country-specific requirements apply: NAMR/APAC-India/LATAM/Hong Kong; EMEA/LATAM-Brazil; Singapore follows local country guidance).
  • Proficiency in site reliability engineering culture and principles, including how to implement site reliability engineering within an application or platform.
  • Proficiency in at least one programming language such as Python, Java/Spring Boot, and .NET, with experience developing, debugging, and maintaining code in a large corporate environment.
  • Working knowledge of using enterprise-authorized AI capabilities within the work environment to support site reliability engineering workflows, including strong validation habits and awareness of data sensitivity.
  • Ability to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements.
  • Experience with observability practices such as white-box and black-box monitoring, SLO alerting, and telemetry collection, with familiarity troubleshooting common networking technologies and issues.
  • Experience with continuous integration and continuous delivery tooling, plus familiarity with containers and container orchestration.

Preferred qualifications, capabilities, and skills

  • Experience in site reliability engineering / production support / DevOps / platform roles with real on-call exposure, including improving reliability through incident response, root-cause analysis (RCA)/postmortems, problem management, and measurable toil reduction (reduced pages, automated repetitive tasks, improved MTTR/MTBF); ability to lead parts of incident calls, write clear RCAs, mentor others, and drive operational standards (runbooks, alerts, SLO definitions).
  • Strong observability experience aligned to your stack, including Dynatrace (APM, alerting, dashboards) and/or Grafana + Prometheus; log analysis in Splunk (queries, dashboards, troubleshooting); ability to correlate metrics, logs, and traces to isolate issues in microservices; plus advanced observability concepts such as OpenTelemetry/distributed tracing, alert-as-code, SLO tooling, synthetic monitoring, and capacity planning using telemetry.
  • Hands-on Kubernetes operations/troubleshooting (deployments, services/ingress, configmaps/secrets, HPA, node/pod debugging) and solid AWS experience supporting containerized workloads (EKS preferred where applicable, plus IAM/VPC basics); experience with CI/CD and infrastructure as code (Terraform modules, state management, pipeline integration; Jenkins/GitLab CI; blue/green and canary deployments) and advanced platform tooling (Helm/Kustomize, service mesh such as Istio/Linkerd, policy such as OPA/Gatekeeper, secrets tools such as Vault, GitOps such as ArgoCD/Flux).
  • Strong automation mindset and microservices reliability depth: Java microservices plus scripting (Python and/or shell) to automate operational tasks (self-healing, runbook automation, unit tests) while following secure coding practices; Linux/Unix fundamentals (process/memory/disk troubleshooting, networking basics, log analysis, performance triage); operational experience supporting Kafka and/or MQ (including IBM MQ), and databases (Oracle and/or MongoDB) with awareness of performance symptoms and connection pools; certificate management in distributed systems (TLS/SSL, rotation/renewal, keystores/truststores, outage prevention); understanding of microservice failure modes (retries/timeouts, circuit breakers, rate limiting, backpressure, dependency mapping);
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
698,699 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Mumbai
Product Test Engineer 5 hours ago
$31k – $80k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Shenzhen
Python
Java
DevOps
RTOS
Apply
In office • Full-Time • Bachelor's Degree
Python
Java
SQL
Bash
Java
Apache Tomcat
Databases
MySQL
PostgreSQL
DevOps
Rest API
Terraform
Puppet
Ansible
Prometheus
CI/CD
Kubernetes
Nginx
Configuration Management
Proxmox VE
Incident Management
IAM
Linux
Apache HTTP Server
Apply
In office • Full-Time • Pune
Python
DevOps
Terraform
GCP
Azure
CI/CD
AWS
GitLab
IAM
Apply
$16k – $34k per year (Estimated) • In office • 1+ year exp • Moscow
Python
Databases
ElasticSearch
AI/ML
Weights & Biases
Jupyter Notebook
MLFlow
LLM
DevOps
Docker
TCP/IP
Apply
$32k – $69k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Bengaluru
Python
Go
DevOps
Rest API
Git
Docker
Kubernetes
DNS
Cybersecurity
Wireshark
Management
Jira
Apply
$33k – $72k per year (Estimated) • In office • 5+ years exp • Bengaluru
Python
JavaScript
TypeScript
AI/ML
Spark
AI Agents
Time Series Forecasting
Frontend
React.js
DevOps
Splunk
OpenTelemetry
Dynatrace
Prometheus
CI/CD
Grafana
Management
Agile
Apply
$24k – $59k per year (Estimated) • In office • 3+ years exp • Hyderabad
Python
JavaScript
Java
SQL
Node JS
Scala
Java
Spring Boot
Databases
PostgreSQL
DynamoDB
DevOps
Terraform
AWS CDK
CloudFormation
CI/CD
AWS
Docker
Kubernetes
Amazon EKS
AWS Lambda
Amazon EC2
Amazon S3
IAM
Amazon ECS
Amazon CloudWatch
Management
Agile
Apply
$95k – $245k per year (Estimated) • In office • London
Design
Figma
Management
Confluence
Jira
Apply
$26k – $63k per year (Estimated) • In office • 3+ years exp • Bengaluru
JavaScript
Java
Kotlin
Java
Spring Boot
Kotlin
Mockito
Databases
Cassandra
Apache Kafka
AI/ML
Copilot
Frontend
React.js
DevOps
AWS
Management
Agile
QA
Playwright
Apply
$28k – $70k per year (Estimated) • In office • 5+ years exp • Bengaluru
Python
Java
PowerShell
AI/ML
Human-in-the-Loop
DevOps
Splunk
WebRTC
Linux
Management
ITIL
Apply
$16k – $38k per year (Estimated) • In office • 7+ years exp • Bachelor's Degree • Mumbai
Apply
$17k – $38k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Mumbai
Apply
$27k – $62k per year (Estimated) • In office • Full-Time • 15+ years exp • Mumbai
Design
AutoCAD
Apply
In office • Full-Time • 2+ years exp • Master's Degree • Hyderabad • New Delhi • Bengaluru • Mumbai • Chennai
Analytics
A/B Testing
Marketing
Meta Ads
LinkedIn Ads
Google Ads
HubSpot
Marketo
GA4
LinkedIn
Apply
$48k – $101k per year (Estimated) • Remote/Hybrid • Full-Time • 15+ years exp • Bengaluru • Nagpur • Jaipur • Coimbatore • Gurgaon
ABAP
ABAP
SAP BTP
Databases
SAP BW
AI/ML
AI Agents
LLM
RAG
DevOps
CI/CD
Management
Agile
Apply
See all jobs
This is one of many
698,699 more open roles from verified company boards, updated every day.