Salary
≈ $130k – $262k per year (Estimated)
Location
In office (Houston)
Seniority
Staff · 5+ years exp
Overview
Company
Impact
Profile match
JPMorganChase is the largest bank in the United States by assets and one of the most systemically important financial institutions in the world, with a lineage running back through more than a thousand predecessor firms to the 1799 founding of the Bank of the Manhattan Company. It combines a dominant investment bank and markets business with Chase, the largest retail banking franchise in America, plus commercial banking and asset and wealth management. Headquartered in New York, the group is unusual among banks for the scale of its technology spending, running one of the largest engineering organisations of any financial institution and deploying its own internal AI platform across the firm.
Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability.
As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate & Investment Bank (CIB) Management and Support Functions Digital & Platform Services team, you hold a leadership role in your team, demonstrate strong knowledge across multiple technical domains, and advise others on the technical and business issues facing them.
Job responsibilities
- Consistently models and champions site reliability culture and practices, documents and shares knowledge within your organization via internal forums and communities of practice
- Leads initiatives to improve the reliability and stability of your team’s applications and platforms using data-driven analytics to improve service levels, proactively identifying and solving technology-related bottlenecks in areas of expertise
- Drives collaboration with your team to identify comprehensive service level indicators and the stakeholder partners to establish reasonable service level objectives and error budgets with your customers
- Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
- Serves as the main point of contact during major incidents for your application and has the skills to identify and solve the issue quickly to avoid financial loss to the business
- Offers a high level of technical expertise within one or more technical domains and provides advice and mentorship to other engineers
- Leads reuse-first adoption of AI-assisted reliability workflows across SDLC/toolchain practices (e.g., CI/CD quality checks, test/validation automation, and operational readiness), ensuring traceability/auditability, resiliency, and security controls.
Required qualifications, capabilities, and skills
- Formal training or certification on site reliability engineering concepts and 5+ years applied experience ( NAMR/APAC - India/ LATAM/ Hong Kong)
- Demonstrated proficiency in reliability, scalability, performance, security, enterprise system architecture, toil reduction, and other site reliability best practices
- Fluent in at least one programming language such as: Python, Java/Spring Boot, .Net
- Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge capture) with strong validation habits and awareness of data sensitivity.
- Ability to evaluate AI-assisted operational recommendations for correctness and risk, define appropriate guardrails for team usage, and ensure outcomes align to resiliency and security expectations.
- Proficient knowledge and experience in observability such as white and black box monitoring, service level objective alerting, and telemetry collection
Proficient with continuous integration and continuous delivery practices and tooling
Preferred qualifications, capabilities, and skills
- Proficient with container and container orchestration
- Experience with troubleshooting common networking technologies and issues
- Advanced knowledge of software applications and technical processes with emerging depth in one or more technical disciplines, and actively self-educates to evaluate and recommend suitable new technologies
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
677,730 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
Similar stack
Same company
Houston
Integration Developer- Apache/Kafka stack
3 hours ago
≈ $72k – $151k per year (Estimated) • Remote • Secret • Full-Time • 3+ years exp
Python
Java
Kotlin
Java
Spring Boot
Quarkus
Apache Camel
Databases
ActiveMQ
Apache Kafka
DevOps
Rest API
OpenShift
Prometheus
CI/CD
Jenkins
Git
Kubernetes
Grafana
Tekton
Management
Agile
QA
Swagger
Apply
DevOps Engineer (Azure ARO)
3 hours ago
≈ $81k – $164k per year (Estimated) • Remote • Secret • Contractor • 5+ years exp • PhD
Python
Java
PowerShell
Java
Spring Boot
DevOps
Terraform
OpenShift
Azure DevOps
Azure
CI/CD
Git
Docker
Kubernetes
Bicep
Management
Power Automate
Power Apps
Apply
Senior Kubernetes Platform Engineer
3 hours ago
≈ $81k – $156k per year (Estimated) • Remote • Full-Time • 1+ year exp
Python
Go
DevOps
Terraform
OpenShift
Helm
OpenTelemetry
cert-manager
Kustomize
Prometheus
CI/CD
GitOps
ArgoCD
Kubernetes
Grafana
Platform Engineering
Configuration Management
SLI/SLO/SLA
Apply
Mid QA Mobile Automation Engineer
3 hours ago
≈ $49k – $104k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
C#
C#
.NET
DevOps
Azure DevOps
Azure
CI/CD
Git
QA
Appium
XCUITest
Apply
Assistant Vice President - AI Operations
2 hours ago
≈ $44k – $90k per year (Estimated) • Remote/Hybrid • Full-Time • Pune
Python
AI/ML
NLP
LLM
RAG
Apply
≈ $14k – $30k per year (Estimated) • In office • Manila
SQL
Analytics
Alteryx
Management
UiPath
Apply
≈ $102k – $205k per year (Estimated) • In office • 3+ years exp • Wilmington
Apply
≈ $74k – $163k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Irvine
Apply
Apply
Lead Site Reliability Engineer
6 hours ago
≈ $138k – $278k per year (Estimated) • In office • 5+ years exp • Jersey City
Python
AI/ML
LLM
Agentic Workflows
DevOps
Terraform
GCP
Azure
CI/CD
AWS
AIOps
Error Budget
SLI/SLO/SLA
Apply
Pediatric SLP Clinical Fellow (CFY)
8 hours ago
≈ $89k – $202k per year (Estimated) • In office • 1+ year exp • Houston
Apply
$56k – $94k per year • In office • Full-Time • Arlington • New York • Austin • Chicago • Sacramento
Apply
$101k – $115k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • Houston • McLean
SQL
Apply
Business Card Product Director
9 hours ago
$155k – $255k per year • Remote • Full-Time • Bachelor's Degree • Columbus • Atlanta • Charlotte • Dallas • Houston
Apply
Sr Archer/OpenPages GRC Programmer/Analyst
9 hours ago
$57k – $113k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Chicago • Detroit • Charlotte • Dallas • Atlanta
Python
Java
C#
C#
.NET
DevOps
OpenShift
AWS
Management
ServiceNow
Agile
Apply
This is one of many
677,730 more open roles from verified company boards, updated every day.

