380,379open jobs
9,957companies
48,275added this week
Browse all
Location
In office
Seniority
Staff · 5+ years exp
Overview
Company
Impact
Profile match
JPMorgan Chase & Co. is a leading global financial services firm and the largest banking institution in the United States by assets. Headquartered in New York City, the company offers a comprehensive range of financial solutions, including investment banking, asset management, treasury services, and commercial banking. Through its widely recognized consumer division, Chase, it delivers retail banking, credit card, and mortgage services to tens of millions of households across the globe.

Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability.

As a Lead Site Reliability Engineer at JPMorgan Chase within the Enterprise technology, engineering services and platform team, youhold a leadership role in your team, demonstrate strong knowledge across multiple technical domains, and advise others on the technical and business issues facing them. Take lead and conduct resiliency design reviews, break up complex problems into digestible work for other engineers, act as a technical lead for medium to large-sized products, and provide advice and mentoring to other engineers.

Job responsibilities

  • Guides and assists others in the areas of building appropriate level designs and gaining consensus from peers where appropriate
  • Collaborates with other software engineers and teams to design and implement deployment approaches using automated continuous integration and continuous delivery pipelines
  • Collaborates with other software engineers and teams to design, develop, test, and implement availability, reliability, scalability, and solutions in their applications
  • Implements infrastructure, configuration, and network as code for the applications and platforms in your remit
  • Collaborates with technical experts, key stakeholders, and team members to resolve complex problems
  • Understands service level indicators and utilizes service level objectives to proactively resolve issues before they impact customers
  • Supports the adoption of site reliability engineering best practices within your team
  • Production 24*7 support for business-critical applications
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
  • Leads reuse-first adoption of AI-assisted reliability workflows across SDLC/toolchain practices (e.g., CI/CD quality checks, test/validation automation, and operational readiness), ensuring traceability/auditability, resiliency, and security controls.

Required qualifications, capabilities, and skills

  • Formal training or certification on site reliability engineering concepts and 5+ years applied experience
  • Proficient in site reliability engineering (SRE) culture and principles, with experience implementing SRE practices within applications and platforms; strong observability background including white/black-box monitoring, SLO-based alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, and similar.
  • Proficient in at least one programming language (e.g., Python, Java/Spring Boot,.NET) with strong knowledge of software applications and technical processes within a technical discipline such as cloud, artificial intelligence, Android, or related areas.
  • Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge capture) with strong validation habits and awareness of data sensitivity.
  • Ability to evaluate AI-assisted operational recommendations for correctness and risk, define appropriate guardrails for team usage, and ensure outcomes align to resiliency and security expectations.
  • Hands-on experience with CI/CD tooling (e.g., Jenkins, GitLab) and infrastructure automation using Terraform to build reliable, repeatable delivery pipelines.
  • Strong familiarity with containers and orchestration platforms (Docker, Kubernetes, ECS), including deploying, scaling, and operating containerized services in production.
  • Proven ability to troubleshoot and resolve common networking issues (DNS, TCP/IP, routing, TLS, load balancing), applying structured debugging to restore service quickly.
  • Collaborative, proactive team contributor: communicates clearly and persuasively with minimal supervision, identifies roadblocks early, learns new technologies quickly, and has experience with event streaming platforms such as Kafka.

Preferred qualifications, capabilities, and skills

  • Ability to identify new technologies and relevant solutions to ensure design constraints are met by the software team
  • Proven track record of initiating and executing ideas that address complex business challenges
  • Deep expertise in networking and systems, including TCP/IP, DNS, load balancing, firewalls, and VPN technologies; strong Linux performance tuning and system-level troubleshooting skills
  • Certifications a plus: AWS Certified SysOps Administrator or AWS Professional, Certified Kubernetes Administrator (CKA), Terraform Associate (or equivalent)
  • Collaborative leader with a proven track record mentoring junior engineers, driving SRE best-practice adoption across teams, and communicating clearly to both technical and non-technical stakeholders (including presentations)
  • Experience in handling critical incident and change management - be part of critical incident taskforce call.
  • Familiarity of agile practices - preferably, scrum and Kanban
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
380,379 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$115k – $175k per year • In office • Full-Time • New York
C++
Java
Python
SQL
C#
C#
.NET
DevOps
Azure
Azure DevOps
CI/CD
Management
Jira
Apply
$70k – $150k per year • In office • Full-Time • 7+ years exp • Master's Degree • Toronto
Python
Python
FastAPI
AI/ML
Amazon SageMaker
NumPy
Pandas
PyTorch
RAG
Scikit-learn
TensorFlow
DevOps
Amazon CloudWatch
Amazon ECS
Amazon EKS
Amazon EventBridge
Amazon S3
API Gateway
AWS
AWS Lambda
CI/CD
Git
IAM
Kubernetes
Platform Engineering
Apply
Backend Developer 8 hours ago
In office • Full-Time • 5+ years exp • Lahore • Islamabad
C#
SQL
JavaScript
C#
.NET
Frontend
React.js
DevOps
AWS
Apply
$142k – $221k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Seattle
PowerShell
Python
DevOps
AWS
Azure
CI/CD
Git
IAM
Platform Engineering
Cybersecurity
Crowdstrike
MISP
MITRE ATT&CK
Recorded Future
STIX
TAXII
Threat Modeling
ThreatConnect
Apply
$90k – $180k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Arlington
C++
MATLAB
Python
C
C++
CMake
C
GCC
DevOps
CI/CD
GitLab
Jenkins
Chips/EDA
Ansys RedHawk
Apply
$94k – $189k per year (Estimated) • In office • Tampa
Apply
$98k – $205k per year (Estimated) • In office • 3+ years exp • Plano
Python
SQL
Python
pySpark
Databases
Amazon Aurora
Apache Kafka
Databricks
Delta Lake
Kafka
Snowflake
AI/ML
Copilot
Feature Store
LLM
Spark
DevOps
Amazon ECS
Amazon EKS
Amazon S3
AWS
AWS Lambda
CI/CD
GitHub
Terraform
Kubernetes
Apply
$195k – $395k per year (Estimated) • In office • 8+ years exp • Palo Alto
Apply
$108k – $218k per year (Estimated) • In office • 5+ years exp • Austin
Apply
$143k – $258k per year (Estimated) • In office • 5+ years exp • Atlanta
Apply
See all jobs
This is one of many
380,379 more open roles from verified company boards, updated every day.