377,351open jobs
9,828companies
48,306added this week
Browse all
Salary
$96k – $187k per year (Estimated)
Location
In office (Atlanta, Raleigh, Charlotte)
Seniority
Senior · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
For our clients: Provide distinctive, secure and successful client experiences through touch and technology. For our teammates: Create an inclusive and energizing environment that empowers teammates to learn, grow and have meaningful careers. Fo...

The position is described below. If you want to apply, click the Apply Now button at the top or bottom of this page. After you click Apply Now and complete your application, you'll be invited to create a profile, which will let you see your application status and any communications. If you already have a profile with us, you can log in to check status.

Need Help?

If you have a disability and need assistance with the application, you can request a reasonable accommodation. Send an email to Accessibility (accommodation requests only; other inquiries won't receive a response).

Regular or Temporary:

Regular

Language Fluency: English (Required)

Work Shift:

1st shift (United States of America)

Please review the following job description:

The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation, observability, and incident management while collaborating across multiple business and technology teams.

Responsibilities include leading major incident responses, driving problem management, and implementing automation to reduce service downtime.

The role involves standardizing observability practices, mentoring SRE team members, and contributing to enterprise-wide reliability frameworks.

Candidates require 7+ years of experience, expertise in distributed systems, Kubernetes, automation scripting, and strong leadership in incident management.

ESSENTIAL DUTIES AND RESPONSIBILITIES

Following is a summary of the essential functions for this job. Other duties may be performed, both major and minor, which are not mentioned below. Specific activities may change from time to time.

1. Implements software architecture and engineering approaches for complex initiatives within the job area, contributing to technical plans and working to achieve operational targets with major impact on results.

2. Adopts and refines advanced software engineering standards, practices, and governance mechanisms for the job area, influencing how multiple teams improve quality, reliability, and delivery.

3. Collaborates with senior engineers, product partners, and architecture teammates to shape technology approaches for the domain, providing deep technical insight and proposing solution patterns that inform local roadmaps and priorities.

4. Leads the end-to-end technical design and implementation of scalable, secure, and highly available software solutions for the job area, producing patterns and examples that other technical professionals can follow.

5. Independently troubleshoots and resolves complex technical issues in the area of responsibility, designing innovative architectures and performance, reliability, and scalability improvements that advance business objectives.

6. Provides ongoing technical guidance, coaching, and training to other engineers, delegating and reviewing work from lower-level technical professionals and raising the technical bar through design reviews and knowledge sharing.

7. Evaluates emerging technologies and techniques relevant to the job area, building prototypes and solution concepts that contribute measurable input into new features, products, or capabilities.

8. Contributes to the development of long-term technical goals and plans for the area of responsibility through well-reasoned recommendations, design proposals, and implementation experience.

9. Leads large or complex initiatives within the job area, coordinating and delegating technical work that may span outside the immediate team, and ensuring cohesive, high-quality outcomes with limited supervision.

Qualifications

Required Qualifications

The requirements listed below are representative of the knowledge, skill and/or ability required. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.

1. Bachelor’s degree in Computer Science, Software Engineering, or related field.

2. Minimum of 7 years of professional experience in software development.

3. Deep knowledge of multiple programming languages, software architecture, and design principles.

4. Deep understanding of software development lifecycle, testing, deployment, and security practices.

Preferred Qualifications

1. Advanced degree in Computer Science or related technical discipline.

2. Professional certifications such as Certified Software Development Professional (CSDP) or equivalent.

3. Deep expertise in cloud-native architectures, microservices, container orchestration, and DevOps.

4. Strong familiarity with Agile frameworks, continuous integration/continuous deployment (CI/CD), and enterprise innovation management.

5. 7+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Infrastructure Operations.

6. Deep hands-on experience with distributed systems, container orchestration (Kubernetes), and cloud-native operational tooling.

7. Proficiency with automation and scripting languages (Python, Go, PowerShell, Ansible).

8. Strong understanding of observability platforms (Splunk, Dynatrace) and event-driven monitoring.

9. Proven leadership in major incident management and cross-team technical coordination.

10. Strong grasp of networking, Linux/Unix internals, and modern infrastructure patterns.

11. Excellent communication skills, including executive-level situational awareness during critical incidents.

12. Demonstrated ability to influence technical roadmaps and drive adoption of reliability best practices.

Preferred Qualifications

  • Financial services or regulated industry experience.

  • Experience enabling large-scale SRE transformations or modernization initiatives.

  • Familiarity with chaos engineering, resilience assessments, and service failure modeling.

  • Exposure to hybrid-cloud and multi-cloud operational frameworks.

  • Experience contributing to or leading Center for Enablement functions or Communities of Practice.

Key Responsibilities

Incident & Problem Management Leadership

  • Lead major and high-severity incident response efforts,focusing on diagnosing technical rootcausestherein, and drivingmulti-team technical resolution.

  • Drive problem management to closure, ensuring systemic fixes replace recurring operational risks.

  • Establish and maintain standardized incident playbooks, escalation paths, and communication frameworks.

Reliability Engineering & Automation

  • Architect and deliver automation solutions that eliminate toil, reduce MTTR, and increase service resilience.

  • Implement intelligent alerting, anomaly detection, and event correlation leveragingAI andAIOps tools.

  • Guide and enforce SLO/SLI adoption across product teams, ensuring metrics inform decision-making and prioritization.

Observability & Operational Excellence

  • Enhance telemetry coverage across logs, metrics, traces, and events using platforms such as Dynatrace and Splunk.

  • Define and standardize enterprise observability practices, dashboards, and KPIs.

  • Ensure operational readiness of applications and platforms through resiliency testing, chaos engineering, and failure-mode validation.

Cross-Functional Leadership & Influence

  • Partner with Delivery, Architecture, Security, and Risk teams to embed reliability and resilience into design and execution.

  • Act as a change agent to elevate operational maturity and drive transformative improvements acrossWholesale.

  • Lead workshops, maturity assessments, and enablement sessions through the SRE C4E and Communities of Practice.

Standardization & Documentation

  • Develop, maintain, and enforce runbooks, response playbooks, and automated recovery patterns.

  • Contribute to enterprise SRE frameworks, templates, and maturity models.

  • Promote consistent adoption of best practices across domains and lines of business.

Mentorship & Technical Development

  • Coach and mentor Associate, Professional, and Senior SREs to build technical depth and operational discipline.

  • Provide thought leadership in SRE methodologies, cloud-native operational patterns, and automated reliability engineering.

For this opportunity, Truist will not sponsor an applicant for work visa status or employment authorization, nor will we offer any immigration-related support for this position (including, but not limited to H-1B, F-1 OPT, F-1 STEM OPT, F-1 CPT, J-1, TN-1 or TN-2, E-3, O-1, or future sponsorship for U.S. lawful permanent residence status.)

Candidate must be willing to work onsite Monday - Friday at either office in Charlotte NC, Raleigh NC, or Atlanta, GA.

General Description of Available Benefits for Eligible Employees of Truist Financial Corporation: All regular teammates (not temporary or contingent workers) working 20 hours or more per week are eligible for benefits, though eligibility for specific benefits may be determined by the division of Truist offering the position. Truistoffers medical, dental, vision, life insurance, disability, accidental death and dismemberment, tax-preferred savings accounts, and a 401k plan to teammates. Teammates also receive no less than 10 days of vacation (prorated based on date of hire and by full-time or part-time status) during their first year of employment, along with 10 sick days (also prorated), and paid holidays. For more details on Truist’s generous benefit plans, please visit our Benefits site. Depending on the position and division, this job may also be eligible for Truist’s defined benefit pension plan, restricted stock units, and/or a deferred compensation plan. As you advance through the hiring process, you will also learn more about the specific benefits available for any non-temporary position for which you apply, based on full-time or part-time status, position, and division of work.

Truist is an Equal Opportunity Employer that does not discriminate on the basis of race, gender, color, religion, citizenship or national origin, age, sexual orientation, gender identity, disability, veteran status, or other classification protected by law. Truist is a Drug Free Workplace.

EEO is the Law E-Verify IER Right to Work

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
377,351 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Atlanta
$93k – $183k per year (Estimated) • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • Atlanta • Raleigh • Charlotte
AI/ML
Anomaly Detection
DevOps
AIOps
Amazon EC2
AWS
Chaos Engineering
CloudFormation
Dynatrace
Incident Management
SLI/SLO/SLA
Splunk
Terraform
Apply
$125k – $150k per year • Equity • In office • Full-Time • 10+ years exp • Bachelor's Degree • Charlotte • Richmond • Atlanta • Raleigh
Java
Node JS
SQL
TypeScript
JavaScript
Java
Hibernate
Maven
Spring Boot
Spring Cloud
Databases
Apache Kafka
Db2
AI/ML
ChatGPT
Copilot
Frontend
React.js
DevOps
AWS
CI/CD
GitHub
GitLab
GitLab CI
Helm
JFrog Artifactory
Kubernetes
OpenShift
Platform Engineering
Rest API
Cybersecurity
SonarQube
Threat Modeling
Apply
$94k – $142k per year • Remote/Hybrid • Full-Time • 5+ years exp • Mississauga
JavaScript
SQL
Databases
Apache Kafka
DynamoDB
Frontend
React.js
DevOps
Amazon S3
AWS
Kubernetes
Apply
$107k – $161k per year • Remote/Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • New Castle
Python
Java
JavaScript
Java
Liquibase
Databases
Apache Kafka
PostgreSQL
AI/ML
LLM
Frontend
React.js
DevOps
Blue-Green Deployment
CI/CD
Docker
GitHub
Grafana
Helm
JFrog Artifactory
Kubernetes
OpenShift
Platform Engineering
Prometheus
Splunk
Tekton
Cybersecurity
Checkmarx
Snyk
Apply
$142k – $213k per year • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • New York
Java
Java
Spring Boot
DevOps
CI/CD
Docker
gRPC
Kubernetes
OpenShift
Web3
Smart Contracts
Apply
$62k – $147k per year (Estimated) • Equity • In office • Full-Time • Bachelor's Degree • Atlanta • Charlotte
Apply
$64k – $143k per year (Estimated) • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree • Charlotte • Atlanta
Python
SQL
Apply
IT Support Analyst 1 day ago
$44k – $88k per year (Estimated) • Equity • In office • Full-Time • 2+ years exp • High School Diploma • Charlotte
Apply
$93k – $183k per year (Estimated) • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • Atlanta • Raleigh • Charlotte
AI/ML
Anomaly Detection
DevOps
AIOps
Amazon EC2
AWS
Chaos Engineering
CloudFormation
Dynatrace
Incident Management
SLI/SLO/SLA
Splunk
Terraform
Apply
$100k – $115k per year • Equity • In office • Full-Time • 5+ years exp • Atlanta • Charlotte • Richmond
Apply
$133k – $338k per year • Remote/Hybrid • Full-Time • 8+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
AI/ML
Knowledge Graph
DevOps
AWS
Azure
GCP
Platform Engineering
SLI/SLO/SLA
Apply
$94k – $266k per year • Remote/Hybrid • Full-Time • 5+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
Databases
Databricks
Google BigQuery
SAP HANA
Snowflake
AI/ML
Knowledge Graph
DevOps
Azure
Apply
$133k – $366k per year • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
JavaScript
Python
TypeScript
AI/ML
Anthropic
Edge AI
Gemini
OpenAI
Management
ServiceNow
Marketing
Salesforce
Apply
$172k – $237k per year • Remote • Full-Time • 10+ years exp • PhD • Charlotte • Louisville • Tampa • Fort Lauderdale • Washington
Databases
Databricks
Snowflake
AI/ML
Claude
Claude Code
Copilot
AI Agents
DevOps
AWS
Azure
GCP
Cybersecurity
HIPAA
Apply
$133k – $338k per year • In office • Full-Time • 12+ years exp • Associate's Degree • Boston • Milwaukee • Dallas • Columbus • Kirkland
Apply
See all jobs
This is one of many
377,351 more open roles from verified company boards, updated every day.