368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$77k – $183k per year (Estimated)
Location
In office (Dublin)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Mastercard is a global technology leader in the financial payments industry, connecting consumers, financial institutions, merchants, and governments across the globe. The company processes electronic funds transfers through its expansive network, enabling credit, debit, prepaid, and commercial card transactions worldwide.

Our Purpose

Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.

Title and Summary

Lead Site Reliability EngineerWho is Mastercard?

At Mastercard technology, we work to connect and power an inclusive, digital economy that benefits everyone, everywhere, by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships, and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. We cultivate a culture of inclusion for all employees that respects their individual strengths, views, and experiences. We believe that our differences enable us to be a better team - one that makes better decisions, drives innovation, and delivers better business results.

About the Role

The Business Operations team is seeking a highly motivated and experienced Lead Site Reliability Engineer (SRE) to join our team. You will play a critical role in ensuring the reliability, scalability, and performance of our applications, supporting essential services that power Mastercard's global operations. As a thought leader in your field, you will bring technical expertise, a passion for automation, and the ability to mentor.

The role of the Business Operations Site Reliability Engineer is to be the production readiness steward for Mastercard products. As Business Operations SRE, we are responsible for ensuring that our platform is stable and healthy. We break down barriers to running our products by fostering developer run ownership and empowering developers to build resilient products. We support our developers during the application build phase in software run principles that include operational design, automation, capacity planning, and monitoring that leads to fault-tolerant, scalable products. We see the big picture and help create and enforce operations standards while facilitating an agile and learning culture.

We support daily operations with a hyper focus on triage, root cause by understanding the business impact of our products and subsequently performing blameless post-mortems. The goal of every Business Operations team is to engage early in the development lifecycle to be more proactive and upfront in the development process, and to proactively manage production and change activities to maximize customer experience and increase the overall value of supported applications.

Business Operations teams also focus on risk management by tying all our activities together with an overarching responsibility for compliance and risk mitigation across all our environments. Ultimately, the role of Business Operations is to align Product and Customer Focused priorities with Operational needs by providing continuous feedback throughout the lifecycle.

As part of the Business Operations team, you will:

  • Be a developing subject matter expert in the Site Reliability Engineering area, influencing stakeholders and applying advanced knowledge to drive achievement of area goals and initiatives by contributing to solution development and improvements for existing products, services, and/or processes.
  • Implement and maintain high-availability system solutions, ensuring stability, performance, and operational continuity.
  • Evaluate operational requirements to develop effective technical solutions within existing frameworks.
  • Lead automation and scripting efforts to streamline operational processes and incident response workflows.
  • Troubleshoot and resolve complex system issues, escalating as necessary to maintain system health and proactively address risks.
  • Contribute to documentation, knowledge sharing, and best practices to improve team operational procedures.
  • Conduct reviews and quality assurance activities to uphold organizational standards for system stability.
  • Keep current with industry trends and emerging technologies relevant to system reliability and operational automation.
  • Guide and mentor junior team members through on-the-job experiences, reviewing work and fostering a culture of continuous improvement to grow expertise around their discipline.

Role qualifications:

The ideal candidate will apply the following skills independently and consistently in complex or nuanced situations, begin using the skills to support broader goals, and be recognized as a key contributor who may coach or support others informally.

  • Observability - Ability to use scripting and tooling to implement observability solutions, enabling the collection, analysis, and visualization of metrics, logs, and traces to support incident detection, diagnosis, and continuous service improvement.
  • Programming and Scripting - Ability to write and maintain code and scripts to automate tasks, build operational tools, and support monitoring, deployment, and incident response using languages such as Python, Go, Bash, or similar.
  • Systems and Network Administration - Ability to configure, operate, and troubleshoot Linux/Unix systems and network components, applying knowledge of networking concepts, protocols, security, and system reliability.
  • Cloud Computing and Infrastructure - Ability to design, deploy, and manage applications and infrastructure on cloud platforms (e.g., AWS, Azure, GCP), ensuring scalability, security, availability, and operational efficiency.
  • Reliability and Scalability - Ability to design and operate systems for high availability, fault tolerance, and disaster recovery, while ensuring systems can scale to meet current and future demand
  • DevOps Practices - Ability to apply DevOps principles and practices, including CI/CD pipelines, containerization, and orchestration, to enable faster, more reliable software delivery and operations.
  • Troubleshooting - Capability to systematically identify, diagnose, and resolve technical issues across systems, applications, and networks, using analytical methods and tools to restore functionality, minimize disruption, and ensure stable operations.
  • Capacity Planning and Performance Optimization - Ability to monitor resource utilization, forecast future capacity needs, and optimize system performance to support growth, scalability, and efficient infrastructure usage.
  • IT Service Management - Ability to apply IT service management principles to incident, problem, and change management, ensuring reliable service delivery, effective incident response, and continuous service improvement aligned to business needs.

• Proactive Monitoring and Improvement (SRE Applications) - The ability to use application reliability signals to anticipate issues, identify risks, and drive preventative improvements that enhance application performance and availability.

Corporate Security Responsibility

All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:

  • Abide by Mastercard’s security policies and practices;

  • Ensure the confidentiality and integrity of the information being accessed;

  • Report any suspected information security violation or breach, and

  • Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Dublin
$185k – $260k per year • Remote • Full-Time • 8+ years exp • Bachelor's Degree
DevOps
AWS
CI/CD
GCP
Kubernetes
GitHub
Cybersecurity
Clair
Dependabot
OWASP Top 10
OWASP ZAP
Snyk
Trivy
Apply
$110k – $131k per year • Remote • Full-Time • 10+ years exp • Bachelor's Degree
DevOps
AWS
Incident Management
VMWare
Apply
$118k – $142k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • El Segundo
C++
MATLAB
Python
SystemC
SystemVerilog
Verilog
VHDL
DevOps
CI/CD
QEMU
Chips/EDA
UVM
Apply
$129k – $232k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Austin
C++
DevOps
CI/CD
Git
RTOS
Apply
$169k – $321k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Phoenix
AI/ML
AI Agents
Anomaly Detection
LLM Guardrails
DevOps
AWS
Kong
Amazon S3
API Gateway
Cybersecurity
Zero Trust
Apply
In office • Full-Time • 7+ years exp • Bachelor's Degree • Istanbul
Apply
$89k – $183k per year (Estimated) • In office • Full-Time • 6+ years exp • Dublin
AI/ML
AI Agents
Design
Figma
Apply
$88k – $141k per year • In office • Full-Time • O'Fallon
Analytics
Tableau
Apply
$20k – $112k per year (Estimated) • In office • Full-Time • São Paulo
Java
Analytics
Power BI
Tableau
Apply
$35k – $81k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Mexico City
Apply
$59k – $103k per year • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Dublin
Apply
$88k – $181k per year (Estimated) • In office • Full-Time • Dublin
Apply
$111k – $201k per year (Estimated) • In office • 10+ years exp • Dublin
Ruby
Apply
$74k – $196k per year (Estimated) • In office • Full-Time • 8+ years exp • Dublin
C#
Java
Node JS
Python
JavaScript
Java
Hibernate
Spring Boot
DevOps
AWS
Azure
GCP
Cybersecurity
CWE
Apply
$95k – $171k per year (Estimated) • Remote/Hybrid • 10+ years exp • Bachelor's Degree • Dublin
Java
Java
Spring Boot
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.