Lead, Site Reliability Engineering

Posted An Hour Ago
Be an Early Applicant
Hiring Remotely in Dublin, IRL
Remote or Hybrid
Entry level
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
We are a global technology company in the payments industry.
The Role
Leads proactive site reliability engineering across application, platform, and infrastructure teams. Analyzes incidents and operational signals to identify failure patterns, improve observability, automate repetitive work, reduce toil, and strengthen system resilience. Designs monitoring and alerting, leads high-severity troubleshooting and root-cause analysis, drives preventative architectural improvements, and mentors engineers in systems thinking and data-driven reliability practices.
Summary Generated by Built In
Our Purpose
Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.
Title and Summary
Lead, Site Reliability Engineering
Site Reliability Engineer (SRE) - Generalist
Role Summary
The Site Reliability Engineer (SRE) - Generalist is a senior level engineer and cross stack reliability expert who proactively ensures system stability, performance, and operational resilience by deeply understanding application behavior and how it manifests across infrastructure.
This role emphasizes anticipation over reaction. While the SRE Generalist participates in incident response, their primary value is in converting operational signals, incidents, and patterns into preventative actions-improving observability, reducing risk, and eliminating classes of failure before they impact customers. They partner closely with application, platform, and infrastructure teams to continuously reduce mean time to detect (MTTD), mean time to resolve (MTTR), and overall incident frequency through data driven insight, automation, and engineering rigor.
Key Responsibilities
Proactive Reliability Engineering• Anticipate reliability risks by analyzing application behavior, system signals, and historical incidents to identify failure patterns and systemic weaknesses before they result in outages.• Translate deep application knowledge into reliability requirements, architectural guidance, and infrastructure improvements that prevent incidents rather than simply respond to them.• Continuously assess system health, resiliency gaps, and operational debt, driving improvements that increase service robustness over time.
Incident Response as an Input to Prevention• Participate in and lead troubleshooting efforts during high severity and cross domain incidents, applying structured, data driven investigation techniques.• Use incidents as learning opportunities-performing root cause analysis that focuses on why systems allowed failure, not just what broke.• Ensure incident outcomes result in concrete, measurable improvements such as better instrumentation, safer defaults, automation, or architectural changes.
Observability, Monitoring & Signal Quality• Proactively design and evolve observability strategies by onboarding new data sources and improving signal quality across logs, metrics, traces, and events.• Build dashboards, alerts, and monitors that surface early indicators of degradation, not just failure states.• Apply analytical techniques to detect emerging trends, weak signals, and anomalous behavior before customers are impacted.• Communicate insights through clear data storytelling that enables engineering teams and leaders to act decisively and early.
Automation & Continuous Improvement• Lead automation efforts that reduce manual intervention, shorten feedback loops, and eliminate repetitive operational work.• Convert operational learnings into reusable tools, standards, documentation, and patterns that raise the reliability baseline across teams.• Actively reduce operational toil and risk by improving system defaults, guardrails, and self healing capabilities.
Collaboration, Influence & Mentorship• Partner across application, infrastructure, and platform teams to drive shared ownership of reliability outcomes and proactive operational thinking.• Influence design and delivery decisions by representing the reliability perspective early in the development lifecycle.• Mentor engineers by modeling proactive troubleshooting, systems thinking, and data driven decision making.
Knowledge, Skills & Abilities• Strong ability to reason about systems end to end, connecting application behavior to infrastructure performance and failure modes.• Expertise in observability, monitoring, and troubleshooting tools, with a focus on signal quality and actionable insight.• Proficiency in scripting and automation to operationalize reliability improvements and accelerate learning.• Broad infrastructure knowledge (networking, Linux, databases, containers, storage), with depth in at least one domain.• Strong data analysis and storytelling skills, enabling proactive identification of risks and clear communication of technical insights.• Working knowledge of machine learning concepts and their application to predictive and proactive operational problem solving.• Curiosity, ownership, and a mindset oriented toward preventing tomorrow's incidents, not just fixing today's.
What Defines Success in This Role
A successful SRE Generalist:• Sees incidents as signals, not endpoints.• Uses observability and data to shift reliability work left and upstream.• Reduces incident frequency and impact over time-not just MTTR.• Acts as a connective force across teams, turning complexity into clarity and prevention.
Corporate Security Responsibility
All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
  • Abide by Mastercard's security policies and practices;
  • Ensure the confidentiality and integrity of the information being accessed;
  • Report any suspected information security violation or breach, and
  • Complete all periodic mandatory security trainings in accordance with Mastercard's guidelines.

Skills Required

  • Strong ability to reason about systems end to end, connecting application behavior to infrastructure performance and failure modes.
  • Expertise in observability, monitoring, and troubleshooting tools, with a focus on signal quality and actionable insight.
  • Proficiency in scripting and automation.
  • Broad infrastructure knowledge including networking, Linux, databases, containers, and storage.
  • Depth in at least one infrastructure domain.
  • Strong data analysis and storytelling skills.
  • Working knowledge of machine learning concepts and their application to predictive and proactive operational problem solving.
  • Experience leading high-severity incident troubleshooting and root-cause analysis.
  • Ability to collaborate across application, infrastructure, and platform teams.
  • Ability to mentor engineers and influence reliability-focused design and delivery decisions.

What the Team is Saying

Jenny
Mastercard

Mastercard Compensation & Benefits Highlights

  • Retirement Support — Retirement plans are highlighted as especially strong, featuring a notably generous company match and added financial-planning resources. Feedback suggests this component stands out as a key strength of the overall package.
  • Parental & Family Support — Parental and family benefits are consistently portrayed as robust, including extended new-parent leave and assistance for fertility, adoption, and surrogacy. Feedback suggests these programs are a signature part of the offering.
  • Leave & Time Off Breadth — Paid time off is described as substantial, with multiple leave types and generous vacation and personal days noted in U.S. materials. Feedback suggests time off is frequently praised as a differentiator.

Mastercard Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Purchase, NY
38,800 Employees
Year Founded: 1966

What We Do

Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re building a resilient economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.

Why Work With Us

We live the Mastercard Way: creating value in the communities we touch, growing together through the opportunities we see, and moving fast to innovate and scale. Our collaborative culture and our passionate people are the key to what we do, driving meaningful change as one team and connecting everyone to priceless possibilities.

Gallery

Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery

Mastercard Teams

Team
Technology
Team
Cybersecurity and Threat Intelligence
Team
Consulting
Team
AI and Data
About our Teams

Mastercard Offices

Hybrid Workspace

Employees engage in a combination of remote and on-site work.

In our ongoing workplace evolution, we’ve introduced hybrid work, Work-From-Elsewhere Weeks and Meeting-Free Days.

Typical time on-site: 3 days a week
Company Office Image
HQPurchase, NY | Global Headquarters
Company Office Image
Arlington Tech Hub
Company Office Image
Atlanta, GA
Company Office Image
Bogotá, Colombia
Boston, MA
Chicago, IL
Company Office Image
Dublin Tech Hub
Gurugram, India
Company Office Image
London, UK
Company Office Image
Miami, FL | Latin America & Caribbean HQ
Mumbai, India
Company Office Image
NYC Tech Hub
Company Office Image
St. Louis Tech Hub
Company Office Image
Pune Tech Hub
Tel Aviv, Israel
Company Office Image
Sydney Tech Hub
San Francisco, CA
São Paulo, Brazil
Seattle, WA
Company Office Image
Singapore
Company Office Image
Toronto, Canada
Vancouver Tech Hub
Learn more

Similar Jobs

Mastercard Logo Mastercard

Lead Software Engineer

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Remote or Hybrid
Dublin, IRL
38800 Employees

Mastercard Logo Mastercard

Senior Enterprise Operations Engineer

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Remote or Hybrid
Dublin, IRL
38800 Employees

Mastercard Logo Mastercard

Senior Product Manager

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Remote or Hybrid
Dublin, IRL
38800 Employees

Mastercard Logo Mastercard

Manager, Software Engineering

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Remote or Hybrid
Dublin, IRL
38800 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account