Lead, Site Reliability Engineer (Infrastructure operations)
Full-time
Mastercard
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary
Mastercard powers economies and empowers people across more than 200 countries and territories worldwide.
We are committed to building an inclusive, digital economy that benefits everyone, everywhere—by making transactions safe, simple, smart, and accessible. Through secure data, trusted networks, strong partnerships, and relentless innovation, we help individuals, financial institutions, governments, and businesses unlock their greatest potential. About the Role:
Mastercard’s Program aligned Site Reliability Engineering (SRE) teams are dedicated to delivering a seamless experience for our customers. We achieve this by maintaining every aspect of our Programs infrastructure and technology ecosystem to the highest standards, ensuring compliance with rigorous security requirements.
Within Mastercard, SRE focuses on the reliability and performance of core infrastructure, networks, and foundational services that power our applications. Our mission is to ensure these components operate with excellence, enabling applications to deliver an outstanding customer experience.
In this role, you will join our Payments Network SRE team and take ownership of continuously assessing and elevating the end to end service quality of our platform. You will leverage data to drive root cause analysis and deliver strategic insights to key stakeholders on resource utilization, capacity forecasting, and performance trends—ensuring the availability, scalability, and resilience of our network. Key Responsibilities: Lead continuous assessments of the application infrastructure supporting critical Mastercard applications, focusing on health, performance, monitoring and alerting, and capacity analysis. Collaborate with Product and Development teams to forecast growth requirements and ensure scalability and resiliency. Champion observability as a core principle for infrastructure services by assessing environments and technologies to uncover gaps in monitoring and alerting. Design and implement strategies to close these gaps, ensuring all infrastructure telemetry is integrated into a unified, single-pane-of-glass view. Build custom dashboards to investigate and perform root cause analysis on complex issues. Lead regular incident reviews with internal support teams to ensure root causes are identified. When patterns of failure or compatibility issues between software and infrastructure emerge, develop and implement strategies to remediate or mitigate risks. Leverage automation and AI technologies to enhance proactive issue detection, enable self-healing capabilities, reducing Mean Time to Detect (MTTD) and Mean Time to Mitigate (MTTM). Develop testing and validation plans for new environment builds, disaster recovery exercises and post-maintenance activities to certify environment readiness before customer traffic is routed to it. Champion continuous learning, development, and knowledge sharing across networking and other infrastructure disciplines to strengthen multi-disciplinary SRE team capabilities. Lead training initiatives for team members and Product and Development on networking aspects of the platforms. Evaluate vendor hardware, firmware, and software upgrade roadmaps, and conduct proof-of-concept (POC) testing to identify potential risks and opportunities for improvement in upcoming releases. All about you: • 5–10 years of experience in an SRE or SRE related operations role, including 3+ years supporting e commerce, financial services, or large scale SaaS platforms.
• Excellent infrastructure troubleshooting and analytical problem solving skills.
• Strong hands on experience with observability and monitoring tools such as Splunk, Dynatrace, or equivalent, with a proven ability to triage and investigate complex issues.
• Familiarity with network telemetry tools such as SolarWinds and NetScout.
• Proficiency in packet level debugging, including capturing traffic with tools like tcpdump and analyzing packets using Wireshark.
• Broad understanding of end to end infrastructure supporting payment platforms—spanning platform services, networking, databases, and storage.
• Experience with automation and Infrastructure as Code tools such as Chef, Ansible, and Terraform, as well as structured data formats (JSON/YAML).
• Excellent communication skills with the ability to coordinate cross functional troubleshooting efforts and lead RCA processes to closure.
• Demonstrated ability to troubleshoot complex production issues, perform root cause analysis, and drive long term corrective actions.
• Experience partnering with development teams to shape architecture, define SLIs/SLOs, and embed reliability into services from design through operation.
• Strong understanding of monitoring and observability ecosystems, including Prometheus, Grafana, ELK/EFK, Splunk, Dyantrace, and OpenTelemetry.
• Effective incident management skills with a structured, analytical approach to problem solving. The Payments Network SRE team is responsible for the runtime availability of some of Mastercard’s most critical core payment systems, which support national infrastructure and operate 24/7 year‑round. As a result, this role will include periodic on‑call responsibilities when required. Corporate Security Responsibility All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must: • Abide by Mastercard’s security policies and practices;
• Ensure the confidentiality and integrity of the information being accessed;
• Report any suspected information security violation or breach, and
• Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines. Corporate Security Responsibility
Lead, Site Reliability Engineer (Infrastructure operations)
Lead SRE Engineer, Site Reliability Engineering
Our Purpose:Mastercard powers economies and empowers people across more than 200 countries and territories worldwide.
We are committed to building an inclusive, digital economy that benefits everyone, everywhere—by making transactions safe, simple, smart, and accessible. Through secure data, trusted networks, strong partnerships, and relentless innovation, we help individuals, financial institutions, governments, and businesses unlock their greatest potential. About the Role:
Mastercard’s Program aligned Site Reliability Engineering (SRE) teams are dedicated to delivering a seamless experience for our customers. We achieve this by maintaining every aspect of our Programs infrastructure and technology ecosystem to the highest standards, ensuring compliance with rigorous security requirements.
Within Mastercard, SRE focuses on the reliability and performance of core infrastructure, networks, and foundational services that power our applications. Our mission is to ensure these components operate with excellence, enabling applications to deliver an outstanding customer experience.
In this role, you will join our Payments Network SRE team and take ownership of continuously assessing and elevating the end to end service quality of our platform. You will leverage data to drive root cause analysis and deliver strategic insights to key stakeholders on resource utilization, capacity forecasting, and performance trends—ensuring the availability, scalability, and resilience of our network. Key Responsibilities: Lead continuous assessments of the application infrastructure supporting critical Mastercard applications, focusing on health, performance, monitoring and alerting, and capacity analysis. Collaborate with Product and Development teams to forecast growth requirements and ensure scalability and resiliency. Champion observability as a core principle for infrastructure services by assessing environments and technologies to uncover gaps in monitoring and alerting. Design and implement strategies to close these gaps, ensuring all infrastructure telemetry is integrated into a unified, single-pane-of-glass view. Build custom dashboards to investigate and perform root cause analysis on complex issues. Lead regular incident reviews with internal support teams to ensure root causes are identified. When patterns of failure or compatibility issues between software and infrastructure emerge, develop and implement strategies to remediate or mitigate risks. Leverage automation and AI technologies to enhance proactive issue detection, enable self-healing capabilities, reducing Mean Time to Detect (MTTD) and Mean Time to Mitigate (MTTM). Develop testing and validation plans for new environment builds, disaster recovery exercises and post-maintenance activities to certify environment readiness before customer traffic is routed to it. Champion continuous learning, development, and knowledge sharing across networking and other infrastructure disciplines to strengthen multi-disciplinary SRE team capabilities. Lead training initiatives for team members and Product and Development on networking aspects of the platforms. Evaluate vendor hardware, firmware, and software upgrade roadmaps, and conduct proof-of-concept (POC) testing to identify potential risks and opportunities for improvement in upcoming releases. All about you: • 5–10 years of experience in an SRE or SRE related operations role, including 3+ years supporting e commerce, financial services, or large scale SaaS platforms.
• Excellent infrastructure troubleshooting and analytical problem solving skills.
• Strong hands on experience with observability and monitoring tools such as Splunk, Dynatrace, or equivalent, with a proven ability to triage and investigate complex issues.
• Familiarity with network telemetry tools such as SolarWinds and NetScout.
• Proficiency in packet level debugging, including capturing traffic with tools like tcpdump and analyzing packets using Wireshark.
• Broad understanding of end to end infrastructure supporting payment platforms—spanning platform services, networking, databases, and storage.
• Experience with automation and Infrastructure as Code tools such as Chef, Ansible, and Terraform, as well as structured data formats (JSON/YAML).
• Excellent communication skills with the ability to coordinate cross functional troubleshooting efforts and lead RCA processes to closure.
• Demonstrated ability to troubleshoot complex production issues, perform root cause analysis, and drive long term corrective actions.
• Experience partnering with development teams to shape architecture, define SLIs/SLOs, and embed reliability into services from design through operation.
• Strong understanding of monitoring and observability ecosystems, including Prometheus, Grafana, ELK/EFK, Splunk, Dyantrace, and OpenTelemetry.
• Effective incident management skills with a structured, analytical approach to problem solving. The Payments Network SRE team is responsible for the runtime availability of some of Mastercard’s most critical core payment systems, which support national infrastructure and operate 24/7 year‑round. As a result, this role will include periodic on‑call responsibilities when required. Corporate Security Responsibility All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must: • Abide by Mastercard’s security policies and practices;
• Ensure the confidentiality and integrity of the information being accessed;
• Report any suspected information security violation or breach, and
• Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines. Corporate Security Responsibility
All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
- Abide by Mastercard’s security policies and practices;
- Ensure the confidentiality and integrity of the information being accessed;
- Report any suspected information security violation or breach, and
- Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines.
Vacancy posted 10 days ago
Similar jobs that could be interesting for youBased on the Lead, Site Reliability Engineer (Infrastructure operations) in Ireland vacancy
- ...ABOUT GREYSTAR Greystar is a leading, fully integrated global real estate platform offering... ..., South Carolina, Greystar manages and operates over $300 billion of real estate in more... ...the senior maintenance professional on site, the Lead Maintenance Technician oversees...OperationsContract workFor contractorsImmediate start
- ...collaborative Senior Software Engineer to join our Revenue... ...closely with Product, Revenue Operations, Finance, and Salesforce implementation... ...partners to deliver reliable, scalable solutions. As a senior... ...Rack's integration platform, lead complex technical initiatives...OperationsFull timeRemote work
- ...Summary Senior Software Engineer Senior Software... ..., testing, and operating highly scalable, resilient... ..., capacity planning, reliability, and continuous improvement... ..., network, and infrastructure performance characteristics... ..., scalability, or site reliability. •...OperationsFull timeWorldwide
- ...environments for security, IT operations, and business analytics. You will lead Splunk architecture... ...Mentor junior Splunk engineers and contribute to... ...leading AI and digital infrastructure providers, with unmatched... ...DATA offices or client sites. This ensures we can provide...OperationsWork at officeRemote workFlexible hours
- ...currently looking for a Lead Databricks Architect... ...data architects, data engineers, platform specialists... ...opportunities, leadership and reliable delivery.... ...leading AI and digital infrastructure providers, with unmatched... ...DATA offices or client sites. This ensures we can provide...SuggestedWork at officeRemote workFlexible hours
- ...next Senior Software Engineer. About the job... ...services and products that operate at a massive scale.... ...delivering industry leading availability. We do this... ...include AWS cloud infrastructure and APIs, Apache... ...architecture, design patterns, reliability and scaling) of new...Full timeWork experience placementRemote workWorldwide
- ...QCP is Asia's leading digital asset partner, empowering clients to seamlessly integrate digital assets into their portfolios. We offer... ..., and feeds to fulfil business requirements from the trading, operations, and risk teams of the company. ● To be responsible for...OperationsRemote jobFull timeCasual workFlexible hours
- ...Virtu Virtu is a leading financial firm that leverages cutting edge technology to deliver... ...a talented group of versatile software engineers who collaborate with the entire Virtu... ...team collaborate closely with traders, operations, and other developers to gather and understand...OperationsFull timeWork at officeRemote workWorldwide
- ...About the role Your goal will be to build and lead a world-class, AI-native Solutions Engineering organisation that enables n8n to win, deploy, and expand... ...Engineering, or Solutions Architecture function and can operate as a peer within a senior sales leadership team. ~...Temporary workImmediate start
- ...About the Role: As a DevOps Engineer at Sardine, you'll play a critical role in evolving our infrastructure and platform tooling to support a growing engineering team.... ...Security and Finance to ensure our systems are reliable, scalable, and cost-efficient. The role involves...Full time
- ...greatest potential. Title and Summary Lead Software Engineer - Frontend (UI) Overview Who We... ...company in the payments industry, operating in over 210+ countries and territories... ...architecture reviews, design standards, reliability practices, and secure coding patterns...Full timeWorldwide
- ...prestigious awards—including Best Engineering Team, Best Company for... ...of Dublin, Ireland. In this infrastructure-critical engineering seat, you... ...entries, you will lead an active packet forwarding,... ...graduate professional history operating inside a C++ Software Engineer...Permanent employmentFull timeRemote workHome officeShift work
- ...governments realize their greatest potential. Title and Summary Lead Product Manager, Emerging Data Signals Overview... ...real time. The successful candidate will work across product, engineering, commercialization, privacy, legal, and customer teams to build...Full timeWorldwide
- ...This role is for strong mobile builders who can ship clean, reliable and intuitive mobile experiences, whether their main strength... ...contribution clearly. People who need heavy guidance for every task. Engineers who move slowly in a high-growth environment....Full time
- ...and Summary Manager, Software Engineering Who is Mastercard?... ...global B2B technology platform, operating at-scale, requiring focus on performance, security, and reliability. Role As Manager, Software... ...requirements to production • Lead and guide an agile teams of...OperationsFull timeRemote workWorldwide
- ...Summary Manager, Software Engineering Overview Authentication... ...Token Provisioning use cases, operating at Mastercard scale with... ...Software Engineering, you will lead a team of engineers responsible... ...through strong ownership of reliability, performance, availability...OperationsFull timeWorldwide
- ...their greatest potential. Title and Summary Manager, Software Engineering Who is Mastercard? Mastercard is a global technology... ...an entrepreneurial mindset. Role: • People and Technology Lead for modernized Java Tech Stack • Manage deliverables and timelines...Full timeWorldwide
- ...their greatest potential. Title and Summary Director, Software Engineering Who is Mastercard? Mastercard is a global technology... ...Tech-hub in Leopardstown, Dublin. In this role you will be a leading a highly agile team building exiting and innovative products,...Full timeWorldwide
- ...Why join us? As our first hire dedicated to this area, you’ll have the opportunity to build the foundational campaign processes, operating rhythms, and growth levers that shape how n8n generates pipeline. Your work will be highly visible across Marketing and Sales...Full timeRemote work
- ...organization, apply now. Banking Practice Lead -AI Location: Dublin, Ireland... ...one of the world's leading AI and digital infrastructure providers, with unmatched capabilities... ...hire locally to NTT DATA offices or client sites. This ensures we can provide timely and...Work at officeRemote workFlexible hours
- ...the core of how our GTM teams operate and make decisions. You'll... ...thinks. Partner with data engineering to optimise performance, and... ..., and build the forecasting infrastructure that improves the accuracy... ...’s AI Strategy & Enablement Lead to spot processes where in-house...Full time
- ...As a Principal Cloud Security Engineer at LastPass, you will partner with DevOps and... ...practices are embedded across our cloud infrastructure. We are looking for an experienced security... ...to help engineers build and securely operate products and services from the ground...Full timeShift work
- ...potential. Title and Summary Vice President & Founding Technology Lead, AI Engineering & Consulting Platforms Overview: At Mastercard, our... ..., and experienced individual contributors. • Ability to operate as a senior technology leader: shaping vision, establishing...Full timeWorldwide
- ...Virtu is an industry-leading financial technology firm that operates both proprietary trading and client-facing businesses... ..., fully stocked kitchen, on-site gym & workout classes, games room,... ...talented group of versatile software engineers. These engineers are responsible...Full timeSummer workInternship
- About the role 🎯 Your goal will be to build and lead a world-class, AI-native Solutions Engineering organisation that enables n8n to win, deploy, and expand increasingly complex enterprise customers. To make that happen, you’ll set the strategy, strengthen our technical...Full time
- ...development, implementation, and operation of Kraken’s IT General... ...environments. Partner with Product, Engineering, Data, Finance, Compliance,... ...risks arising from system, infrastructure, process, or data changes... .... ~ Proven experience leading complex, multi-stakeholder...OperationsFull time
- ...order to execute against mutual agreed objectives across all operational functions (fraud, marketing, operations, payments, technology... ...negotiation skills. Demonstrate leadership capabilities in building, leading, and developing high performance teams. Demonstrate proven...OperationsFull timeWorldwide
- ...succeed. What You’ll Do Build trusted relationships with Sales, Deal Desk, Revenue Operations and other business stakeholders to support customer and channel transactions. Lead the review, drafting and negotiation of a wide range of commercial agreements,...OperationsFull timeContract work
- ...Senior Vice President – Software Engineering, Commercial Acceptance... ...for designing, building, and operating the technology that powers Mastercard... ...the Commercial business • Lead globally distributed teams... ...to public cloud infrastructure. • Deep knowledge of secure...Full timeWorldwide
- ...The Role We are looking for Android engineers to build the native Android experience for KIRA's AI Neobank App. This role is for... ...clean user flows, fast performance, product quality and shipping reliable mobile features used by real customers. What You'll Own...Full timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lead, Site Reliability Engineer (Infrastructure operations). Be the first to apply!
Related searches
- on-site clinical research associate (traveling/remote) Ireland
- business operations Ireland
- revenue operations Ireland
- senior vice president of operations Ireland
- business operations intern Ireland
- operations tech Ireland
- vice president of field operations Ireland
- senior operations technician Ireland
- administrative & business operations Ireland
- travel operations Ireland




