Principal Site Reliability Engineer
Mastercard
Principal Site Reliability Engineer
Who is Mastercard?
At Mastercard technology, we work to connect and power an inclusive, digital economy that benefits everyone, everywhere, by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships, and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. We cultivate a culture of inclusion for all employees that respects their individual strengths, views, and experiences. We believe that our differences enable us to be a better team – one that makes better decisions, drives innovation, and delivers better business results.
What we create today will define tomorrow. Revolutionary technologies that reshape the digital economy to be more connected and inclusive than ever before. Safer, faster, more sustainable. And we need the best people to do it. Technologists who are energized by the challenges of a truly global network. With the talent and vision to create the critical systems and products that power global commerce and connect people everywhere to the vital goods and services they need every day. Working at Mastercard means being part of a unique culture. Inclusive and diverse, a rich collaboration of ideas and perspectives. A place that celebrates your strengths, values your experiences, and offers you the flexibility to shape a career across disciplines and continents. And the opportunity to work alongside experts and leaders at every level of the business, improving what exists, and inventing what’s next. About the Role
The Business Operations team is seeking a Principal Site Reliability Engineer.
The role of Business Operations Organization is to be the production readiness steward for Mastercard products. As a Business Operations SRE, we are responsible for ensuring that our platform is stable and healthy. We break down barriers to run our products by fostering developer run ownership and empowering developers to build resilient products. We support our developers during the application build phase in software run principals that includes operational design, automation, capacity planning, monitoring that leads to fault-tolerant, scalable products. We see the big picture and help create and enforce operations standards while facilitating an agile and learning culture. We are seeking a highly motivated and experienced Principal Site Reliability Engineer to join our growing team. You will play a critical role in ensuring the reliability, scalability, and performance of our applications, supporting essential services that power Mastercard's global operations. As a thought leader in your field, you will bring technical expertise, a passion for automation, and the ability to mentor. We support daily operations with a hyper focus on triage, root cause by understanding the business impact of our products and subsequently performing blameless post-mortems. The goal of every Business Operations team is to engage early in the development lifecycle to be more proactive and upfront in the development process, and to proactively manage production and change activities to maximize customer experience and increase the overall value of supported applications. Business Operations teams also focus on risk management by tying all our activities together with an overarching responsibility for compliance and risk mitigation across all our environments. Ultimately, the role of Business Operations is to align Product and Customer Focused priorities with Operational needs by providing continuous feedback throughout the lifecycle. Team Specific Skills:
It is not expected that any single candidate would have expertise across all these areas, but a Site Reliability
engineer will spend time throughout their career with all of these aspects of the role:
• Operational Readiness Architect:
o Serve as the primary contact responsible for the overall application health,
performance, and capacity
o Support services before they go live through activities such as system design
consulting, capacity planning and launch reviews.
o Partner with the development and product team of a new application to establish
the right monitoring and alerting strategy and create the framework to achieve zero
downtime during deployment.
• Site Reliability Engineering:
o Performs operability and resilience design and implements and maintains highly
reliable and scalable infrastructure.
o Perform root cause analysis of incidents and collaborate with development teams to
resolve issues.
o Stay up to date with the latest technologies and trends in SRE and cloud computing.
o Participate in on-call rotations and be available to respond to critical incidents.
o Complete end-to-end run ownership of the product.
Practice sustainable incident response and blameless post-mortems while taking a
holistic approach to problem solving and optimizing time to recover.
o Automate data-driven alerts to proactively escalate issues. Work with development
teams to establish SLOs and improve reliability.
• DevOps/Automation:
o Tackle complex development, automation, and business process problems. o Support the application CI/CD pipeline for promoting software into higher
environments through validation and operational gating, and lead Mastercard in
DevOps automation and best practices.
o Performs operational and resilience Design and implements solutions for capacity
planning and performance optimization.
o Increase automation and tooling to reduce toil and manual intervention
• ITSM Practices:
o Analyses ITSM activities of the platform and provide feedback loop to development
teams on operational gaps or resiliency concerns Role qualifications:
The ideal candidate will have experience in many of these areas:
• BS degree in Computer Science or related technical field involving coding (e.g., physics or
mathematics), or equivalent practical experience.
• Strong understanding of DevOps principles, practices along with configuration management.
• Experience in operational and resilience designing, building, and operating large-scale,
distributed systems.
• Appetite for change and pushing the boundaries of what can be done with automation. Be
curious about new technology, infrastructure, and practices to scale our architecture and
prepare for future growth.
• Experience with algorithms, data structures, scripting, pipeline management, and software
design.
• Systematic problem-solving approach, analytical, coupled with strong communication skills
and a sense of ownership and drive.
• Interest in designing, analysing, and troubleshooting large-scale distributed systems.
• Strong leadership and mentoring skills.
• A passion for observability, automation and continuous improvement.
• Willingness and ability to learn and take on challenging opportunities and to work as a
member of matrix based diverse and geographically distributed project team.
• Ability to balance doing things right with fixing things quickly. Flexible and pragmatic, while
working towards improving the long-term health of the system.
• Comfortable collaborating with cross-functional teams to ensure that expected system
behaviour is understood and monitoring exists to detect anomalies. Corporate Security Responsibility
All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
- Abide by Mastercard’s security policies and practices;
- Ensure the confidentiality and integrity of the information being accessed;
- Report any suspected information security violation or breach, and
- Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines.
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...SuggestedFull timeWorldwide
- ...and services that help people, businesses and governments realize their greatest potential. Title and Summary Director, Site Reliability Engineering Director, Site Reliability Engineering Our Purpose: Mastercard powers economies and empowers people across more...SuggestedFull timeWorldwide
- ...and services that help people, businesses and governments realize their greatest potential. Title and Summary Manager, Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that benefits...SuggestedFull timeWorldwideShift work
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Lead, Site Reliability Engineer (Infrastructure operations) Lead SRE Engineer, Site Reliability Engineering Our Purpose: Mastercard powers...SuggestedFull timeWorldwide
- ...governments realize their greatest potential. Title and Summary Principal Software Engineer About Mastercard Mastercard powers a global,... ...excellence across development, testing, security, and reliability Embed strong development lifecycle discipline (PDLC) to...PrincipalFull timeWorldwide
- As a Principal Cloud Security Engineer at LastPass, you will partner with DevOps and CI/CD engineers and our Architects team to ensure security best practices are embedded across our cloud infrastructure. We are looking for an experienced security engineering leader with...PrincipalFull time
- ...governments realize their greatest potential. Title and Summary Principal DevOps Engineer - Decision Management Platform Overview Join... ...Terraform and CloudFormation. • Drive observability and reliability through monitoring, logging, and alerting systems (Prometheus...PrincipalFull timeWorldwide
- ...businesses and governments realize their greatest potential. Title and Summary Principal Platform Architect Overview Infrastructure Design Services is a team of Technology Architects and Engineers responsible for defining and governing the enterprise infrastructure...PrincipalFull timeWorldwide
- At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote...Full timeRemote workWorldwide
- QCP is Asia's leading digital asset partner, empowering clients to seamlessly integrate digital assets into their portfolios. We offer a comprehensive range of solutions - from spot on/off ramping and fixed income strategies to vanilla options and bespoke exotics. ...Remote jobFull timeCasual workFlexible hours
- ...their greatest potential. Title and Summary Senior Software Engineer Overview The CNPF Data & AI organisation is looking for an... ...platform innovation—turning emerging technologies into secure, reliable, and reusable capabilities that create measurable business value...Full timeWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Principal Information Security Engineer - Blockchain Technology Who is Mastercard? Mastercard is a global technology company in the payments industry...PrincipalFull timeWork experience placementWorldwide
- ...yourself at Twilio Join the team as Twilio’s next Senior Software Engineer. About the job This position will be part of a team of... ...to architecture and design (architecture, design patterns, reliability and scaling) of new and current systems. ~ Experience...Full timeWork experience placementRemote workWorldwide
- ...their greatest potential. Title and Summary Senior Software Engineer Overview Mastercard is a global technology company in the... ...architecture, translating business and system requirements into scalable, reliable, and maintainable solutions Define high- and low-level...Full timeWorldwide
- ...that help people, businesses and governments realize their greatest potential. Title and Summary Senior Vice President – Software Engineering, Commercial Acceptance Senior Vice President – Software Engineering, Commercial Acceptance Overview The Commercial...Full timeWorldwide
- ...their greatest potential. Title and Summary Senior Platform Engineer Overview: As a Senior Platform Engineer (specializing in... ...infrastructure. You will play a crucial role in ensuring the reliability, security, and efficiency of our z/OS communication environment...Full timeWorldwideWeekend work
- ...Arista provides global enterprises with a definitive competitive edge. Consistently decorated with prestigious awards—including Best Engineering Team, Best Company for Diversity, and top rankings for Work-Life Balance—Arista fosters a highly inclusive environment where...Permanent employmentFull timeRemote workHome officeShift work
- ...realize their greatest potential. Title and Summary Senior AI Engineer-1 Who is Mastercard? Mastercard is a global technology... ...Ensure AI solutions meet Mastercard standards for performance, reliability, security and governance Collaborate with platform, security...Full timeWorldwide
- ...realize their greatest potential. Title and Summary Senior AI Engineer Overview Mastercard is seeking a Senior AI Engineer to... ...engineering practices to move models from experimentation into reliable, performant production systems. This role represents a critical...Full timeWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Software Engineer II Overview The Virtual Card Management (VCM) team is part of the Commercial Transaction Management & Controls (CTMC) program...Full timeWorldwide
- ...realize their greatest potential. Title and Summary Software Engineer – DevOps / SRE Overview The Mastercard Authentication... ...in Dublin is hiring a Software Engineer II with emphasis on site reliability to support and evolve our Authentication and Post-Authentication...Full timeWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Software Engineer II Who is Mastercard? Mastercard is a global technology company in the payments industry. Our mission is to connect and power...Full timeWorldwide
- ...realize their greatest potential. Title and Summary Software Engineer II The Business Experimentation and Optimization (BE&O) teams... ...grow your technical skills while helping the team deliver reliable, high-quality software. Our teams are small, agile, and empowered...Full timeImmediate startWorldwide
- .... Perks : Daily catered lunches, fully stocked kitchen, on-site gym & workout classes, games room, social events and laundry service... ...team is a highly talented group of versatile software engineers. These engineers are responsible for the design, development and...Full timeSummer workInternship
- ...investigations and compliance audits. Mentor junior Splunk engineers and contribute to Splunk‑related practices. Required Skills... ...Whenever possible, we hire locally to NTT DATA offices or client sites. This ensures we can provide timely and effective support tailored...Work at officeRemote workFlexible hours
- ...language queries, whether through Omni directly or via Claude, return accurate, contextually aware results. You are the steward of how well our BI tool thinks. Partner with data engineering to optimise performance, and build what's missing Own and maintain the semantic...Full time
- About the role 🎯 Your goal will be to build and lead a world-class, AI-native Solutions Engineering organisation that enables n8n to win, deploy, and expand increasingly complex enterprise customers. To make that happen, you’ll set the strategy, strengthen our technical...Full time
- ...About the role Your goal will be to build and lead a world-class, AI-native Solutions Engineering organisation that enables n8n to win, deploy, and expand increasingly complex enterprise customers. To make that happen, you’ll set the strategy, strengthen our technical...Full timeTemporary workImmediate start
- About the Role: As a DevOps Engineer at Sardine, you'll play a critical role in evolving our infrastructure and platform tooling to support... ...like Security and Finance to ensure our systems are reliable, scalable, and cost-efficient. The role involves building and maintaining...Full time
€115k - €130k per year
...productive zone and work from there. About the Role: As a DevOps Engineer at Sardine, you'll play a critical role in evolving our... ...stakeholders like Security and Finance to ensure our systems are reliable, scalable, and cost-efficient. The role involves building and maintaining...Remote jobFull timeWorldwideHome officeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!
- principal developer Ireland
- principal engineer Ireland
- director software engineering Ireland
- general engineer Ireland
- engineering director Ireland
- chief engineer Ireland
- hotel chief engineer Ireland
- data center chief engineer Ireland
- on-site clinical research associate (traveling/remote) Ireland
- chief engineer tug






