Manager, Site Reliability Engineering
$121k - $169kMastercard
Manager, Site Reliability Engineering
Who is Mastercard?
At Mastercard technology, we work to connect and power an inclusive, digital economy that benefits everyone, everywhere, by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships, and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. We cultivate a culture of inclusion for all employees that respects their individual strengths, views, and experiences. We believe that our differences enable us to be a better team – one that makes better decisions, drives innovation, and delivers better business results.
What we create today will define tomorrow. Revolutionary technologies that reshape the digital economy to be more connected and inclusive than ever before. Safer, faster, more sustainable and we need the best people to do it. Technologists who are energized by the challenges of a truly global network. With the talent and vision to create the critical systems and products that power global commerce and connect people everywhere to the vital goods and services they need every day. About the Role
The Business Operations team is seeking a highly motivated and experienced Manager, Site Reliability Engineering (SRE) to join our team. You will play a critical role in ensuring the reliability, scalability, and performance of our applications, supporting essential services that power Mastercard's global operations. As a thought leader in your field, you will bring technical expertise, a passion for automation, and the ability to mentor. The role of the Business Operations Site Reliability Engineer is to be the production readiness steward for Mastercard products. As Business Operations SRE, we are responsible for ensuring that our platform is stable and healthy. We break down barriers to running our products by fostering developer run ownership and empowering developers to build resilient products. We support our developers during the application build phase in software run principles that include operational design, automation, capacity planning, and monitoring that leads to fault-tolerant, scalable products. We see the big picture and help create and enforce operations standards while facilitating an agile and learning culture. As part of the Business Operations team, you will:
• Oversee a team of individual contributors, supporting the execution of strategic initiatives by providing technical expertise and leadership within the Site Reliability Engineering discipline to analyze complex problems and provide novel solutions and/or improvements.
• Guide the team in automating routine tasks, troubleshooting complex issues, and optimizing system performance.
• Collaborate with cross-functional teams to develop strategies for system scalability and resilience, training team members on technical skills, operational best practices, and incident management.
• Oversee incident response efforts, ensuring timely resolution and comprehensive root cause analysis.
• Cultivate a culture of continuous improvement by promoting best practices, innovation, and proactive risk management.
• Support the implementation and maintenance of high-availability systems to ensure operational stability.
• Contribute to documentation, knowledge sharing, and best practices to improve team operational procedures.
• Lead automation and scripting efforts to streamline operational processes and incident response workflows.
• Manage a team of individual contributors(s) and/or technical lead(s), directing area processes and work to ensure that they align with functional best practices and organizational standards; conduct goal setting and performance appraisal processes to coach team members and support their professional development. Role:
• Serve as the primary contact responsible for the overall application health, performance, and capacity
• Support services before they go live through activities such as system design consulting, capacity planning and launch reviews.
• Partner with the development and product team of a new application to establish the right monitoring and alerting strategy and create the framework to achieve zero downtime during deployment.
• Serve as the primary contact responsible for ensuring application scalability, performance, and resilience.
• Practice sustainable incident response and blameless post-mortems while taking a holistic approach to problem-solving and optimizing time to recover.
• Automate data-driven alerts to proactively escalate issues. Work with development teams to establish SLOs and improve reliability.
• Tackle complex development, automation, and business process problems. Engage in and improve the whole lifecycle of services—from inception and design, through deployment, operation, and refinement.
• Support the application CI/CD pipeline for promoting software into higher environments through validation and operational gating, and lead Mastercard in DevOps automation and best practices.
• Increase automation and tooling to reduce toil and manual interventiono Analyses ITSM activities of the platform and provide feedback loop to development teams on operational gaps or resiliency concerns All about you:
The ideal candidate will have experience with leading teams in many of these areas:
• BS degree in Computer Science or related technical field involving coding (e.g., physics or mathematics), or equivalent practical experience.
• Coding or scripting exposure.
• Appetite for change and pushing the boundaries of what can be done with automation. Be curious about new technology, infrastructure, and practices to scale our architecture and prepare for future growth.
• Experience with algorithms, data structures, scripting, pipeline management, and software design• Systematic problem-solving approach, coupled with strong communication skills and a sense of ownership and drive.
• Interest in designing, analyzing, and troubleshooting large-scale distributed systems.
• Willingness and ability to learn and take on challenging opportunities and to work as a member of a matrix-based, diverse and geographically distributed project team.
• Ability to balance doing things right with fixing things quickly. Flexible and pragmatic, while working towards improving the long-term health of the system.
• Comfortable collaborating with cross-functional teams to ensure that expected system behaviour is understood and monitoring exists to detect anomalies.Experience with DevOps practices and tools such as Chef, Ansible, Artifactory, GitHub, Bitbucket, Jenkins, XLR, and Remedy
Mastercard is a merit-based, inclusive, equal opportunity employer that considers applicants without regard to gender, gender identity, sexual orientation, race, ethnicity, disabled or veteran status, or any other characteristic protected by law. We hire the most qualified candidate for the role. In the US or Canada, if you require accommodations or assistance to complete the online application process or during the recruitment process, please contact View email address on decentrajobs.com and identify the type of accommodation or assistance you are requesting. Do not include any medical or health information in this email. The Reasonable Accommodations team will respond to your email promptly.
Corporate Security Responsibility All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
- Abide by Mastercard’s security policies and practices;
- Ensure the confidentiality and integrity of the information being accessed;
- Report any suspected information security violation or breach, and
- Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines.
In line with Mastercard’s total compensation philosophy and assuming that the job will be performed in Canada, the successful candidate will be offered a competitive pay based on location, experience and other qualifications for the role and may be eligible to participate in a discretionary annual incentive program.
Pay RangesVancouver, Canada: $121,000 - $169,000 CAD
- ...digital experiences, and identity and access management. You will: Manage multiple cloud... ...health—utilization, performance, and reliability—across our infrastructure Understand... ...code review—that automates reliability engineering work: Deployment tooling Fault-...SuggestedFull timeCasual workLocal areaWorldwideFlexible hoursShift work
$95k - $134k
...10/31/2026 The Opportunity DAT is looking for a Site Reliability Engineer to join our SRE platform team. This position will work hybrid... ...closely with peer teams, platform/software architects and management to drive key reliability improvements. Willingness to...SuggestedTemporary workFor contractorsWork experience placementWork at officeLocal areaImmediate startFlexible hours- ...Digital, and many more. ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems... ...capacity and performance to uphold these standards. As Manager of SRE, you'll lead a team of engineers responsible for...SuggestedFull time
$90k - $105k
...work directly with customers to troubleshoot issues, dig into how Digs works under the hood, identify root causes, and partner with Engineering and Product on complex problems. This isn't traditional support. You'll combine strong customer instincts with technical...SuggestedFull timeWork at officeWork visa3 days per week$46.92k
...development so they can reach their full potential. Responsibilities include: Providing daily supervision and mentorship Managing household routines and student schedules Administering medications and ensuring student wellness Driving students to...SuggestedFull timeWork from homeRelocationRelocation packageFlexible hoursWeekday work$77k - $105k
...experience in design or architecture, including design patterns, reliability, and scaling of new and existing systems ~ Experience... ...experience, including coding standards, code reviews, source control management, build processes, testing, and operations ~ Bachelors...Full timeInternshipWorldwide$130.7k - $205.2k
...AI Solution Engineer Description - We are seeking an innovative and results-driven AI... ...team, you will work closely with product management, imaging and software architects, and development... ...solution quality, performance, and reliability. Troubleshoot, debug, and resolve...Full timeTemporary workLocal areaRelocationFlexible hoursShift work$77k - $105k
...system design or architecture, including reliability and scaling for new or existing systems... ...We require familiarity with software engineering best practices across the full development... ...capabilities for Amazon RDS and Aurora managed databases. Our mission is to lower...Full timeInternship$77k - $105k
...design or architecture experience, including design patterns, reliability, and scaling of new and existing systems. We require... ...experience, including coding standards, code reviews, source control management, build processes, testing, and operations. Preferred: a...Full timeInternshipFlexible hours$77k - $105k
...architecture, including design patterns, reliability, and scaling for new or existing... ...standards, code reviews, source control management, build processes, testing, and operations... ...thoughtful feedback to peers, including senior engineers, and seeking early input on our own...Full timeInternshipImmediate startWorldwide$77k - $105k
...architecting new and existing systems, including design patterns, reliability, and scaling. We require experience programming in at... ...experience, including coding standards, code reviews, source control management, build processes, testing, and operations. We prefer a...Full timeInternshipWork at office$77k - $105k
...architecture, including design patterns, reliability, and scaling, for new and existing... ...standards, code reviews, source control management, build processes, testing, and operations... ...direction for features, growing into a broader engineering voice over time. Develop project...Full timeInternship$77k - $105k
...architecture, including design patterns, reliability, and scaling for new and existing... ...closely with Principal and Distinguished Engineers to advance our technical depth and leadership... ...optimization through data lifecycle management. This role offers exposure to some of...Full timeInternship$77k - $105k
...architecture, including design patterns, reliability, and scaling of new and existing... ...foreign equivalent in Computer Science, Engineering, Mathematics, or a related field. Preferred... ..., code reviews, source control management, build processes, testing, and operations...Full timeInternship$77k - $105k
...existing systems, including design patterns, reliability, and scaling. We require experience... ..., code reviews, source control management, build processes, testing, and operations... ...Responsibilities: We operate and engineer distributed systems at scale for the Amazon...Full timeInternship$77k - $105k
...architecting new and existing systems, including patterns for reliability and scale ~ Proficiency in at least one programming... ...communication experiences Partner with product, science, design, and engineering teams to understand requirements and deliver effective...Full timeInternship$77k - $105k
...architecting new and existing systems, including design patterns, reliability, and scaling. We require experience programming in at... ...experience, including coding standards, code reviews, source control management, build processes, testing, and operations. We prefer a...Full timeInternshipWorldwide$110k - $150k
...leading design or architecture for new and existing systems, including design patterns, reliability, and scaling. We require experience as a mentor, tech lead, or engineering manager. We prefer 5+ years of full software development life cycle experience, including...Full timeInternshipImmediate startWorldwide$73.5k - $212.28k
...At PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to design and develop robust... ...for coaching, leveraging team member's unique strengths, and managing performance to deliver on client expectations. With your...H1b$90k - $130k
...least 2 years of experience in backend engineering, AI automation, or integrating complex systems... ...in live, real-world environments and managing multi-step system interactions. We... ...in refining agent workflows so they can reliably complete multi-step tasks using external...Full timeWork at office- ...Columbia Sportswear Company seeks a Manager, Software Engineering (Marketing Technology) to lead a team delivering and evolving Marketing Technology platforms, integrations, and digital capabilities. You will partner with Marketing, CRM, Digital Commerce, and Architecture...
$95k - $154k
...Currently, We are looking for entry-level software programmers, Java Full stack developers, Python/Java developers, Data analysts/Data Engineers/ Data Scientists, Machine Learning engineers for full time positions with clients. We Focus on Java /Full stack/Devops and Data...Full timeH1b$90k - $129.15k
Software Engineer (New Graduate) - North America Software Center At... ...including hardware, product management, and QA, to ensure seamless project... ...debugging software to ensure reliability, performance, and scalability... ...: This role is based on-site in Vancouver, Washington, operating...WorldwideMonday to Friday$28.85 per hour
...We are seeking a driven and dynamic Canvass Manager to lead, develop, and inspire a team of field canvassers. This role is ideal for a results-oriented leader who thrives in a fast-paced, people-focused environment and is passionate about coaching others to succeed....Relocation$95k - $115k
Senior Salesforce Developer (Remote) Ateko is seeking a Senior Salesforce Developer to work with our clients; understanding their business needs and translating them into technical solutions. In this role you will demonstrate in-depth knowledge of Salesforce development...Full timeWork at officeRemote workFlexible hours$21 - $24 per hour
...Assistant General Manager $21 - $24 / hour PLUS OT! Come join Panera Bread– an award-winning leader in the restaurant industry and employer of choice for 2022 and 2023! We are also proud to be named a Top Workplace for 2024! What's in it for you? ~ A comprehensive...Daily paidFlexible hoursShift workNight shift- ...JOB We’re looking for an experienced Full-Stack Go Software Engineer to join our innovative technology team at Trend Capital. In this... ...stand-ups all day. We pride ourselves that each software engineer works in autonomous roles with little management or guidance!...Full timeFlexible hours
$62k - $102k
Salary: $62,000 - 102,000 per year Requirements: We want someone with 6+ years of hands-on experience across Salesforce Sales Cloud and Service Cloud, including at least 3 years working with Lightning Web Components. We need proven experience in requirements gathering...Full timeWork at officeRemote workFlexible hours- ...Massage Envy - Manager Position Available! Are you looking for a dynamic workplace where you can truly make a difference? Do you thrive in an environment that values empathy , excellence , and consistency ? If so, we want you to join our team as a highly skilled...Full timeMonday to FridayFlexible hoursShift workWeekend workDay shift
$124k - $280k
At PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to design and develop robust... ...data storage solutions using cloud services Designing and managing data warehouses and data lakes Implementing IAM roles and policies...H1b
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!
- site safety Vancouver, WA
- on-site clinical research associate (traveling/remote) Vancouver, WA
- site services specialist Vancouver, WA
- construction site safety Vancouver, WA
- junior website developer Vancouver, WA
- historic site Vancouver, WA
- IT site lead Vancouver, WA
- site leader Vancouver, WA
- official site Vancouver, WA
- junior site reliability engineer




