Manager, Site Reliability Engineering
$121k - $169kMastercard
Manager, Site Reliability Engineering
Overview:
Ethoca and Mastercard are seeking a Manager, Site Reliability Engineering, to lead globally distributed application operations teams supporting mission-critical Java applications across on-premises infrastructure and Azure public cloud environments. You will provide leadership, coaching, and technical oversight for teams operating in a 24/7 support model while partnering with Ethoca and Mastercard stakeholders to deliver secure, resilient technology solutions and drive observability, incident management, operational excellence, and continuous improvement. The ideal candidate is customer- and employee-focused, intellectually curious, analytical, and comfortable leading through ambiguity in a fast-paced, highly secure environment. Role: • Lead and manage globally distributed application operations teams supporting mission-critical Java-based applications across on-premises infrastructure and Azure public cloud environments in a 24/7 support model.• Provide strong leadership, coaching, and mentorship to application operations teams, driving employee development and a collaborative, high-performance culture.
• Provide escalation support to the team, including in-depth problem investigation and resolution of complex technical issues.
• Establish and promote sustainable incident management practices, including blameless post-incident reviews, escalation management, and continuous operational and process improvement.
• Champion application observability across logging, monitoring, and alerting through APIs and tools such as Azure Monitor, Dynatrace, Splunk, Zabbix, or similar platforms.
• Partner with Ethoca and Mastercard stakeholders to design, implement, operate, configure, and optimize secure, resilient, end-to-end technology solutions within a highly secure computing environment.
• Lead customer and regulatory audits, including PCI-DSS and SOC 2 compliance and annual disaster recovery exercises, while developing and maintaining high-quality technical documentation, procedures, and standards.
• Manage multiple concurrent projects while ensuring business-as-usual operational activities are consistently delivered. All About You: • Proven experience managing Azure public cloud services, Linux or Unix servers, and Java-based applications in secure, large-scale environments.
• Demonstrated technical leadership, coaching, mentorship, delegation, conflict management, and negotiation skills, with excellent written and verbal communication.
• Strong familiarity with observability and monitoring platforms such as Zabbix, Dynatrace, Splunk, Azure Monitor, or similar tools.
• Strong security and networking knowledge, including encryption, hardening, certificates, vulnerability management, PKI, SSH, VPN, and core networking concepts.
• Working knowledge of ITIL, ITSM, and operational best practices, with hands-on experience using tools such as Remedy, Jira, and Confluence.
• Strong problem-solving skills, including the ability to guide complex technical issue resolution, think quickly, and deliver effective solutions under pressure.
• Highly accountable and outcome-focused, with the ability to manage multiple initiatives, projects, and team activities in a fast-paced environment while delivering high-quality results on time.
• Proven ability to create, follow, and guide others in adhering to documented processes and procedures while staying current with emerging technologies.
• Bachelor’s degree in Computer Science, Engineering, or equivalent practical experience. Preferred: • Experience supporting highly secure, regulated, or compliance-driven environments, including customer or regulatory audits such as PCI-DSS, SOC 2, or disaster recovery exercises.
• Relevant technical certifications in Azure, cloud operations, site reliability engineering, ITIL, security, or related disciplines.
Mastercard is a merit-based, inclusive, equal opportunity employer that considers applicants without regard to gender, gender identity, sexual orientation, race, ethnicity, disabled or veteran status, or any other characteristic protected by law. We hire the most qualified candidate for the role. In the US or Canada, if you require accommodations or assistance to complete the online application process or during the recruitment process, please contact View email address on us.fitly.work and identify the type of accommodation or assistance you are requesting. Do not include any medical or health information in this email. The Reasonable Accommodations team will respond to your email promptly.
Corporate Security Responsibility All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
- Abide by Mastercard’s security policies and practices;
- Ensure the confidentiality and integrity of the information being accessed;
- Report any suspected information security violation or breach, and
- Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines.
In line with Mastercard’s total compensation philosophy and assuming that the job will be performed in Canada, the successful candidate will be offered a competitive pay based on location, experience and other qualifications for the role and may be eligible to participate in a discretionary annual incentive program. This posting reflects one or more current openings on our team.
Pay RangesToronto, Canada: $121,000 - $169,000 CAD
$87k - $105k
...digital experiences, and identity and access management. You will: Manage multiple cloud... ...health—utilization, performance, and reliability—across our infrastructure Understand... ...code review—that automates reliability engineering work: Deployment tooling Fault-...SuggestedFull timeCasual workLocal areaWorldwideFlexible hoursShift work$90k - $125k
...Reserve]( Learn more at [chain.link]( About The Role The Engineering Team As Chainlink Labs’ engineering org scales, the DevEx... ...define how engineering teams build, test, and deploy, shaping the reliability and scalability of the developer experience org-wide....SuggestedRemote jobFull time$120k - $200k
...Coinbase Ventures, Uniswap Labs, Circle Ventures, Delphi Digital, and many more. ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems engineering, dedicated to crafting and maintaining large-scale...SuggestedFull time$158.3k - $195.24k
...People Business Partner to support our Engineering, Data Science, and Security/IT organizations... ...early, and navigating performance management, promotions, leveling, and employee relations... ...other than a Gusto office, a secure, reliable, and consistent internet connection is...SuggestedFull timeWork at officeLocal area2 days per week3 days per week$120k - $150k
...-incident reviews. Strong incident management experience, ITIL knowledge, and observability... ...operational excellence and platform reliability at scale and love ensuring the... ...findings and recommendations to product and engineering teams on a regular cadence. Enforce...SuggestedFull timeShift workNight shiftWeekend work$164.49k - $197.39k
...open culture. Grafana Cloud, our fully managed observability platform, is flexible and... ...Salesforce – trust Grafana Labs to ensure reliability of their applications and systems,... ...tool includes a rule-based recommendation engine that provides useful contextual recommendations...Full timeLocal areaRemote workFlexible hours- ...About the AI Platform Foundations Team AI Platform Foundations is a team within Platform Engineering on a mission to make AI a reliable, governed, and productivity-multiplying force across Wealthsimple's engineering organization. We build and maintain the shared AI tooling...Full timeImmediate start
$115k
...experience in solutions architecture, platform engineering, or enterprise SaaS implementations.... ...advantage. We need the ability to reliably commute to or relocate to Mississauga, Ontario... ...We mentor administrators so they can manage more complex configurations and...Full timeWork at officeRemote workWorldwideRelocation$100k - $140k
...systems-minded Senior Platform Engineer to join our core Platform... ...infrastructure maturity, driving reliability, system scaling parameters,... ...Kubernetes (AKS) clusters, managing complex multi-environment configurations... ...DevOps pipeline automation, site reliability engineering (SRE)...Permanent employmentFull timeWork at officeLocal areaRemote workWork from homeShift work- ...Job Title: Platform Modernization Engineer Location: Montreal, QB Job Type:... ...cloud technologies to improve platform reliability, resilience, and operational excellence... ...Software Development Vulnerability Management Application Security High Availability...Full time
- ...of this role As a Lead People Systems Engineer, you'll drive how GitLab's People systems... ...GitLab's People automation ecosystem running reliably at scale across Workato, Google Cloud,... ..., code reviews, and disciplined change management practices. Develop and manage Slack...Full timeRemote work
- An overview of this role As a Lead People Systems Engineer, you'll drive how GitLab's People systems evolve through an AI-first transformation... ...that keep GitLab's People automation ecosystem running reliably at scale across Workato, Google Cloud, Claude, Workday, and the...Full time
- ...is looking for an experienced AWS Cloud Engineer – Platform Operations to support and... ...This role combines AWS Cloud Operations, Site Reliability Engineering (SRE), Infrastructure Automation... ...and related platform services. Manage Infrastructure as Code (IaC) using Terraform...Full timeInternshipRemote workRelocation
- ...provision, deploy, observe, and optimize infrastructure. We give engineering teams unprecedented velocity without ever compromising on... ...# Screening with Marie, Head of HR(45 min to 1hr) # Hiring Manager interview to deep dive into your tech skills (45 min to 1hr)...Remote work
- ...The Forward Deployment Engineer (FDE) drives the on-site deployment, integration, and scaling of our enterprise Generative AI solutions. This role... ...● Leverage the Agent Developer Kit (ADK) to build and manage multi-agent systems that collaborate to solve end-to-end...Full time
- ...****@*****.*** Job Title: Senior Avionics Systems Engineer Open IMA Platform Location: Montreal, Quebec, Canada Work... ...Define and maintain Interface Control Documents (ICDs) and manage interfaces between hosted applications, platform services, and...Full time
$14 - $36 per hour
...Role Overview Contribute detailed, real-world software engineering workflows that teach and evaluate AI systems how to perform multi... ...that mirror software engineering work, for example pull request management, issue triage, merge conflict resolution, and CI...Hourly payFor contractorsRemote work$210k - $278k
...We're hiring a Senior or Staff Software Engineer to work across our product teams. This is... ...with senior engineers and engineering managers, influence how we build across the organization... ...in over time Identify and address reliability, performance, or scalability issues...Full timeRemote workFlexible hours$200.7k - $250.9k
...In 1965, an engineer in Scotland was given a mundane task with a tricky problem to it -... ...to sending payments, issuing cards, and managing invoices. With the product now in customers... ...the tests that prove it Own the reliability and quality of what you ship, from initial...Full timeShift work- About the Team The Leverage team lives inside Book of Record (BOR) - the ledger of truth for every dollar that moves through Wealthsimple. Leverage owns the borrowing side of that ledger: the systems behind margin, portfolio lines of credit, and the lending products that...Full time
$133k - $183k
...fluent, and development-oriented Software Engineer II (Money Movement & Card Ledger) to... ...Partnering cross-functionally alongside Product Managers, Data Scientists, and global engineering... ...primitives, targets strict software reliability targets, and participates in on-call...Permanent employmentFull timeRemote workFlexible hoursShift work$111k - $160k
...greatest potential. Title and Summary Senior Software Engineer Overview: The Decision Management Program (DMP) team is looking for a Senior Software... ...and help deliver highly scalable, secure, and reliable solutions across Mastercard's global payments network...Full timeImmediate startWorldwide$91k - $113k
...digital experiences, and identity and access management. Join our team: Join the PingOne SSO... ...We are looking for a Senior Software Engineer who can deliver high-quality production... ...SaaS environment. Implement secure, reliable Java-based microservices and APIs that...Full timeLocal areaWorldwideFlexible hours- ...financial infrastructure with hands-on support so founders can operate efficiently and scale with confidence. We’re hiring a Software Engineer who wants real ownership in a fast-moving startup environment. What You’ll Do Build full-stack web features (front-end,...Full time
- ...detect errors and measure model capabilities across the software engineering lifecycle. Key Responsibilities Curate and author code... ...generated code for correctness, efficiency, scalability, and reliability, and provide structured feedback and improvements....For contractorsRemote work10 hours per weekFlexible hours
$61.6k - $113.9k
...both development and production settings. Assist in release management, version control, and ongoing improvement initiatives.... ...resource to stay updated with new technologies, frameworks, and engineering best practices. Technologies: AI Cloud Copilot...Full timeWork at office2 days per week3 days per week- ...BeyondTrust is seeking a Software Development Engineer to help build the NHI Governance domain within our Pathfinder products. You... ...Investigate and fix bugs, and help improve the quality and reliability of the services your team owns. Collaborate with product, design...Full time
- ...signal, at low latency and high reliability in a trading environment... ...harnesses — tools, context management, evals, guardrails — that do... ...including what broke. ~ Strong engineering fundamentals in Python, Go,... ..., monthly wellness plan, on-site weekly massages, and games...Full timeLocal area
- ...ownership of quality, testing, security, and reliability from the earliest stages of the... ...workflows. Collaborate with product managers, designers, and other developers to define... ...Bachelor’s degree in computer science, Engineering, or a related field (or equivalent...Full timeLocal areaRemote workShift work
- ...Job Title: Software Engineer Location: Calgary, AB (Onsite) Job Type: Full-Time Java Software Engineer Kafka & Microservices Requirements: ~5+ years of Java development experience ~ Strong experience with Spring Boot and Microservices...Full timeWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!
- on-site clinical research associate (traveling/remote) Canada
- junior website developer Canada
- historic site Canada
- construction site safety Canada
- official site Canada
- site safety Canada
- site reliability engineer remote
- site reliability engineer sre
- site reliability engineering manager
- site reliability engineer
