Site Reliability Engineering Manager
Mastercard
Site Reliability Engineering Manager
The Xborder team is looking for a Site Reliability Engineering Manager who can help us solve problems, implement automation, and leverage best practices.
· Are you a born problem solver who loves to figure out how something works?
· Are you a detail -oriented individual who enjoys complex problem solving?
· Do you love determining the correct actions required to fix a problem?
· Do you have a low tolerance for manual work and look to automate everything you can?
• Lead and drive the end-to-end service lifecycle, ensuring teams effectively engage from inception and design through deployment, operations, and continuous improvement, while aligning with business objectives.
• Oversee ITSM practices across the platform, establishing governance and ensuring teams proactively identify operational gaps and resiliency risks, while driving action plans in partnership with engineering teams.
• Provide strategic direction for production readiness, guiding teams on system design consulting, capacity planning, and launch readiness reviews to ensure scalable and reliable service delivery.
• Own service reliability outcomes, by defining KPIs/SLOs and leading the monitoring of availability, latency, and system health, ensuring accountability across the team.
• Drive scalability and operational efficiency, promoting automation, standardization, and continuous improvement initiatives to enhance reliability, reduce toil, and accelerate delivery velocity.
• Lead incident management excellence, establishing best practices for sustainable incident response, ensuring blameless postmortems, and driving root cause remediation and preventive actions at scale.
• Champion a holistic, cross-stack problem-solving approach, enabling teams to effectively manage complex production incidents and improve mean time to recovery (MTTR).
• Manage and develop a high-performing global team, fostering collaboration across geographies and time zones while ensuring alignment, engagement, and productivity.
• Build and nurture talent, through coaching, mentoring, and career development, while promoting a strong culture of knowledge sharing and continuous learning. All about you • Bachelor’s degree in computer science, Information Technology, or a related technical field (e.g., Engineering, Physics, Mathematics), or equivalent practical experience. Experience in financial services is preferred.
• 8–15 years of relevant experience in Site Reliability Engineering, Infrastructure, or DevOps roles, with a combination of hands-on technical expertise and early leadership responsibilities.
• Strong technical foundation across enterprise platforms, Linux/UNIX systems, operating systems, and database environments (Oracle/SQL, DBA), with the ability to provide technical guidance and support to the team.
• Experience with observability and monitoring tools (e.g., Splunk, Dynatrace), driving improved system visibility, performance, and reliability.
• Solid experience in DevOps and CI/CD practices, with the ability to support and guide automation, deployment pipelines, and operational improvements.
• Proficiency in one or more programming or scripting languages such as Python, Java, Go, C/C++, Perl, or Ruby, with practical application in automation or system improvements.
• Proven exposure to automation initiatives, with the ability to contribute to and help scale solutions that reduce operational toil and improve efficiency.
• Working knowledge of ITSM processes, including incident, problem, and change management, with experience applying these practices in production environments.
• Experience supporting customer-facing platforms, ensuring service reliability, availability, and effective issue resolution.
• Strong analytical and problem-solving skills, with the ability to troubleshoot complex issues and support the team during high-severity incidents.
• Ability to prioritize, organize, and manage multiple workstreams, balancing operational needs with ongoing improvements.
• Effective communication and collaboration skills, with experience working across engineering, product, and operations teams in a global environment.
• Demonstrated experience mentoring and supporting junior engineers, contributing to team development and knowledge sharing (formal people management experience is a plus but not mandatory).
• Understanding of large-scale distributed systems, including basic design principles, performance considerations, and troubleshooting approaches.
• Exposure to Artificial Intelligence use cases and implementation is a plus, particularly in relation to automation, observability, or operational insights. Corporate Security Responsibility
All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
- Abide by Mastercard’s security policies and practices;
- Ensure the confidentiality and integrity of the information being accessed;
- Report any suspected information security violation or breach, and
- Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines.
- ...potential. Title and Summary Director, Infrastructure & Site Reliability Engineering Who is Mastercard? Mastercard is a global... ...Lead modernization efforts including hardware lifecycle management, virtualization upgrades, and infrastructure optimization...SuggestedFull timeWorldwide
- ...governments realize their greatest potential. Title and Summary Site Reliability Engineer II Who is Mastercard? At Mastercard technology, we... ...and upfront in the development process, and to proactively manage production and change activities to maximize customer...SuggestedRemote jobFull timeWorldwide
- ...as the development, implementation and management of applications, infrastructure and... ...locally to NTT DATA offices or client sites. This ensures we can provide timely and... ...looking for a strong, hands-on Senior Site Reliability Engineer to support and improve cloud...SuggestedWork at officeRemote workFlexible hours
- ...realize their greatest potential. Title and Summary Senior Site Reliability Engineer The Xborder team is looking for a Senior Site... ...Solid understanding of ITSM processes (Change and Problem Management). • Experience with observability and monitoring tools such...SuggestedFull timeWorldwide
- ...thinking organization, apply now. We are currently seeking a Site Reliability Engineer to join our team in Guadalajara, Jalisco (MX-JAL), Mexico... ...cloud systems to prevent outages and initiate an Incident Management bridge in case of an outage. Troubleshoot Azure resources,...SuggestedWork at officeRemote workMonday to FridayFlexible hoursRotating shiftDay shift
- ...realize their greatest potential. Title and Summary Lead Site Reliability Engineer Overview: The role of Business Operations Organization... ...and upfront in the development process, and to proactively manage production and change activities to maximize customer...Full timeWorldwideShift work
- ...Reforma Latino (97001), Mexico, Ciudad de Mexico, Ciudad de Mexico Lead Site Reliability Engineer We're building a Site Reliability Engineering center in Mexico City, and we're hiring a Manager-level Backend Engineer to own the reliability and operational maturity...InternshipLocal area
- ...Our trading platform powers every customer interaction, making reliability a first-class product concern. You will be responsible for... ...observable, and resilient. You'll collaborate closely with software engineers to improve monitoring, deployment safety, automation, fault...Full time
- ...greatest potential. Title and Summary Business Operations Site Reliability Engineer Overview: The role of Business Operations... ...and upfront in the development process, and to proactively manage production and change activities to maximize customer experience...Full timeWorldwideShift work
- ...About the Role: We are looking for an Engineering Manager to lead and grow an existing RPA team while contributing directly as a hands-on backend engineer at an AI-native healthcare platform. This is a dual role where you will manage people and write production-quality...Full timeRemote work
- ...JOB SUMMARY Manages all engineering/maintenance operations, including maintaining the building, grounds and physical plant with particular attention towards safety, security and asset protection. Accountable for managing the budget, capital expenditure projects,...Full timeFor contractorsWork at office
- ...Title and Summary Director, Software Engineering Overview The CNPF Data & AI organization... ...AI and agentic concepts into secure, reliable, observable, and production-grade... ...grade agentic systems, including context management, memory, tool integration, workflow orchestration...Full timeWorldwide
- ...de Mexico Senior Director, Software Engineering Capital One is seeking an experienced... ...help us build and grow our Technology Site in Mexico City. Based in Mexico City, the... ...native architectures; who focus on well managed experiences to empower leaders and application...Local areaShift work
- ...Ciudad de Mexico Director, Software Engineering Capital One is seeking a Director of Software Engineering to lead, manage, mentor, and build extremely talented software... ...~7+ years of experience with Site Reliability Engineering (SRE) At Capital One,...Local area
- ...City, Mex, Mexico, Ciudad de Mexico, Ciudad de Mexico Sr. Manager Software Engineering IC Do you love building and pioneering in the... ...educational tools or other information available through this site. Capital One Financial is made up of several different...Local area
$15 per hour
...Summary The Wikimedia Foundation is seeking a Senior Software Engineer to join the team supporting the Wikidata Platform — the... ...-scale, production-grade services while ensuring performance, reliability, and maintainability. Working closely with the technical and product...Full timeRemote work- ...potential. Title and Summary Software Engineer II Overview The CNPF Data & AI organization... ...emerging technologies into secure, reliable, and reusable capabilities that create... ...native development using Kubernetes and managed cloud platforms such as AWS or Azure •...Full timeWorldwide
- ...potential. Title and Summary Senior Software Engineer Overview: The Mastercard... ...lifecycle, with a focus on usability and reliability. Contribute to the governance of the... ...— supporting version lifecycle management, tool evaluations, and deprecation of redundant...Full timeWorldwide
- ...Mexico, Ciudad de Mexico Senior Software Engineer - Full Stack Do you love building... ...community Collaborate with digital product managers, and deliver robust cloud-based solutions... ...other information available through this site. Capital One Financial is made up of...InternshipLocal area
- ...currently seeking a SAP Operations Project Manager to join our team in Guadalajara, Jalisco... ..., Computer Science, Technical Engineering, Economics or related field Nice to Have... ...hire locally to NTT DATA offices or client sites. This ensures we can provide timely and...Contract workWork at officeRemote workFlexible hours
- ...EdTech company to find a Senior AI Systems Engineer who will play a key role in designing,... ...data, and developing scalable, reliable systems with a strong focus on quality,... ...continuously improve AI performance. Manage and integrate structured sports data from...Contract workPart timeRemote workMonday to Friday
- ...Mex, Mexico, Ciudad de Mexico, Ciudad de Mexico Senior Manager, Software Engineering (People Leader) Do you love building and pioneering in... ...educational tools or other information available through this site. Capital One Financial is made up of several...InternshipLocal area
- ...potential. Title and Summary Lead Software Engineer Overview The CNPF Data & AI... ...design and delivery of secure, scalable, and reliable agentic applications that can reason,... ...native environments using Kubernetes and managed cloud services on AWS, Azure, or GCP •...Full timeTemporary workWorldwide
- ...part of NTT Group, which invests over $3 billion each year in R&D. Whenever possible, we hire locally to NTT DATA offices or client sites. This ensures we can provide timely and effective support tailored to each client’s needs. While many positions offer remote or...Work at officeRemote workFlexible hours
- ...development, implementation and management of applications,... ...to NTT DATA offices or client sites. This ensures we can provide... ...looking for a Senior DevOps Engineer with strong experience in infrastructure... ..., security, and application reliability. The candidate should be...Work at officeRemote workFlexible hours
- ...About the Team: We’re a small DevOps team supporting the whole Engineering organization, building applications on top of EC2 and AWS. We... ...You are an experienced cloud engineer with at least 5+ years managing Amazon Web Services (AWS) including services such as EC2,...Full timeRemote work
- ...are looking for a skilled and technically driven Senior Software Engineer (Python) to join a fast-paced cybersecurity environment. You... ...Proficiency with generators, iterators, decorators, and context managers ~ Solid understanding of the Global Interpreter Lock (GIL) and...Full timeRemote workFlexible hours
- ...Title and Summary Principal Software Engineer, SE Guild Mastercard is a global technology... ...that enable SE Guild programs to scale reliably and deliver value across the engineering... ...SE Guild productivity (e.g., roster management, document publishing, reporting). Design...Full timeWorldwide
- ...administration, monitoring, configurations, apply patches, change management, security and compliance activities ~ Test and deploy... ...Whenever possible, we hire locally to NTT DATA offices or client sites. This ensures we can provide timely and effective support tailored...Work at officeRemote workFlexible hours
- ...performance, scalability, and reliability Write clean, maintainable... ...architecture, and state management Backend Expertise in... ...with ETL pipelines or data engineering concepts Who You Are Energetic... ...NTT DATA offices or client sites. This ensures we can provide...Work at officeRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineering Manager. Be the first to apply!
- site reliability engineer Mexico
- junior website developer Mexico
- on-site clinical research associate (traveling/remote) Mexico
- site reliability engineer sre
- site reliability engineering manager
- junior site reliability engineer
- site reliability engineer
- site reliability engineer remote
- lead site reliability engineer
- senior site operations manager






