Site Reliability Engineer II
MasterCard
Site Reliability Engineer II
Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart, and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.
About the Role
The Business Operations team is seeking a highly motivated and experienced Site Reliability Engineer II (SRE) to join our team. You will play a critical role in ensuring the reliability, scalability, and performance of our applications, supporting essential services that power Mastercard's global operations. As a thought leader in your field, you will bring technical expertise, a passion for automation, and the ability to mentor.
The role of the Business Operations Site Reliability Engineer is to be the production readiness steward for Mastercard products. As Business Operations SRE, we are responsible for ensuring that our platform is stable and healthy. We break down barriers to running our products by fostering developer run ownership and empowering developers to build resilient products. We support our developers during the application build phase in software run principles that include operational design, automation, capacity planning, and monitoring that leads to fault-tolerant, scalable products. We see the big picture and help create and enforce operations standards while facilitating an agile and learning culture. We support daily operations with a hyper focus on triage, root cause by understanding the business impact of our products and subsequently performing blameless post-mortems. The goal of every Business Operations team is to engage early in the development lifecycle to be more proactive and upfront in the development process, and to proactively manage production and change activities to maximize customer experience and increase the overall value of supported applications. Business Operations teams also focus on risk management by tying all our activities together with an overarching responsibility for compliance and risk mitigation across all our environments. Ultimately, the role of Business Operations is to align Product and Customer Focused priorities with Operational needs by providing continuous feedback throughout the lifecycle. As part of the Business Operations team, you will:
- Work independently on elements of projects/processes within the Site Reliability Engineering area by applying intermediate/practical knowledge and area best practices to meet organizational standards of quality and excellence.
- Support the implementation and maintenance of high-availability systems to ensure operational stability.
- Assist in evaluating operational needs and developing technical solutions under guidance.
- Contribute to automation and scripting projects to streamline routine operational tasks.
- Troubleshoot and resolve basic to moderate system issues, escalating more complex problems as needed.
- Document operational procedures and shares knowledge with team members.
- Participate in quality checks and reviews to ensure system stability and reliability.
- Utilize experience and a comprehensive understanding of area processes and tools to make minor adjustments or enhancements to resolve identifiable issues. May manage smaller project/initiatives as an experienced individual contributor with specialized knowledge within the Site Reliability Engineering area.
Role qualifications:
The ideal candidate will apply the following skills independently in routine and moderately complex situations, requiring occasional guidance typically only in unfamiliar or highly complex scenarios. They will demonstrate growing consistency and reliability in applying the skills.
- Observability - Ability to use scripting and tooling to implement observability solutions, enabling the collection, analysis, and visualization of metrics, logs, and traces to support incident detection, diagnosis, and continuous service improvement.
- Programming and Scripting - Ability to write and maintain code and scripts to automate tasks, build operational tools, and support monitoring, deployment, and incident response using languages such as Python, Go, Bash, or similar.
- Systems and Network Administration - Ability to configure, operate, and troubleshoot Linux/Unix systems and network components, applying knowledge of networking concepts, protocols, security, and system reliability.
- Cloud Computing and Infrastructure - Ability to design, deploy, and manage applications and infrastructure on cloud platforms (e.g., AWS, Azure, GCP), ensuring scalability, security, availability, and operational efficiency.
- Reliability and Scalability - Ability to design and operate systems for high availability, fault tolerance, and disaster recovery, while ensuring systems can scale to meet current and future demand.
- DevOps Practices - Ability to apply DevOps principles and practices, including CI/CD pipelines, containerization, and orchestration, to enable faster, more reliable software delivery and operations.
- Troubleshooting - Capability to systematically identify, diagnose, and resolve technical issues across systems, applications, and networks, using analytical methods and tools to restore functionality, minimize disruption, and ensure stable operations.
- Capacity Planning and Performance Optimization - Ability to monitor resource utilization, forecast future capacity needs, and optimize system performance to support growth, scalability, and efficient infrastructure usage.
- IT Service Management - Ability to apply IT service management principles to incident, problem, and change management, ensuring reliable service delivery, effective incident response, and continuous service improvement aligned to business needs.
- Proactive Monitoring and Improvement (SRE Applications) - The ability to use application reliability signals to anticipate issues, identify risks, and drive preventative improvements that enhance application performance and availability.
Corporate Security Responsibility
All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
- Abide by Mastercard's security policies and practices;
- Ensure the confidentiality and integrity of the information being accessed;
- Report any suspected information security violation or breach, and
- Complete all periodic mandatory security trainings in accordance with Mastercard's guidelines.
$76k - $127k
...and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II The BizOps team at Mastercard is looking for a Site Reliability Engineer (SRE) who thrives on solving complex problems...SuggestedFull timePart timeWorldwideFlexible hoursEarly shift$76k - $127k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...SuggestedFull timePart timeWorldwideFlexible hours$98.58k - $138.02k
...Northern California / Silicon Valley Region / Denver, COProduct Engineering - DevOps /Full Time /HybridRestaurant365 is a SaaS... ...office locations: Austin, TX; Irvine, CA; or Akron, OH. The Site Reliability Engineer II will be responsible for supporting, enhancing, and...SuggestedFull timeWork at office$102.1k - $202.2k
...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole... ...EngineeringDiscipline: Site Reliability EngineeringCompany:... ...further than the Microsoft Defender engineering team. We are looking for a Site Reliability Engineer II who will be building and delivering...SuggestedOngoing contractLocal area3 days per week$102.1k - $202.2k
...per yearEmployment type: Full-TimeWork site: 0 days / week in-office - remoteRole type... ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewMicrosoft... ...workloads. As a Site Reliability Engineer II, you will take ownership of reliability...SuggestedOngoing contractWork at officeLocal area- ...big ideas, and your desire to team up with some of the best and brightest in technology and entertainment. The RoleThe Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus on...Full timeLocal areaWorldwideFlexible hours
- Play a key role in ensuring system reliability at one of the world’s most iconic and largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will use technology to...
$113k - $171.6k
...Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization.As a Site Reliability Engineer II on the Core Infrastructure team in our Atlanta office,you'll help build and operate the foundational infrastructure that...Work at officeLocal areaFlexible hours$135k - $165k
...also making it easy for buyers at Fortune 1000 companies to tap into global manufacturing capacity.Xometry is seeking a Site Reliability Engineer II to join our Site Reliability Engineering (SRE) Organization. In this role as an individual contributor, you will guide the...Flexible hours- ...Site Reliability Engineer II Remote - Bangalore Backblaze is the object storage leader in the open cloud movement, fueling customer success with cloud storage built purposefully to unlock budgets, unburden administrators, and unleash innovators. Together with our...Remote work
$95k - $171k
...Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$103.5k - $150k
...experiences together. Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure... ...power a highly reliable global SaaS platform. As an SRE II, you will help operate and improve the reliability,...Temporary workWork experience placementLocal area3 days per week$165k - $195k
...for employees who prefer to work in an office some or all of the time. About Your Role We're looking for a Senior Site Reliability Engineer II to help us scale our infrastructure and reliability practices as we grow our engineering org by 90+ people this year. You...Full timeWork at officeLocal areaRemote workWork from homeFlexible hours- ...Site Reliability Engineer II (SRE) NationsBenefits is recognized as one of the fastest-growing companies in America and a Healthcare Fintech provider of supplemental benefits, flex cards, and member engagement solutions. We partner with managed care organizations to...Remote workFlexible hoursShift workWeekday work
- ..., Hope all is well, Please find the job description given below and let me know your interest. Position: Site Reliability Engineer II Location: Pennington, NJ | Onsite (2 Virtual Rounds | Onsite Interview May Be Requested) Contract: 12+ Months Contract...Contract work
$138.24k - $171k
Mon, 08/31/2026 - 04:40 Job Title: Site Reliability Engineer II Work Location: 145 Broadway, Cambridge, MA 02142 Job Description: Akamai Technologies, Inc. is hiring for the following role in Cambridge, MA (multiple openings): Site Reliability Engineer II. Perform...Work experience placementWork at officeRemote work$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week$104.9k - $174.7k
Are you passionate about improving reliability, scalability, and resilience in complex database... ....Own prioritization of reliability engineering tasks within team backlogs.Lead incident... ...a Service (IaaS).Background in DevOps, site reliability engineering practices, or related...Full timeLocal area$102.1k - $202.2k
...per yearEmployment type: Full-TimeWork site: Fully on-siteRole type: Individual ContributorTravel... ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:... ...at the intersection of large-scale cloud engineering, service reliability, and operational excellence...Ongoing contractLocal areaWorldwide$145.7k - $218.5k
...synonymous with entertainment excellence and creativity.Service Reliability EngineerDo you want to use transformative technologies to... ...scalability and efficiency? Do you want a career that combines your engineering skills and your passion for video gaming? Are you fascinated...Work experience placementShift work- Site Reliability Engineer IIJob#: 3049111Job Description:Site Reliability Engineer IILocation: Plano, Texas (Onsite)Role OverviewWe are seeking a Site Reliability Engineer (SRE) to join a newly forming team. This is an opportunity to establish SRE practices from the ground...
$115.5k - $164.8k
...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant... ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations...Work experience placementWork at officeRemote work- ...high traffic, business critical internet site communications and/or network-based (... ...teams to ensure software is designed for reliability, scalability, and operational efficiency... ...Bachelor's degree in Computer Science, Engineering, or a related technical field, or equivalent...Full timeLive outLocal areaFlexible hours
$114.3k - $235.32k
...verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business.We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven...Work at officeLocal areaRelocationRelocation package$104.9k - $174.7k
About the role:A FinOps Site Reliability Engineer (SRE) bridges the gap between engineering, operations, and financial governance by embedding cost optimization into infrastructure design, automation, monitoring, and operational processes. A FinOps SRE proactively identifies...Full timeLocal area- ...generative AI and cloud-native platforms to advanced release engineering practices, our teams are redefining how financial technology... ...AI-driven solutions that accelerate development and improve reliability. Your work will directly influence how GM Financial leverages...H1bWork at officeRemote workVisa sponsorshipFlexible hours2 days per week
- ...Site Reliability Engineer (SRE) Step into the world of Mrsool where convenience meets innovation! As one of the largest delivery platforms in the Middle East and North Africa (MENA) region, Mrsool has captivated users with its unique and seamless experience, earning...Remote workNight shift
$71.6k - $119.4k
...deployment support, and security improvements. You'll help implement automation, troubleshoot issues, and work closely with senior engineers to learn and apply best practices. You'll gain exposure to a wide range of cloud technologies, automation tools, and data...Temporary workInternshipLocal areaRemote work- ...Eyes on glass. Hands on the pipeline. Real ownership from day one. This isn't a watch-and-wait monitoring seat. Our client needs engineers who can read a Kibana query at 3am, know the difference between a blip and a breach, and act on it, on a FedRAMP-authorised cloud...Hourly payFor contractorsShift workNight shiftWeekend work
$95.6k - $119.5k
...Honda’s, we want you to join our team to Bring the Future! Job Purpose The Digital Product Management (DPM) Development Systems Engineer II is responsible for projects, system enhancements and portfolio support of the Honda DPM systems. This includes tools, methods and...Full timeTemporary workWork experience placementRemote workRelocation package
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer II. Be the first to apply!
- site reliability engineer United States
- site reliability engineering manager United States
- site reliability engineer sre United States
- site reliability engineer remote United States
- lead site reliability engineer United States
- website coordinator United States
- on-site clinical research associate (traveling/remote) United States
- site safety United States
- after school site coordinator United States
- junior website developer United States




