Site Reliability Engineer II
MasterCard
Site Reliability Engineer II
Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart, and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.
About the Role
The Business Operations team is seeking a highly motivated and experienced Site Reliability Engineer II (SRE) to join our team. You will play a critical role in ensuring the reliability, scalability, and performance of our applications, supporting essential services that power Mastercard's global operations. As a thought leader in your field, you will bring technical expertise, a passion for automation, and the ability to mentor.
The role of the Business Operations Site Reliability Engineer is to be the production readiness steward for Mastercard products. As Business Operations SRE, we are responsible for ensuring that our platform is stable and healthy. We break down barriers to running our products by fostering developer run ownership and empowering developers to build resilient products. We support our developers during the application build phase in software run principles that include operational design, automation, capacity planning, and monitoring that leads to fault-tolerant, scalable products. We see the big picture and help create and enforce operations standards while facilitating an agile and learning culture. We support daily operations with a hyper focus on triage, root cause by understanding the business impact of our products and subsequently performing blameless post-mortems. The goal of every Business Operations team is to engage early in the development lifecycle to be more proactive and upfront in the development process, and to proactively manage production and change activities to maximize customer experience and increase the overall value of supported applications. Business Operations teams also focus on risk management by tying all our activities together with an overarching responsibility for compliance and risk mitigation across all our environments. Ultimately, the role of Business Operations is to align Product and Customer Focused priorities with Operational needs by providing continuous feedback throughout the lifecycle. As part of the Business Operations team, you will:
- Work independently on elements of projects/processes within the Site Reliability Engineering area by applying intermediate/practical knowledge and area best practices to meet organizational standards of quality and excellence.
- Support the implementation and maintenance of high-availability systems to ensure operational stability.
- Assist in evaluating operational needs and developing technical solutions under guidance.
- Contribute to automation and scripting projects to streamline routine operational tasks.
- Troubleshoot and resolve basic to moderate system issues, escalating more complex problems as needed.
- Document operational procedures and shares knowledge with team members.
- Participate in quality checks and reviews to ensure system stability and reliability.
- Utilize experience and a comprehensive understanding of area processes and tools to make minor adjustments or enhancements to resolve identifiable issues. May manage smaller project/initiatives as an experienced individual contributor with specialized knowledge within the Site Reliability Engineering area.
Role qualifications:
The ideal candidate will apply the following skills independently in routine and moderately complex situations, requiring occasional guidance typically only in unfamiliar or highly complex scenarios. They will demonstrate growing consistency and reliability in applying the skills.
- Observability - Ability to use scripting and tooling to implement observability solutions, enabling the collection, analysis, and visualization of metrics, logs, and traces to support incident detection, diagnosis, and continuous service improvement.
- Programming and Scripting - Ability to write and maintain code and scripts to automate tasks, build operational tools, and support monitoring, deployment, and incident response using languages such as Python, Go, Bash, or similar.
- Systems and Network Administration - Ability to configure, operate, and troubleshoot Linux/Unix systems and network components, applying knowledge of networking concepts, protocols, security, and system reliability.
- Cloud Computing and Infrastructure - Ability to design, deploy, and manage applications and infrastructure on cloud platforms (e.g., AWS, Azure, GCP), ensuring scalability, security, availability, and operational efficiency.
- Reliability and Scalability - Ability to design and operate systems for high availability, fault tolerance, and disaster recovery, while ensuring systems can scale to meet current and future demand.
- DevOps Practices - Ability to apply DevOps principles and practices, including CI/CD pipelines, containerization, and orchestration, to enable faster, more reliable software delivery and operations.
- Troubleshooting - Capability to systematically identify, diagnose, and resolve technical issues across systems, applications, and networks, using analytical methods and tools to restore functionality, minimize disruption, and ensure stable operations.
- Capacity Planning and Performance Optimization - Ability to monitor resource utilization, forecast future capacity needs, and optimize system performance to support growth, scalability, and efficient infrastructure usage.
- IT Service Management - Ability to apply IT service management principles to incident, problem, and change management, ensuring reliable service delivery, effective incident response, and continuous service improvement aligned to business needs.
- Proactive Monitoring and Improvement (SRE Applications) - The ability to use application reliability signals to anticipate issues, identify risks, and drive preventative improvements that enhance application performance and availability.
Corporate Security Responsibility
All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
- Abide by Mastercard's security policies and practices;
- Ensure the confidentiality and integrity of the information being accessed;
- Report any suspected information security violation or breach, and
- Complete all periodic mandatory security trainings in accordance with Mastercard's guidelines.
$113k - $171.6k
...growing rapidly and hiring top talent with leading AI skills across engineering, sales, product, marketing, and beyond as we build the leading digital operations platform.As a Site Reliability Engineer II on the Core Infrastructure team in our Atlanta office,you'll help...SuggestedWork at officeLocal areaFlexible hours- Play a key role in ensuring system reliability at one of the world’s most iconic and largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will use technology to...Suggested
$135k - $155k
...also making it easy for buyers at Fortune 1000 companies to tap into global manufacturing capacity.Xometry is seeking a Site Reliability Engineer II to join our Site Reliability Engineering (SRE) Organization. In this role as an individual contributor, you will guide the...SuggestedFlexible hours- ...big ideas, and your desire to team up with some of the best and brightest in technology and entertainment. The RoleThe Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus on...SuggestedFull timeLocal areaWorldwideFlexible hours
- ...India; Chennai, TN, IndiaJob Function: Engineering & ArchitectureSchedule: Full timeShift... ...: American ExpressDescriptionSite Reliability Engineer II collaborates with engineering teams to... ...analysis platformsKnowledge of cloud-based Site Reliability Engineering (SRE)...Suggested
$98.58k - $138.02k
...Northern California / Silicon Valley Region / Denver, COProduct Engineering - DevOps /Full Time /HybridRestaurant365 is a SaaS... ...office locations: Austin, TX; Irvine, CA; or Akron, OH. The Site Reliability Engineer II will be responsible for supporting, enhancing, and...Full timeWork at office- Play a key role in ensuring system reliability at one of the world’s most iconic and largest financial institutions. As a Site Reliability Engineer II at JPMorgan Chase within the within the Corporate and Investment Bank, Payments Technology Team, you will use technology...
$76k - $127k
...and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II The BizOps team at Mastercard is looking for a Site Reliability Engineer (SRE) who thrives on solving complex problems...Full timePart timeWorldwideFlexible hoursEarly shift- ...Site Reliability Engineer II Do you like collaborating across teams to solve complex problems? Do you enjoy solving large scale distributed content delivery challenges? Join our highly skilled Site Reliability Team The Platform & Reliability Engineering team...Work at officeRemote work
£62.4k - £93.6k per year
...Site Reliability Engineer II London, UK About Us GoCardless, a Mollie company, is a global leader in bank payments. Over 100,000 businesses, from start-ups to household names, use GoCardless to collect, manage and send bank payments through Direct Debit, real...Remote work$95k - $171k
...Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts,...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$53 - $57 per hour
...A client of Innova Solutions is immediately hiring for the Site Reliability Engineer II Position Type: Full time, Hybrid Onsite Contract --- (***Hybrid Position - 3 days onsite and 2 days remote work in a week!!!) Duration: 12-18 Months contract Location:...Hourly payFull timeContract workTemporary workWork experience placementLocal areaImmediate startRemote workWorldwideFlexible hours- ...Site Reliability Engineer II Overview / Summary We are seeking a Site Reliability Engineer to join an SRE team focused on observability, monitoring, and technical consulting across GCP-based data platforms. This role is responsible for ensuring the availability...
- ...Site Reliability Engineer II At Mastercard technology, we work to connect and power an inclusive, digital economy that benefits everyone, everywhere, by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships, and passion...
- ...Site Reliability Engineer II Play a key role in ensuring system reliability at one of the world's most iconic and largest financial institutions. As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology...Local area
$87.5k - $143.75k
...Observability Engineer DISH is transforming the future of connectivity. We're doing it... ...Personal responsibility for the quality, reliability, and usability of the NOC Observability tools... ...Coordinate training for Tier I and Tier II NOC Engineers Other duties as assigned...Flexible hoursNight shift- ...organization across multiple locations in the US, South America, and India. Location: Remote (US-Based Candidates Only) Site Reliability Engineer II (SRE) Position Overview We are seeking a Site Reliability Engineer II (SRE) to join our growing Site Reliability...Remote workFlexible hoursShift workWeekday work
- ...Play a key role in ensuring system reliability at one of the world's most iconic and largest financial institutions. As a Site Reliability Engineer II at JPMorgan Chase within the within the Corporate and Investment Bank, Payments Technology Team , you will use...
- ...world running. Location: 5 On-Site Days a Week in Sunnyvale, CA Headquarters Our Engineering team is driven by a culture... ...Your Impact As an SRE Engineer II, you will be responsible for... ...will work on enhancing system reliability and scalability of Illumio SaaS...Work experience placementImmediate start
$165k - $195k
...for employees who prefer to work in an office some or all of the time. About Your Role We're looking for a Senior Site Reliability Engineer II to help us scale our infrastructure and reliability practices as we grow our engineering org by 90+ people this year. You...Full timeWork at officeLocal areaRemote workWork from homeFlexible hours$115.5k - $164.8k
...Site Reliability Engineer II Washington, United States Join Axon and be a Force for Good. At Axon, we're on a mission to Protect Life. We're explorers, pursuing society's most critical safety and justice issues with our ecosystem of devices and cloud software...Work experience placementWork at officeRemote work- ...Site Reliability Engineer II This role sits at the intersection of software engineering, infrastructure, and network operations, helping scale and strengthen a globally distributed platform. You'll design and build automation that reduces operational toil and improves...Work at officeRemote workWork from homeFlexible hours
- ...Site Reliability Engineer II Do you like collaborating across teams to solve complex problems? Do you enjoy solving large scale distributed content delivery challenges? Join our critical Platform and Reliability Engineering Team! The Platform & Reliability Engineering...Work at officeRemote work
$103.5k - $150k
...Site Reliability Engineering II Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS platform, Medallia Experience Cloud, leads the market in the management of experiences, insights, and actions for candidates, customers, employees...Temporary workLocal area3 days per week$95k - $171k
...professional who thrives in a dynamic environment? Join our Site Reliability team Our Team builds and delivers highly secure network... ...the Akamai Zero Trust platform. As a Site Reliability Engineer II, you will be responsible for: Ensuring the reliability...Work experience placementWork at officeRemote work- ...Site Reliability Engineer II Remote - Bangalore Backblaze is the object storage leader in the open cloud movement, fueling customer success with cloud storage built purposefully to unlock budgets, unburden administrators, and unleash innovators. Together with our...Remote work
- ...Site Reliability Engineer II All IT Solutions United States · Chandler, Arizona Workplace Type — Remote Employment Type — Contract We are currently seeking a qualified Site Reliability Engineer II to support this engagement. Please review the complete requirements...Contract workTemporary workLocal areaRemote work3 days per week
- Position Summary Applied AI Site Reliability Engineer II Role Overview: As an Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity of high...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week
$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...Full timeWork at officeLocal areaRemote workWork from home$114.3k - $235.32k
...verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business.We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven...Work at officeLocal areaRelocationRelocation package
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer II. Be the first to apply!
- site reliability engineer sre United States
- site reliability engineer United States
- site reliability engineering manager United States
- site reliability engineer remote United States
- lead site reliability engineer United States
- IT site lead United States
- site safety United States
- site merchandiser United States
- website content developer United States
- site leader United States



