Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer II

MasterCard

Site Reliability Engineer II

Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart, and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.

About the Role

The Business Operations team is seeking a highly motivated and experienced Site Reliability Engineer II (SRE) to join our team. You will play a critical role in ensuring the reliability, scalability, and performance of our applications, supporting essential services that power Mastercard's global operations. As a thought leader in your field, you will bring technical expertise, a passion for automation, and the ability to mentor.

The role of the Business Operations Site Reliability Engineer is to be the production readiness steward for Mastercard products. As Business Operations SRE, we are responsible for ensuring that our platform is stable and healthy. We break down barriers to running our products by fostering developer run ownership and empowering developers to build resilient products. We support our developers during the application build phase in software run principles that include operational design, automation, capacity planning, and monitoring that leads to fault-tolerant, scalable products. We see the big picture and help create and enforce operations standards while facilitating an agile and learning culture. We support daily operations with a hyper focus on triage, root cause by understanding the business impact of our products and subsequently performing blameless post-mortems. The goal of every Business Operations team is to engage early in the development lifecycle to be more proactive and upfront in the development process, and to proactively manage production and change activities to maximize customer experience and increase the overall value of supported applications. Business Operations teams also focus on risk management by tying all our activities together with an overarching responsibility for compliance and risk mitigation across all our environments. Ultimately, the role of Business Operations is to align Product and Customer Focused priorities with Operational needs by providing continuous feedback throughout the lifecycle. As part of the Business Operations team, you will:

  • Work independently on elements of projects/processes within the Site Reliability Engineering area by applying intermediate/practical knowledge and area best practices to meet organizational standards of quality and excellence.
  • Support the implementation and maintenance of high-availability systems to ensure operational stability.
  • Assist in evaluating operational needs and developing technical solutions under guidance.
  • Contribute to automation and scripting projects to streamline routine operational tasks.
  • Troubleshoot and resolve basic to moderate system issues, escalating more complex problems as needed.
  • Document operational procedures and shares knowledge with team members.
  • Participate in quality checks and reviews to ensure system stability and reliability.
  • Utilize experience and a comprehensive understanding of area processes and tools to make minor adjustments or enhancements to resolve identifiable issues. May manage smaller project/initiatives as an experienced individual contributor with specialized knowledge within the Site Reliability Engineering area.

Role qualifications:

The ideal candidate will apply the following skills independently in routine and moderately complex situations, requiring occasional guidance typically only in unfamiliar or highly complex scenarios. They will demonstrate growing consistency and reliability in applying the skills.

  • Observability - Ability to use scripting and tooling to implement observability solutions, enabling the collection, analysis, and visualization of metrics, logs, and traces to support incident detection, diagnosis, and continuous service improvement.
  • Programming and Scripting - Ability to write and maintain code and scripts to automate tasks, build operational tools, and support monitoring, deployment, and incident response using languages such as Python, Go, Bash, or similar.
  • Systems and Network Administration - Ability to configure, operate, and troubleshoot Linux/Unix systems and network components, applying knowledge of networking concepts, protocols, security, and system reliability.
  • Cloud Computing and Infrastructure - Ability to design, deploy, and manage applications and infrastructure on cloud platforms (e.g., AWS, Azure, GCP), ensuring scalability, security, availability, and operational efficiency.
  • Reliability and Scalability - Ability to design and operate systems for high availability, fault tolerance, and disaster recovery, while ensuring systems can scale to meet current and future demand.
  • DevOps Practices - Ability to apply DevOps principles and practices, including CI/CD pipelines, containerization, and orchestration, to enable faster, more reliable software delivery and operations.
  • Troubleshooting - Capability to systematically identify, diagnose, and resolve technical issues across systems, applications, and networks, using analytical methods and tools to restore functionality, minimize disruption, and ensure stable operations.
  • Capacity Planning and Performance Optimization - Ability to monitor resource utilization, forecast future capacity needs, and optimize system performance to support growth, scalability, and efficient infrastructure usage.
  • IT Service Management - Ability to apply IT service management principles to incident, problem, and change management, ensuring reliable service delivery, effective incident response, and continuous service improvement aligned to business needs.
  • Proactive Monitoring and Improvement (SRE Applications) - The ability to use application reliability signals to anticipate issues, identify risks, and drive preventative improvements that enhance application performance and availability.

Corporate Security Responsibility

All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:

  • Abide by Mastercard's security policies and practices;
  • Ensure the confidentiality and integrity of the information being accessed;
  • Report any suspected information security violation or breach, and
  • Complete all periodic mandatory security trainings in accordance with Mastercard's guidelines.
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer II in United States vacancy
  • $76k - $127k

     ...and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II The BizOps team at Mastercard is looking for a Site Reliability Engineer (SRE) who thrives on solving complex problems... 
    Suggested
    Full time
    Part time
    Worldwide
    Flexible hours
    Early shift

    Mastercard

    O Fallon, MO
    18 hours ago
  • $76k - $127k

     ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that... 
    Suggested
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    18 hours ago
  • $98.58k - $138.02k

     ...Northern California / Silicon Valley Region / Denver, COProduct Engineering - DevOps /Full Time /HybridRestaurant365 is a SaaS...  ...office locations: Austin, TX; Irvine, CA; or Akron, OH. The Site Reliability Engineer II will be responsible for supporting, enhancing, and... 
    Suggested
    Full time
    Work at office

    Restaurant 365

    Irvine, CA
    4 days ago
  • $102.1k - $202.2k

     ...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole...  ...EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...further than the Microsoft Defender engineering team. We are looking for a Site Reliability Engineer II who will be building and delivering... 
    Suggested
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Redmond, WA
    2 days ago
  • $102.1k - $202.2k

     ...per yearEmployment type: Full-TimeWork site: 0 days / week in-office - remoteRole type...  ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewMicrosoft...  ...workloads. As a Site Reliability Engineer II, you will take ownership of reliability... 
    Suggested
    Ongoing contract
    Work at office
    Local area

    Microsoft

    Redmond, WA
    4 days ago
  •  ...big ideas, and your desire to team up with some of the best and brightest in technology and entertainment. The RoleThe Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus on... 
    Full time
    Local area
    Worldwide
    Flexible hours

    AXS Group

    Los Angeles, CA
    18 hours ago
  • Play a key role in ensuring system reliability at one of the world’s most iconic and largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will use technology to... 

    JP Morgan Chase

    Chicago, IL
    1 day ago
  • $113k - $171.6k

     ...Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization.As a Site Reliability Engineer II on the Core Infrastructure team in our Atlanta office,you'll help build and operate the foundational infrastructure that... 
    Work at office
    Local area
    Flexible hours

    PagerDuty

    Atlanta, GA
    16 hours ago
  • $135k - $165k

     ...also making it easy for buyers at Fortune 1000 companies to tap into global manufacturing capacity.Xometry is seeking a Site Reliability Engineer II to join our Site Reliability Engineering (SRE) Organization. In this role as an individual contributor, you will guide the... 
    Flexible hours

    Xometry

    Boston, MA
    16 hours ago
  •  ...Site Reliability Engineer II Remote - Bangalore Backblaze is the object storage leader in the open cloud movement, fueling customer success with cloud storage built purposefully to unlock budgets, unburden administrators, and unleash innovators. Together with our... 
    Remote work

    Backblaze

    United States
    2 days ago
  • $95k - $171k

     ...Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    United States
    3 days ago
  • $103.5k - $150k

     ...experiences together. Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure...  ...power a highly reliable global SaaS platform. As an SRE II, you will help operate and improve the reliability,... 
    Temporary work
    Work experience placement
    Local area
    3 days per week

    Medallia

    McLean, VA
    3 days ago
  • $165k - $195k

     ...for employees who prefer to work in an office some or all of the time. About Your Role We're looking for a Senior Site Reliability Engineer II to help us scale our infrastructure and reliability practices as we grow our engineering org by 90+ people this year. You... 
    Full time
    Work at office
    Local area
    Remote work
    Work from home
    Flexible hours

    Juniper Square

    United States
    18 hours ago
  •  ...Site Reliability Engineer II (SRE) NationsBenefits is recognized as one of the fastest-growing companies in America and a Healthcare Fintech provider of supplemental benefits, flex cards, and member engagement solutions. We partner with managed care organizations to... 
    Remote work
    Flexible hours
    Shift work
    Weekday work

    NationsBenefits

    United States
    1 day ago
  •  ..., Hope all is well, Please find the job description given below and let me know your interest. Position: Site Reliability Engineer II Location: Pennington, NJ | Onsite (2 Virtual Rounds | Onsite Interview May Be Requested) Contract: 12+ Months Contract... 
    Contract work

    DMS Vision Inc

    Pennington, NJ
    3 days ago
  • $138.24k - $171k

    Mon, 08/31/2026 - 04:40 Job Title: Site Reliability Engineer II Work Location: 145 Broadway, Cambridge, MA 02142 Job Description: Akamai Technologies, Inc. is hiring for the following role in Cambridge, MA (multiple openings): Site Reliability Engineer II. Perform... 
    Work experience placement
    Work at office
    Remote work

    Akamai Technologies

    United States
    3 days ago
  • $113.1k - $232.3k

    Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity... 
    Work at office
    Local area
    Visa sponsorship
    Flexible hours
    3 days per week

    Deloitte

    New York, NY
    18 hours ago
  • $104.9k - $174.7k

    Are you passionate about improving reliability, scalability, and resilience in complex database...  ....Own prioritization of reliability engineering tasks within team backlogs.Lead incident...  ...a Service (IaaS).Background in DevOps, site reliability engineering practices, or related... 
    Full time
    Local area

    RELX Group

    Georgia
    2 days ago
  • $102.1k - $202.2k

     ...per yearEmployment type: Full-TimeWork site: Fully on-siteRole type: Individual ContributorTravel...  ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...at the intersection of large-scale cloud engineering, service reliability, and operational excellence... 
    Ongoing contract
    Local area
    Worldwide

    Microsoft

    Reston, VA
    4 days ago
  • $145.7k - $218.5k

     ...synonymous with entertainment excellence and creativity.Service Reliability EngineerDo you want to use transformative technologies to...  ...scalability and efficiency? Do you want a career that combines your engineering skills and your passion for video gaming? Are you fascinated... 
    Work experience placement
    Shift work

    Sony Interactive Entertainment America

    Aliso Viejo, CA
    1 day ago
  • Site Reliability Engineer IIJob#: 3049111Job Description:Site Reliability Engineer IILocation: Plano, Texas (Onsite)Role OverviewWe are seeking a Site Reliability Engineer (SRE) to join a newly forming team. This is an opportunity to establish SRE practices from the ground... 

    Apex Systems

    Plano, TX
    1 day ago
  • $115.5k - $164.8k

     ...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant...  ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations... 
    Work experience placement
    Work at office
    Remote work

    Axon

    Washington DC
    2 days ago
  •  ...high traffic, business critical internet site communications and/or network-based (...  ...teams to ensure software is designed for reliability, scalability, and operational efficiency...  ...Bachelor's degree in Computer Science, Engineering, or a related technical field, or equivalent... 
    Full time
    Live out
    Local area
    Flexible hours

    Waystar

    Louisville, KY
    4 days ago
  • $114.3k - $235.32k

     ...verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business.We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven... 
    Work at office
    Local area
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    1 day ago
  • $104.9k - $174.7k

    About the role:A FinOps Site Reliability Engineer (SRE) bridges the gap between engineering, operations, and financial governance by embedding cost optimization into infrastructure design, automation, monitoring, and operational processes. A FinOps SRE proactively identifies... 
    Full time
    Local area

    LexisNexis Risk Solutions Group

    Boca Raton, FL
    16 hours ago
  •  ...generative AI and cloud-native platforms to advanced release engineering practices, our teams are redefining how financial technology...  ...AI-driven solutions that accelerate development and improve reliability. Your work will directly influence how GM Financial leverages... 
    H1b
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours
    2 days per week

    GM Financial

    Arlington, TX
    18 hours ago
  •  ...Site Reliability Engineer (SRE) Step into the world of Mrsool where convenience meets innovation! As one of the largest delivery platforms in the Middle East and North Africa (MENA) region, Mrsool has captivated users with its unique and seamless experience, earning... 
    Remote work
    Night shift

    Mrsool

    United States
    1 day ago
  • $71.6k - $119.4k

     ...deployment support, and security improvements. You'll help implement automation, troubleshoot issues, and work closely with senior engineers to learn and apply best practices. You'll gain exposure to a wide range of cloud technologies, automation tools, and data... 
    Temporary work
    Internship
    Local area
    Remote work

    RELX

    United States
    18 hours ago
  •  ...Eyes on glass. Hands on the pipeline. Real ownership from day one. This isn't a watch-and-wait monitoring seat. Our client needs engineers who can read a Kibana query at 3am, know the difference between a blip and a breach, and act on it, on a FedRAMP-authorised cloud... 
    Hourly pay
    For contractors
    Shift work
    Night shift
    Weekend work

    C-Serv

    Reston, VA
    5 days ago
  • $95.6k - $119.5k

     ...Honda’s, we want you to join our team to Bring the Future! Job Purpose The Digital Product Management (DPM) Development Systems Engineer II is responsible for projects, system enhancements and portfolio support of the Honda DPM systems. This includes tools, methods and... 
    Full time
    Temporary work
    Work experience placement
    Remote work
    Relocation package

    American Honda Motor Co., Inc.

    Raymond, OH
    a month ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer II. Be the first to apply!