Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

Sphere

Site Reliability Engineer

Our client is a fast-growing global fintech company building a modern digital payments platform that enables fast, secure, and compliant international money transfers. Serving millions of customers across multiple markets, the company operates a highly available cloud-native infrastructure where reliability, scalability, and operational excellence are core business priorities.

We are looking for an experienced Site Reliability Engineer to help strengthen platform reliability, improve production operations, and drive automation across the engineering organization.

As a Site Reliability Engineer, you will own and improve the reliability, availability, and operational health of a large-scale cloud platform. You'll collaborate closely with Software Engineers, Infrastructure, Customer Operations, and Product teams while helping evolve production support processes and operational standards.

The role combines traditional SRE responsibilities with modern AI-assisted engineering practices, leveraging AI tools to improve incident response, documentation, operational workflows, and engineering productivity.

This position is ideal for someone who enjoys solving complex production challenges, improving observability, automating repetitive operational work, and building scalable reliability practices.

Responsibilities:

  • Act as the primary technical escalation point for critical production incidents, providing hands-on support during high-severity outages.
  • Improve platform reliability by reviewing new product launches, infrastructure changes, and production readiness before release.
  • Design, implement, and optimize monitoring, alerting, and observability solutions across cloud infrastructure and applications.
  • Analyze operational metrics, recurring alerts, and incident trends to reduce alert fatigue and improve overall system health.
  • Lead incident investigations and post-mortems, ensuring root causes are identified and preventative actions are implemented.
  • Collaborate with Engineering, Infrastructure, Customer Operations, and external support teams to coordinate incident response and customer communications.
  • Participate in capacity planning, peak traffic readiness, disaster recovery exercises, and system performance reviews.
  • Develop and maintain operational runbooks, documentation, and incident response procedures.
  • Improve internal reliability tooling and automate operational workflows using modern AI-assisted development tools.
  • Drive continuous improvements in operational excellence through automation, standardization, and proactive reliability initiatives.

Requirements:

  • 4+ years of experience as a Site Reliability Engineer, DevOps Engineer, Production Engineer, or a similar infrastructure-focused role.
  • Strong experience supporting production systems running on AWS.
  • Hands-on experience with monitoring and observability platforms such as Datadog, AWS CloudWatch, New Relic, or similar.
  • Experience with incident management platforms such as PagerDuty.
  • Strong understanding of production incident management, root cause analysis, and post-incident review processes.
  • Experience working with ticketing and documentation platforms such as Jira and Confluence.
  • Familiarity with operational dashboards and reporting tools (Looker or similar BI platforms).
  • Experience building operational documentation, runbooks, and support processes.
  • Comfortable working outside regular business hours when critical production incidents require senior engineering support.
  • Experience using AI-assisted engineering tools (Claude, GitHub Copilot, Cursor, or similar) to improve engineering workflows, automate documentation, incident triage, reporting, or operational tasks.
  • Strong scripting or automation skills (Python, Bash, or similar) are considered a plus.

Nice to Have:

  • Experience working in fintech, payments, financial services, or other high-availability environments.
  • Experience with Infrastructure as Code (Terraform, CloudFormation, or similar).
  • Familiarity with Kubernetes and containerized environments.
Vacancy posted 7 hours ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in United States vacancy
  • $160k - $180k

     ...expertise, and world-class customer satisfaction. The Platform Engineering group at CentralReach builds the underlying technologies that...  ...in Software Engineering to drive adoption of modern reliability practices like SLOs, error budget policies, actionable alerts... 
    Suggested
    Full time
    Worldwide

    CentralReach

    Holmdel, NJ
    3 days ago
  • $160k - $240k

     ...one another millions of times a day - quickly, reliably, and securely. Any time you swipe your credit...  ...come make a difference at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our global team in... 
    Suggested
    Full time

    Fiserv

    Sunnyvale, CA
    2 days ago
  • $104.9k - $174.7k

     ...SRE role is responsible for improving the reliability, availability, performance, and...  ...actions through completion.Follow up with engineering, development, security, support, and business...  ...Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering... 
    Suggested
    Full time
    Local area

    RELX Group

    San Jose, CA
    1 day ago
  •  ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS... 
    Suggested

    Kong

    Washington DC
    3 days ago
  • $135k - $155k

     ...while also making it easy for buyers at Fortune 1000 companies to tap into global manufacturing capacity.Xometry is seeking a Site Reliability Engineer II to join our Site Reliability Engineering (SRE) Organization. In this role as an individual contributor, you will guide... 
    Suggested
    Flexible hours

    Thomas

    Denver, CO
    1 day ago
  • $130k - $153k

     ...our customers, and in our growing commitment to land stewardship and recreational access.WHAT YOU WILL DOonX is seeking a Site Reliability Engineer to build and maintain the infrastructure that enables our developers to ship reliably at scale. You'll manage onX's infrastructure... 
    Full time
    Local area
    Remote work

    onXmaps

    Missoula, MT
    1 hour ago
  • Total Number of Openings1Chevron is accepting online applications for the position Site Reliability Engineer (SRE) through October 14, 2026 at 11:59 p.m. (Central Time).Overview:The Site Reliability Engineer (SRE) is responsible for ensuring the integrity, reliability,... 
    Full time
    Relocation
    Visa sponsorship
    Work visa

    Chevron Oil Company

    Houston, TX
    2 days ago
  • $95.3k - $158.8k

    Senior Site Reliability Engineer Are you passionate about building resilient, scalable systems that power mission-critical applications?Do you thrive on automating operations, improving reliability, and ensuring exceptional system performance?About the team:Embedded Innovation... 
    Full time
    Local area
    Remote work
    Work from home

    Elsevier

    Philadelphia, PA
    2 days ago
  • $155k - $195k

     ...you to join us on our mission of providing humankind access to the galaxy beyond our planet. About the RoleWe are seeking a Site Reliability Engineer to join our Ground Software team. As a Site Reliability Engineer, you will design, build, and operate the ground and site... 
    Permanent employment
    Full time
    Work at office

    Apex Technology

    Los Angeles, CA
    4 days ago
  • $70.8k - $131.4k

    Job DescriptionThomson Reuters is strengthening its Site Reliability Engineering capability to help engineering and operations teams build, operate, and improve reliable production services.The Site Reliability Engineer will support the tools, processes, and operational... 
    Full time
    Work at office
    Local area
    Flexible hours

    Thomson Reuters

    Eagan, MN
    2 days ago
  •  ...We are seeking a Senior SRE / DevSecOps Engineer with strong experience in Kubernetes, AWS...  .... The role will focus on platform reliability, incident management, SLO/SLI governance...  ...Experience with SLO/SLI governance and site reliability practices. ~ Strong understanding... 
    Contract work

    PB consulting

    Charlotte, NC
    a month ago
  •  ...in Atlanta, Georgia, and serves customers in more than 35 countries worldwide.Site Reliability EngineerOnsite: Atlanta, GAJob SummaryAt NCR Voyix, we're looking for a Site Reliability Engineer II to help build, support, and scale the cloud platforms that power our... 
    Full time
    Worldwide
    Flexible hours

    NCR

    Atlanta, GA
    2 days ago
  • The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams... 
    Full time
    Work at office
    Local area

    Castleton Commodities International

    Houston, TX
    2 days ago
  • $175k - $215k

     ...unforgettable experiences — and we’re constantly looking for new ways to enhance these exciting experiences.Sr. Manager, Site Reliability Engineer provides strategic leadership across multiple SRE teams and their managers, ensuring alignment with organizational priorities... 

    Disney Interactive

    Orlando, FL
    2 hours ago
  •  ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises... 
    Permanent employment
    Full time
    Part time
    H1b
    Work at office
    Local area
    Immediate start
    Work visa
    Monday to Friday
    Shift work
    Day shift

    Truist

    Raleigh, NC
    3 days ago
  • $128.6k - $184.9k

     ...FedRAMP team builds and operates secure, reliable cloud services for U.S. government...  ...customers. We partner closely with application engineering, security, compliance, and...  ...services, DevOps, platform engineering, or Site Reliability Engineering in AWS.Experience... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours

    CISCO Systems

    Boxborough, MA
    2 days ago
  •  ...enterprise solutions within the Information Security team. The successful candidate will apply an engineering-based approach to solving complex security and reliability challenges, leveraging machine data analytics and log analysis across the enterprise. Designing systems... 
    Full time
    Internship
    Summer internship
    Work at office
    Local area
    Remote work
    Flexible hours

    F5 Networks

    Seattle, WA
    1 day ago
  • $135k - $154k

     ...where you matter.Your ImpactAs a contributor in the APX platform engineering organization on the CloudNet team, you are passionate about...  .... You are also obsessed about achieving the high quality and reliability our customers demand. You will work closely with sovereign... 
    Work experience placement
    Work at office
    Remote work

    Axon

    Washington DC
    4 days ago
  • $108k - $216k

     ...PermanentCompany: VizioBusiness Segment: Home OfficePosition: Senior Site Reliability EngineerJob Location: 39 Tesla, Irvine, CA 92618Duties:...  ...to customer support experiences. Collaborate with engineering teams to embed reliability into the software development lifecycle... 
    Full time
    Temporary work

    DCC Technology

    Irvine, CA
    1 day ago
  • $75 - $84 per hour

    DescriptionKforce has a client in Irvine, CA that is seeking a Site Reliability Engineer for an onsite role - must work with PST hours.Summary:We are seeking a Site Reliability Engineer (SRE) to support the reliability, scalability, and performance of a cloud-native platform... 

    KForce

    Irvine, CA
    1 day ago
  •  ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production support, incident management, monitoring, observability, troubleshooting, and improving system reliability and availability... 
    Contract work

    2T Consulting

    Atlanta, GA
    a month ago
  •  ...The Role:GIPHY is seeking a highly experienced Site Reliability Engineer to join our SRE team. You will help design, build, operate, and evolve the infrastructure that powers GIPHY, including our cloud environment, Kubernetes clusters, and CI/CD platforms.You will also... 
    Full time
    Work experience placement
    Remote work

    Shutterstock

    New York, NY
    3 days ago
  • $96.8k - $145.2k

     ...If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a Site Reliability Engineer (Onsite Hybrid) to join our team in Plano, Texas (US-TX), United States (US).Job Responsibilities Include: Own and manage... 
    Temporary work
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Plano, TX
    3 days ago
  • $81.1k - $187k

     .... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers...  ...posted.Career Level - IC3Escalation points for junior Site Reliability Engineers during complex or high-impact incidents.Manage and... 
    Temporary work
    Monday to Friday
    Flexible hours
    Shift work
    Night shift

    Oracle Corporation

    Reston, VA
    2 hours ago
  • $143k - $194k

     ...mission critical capabilities to our customers. System Deployment Engineers work in complex environments with shared environmental...  ...customersParticipate in customer demonstrations and exercisesWork with site reliability engineers to provide and refine requirements for tooling and... 
    Full time
    Temporary work
    Work experience placement
    Immediate start

    Anduril Industries

    Seattle, WA
    2 days ago
  • Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firm’s most critical customer-facing microservices that power all eCommerce channels. This role applies Google-inspired SRE principles to balance... 
    Local area
    Remote work
    Flexible hours
    Shift work

    O'Reilly Auto Parts

    Springfield, MO
    3 days ago
  • $119.8k - $234.7k

     ...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual...  ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...opportunity for a Senior Site Reliability Engineer (SRE) to join the Azure Silver and Sovereign... 
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Reston, VA
    3 days ago
  •  ...importance of in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location(s).As a Site Reliability Engineer supporting the Cashiering organization, you will play a critical role in ensuring the stability,... 
    Full time
    Work at office

    The Charles Schwab Corporation

    Southlake, TX
    1 day ago
  • $152.5k - $205k

     ...work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind critical... 
    Flexible hours

    Circle

    San Francisco, CA
    1 day ago
  •  ...principles to see it in full.About the teamThe Engineering team at Airwallex is a diverse group of...  ..., working together to build scalable, reliable, and secure products that empower...  ...Global services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work closely... 
    Temporary work
    Local area

    Airwallex

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!