Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

System Automation

Mid-Level Site Reliability Engineer

We're looking for a mid-level Site Reliability Engineer to help build and operate the critical cloud infrastructure behind our platform in Microsoft Azure. You'll define the observability standards (SLOs, SLIs, dashboards, alerting) that tell us whether our systems are healthy, automate away manual toil, and help the team ship changes safely and often through solid CI/CD practices. You'll also share in an on-call rotation, responding to incidents and driving blameless postmortems that make the platform more resilient over time.

This role suits someone with hands-on IT Operations or SRE experience, comfort working inside an agile team, and a genuine cloud-native understanding of how to design for reliability, security, and scale.

Key Responsibilities

Reliability & Operations

  • Build, operate, and scale production systems in Microsoft Azure (App Service, Networking, WAF, CosmosDB and related infrastructure) to meet availability and performance targets.
  • Participate in an on-call rotation; triage, respond to, and resolve production incidents, and lead or contribute to blameless postmortems.
  • Define and track SLOs/SLIs and error budgets in partnership with engineering and product teams.

Observability & Monitoring

  • Design and maintain observability and APM tooling (metrics, logs, tracing, dashboards, alerting) so issues are caught before they impact customers.
  • Continuously refine alert thresholds and runbooks to reduce noise and mean time to resolution.

Automation & CI/CD

  • Reduce operational toil through automation — scripting, self-healing systems, and repeatable processes.
  • Build and maintain CI/CD pipelines that let the development team ship safely and frequently.
  • Provision and manage infrastructure as code (Bicep) and follow standard change control and version control practices.

Security & Compliance

  • Ensure application infrastructure meets security and compliance requirements (e.g., SOC 2, GovRAMP) in partnership with the compliance team.
  • Apply security best practices to infrastructure design and change management.

Collaboration & Documentation

  • Partner with the agile development team to translate business requirements into reliable technical solutions.
  • Participate in technical design sessions and produce clear documentation (diagrams, runbooks, architecture notes).
  • Stay current on new Azure capabilities, industry standards, and SRE best practices, and bring recommendations back to the team.
  • Other duties as assigned.
Knowledge, Skills, and Abilities

• Solid understanding of networking fundamentals, and observability principles.

• Ability to evaluate multiple technical approaches and recommend the most effective solution for the context.

• Strong independent problem-solving skills balanced with effective collaboration in a team environment.

• Familiarity with software development lifecycle and programming/coding standards.

• Clear, professional communication, especially under incident pressure.

Qualifications

Required

  • 3+ years of experience in an IT Operations, DevOps, or SRE role.
  • Hands-on technical experience with Microsoft Azure in a production environment.
  • Experience with infrastructure as code — Terraform and/or Bicep.
  • Proficiency in at least one scripting/programming language — Python or TypeScript.
  • Experience working with REST and/or GraphQL APIs.
  • Experience defining and tracking KPIs/SLOs for a web-based application.
  • Comfortable participating in an on-call rotation.

Preferred

  • Experience with compliance audits (SOC 2 Type 2, GovRAMP).
  • Familiarity with security frameworks (NIST, ISO 27001).
  • AZ-104 certification, or equivalent Azure networking experience.
  • Experience with Node.js.
  • Experience with low-code platforms (Power Apps, Logic Apps).
  • Familiarity with Scrum/Agile methodology and supporting tools (Confluence, JIRA, Git, Jenkins, Bamboo, TFS).
  • Ability to translate business requirements directly into application/site behavior changes.

What to Expect

Hiring Process: Application review ? interviews and technical assessment ? conditional offer

Security Screening: Due to the sensitive nature of IT systems and data access, final candidates will receive a conditional employment offer contingent upon successful completion of a drug screening and a fingerprint-based background investigation.

We may use technology-assisted tools, including artificial intelligence tools, to assist recruiters and hiring managers in reviewing application materials and identifying qualifications relevant to a position. These tools support, but do not replace human decision-making. All employment decisions are made by qualified human reviewers. Applicants requiring accommodation during the application or selection process may contact us at View email address on click.appcast.io.

We are an Equal Opportunity Employer and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, disability, protected veteran status, or any other characteristic protected by applicable law.

Salary Description 120,000 - 140,000

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in United States vacancy
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex... 
    Suggested
    Work at office

    JP Morgan Chase

    Jersey City, NJ
    4 days ago
  •  ...to physicians, providing critical information about the right treatments for the right patients, at the right time.The Site Reliability Engineering team works with all departments and business units to provide dependable cloud infrastructure solutions, along with support... 
    Suggested
    Full time

    Tempus

    Chicago, IL
    4 days ago
  • $80k - $133k

     ...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and... 
    Suggested
    Permanent employment
    Full time
    Contract work
    Remote work
    Flexible hours

    Guidehouse

    San Antonio, TX
    4 days ago
  • $165k - $280k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most... 
    Suggested
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Palo Alto, CA
    3 days ago
  • Recognized as the No. 1 site trusted by real estate professionals, Realtor.com has been at the forefront of online real estate...  ...confidence through expert guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization,... 
    Suggested
    Work at office
    Local area

    Realtor.com

    Austin, TX
    3 days ago
  • Reliability Engineering Design, implement, and operate scalable, resilient, and highly available systems on Google Cloud Platform. Improve service...  ...Skills, and Abilities Three or more years of experience in Site Reliability Engineering, platform engineering, DevOps, cloud... 
    Remote work

    Patterson-UTI

    Houston, TX
    20 hours ago
  • $139k - $257.55k

    The ChallengeThe Adobe Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning, autonomous AI workflows, and cloud-native infrastructure. Adobe Stock gives designers and businesses... 
    Full time
    Temporary work
    Local area
    Remote work
    Worldwide

    Adobe Systems

    New York, NY
    4 days ago
  • $165k - $225.6k

     ...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability Engineering, this role will help build,... 
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    2 days ago
  • $158.5k - $172k

     ...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and...  .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology... 
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    Chicago, IL
    2 days ago
  • $138.4k - $173k

     ...infrastructure as well as help improve the reliability, quality of services and overall...  ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability...  ...about our locations by visiting our site.Compensation & BenefitsThe base salary that... 
    Full time
    Flexible hours

    AppFolio

    Santa Barbara, CA
    2 days ago
  • $125k - $150k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (RAPTOR)SpaceX is looking for a Site Reliability Engineer with a strong drive to solve challenging problems in the Raptor... 
    Permanent employment
    Temporary work

    SpaceX

    Hawthorne, CA
    2 days ago
  • $230k - $250k

    GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is... 
    Remote work

    Govcio

    Arlington, VA
    2 days ago
  • $104.9k - $174.7k

     ...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    LexisNexis Risk Solutions Group

    Atlanta, GA
    1 day ago
  • $128.6k - $184.9k

     ...global cloud platform. As a team of six engineers distributed across the US, Canada, and the...  ...with a strong focus on automation, reliability, and operational excellence. We are one...  ...Qualifications7+ years of experience in Site Reliability Engineering, DevOps, Infrastructure... 
    Permanent employment
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    CISCO Systems

    Richardson, TX
    1 day ago
  • $130k - $200k

    IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion... 
    Full time
    Work at office
    Immediate start

    IXL Learning

    San Mateo, CA
    20 hours ago
  • $165k - $190k

     ...DevOps / SRE TeamThe DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable, scalable, and high-...  ...security platformAddress complex challenges around scalability, reliability, observability, and cost efficiencyCollaborate with Engineering... 
    Work from home

    Obsidian Security

    Palo Alto, CA
    4 days ago
  • $130k - $153k

     ...our customers, and in our growing commitment to land stewardship and recreational access.WHAT YOU WILL DOonX is seeking a Site Reliability Engineer to build and maintain the infrastructure that enables our developers to ship reliably at scale. You'll manage onX's infrastructure... 
    Full time
    Part time
    Work at office

    onXmaps

    Bozeman, MT
    4 days ago
  • $45 - $85 per hour

    DescriptionThe Resy Site Reliability Engineering groups goal is to ensure Resy Customers can always use the service reliably. We're looking for engineers to be part of an empowered, self-organizing group, with the opportunity to use modern languages and tools and to operate... 
    Contract work
    Temporary work

    TEKsystems

    Phoenix, AZ
    1 day ago
  • Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems thinking, and operational excellence... 
    Worldwide

    Inspire Brands

    Atlanta, GA
    1 day ago
  • $104.9k - $174.7k

    About the role:A FinOps Site Reliability Engineer (SRE) bridges the gap between engineering, operations, and financial governance by embedding cost optimization into infrastructure design, automation, monitoring, and operational processes. A FinOps SRE proactively identifies... 
    Full time
    Local area

    LexisNexis Risk Solutions Group

    Boca Raton, FL
    1 day ago
  • $210k - $230k

    GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation... 
    Currently hiring
    Remote work

    Govcio

    Arlington, VA
    3 days ago
  • About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure... 

    Alembic

    San Francisco, CA
    2 days ago
  • $107.9k - $195.05k

    Site Reliability EngineerLocation: Hickam Air Force Base, HawaiiClearance: TS/SCILeidos has an opening for a highly qualified TS/SCI cleared Site Reliability Engineer at Hickam Air Force Base, Hawaii for the Decision Advantage Business Area in Defense sector. This is an... 
    Full time

    Leidos

    Honolulu, HI
    2 days ago
  • $230k - $250k

     ...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change...  ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"... 
    Night shift

    Forward Networks

    Santa Clara, CA
    2 days ago
  • Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through...  ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform...  ...vendor resources Willingness to work on-site at stated location in the job openingDepartment... 
    Contract work
    For contractors
    Work experience placement

    Cedent Consulting

    Dallas, TX
    2 days ago
  • $143k - $191k

     ...mission critical capabilities to our customers. System Deployment Engineers work in complex environments with shared environmental...  ...customersParticipate in customer demonstrations and exercisesWork with site reliability engineers to provide and refine requirements for tooling and... 
    Full time
    Temporary work
    Work experience placement
    Immediate start

    Anduril Industries

    Seattle, WA
    2 days ago
  • $62k - $141k

    Site Reliability EngineerThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if you have... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Chantilly, Loudoun County, VA
    1 day ago
  • $148.5k - $223.9k

     ...right place! Agentforce is the future of AI, and you are the future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations,... 
    Full time
    Worldwide
    Weekend work

    Salesforce

    San Francisco, CA
    1 day ago
  • $167.7k - $245.2k

     ...very effective.We’re looking for talented engineers with a software or operations background...  ...development teams to ensure the reliability, performance and security of our infrastructure...  ...insurance. Please see the Cisco careers site to discover more benefits and perks.... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    1 day per week

    CISCO Systems

    New York, NY
    1 day ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and broad business problems with simple and straightforward solutions... 

    JP Morgan Chase

    Plano, TX
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!