Site Reliability Engineer
System Automation
Mid-Level Site Reliability Engineer
We're looking for a mid-level Site Reliability Engineer to help build and operate the critical cloud infrastructure behind our platform in Microsoft Azure. You'll define the observability standards (SLOs, SLIs, dashboards, alerting) that tell us whether our systems are healthy, automate away manual toil, and help the team ship changes safely and often through solid CI/CD practices. You'll also share in an on-call rotation, responding to incidents and driving blameless postmortems that make the platform more resilient over time.
This role suits someone with hands-on IT Operations or SRE experience, comfort working inside an agile team, and a genuine cloud-native understanding of how to design for reliability, security, and scale.
Key Responsibilities
Reliability & Operations
- Build, operate, and scale production systems in Microsoft Azure (App Service, Networking, WAF, CosmosDB and related infrastructure) to meet availability and performance targets.
- Participate in an on-call rotation; triage, respond to, and resolve production incidents, and lead or contribute to blameless postmortems.
- Define and track SLOs/SLIs and error budgets in partnership with engineering and product teams.
Observability & Monitoring
- Design and maintain observability and APM tooling (metrics, logs, tracing, dashboards, alerting) so issues are caught before they impact customers.
- Continuously refine alert thresholds and runbooks to reduce noise and mean time to resolution.
Automation & CI/CD
- Reduce operational toil through automation — scripting, self-healing systems, and repeatable processes.
- Build and maintain CI/CD pipelines that let the development team ship safely and frequently.
- Provision and manage infrastructure as code (Bicep) and follow standard change control and version control practices.
Security & Compliance
- Ensure application infrastructure meets security and compliance requirements (e.g., SOC 2, GovRAMP) in partnership with the compliance team.
- Apply security best practices to infrastructure design and change management.
Collaboration & Documentation
- Partner with the agile development team to translate business requirements into reliable technical solutions.
- Participate in technical design sessions and produce clear documentation (diagrams, runbooks, architecture notes).
- Stay current on new Azure capabilities, industry standards, and SRE best practices, and bring recommendations back to the team.
- Other duties as assigned.
Knowledge, Skills, and Abilities
• Solid understanding of networking fundamentals, and observability principles.
• Ability to evaluate multiple technical approaches and recommend the most effective solution for the context.
• Strong independent problem-solving skills balanced with effective collaboration in a team environment.
• Familiarity with software development lifecycle and programming/coding standards.
• Clear, professional communication, especially under incident pressure.
Qualifications
Required
- 3+ years of experience in an IT Operations, DevOps, or SRE role.
- Hands-on technical experience with Microsoft Azure in a production environment.
- Experience with infrastructure as code — Terraform and/or Bicep.
- Proficiency in at least one scripting/programming language — Python or TypeScript.
- Experience working with REST and/or GraphQL APIs.
- Experience defining and tracking KPIs/SLOs for a web-based application.
- Comfortable participating in an on-call rotation.
Preferred
- Experience with compliance audits (SOC 2 Type 2, GovRAMP).
- Familiarity with security frameworks (NIST, ISO 27001).
- AZ-104 certification, or equivalent Azure networking experience.
- Experience with Node.js.
- Experience with low-code platforms (Power Apps, Logic Apps).
- Familiarity with Scrum/Agile methodology and supporting tools (Confluence, JIRA, Git, Jenkins, Bamboo, TFS).
- Ability to translate business requirements directly into application/site behavior changes.
What to Expect
Hiring Process: Application review ? interviews and technical assessment ? conditional offer
Security Screening: Due to the sensitive nature of IT systems and data access, final candidates will receive a conditional employment offer contingent upon successful completion of a drug screening and a fingerprint-based background investigation.
We may use technology-assisted tools, including artificial intelligence tools, to assist recruiters and hiring managers in reviewing application materials and identifying qualifications relevant to a position. These tools support, but do not replace human decision-making. All employment decisions are made by qualified human reviewers. Applicants requiring accommodation during the application or selection process may contact us at View email address on click.appcast.io.
We are an Equal Opportunity Employer and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, disability, protected veteran status, or any other characteristic protected by applicable law.
Salary Description 120,000 - 140,000
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex...SuggestedWork at office
- ...to physicians, providing critical information about the right treatments for the right patients, at the right time.The Site Reliability Engineering team works with all departments and business units to provide dependable cloud infrastructure solutions, along with support...SuggestedFull time
$80k - $133k
...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and...SuggestedPermanent employmentFull timeContract workRemote workFlexible hours$165k - $280k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most...SuggestedPermanent employmentTemporary workWorldwideWeekend work- Recognized as the No. 1 site trusted by real estate professionals, Realtor.com has been at the forefront of online real estate... ...confidence through expert guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization,...SuggestedWork at officeLocal area
- Reliability Engineering Design, implement, and operate scalable, resilient, and highly available systems on Google Cloud Platform. Improve service... ...Skills, and Abilities Three or more years of experience in Site Reliability Engineering, platform engineering, DevOps, cloud...Remote work
$139k - $257.55k
The ChallengeThe Adobe Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning, autonomous AI workflows, and cloud-native infrastructure. Adobe Stock gives designers and businesses...Full timeTemporary workLocal areaRemote workWorldwide$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability Engineering, this role will help build,...Permanent employmentLocal areaWorldwideFlexible hours$158.5k - $172k
...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and... .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology...Full timeTemporary workWork at officeFlexible hours3 days per week$138.4k - $173k
...infrastructure as well as help improve the reliability, quality of services and overall... ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability... ...about our locations by visiting our site.Compensation & BenefitsThe base salary that...Full timeFlexible hours$125k - $150k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (RAPTOR)SpaceX is looking for a Site Reliability Engineer with a strong drive to solve challenging problems in the Raptor...Permanent employmentTemporary work$230k - $250k
GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is...Remote work$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...Full timeWork at officeLocal areaRemote workWork from home$128.6k - $184.9k
...global cloud platform. As a team of six engineers distributed across the US, Canada, and the... ...with a strong focus on automation, reliability, and operational excellence. We are one... ...Qualifications7+ years of experience in Site Reliability Engineering, DevOps, Infrastructure...Permanent employmentFull timeTemporary workLocal areaWorldwideFlexible hours$130k - $200k
IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion...Full timeWork at officeImmediate start$165k - $190k
...DevOps / SRE TeamThe DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable, scalable, and high-... ...security platformAddress complex challenges around scalability, reliability, observability, and cost efficiencyCollaborate with Engineering...Work from home$130k - $153k
...our customers, and in our growing commitment to land stewardship and recreational access.WHAT YOU WILL DOonX is seeking a Site Reliability Engineer to build and maintain the infrastructure that enables our developers to ship reliably at scale. You'll manage onX's infrastructure...Full timePart timeWork at office$45 - $85 per hour
DescriptionThe Resy Site Reliability Engineering groups goal is to ensure Resy Customers can always use the service reliably. We're looking for engineers to be part of an empowered, self-organizing group, with the opportunity to use modern languages and tools and to operate...Contract workTemporary work- Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems thinking, and operational excellence...Worldwide
$104.9k - $174.7k
About the role:A FinOps Site Reliability Engineer (SRE) bridges the gap between engineering, operations, and financial governance by embedding cost optimization into infrastructure design, automation, monitoring, and operational processes. A FinOps SRE proactively identifies...Full timeLocal area$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...Currently hiringRemote work- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure...
$107.9k - $195.05k
Site Reliability EngineerLocation: Hickam Air Force Base, HawaiiClearance: TS/SCILeidos has an opening for a highly qualified TS/SCI cleared Site Reliability Engineer at Hickam Air Force Base, Hawaii for the Decision Advantage Business Area in Defense sector. This is an...Full time$230k - $250k
...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change... ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"...Night shift- Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources Willingness to work on-site at stated location in the job openingDepartment...Contract workFor contractorsWork experience placement
$143k - $191k
...mission critical capabilities to our customers. System Deployment Engineers work in complex environments with shared environmental... ...customersParticipate in customer demonstrations and exercisesWork with site reliability engineers to provide and refine requirements for tooling and...Full timeTemporary workWork experience placementImmediate start$62k - $141k
Site Reliability EngineerThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if you have...Full timeContract workPart timeWork at officeLocal areaRemote work$148.5k - $223.9k
...right place! Agentforce is the future of AI, and you are the future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations,...Full timeWorldwideWeekend work$167.7k - $245.2k
...very effective.We’re looking for talented engineers with a software or operations background... ...development teams to ensure the reliability, performance and security of our infrastructure... ...insurance. Please see the Cisco careers site to discover more benefits and perks....Full timeTemporary workWork at officeLocal areaFlexible hours1 day per week- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and broad business problems with simple and straightforward solutions...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer remote United States
- lead site reliability engineer United States
- site reliability engineer United States
- site reliability engineer sre United States
- site reliability engineering manager United States
- junior website developer United States
- website content developer United States
- on site coordinator United States
- after school site coordinator United States
- website coordinator United States
