Staff Site Reliability Engineer
ASSURED
Staff Site Reliability Engineer
Assured is on a mission to modernize insurance. Claims processing (i.e. should we pay this claim?), while often overlooked, is the foundation of the entire industry. It's currently highly manual, involving phone calls, faxes, and gut instinct, costing tens of billions of dollars a year. We can do better.
At Assured, we provide large insurers with the software solutions they need to win in a modern, technology-driven world. From self-service claim-filing software to backend fraud detection, we're the engine that powers claims processing for some of the largest insurers in the world.
The challenges we face are deep and diverse, from creating digital experiences that provide comfort and clarity to claimants at their most stressed and vulnerable to orchestrating large-scale ML-driven decision-making on billions of dollars of claims payments, life at Assured is dynamic, collaborative, and rewarding.
As a Staff Site Reliability Engineer, you'll define the standards, patterns, and platforms that let every engineering team at Assured run their services reliably — and partner directly with the teams adopting them.
Some of the Problems You'll Solve:
Define what reliability means across a growing engineering organization.
Set the standards, patterns, and reference implementations teams adopt for SLOs, error budgets, instrumentation, alerting, and incident practice — and make them easy enough to adopt that teams actually do.
Measure reliability the way customers experience it.
Move us beyond component-level availability targets to end-to-end SLOs for the claims journeys insurers and claimants depend on, spanning many services and teams, and extend that into how we measure and report SLA compliance.
Unify a fragmented observability picture.
Help drive our consolidation onto OpenTelemetry as a single instrumentation standard across shared services and product applications, so signal is consistent and comparable wherever it comes from.
Turn scattered reliability signal into decisions.
Build on and refine the reporting layer that pulls incident, alerting, and coverage data into one place, surfacing where risk actually lives across the platform and where we're flying blind.
Shorten the distance between an incident and a lasting improvement.
Improve how we detect, respond to, and learn from failure — incident tooling and automation, post-incident review practice, and making sure action items get closed rather than quietly aging out.
How You'll Make an Impact:
Help other teams run their own systems well.
Embed with product teams for a period at a time: specify what reliability looks like for their most critical paths, help them build it, then hand it over with them as the durable owner.
Find the risk before it finds us.
Surface coverage gaps, weak signals, and single points of failure across the platform, and make the case for fixing them before they become incidents.
Support engineering when things go wrong.
Share an interrupt rotation with the rest of the SRE team, triaging reliability escalations and requests from across the organization.
Raise the technical bar around you.
Mentor engineers across the organization through design review, written guidance, and hands-on collaboration on the problems they own.
Use AI to work faster and more effectively.
Use tools such as Claude, Codex, Cursor, and similar platforms to support tooling development, incident analysis, debugging, documentation, and operational work.
You'll Probably Thrive Here If You:
Have deep site reliability and systems expertise.
You bring 10+ years of site reliability, production, or platform engineering experience, ideally within SaaS platforms or high-scale distributed systems environments.
Work across a modern reliability and observability stack.
You're comfortable with OpenTelemetry, metrics, traces and logs, AWS, Kubernetes, PostgreSQL, and modern incident tooling. Experience with every tool isn't required — we value strong fundamentals and the ability to learn quickly. Platform provisioning and cloud infrastructure sits with a separate Infrastructure team, and you'll work closely with a dedicated Database Reliability Engineering function.
Know how to make SLOs stick.
You've designed and landed SLOs and error budgets that teams genuinely use to make decisions, rather than dashboards nobody opens.
Lead through influence rather than ownership.
Our SRE function advises and enables rather than executing on other teams' behalf. You can bring a product team along with you, and redirect work that genuinely belongs elsewhere.
Stay effective when both systems and people are under pressure.
You've run incidents and post-incident reviews, and you're as comfortable coordinating people mid-incident as you are debugging the failure itself.
Build, rather than only configure.
You write real tooling and services, and you can reason about failure modes in systems you didn't build and debug across team boundaries.
Adapt quickly to new tools and technologies.
Your engineering judgment and ability to learn matter more than experience with a particular framework or platform. Great engineers learn new technologies. Great systems thinking is harder to teach.
Benefits:
Competitive Compensation: Competitive salary and equity packages for all employees
Healthcare Plan: Platinum medical, dental, and vision
Free life insurance: Including long-term disability & short-term disability
Unlimited PTO: Uncapped vacation days & paid holidays
Family Leave: Maternity & paternity
401(k) Contribution: Assured contributes 3% of your income, even if you don't contribute
WFH Benefits: Lunch on us 2x/week, monthly phone stipend & other home office perks
Health FSAs & HSAs: Pre-tax accounts for out-of-pocket medical expenses
Team events & Offsites: We're remote, but we regularly get together
**We have been made aware of individuals falsely posing as recruiters from Assured Insurance Technologies Inc. Please note that we only contact candidates from official @assured.claims email addresses and all interviews are conducted through verified company channels. If you are unsure whether a message is legitimate, please contact us directly at View email address on click.appcast.io before sharing any personal information **
Our Commitment: We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, sex, gender, gender expression, sexual orientation, age, marital status, veteran status, or disability status. We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform essential job functions, and receive other benefits and privileges of employment. Please contact us to request accommodation.
- ...The Team Platform Engineering is the department within SRE that is responsible for a range... ...role in developing and maintaining the reliable and globally connected multi-cloud network... ...Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong...SuggestedFull timeWork at officeRemote workWorldwide
- ...: Lovelace is the only provider of enterprise-scale context engines capable of analyzing trillions of real-time data points to create... ...: ~ Lovelace AI is seeking a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team. As an SRE at...SuggestedFull time
- ...to meet you. Our Enterprise Information Technology (EIT) organization is expanding, and we are seeking a Senior Site Reliability Engineer to help drive a major architectural modernization. In this role, you will move beyond traditional infrastructure maintenance...SuggestedPermanent employmentFull timeH1bLocal areaRemote workShift work
$153k - $210k
...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating...SuggestedFull time$140k - $230k
...Zoox is seeking a Site Reliability Engineer to help ensure the availability, performance, and resilience of the services that power the development and operation of our autonomous vehicles. In this role, you will own the full lifecycle of our services—from designing fault...SuggestedFull time- ...Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems... ...make safe shipping the default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale...Permanent employmentWork experience placementWork at officeLocal area
- ...and best in class outcomesVisionary in future focused problem-solvingExceptional in execution and impactThe RoleAs a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to deliver...Full timeFlexible hours
- ...cloud-native platforms to advanced release engineering practices, our teams are redefining how... ...alert behavior preferred Exposure to reliability engineering concepts such as SLOs/SLIs and... ...office#LI-KC1#GMFjobsAbout The Role:The Site Reliability Engineer under the general...Work experience placementH1bWork at officeRemote workVisa sponsorshipFlexible hoursShift work2 days per week
$15k
...office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to... ..., and work to provide the best experience to our technical staff. You will leverage IaC, Automation, and SRE principles to refine...Work at officeLocal areaRemote work$113.4k - $162k
...break down barriers to communication and free the flow of conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd, reliability and everything in between!This role is about impact at...Temporary work$118.6k - $195.68k
Job SummaryThe Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid OpenShift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat OpenShift at scale...Permanent employmentFull timeContract workWork experience placementWork at officeRemote workFlexible hours$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....Work at officeLocal areaRemote workWorldwideFlexible hours$165k - $230k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts....Permanent employmentTemporary workImmediate startWeekend work- IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion...Work at officeImmediate start
- ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and... ...and networking teams to improve service reliability and deployment workflowsDeploy and... ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering...Work at officeLocal areaWork from homeFlexible hours
$130k - $180k
...of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform guardrails...Work at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours$81.1k - $187k
.... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers... ...posted.Career Level - IC3Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and...Temporary workMonday to FridayFlexible hoursShift workNight shift- ...United States of America / Alberta / British ColumbiaTechnology - Engineering /Full-time - Permanent /RemoteAbout MegaportWe’re not your... ...goals are met.What You Will Be DoingImproving production reliability and system resilience within an SRE scoped teamChampioning high...Permanent employmentFull timeRemote workFlexible hours
- ...us a leader in the industry, and we're searching for exceptional talent to help us stay at the cutting edge. As a DevOps Site Reliability Engineer, you’ll have the chance to contribute to the continuous evolution and enhancement of our customer-facing products. Our DevOps...Work from home2 days per week
$104.9k - $174.7k
About the role:A FinOps Site Reliability Engineer (SRE) bridges the gap between engineering, operations, and financial governance by embedding cost optimization into infrastructure design, automation, monitoring, and operational processes. A FinOps SRE proactively identifies...Full timeLocal area$130k - $150k
...Cloud SolutionsInformation SecurityInformation Technology staff are based in the Boston, Chicago, London, Munich, New York,... ...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are...Work at officeWork from home3 days per week- ...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient...Ongoing contractFull timeTemporary workWork experience placementWorldwide
- ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering... ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or...Work at officeLocal areaWork from homeFlexible hours
$167.7k - $245.2k
...very effective.We’re looking for talented engineers with a software or operations background... ...development teams to ensure the reliability, performance and security of our infrastructure... ...insurance. Please see the Cisco careers site to discover more benefits and perks....Full timeTemporary workWork at officeLocal areaFlexible hours1 day per week$114.3k - $235.32k
...verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business.We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven...Work at officeLocal areaRelocationRelocation package$119.8k - $234.7k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewAre... ...no further than the Microsoft Defender engineering team. We are looking for a Senior Site...Ongoing contractLocal area3 days per week- Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firm’s most critical customer-facing microservices that power all eCommerce channels. This role applies Google-inspired SRE principles to balance...Local areaRemote workFlexible hoursShift work
- ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
- Play a key role in ensuring system reliability at one of the world’s most iconic and largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will use technology to solve...
$119.8k - $234.7k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:... ...opportunity for a Senior Site Reliability Engineer (SRE) to join the Azure Silver and Sovereign...Ongoing contractLocal area3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!
- engineering aide United States
- technology administrator United States
- research assistant engineering United States
- staff security engineer United States
- information technology administrative assistant United States
- assistant engineering manager United States
- assistant building engineer United States
- project engineer assistant project manager United States
- staff qa engineer United States
- staff devops engineer United States


