Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Site Reliability Engineer

ASSURED

Staff Site Reliability Engineer

Assured is on a mission to modernize insurance. Claims processing (i.e. should we pay this claim?), while often overlooked, is the foundation of the entire industry. It's currently highly manual, involving phone calls, faxes, and gut instinct, costing tens of billions of dollars a year. We can do better.

At Assured, we provide large insurers with the software solutions they need to win in a modern, technology-driven world. From self-service claim-filing software to backend fraud detection, we're the engine that powers claims processing for some of the largest insurers in the world.

The challenges we face are deep and diverse, from creating digital experiences that provide comfort and clarity to claimants at their most stressed and vulnerable to orchestrating large-scale ML-driven decision-making on billions of dollars of claims payments, life at Assured is dynamic, collaborative, and rewarding.

As a Staff Site Reliability Engineer, you'll define the standards, patterns, and platforms that let every engineering team at Assured run their services reliably — and partner directly with the teams adopting them.

Some of the Problems You'll Solve:

Define what reliability means across a growing engineering organization.

Set the standards, patterns, and reference implementations teams adopt for SLOs, error budgets, instrumentation, alerting, and incident practice — and make them easy enough to adopt that teams actually do.

Measure reliability the way customers experience it.

Move us beyond component-level availability targets to end-to-end SLOs for the claims journeys insurers and claimants depend on, spanning many services and teams, and extend that into how we measure and report SLA compliance.

Unify a fragmented observability picture.

Help drive our consolidation onto OpenTelemetry as a single instrumentation standard across shared services and product applications, so signal is consistent and comparable wherever it comes from.

Turn scattered reliability signal into decisions.

Build on and refine the reporting layer that pulls incident, alerting, and coverage data into one place, surfacing where risk actually lives across the platform and where we're flying blind.

Shorten the distance between an incident and a lasting improvement.

Improve how we detect, respond to, and learn from failure — incident tooling and automation, post-incident review practice, and making sure action items get closed rather than quietly aging out.

How You'll Make an Impact:

Help other teams run their own systems well.

Embed with product teams for a period at a time: specify what reliability looks like for their most critical paths, help them build it, then hand it over with them as the durable owner.

Find the risk before it finds us.

Surface coverage gaps, weak signals, and single points of failure across the platform, and make the case for fixing them before they become incidents.

Support engineering when things go wrong.

Share an interrupt rotation with the rest of the SRE team, triaging reliability escalations and requests from across the organization.

Raise the technical bar around you.

Mentor engineers across the organization through design review, written guidance, and hands-on collaboration on the problems they own.

Use AI to work faster and more effectively.

Use tools such as Claude, Codex, Cursor, and similar platforms to support tooling development, incident analysis, debugging, documentation, and operational work.

You'll Probably Thrive Here If You:

Have deep site reliability and systems expertise.

You bring 10+ years of site reliability, production, or platform engineering experience, ideally within SaaS platforms or high-scale distributed systems environments.

Work across a modern reliability and observability stack.

You're comfortable with OpenTelemetry, metrics, traces and logs, AWS, Kubernetes, PostgreSQL, and modern incident tooling. Experience with every tool isn't required — we value strong fundamentals and the ability to learn quickly. Platform provisioning and cloud infrastructure sits with a separate Infrastructure team, and you'll work closely with a dedicated Database Reliability Engineering function.

Know how to make SLOs stick.

You've designed and landed SLOs and error budgets that teams genuinely use to make decisions, rather than dashboards nobody opens.

Lead through influence rather than ownership.

Our SRE function advises and enables rather than executing on other teams' behalf. You can bring a product team along with you, and redirect work that genuinely belongs elsewhere.

Stay effective when both systems and people are under pressure.

You've run incidents and post-incident reviews, and you're as comfortable coordinating people mid-incident as you are debugging the failure itself.

Build, rather than only configure.

You write real tooling and services, and you can reason about failure modes in systems you didn't build and debug across team boundaries.

Adapt quickly to new tools and technologies.

Your engineering judgment and ability to learn matter more than experience with a particular framework or platform. Great engineers learn new technologies. Great systems thinking is harder to teach.

Benefits:

Competitive Compensation: Competitive salary and equity packages for all employees

Healthcare Plan: Platinum medical, dental, and vision

Free life insurance: Including long-term disability & short-term disability

Unlimited PTO: Uncapped vacation days & paid holidays

Family Leave: Maternity & paternity

401(k) Contribution: Assured contributes 3% of your income, even if you don't contribute

WFH Benefits: Lunch on us 2x/week, monthly phone stipend & other home office perks

Health FSAs & HSAs: Pre-tax accounts for out-of-pocket medical expenses

Team events & Offsites: We're remote, but we regularly get together

**We have been made aware of individuals falsely posing as recruiters from Assured Insurance Technologies Inc. Please note that we only contact candidates from official @assured.claims email addresses and all interviews are conducted through verified company channels. If you are unsure whether a message is legitimate, please contact us directly at View email address on click.appcast.io before sharing any personal information **

Our Commitment: We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, sex, gender, gender expression, sexual orientation, age, marital status, veteran status, or disability status. We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform essential job functions, and receive other benefits and privileges of employment. Please contact us to request accommodation.

Vacancy posted 16 hours ago
Similar jobs that could be interesting for youBased on the Staff Site Reliability Engineer in United States vacancy
  •  ...The Team Platform Engineering is the department within SRE that is responsible for a range...  ...role in developing and maintaining the reliable and globally connected multi-cloud network...  ...Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong... 
    Suggested
    Full time
    Work at office
    Remote work
    Worldwide

    Mongodb

    United States
    7 days ago
  •  ...: Lovelace is the only provider of enterprise-scale context engines capable of analyzing trillions of real-time data points to create...  ...: ~ Lovelace AI is seeking a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team. As an SRE at... 
    Suggested
    Full time

    Lovelace Ai

    Pittsburgh, PA
    11 hours ago
  •  ...to meet you. Our Enterprise Information Technology (EIT) organization is expanding, and we are seeking a Senior Site Reliability Engineer to help drive a major architectural modernization. In this role, you will move beyond traditional infrastructure maintenance... 
    Suggested
    Permanent employment
    Full time
    H1b
    Local area
    Remote work
    Shift work

    Jack Henry & Associates

    New York, NY
    1 day ago
  • $153k - $210k

     ...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating... 
    Suggested
    Full time

    Ridgeline

    New York, NY
    11 hours ago
  • $140k - $230k

     ...Zoox is seeking a Site Reliability Engineer to help ensure the availability, performance, and resilience of the services that power the development and operation of our autonomous vehicles. In this role, you will own the full lifecycle of our services—from designing fault... 
    Suggested
    Full time

    Zoox

    Remote
    11 hours ago
  •  ...Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems...  ...make safe shipping the default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale... 
    Permanent employment
    Work experience placement
    Work at office
    Local area

    Plaid Financial

    San Francisco, CA
    3 days ago
  •  ...and best in class outcomesVisionary in future focused problem-solvingExceptional in execution and impactThe RoleAs a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to deliver... 
    Full time
    Flexible hours

    Proofpoint

    Florida
    4 days ago
  •  ...cloud-native platforms to advanced release engineering practices, our teams are redefining how...  ...alert behavior preferred Exposure to reliability engineering concepts such as SLOs/SLIs and...  ...office#LI-KC1#GMFjobsAbout The Role:The Site Reliability Engineer under the general... 
    Work experience placement
    H1b
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours
    Shift work
    2 days per week

    GM Financial

    Arlington, TX
    2 days ago
  • $15k

     ...office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to...  ..., and work to provide the best experience to our technical staff. You will leverage IaC, Automation, and SRE principles to refine... 
    Work at office
    Local area
    Remote work

    The Voleon Group

    Berkeley, CA
    2 days ago
  • $113.4k - $162k

     ...break down barriers to communication and free the flow of conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd, reliability and everything in between!This role is about impact at... 
    Temporary work

    TextNow

    San Francisco, CA
    2 days ago
  • $118.6k - $195.68k

    Job SummaryThe Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid OpenShift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat OpenShift at scale... 
    Permanent employment
    Full time
    Contract work
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Red Hat

    Raleigh, NC
    2 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    1 day ago
  • $165k - $230k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts.... 
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    1 day ago
  • IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion... 
    Work at office
    Immediate start

    IXL Learning

    Raleigh, NC
    1 day ago
  •  ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and...  ...and networking teams to improve service reliability and deployment workflowsDeploy and...  ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    1 day ago
  • $130k - $180k

     ...of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform guardrails... 
    Work at office
    Local area
    Remote work
    Worldwide
    Monday to Friday
    Flexible hours

    Imanage

    Chicago, IL
    1 day ago
  • $81.1k - $187k

     .... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers...  ...posted.Career Level - IC3Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and... 
    Temporary work
    Monday to Friday
    Flexible hours
    Shift work
    Night shift

    Oracle Corporation

    Reston, VA
    1 day ago
  •  ...United States of America / Alberta / British ColumbiaTechnology - Engineering /Full-time - Permanent /RemoteAbout MegaportWe’re not your...  ...goals are met.What You Will Be DoingImproving production reliability and system resilience within an SRE scoped teamChampioning high... 
    Permanent employment
    Full time
    Remote work
    Flexible hours

    Megaport

    Texas
    3 hours ago
  •  ...us a leader in the industry, and we're searching for exceptional talent to help us stay at the cutting edge. As a DevOps Site Reliability Engineer, you’ll have the chance to contribute to the continuous evolution and enhancement of our customer-facing products. Our DevOps... 
    Work from home
    2 days per week

    Reynolds & Reynolds

    Tallahassee, FL
    3 days ago
  • $104.9k - $174.7k

    About the role:A FinOps Site Reliability Engineer (SRE) bridges the gap between engineering, operations, and financial governance by embedding cost optimization into infrastructure design, automation, monitoring, and operational processes. A FinOps SRE proactively identifies... 
    Full time
    Local area

    LexisNexis Risk Solutions Group

    Boca Raton, FL
    4 days ago
  • $130k - $150k

     ...Cloud SolutionsInformation SecurityInformation Technology staff are based in the Boston, Chicago, London, Munich, New York,...  ...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are... 
    Work at office
    Work from home
    3 days per week

    CRA International

    Boston, MA
    3 days ago
  •  ...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient... 
    Ongoing contract
    Full time
    Temporary work
    Work experience placement
    Worldwide

    SS&C Technologies

    Waltham, MA
    4 days ago
  •  ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering...  ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago
  • $167.7k - $245.2k

     ...very effective.We’re looking for talented engineers with a software or operations background...  ...development teams to ensure the reliability, performance and security of our infrastructure...  ...insurance. Please see the Cisco careers site to discover more benefits and perks.... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    1 day per week

    CISCO Systems

    Austin, TX
    3 days ago
  • $114.3k - $235.32k

     ...verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business.We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven... 
    Work at office
    Local area
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    11 hours ago
  • $119.8k - $234.7k

     ...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual...  ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewAre...  ...no further than the Microsoft Defender engineering team. We are looking for a Senior Site... 
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Redmond, WA
    11 hours ago
  • Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firm’s most critical customer-facing microservices that power all eCommerce channels. This role applies Google-inspired SRE principles to balance... 
    Local area
    Remote work
    Flexible hours
    Shift work

    O'Reilly Auto Parts

    Springfield, MO
    11 hours ago
  •  ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises... 
    Permanent employment
    Full time
    Part time
    H1b
    Work at office
    Local area
    Immediate start
    Work visa
    Monday to Friday
    Shift work
    Day shift

    Truist

    Atlanta, GA
    3 hours ago
  • Play a key role in ensuring system reliability at one of the world’s most iconic and largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will use technology to solve... 

    JP Morgan Chase

    Chicago, IL
    11 hours ago
  • $119.8k - $234.7k

     ...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual...  ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...opportunity for a Senior Site Reliability Engineer (SRE) to join the Azure Silver and Sovereign... 
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Reston, VA
    11 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!