Senior Site Reliability Engineer
$153k - $190kb.well Connected Health
Senior Site Reliability Engineer
This is a full-time role and not open to contract work. This role is open to fully remote work.
Company Overview: b.well is solving healthcare's fragmentation problem with our FHIR-based health data management platform. The platform connects data from EHRs, wearables, portals, and other sources, while our intelligence engine personalizes the consumer experience. By simplifying the complex healthcare ecosystem, we make it easy and convenient for consumers to engage and take action—whether it's scheduling care, setting reminders, accessing health data, and more. For our clients, this means better health outcomes, operational efficiency, and stronger consumer engagement.
Job Description
As a Senior Site Reliability Engineer, you'll make sure we build a reliable, secure, and efficient platform for the b.well network. We run a production healthcare platform entirely on AWS, orchestrated on EKS, and reliability at that scale is an engineering discipline—SLIs, SLOs, error budgets, and relentless automation—not a firefighting hobby.
You'll be encouraged to blog, speak, and join events to talk about the work you're doing and encourage other companies to follow our lead. Automation is a core value here, from reliability to scaling, and you'll be one of the people who sets the bar for it.
What You Will Do:
- Let the robots do the work! Drive down toil by automating everything—infrastructure provisioning, deployments, monitoring, and disaster recovery. Reliability at scale isn't something you do by hand.
- Run the fleet. We're all-in on AWS and orchestrate everything on EKS. You'll keep Kubernetes humming—scaling, upgrades, and all the sharp edges in between.
- Python, Go, Bash—pick your poison. The team ships in all three. We care more about clean, reliable automation than which one you reach for.
- Define what "reliable" means and hold the line on it. SLIs, SLOs, error budgets, incident response, and blameless postmortems—so there are steadily fewer 3am pages for everyone.
- Like the Eye of Sauron, you'll keep watch. Build the dashboards, alerts, and SLOs that catch problems before our customers ever notice.
- CI/CD is your love language. Keep our GitHub Actions pipelines fast, safe, and boring in the best possible way.
- New platform, who dis? Design and build new solutions that move the business forward—then pass the knowledge on to your peers.
- Must feed, water, and provide high-fidelity, low-friction platform tooling to engineers to keep them happy and shipping.
- Write the ancient artifacts of documentation —runbooks, policies, and procedures—so your peers (and future you) know how the environment actually works.
- Share the on-call load. This role includes an on-call rotation for a production healthcare platform. A big part of the job is making on-call suck less through better automation and smarter alerting.
- Get paid to be curious. You like reading about the latest tech and trying it out? Come do it here—and blog, speak, and share what you learn.
Who You Are:
- Deep in the penguin. Whatever your flavor of Linux, you know it cold—from kernel knobs to networking internals.
- Kubernetes is home turf. You've run real production workloads on it (EKS ideally), packaged with Helm, and understand what's happening below kubectl apply.
- Terraforming isn't just sci-fi—it's how you build infrastructure (Terraform, OpenTofu, or Pulumi). Infrastructure-as-code is the only way you'd have it.
- Git is the source of truth. GitOps is core to how we operate—infrastructure, deployments, and config all flow through code and version control (Terraform, Argo, Helm). If something can only be done through ClickOps, you see it as a gap to remediate, not a shortcut to take.
- Cloud native, literally. Strong AWS experience—or equivalent depth in GCP/Azure that you can bring across.
- Observability is second nature. You've run distributed monitoring in a microservices world (Grafana, Datadog, or similar) and can turn a wall of metrics into an actual answer. Our ingest is OpenTelemetry-based with some Prometheus in the mix, so love for either is a real plus.
- RFC isn't just three letters to you. Networking, DNS, firewalls, TLS—the fundamentals are, well, fundamental.
- Watson, is that you? You're a relentless troubleshooter who can chase a problem across services and layers until it confesses.
- Security-minded by default. You think about compliance and least privilege before anyone makes you—bonus if you've worked in a HIPAA, HITRUST, or SOC 2 environment, since we live in healthcare.
- An open source citizen. We love the community and would love for you to contribute back to it.
- Be Awesome. People come to you with problems while they're frustrated. You meet them with patience, clarity, and a fix—and you explain the WHY, not just the what.
Do You Have a Special Set of Skills?
- Hands-on OpenTelemetry instrumentation experience
- HIPAA / HITRUST / SOC 2 or other regulated/compliance experience
- Meaningful open source contributions to infrastructure tooling
The target salary range for this position is $153,000 - $190,000 and is part of a competitive total rewards package including stock options, benefits, and incentive pay for eligible roles. Individual pay may vary from the target range and is determined by a number of factors including experience, location, internal pay equity, and other relevant business considerations. We review all employee pay and compensation programs annually at minimum to ensure competitive and fair pay.
Data shows that women, people of color, and other underrepresented groups may be less likely to apply for jobs unless they believe they are a perfect match. But b.well holds diversity amongst its key values, and we have a strong commitment to building our workforce and products through that lens. You don't have to check every box in this job description to be a great fit for the role! If you're excited about this position and the prospect of working for b.well, please apply. If it turns out this role isn't for you, there may be other openings that could align with your experience and expertise! We are committed to an inclusive and diverse b.well. We are an equal opportunity employer. We do not discriminate based on race, ethnicity, color, ancestry, national origin, religion, sex, sexual orientation, gender identity, age, disability, veteran, genetic information, marital status or any other legally protected status.
$153k - $210k
...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating...SeniorFull time- ...A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure. The ideal candidate has over 8 years of experience in wireless network operations and a strong background in wireless...Senior
- ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, Kubernetes, Terraform, CloudFormation, Ansible, CloudWatch, Grafana, Datadog, OpenSearch, PagerDuty Requirements: Senior SRE...SeniorFull timeRemote work
- ...A tech startup in San Francisco is looking for Site Reliability Engineers to enhance system reliability and performance. Ideal candidates have over 5 years of relevant experience and strong expertise in cloud infrastructure, including AWS and Kubernetes. The role involves...Senior
- ...Apple Service Engineering (ASE) seeks a senior SRE software engineer to own the architectural direction of Kubernetes internals powering Apple services... ...will define controllers and namespace management, raise reliability, and contribute to upstream Kubernetes. The role...Senior
- ...Lambda Inc. in San Francisco is seeking a Storage Engineer to own the reliability, performance, and capacity of our production storage fleet across multiple data centers, using a software-defined data plane. You will build monitoring, dashboards, and alerting for storage...Senior
- ...Sight Machine is seeking a senior Cloud Infrastructure IC to lead reliability, automation, and scale across our platform. You will drive IaC, CI/CD, observability and operate agentic AI systems, mentoring engineers and guiding architectural decisions while staying hands...Senior
- ...data and AI infrastructure provider is seeking a Senior Staff Technical Program Manager for Reliability to enhance the reliability and performance of their... ...involves leading programs in partnership with senior engineering leaders, requiring over 10 years of experience in...Senior
- ...Oracle Cloud Infrastructure (OCI) seeks a Senior Principal Engineer to lead the design and implementation of reliability validation for OCI control plane services, focusing on a high-performance, low-level systems approach. You will mentor engineers, define validation...Senior
- ...Avalara, Inc. is seeking a senior reliability engineer to lead how reliability is engineered across Avalara's global SaaS platform as we move toward an AI-first operating model. You will build a modern, automation-first reliability ecosystem that improves stability,...Senior
- ...Senior Site Reliability Engineer Company: Sphera Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Terraform, ARM templates, Kubernetes, Azure, SonarCloud, CheckPoint, Hadoop, Kafka, Presto, NewRelic, CI/CD, Linux, Windows, Redis...SeniorFull timeRemote work
- ...Senior Sre We're hiring a Senior SRE based in Latin America to work alongside our US-based engineering team, building out observability, on-call coverage, and deployment automation for a client with strict compliance requirements. We're specifically looking for someone...SeniorFull timeRemote work
- ...We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available database... ...data platforms. The role combines database engineering, site reliability engineering, Linux systems administration, and...Senior
$166k - $244k
# Senior Software Engineer, Site Reliability EngineeringGoogle • onsite • 601 N 34th St, Seattle, WA 98103, USA • full\_timePay: USD 166000.00 - USD 244000.00 / unspecifiedBusinesses of all shapes and sizes rely on Google’s unparalleled advertising solutions to help them...SeniorTemporary work- ...compute, storage and build continuous integration, continuous delivery pipeline (CI/CD). Implement strategies that increase system reliability and performance through on-call rotation and process optimization. Add necessary automations for improved collaborative...SeniorRemote work
$152k - $195k
...investors including Silver Lake Waterman, Moody’s, Sequoia Capital, GV and Riverwood Capital. About the Team: As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization of our Kubernetes-based infrastructure and CI/...SeniorRemote work- ...Senior Site Reliability Engineer Teikametrics is seeking a Senior Site Reliability Engineer to enhance and maintain our cloud infrastructure for hosting applications and platforms. This position involves collaborating within a DevOps model to design and deploy automation...SeniorRemote work
- ...Senior Site Reliability Engineer We are looking for a Senior Site Reliability Engineer with Cloud platform experience. This individual will be part of a team responsible for operating and maintaining production clusters and developing our observability solutions; they...SeniorRemote work
- ...and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview-The ProCOM team is looking for a Site Reliability Engineering (SRE) who can help us solve problems, build...SeniorFull timePart timeImmediate startWorldwideFlexible hours
$182.8k - $247.3k
...mission to develop education for our half a billion (and growing!) learners around the world. About the role... As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure Duolingo’s sophisticated distributed...SeniorWork experience placement- ...automated detection, drain/cordon/taint, workload rescheduling. Feed the AIOps substrate The remediation-actuator and workflow engine land here — you make the control plane safe for automated action. Your CRDs are the schema the platform's predictors and...SeniorLocal area
$120k - $175k
...level of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's possible...SeniorFull timeRemote workWork visaFlexible hours- ...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner... ...through efficient, data-driven solutions. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...SeniorWork at officeRemote work
$152.5k - $205k
...environment where new ideas are encouraged and everyone is a stakeholder. What you’ll be responsible for: As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind...SeniorRemote workFlexible hours- Job Posting Datavant recognizes the importance of information security and data privacy, including in its hiring processes and recruitment. Datavant encourages all potential job applicants to take precautions against potential phishing schemes or other scams that improperly...SeniorFixed term contractWork at officeLocal area
- ...Senior Site Reliability Engineer We're looking for an experienced Site Reliability Engineer to help scale and maintain our hosting infrastructure in AWS. You'll work closely with our engineering teams to build secure, scalable systems that keep Framer's websites running...SeniorRemote work
- ...AI Senior SRE MOZN is a leading Enterprise AI company enabling organizations to make... ...doing it. This role exists because most reliability toil (triage, root-causing, remediation,... ...movement, false positive/negative rates, and engineer-hours of toil removed. Keep a...SeniorRemote work
$145k - $175k
...straightforward communication and clinical domain expertise, Commence cuts straight to better care. Requirements As a Senior Site Reliability Engineer at Commence, you will own the reliability, scalability, and operational health of our mission-critical healthcare data...SeniorFull timeRemote work$160k - $180k
...have a big impact. See Arkestro in action at arkestro.com. About the Role Arkestro is hiring for a Senior SRE Engineer to manage our performance and reliability for our software platform and infrastructure. The right candidate will own and develop our...SeniorLocal areaRemote work$115k - $160k
...collaborating with Barclays to connect them with exceptional professionals for this role. Embark on a transformative journey as a Senior Site Reliability Engineer - AVP - Credit Trade Floor. At Barclays, our vision is clear –to redefine the future of banking and help craft...SeniorHourly payWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre United States
- site reliability engineering manager United States
- site reliability engineer United States
- site reliability engineer remote United States
- senior technical service engineer United States
- senior functional safety engineer United States
- senior technical consultant United States
- senior director product management United States
- senior vendor manager United States
- senior vice president human resources United States


