Senior Site Reliability Engineer
$140k - $180kEndear
Senior Site Reliability Engineer
We're hiring a Senior Site Reliability Engineer to become Endear's first dedicated reliability hire. This is a hands-on builder role for someone who wants to solve the underlying systems problems that create on-call burden—not simply respond to pages.
You will own the work of making Endear's systems more observable, reliable, and scalable. You'll investigate recurring incidents, drive root-cause fixes, improve alerting and incident playbooks, and partner with engineers on database, queue, and event-processing reliability.
You will also help establish a nearshore triage layer for routine, well-understood issues, so product engineers can spend more time building product and less time on operational interruptions.
What You'll Accomplish
In your first 6 months, you'll…
- Build a clear view of Endear's highest-impact reliability risks, recurring incidents, and on-call pain points.
- Establish a prioritized reliability backlog and drive root-cause fixes for the most important issues.
- Improve alert quality, severity definitions, escalation paths, and runbooks for common incidents.
- Strengthen observability, queue health, database capacity planning, and operational readiness ahead of peak retail periods.
- Create the foundation for a nearshore triage process for low-priority, repeatable issues.
In your first year, you'll…
- Make on-call materially quieter and less disruptive for product engineers.
- Build scalable systems for observability, alerting, incident response, database reliability, and queue/event-processing health.
- Own and improve the nearshore triage relationship, playbooks, and escalation process.
- Help establish the technical roadmap and future resourcing plan for Endear's broader platform and reliability function.
You'll Thrive in This Role If You…
- Have deep hands-on experience with Kubernetes, production databases, and event-driven systems.
- Have operated and improved high-volume production systems with meaningful reliability, performance, and data-scale requirements.
- Enjoy finding root causes, fixing repeat incidents, and building tooling that makes engineers' lives easier.
- Have experience with observability, alerting, incident response, capacity planning, and operational runbooks.
- Can work effectively as a senior IC: owning complex technical work directly while coordinating across teams.
- Are comfortable in a lean environment where priorities move quickly and you will need to make practical trade-offs.
- Bring experience from a scaling, mid-size company rather than only an early-stage startup or hyperscaler environment.
- Have GCP or ClickHouse experience, which are strong pluses.
About the Team
You'll partner with:
JP Grace, CTO: Align on reliability priorities, technical risks, and the roadmap for improving on-call health.
Engineering team: Partner on root-cause fixes, platform improvements, and architecture decisions that improve reliability.
Nearshore triage partner: Build and maintain playbooks, escalation paths, and expectations for routine incident handling.
Product and Support: Help ensure issues are surfaced, prioritized, and resolved with the right level of urgency.
Endear is a lean, remote team where individuals have broad ownership. This role will directly shape how the company handles production reliability as it grows.
Our Hiring Process
Recruiter screen — 30 minutes
Behavioral interview with CTO — 60 minutes
Technical panel with Engineering — 60 minutes
Final conversation with Co-Founders
Offer
Compensation & Benefits
Base salary: $140,000-180,000
Fully remote, U.S.-based role
Comprehensive healthcare, including medical, dental, and vision, plus a 401(k) plan
Monthly stipend for co-working and home-office setup
Flexible PTO and unlimited vacation
Opportunity to build Endear's first dedicated reliability function from the ground up
Apply Even If You Don't Check Every Box
If this role excites you but you're unsure if you meet every requirement, reach out anyway. We care about skills, motivation, and how you think more than perfect resumes.
$163k
...and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview-The ProCOM team is looking for a Site Reliability Engineering (SRE) who can help us solve problems, build...SeniorFull timePart timeImmediate startWorldwideFlexible hours$96k - $163k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer, Performance Engineering Senior Site Reliability Engineer, Performance Engineering Payment Optimization unifies...SeniorFull timePart timeWorldwideFlexible hours$96k - $163k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...SeniorFull timePart timeWorldwideFlexible hours$153k - $210k
...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating...SeniorFull time- ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to...SeniorFull time
$170k - $220k
Who We're Looking ForWe’re looking for a hands-on, high-agency Site Reliability Engineer to help shape and scale the reliability layer of our stack. You'll own the release pipeline end-to-end — managing daily releases, weekly deploys, and hotfixes — while also automating...Senior- ...TechMContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8... ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will...SeniorRemote work
- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure...Senior
- Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems thinking, and operational excellence...SeniorWorldwide
- Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with...SeniorFlexible hours
- IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion...SeniorWork at officeImmediate start
$104.9k - $174.7k
About the role:A FinOps Site Reliability Engineer (SRE) bridges the gap between engineering, operations, and financial governance by embedding cost optimization into infrastructure design, automation, monitoring, and operational processes. A FinOps SRE proactively identifies...SeniorFull timeLocal area$152.6k - $191.5k
...responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include... ...and continuous improvement.Position Summary:The Senior GCP Site Reliability Engineer acts as an advanced senior...SeniorFull timeWork at officeDay shift- ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and... ...and networking teams to improve service reliability and deployment workflowsDeploy and... ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering...SeniorWork at officeLocal areaWork from homeFlexible hours
$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...SeniorCurrently hiringRemote work$168k - $270.25k
...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance...SeniorFull time$174k - $252k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you...Senior- ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering... ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or...SeniorWork at officeLocal areaWork from homeFlexible hours
$152.5k - $205k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind...SeniorFlexible hours- We are looking for a Senior or Staff level Site Reliability Engineer to strengthen the reliability, scalability, and operational maturity of our platform in San Francisco, California. This role will focus on improving service health, refining observability, and partnering...Senior
- LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is...SeniorFull timeWork at office2 days per week
$160k - $240k
...millions of times a day - quickly, reliably, and securely. Any time you... ...at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our... ...operations or DevOps at a mid-to-senior level.Strong shell scripting...SeniorFull time- ...professionalism. We are seeking an experienced AWS solution design engineer/architect to join our infrastructure cloud team. The... ...product features efficiently and confidently them into production.As Senior SRE, you will be responsible for providing leadership, design and...Senior
- ...and foster a dynamic work environment where new ideas thrive. Are you ready to join our team and make an impact?As a Senior Site Reliability Engineer at TeamViewer, you’ll be a key player in ensuring the reliability, scalability, and performance of our Azure-based SaaS...SeniorTemporary workCasual workWorldwide
$119.8k - $234.7k
...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewMicrosoft... ...’s most demanding workloads. As a Senior Site Reliability Engineer, you will lead reliability...SeniorOngoing contractLocal area3 days per week$90k - $180k
...nutritionals and branded generic medicines. Our 122,000 colleagues serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We...SeniorRemote work- ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS...SeniorTemporary work
$104.9k - $174.7k
Are you passionate about improving reliability, scalability, and resilience in complex database... ....Own prioritization of reliability engineering tasks within team backlogs.Lead incident... ...a Service (IaaS).Background in DevOps, site reliability engineering practices, or related...SeniorFull timeLocal area$160k - $200k
...data, ideally using promQLKey Responsibilities:Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coachPerform incident...SeniorTemporary workWork at officeLocal areaFlexible hours3 days per week$267k - $356k
...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-... ...workloads in the industry, which means reliability and performance aren't just goals—they're... ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc...SeniorWork experience placementWork at officeLocal areaWork from homeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer United States
- site reliability engineering manager United States
- site reliability engineer sre United States
- site reliability engineer remote United States
- lead site reliability engineer United States
- senior human resources associate United States
- senior network engineer remote United States
- senior education consultant United States
- senior benefits manager United States
- senior app developer United States


