Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

$140k - $180k

Endear

Our Story

At Endear, we’re building a modern CRM for retail teams—starting with the frontline. Our software helps sales associates have more personal, effective customer conversations through AI-powered tools that drive measurable revenue.

Despite retail being significantly larger than eCommerce, most software overlooks the in-store experience. Endear helps brands turn real customer relationships into growth through intuitive software and thoughtful design.

As Endear grows, we are investing in the reliability and platform systems that keep our product fast, resilient, and easy for engineering teams to operate.

Position Overview

We’re hiring a Senior Site Reliability Engineer to become Endear’s first dedicated reliability hire. This is a hands-on builder role for someone who wants to solve the underlying systems problems that create on-call burden—not simply respond to pages.

You will own the work of making Endear’s systems more observable, reliable, and scalable. You’ll investigate recurring incidents, drive root-cause fixes, improve alerting and incident playbooks, and partner with engineers on database, queue, and event-processing reliability.

You will also help establish a nearshore triage layer for routine, well-understood issues, so product engineers can spend more time building product and less time on operational interruptions.

What You’ll Accomplish
In your first 6 months, you’ll…
  • Build a clear view of Endear’s highest-impact reliability risks, recurring incidents, and on-call pain points.

  • Establish a prioritized reliability backlog and drive root-cause fixes for the most important issues.

  • Improve alert quality, severity definitions, escalation paths, and runbooks for common incidents.

  • Strengthen observability, queue health, database capacity planning, and operational readiness ahead of peak retail periods.

  • Create the foundation for a nearshore triage process for low-priority, repeatable issues.

In your first year, you’ll…
  • Make on-call materially quieter and less disruptive for product engineers.

  • Build scalable systems for observability, alerting, incident response, database reliability, and queue/event-processing health.

  • Own and improve the nearshore triage relationship, playbooks, and escalation process.

  • Help establish the technical roadmap and future resourcing plan for Endear’s broader platform and reliability function.

You’ll Thrive in This Role If You…
  • Have deep hands-on experience with Kubernetes, production databases, and event-driven systems.

  • Have operated and improved high-volume production systems with meaningful reliability, performance, and data-scale requirements.

  • Enjoy finding root causes, fixing repeat incidents, and building tooling that makes engineers’ lives easier.

  • Have experience with observability, alerting, incident response, capacity planning, and operational runbooks.

  • Can work effectively as a senior IC: owning complex technical work directly while coordinating across teams.

  • Are comfortable in a lean environment where priorities move quickly and you will need to make practical trade-offs.

  • Bring experience from a scaling, mid-size company rather than only an early-stage startup or hyperscaler environment.

  • Have GCP or ClickHouse experience, which are strong pluses.

About the Team

You’ll partner with:

  • JP Grace, CTO: Align on reliability priorities, technical risks, and the roadmap for improving on-call health.

  • Engineering team: Partner on root-cause fixes, platform improvements, and architecture decisions that improve reliability.

  • Nearshore triage partner: Build and maintain playbooks, escalation paths, and expectations for routine incident handling.

  • Product and Support: Help ensure issues are surfaced, prioritized, and resolved with the right level of urgency.

Endear is a lean, remote team where individuals have broad ownership. This role will directly shape how the company handles production reliability as it grows.

Our Hiring Process
  1. Recruiter screen — 30 minutes

  2. Behavioral interview with CTO — 60 minutes

  3. Technical panel with Engineering — 60 minutes

  4. Final conversation with Co-Founders

  5. Offer

Compensation & Benefits
  • Base salary: $140,000-180,000

  • Fully remote, U.S.-based role

  • Comprehensive healthcare, including medical, dental, and vision, plus a 401(k) plan

  • Monthly stipend for co-working and home-office setup

  • Flexible PTO and unlimited vacation

  • Opportunity to build Endear’s first dedicated reliability function from the ground up

Apply Even If You Don’t Check Every Box

If this role excites you but you’re unsure if you meet every requirement, reach out anyway. We care about skills, motivation, and how you think more than perfect resumes.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in United States vacancy
  •  ...Job Summary We are seeking a Senior SRE / DevSecOps Engineer with strong experience in Kubernetes, AWS...  ...troubleshooting. The role will focus on platform reliability, incident management, SLO/SLI...  ...with SLO/SLI governance and site reliability practices. ~ Strong understanding... 
    Senior
    Contract work

    PB consulting

    Charlotte, NC
    a month ago
  • $174k - $252k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 
    Senior

    Google

    Sunnyvale, TX
    3 days ago
  • $152.5k - $205k

     ...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind... 
    Senior
    Flexible hours

    Circle

    San Francisco, CA
    2 days ago
  •  ...importance of in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location(s).As a Site Reliability Engineer supporting the Cashiering organization, you will play a critical role in ensuring the stability,... 
    Senior
    Full time
    Work at office

    The Charles Schwab Corporation

    Southlake, TX
    2 days ago
  • $160k - $240k

     ...millions of times a day - quickly, reliably, and securely. Any time you...  ...at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our...  ...operations or DevOps at a mid-to-senior level.Strong shell scripting... 
    Senior
    Full time

    Fiserv

    Sunnyvale, CA
    2 days ago
  •  ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, Kubernetes, Terraform, CloudFormation, Ansible, CloudWatch, Grafana, Datadog, OpenSearch, PagerDuty Requirements: Senior SRE... 
    Senior
    Full time
    Remote work

    CyberArk

    United States
    4 days ago
  • The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams... 
    Senior
    Full time
    Work at office
    Local area

    Castleton Commodities International

    Houston, TX
    2 days ago
  •  ...The Role:GIPHY is seeking a highly experienced Site Reliability Engineer to join our SRE team. You will help design, build, operate, and evolve the infrastructure that powers GIPHY, including our cloud environment, Kubernetes clusters, and CI/CD platforms.You will also... 
    Senior
    Full time
    Work experience placement
    Remote work

    Shutterstock

    New York, NY
    4 days ago
  • $80k - $140k

    Job DescriptionRBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and... 
    Senior
    Full time
    Flexible hours
    Shift work

    Royal Bank of Canada

    Minneapolis, MN
    3 days ago
  • $106.7k - $177.9k

    OverviewJoin M&T Bank's Digital Banking organization and help drive the reliability, resiliency, and performance of the platforms our customers depend on every day. As a Senior Software Engineer, Site Reliability Engineering (SRE), you will play a key role in supporting... 
    Senior
    Permanent employment
    Full time
    Work experience placement

    M&T Bank

    Wilmington, DE
    4 days ago
  • $108k - $216k

     ...PermanentCompany: VizioBusiness Segment: Home OfficePosition: Senior Site Reliability EngineerJob Location: 39 Tesla, Irvine, CA 92618Duties:...  ...to customer support experiences. Collaborate with engineering teams to embed reliability into the software development lifecycle... 
    Senior
    Full time
    Temporary work

    DCC Technology

    Irvine, CA
    1 day ago
  • $104.9k - $174.7k

     ...SRE role is responsible for improving the reliability, availability, performance, and...  ...actions through completion.Follow up with engineering, development, security, support, and business...  ...Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering... 
    Senior
    Full time
    Local area

    RELX Group

    San Jose, CA
    1 day ago
  •  ...US Corp. is seeking a Lead Site Reliability Engineer to spearhead our mission of delivering highly available and performant systems. With an average of over 12 years of industry experience, the successful candidate will bridge the gap between software development and systems... 
    Senior

    Axiom Pursuits

    San Francisco, CA
    22 hours ago
  •  ...principles to see it in full.About the teamThe Engineering team at Airwallex is a diverse group of...  ..., working together to build scalable, reliable, and secure products that empower...  ...our Global services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work... 
    Senior
    Temporary work
    Local area

    Airwallex

    San Francisco, CA
    2 days ago
  •  ...in Ausin, TX** Our Opportunity: We are looking for a skilled engineer with disciplines that incorporate aspects of software systems...  ...applications — including AI/ML-driven approaches to observability and reliability. What you’ll do: • Evangelize SRE mindset and solve problems... 
    Senior

    Mindlance

    Austin, TX
    1 day ago
  • $134k - $170k

     ...independent system operator responsible for ensuring the safe and reliable flow of electricity in our region and planning for the...  ...of New England’s ongoing transition to clean energy. The Senior Site Reliability Engineer (SRE) is a hands-on engineering role responsible for... 
    Senior
    Permanent employment
    H1b
    Relocation
    Visa sponsorship
    Work visa
    Relocation package
    Flexible hours
    3 days per week

    ISO New England

    Holyoke, MA
    3 days ago
  • Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firm’s most critical customer-facing microservices that power all eCommerce channels. This role applies Google-inspired SRE principles to balance... 
    Senior
    Local area
    Remote work
    Flexible hours
    Shift work

    O'Reilly Auto Parts

    Springfield, MO
    3 days ago
  •  ...Senior Site Reliability Engineer Company: Sphera Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Terraform, ARM templates, Kubernetes, Azure, SonarCloud, CheckPoint, Hadoop, Kafka, Presto, NewRelic, CI/CD, Linux, Windows, Redis... 
    Senior
    Full time
    Remote work

    Sphera

    United States
    4 days ago
  • $95.3k - $158.8k

    Senior Site Reliability Engineer Are you passionate about building resilient, scalable systems that power mission-critical applications?Do you thrive on automating operations, improving reliability, and ensuring exceptional system performance?About the team:Embedded Innovation... 
    Senior
    Full time
    Local area
    Remote work
    Work from home

    Elsevier

    Philadelphia, PA
    3 days ago
  •  ...As a Senior Site Reliability Engineer, you will help design, deploy, maintain, and improve reliable, secure, and scalable infrastructure and services. You’ll proactively identify operational risks and potential failure points, troubleshoot system and application issues... 
    Senior

    Ll Oefentherapie

    Nashville, TN
    2 days ago
  • $139.3k - $203.6k

     ...FedRAMP team builds and operates secure, reliable cloud services for U.S. government...  ...customers. We partner closely with application engineering, security, compliance, and...  ...insurance. Please see the Cisco careers site to discover more benefits and perks. Employees... 
    Senior
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours

    CISCO Systems

    Boxborough, MA
    2 days ago
  •  ...Discover exciting DevOps job opportunities and connect with 28,396 DevOps professionals. The Senior Site Reliability Engineer role at Jobicy is designed for experienced professionals who are passionate about enhancing system reliability and operational efficiency.... 
    Senior
    Remote work
    Flexible hours

    DevOpsChat

    Eastern, KY
    22 hours ago
  • $81.1k - $187k

     .... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers...  ...posted.Career Level - IC3Escalation points for junior Site Reliability Engineers during complex or high-impact incidents.Manage and... 
    Senior
    Temporary work
    Monday to Friday
    Flexible hours
    Shift work
    Night shift

    Oracle Corporation

    Reston, VA
    22 hours ago
  •  ...developer-tooling company whose product is used by engineering teams at thousands of software companies for...  ...commitments and the SRE team is a senior, well-resourced group of nine. As Senior SRE you will lead reliability initiatives across the platform — from defining... 
    Senior

    Kovoro

    Eastern, KY
    22 hours ago
  •  ...Job Title Location Remote - United States Job Category Information Technology, Platform Engineering, Site Reliability Engineering Industry Computer Software, SaaS, National Security Employee Type FT Exempt Manage Others No Minimum Experience 5 Years... 
    Senior
    Remote work

    CenCore

    United States
    22 hours ago
  • $108k - $216k

     ...Senior Site Reliability Engineer Senior Site Reliability Engineer professional opening available at Wal-Mart in Irvine, CA. Qualifications Master's or equiv in CS, Comp Eng'g, Comp Info Systs, SW Eng'g, Electrical Eng'g, Info Systs Security, or rel.... 
    Senior
    Temporary work

    Energy Jobs Network

    Irvine, CA
    1 day ago
  •  ...We need a Senior SRE to ensure Xident's verification platform runs at 99.99% uptime....  ...SLOs, incident response, and production reliability for a system that processes millions of...  ...structured logging Implement chaos engineering practices to proactively identify failure... 
    Senior
    Remote work

    Xident B.V.

    Union, NJ
    22 hours ago
  • Senior Cloud Engineer Long term contract- 2+ years 100% remote in the continental US Our client, a premier national healthcare provider, is currently looking for a Senior Site Reliability Engineer on their Cloud Infrastructure team. You will be helping to maintain... 
    Senior
    Long term contract
    Remote work

    ASCENDING LLC

    United States
    4 days ago
  • $262k - $364k

     ...infrastructure from SRE side, ensuring it is reliable, scalable, cost effective and performant, while working closely with senior technical leads in the development teams....  ...:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software... 
    Senior

    Google

    Sunnyvale, CA
    4 days ago
  •  ...Senior Site Reliability Engineer Company: ZetaChain Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Go, Python, Bash, Terraform, Ansible, Kubernetes, Docker, Linux, Prometheus, Grafana, Datadog, Loki, incident.io, AWS, GCP, Bare... 
    Senior
    Full time
    Remote work

    ZetaChain

    United States
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!