Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Sr. Reliability Engineer

HIMS Inc.

Hims & Hers is the leading health and wellness platform, on a mission to help the world feel great through the power of better health. We are redefining healthcare by putting the customer first and delivering access to care that is affordable, accessible, and personal, from diagnosis to treatment to delivery. No two people are the same, so we provide access to personalized care designed for results. By normalizing health & wellness challenges and innovating on their solutions, we’re making better health outcomes easier to achieve.

Hims & Hers is a public company, traded on the NYSE under the ticker symbol “HIMS.” To learn more about the brand and offerings, you can visit hims.com/about and hims.com/how-it-works . For information on the company’s outstanding benefits, culture, and its talent-first flexible/remote work approach, see below and visit

About the Role:

We're looking for a Senior Reliability Engineer to make Hims & Hers systems measurably more reliable: building the observability, tooling, and automation that catch problems before people do. Recent work at this level includes preparing our stack for 32x baseline traffic during high-stakes seasonal surges with zero customer-facing issues and improved P95 latency, turning databases into a paved platform with consistent observability and guardrails, and building AI agents that pull FireHydrant, Datadog, Jira, and Confluence into a single incident picture. We use AI first: every engineer gets a Claude Enterprise license, and we expect you to use it as a core part of how you investigate, build, and write, not as an afterthought.

You Will:
  • Own reliability for Tier 1 customer journeys: Define and instrument SLOs, golden signals, and business-level monitors for the journeys that matter most (checkout, telehealth visits, prescription fulfillment), so that degradation is detected by systems rather than by customers or support tickets. Move our "detected by monitors vs. detected by humans" ratio in the right direction and be able to show it.

  • Engineer for peak load and failure: Lead capacity and resilience work for high-stakes events and steady-state growth: load-test design, second-by-second analysis of prior events, database and backend bottleneck investigation, and hardening across caching, GraphQL, VPC capacity, and vendor rate limits. Validate before game day, not during it.

  • Build incident response as software: Mature FireHydrant, Datadog, and Jira into one connected pipeline: automated incident and RCA ticket creation, SLO burn-rate and composite alerting with deduplication, Tier 1 alert routing, and enforced post-mortem action tracking. Author and maintain the runbooks and severity standards that make any responder effective on any service.

  • Automate operational excellence with AI: Build and operate agents and tooling that reduce manual OE work: OER report generation, RCA drafting, stale action-item detection, monitor and runbook gap detection, and OpenClaw agents wired to Datadog and FireHydrant for first-pass incident triage. Ship these as reusable capabilities, not personal scripts.

  • Be the deep-debugging expert, and make teams better at it: Teams own debugging their own services, but you are the person they pull in when a problem crosses boundaries: silent service-to-service failures, anomalous traffic, or regressions that span frontend, API, mesh, and database layers. Bring that depth to the hardest cases yourself, then turn what you learn into runbooks, tooling, and pairing so teams can catch and resolve the next one on their own. Partner with Security and product teams on anomaly detection and response, weighing engineering cost and user impact alongside the benefit of any control.

  • Maintain the tooling that informs Operational Excellence reviews: Produce the tooling that powers the metrics and narrative that go into bi-weekly VP-level OE reviews and the monthly cross-engineering OER, and drive the resulting action items to closure.

  • Raise the bar for others: Document what you build, onboard teammates to roll it out, and coach engineers across squads on SLOs, blameless post-mortems, and on-call practice.

You Have:

  • 5+ years as a Software, SRE, Platform, or Infrastructure Engineer, with a track record of owning reliability outcomes for production systems that customers depend on.

  • Strong software engineering fundamentals. You solve reliability problems by writing code and building tooling, and you're comfortable reading application code across the stack to find the real cause.

  • Hands-on depth in observability and SLO engineering: golden signals, burn-rate alerting, journey-level monitors, and turning noisy alert streams into actionable pages (Datadog preferred; Prometheus/Grafana and OpenTelemetry welcome).

  • Production experience with AWS, Kubernetes/EKS, Terraform, and PostgreSQL (RDS/Aurora).

  • Experience running or maturing incident management end to end: on-call design, escalation policies, incident command, blameless post-mortems, and action-item follow-through in a tool like FireHydrant or PagerDuty.

  • Daily, practical use of AI coding and analysis tools (Claude, Cursor, or similar) to accelerate investigation, code, documentation, and reporting, and clear judgment about when to trust the output and when to verify it.

  • Communication skills to explain risk, tradeoffs, and post-incident learnings to engineers and leadership alike, in writing and in review meetings.

Preferred Qualifications:
  • Experience building AI agents or LLM-backed automation for operations: incident triage, RCA generation, alert correlation, or observability data analysis, ideally with MCP or tool-calling integrations against Datadog, FireHydrant, or Jira.

  • Load-testing and performance-engineering experience at meaningful scale (k6 or similar), including validating systems to well above expected peak.

  • Familiarity with service mesh (Istio) and its observability and traffic-management features.

  • Background in a regulated or healthcare environment, where reliability and data handling carry patient-safety and compliance weight.

  • Experience designing vendor and partner escalation frameworks with defined severities and response SLAs.

Our Benefits (there are more but here are some highlights):
  • Competitive salary & equity compensation for full-time roles

  • Unlimited PTO, company holidays, and quarterly mental health days

  • Comprehensive health benefits including medical, dental & vision, and parental leave

  • Employee Stock Purchase Program (ESPP)

  • 401k benefits with employer matching contribution

  • Offsite team retreats

We are committed to building a workforce that reflects diverse perspectives and prioritizes ethics, wellness, and a strong sense of belonging.

Hims considers all qualified applicants for employment, including applicants with arrest or conviction records, in accordance with the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance, the California Fair Chance Act, and any similar state or local fair chance laws.

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Hims & Hers is committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, please contact us at View email address on click.appcast.io and describe the needed accommodation. Your privacy is important to us, and any information you share will only be used for the legitimate purpose of considering your request for accommodation. Hims & Hers gives consideration to all qualified applicants without regard to any protected status, including disability. Please do not send resumes to this email address.

To learn more about how we collect, use, retain, and disclose Personal Information, please visit our Global Candidate Privacy Statement.

#J-18808-Ljbffr
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Sr. Reliability Engineer in Eastern, KY vacancy
  •  ...This is an engineering-first Senior SRE role. We’re looking for senior engineers who have: Built and shipped significant backend...  ...services end-to-end in production (design → launch → on-call → reliability improvements) Led incident response and driven durable... 
    Senior

    Practice by Numbers

    Eastern, KY
    3 days ago
  • $110k - $145k

     ...global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is... 
    Senior
    Flexible hours

    Hirebridge

    Eastern, KY
    1 day ago
  •  ...Blackpoint Cyber is in hyper-growth mode, fueled by a recent $190m series C round. SUMMARY We're hiring a Senior Site Reliability Engineer to design, implement, and maintain our cloud and on-premise infrastructure and CI/CD pipelines, with a focus on automation,... 
    Senior
    Local area

    Blackpoint Cyber

    Eastern, KY
    1 day ago
  • $160k - $180k

    # Sr. Site Reliability EngineerUnited States18 hours agoID 1399715Price on request## DetailsEmployment type: Full-timeRemote: YesCompany: CentralReachLevel...  ..., and world-class customer satisfaction. The Platform Engineering group at CentralReach builds the underlying technologies... 
    Senior
    Full time
    Worldwide

    Bazeta

    Eastern, KY
    2 days ago
  • $104k - $178k

    ## Sr. Site Reliability Engineer IApply: Hybrid: NYC Global HQ: Full time: Posted 12 Days Ago: JR00000779# ****Who We Are****DV is the leader in digital performance solutions, helping our advertiser and agency partners Verify the quality of their digital campaigns, Optimise... 
    Senior
    Full time

    DoubleVerify

    Eastern, KY
    1 day ago
  • $130k - $180k

     ...platform in the world. What you’ll do: As a Sr. Database ReliabilityEngineer II, you...  ...(DBA) role. We’re looking for a software engineer with deep MySQL/Postgres expertise who enjoys...  ...schemas, and improving application reliability. In this role you will: Design and evolve... 
    Senior
    Remote job
    Work at office
    Worldwide
    Monday to Friday
    Flexible hours

    NinjaTrader Clearing, LLC

    Eastern, KY
    2 days ago
  • ## Sr Network Reliability EngineerApplylocations: Los Angeles, CA: Space Coast, FL: Denver, CO: Reston, VAtime type: Full timeposted on: Posted...  ...to safe, reliable spaceflight.As a Ground Network Engineer, you will be at the forefront of designing, building, and maintaining... 
    Senior
    Permanent employment
    Temporary work
    Local area
    Flexible hours
    Night shift

    Blue Origin

    Eastern, KY
    5 days ago
  •  ...research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality. As the Senior Reliability Engineer, you own reliability as an engineering discipline, not just a test outcome. Every mission profile we define, every... 
    Senior

    RHODA

    Eastern, KY
    1 day ago
  • $170k

     ...harnesses, structural limbs, and dynamic end effectors—to execute complex tasks in demanding environments. We are seeking a Senior Reliability Engineer with 5–8+ years of hands‑on experience in mechatronics and electromechanical systems to join our team in the San... 
    Senior
    Full time
    Temporary work
    Work at office
    Relocation package
    Flexible hours

    Hardware FYI

    Eastern, KY
    2 days ago
  • $120k - $150k

     ...query tuning high availability This job requires a minimum of 3 years of experience. About the Job Senior Database Reliability Engineer: Job Type: Full-time Location: Remote Job Summary: Join our team as a Senior Database Reliability Engineer, where... 
    Senior
    Full time
    Local area
    Remote work

    Artha Nexgen

    Eastern, KY
    1 day ago
  •  ...ID.me, Inc. is seeking a Site Reliability Engineer to join the Core Platform Engineering team. The role focuses on automation, observability, and reliability of ID.me services with automated guardrails and scalable architectures. The position is based in Mountain View... 
    Senior
    Full time
    Work at office

    ID.me Inc

    Eastern, KY
    2 days ago
  •  ...Vida is seeking a dedicated Site Reliability Engineer to join the Enablement Team. You’ll own production systems, consolidate Terraform patterns, and scale infrastructure for enterprise launches in a fully remote role. You’ll work with Terraform, GCP (GKE, Cloud SQL... 
    Senior
    Remote work

    Lever

    Eastern, KY
    4 days ago
  •  ...Discover exciting DevOps job opportunities and connect with 28,396 DevOps professionals. The Senior Site Reliability Engineer role at Jobicy is designed for experienced professionals who are passionate about enhancing system reliability and operational efficiency.... 
    Senior
    Remote work
    Flexible hours

    DevOpsChat

    Eastern, KY
    1 day ago
  •  ...profitable developer-tooling company whose product is used by engineering teams at thousands of software companies for application...  ...well-resourced group of nine. As Senior SRE you will lead reliability initiatives across the platform — from defining and driving SLOs... 
    Senior

    Kovoro

    Eastern, KY
    1 day ago
  •  ...% uptime. You'll own SLOs, incident response, and production reliability for a system that processes millions of identity verifications...  ...Sentry error tracking, structured logging Implement chaos engineering practices to proactively identify failure modes Optimize... 
    Senior
    Remote work

    Xident B.V.

    Eastern, KY
    1 day ago
  •  ...match. The role We're looking for a Senior SRE to own the reliability, scalability, and operational posture of Satsuma's multi-...  ...using AI-assisted development workflows Partner closely with engineering on reliability reviews and architecture decisions ~5-8... 
    Senior

    Satsuma AI, Inc.

    Eastern, KY
    1 day ago
  •  ...As a Senior Site Reliability Engineer on our cloud engineering team, you'll keep our production environment healthy, secure, and running smoothly. This is an operations-focused role: you'll own the day-to-day administration of our AWS accounts and databases, backup posture... 
    Senior
    Work experience placement

    MeridianLink

    Eastern, KY
    1 day ago
  •  ...Cloudflare, GitHub Actions, PostgreSQL, Redis/BullMQ, Node.js/NestJS, Datadog, TypeScript, React, SQL Position: Senior Site Reliability Engineer Engagement period: Ongoing Interview timeline: ASAP Interview process: 1) CV review 2) Interview with our CTO 3)... 
    Senior
    Contract work
    Immediate start

    Devspace

    Eastern, KY
    1 day ago
  •  ...practical to use as money at a global scale. Our team includes engineers and researchers from organizations such as Blockstream,...  ...and blockchain projects. Role Overview As a Senior Site Reliability Engineer (SRE), you will lead the design, scalability, and reliability... 
    Senior

    Alpen Labs Inc.

    Eastern, KY
    1 day ago
  •  ...remoteseniorfull-time# Senior Site Reliability EngineerLaravelRemote / Argentina, Brazil, Denmark, Portugal, United Kingdom Salary not...  ...ship their dreams. We are looking for a Senior Site Reliability Engineer to help us scale that mission by ensuring our global infrastructure... 
    Senior
    Remote work

    LaraBench

    Eastern, KY
    1 day ago
  •  ...democratizing software development by removing traditional barriers to application creation. About the role: Join our Site Reliability Engineering team and help ensure the reliability, scalability, and performance of Replit's infrastructure that serves millions of... 
    Senior
    Full time
    Temporary work
    Work at office
    Worldwide
    Flexible hours

    Replit

    Eastern, KY
    1 day ago
  •  ...Quarterhill is seeking a Senior Site Reliability Engineer (SRE) to join our growing team. This role is an exciting opportunity to contribute to the reliability and performance of smart transportation systems, including a next-generation, cloud-native tolling platform... 
    Senior
    Local area

    Electronic Transaction Consultants

    Eastern, KY
    1 day ago
  • $180k - $230k

     ...Acceleration Job Description We're looking for a Senior SRE to own the reliability, scalability, and observability of our production systems. You'll work closely with platform and data engineering to keep high-throughput, data-intensive services running at the... 
    Senior
    Work at office
    Local area
    Immediate start
    Remote work
    3 days per week

    GridCARE

    Eastern, KY
    1 day ago
  • $104.9k - $174.7k

     ...Technology Senior Site Reliability Engineer II The SRE role is responsible for improving the reliability, availability, performance, and operational quality of production systems. This role provides technical input into project plans, schedules, methodologies, and... 
    Senior
    Temporary work
    Local area

    LexisNexis Risk Solutions

    Eastern, KY
    4 days ago
  • $150k - $220k

     ...Senior Site Reliability EngineerJob detailsDepartment / EngineeringRemoteFull-time$150,000 USD - $220,000 USD## About UsMetaRouter is a...  ...their architecture.## About The RoleAs a Senior Site Reliability Engineer, you own significant pieces of our infrastructure and... 
    Senior
    Full time
    Remote work

    Deel

    Eastern, KY
    2 days ago
  • $148.5k - $223.9k

     ...Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations, this organization provides a global team of engineers monitoring cloud service... 
    Senior
    Worldwide
    Weekend work

    Salesforce.Com Inc

    Eastern, KY
    1 day ago
  • $180k - $200k

     ...Summary Come join tastytrade, part of IG Group, as we build the reliability practice behind the brokerage platform that active options,...  ...on every market day. As our first Senior Site Reliability Engineer, you'll define what reliability means at tastytrade, from customer... 
    Senior
    Work at office
    3 days per week

    tastyworks

    Eastern, KY
    1 day ago
  •  ...company: ~ Fleetio overview video: ( ~ Our careers page: ( Description Our Platform Engineering team is looking for a Senior Site Reliability Engineer to help run, maintain, and improve the performance of our Ruby on Rails Stack and Infrastructure.... 
    Senior
    Contract work
    Temporary work
    Remote work
    Worldwide

    Wwshemi

    Eastern, KY
    2 days ago
  •  ...phase balancing, and grounding meet the standards required for reliable charger and vehicle operation. Assess as-built construction...  ...remediation. Work with Highland's Construction Electrical Engineering and Procurement teams to build field feedback and reliability... 
    Temporary work
    Remote work

    NextGenEnergyJobs

    Eastern, KY
    3 days ago
  •  ...are the foundation that make us successful for ourselves, our customers and the planet.Albemarle is hiring for a **Brinefield Reliability Engineer**. This position is located in Magnolia, AR, on-site/in-office.The core work for this role involves serving as key technical... 
    Work at office

    Albemarle

    Eastern, KY
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Sr. Reliability Engineer. Be the first to apply!