Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Lead Site Reliability Engineer

$123k - $154k

Gifthealth

Lead Site Reliability Engineer (SRE)

At Gifthealth, we're revolutionizing the way people experience healthcare by simplifying the process of managing prescriptions and health services. Our mission is to provide a seamless, personalized, and efficient healthcare experience for all our customers. We're a dynamic, innovative, and customer-centric company dedicated to making a positive impact on people's lives.

Position Summary

Reporting to the Director of Engineering, the Lead Site Reliability Engineer (SRE) is a senior technical contributor responsible for building reliable, scalable software systems and the DevOps practices that support them. This role blends software engineering, operational excellence, and automation to improve the performance and resilience of Gifthealth's applications.

We are seeking a Lead SRE to play a key part in enabling fast, safe delivery of customer-facing features. This role partners closely with product and application engineers to embed reliability, observability, and operational ownership directly into the development lifecycle, ensuring alignment with organizational goals, operational excellence, and compliance standards.

Key Responsibilities
  • Designs, builds, and maintains reliable, scalable software systems supporting Ruby on Rails applications
  • Embs reliability, performance, and operational best practices into application code and development workflows
  • Owns DevOps practices including CI/CD reliability, deployment strategies, and release safety
  • Leads incident response, debugging, and root cause analysis across application and platform layers
  • Implements and evolves observability (logging, metrics, tracing) within application and service code
  • Partners with engineering teams on architecture, capacity planning, and technical standards
Qualifications
  • Education: Bachelor's degree in computer science, engineering, or related field OR
  • equivalent professional experience in software engineering, SRE, or DevOps roles (Required)

    • Licensure/Certification:
    • Cloud platform certifications (AWS, GCP, Azure) (Preferred)
    • SRE or DevOps-focused certifications (Preferred)
    • Experience:
    • 5+ years of experience in software engineering, SRE, or DevOps roles (Required)
    • Hands-on experience building and operating Ruby on Rails applications in production (Required)
    • Experience in owning production incidents and application-level reliability (Required)
    • Experience in high-growth or scaling engineering organizations (Preferred)
    • Experience working in regulated or customer-impact–sensitive environments (Preferred)
    • Knowledge, Skills, & Abilities:
    • Knowledge of Ruby on Rails application architecture and production operations; software reliability engineering principles (SLOs, SLIs, error budgets); and modern DevOps and CI/CD practices (Required)
    • Knowledge of security and compliance considerations in production systems (Preferred)
    • Strong software engineering skills (Ruby and/or comparable backend languages) (Required)
    • Debugging and performance optimization of production applications skills (Required)
    • CI/CD pipelines, deployment automation, and release tooling skills (Required)
    • Monitoring and observability tooling (Datadog, New Relic, Prometheus, etc.) skills (Required)
    • Infrastructure as Code (Terraform or similar) skills (Preferred)
    • Containerization and orchestration (Docker) skills (Preferred)
    • Ability to write production-quality code that improves system reliability (Required)
    • Ability to collaborate with product and engineering teams to influence design decisions (Required)
    • Ability to troubleshoot complex, cross-system failures (Required)
    • Ability to mentor engineers on operational ownership and reliability practices (Preferred)
    • Ability to balance speed of delivery with long-term system health (Preferred)

    Work Environment

    • Location: Remote
    • Schedule: 8:00 A.M. to 5:00 P.M. Monday through Friday with night and weekend hours on occasion as determined by the needs of the business.
    • Regular meetings with internal Backend and Full-Stack Engineers, Engineering Managers, and Product and Security teams. This role may also have meetings with external cloud and tooling vendor representatives.
    Key Essential Functions
    • Must be able to remain in a stationary position for extended periods while writing or reviewing documentation
    • Must be able to work on a computer for the entire shift
    • Must be able to attend virtual meetings with cross-functional teams.
    Employment Classification

    Status: Full-time FLSA: Exempt

    Equal Employment Opportunity (EEO) Statement

    Gifthealth is an Equal Opportunity Employer and prohibits discrimination and harassment of any kind. All employment decisions are made without regard to race, color, religion, sex, sexual orientation, gender identity, transgender status, national origin, age, disability, veteran status, or any other legally protected status.

    We celebrate diversity and are committed to creating an inclusive environment for all employees. If you do not meet every requirement but still feel you would be a great fit for this role, we encourage you to apply!

    Disclaimer

    This job description is intended to describe the general nature and level of work being performed. It is not intended to be an exhaustive list of all responsibilities, duties, or skills required of personnel. Gifthealth reserves the right to modify job duties or descriptions at any time.

    Salary Description $123,000- $154,000

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Lead Site Reliability Engineer in United States vacancy
  • $99k - $225k

    Site Reliability Engineer, LeadThe Opportunity:  As a Lead Site Reliability Engineer (SRE) on our team, you’ll be responsible for ensuring the reliability, performance, scalability, and security of critical production systems and platforms. This role leads the design and... 
    Suggested
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Chantilly, Loudoun County, VA
    4 days ago
  • $113.1k - $232.3k

    Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity... 
    Suggested
    Work at office
    Local area
    Visa sponsorship
    Flexible hours
    3 days per week

    Deloitte

    Tampa, FL
    2 days ago
  • Job Summary Job Summary The Support Lead (SRE) is responsible for overseeing the support operations and site reliability engineering tasks, ensuring the effective functioning of systems and applications. The primary goal is to enhance system performance, availability,... 
    Suggested

    TechDigital Group

    Fairfax, VA
    3 days ago
  •  ...build a successful career with opportunities to learn, grow, and make an impact. Join us! Position Summary: The IKCP Site Reliability Engineer Lead is responsible for ensuring the reliability, scalability, performance, security, and operational excellence of the... 
    Suggested
    Work at office
    Flexible hours
    Shift work
    Day shift

    Bank of America Corporation

    Chandler, AZ
    5 days ago
  • Google is hiring Site Reliability Engineers (SRE) in Sunnyvale, CA, to ensure reliability and performance across Google’s services. The role blends software and systems engineering, allowing code fixes to improve systems while maintaining production reliability at scale... 
    Suggested

    Google

    Sunnyvale, CA
    4 days ago
  •  ...only provider of enterprise-scale context engines capable of analyzing trillions of real-...  ...seeking a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team....  ...issues before they impact end-users. Lead troubleshooting efforts for complex production... 
    Full time

    Lovelace Ai

    Pittsburgh, PA
    1 day ago
  • $140k - $230k

     ...Zoox is seeking a Site Reliability Engineer to help ensure the availability, performance, and resilience of the services that power the development...  ...deployment processes, and drive automation initiatives. Lead incident resolution: You will conduct thorough root cause... 
    Full time

    Zoox

    Remote
    1 day ago
  • $175k - $250k

     ...developed by our expert team of lawyers, engineers and research scientists. We’ve found...  ...Overview As a Software Engineer on the Site Reliability team at Harvey, you will ensure the...  ...networking) across 50+ global regions Lead incident management processes, including... 
    Full time
    Relocation package

    Harvey

    Remote
    1 day ago
  •  ...Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems safely...  ...safe shipping the default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale... 
    Permanent employment
    Work experience placement
    Work at office
    Local area

    Plaid Financial

    San Francisco, CA
    1 day ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex... 
    Work at office

    JP Morgan Chase

    Jersey City, NJ
    4 days ago
  •  ...to physicians, providing critical information about the right treatments for the right patients, at the right time.The Site Reliability Engineering team works with all departments and business units to provide dependable cloud infrastructure solutions, along with support... 
    Full time

    Tempus

    Chicago, IL
    4 days ago
  • $80k - $133k

     ...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and... 
    Permanent employment
    Full time
    Contract work
    Remote work
    Flexible hours

    Guidehouse

    San Antonio, TX
    4 days ago
  • $165k - $280k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most... 
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Palo Alto, CA
    3 days ago
  • Recognized as the No. 1 site trusted by real estate professionals, Realtor.com has...  ...guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence...  ...a diverse team of experts as you use leading-edge tech to empower everyone to meet a... 
    Work at office
    Local area

    Realtor.com

    Austin, TX
    3 days ago
  • $147k - $210k

     ...development code.Review code developed by other engineers and provide feedback to ensure best...  ...and quality. Participate in, or lead design reviews with peers and stakeholders...  ...large-scale distributed systems. Site Reliability Engineering (SRE) is what you get when... 

    Google

    Sunnyvale, TX
    2 days ago
  • $112k - $137k

     ...Financial Group (MUFG), one of the world’s leading financial groups. Across the globe, we’...  ...will work at an MUFG office or client sites four days per week and work remotely...  ...highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and... 
    Full time
    Work at office
    Local area
    Remote work

    MUFG

    Tampa, FL
    2 days ago
  • Reliability Engineering Design, implement, and operate scalable, resilient, and highly available systems...  ..., coordinate service restoration, and lead incident response when appropriate....  ...Abilities Three or more years of experience in Site Reliability Engineering, platform... 
    Remote work

    Patterson-UTI

    Houston, TX
    11 hours ago
  • $139k - $257.55k

     ...Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning,...  ...productivity and personalized customer experiences. Adobe’s industry-leading offerings including Adobe Acrobat Studio, Adobe Express,... 
    Full time
    Temporary work
    Local area
    Remote work
    Worldwide

    Adobe Systems

    New York, NY
    4 days ago
  • $165k - $225.6k

     ...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability Engineering, this role will help build,... 
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    2 days ago
  • $125k - $185k

     ...HybridA World-Changing CompanyPalantir builds the world’s leading software for data-driven decisions and operations. By bringing...  ..., and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-performance... 
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Washington DC
    2 days ago
  • $158.5k - $172k

     ...velocity energy of a powerhouse startup.As a leading U.S. ordering and delivery marketplace,...  ....About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will...  ...high-impact position driving continuous reliability, deep system optimization, and automation... 
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    Chicago, IL
    2 days ago
  • $138.4k - $173k

     ...infrastructure as well as help improve the reliability, quality of services and overall...  ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability...  ...about our locations by visiting our site.Compensation & BenefitsThe base salary that... 
    Full time
    Flexible hours

    AppFolio

    Santa Barbara, CA
    2 days ago
  • $125k - $150k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (RAPTOR)SpaceX is looking for a Site Reliability Engineer with a strong drive to solve challenging problems in the Raptor... 
    Permanent employment
    Temporary work

    SpaceX

    Hawthorne, CA
    2 days ago
  • $165k - $230k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCE (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink... 
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Redmond, WA
    2 days ago
  • $230k - $250k

    GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of...  ....Improve system performance, capacity, and resilience.Lead incident response and root cause analysis.Implement disaster... 
    Remote work

    Govcio

    Arlington, VA
    2 days ago
  • $174k - $252k

     ...pushing for changes that improve reliability and velocity.Practice...  ...degree in Computer Science, Engineering, a related field, or equivalent...  ...systems.2 years of experience leading projects and providing technical...  ...Science or Engineering.Site Reliability Engineering (SRE)... 

    Google

    Sunnyvale, TX
    2 days ago
  • $104.9k - $174.7k

     ...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    LexisNexis Risk Solutions Group

    Atlanta, GA
    1 day ago
  • $128.6k - $184.9k

     ...global cloud platform. As a team of six engineers distributed across the US, Canada, and the...  ...with a strong focus on automation, reliability, and operational excellence. We are one...  ...Qualifications7+ years of experience in Site Reliability Engineering, DevOps, Infrastructure... 
    Permanent employment
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    CISCO Systems

    Richardson, TX
    1 day ago
  • $130k - $200k

    IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion... 
    Full time
    Work at office
    Immediate start

    IXL Learning

    San Mateo, CA
    11 hours ago
  • $165k - $190k

    Obsidian Security is the leading SaaS security platform, trusted by global enterprises...  ...DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable,...  ...complex challenges around scalability, reliability, observability, and cost efficiencyCollaborate... 
    Work from home

    Obsidian Security

    Palo Alto, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Lead Site Reliability Engineer. Be the first to apply!