Lead Site Reliability Engineer
$123k - $154kGifthealth
Lead Site Reliability Engineer (SRE)
At Gifthealth, we're revolutionizing the way people experience healthcare by simplifying the process of managing prescriptions and health services. Our mission is to provide a seamless, personalized, and efficient healthcare experience for all our customers. We're a dynamic, innovative, and customer-centric company dedicated to making a positive impact on people's lives.
Position Summary
Reporting to the Director of Engineering, the Lead Site Reliability Engineer (SRE) is a senior technical contributor responsible for building reliable, scalable software systems and the DevOps practices that support them. This role blends software engineering, operational excellence, and automation to improve the performance and resilience of Gifthealth's applications.
We are seeking a Lead SRE to play a key part in enabling fast, safe delivery of customer-facing features. This role partners closely with product and application engineers to embed reliability, observability, and operational ownership directly into the development lifecycle, ensuring alignment with organizational goals, operational excellence, and compliance standards.
Key Responsibilities
- Designs, builds, and maintains reliable, scalable software systems supporting Ruby on Rails applications
- Embs reliability, performance, and operational best practices into application code and development workflows
- Owns DevOps practices including CI/CD reliability, deployment strategies, and release safety
- Leads incident response, debugging, and root cause analysis across application and platform layers
- Implements and evolves observability (logging, metrics, tracing) within application and service code
- Partners with engineering teams on architecture, capacity planning, and technical standards
Qualifications
- Education: Bachelor's degree in computer science, engineering, or related field OR
- Licensure/Certification:
- Cloud platform certifications (AWS, GCP, Azure) (Preferred)
- SRE or DevOps-focused certifications (Preferred)
- Experience:
- 5+ years of experience in software engineering, SRE, or DevOps roles (Required)
- Hands-on experience building and operating Ruby on Rails applications in production (Required)
- Experience in owning production incidents and application-level reliability (Required)
- Experience in high-growth or scaling engineering organizations (Preferred)
- Experience working in regulated or customer-impact–sensitive environments (Preferred)
- Knowledge, Skills, & Abilities:
- Knowledge of Ruby on Rails application architecture and production operations; software reliability engineering principles (SLOs, SLIs, error budgets); and modern DevOps and CI/CD practices (Required)
- Knowledge of security and compliance considerations in production systems (Preferred)
- Strong software engineering skills (Ruby and/or comparable backend languages) (Required)
- Debugging and performance optimization of production applications skills (Required)
- CI/CD pipelines, deployment automation, and release tooling skills (Required)
- Monitoring and observability tooling (Datadog, New Relic, Prometheus, etc.) skills (Required)
- Infrastructure as Code (Terraform or similar) skills (Preferred)
- Containerization and orchestration (Docker) skills (Preferred)
- Ability to write production-quality code that improves system reliability (Required)
- Ability to collaborate with product and engineering teams to influence design decisions (Required)
- Ability to troubleshoot complex, cross-system failures (Required)
- Ability to mentor engineers on operational ownership and reliability practices (Preferred)
- Ability to balance speed of delivery with long-term system health (Preferred)
- Location: Remote
- Schedule: 8:00 A.M. to 5:00 P.M. Monday through Friday with night and weekend hours on occasion as determined by the needs of the business.
- Regular meetings with internal Backend and Full-Stack Engineers, Engineering Managers, and Product and Security teams. This role may also have meetings with external cloud and tooling vendor representatives.
- Must be able to remain in a stationary position for extended periods while writing or reviewing documentation
- Must be able to work on a computer for the entire shift
- Must be able to attend virtual meetings with cross-functional teams.
equivalent professional experience in software engineering, SRE, or DevOps roles (Required)
Work Environment
Key Essential Functions
Employment Classification
Status: Full-time FLSA: Exempt
Equal Employment Opportunity (EEO) Statement
Gifthealth is an Equal Opportunity Employer and prohibits discrimination and harassment of any kind. All employment decisions are made without regard to race, color, religion, sex, sexual orientation, gender identity, transgender status, national origin, age, disability, veteran status, or any other legally protected status.
We celebrate diversity and are committed to creating an inclusive environment for all employees. If you do not meet every requirement but still feel you would be a great fit for this role, we encourage you to apply!
Disclaimer
This job description is intended to describe the general nature and level of work being performed. It is not intended to be an exhaustive list of all responsibilities, duties, or skills required of personnel. Gifthealth reserves the right to modify job duties or descriptions at any time.
Salary Description $123,000- $154,000
$99k - $225k
Site Reliability Engineer, LeadThe Opportunity: As a Lead Site Reliability Engineer (SRE) on our team, you’ll be responsible for ensuring the reliability, performance, scalability, and security of critical production systems and platforms. This role leads the design and...SuggestedFull timeContract workPart timeWork at officeLocal areaRemote work$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity...SuggestedWork at officeLocal areaVisa sponsorshipFlexible hours3 days per week- Job Summary Job Summary The Support Lead (SRE) is responsible for overseeing the support operations and site reliability engineering tasks, ensuring the effective functioning of systems and applications. The primary goal is to enhance system performance, availability,...Suggested
- ...build a successful career with opportunities to learn, grow, and make an impact. Join us! Position Summary: The IKCP Site Reliability Engineer Lead is responsible for ensuring the reliability, scalability, performance, security, and operational excellence of the...SuggestedWork at officeFlexible hoursShift workDay shift
- Google is hiring Site Reliability Engineers (SRE) in Sunnyvale, CA, to ensure reliability and performance across Google’s services. The role blends software and systems engineering, allowing code fixes to improve systems while maintaining production reliability at scale...Suggested
- ...only provider of enterprise-scale context engines capable of analyzing trillions of real-... ...seeking a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team.... ...issues before they impact end-users. Lead troubleshooting efforts for complex production...Full time
$140k - $230k
...Zoox is seeking a Site Reliability Engineer to help ensure the availability, performance, and resilience of the services that power the development... ...deployment processes, and drive automation initiatives. Lead incident resolution: You will conduct thorough root cause...Full time$175k - $250k
...developed by our expert team of lawyers, engineers and research scientists. We’ve found... ...Overview As a Software Engineer on the Site Reliability team at Harvey, you will ensure the... ...networking) across 50+ global regions Lead incident management processes, including...Full timeRelocation package- ...Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems safely... ...safe shipping the default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale...Permanent employmentWork experience placementWork at officeLocal area
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex...Work at office
- ...to physicians, providing critical information about the right treatments for the right patients, at the right time.The Site Reliability Engineering team works with all departments and business units to provide dependable cloud infrastructure solutions, along with support...Full time
$80k - $133k
...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and...Permanent employmentFull timeContract workRemote workFlexible hours$165k - $280k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most...Permanent employmentTemporary workWorldwideWeekend work- Recognized as the No. 1 site trusted by real estate professionals, Realtor.com has... ...guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence... ...a diverse team of experts as you use leading-edge tech to empower everyone to meet a...Work at officeLocal area
$147k - $210k
...development code.Review code developed by other engineers and provide feedback to ensure best... ...and quality. Participate in, or lead design reviews with peers and stakeholders... ...large-scale distributed systems. Site Reliability Engineering (SRE) is what you get when...$112k - $137k
...Financial Group (MUFG), one of the world’s leading financial groups. Across the globe, we’... ...will work at an MUFG office or client sites four days per week and work remotely... ...highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and...Full timeWork at officeLocal areaRemote work- Reliability Engineering Design, implement, and operate scalable, resilient, and highly available systems... ..., coordinate service restoration, and lead incident response when appropriate.... ...Abilities Three or more years of experience in Site Reliability Engineering, platform...Remote work
$139k - $257.55k
...Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning,... ...productivity and personalized customer experiences. Adobe’s industry-leading offerings including Adobe Acrobat Studio, Adobe Express,...Full timeTemporary workLocal areaRemote workWorldwide$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability Engineering, this role will help build,...Permanent employmentLocal areaWorldwideFlexible hours$125k - $185k
...HybridA World-Changing CompanyPalantir builds the world’s leading software for data-driven decisions and operations. By bringing... ..., and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-performance...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package$158.5k - $172k
...velocity energy of a powerhouse startup.As a leading U.S. ordering and delivery marketplace,... ....About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will... ...high-impact position driving continuous reliability, deep system optimization, and automation...Full timeTemporary workWork at officeFlexible hours3 days per week$138.4k - $173k
...infrastructure as well as help improve the reliability, quality of services and overall... ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability... ...about our locations by visiting our site.Compensation & BenefitsThe base salary that...Full timeFlexible hours$125k - $150k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (RAPTOR)SpaceX is looking for a Site Reliability Engineer with a strong drive to solve challenging problems in the Raptor...Permanent employmentTemporary work$165k - $230k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCE (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink...Permanent employmentTemporary workWorldwideWeekend work$230k - $250k
GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of... ....Improve system performance, capacity, and resilience.Lead incident response and root cause analysis.Implement disaster...Remote work$174k - $252k
...pushing for changes that improve reliability and velocity.Practice... ...degree in Computer Science, Engineering, a related field, or equivalent... ...systems.2 years of experience leading projects and providing technical... ...Science or Engineering.Site Reliability Engineering (SRE)...$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...Full timeWork at officeLocal areaRemote workWork from home$128.6k - $184.9k
...global cloud platform. As a team of six engineers distributed across the US, Canada, and the... ...with a strong focus on automation, reliability, and operational excellence. We are one... ...Qualifications7+ years of experience in Site Reliability Engineering, DevOps, Infrastructure...Permanent employmentFull timeTemporary workLocal areaWorldwideFlexible hours$130k - $200k
IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion...Full timeWork at officeImmediate start$165k - $190k
Obsidian Security is the leading SaaS security platform, trusted by global enterprises... ...DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable,... ...complex challenges around scalability, reliability, observability, and cost efficiencyCollaborate...Work from home
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lead Site Reliability Engineer. Be the first to apply!
- lead ios engineer United States
- lead product engineer United States
- lead integration engineer United States
- lead firmware engineer United States
- lead sales engineer United States
- lead software test engineer United States
- lead mobile developer United States
- lead backend developer United States
- lead quality engineer United States
- lead support engineer United States


