Site Reliability Engineering Manager
NationsBenefits, LLC
NationsBenefits is recognized as one of the fastest-growing companies in America and a Healthcare Fintech provider of supplemental benefits, flex cards, and member engagement solutions. We partner with managed care organizations to provide innovative healthcare solutions that drive growth, improve outcomes, reduce costs, and bring value to their members.Through our comprehensive suite of innovative supplemental benefits, fintech payment platforms, and member engagement solutions, we help health plans deliver high-quality benefits to their members that address the social determinants of health and improve member health outcomes and satisfaction.Our compliance-focused infrastructure, proprietary technology systems, and premier service delivery model allow our health plan partners to deliver high-quality, value-based care to millions of members.We offer a fulfilling work environment that attracts top talent and encourages all associates to contribute to delivering premier service to internal and external customers alike. Our goal is to transform the healthcare industry for the better! We provide career advancement opportunities from within the organization across multiple locations in the US, South America, and India.
Location: Remote (US-based candidates only) Manager, Site Reliability Engineering (SRE)Position Overview
We are seeking a Manager, Site Reliability Engineering (SRE) to lead our US-based SRE team and drive operational excellence across our production platforms.
This is a player-coach leadership role that combines people management with hands-on technical leadership. You will mentor and grow a team of Site Reliability Engineers while actively participating in major incident response, reliability initiatives, and operational reviews. The role is a key part of our global follow-the-sun support model and requires close collaboration with SRE leadership in India.
Key Responsibilities
Team Leadership & Development
- Lead, mentor, and develop a US-based team of Site Reliability Engineers.
- Conduct regular 1:1s, performance reviews, and career development discussions.
- Own hiring, onboarding, and retention efforts as the team scales.
- Foster a culture of ownership, blameless postmortems, and continuous improvement.
Operational Excellence & Incident Management
- Lead day-to-day production operations and ensure timely incident triage, resolution, and escalation.
- Serve as an escalation point and incident commander for major production incidents.
- Drive problem management and root cause analysis processes.
- Carry PagerDuty on-call escalation responsibilities for critical issues.
- Track and report operational KPIs, SLAs, and SLOs, including availability, MTTR, and incident trends.
Reliability & Automation
- Improve system reliability, observability, and resilience using Datadog and related tooling.
- Drive automation, self-healing capabilities, and runbook maturity.
- Partner with Development, DevOps, DevSecOps, and Engineering teams to embed reliability into the SDLC.
- Contribute hands-on to tooling, automation, and technical reviews as needed.
Collaboration & Global Alignment
- Coordinate closely with SRE leadership in India to ensure seamless follow-the-sun coverage.
- Represent the US SRE organization in cross-functional planning and operational reviews.
- Communicate effectively with both technical and non-technical stakeholders.
Documentation & Compliance
- Maintain high-quality documentation for incidents, postmortems, runbooks, and operational procedures.
- Ensure adherence to healthcare and fintech compliance standards, including HIPAA, PCI DSS, SOC 2, ISO 27001, and HITRUST.
Required Qualifications
- 5–8 years of experience in Site Reliability Engineering, DevOps, Production Support, or Platform Engineering.
- 1–2+ years of experience leading, mentoring, or managing engineers.
- Demonstrated success operating in a player-coach leadership model.
- Strong hands-on experience with production incident management and escalation processes.
- Proficiency with Datadog or similar observability platforms.
- Hands-on experience with Kubernetes and Docker in production environments.
- Strong scripting or programming skills in PowerShell, Bash, Python, Java, or C#.
- Experience with Helm, CI/CD pipelines, and deployment automation.
- Working knowledge of ITIL processes and Agile methodologies.
- Experience working with SQL, MySQL, or NoSQL databases.
- Excellent communication and stakeholder management skills.
- Willingness to participate in PagerDuty on-call escalation and work within a global follow-the-sun operating model.
Preferred Qualifications
- Experience with cloud platforms such as AWS, Azure, or GCP.
- Experience building or scaling SRE teams and on-call programs.
- Experience defining and managing SLOs, SLIs, and error budgets.
- Prior experience in the healthcare or fintech industry.
- Knowledge of security and compliance frameworks relevant to regulated environments.
Why Join NationsBenefits?
- Competitive compensation and comprehensive benefits.
- Unlimited PTO.
- Fully remote work environment (US-based).
- Opportunity to lead and grow a high-impact SRE organization.
- Exposure to modern cloud-native technologies and large-scale reliability challenges.
- Collaborative culture focused on innovation, learning, and continuous improvement.
- Meaningful work that directly impacts healthcare technology and millions of members.
Ideal Candidate
We are looking for a technically strong SRE leader who enjoys building teams, improving operational maturity, and remaining hands-on during critical production events. The ideal candidate combines leadership, systems thinking, and automation expertise to help scale reliability practices across a fast-growing Healthcare FinTech organization.
NationsBenefits is an Equal Opportunity Employer.- ...Site Reliability Engineering Manager Home Based - APAC; Home based - EMEA Canonical is a leading provider of open-source software and operating systems for global enterprise and technology markets. Our platform, Ubuntu, is very widely used in breakthrough enterprise...SuggestedWork at officeLocal areaRemote workWork from homeWorldwide
- ...flex cards, and member engagement solutions. We partner with managed care organizations to provide innovative healthcare... ...Location: Remote (US-based candidates only) Manager, Site Reliability Engineering (SRE) Position Overview We are seeking a Manager, Site...SuggestedRemote workFlexible hours
$100 per hour
...team of platform-focused SRE engineers, providing technical mentorship... ...development, and performance management while fostering a culture of... ..., and operate their services reliably Manage observability... ...regular team and company off-sites throughout the year as well as...SuggestedWork experience placementLive inLocal areaRemote workWork from homeHome officeFlexible hours$139.7k - $232.9k
...Wilmington Center Wilmington, DE, location with the flexibility to work from home one day per week Overview The Site Reliability Engineering (SRE) Manager leads teams responsible for the reliability, availability, performance, and operational excellence of critical...SuggestedFull timeWork from home1 day per week$276.1k - $311.4k
...defines your career! The role As SRE Manager, you'll build the Vehicle Software SRE... ...defining its charter, hiring its founding engineers, establishing the operating model, and... ...creating the technical strategy that makes reliability a first-class property of the software...SuggestedPermanent employmentFull timeWork at officeWork from home$195k - $285k
...the infrastructure underpinning our engineering organization must be as reliable and scalable as the chips we build... ...role builds and leads d-Matrix's Site Reliability Engineering function... ...), and build the incident and RCA management framework.Own the full observability...Remote work$243.8k
...Windows, iOS, and Android, our search engine, and the DuckDuckGo subscription. We also... ..., inclusivity, and empowered project management underpins everything we do, where each... ...Your Team and Role Working on the Site Reliability Team, you'll help build and maintain world...Full timeWork at officeLocal areaRemote workFlexible hours$175k - $220k
...across the U.S., Canada, and India. The Director, Site Reliability Engineering (SRE) will lead reliability, performance, and... ...CD practices for assigned product families. Directors will manage multiple teams and collaborate with Product Development, Architecture...Contract workTemporary workWork at officeWork from homeFlexible hours- ...Site Reliability Engineering (SRE) Team Lead The Site Reliability Engineering (SRE) team is foundational to the growth and scale of our platform... ...pipelines, using CI/CD systems and configuration management Why You'll Like It Here We are collaborative at...Shift work
- ...Service (S3), and Auto Scaling Groups for dynamic resource management. Designs, develops, and executes performance tests using... ...Delivery (CI/CD) pipelines and Kubernetes. Supports Site Reliability Engineering (SRE) functions by establishing Service Level Objectives...Full time
$140k - $210k
...Your job As an Engineering Manager in Site Reliability Engineering at Indeed, you will manage and grow a team that applies software engineering principles to make systems more reliable and scalable. These systems support Indeed’s mission to help people get jobs. You...Work experience placementLocal areaRemote work$150k - $170k
...compliance in everything we do. Join us and help build the future of global investing! About The Role As the Manager of Site Reliability Engineering, you'll lead a team of SRE Automation Engineers while remaining a hands‑on technical authority for our Brokerage‑as...Full timeWork at officeRemote workWorldwideFlexible hours- ...I.T. is actively seeking a Principal Engineer for an immediate full-time opportunity... ...technology company is seeking experienced Site Reliability Engineers to take ownership of... ...including SLI/SLO frameworks and error budget management Establish escalation protocols and...Full timeContract workImmediate startWork from homeFlexible hours
$200k - $250k
...progress, come build the future together. As a Principal Site Reliability Engineer, you'll shape the long-term strategy for the... ...across critical infrastructure, including cluster lifecycle management, networking, identity and access management, observability...Full timeImmediate startRemote work$169.3k - $304.7k
...Our team designs, develops, and manages applications and infrastructure that... ...maintaining fast, efficient, scalable, and reliable routing software and infrastructure that... ...global platform. As a Principal Site Reliability Engineer - Network, you will be responsible...Work experience placementWork at officeRemote work$189.59k - $220k
Director, Site Reliability EngineeringNBCUniversal is one of the world's leading media and entertainment... ...of NBCUniversal's Production Software Engineering team, responsible for leading and... ...configuration and support as well as managing the work of other architects and...For contractorsRemote work$160k - $180k
..., Canada, and India. We are seeking a Principal Site Reliability Engineer to define the strategic vision and own the enterprise-wide... ...: Establish the governance models for defining and managing SLIs and SLOs across multiple product lines. ~ Delivery...Contract workTemporary workWork at officeWork from homeFlexible hours- ...Setting the reliability strategy for the platform, the full-time Principal Site Reliability Engineer will define deployment and operational standards for distributed systems, ensuring reliability and automation across customer environments while working remotely. Key...Full timeRemote work
$165k - $185k
...Principal Site Reliability Engineer (SRE) GROW WITH US: Tandem Diabetes Care creates new possibilities for people living with diabetes, their... ...also responsible for: Production Support & Incident Management Reliability Engineering & Observability...Local areaRemote workFlexible hours- ...Principal Site Reliability Engineer LivePerson transforms customer care from voice calls to mobile messaging. Our cloud-based software platform... ..., including networking, IAM, compute, storage, and managed services. ~ Extensive experience with Kubernetes and containerization...Local areaRemote work
$122k - $207k
...services that help people, businesses and governments realize their greatest potential. Title and Summary Manager, Site Reliability Engineering Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that benefits...Full timePart timeWorldwideFlexible hours- ...technology designed to keep the electric grid secure and reliable, even during extended periods of stress. By... ...right place. Role Description Form Energy is hiring a Manager, Site Reliability Engineer to lead the operational function responsible for maintaining...Full timeRemote workRelocation package
$182k - $250.8k
...at Okta is the backbone of our platform's reliability and operational excellence. We are a forward-thinking group of engineers and leaders who believe that great infrastructure... ...for millions of users worldwide. As a Manager, Site Reliability Engineer, you'll lead this...Permanent employmentLocal areaRemote workWorldwideFlexible hoursWeekend workWeekday work$120k - $170k
Sr. Manager/Manager Site Reliability Engineering Join to apply for the Sr. Manager/Manager Site Reliability Engineering role at Aritzia Sr. Manager/Manager Site Reliability Engineering 1 day ago Be among the first 25 applicants Join to apply for the Sr. Manager/Manager...Full timeWork at officeRemote workFlexible hours- ...movement with the scale of a leader and the energy of a high-growth company. The Role Nium is looking for a Senior Manager, Site Reliability Engineering to lead the teams responsible for the availability, performance, scalability, and operational excellence of our...Full timeWork at officeLocal areaWorldwideFlexible hours3 days per week
- ...Digital, and many more. ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems... ...capacity and performance to uphold these standards. As Manager of SRE, you'll lead a team of engineers responsible for...Full time
- ...This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Director, Site Reliability Engineering based in United States. As Director, Site Reliability Engineering, you will lead the teams and technical...Full timeRemote work
- ...Sophos seeks an experienced Manager, Software Engineering (SRE) to lead a distributed team across the U.S. and Canada, focusing on reliability, scalability, and efficient cloud operations. You will guide AWS, Kubernetes/EKS, Terraform/IaC, and automation efforts while...Remote job
$118.1k - $200.76k
Principal Networking Site Reliability Engineer See what you're missing. Our employees work on the world's most advanced electronics - from detecting... ...understanding of cloud technologies and service lifecycle management Experience in deploying capabilities and services through...Full timeLocal areaRemote workWorldwideFlexible hoursNight shift$81.5k - $141.3k
...in more than 160 countries.JOB DESCRIPTION:Position Title: Site Reliability Engineer IITeam: CRM DevOps Employment Type: Full-TimeAbout the... ...Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We are seeking a highly skilled and mission-driven...Remote workShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineering Manager. Be the first to apply!



