Manager Site Reliability Engineering
$51.9 per hourHighmark Health
Company :
Allegheny Health Network
Job Description :
GENERAL OVERVIEW:
This job is responsible for the reliability, availability, and performance of critical healthcare IT systems, principally in the Environment of Care (EOC), enabling seamless access to essential services for patients, providers, and the people we serve. Proactively identifies and mitigates potential disruptions to maintain the highest standards of care and operational efficiency. This role blends software engineering, clinical engineering, and security principles with a deep understanding of healthcare operations to minimize downtime, improve system resilience, and to support clinical workflows and continuity of hospital operations. Works cross-functionally with AHN site leaders and teams to navigate and to monitor and support building automation and facility systems, clinical engineering / IoT, healthcare delivery technology architecture, infrastructure and platform operations, and cybersecurity. Fosters a culture of automation, continuous improvement, collaboration, and patient safety. Develops core metrics for monitoring and maintaining system health for SRE practitioners (e.g., latency, traffic, errors, and saturation) leveraging industry practices, manufacturer guidance, and other service delivery metrics.
ESSENTIAL RESPONSIBILITIES
Perform management responsibilities to include, but are not limited to: involved in hiring and termination decisions, coaching and development, rewards and recognition, performance management and staff productivity.Plan, organize, staff, direct and control the day-to-day operations of the department; develop and implement policies and programs as necessary; may have budgetary responsibility and authority. (25%)
Oversees the partnership with clinical engineering, cybersecurity, device manufacturers, suppliers, and Information Technology SMEs to oversee and to implement strategies for managing, monitoring, and securing a diverse range of clinical devices and other technology equipment (e.g., IoT), ensuring compliance with HIPAA and other relevant regulations (e.g., FDA, TJC, PCI). Keeps current on healthcare IT trends, including AI, security patching, and best practices for device hardening. Oversees and assists with network segmentation and access controls to isolate and to protect clinical and other critical devices. Automates monitoring tasks to improve efficiency and reduce errors. Identifies and remediates vulnerabilities in clinical devices and related infrastructure. Manages and reports issues with assets, devices, integration services, and other equipment. Engages the appropriate parties to develop and deploy a fix/solution or oversees ownership of resolution actions. Utilizes observability practices to gain deep insights into system behavior, enabling faster identification and resolution of issues. (15%)
Oversees the SRE partnership with Clinical Engineering and Cybersecurity Engineering to troubleshoot technical issues related to medical equipment and systems. Participates in the medical device technology lifecycle – from product/device evaluation, discovery, to implementation, maintenance, and through retirement. Develops the framework and structure to maintain documentation related to the IT infrastructure supporting clinicaland other critical devices. Participates in the planning and oversees the execution of preventative maintenance activities. Provides direction and guidance to team members on how to analyze complex problems and develop effective solutions, how to troubleshoot system outages and performance issues, and how to work collaboratively with other IT, cybersecurity, facility, AI and application teams to resolve issues and to conduct root cause analyses. (15%)
Oversees the SRE partnership with facility leaders to optimize the performance and monitoring of building automation systems (BAS), including HVAC, lighting, fire suppression, security systems, etc. Manages processes and procedures used to monitor BAS performance metrics and proactively identifies potential issues. Works with facilities management to implement improvements to the BAS infrastructure. Works with cybersecurity, vendors/manufacturers, et. al. to ensure the security of building automation systems and oversees monitoring of performance, service delivery, and support. (15%)
Oversees the SRE partnership with IT teams including, but not limited to platform / product management, disaster recovery services, infrastructure and architecture, storage management, and release management. Participates in the planning and execution of downtime drills and system / device recovery exercises. Supports other emergency preparedness drills and exercises, as needed. Leads or participates in post-incident reviews to identify root causes and implement corrective actions. Works with cross-functional stakeholders to Implement and to maintain redundant systems and failover mechanisms to minimize downtime. Reviews and provides feedback on emergency operations plans and other materials which are used to respond to emergency situations (e.g., Continuity of Operations Plans, Incident Response Guides, Downtime Procedures). Manages team members who are supporting the planning and execution of system migrations, releases, and upgrades to ensure minimal disruption to clinical operations. Oversees detailed migration or installation plans, including risk assessments, rollback procedures, and communication strategies. Assists local site leaders with navigating shared services (e.g., AI, IT, Information Security, Clinical Engineering, Platform Operations, Technology Acquisition). (15%)
Establishes core metrics for monitoring and maintaining system health for SRE practitioners (e.g., latency, traffic, errors, and saturation). Manages the processes and procedures used for documentation and knowledge sharing including maintaining detailed documentation of systems, device inventories, processes, and procedures.Leads by example by sharing knowledge and best practices with other staff and cross-functional teams. Provides training and mentorship to junior or less experienced team members. Stays current with the latest technologies and trends in site reliability engineering. Leads or participates in briefings with cross-functional stakeholders to manage priorities and team assignments, support ticket queues, etc. (10%)
Other duties as assigned or requested. (5%)
Q UALIFICATIONS:
Required
Bachelor’s degree in Computer Science, Engineering, Management Information Systems, IT, or related field or relevant experience and/or education as determined by the company in lieu of bachelor's degree.
3 years with Management or leadership role
Preferred
Master's degree in Computer Science, Engineering, Management Information Systems, IT, or related field
5 years of experience with Site Reliability Engineering (SRE), Systems Administration, or DevOps particularly in healthcare IT
5 years of experience in Medical device management lifecycle, network / device segmentation, vulnerability and patch management
5 years of experience in Healthcare IT experience in architecture, automation, IoT, telemetry, telehealth, security, system development lifecycle, capacity planning, networking, continuous integration / continuous delivery pipelines (CI/CD), incident management, scripting, metrics, monitoring, redundancy, etc.
3 years of experience working in highly regulated environments
3 years of experience with Progressive leadership roles, preferably inclinical engineering, IT, business continuity, backup and storage management, building automation, or cybersecurity discipline in healthcare
SKILLS:
Problem-Solving: Excellent analytical and troubleshooting skills; High capacity to think analytically, interpret information / observations, apply judgment and to assist with making effective, strategic decisions.
Collaboration: Ability to work effectively in a team environment; demonstrated ability to support multiple sites and locations while maintaining consistency in service delivery processes and procedures.
Communication: Strong written and verbal communication skills.
Flexibility: Willingness to participate in activities or incidents which may occur outside of regular work schedules.
Leadership: Demonstrated resource and project planning capabilities, decision making skills, history of results-oriented delivery, and effective team building across multiple locations and a diverse team of staff, partners, and stakeholders.
Security Awareness: Understanding of security best practices and how to apply them in a healthcare IT environment.
Delivery and Execution: Demonstrated competency in the execution of multiple projects, including managing resources across multiple projects to meet goals.
Relationships: Strong relationship building skills and ability to influence with and without authority in a matrixed organization.
Disclaimer: The job description has been designed to indicate the general nature and essential duties and responsibilities of work performed by employees within this job title. It may not contain a comprehensive inventory of all duties, responsibilities, and qualifications required of employees to do this job.
Compliance Requirement : This job adheres to the ethical and legal standards and behavioral expectations as set forth in the code of business conduct and company policies.
As a component of job responsibilities, employees may have access to covered information, cardholder data, or other confidential customer information that must be protected at all times. In connection with this, all employees must comply with both the Health Insurance Portability Accountability Act of 1996 (HIPAA) as described in the Notice of Privacy Practices and Privacy Policies and Procedures as well as all data security guidelines established within the Company’s Handbook of Privacy Policies and Practices and Information Security Policy.
Furthermore, it is every employee’s responsibility to comply with the company’s Code of Business Conduct. This includes but is not limited to adherence to applicable federal and state laws, rules, and regulations as well as company policies and training requirements.
Pay Range Minimum:
$51.90
Pay Range Maximum:
$83.84
Base pay is determined by a variety of factors including a candidate’s qualifications, experience, and expected contributions, as well as internal peer equity, market, and business considerations. The displayed salary range does not reflect any geographic differential Highmark may apply for certain locations based upon comparative markets.
Highmark Health and its affiliates prohibit discrimination against qualified individuals based on their status as protected veterans or individuals with disabilities and prohibit discrimination against all individuals based on any category protected by applicable federal, state, or local law.
We endeavor to make this site accessible to any and all users. If you would like to contact us regarding the accessibility of our website or need assistance completing the application process, please contact the email below.
For accommodation requests, please contact HR Services Online at View email address on click.appcast.io
California Consumer Privacy Act Employees, Contractors, and Applicants Notice
Req ID: J280531
$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range... ...observability and alerting systems.The Fleet Management team provides the core runtime... ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager...SuggestedWork at officeLocal areaRemote workWorldwideFlexible hours$93.6k - $170.64k
...currently have a career opportunity for a NOC AI-Ops Engineer to join our team located in Boston, MA. This is a... ...Overview:We are seeking a Senior AIOps and Incident/Site Reliability Engineer to lead incident management, operational resilience, and intelligent automation...SuggestedWork at officeLocal areaNight shift3 days per week$140k - $205k
Senior Technology Site Reliability EngineerCooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operations team.Position summary:... ...high availability and performanceImplement and manage service-level indicators (SLIs), objectives (...SuggestedFull timeTemporary workWork at officeFlexible hoursWeekend work$134.25k - $214.8k
...where you matter.Your ImpactAre you an engineer who gets excited about the challenge... ...the Observability team within Axon's Site Reliability organization — a focused team responsible... ...ArgoCD, and Helm — including capacity management, cybersecurity requirements and...SuggestedWork experience placementWork at officeRemote work$130k - $150k
...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are... ...including Windows Server, VMware vSphere, VMware Site Recovery Manager (SRM), SAN technologies, and the Rubrik ecosystem, with the...SuggestedWork at officeWork from home3 days per week$160k - $200k
...instrumenting distributed systems using OpenTelemetry and managing metrics pipelines with Prometheus at scaleDirect... ...evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage &...Temporary workWork at officeLocal areaFlexible hours3 days per week$138.1k - $198.2k
...technology that simply works. The SRE Engineering Enablement Team supports our CI... ...environments, build tools, code review, artifact management, CI, education, and documentation. Our... ...engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours$166k - $220k
...is critical that Anduril services are reliable and maintainable. This means that all... ...Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will... ...telemetry ingestion, and operate/manage our central monitoring and alerting infrastructure...Full timeWork experience placementImmediate start$160k - $200k
Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems... ...test procedures for QuEra’s Quantum computer operations. Manage science, development and testing infrastructure, including...Local areaRemote work$127k - $249k
...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas... ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This...Local areaRemote workWorldwideFlexible hours- ...our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team,... ...overall responsibility for the design, management and execution of operations required to... ...someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and...Full time
$130k - $180k
...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and... ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to...Temporary workWork at officeImmediate startRemote workFlexible hours$130k - $140k
...scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring... ...and operate solutions that help clients manage data, automate processes, and scale their...Ongoing contractFull timeTemporary workWork experience placement$124k - $280k
...Description & SummaryAt PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to design... ...data storage solutions using cloud services- Designing and managing data warehouses and data lakes- Implementing IAM roles and...Full timeH1b$99k - $232k
...SectorNot ApplicableSpecialismSAPManagement LevelManagerJob Description & SummaryThe OpportunityAs a SAP Order to Cash Consultant, Manager, you will lead our clients in their customer transformation journey by reimagining exceptional experiences for their customers and...Full timeH1b- ...strategy, digital identity, cyber defense, application security, and managed service solutions to rethink the entire security lifecycle. Do... ...is not a pen test engagement. A Cybersecurity Forward Deployed Engineer is a production engineer who works embedded inside a client’s...Full timeWork experience placementLive inWork at officeLocal area
- ...then consider a career in Advisory.KPMG is currently seeking a Manager, AI Engineer to join our Advisory Services practice.Responsibilities:End-... ...can be found towards the bottom of our KPMG US Careers site at Benefits & How We Work.Follow this link to obtain salary ranges...H1bLocal area
$99k - $232k
...for coaching, leveraging team member’s unique strengths, and managing performance to deliver on client expectations. With your growing... ...requirements.The OpportunityAs part of the Data and Analytics Engineering team, you will serve as both a technical leader and a trusted...Full timeH1b$99k - $252.45k
...Description & SummaryThe OpportunityAs an AI Senior Developer Manager, you will lead the charge in developing innovative software solutions... ...transformation- Managing and mentoring a team of software engineers to deliver quality software products- Designing, coding, and testing...Full timeH1b$73.5k - $212.28k
Industry/SectorNot ApplicableSpecialismIFS - Information Technology (IT)Management LevelManagerJob Description & SummaryAt PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to design and develop robust data solutions...Full timeH1b- ...and game changing solutions in our inclusive culture that values diversity of ideas, experiences and backgrounds.You are a confident Manager who spots and stays ahead of the SAP platform , industry and Finance trends and knows how to translate client goals into clear and...Full timeWork experience placementLive inWork at officeLocal area
$128.03k - $261.63k
Position Summary What You’ll Do As a Deloitte Tax, AI Engineer Manager, you will oversee the design, development, deployment, and... ...solutions.Identify and resolve technical issues, ensuring the reliability and performance of applications.Create and maintain...Work at officeLocal areaVisa sponsorship2 days per week3 days per week$150k - $215k
...of their bodies and daily lives. Protecting our members’ data and ensuring our systems scale securely and reliably is core to this mission.As an Engineering Manager on the Client Platform team, you will lead a team responsible for building and scaling the systems that...Full timeWork at officeRelocation$122.5k - $423.78k
...Description & SummaryAt PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to design... ...Large Language Models (LLMs) and orchestrated applications- Manage and own the full lifecycle of AI models and solutions, from...Full timeTemporary workH1bRemote work$73.5k - $212.28k
...At PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to design and develop robust... ...for coaching, leveraging team member's unique strengths, and managing performance to deliver on client expectations. With your...H1b$99k - $232k
...a Delivery Excellence - Tech Enablement - Senior Developer - Manager, you will play a pivotal role in driving software and product... ...drive digital transformation- Managing and mentoring software engineering teams to enhance performance and deliver quality outcomes- Designing...Full timeH1b$60 - $70 per hour
Boston, MAHybridContract$60/hr - $70/hrThis client is hiring a Lead Quality Assurance Engineer in Boston, MA (4 days onsite) to join a global investment management firm building software that powers critical financial systems. This is a long-term contract with strong visibility...Long term contractFull timeTemporary workFlexible hours$95k - $245k
...space exploration to biomedical engineering, lives often depend on the... ...design specificationsAbility to manage small technical... ...fabrication, assembly & test sites to guarantee schedule, cost,... ..., assembly, and testingDrive reliability and characterization of integrated...Full timeLocal area$124k - $280k
...Description & SummaryAt PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to design... ...complex needs of health system and health plans. As a Senior Manager, you will drive use case development across clinical decision...Full timeH1b$140.6k - $198k
We anticipate the application window for this opening will close on - 12 Aug 2026Position Description:Principal Reliability Engineer- Software Quality for Covidien, LP. Responsible for various software quality assurance functions during the Product Development Process...Relocation package
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager Site Reliability Engineering. Be the first to apply!
- site reliability engineer Boston, MA
- site reliability engineer sre Boston, MA
- site services specialist Boston, MA
- construction site safety Boston, MA
- site leader Boston, MA
- official site Boston, MA
- website content developer Boston, MA
- on site coordinator Boston, MA
- IT site lead Boston, MA
- site safety Boston, MA

