Site Reliability Engineering (SRE) Manager
$139.7k - $232.9kM&T Bank
OverviewThe Site Reliability Engineering (SRE) Manager leads teams responsible for the reliability, availability, performance, and operational excellence of critical business applications and platforms. This role combines engineering leadership with deep expertise in production operations, observability, automation, incident management, and cloud technologies.The SRE Manager partners with Engineering, Architecture, Infrastructure, Security, Product, and Business stakeholders to ensure systems are resilient, scalable, secure, and supportable. The role is accountable for driving operational excellence through automation, reliability engineering practices, and continuous improvement while developing high-performing SRE and Production Support teams.Primary ResponsibilitiesReliability & Operational ExcellenceDefine and execute SRE strategies that improve system reliability, availability, scalability, and performance.Establish and govern Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational health metrics.Lead production readiness reviews, disaster recovery testing, resilience assessments, and operational risk mitigation activities.Drive continuous improvement of application stability, service availability, and customer experience.Incident & Problem ManagementLead major incident response and escalation management for critical production issues.Oversee root cause analysis (RCA) processes and ensure corrective actions are implemented and tracked to completion.Drive reduction of recurring incidents through engineering improvements, automation, and proactive monitoring.Provide executive-level communication during significant incidents and service disruptions.Observability & AutomationEstablish monitoring, alerting, logging, tracing, and observability standards across supported platforms.Lead implementation of dashboards and operational metrics that provide visibility into service health and customer impact.Drive automation initiatives that reduce manual operational effort, improve recovery times, and increase engineering efficiency.Promote Infrastructure as Code (IaC), CI/CD integration, automated remediation, and self-service operational capabilities.Cloud & Platform ReliabilityPartner with Engineering and Infrastructure teams to support cloud-native and hybrid application environments.Ensure applications are designed and operated using resilient, scalable, and supportable architectures.Support modernization initiatives involving Azure cloud services, containers, APIs, microservices, and platform engineering practices.Evaluate vendor platforms and third-party services to ensure reliability and operational readiness.AI & Modern OperationsDrive adoption of AI and Generative AI capabilities to improve incident response, troubleshooting, observability, and operational efficiency.Identify opportunities for intelligent automation, anomaly detection, automated diagnostics, and AI-assisted knowledge management.Promote responsible AI adoption aligned with enterprise security, governance, and risk standards.People LeadershipRecruit, develop, coach, and retain high-performing Site Reliability Engineers, Production Engineers, Automation Engineers, and Observability Engineers.Establish career paths, skill development plans, and succession strategies.Foster a culture of ownership, accountability, innovation, collaboration, and continuous learning.Manage staffing, performance management, compensation recommendations, and organizational development activities.Risk & GovernanceEnsure adherence to enterprise risk, cybersecurity, regulatory, and operational control standards.Identify and escalate operational risks impacting critical services or customer experiences.Support audits, regulatory reviews, disaster recovery exercises, and operational governance programs.Scope of ResponsibilitiesLeads teams responsible for:Site Reliability Engineering (SRE)Production SupportObservability EngineeringIncident ManagementOperational AutomationCloud ReliabilityPlatform OperationsResponsible for reliability and operational health across multiple applications, platforms, cloud services, and vendor-supported solutions.Supervisory ResponsibilitiesTypically manages 10-20 direct and indirect reports including SRE Engineers, Production Engineers, Technical Leads, and Engineering Managers.Education & Experience Required10+ years of technology experience with application support, infrastructure, cloud, software engineering, or reliability engineering responsibilities.5+ years of leadership experience managing engineering, operations, or SRE teams.Experience managing production systems supporting critical business functions.Strong knowledge of Site Reliability Engineering principles, including SLOs, observability, automation, incident management, and operational excellence.Experience leading major incident response, root cause analysis, and service restoration efforts.Experience with cloud platforms, distributed systems, APIs, and modern application architectures.Strong communication, analytical, decision-making, and stakeholder management skills.Preferred QualificationsBachelor's degree in Computer Science, Engineering, Information Technology, or related field.Experience leading SRE or Production Engineering organizations.Experience with Azure cloud technologies and cloud-native architectures.Experience with observability platforms such as Dynatrace, Splunk, Datadog, Grafana, Azure Monitor, or OpenTelemetry.Experience with scripting and automation technologies including PowerShell, Python, Bash, and APIs.Experience with CI/CD, Infrastructure as Code, DevOps, and Platform Engineering practices.Experience implementing operational AI use cases including incident analysis, observability analytics, and automated diagnostics.Financial services or other highly regulated industry experience preferred.What Great Looks LikeA successful SRE Manager at M&T:Delivers highly available and resilient customer-facing platforms.Uses automation to eliminate operational toil and improve efficiency.Reduces mean time to detect (MTTD) and mean time to restore (MTTR).Establishes strong observability and operational intelligence capabilities.Builds a culture of reliability, accountability, and continuous improvement.Successfully integrates AI-assisted operations and automation into support workflows.develops high-performing teams that balance reliability, speed, risk management, and customer experience.M&T Bank is committed to fair, competitive, and market-informed pay for our employees. The pay range for this position is $139,700.00 - $232,900.00 Annual (USD). The successful candidate’s particular combination of knowledge, skills, and experience will inform their specific compensation.LocationBuffalo, New York, United States of AmericaSummaryLocation: Buffalo, NYType: Full time
$116.4k - $194k
...SummaryResponsible for the reliability, availability,... ...production support, release engineering, observability,... ...stability.Lead incident management, root cause analysis (... ...initiatives.Understanding of SRE principles,... ...in Production Support, Site Reliability Engineering...SuggestedFull timeWork experience placementWork from home$139.7k - $232.9k
Manager, Site Reliability Engineering 62 M Overview Responsible for leading the Site Reliability Engineering Center of Excellence and the Forward Deployed SRE program supporting critical banking platforms, applications, and technology services. Manages an organization...SuggestedFull timeWork experience placement$139.7k - $232.9k
...continuously improving highly reliable, scalable, and resilient... ...subject matter expert (SME) in Site Reliability Engineering, driving reliability... ...requirements. • Lead incident management practices, including detection... ...: Applies expert-level SRE practices across multiple platforms...SuggestedFull timeWork experience placement- ...Seeking experienced Site Reliability Engineers focused on observability, monitoring, automation, and platform reliability. Candidates should have... ...Dynatrace 1. Observability 1. Site Reliability Engineering (SRE) 1. Monitoring & Alerting 1. Application Performance...Suggested
- ...Position Title: Site Reliability Engineer Location: Buffalo, NY Duration: 6 Months with possible... ...Site Reliability Engineering (SRE) practices across the software development... ..., testing, and proactive operational management while coaching and influencing others....SuggestedWork experience placement
$167.6k - $279.4k
...platforms.Extensive experience leading engineering teams responsible for developing... ...risk scoring, investigations, case management, suspicious activity detection,... ...including CI/CD, Infrastructure as Code, Site Reliability Engineering (SRE), platform observability, and...Full timeTemporary workWork experience placement$139.7k - $232.9k
Overview The Engineering Manager leads high-performing software engineering teams responsible for delivering secure, scalable, and... ..., Agile delivery practices, CI/CD automation, and Site Reliability Engineering (SRE) principles.Primary ResponsibilitiesEngineering Leadership...Full time- .... PRIMARY RESPONSIBILITIES Work closely with Technology management, customers, and support teams on a regular basis to lead the design... ...vendor resources as needed. Mentor and coach less experienced engineers, technicians, and integrators. Review documentation, proposals...Contract workWork experience placement
$167.6k - $279.4k
...platforms.Extensive experience leading engineering teams responsible for developing... ...prevention technologies, risk management, real-time transaction processing,... ...including CI/CD, Infrastructure as Code, Site Reliability Engineering (SRE), platform observability, and...Full timeTemporary workWork experience placement$116.4k - $194k
OverviewThe Technical Engineer serves as a senior production support and reliability engineering professional responsible... ...expertise with modern Site Reliability Engineering (SRE), observability, automation... ...cloud operations, and incident management practices. The Technical...Full time$208.94k - $218.94k
* * * * * *Title: Engineering Manager Job Location: 1 Seneca St, Buffalo, NY 14203. Position requires in-office work four (4) days every week.Job Description: Oversee application development, microservices, ETL processes, and data modeling using SQL, Oracle, NoSQL, and...Full timeWork at office$139.7k - $232.9k
...location, with the flexibility to work from home one day per weekOverview: We are seeking an experienced Technology Manager to lead our Database Platform Engineering team and drive the development, optimization, and management of cutting-edge database platforms. In this role,...Full timeWork experience placementWork from home$99k - $232k
Industry/SectorNot ApplicableSpecialismData, Analytics & AIManagement LevelManagerJob Description & SummaryThe OpportunityAs a Engineering Manager, Google Cloud AI/ML you will lead the design and delivery of Google Cloud AI and machine learning solutions within our Data &...Full timeH1b$124k - $280k
...ManagerJob Description & SummaryThe OpportunityAs a Senior Engineering Manager, Google Cloud AI/ML you will lead engineering efforts that design... ..., model evaluation, and guardrail design to support reliable AI behavior- Shaping LLMOps practices for versioning, monitoring...Full timeH1b$120k - $150k
...spanning reach. From production operators to engineers, technicians to specialists, sales to... ...best. Come grow with us!The Engineering Manager will be accountable for all process engineering... ...operations. Acts as voice of site's Process Engineering Group and manages the...Full timeLocal area$150k - $200k
...DegreeSalary Range: $150,000.00 - $200,000.00 Salary/yearTravel Percentage: Up to 30%Job Shift: DayJob Category: Engineering Job Purpose: The Senior Engineering Manager at CONAX TECHNOLOGIES LLC will lead and drive the engineering team to achieve excellence in product...Shift work$90k - $110k
...Job Description ENGINEERING PROJECT MANAGER (Permanent) Our client, a global manufacturing organization in WNY, is looking to add an Engineering Project Manager to their growing team. Responsibilities: Manage cross-functional engineeringprojects...Permanent employmentFull timeVisa sponsorship- ...: AI & Platform Software Engineer Location: Remote within... ...AI gateway on Azure API Management, model access via Azure AI Foundry... ...test data management, and site reliability engineering). Because the... ...engineering, AI engineering, DevOps, SRE, or related disciplines...Contract workFor contractorsWork at officeRemote work
- ...Engineering Manager Location: Onshore (US) — Candidate must be in EST/CST time zones or fully align with EST hours Experience: 10+ Years Key Responsibilities: Lead and mentor a cross-functional engineering team with clear ownership, accountability, and delivery...Shift work
$116.4k - $194k
...and mentor solution developers, software engineers, analysts, and support team members.... ...static data, market data, pricing, position management, collateral processing, confirmations,... ...to improve performance, resiliency, reliability, scalability, supportability, and long-...Full timeWork experience placementWork from home$95.92k - $222.03k
...Level Support for applications that use DB2.Ensure support and management of distributed DB2 connections.Ensure stability and integrity of... ...possible, we hire locally to NTT DATA offices or client sites. This ensures we can provide timely and effective support tailored...Temporary workWork at officeRemote workFlexible hours$125k - $175k
...of life inside and outside of work.Job Title:Quality Engineering ManagerReporting To:Director, AG Site OperationsWork Schedule:Onsite - Buffalo, NYOur team... ...Aircraft Group is looking for a Quality Engineering Manager to join them. You’ll report to the Site Electronic Operations...Full timeContract workRelocation packageFlexible hours- DescriptionWe seek a Director of Engineering with a highway or bridge background to join our... ...and private clients. We want to grow our management team with individuals who possess a... ...membersAble to traverse a construction job site consisting of uneven ground varying in height...Contract workWork at officeLocal area
$80.9k - $134.8k
...Liquidity Coverage Ratio (LCR). Writes reliable code, supports technical requirements gathering... ..., and resilient regulatory and management reporting solutions in accordance with banking... ...tasks, partnering with senior engineers, report owners, Treasury, and other stakeholders...Full timeWork experience placementWork from home$115k - $130k
...deliver their goals with excellence.INSPIRE seeks a Senior Software Engineer to work on our Digital Services team in Buffalo, NY, to design... ...life cycle, web application development, and project management.Experience with source code management and issue tracking systems...Full timeTemporary work$97.1k - $161.8k
...weekOverview: Archer (RSA Archer) specific engineer responsible at the advanced level for... ...CD practices, preferably Git/GitLab, for managing scripts, integration code, and... ...reporting and consumer requirements into reliable, governed datasets while supporting access...Full timeWork experience placementWork from home$97.1k - $161.8k
OverviewJoin a highly visible engineering team responsible for the technology that powers... ...automation initiatives, improve operational reliability, and deliver enhancements that are... ...ResponsibilitiesWork closely with Technology management, senior Engineers, and support teams on...Full timeWork experience placement$80.9k - $134.8k
OverviewJoin a highly visible engineering team responsible for supporting the technology that... ...initiatives, improve operational reliability, and contribute to enhancements that are... ...ResponsibilitiesWork closely with Technology management, more experienced Engineers, and...Full timeWork experience placement$147.12k
...monitor resources, coordinating tasks on projects. Prepare and manage technical components of project plans. Collaborate with teams in... ...five (5) years of experience in the job offered or as Software Engineer, Software Developer, Application Development Analyst, or...Full timeWork at office$97.1k - $161.8k
...review pull requests, provide feedback, and execute on the change management of the request.Author organized, clean, efficient, and secure... ...of one programming language to be verified by a lead software engineer and apply knowledge of appropriate data structure and...Full timeWork experience placementWork from home
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineering (SRE) Manager. Be the first to apply!



