Head of Enterprise Site Reliability Engineering (SRE) - Hybrid
$85 - $95 per hourGenesis10
Job Description
Job Description
Genesis10 is currently seeking a AVP / Head of Enterprise Site Reliability Engineering (SRE) with our consumer finance lender firm client in their Pittsburgh, PA location. This is a Right to hire position.
An experienced Site Reliability Engineering leader is needed to establish, lead, and scale a formal enterprise SRE capability. This is a contract-to-hire opportunity intended to convert to a full-time AVP-level position based on performance, organizational approval, and mutually agreed-upon terms.
This leader will define the enterprise SRE strategy, operating model, governance framework, roadmap, and success measures while building and managing a centralized team focused on reliability, resilience, operational maturity, and engineering velocity across on-premises, hybrid, and cloud environments.
The role will drive the adoption of Service Level Indicators (SLIs), Service Level Objectives (SLOs), error budgets, observability standards, automation frameworks, incident management practices, and production-readiness expectations. The successful candidate will establish reliability as a measurable business and engineering outcome by creating repeatable patterns, paved-road standards, and operating practices that reduce toil and improve service performance.
This is a strategic, hands-on leadership position requiring strong people leadership, executive communication, program development, and cross-functional influence. The leader will partner with senior stakeholders across Technology, Product, Architecture, Security, Risk, Compliance, and Operations to make reliability a core product and platform capability.
Responsibilities:
- Establish the enterprise SRE strategy, vision, roadmap, operating model, and governance framework in alignment with business priorities, technology modernization, cloud adoption, and operational resilience goals.
- Build, manage, and develop a centralized SRE team, including defining roles, creating staffing plans, establishing performance expectations, coaching team members, supporting career development, and planning for succession.
- Define the SRE engagement model, including service tiers, onboarding criteria, production-readiness standards, intake and prioritization processes, and partnership expectations for product, platform, infrastructure, and application teams.
- Define and operationalize enterprise reliability measures, including SLIs, SLOs, error budgets, availability targets, toil-reduction goals, incident metrics, change failure rate, mean time to detect, and mean time to restore.
- Establish error-budget policies and executive-level decision frameworks that guide tradeoffs among delivery velocity, operational risk, and service stability.
- Develop and govern enterprise reliability standards and reusable assets, including golden-signal frameworks, observability patterns, alerting standards, runbook templates, SLO dashboards, incident playbooks, and production-readiness checklists.
- Advance incident management maturity through effective high-severity response, executive communications, blameless post-incident reviews, root-cause analysis, corrective-action tracking, and measurable reliability improvements.
- Oversee operational readiness, capacity planning, disaster-recovery validation, resilience testing, service-health reviews, and reliability risk management for supported platforms and critical business services.
- Drive automation and observability strategies that reduce operational toil, improve visibility, accelerate recovery, and enable scalable support models across hybrid and cloud platforms.
- Identify, sponsor, and govern AI-enabled reliability capabilities, including alert-noise reduction, incident summarization, event correlation, runbook assistance, predictive operations, and approved automated remediation.
- Ensure AI-enabled operational capabilities are implemented with appropriate privacy, security, risk, compliance, auditability, and human-oversight controls.
- Partner with Security, Risk, Compliance, Audit, Architecture, Product, Development, Infrastructure, Cloud Operations, and Technology Operations leaders to embed reliability requirements throughout the software development lifecycle and production support model.
- Communicate reliability posture, risks, progress, business impact, and investment requirements to executive stakeholders using clear metrics and actionable recommendations.
- Promote shared ownership, engineering excellence, accountability, continuous improvement, and blameless learning across technology teams.
- Remain current on SRE, observability, platform engineering, AIOps, resilience engineering, and cloud reliability practices, applying relevant approaches to improve business and operational outcomes.
- Provide leadership during major incidents and critical operational events, including escalation management, cross-functional coordination, stakeholder communications, and executive updates.
- Progressive leadership experience in Site Reliability Engineering, infrastructure engineering, platform engineering, cloud operations, production engineering, technology operations, or a closely related discipline.
- Demonstrated success establishing a new SRE capability or significantly maturing an existing practice, including strategy, operating model, governance, service engagement, SLO management, production readiness, incident management, and roadmap execution.
- Proven experience hiring, managing, coaching, and developing engineers or other technical professionals.
- Strong executive presence and communication skills, with the ability to translate complex reliability, risk, and operational issues into clear business impacts, priorities, and decisions.
- Deep understanding of SRE practices, including SLIs, SLOs, error budgets, observability, incident command, post-incident reviews, toil reduction, automation, capacity planning, resilience testing, and production readiness.
- Experience leading high-severity incident response, coordinating cross-functional teams, communicating with senior stakeholders, and ensuring corrective actions are completed and evaluated for effectiveness.
- Strong knowledge of cloud and hybrid infrastructure, infrastructure as code, CI/CD, containerization, orchestration, monitoring, logging, distributed tracing, and modern production operations tooling.
- Ability to establish measurable programs, manage competing priorities, balance reliability with delivery velocity, and align technical investments with business criticality and risk reduction.
- Experience partnering with Security, Risk, Compliance, Audit, Architecture, Product, Development, Infrastructure, and Operations teams to establish secure, auditable, and sustainable reliability practices.
- Working knowledge of AIOps, predictive operations, AI-assisted incident response, observability automation, or related capabilities, including responsible-use and governance considerations.
- Only candidates available and ready to work directly as Genesis10 employees will be considered for this position.
- Education and Experience
- Bachelor's degree in Computer Science, Engineering, Information Technology, or a related discipline, or equivalent professional experience.
- Master's degree or comparable technology leadership experience preferred.
- At least 10 years of progressive technology experience, including five or more years in SRE, infrastructure engineering, platform engineering, cloud operations, production engineering, or technology operations leadership.
- At least three years of people-management experience leading engineers or technical teams, preferably within reliability, infrastructure, platform, cloud, or technology operations functions.
- Experience establishing an enterprise SRE function within a large, complex, regulated, or highly distributed technology environment.
- Experience supporting business-critical services across on-premises, hybrid, and public-cloud platforms.
- Experience developing executive-level reliability reporting, service-health reviews, and investment recommendations.
- Relevant certifications in AWS, Microsoft Azure, Google Cloud, Terraform, Kubernetes, ITIL, DevOps, reliability engineering, automation, or technology leadership.
- The ideal candidate combines enterprise strategy, technical credibility, people leadership, and operational judgment. This individual can build an SRE function from the ground up, establish practical governance without creating unnecessary friction, and influence senior leaders across multiple technology and business disciplines.
- The successful candidate will be comfortable operating at both strategic and execution levels—defining the long-term reliability model while helping teams address immediate operational risks, improve incident response, implement measurable SLOs, and create repeatable engineering standards.
If you have the described qualifications and are interested in this exciting opportunity, please apply!
Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals.
For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10:
- Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years.
- The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years.
- Access to an experienced, caring recruiting team (more than 7 years of experience, on average.)
- Behavioral Health Platform
- Medical, Dental, Vision
- Health Savings Account
- Voluntary Hospital Indemnity (Critical Illness & Accident)
- Voluntary Term Life Insurance
- 401K
- Sick Pay (for applicable states/municipalities)
- Commuter Benefits (Dallas, NYC, SF)
- Remote opportunities available
For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance, as a Best Company for Career Growth, for Diversity, and for Leadership, amongst others. To learn more and to view all our available career opportunities, please visit us at our website.
Genesis10 is an Equal Opportunity Employer. Candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
- ...Lovelace is the only provider of enterprise-scale context engines capable of analyzing trillions of... ...in 2023 by Andrew Moore, former head of Google Cloud AI, dean of Carnegie... ...a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team. As an...SuggestedFull time
$71.6k - $119.4k
...Our services provide applications with reliability, security, and better customer experiences... ...issues, and work closely with senior engineers to learn and apply best practices. You’ll... ...: ~1–3 years of experience in DevOps, SRE, cloud engineering, or related IT roles...SuggestedFull timeTemporary workInternshipLocal areaWork from home- ...Senior Site Reliability Engineer (SRE) Location: Pittsburgh, PA / Cleveland, OH / Dallas, TX FTE Position Overview We are seeking an... ...Managed GlassBox ITCAM / ITCAMS TrueSight Oracle Enterprise Manager (OEM) Additional Skills Agile methodology...SuggestedFull timeLocal areaShift workWeekend work
- ...This role can sit in our NYC HQ on a hybrid basis, or it can be fully remote while... ...looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and... ...Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure...SuggestedRemote workWorldwide
$132.23k - $176.31k
...performance connectivity across cloud, edge, and AI workloads for enterprises, governments, and communities. At Lumen, you’ll work on... ...are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and...SuggestedFull timeTemporary workRemote work$182.8k - $247.3k
...mission to develop education for our half a billion (and growing!) learners around the world.About the role...As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure Duolingo’s sophisticated distributed systems...Work experience placement$135.2k - $306.4k
...Senior / Principal Software Engineer or Architect Database Engine... ...platform used by demanding enterprise customers at scale. This... ...layer: document processing, hybrid lexical/vector retrieval,... ...correctness, and operational reliability. This area is especially...Full timeTemporary workFlexible hours- ...Job Title: Site Reliability Engineer Duration- Fulltime Permanent Location: Pittsburgh, PA/Strongsville, OH (Onsite From Day 1) Job Description: Skill: Site Reliability Engineer (Full Stack Developer) Must skills: Part of a full stack agile...Permanent employmentFull timeFlexible hours
- ...Site Reliability Engineer We are looking for a Site Reliability Engineer who can partner with us to bring the team and site stability to the next level. This candidate will help us ensure site stability through proactive and reactive analysis of site performance as...
$148k - $249k
...stakeholders and leadership. Qualifications: - 5+ years software engineering or systems/performance engineering experience (BS in CS/EE or... ...scheduled team building activities and social events both on-site, off-site & virtually. - As we grow, this list continues to...Full timeWork at officeWork from homeFlexible hours$70.8k - $156.7k
Senior Site Reliability Engineer - Local to Cleveland, Pittsburgh, or Dallas Position Description This role will require someone onsite at our client office in Cleveland, OH, Pittsburgh, PA, or Dallas, TX. Love technology? We do too. CGI is looking for a Site...Work at officeLocal areaFlexible hoursShift workWeekend work- ...Job Description Job Description Job Title: Senior Site Reliability Engineer Job Category: Infrastructure/Cloud Job Type: Permanent Full Time Location: Pittsburgh, Pennsylvania, United States Position Description This role will require someone onsite...Permanent employmentFull timeWork at officeFlexible hoursShift workWeekend work
$100.22k - $111.18k
...Qualifications Requires a Bachelor’s degree in Systems Engineering, or a related Science, Engineering, Technology or... ...benefitsWorkplace Options:This position is located in Pittsburgh PA. On site work is required and Hybrid/Flex work schedule is permitted. Please note that this...For subcontractorSecond jobWork at officeFlexible hours- ...future.**Please note: this on-site position is based at our... ...practices and apply them to a hybrid cloud/datacenter model are strongly... ...Overview:The Senior Platform Engineer is a senior‑level technical... ...engineering, and optimizing enterprise‑grade cloud platforms that support...Full timeWork at officeLocal areaRelocation
$126k - $201k
...efficient and accessible for all. We’re searching for a Software Engineer to Vehicle Platform Integrations Team.In this role, you... ...our ability to lead effectively. As a result, we operate in a hybrid work environment where Aurorans are in office at least 3 days per...Work at officeLocal area3 days per week$159k - $207k
...We work at the intersection of software engineering, machine learning, sensors, and hardware... ...positive impact on the world.This role is hybrid from our Boston office. It requires two... ...making autonomous vehicles a safe, reliable, and accessible reality. We’re driven by...Work at office2 days per week$147k - $211k
...Bachelor’s degree in Computer Science, Engineering, Mathematics, Information Systems or a... ...including validity, verification, performance, reliability, usability, and stress testing; and... ...to the Google office & may allow for a hybrid schedule as per Google policy.Bachelor’...Full timeWork at office$139k - $223k
...This RoleWe are looking for a Software Engineer to partner with our Mapping Infrastructure... ...a proven ability to deliver scalable, reliable backend systemsA commitment to writing robust... .... As a result, we operate in a hybrid work environment where Aurorans are in office...Work at officeLocal area3 days per week- ...opportunity for you! We are seeking a talented Senior Software Engineer to join our team and help us enhance our digital presence and improve... ...Leave, Fertility and Adoption Assistance Program Flexibility: Hybrid Work Model (For most professional roles) Training: Hands-On,...Full timeFlexible hours
$172k - $229k
...driverless revolution. We are seeking a Staff Engineer Team Lead to grow and lead our Release... ...infrastructure that ensures safe and reliable software releases across our autonomous... ...applications and environments.This role is hybrid from our Pittsburgh or Las Vegas office....Work experience placementWork at officeShift work- Senior Software Engineer - Build Cutting-Edge Solutions ? Location: [Specify Remote, Hybrid, or Onsite Location] ? Job Type: [Full-Time]About the RoleWe are seeking a Senior Software Engineer to play a key role in designing and developing innovative software solutions.....Full timeWork at officeRemote workFlexible hours
- ...electric energy, providing a secure supply of reliable power to more than half a million... ...you to join our team! The Reliability Engineer will provide technical support to the Operations... ...other duties as assigned. Location: Hybrid, Pittsburgh, PA at Woods Run Complex...Permanent employmentTemporary workLocal area
- ...Site Reliability Maintenance Engineer This is an exciting new position meant to be a key player in our newly created Reliability Program. The role is responsible for identifying and managing reliability improvements to steel producing equipment and facilities, minimizing...
$171k - $273k
...LinkedIn.What we are looking forWe’re searching for a Staff Software Engineer to join our Aurora Services team. The Aurora Services team sits... ...our ability to lead effectively. As a result, we operate in a hybrid work environment where Aurorans are in office at least 3 days...Work at officeLocal area3 days per week$179.2k - $268.8k
...systems, test operations, systems and safety engineering - all dedicated to redefining the... ...teams across the company to deploy software reliably to autonomous vehicles.This role sits at... ...to have: Experience in a DevOps, SRE, or infrastructure role in autonomous vehicles...Permanent employmentFull timeWork at officeImmediate startVisa sponsorship$171k - $273k
...accessible for all. We are searching for a Staff Software Assurance Engineer on the Safety Case and Quality Management team who is a... ...our ability to lead effectively. As a result, we operate in a hybrid work environment where Aurorans are in office at least 3 days per...Work at officeLocal area3 days per week$118.57k - $125.05k
...Requirements:Requires a Bachelor’s degree in Software Engineering, or a related Science, Engineering, Technology... ..., or Boulder. If located in Boulder, the flex site will be a customer site with 2-3 days onsite per week.#LI-Hybrid#LI-JH1#CJ3#SGS Salary Note This estimate...Relocation packageFlexible hours2 days per week3 days per week- ...At Deloitte, Forward Deployed Engineers (FDE) don’t just build AI... ...clients turn AI ambition into enterprise-scale impact, pairing leading... ...deployed onshore with clients or in hybrid onshore/offshore... ...workstream engagements, ensuring reliable architecture and consistent client...Local areaShift work
$155k - $241k
...Senior Software Engineer, Navigation Agility's commercially deployed humanoids operate... ...optimization-based methods (MPC/LQR), and hybrid A*. ~ Expert proficiency in modern C++... ...a winter shutdown, annually. On-Site Perks: Catered lunches four times a week...Full timeTemporary workLocal areaRelocation packageFlexible hours$147.9k - $220k
...Cloud Infrastructure / Site Reliability Engineer As a Cloud Infrastructure / Site Reliability Engineer, you will operate at the intersection of... ...and Database in cloud-based SaaS/IaaS environments. Implement SRE best practices for effective resolution. Document system...Odd jobLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Head of Enterprise Site Reliability Engineering (SRE) - Hybrid. Be the first to apply!
- engineering director Pittsburgh, PA
- principal engineer Pittsburgh, PA
- general engineer Pittsburgh, PA
- data center chief engineer Pittsburgh, PA
- hotel chief engineer Pittsburgh, PA
- principal developer Pittsburgh, PA
- senior principal engineer Pittsburgh, PA
- senior director engineering Pittsburgh, PA
- senior chief engineer Pittsburgh, PA
- chief engineer Pittsburgh, PA




