Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Head of Enterprise Site Reliability Engineering (SRE) - Hybrid

$85 - $95 per hour

Genesis10

Job Description

Job Description

Genesis10 is currently seeking a AVP / Head of Enterprise Site Reliability Engineering (SRE) with our consumer finance lender firm client in their Pittsburgh, PA location. This is a Right to hire position.

Summary:
An experienced Site Reliability Engineering leader is needed to establish, lead, and scale a formal enterprise SRE capability. This is a contract-to-hire opportunity intended to convert to a full-time AVP-level position based on performance, organizational approval, and mutually agreed-upon terms.

This leader will define the enterprise SRE strategy, operating model, governance framework, roadmap, and success measures while building and managing a centralized team focused on reliability, resilience, operational maturity, and engineering velocity across on-premises, hybrid, and cloud environments.

The role will drive the adoption of Service Level Indicators (SLIs), Service Level Objectives (SLOs), error budgets, observability standards, automation frameworks, incident management practices, and production-readiness expectations. The successful candidate will establish reliability as a measurable business and engineering outcome by creating repeatable patterns, paved-road standards, and operating practices that reduce toil and improve service performance.

This is a strategic, hands-on leadership position requiring strong people leadership, executive communication, program development, and cross-functional influence. The leader will partner with senior stakeholders across Technology, Product, Architecture, Security, Risk, Compliance, and Operations to make reliability a core product and platform capability.

Responsibilities:
  • Establish the enterprise SRE strategy, vision, roadmap, operating model, and governance framework in alignment with business priorities, technology modernization, cloud adoption, and operational resilience goals.
  • Build, manage, and develop a centralized SRE team, including defining roles, creating staffing plans, establishing performance expectations, coaching team members, supporting career development, and planning for succession.
  • Define the SRE engagement model, including service tiers, onboarding criteria, production-readiness standards, intake and prioritization processes, and partnership expectations for product, platform, infrastructure, and application teams.
  • Define and operationalize enterprise reliability measures, including SLIs, SLOs, error budgets, availability targets, toil-reduction goals, incident metrics, change failure rate, mean time to detect, and mean time to restore.
  • Establish error-budget policies and executive-level decision frameworks that guide tradeoffs among delivery velocity, operational risk, and service stability.
  • Develop and govern enterprise reliability standards and reusable assets, including golden-signal frameworks, observability patterns, alerting standards, runbook templates, SLO dashboards, incident playbooks, and production-readiness checklists.
  • Advance incident management maturity through effective high-severity response, executive communications, blameless post-incident reviews, root-cause analysis, corrective-action tracking, and measurable reliability improvements.
  • Oversee operational readiness, capacity planning, disaster-recovery validation, resilience testing, service-health reviews, and reliability risk management for supported platforms and critical business services.
  • Drive automation and observability strategies that reduce operational toil, improve visibility, accelerate recovery, and enable scalable support models across hybrid and cloud platforms.
  • Identify, sponsor, and govern AI-enabled reliability capabilities, including alert-noise reduction, incident summarization, event correlation, runbook assistance, predictive operations, and approved automated remediation.
  • Ensure AI-enabled operational capabilities are implemented with appropriate privacy, security, risk, compliance, auditability, and human-oversight controls.
  • Partner with Security, Risk, Compliance, Audit, Architecture, Product, Development, Infrastructure, Cloud Operations, and Technology Operations leaders to embed reliability requirements throughout the software development lifecycle and production support model.
  • Communicate reliability posture, risks, progress, business impact, and investment requirements to executive stakeholders using clear metrics and actionable recommendations.
  • Promote shared ownership, engineering excellence, accountability, continuous improvement, and blameless learning across technology teams.
  • Remain current on SRE, observability, platform engineering, AIOps, resilience engineering, and cloud reliability practices, applying relevant approaches to improve business and operational outcomes.
  • Provide leadership during major incidents and critical operational events, including escalation management, cross-functional coordination, stakeholder communications, and executive updates.
Requirements:
  • Progressive leadership experience in Site Reliability Engineering, infrastructure engineering, platform engineering, cloud operations, production engineering, technology operations, or a closely related discipline.
  • Demonstrated success establishing a new SRE capability or significantly maturing an existing practice, including strategy, operating model, governance, service engagement, SLO management, production readiness, incident management, and roadmap execution.
  • Proven experience hiring, managing, coaching, and developing engineers or other technical professionals.
  • Strong executive presence and communication skills, with the ability to translate complex reliability, risk, and operational issues into clear business impacts, priorities, and decisions.
  • Deep understanding of SRE practices, including SLIs, SLOs, error budgets, observability, incident command, post-incident reviews, toil reduction, automation, capacity planning, resilience testing, and production readiness.
  • Experience leading high-severity incident response, coordinating cross-functional teams, communicating with senior stakeholders, and ensuring corrective actions are completed and evaluated for effectiveness.
  • Strong knowledge of cloud and hybrid infrastructure, infrastructure as code, CI/CD, containerization, orchestration, monitoring, logging, distributed tracing, and modern production operations tooling.
  • Ability to establish measurable programs, manage competing priorities, balance reliability with delivery velocity, and align technical investments with business criticality and risk reduction.
  • Experience partnering with Security, Risk, Compliance, Audit, Architecture, Product, Development, Infrastructure, and Operations teams to establish secure, auditable, and sustainable reliability practices.
  • Working knowledge of AIOps, predictive operations, AI-assisted incident response, observability automation, or related capabilities, including responsible-use and governance considerations.
  • Only candidates available and ready to work directly as Genesis10 employees will be considered for this position.
  • Education and Experience
    • Bachelor's degree in Computer Science, Engineering, Information Technology, or a related discipline, or equivalent professional experience.
    • Master's degree or comparable technology leadership experience preferred.
    • At least 10 years of progressive technology experience, including five or more years in SRE, infrastructure engineering, platform engineering, cloud operations, production engineering, or technology operations leadership.
    • At least three years of people-management experience leading engineers or technical teams, preferably within reliability, infrastructure, platform, cloud, or technology operations functions.
Preferred Qualifications:
  • Experience establishing an enterprise SRE function within a large, complex, regulated, or highly distributed technology environment.
  • Experience supporting business-critical services across on-premises, hybrid, and public-cloud platforms.
  • Experience developing executive-level reliability reporting, service-health reviews, and investment recommendations.
  • Relevant certifications in AWS, Microsoft Azure, Google Cloud, Terraform, Kubernetes, ITIL, DevOps, reliability engineering, automation, or technology leadership.
Ideal Candidate Profile:
  • The ideal candidate combines enterprise strategy, technical credibility, people leadership, and operational judgment. This individual can build an SRE function from the ground up, establish practical governance without creating unnecessary friction, and influence senior leaders across multiple technology and business disciplines.
  • The successful candidate will be comfortable operating at both strategic and execution levels—defining the long-term reliability model while helping teams address immediate operational risks, improve incident response, implement measurable SLOs, and create repeatable engineering standards.
Pay rate range: $85.00 - $95.00 hourly

If you have the described qualifications and are interested in this exciting opportunity, please apply!

Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals.

For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10:
  • Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years.
  • The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years.
  • Access to an experienced, caring recruiting team (more than 7 years of experience, on average.)
  • Behavioral Health Platform
  • Medical, Dental, Vision
  • Health Savings Account
  • Voluntary Hospital Indemnity (Critical Illness & Accident)
  • Voluntary Term Life Insurance
  • 401K
  • Sick Pay (for applicable states/municipalities)
  • Commuter Benefits (Dallas, NYC, SF)
  • Remote opportunities available

For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance, as a Best Company for Career Growth, for Diversity, and for Leadership, amongst others. To learn more and to view all our available career opportunities, please visit us at our website.
Genesis10 is an Equal Opportunity Employer. Candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Head of Enterprise Site Reliability Engineering (SRE) - Hybrid in Pittsburgh, PA vacancy
  •  ...Lovelace is the only provider of enterprise-scale context engines capable of analyzing trillions of...  ...in 2023 by Andrew Moore, former head of Google Cloud AI, dean of Carnegie...  ...a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team. As an... 
    Suggested
    Full time

    Lovelace Ai

    Pittsburgh, PA
    9 hours ago
  • $71.6k - $119.4k

     ...Our services provide applications with reliability, security, and better customer experiences...  ...issues, and work closely with senior engineers to learn and apply best practices. You’ll...  ...: ~1–3 years of experience in DevOps, SRE, cloud engineering, or related IT roles... 
    Suggested
    Full time
    Temporary work
    Internship
    Local area
    Work from home

    RELX

    Pittsburgh, PA
    23 hours ago
  •  ...Senior Site Reliability Engineer (SRE) Location: Pittsburgh, PA / Cleveland, OH / Dallas, TX FTE Position Overview We are seeking an...  ...Managed GlassBox ITCAM / ITCAMS TrueSight Oracle Enterprise Manager (OEM) Additional Skills Agile methodology... 
    Suggested
    Full time
    Local area
    Shift work
    Weekend work

    System One

    Pittsburgh, PA
    11 days ago
  •  ...This role can sit in our NYC HQ on a hybrid basis, or it can be fully remote while...  ...looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and...  ...Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure... 
    Suggested
    Remote work
    Worldwide

    Mongodb

    Pittsburgh, PA
    a month ago
  • $132.23k - $176.31k

     ...performance connectivity across cloud, edge, and AI workloads for enterprises, governments, and communities. At Lumen, you’ll work on...  ...are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and... 
    Suggested
    Full time
    Temporary work
    Remote work

    Lumen

    Pittsburgh, PA
    3 days ago
  • $182.8k - $247.3k

     ...mission to develop education for our half a billion (and growing!) learners around the world.About the role...As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure Duolingo’s sophisticated distributed systems... 
    Work experience placement

    Duolingo

    Pittsburgh, PA
    4 days ago
  • $135.2k - $306.4k

     ...Senior / Principal Software Engineer or Architect Database Engine...  ...platform used by demanding enterprise customers at scale. This...  ...layer: document processing, hybrid lexical/vector retrieval,...  ...correctness, and operational reliability. This area is especially... 
    Full time
    Temporary work
    Flexible hours

    Oracle

    Pittsburgh, PA
    2 days ago
  •  ...Job Title: Site Reliability Engineer Duration- Fulltime Permanent Location: Pittsburgh, PA/Strongsville, OH (Onsite From Day 1) Job Description: Skill: Site Reliability Engineer (Full Stack Developer) Must skills: Part of a full stack agile... 
    Permanent employment
    Full time
    Flexible hours

    Q1 Technologies

    Pittsburgh, PA
    3 days ago
  •  ...Site Reliability Engineer We are looking for a Site Reliability Engineer who can partner with us to bring the team and site stability to the next level. This candidate will help us ensure site stability through proactive and reactive analysis of site performance as... 

    Software Technology Inc

    Pittsburgh, PA
    1 day ago
  • $148k - $249k

     ...stakeholders and leadership. Qualifications: - 5+ years software engineering or systems/performance engineering experience (BS in CS/EE or...  ...scheduled team building activities and social events both on-site, off-site & virtually. - As we grow, this list continues to... 
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    Pittsburgh, PA
    2 days ago
  • $70.8k - $156.7k

    Senior Site Reliability Engineer - Local to Cleveland, Pittsburgh, or Dallas Position Description This role will require someone onsite at our client office in Cleveland, OH, Pittsburgh, PA, or Dallas, TX. Love technology? We do too. CGI is looking for a Site... 
    Work at office
    Local area
    Flexible hours
    Shift work
    Weekend work
    Pittsburgh, PA
    17 days ago
  •  ...Job Description Job Description Job Title: Senior Site Reliability Engineer Job Category: Infrastructure/Cloud Job Type: Permanent Full Time Location: Pittsburgh, Pennsylvania, United States Position Description This role will require someone onsite... 
    Permanent employment
    Full time
    Work at office
    Flexible hours
    Shift work
    Weekend work

    System One

    Pittsburgh, PA
    15 days ago
  • $100.22k - $111.18k

     ...Qualifications Requires a Bachelor’s degree in Systems Engineering, or a related Science, Engineering, Technology or...  ...benefitsWorkplace Options:This position is located in Pittsburgh PA. On site work is required and Hybrid/Flex work schedule is permitted. Please note that this... 
    For subcontractor
    Second job
    Work at office
    Flexible hours

    General Dynamics Mission Systems

    Pittsburgh, PA
    2 days ago
  •  ...future.**Please note: this on-site position is based at our...  ...practices and apply them to a hybrid cloud/datacenter model are strongly...  ...Overview:The Senior Platform Engineer is a senior‑level technical...  ...engineering, and optimizing enterprise‑grade cloud platforms that support... 
    Full time
    Work at office
    Local area
    Relocation

    First National Bank

    Pittsburgh, PA
    9 hours ago
  • $126k - $201k

     ...efficient and accessible for all. We’re searching for a Software Engineer to Vehicle Platform Integrations Team.In this role, you...  ...our ability to lead effectively. As a result, we operate in a hybrid work environment where Aurorans are in office at least 3 days per... 
    Work at office
    Local area
    3 days per week

    Aurora Innovation

    Pittsburgh, PA
    9 hours ago
  • $159k - $207k

     ...We work at the intersection of software engineering, machine learning, sensors, and hardware...  ...positive impact on the world.This role is hybrid from our Boston office. It requires two...  ...making autonomous vehicles a safe, reliable, and accessible reality. We’re driven by... 
    Work at office
    2 days per week

    Motional

    Pittsburgh, PA
    2 days ago
  • $147k - $211k

     ...Bachelor’s degree in Computer Science, Engineering, Mathematics, Information Systems or a...  ...including validity, verification, performance, reliability, usability, and stress testing; and...  ...to the Google office & may allow for a hybrid schedule as per Google policy.Bachelor’... 
    Full time
    Work at office

    Google

    Pittsburgh, PA
    3 days ago
  • $139k - $223k

     ...This RoleWe are looking for a Software Engineer to partner with our Mapping Infrastructure...  ...a proven ability to deliver scalable, reliable backend systemsA commitment to writing robust...  .... As a result, we operate in a hybrid work environment where Aurorans are in office... 
    Work at office
    Local area
    3 days per week

    Aurora Innovation

    Pittsburgh, PA
    9 hours ago
  •  ...opportunity for you! We are seeking a talented Senior Software Engineer to join our team and help us enhance our digital presence and improve...  ...Leave, Fertility and Adoption Assistance Program Flexibility: Hybrid Work Model (For most professional roles)  Training: Hands-On,... 
    Full time
    Flexible hours

    Smith & Nephew

    Pittsburgh, PA
    3 days ago
  • $172k - $229k

     ...driverless revolution. We are seeking a Staff Engineer Team Lead to grow and lead our Release...  ...infrastructure that ensures safe and reliable software releases across our autonomous...  ...applications and environments.This role is hybrid from our Pittsburgh or Las Vegas office.... 
    Work experience placement
    Work at office
    Shift work

    Motional

    Pittsburgh, PA
    2 days ago
  • Senior Software Engineer - Build Cutting-Edge Solutions ? Location: [Specify Remote, Hybrid, or Onsite Location] ? Job Type: [Full-Time]About the RoleWe are seeking a Senior Software Engineer to play a key role in designing and developing innovative software solutions..... 
    Full time
    Work at office
    Remote work
    Flexible hours

    Technical Search Consultants

    Pittsburgh, PA
    3 days ago
  •  ...electric energy, providing a secure supply of reliable power to more than half a million...  ...you to join our team! The Reliability Engineer will provide technical support to the Operations...  ...other duties as assigned. Location: Hybrid, Pittsburgh, PA at Woods Run Complex... 
    Permanent employment
    Temporary work
    Local area

    Duquesne Light

    Pittsburgh, PA
    3 days ago
  •  ...Site Reliability Maintenance Engineer This is an exciting new position meant to be a key player in our newly created Reliability Program. The role is responsible for identifying and managing reliability improvements to steel producing equipment and facilities, minimizing... 

    Universal Stainless

    Bridgeville, PA
    1 day ago
  • $171k - $273k

     ...LinkedIn.What we are looking forWe’re searching for a Staff Software Engineer to join our Aurora Services team. The Aurora Services team sits...  ...our ability to lead effectively. As a result, we operate in a hybrid work environment where Aurorans are in office at least 3 days... 
    Work at office
    Local area
    3 days per week

    Aurora Innovation

    Pittsburgh, PA
    3 days ago
  • $179.2k - $268.8k

     ...systems, test operations, systems and safety engineering - all dedicated to redefining the...  ...teams across the company to deploy software reliably to autonomous vehicles.This role sits at...  ...to have: Experience in a DevOps, SRE, or infrastructure role in autonomous vehicles... 
    Permanent employment
    Full time
    Work at office
    Immediate start
    Visa sponsorship

    Latitude AI

    Pittsburgh, PA
    2 days ago
  • $171k - $273k

     ...accessible for all. We are searching for a Staff Software Assurance Engineer on the Safety Case and Quality Management team who is a...  ...our ability to lead effectively. As a result, we operate in a hybrid work environment where Aurorans are in office at least 3 days per... 
    Work at office
    Local area
    3 days per week

    Aurora Innovation

    Pittsburgh, PA
    9 hours ago
  • $118.57k - $125.05k

     ...Requirements:Requires a Bachelor’s degree in Software Engineering, or a related Science, Engineering, Technology...  ..., or Boulder. If located in Boulder, the flex site will be a customer site with 2-3 days onsite per week.#LI-Hybrid#LI-JH1#CJ3#SGS Salary Note This estimate... 
    Relocation package
    Flexible hours
    2 days per week
    3 days per week

    General Dynamics Mission Systems

    Pittsburgh, PA
    4 days ago
  •  ...At Deloitte, Forward Deployed Engineers (FDE) don’t just build AI...  ...clients turn AI ambition into enterprise-scale impact, pairing leading...  ...deployed onshore with clients or in hybrid onshore/offshore...  ...workstream engagements, ensuring reliable architecture and consistent client... 
    Local area
    Shift work

    Deloitte

    Pittsburgh, PA
    1 day ago
  • $155k - $241k

     ...Senior Software Engineer, Navigation Agility's commercially deployed humanoids operate...  ...optimization-based methods (MPC/LQR), and hybrid A*. ~ Expert proficiency in modern C++...  ...a winter shutdown, annually. On-Site Perks: Catered lunches four times a week... 
    Full time
    Temporary work
    Local area
    Relocation package
    Flexible hours

    Agility Robotics

    Pittsburgh, PA
    1 day ago
  • $147.9k - $220k

     ...Cloud Infrastructure / Site Reliability Engineer As a Cloud Infrastructure / Site Reliability Engineer, you will operate at the intersection of...  ...and Database in cloud-based SaaS/IaaS environments. Implement SRE best practices for effective resolution. Document system... 
    Odd job
    Local area

    NetApp

    Pittsburgh, PA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Head of Enterprise Site Reliability Engineering (SRE) - Hybrid. Be the first to apply!