Head of Enterprise Site Reliability Engineering (SRE) - Hybrid
$85 - $95 per hourGenesis10
Job Description
Job Description
Genesis10 is currently seeking a AVP / Head of Enterprise Site Reliability Engineering (SRE) with our consumer finance lender firm client in their Pittsburgh, PA location. This is a Right to hire position.
An experienced Site Reliability Engineering leader is needed to establish, lead, and scale a formal enterprise SRE capability. This is a contract-to-hire opportunity intended to convert to a full-time AVP-level position based on performance, organizational approval, and mutually agreed-upon terms.
This leader will define the enterprise SRE strategy, operating model, governance framework, roadmap, and success measures while building and managing a centralized team focused on reliability, resilience, operational maturity, and engineering velocity across on-premises, hybrid, and cloud environments.
The role will drive the adoption of Service Level Indicators (SLIs), Service Level Objectives (SLOs), error budgets, observability standards, automation frameworks, incident management practices, and production-readiness expectations. The successful candidate will establish reliability as a measurable business and engineering outcome by creating repeatable patterns, paved-road standards, and operating practices that reduce toil and improve service performance.
This is a strategic, hands-on leadership position requiring strong people leadership, executive communication, program development, and cross-functional influence. The leader will partner with senior stakeholders across Technology, Product, Architecture, Security, Risk, Compliance, and Operations to make reliability a core product and platform capability.
Responsibilities:
- Establish the enterprise SRE strategy, vision, roadmap, operating model, and governance framework in alignment with business priorities, technology modernization, cloud adoption, and operational resilience goals.
- Build, manage, and develop a centralized SRE team, including defining roles, creating staffing plans, establishing performance expectations, coaching team members, supporting career development, and planning for succession.
- Define the SRE engagement model, including service tiers, onboarding criteria, production-readiness standards, intake and prioritization processes, and partnership expectations for product, platform, infrastructure, and application teams.
- Define and operationalize enterprise reliability measures, including SLIs, SLOs, error budgets, availability targets, toil-reduction goals, incident metrics, change failure rate, mean time to detect, and mean time to restore.
- Establish error-budget policies and executive-level decision frameworks that guide tradeoffs among delivery velocity, operational risk, and service stability.
- Develop and govern enterprise reliability standards and reusable assets, including golden-signal frameworks, observability patterns, alerting standards, runbook templates, SLO dashboards, incident playbooks, and production-readiness checklists.
- Advance incident management maturity through effective high-severity response, executive communications, blameless post-incident reviews, root-cause analysis, corrective-action tracking, and measurable reliability improvements.
- Oversee operational readiness, capacity planning, disaster-recovery validation, resilience testing, service-health reviews, and reliability risk management for supported platforms and critical business services.
- Drive automation and observability strategies that reduce operational toil, improve visibility, accelerate recovery, and enable scalable support models across hybrid and cloud platforms.
- Identify, sponsor, and govern AI-enabled reliability capabilities, including alert-noise reduction, incident summarization, event correlation, runbook assistance, predictive operations, and approved automated remediation.
- Ensure AI-enabled operational capabilities are implemented with appropriate privacy, security, risk, compliance, auditability, and human-oversight controls.
- Partner with Security, Risk, Compliance, Audit, Architecture, Product, Development, Infrastructure, Cloud Operations, and Technology Operations leaders to embed reliability requirements throughout the software development lifecycle and production support model.
- Communicate reliability posture, risks, progress, business impact, and investment requirements to executive stakeholders using clear metrics and actionable recommendations.
- Promote shared ownership, engineering excellence, accountability, continuous improvement, and blameless learning across technology teams.
- Remain current on SRE, observability, platform engineering, AIOps, resilience engineering, and cloud reliability practices, applying relevant approaches to improve business and operational outcomes.
- Provide leadership during major incidents and critical operational events, including escalation management, cross-functional coordination, stakeholder communications, and executive updates.
- Progressive leadership experience in Site Reliability Engineering, infrastructure engineering, platform engineering, cloud operations, production engineering, technology operations, or a closely related discipline.
- Demonstrated success establishing a new SRE capability or significantly maturing an existing practice, including strategy, operating model, governance, service engagement, SLO management, production readiness, incident management, and roadmap execution.
- Proven experience hiring, managing, coaching, and developing engineers or other technical professionals.
- Strong executive presence and communication skills, with the ability to translate complex reliability, risk, and operational issues into clear business impacts, priorities, and decisions.
- Deep understanding of SRE practices, including SLIs, SLOs, error budgets, observability, incident command, post-incident reviews, toil reduction, automation, capacity planning, resilience testing, and production readiness.
- Experience leading high-severity incident response, coordinating cross-functional teams, communicating with senior stakeholders, and ensuring corrective actions are completed and evaluated for effectiveness.
- Strong knowledge of cloud and hybrid infrastructure, infrastructure as code, CI/CD, containerization, orchestration, monitoring, logging, distributed tracing, and modern production operations tooling.
- Ability to establish measurable programs, manage competing priorities, balance reliability with delivery velocity, and align technical investments with business criticality and risk reduction.
- Experience partnering with Security, Risk, Compliance, Audit, Architecture, Product, Development, Infrastructure, and Operations teams to establish secure, auditable, and sustainable reliability practices.
- Working knowledge of AIOps, predictive operations, AI-assisted incident response, observability automation, or related capabilities, including responsible-use and governance considerations.
- Only candidates available and ready to work directly as Genesis10 employees will be considered for this position.
- Education and Experience
- Bachelor's degree in Computer Science, Engineering, Information Technology, or a related discipline, or equivalent professional experience.
- Master's degree or comparable technology leadership experience preferred.
- At least 10 years of progressive technology experience, including five or more years in SRE, infrastructure engineering, platform engineering, cloud operations, production engineering, or technology operations leadership.
- At least three years of people-management experience leading engineers or technical teams, preferably within reliability, infrastructure, platform, cloud, or technology operations functions.
- Experience establishing an enterprise SRE function within a large, complex, regulated, or highly distributed technology environment.
- Experience supporting business-critical services across on-premises, hybrid, and public-cloud platforms.
- Experience developing executive-level reliability reporting, service-health reviews, and investment recommendations.
- Relevant certifications in AWS, Microsoft Azure, Google Cloud, Terraform, Kubernetes, ITIL, DevOps, reliability engineering, automation, or technology leadership.
- The ideal candidate combines enterprise strategy, technical credibility, people leadership, and operational judgment. This individual can build an SRE function from the ground up, establish practical governance without creating unnecessary friction, and influence senior leaders across multiple technology and business disciplines.
- The successful candidate will be comfortable operating at both strategic and execution levels—defining the long-term reliability model while helping teams address immediate operational risks, improve incident response, implement measurable SLOs, and create repeatable engineering standards.
If you have the described qualifications and are interested in this exciting opportunity, please apply!
Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals.
For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10:
- Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years.
- The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years.
- Access to an experienced, caring recruiting team (more than 7 years of experience, on average.)
- Behavioral Health Platform
- Medical, Dental, Vision
- Health Savings Account
- Voluntary Hospital Indemnity (Critical Illness & Accident)
- Voluntary Term Life Insurance
- 401K
- Sick Pay (for applicable states/municipalities)
- Commuter Benefits (Dallas, NYC, SF)
- Remote opportunities available
For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance, as a Best Company for Career Growth, for Diversity, and for Leadership, amongst others. To learn more and to view all our available career opportunities, please visit us at our website.
Genesis10 is an Equal Opportunity Employer. Candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
- ...Lovelace is the only provider of enterprise-scale context engines capable of analyzing trillions of... ...in 2023 by Andrew Moore, former head of Google Cloud AI, dean of Carnegie... ...a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team. As an...SuggestedFull time
- ...Senior Site Reliability Engineer (SRE) Location: Pittsburgh, PA / Cleveland, OH / Dallas, TX FTE Position Overview We are seeking an... ...Managed GlassBox ITCAM / ITCAMS TrueSight Oracle Enterprise Manager (OEM) Additional Skills Agile methodology...SuggestedFull timeLocal areaShift workWeekend work
- ...This role can sit in our NYC HQ on a hybrid basis, or it can be fully remote while... ...looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and... ...Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure...SuggestedFull timeRemote workWorldwide
$135.2k - $306.4k
...Senior / Principal Software Engineer or Architect Database Engine... ...platform used by demanding enterprise customers at scale. This... ...layer: document processing, hybrid lexical/vector retrieval,... ...correctness, and operational reliability. This area is especially...SuggestedFull timeTemporary workFlexible hours- ...Job Description Job Description Job Title: Senior Site Reliability Engineer Job Category: Infrastructure/Cloud Job Type: Permanent Full Time Location: Pittsburgh, Pennsylvania, United States Position Description This role will require someone onsite...SuggestedPermanent employmentFull timeWork at officeFlexible hoursShift workWeekend work
$70.8k - $156.7k
Senior Site Reliability Engineer - Local to Cleveland, Pittsburgh, or Dallas Position Description This role will require someone onsite at our client office in Cleveland, OH, Pittsburgh, PA, or Dallas, TX. Love technology? We do too. CGI is looking for a Site...Work at officeLocal areaFlexible hoursShift workWeekend work$100.22k - $111.18k
...Qualifications Requires a Bachelor's degree in Systems Engineering, or a related Science, Engineering, Technology or... ...Options: This position is located in Pittsburgh PA. On site work is required and Hybrid/Flex work schedule is permitted. Please note that this...For subcontractorSecond jobWork at officeFlexible hours- ...Senior Platform Engineer Primary Office Location: 626 Washington... ...future. Please note: this on-site position is based at our... ...practices and apply them to a hybrid cloud/datacenter model are strongly... ...engineering, and optimizing enterprise-grade cloud platforms that...Work at officeLocal areaRelocation
$120k - $140k
...neuroscientists, and consumer electronics engineers is dedicated to delivering prescription-... .... ***This is a full-time, hybrid position located in our Pittsburgh office... ...involving application behavior, service reliability, data integrity, and customer-facing systems...Full timeWork at office- ...as Azure DevOps and GitHub # Work with enterprise data stored in relational databases,... ...quality, maintainability, performance, and reliability # Support fusion development initiatives... ...Resources offers a competitive salary, hybrid work schedule and a comprehensive...Local areaRelocation
$155k - $241k
...Senior Software Engineer, Navigation Agility's commercially deployed humanoids operate... ...optimization-based methods (MPC/LQR), and hybrid A*. ~ Expert proficiency in modern C++... ...a winter shutdown, annually. On-Site Perks: Catered lunches four times a week...Full timeTemporary workLocal areaRelocation packageFlexible hours- ...the world. Position: Sr. Applications Engineer Location: Headquarters – Moon Township... ...plans, competitive pay, dress for your day, hybrid schedules, paid time off (vacation... ...troubleshooting of adsorption equipment at customer sites is requiredBusiness/marketing experience...Full timeRemote workWorldwideOverseasMonday to Friday
- ...pivotal moment: live systems, real enterprise customers, and a platform being deliberately re-engineered from growth-era tradeoffs into... ...engineering work. This is a hybrid role based at our Pittsburgh,... ...targets for scalability, reliability, availability, latency, performance...Shift work
- ...Senior API Platform Software Engineer Fulltime Hybrid – Knoxville, TN | Columbia, SC | Lafayette, LA | Birmingham, AL Position Overview... ...to work on high-impact cloud data initiatives supporting enterprise risk and analytics with one of world class fortune 500 company...Full timeLocal area
$116.36k - $155.15k
..., edge, and AI workloads for enterprises, governments, and communities... ...experience in system architecture and engineering disciplines. Specific... ...activities including site surveys, design, design review... ...problems ~ Ability to travel ~ Hybrid role where the candidate will...Full timeTemporary workRemote work1 day per week$86.25k - $158.13k
...to the company's success. As a Software Engineer Lead within PNC's Technology organization... ...Modern Development Practices: DevOps and SRE principles Agile/Scrum methodologies... ...ensure they adhere to and support PNC's Enterprise Risk Management Framework. Qualifications...Full timeTemporary workPart timeWork experience placementWork at office$60 - $67 per hour
...you will design, develop, and maintain enterprise-grade applications that support critical... ...closely with architects, product owners, QA engineers, and DevOps teams to deliver scalable,... ...to work independently in a remote or hybrid environment. Preferred Qualifications...Full timeContract workRemote work- ...Sr Devops Engineer Seasoned and impact-driven Sr. DevOps Engineer with 10 years of experience... ...across AWS, multi-cloud, and hybrid environments. Highly proficient in Terraform... ...and compliance enforcement in regulated enterprise environments. Hands-on with usage and...
$142.8k - $274.8k
...enabled world. Microsoft's Azure Data engineering team is leading the transformation of... ...to help us accelerate modernization of reliable and cost-efficient managed telemetry pipelines... .... - Partner with product management, SRE, security, compliance, CSS, and platform...Ongoing contractLocal area- ...Description Description: Who WE Are: ST Engineering Aethon, Inc is a forward-thinking... ...Software Engineering · Work Model: Hybrid. Flexible with minimum of 3 days/week... ...system-level tests to ensure software reliability and performance. · Analyze and solve...Local areaImmediate startFlexible hours3 days per week
$50 - $90 per hour
...Job Title: Sr. Software Perception Engineer Job Description The Sr.... ...ensuring high performance, safety, and reliability. You contribute to the full... ...Environment This role is based on-site in Pittsburgh PA OR Dallas TX with a hybrid schedule available, offering flexibility...Permanent employmentTemporary work- ...Description Job Description Description: Hybrid 4 days onsite in Pittsburgh, PA Our client seeks a Senior Software Engineer to build a secure, fault-tolerant Azure-... ...delivery guarantees. Ensure persistent and reliable input delivery to prevent client...Hourly payContract workLocal area
- ...Were looking for a Senior Software Engineer to join our healthcare-related client in a hybrid-schedule role based in Pittsburgh, PA . This is a strong opportunity... ....  We need someone who can build and support reliable file-processing solutions, strengthen validation...Full time
$40 per hour
...The Software Engineer Intern implements developer tools or product features on a rapid-release... ...Working Conditions Minimal travel Reliable internet access for any period of time... ...hubs, you are welcome to work in a hybrid capacity and utilize our office spaces....Remote jobSummer workInternshipSummer internshipWork at office- ...Solutions is seeking a Senior Embedded Software Engineer for a 6- to 12-month contract... ...systems for the transportation industry. This hybrid position requires candidates to... ...system and engineering requirements into reliable software solutions Participate in software...Contract work
$182.8k - $247.3k
...mission to develop education for our half a billion (and growing!) learners around the world.About the role...As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure Duolingo’s sophisticated distributed systems...Work experience placement$51 - $61 per hour
...onsite at the project, significantly reducing and/or eliminating the demands to travel. Key Responsibilities:As a Release Train Engineer, you will be responsible for facilitating Agile Release Train events and processes including communicating with stakeholders...Hourly payLive inWork at officeLocal areaImmediate startFlexible hoursShift work$58k - $123.8k
...bitbucket, SQL. You will collaborate with engineering, platform, SRE, and security partners to automate delivery workflows, strengthen reliability and controls, and enable teams to ship... ...! Click here to be directed to our site that is dedicated to veterans and transitioning...Work at officeLocal areaRemote work- ...Full Time Location : Chicago, IL or Lafayette, LA Work Mode : Hybrid Summary: The EDI EDIFECS Application Developer will design,... ...bottlenecks • Collaborate with Product Owners, Quality Engineers, Scrum Masters, and Software Engineers on your team as well as...Permanent employmentFull timeWork experience placement
- ...Description Job Description Job Title: Software Engineer (Python) Duration : Contract to Hire... ...will collaborate with engineering, platform, SRE, and security partners to automate delivery workflows, strengthen reliability and controls, and enable teams to ship...Contract workLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Head of Enterprise Site Reliability Engineering (SRE) - Hybrid. Be the first to apply!
- engineering director Pittsburgh, PA
- principal engineer Pittsburgh, PA
- general engineer Pittsburgh, PA
- data center chief engineer Pittsburgh, PA
- hotel chief engineer Pittsburgh, PA
- principal developer Pittsburgh, PA
- senior principal engineer Pittsburgh, PA
- senior director engineering Pittsburgh, PA
- senior chief engineer Pittsburgh, PA
- chief engineer Pittsburgh, PA





