Executive Director, Site Reliability and Cloud Platform Strategy
Crucial Hire
Our client is a well-established Chicago financial services organization operating in one of the most heavily regulated environments in the industry. The platform engineering organization supports mission-critical systems where stability is not negotiable and every architectural decision carries regulatory weight. To protect the confidentiality of this search, additional details will be shared with qualified candidates as the process progresses. About the Role This is an executive seat with four distinct areas of ownership: Site Reliability Engineering, platform governance and compliance, cloud strategy and architecture, and metrics and reporting. The leader in this role scales a maturing SRE practice, sets cloud architectural standards and multi-year strategy, and serves as product manager for the FinOps and SecOps domains. They also own platform engineering's compliance obligations across incidents, risks, and audit findings. The client is looking for someone with real technical credibility who can also sit in front of executives and translate architectural trade-offs into business decisions. Responsibilities Site Reliability Engineering Scale and mature the SRE practice, establishing error budgets, SLOs, SLAs, and incident response frameworks across all platform services Define and enforce reliability standards including on-call models, blameless postmortem processes, and corrective action tracking Partner with platform foundation teams across Kubernetes, Kafka, FinOps, and Security to embed reliability into build and operate models Drive toil reduction through automation so engineering capacity moves from manual operations to higher-value platform work Governance and Compliance Serve as product manager for the FinOps and SecOps domains, owning product vision, prioritization, and stakeholder alignment for governance tooling Build and maintain a governance framework covering incident and problem management, risk tracking, and audit findings Own the end-to-end process for compliance obligations, driving timely resolution and clear accountability Partner with Risk, Compliance, and Security to identify governance gaps and drive remediation Report on compliance posture to senior leadership, surfacing trends, aging items, and residual risk Cloud Strategy and Architecture Define and execute the multi-year cloud architecture strategy against growth, scalability, regulatory, and cost objectives Establish cloud architectural standards, reference architectures, and governance frameworks covering landing zones, identity, network patterns, and service catalog Guide decisions on containers and orchestration, IaaS and PaaS adoption, disaster recovery, and multi-region patterns against regulatory requirements including CIS and NIST Own technology roadmaps and end-of-life planning for cloud platform components Serve as a key technical advisor to senior leadership Metrics and Reporting Own the platform metrics function, building a consistent framework for measuring platform health, engineering velocity, reliability, and cost efficiency Define and track KPIs aligned to internal SLAs, executive reporting, and audit requirements Ensure platform tooling is the single source of truth for work visibility Build reporting cadences including platform health scorecards, capacity forecasting, and risk transparency Delivery and Leadership Serve as the primary engineering leadership partner to platform program management, ensuring initiatives are scoped, sequenced, and resourced Align engineering capacity to roadmap commitments and surface dependency risks early Lead, develop, and retain a team of engineering managers and senior individual contributors Manage budget, schedules, and performance for areas of responsibility Oversee remediation of audit findings, ensuring root cause is addressed and residual risk is reduced Qualifications 15+ years in cloud engineering, platform reliability, or infrastructure, with at least 5 years in senior engineering leadership Executive-level leadership of SRE, cloud engineering, or platform reliability organizations in a regulated industry Proven ability to build and scale SRE practices including SLO and SLA frameworks, on-call models, error budgets, and incident response Deep expertise in cloud architecture strategy and governance, including enterprise-wide architectural standards Experience serving as product manager for technical domains such as FinOps, SecOps, or platform tooling Experience establishing governance and compliance frameworks within a platform or infrastructure organization Ability to design metrics and reporting frameworks for both technical and executive audiences Strong track record partnering with program and product management to turn platform capability into delivery-ready roadmaps Exceptional written and verbal communication Experience leading in Agile and Scrum environments Bachelor's degree, preferably in a technical discipline, or equivalent experience Preferred Skills "Financial services" or similarly regulated industry experience with exposure to CIS, NIST, and related frameworks Experience in production change control and working directly with audit and compliance functions Technical Skills Deep knowledge of SRE tooling and observability platforms such as Prometheus, Grafana, PagerDuty, or Datadog Expert-level knowledge of AWS, Azure, or GCP, with multi-cloud or hybrid experience preferred Strong working knowledge of cloud-native architecture patterns and Infrastructure as Code Familiarity with Kubernetes, Kafka, and CI/CD tooling such as GitHub Actions or Jenkins Working knowledge of FinOps principles and cloud cost governance at organizational scale Familiarity with SecOps tooling and cloud security governance GRC tooling experience such as ServiceNow or Archer preferred Certifications AWS Solutions Architect Associate or higher strongly desired Google Cloud Professional Cloud Architect, Azure Solutions Architect, or equivalent preferred #J-18808-Ljbffr Crucial Hire
$250k - $270k
Executive Director - SRE/Platform Strategy & Governance Salary: $250,000 to $270,000 Location: Chicago, ILTravel: 0%Job Categories... ...engineering governance, compliance, and site reliability. The director will focus on AWS cloud, Kubernetes, Kafka, FinOps, and SecOps within...PlatformCloudWebsitePermanent employment- ...through our careers site by directly... ...This role provides executive-level leadership over Platform Engineering Governance... ...Compliance, Site Reliability Engineering, Strategy & Architecture,... ..., driving cloud architectural standards... ...CARE Executive Director.Coordinate across...PlatformCloudWebsiteFull timeRemote work2 days per week
- About the Role Lead the Platform Engineering organization at OCC, overseeing... ...Engineering Governance & Compliance, Site Reliability Engineering, Cloud Strategy & Architecture, and Metrics &... ...maintain compliance posture. Define and execute multi‑year cloud architecture...PlatformCloudWebsiteLocal areaRemote work2 days per week
- Our client in Chicago seeks an executive leader for Site Reliability Engineering, platform governance, cloud strategy, and metrics. You will scale a mature SRE practice, set cloud standards, and translate architectural trade-offs into business decisions. You will own governance...PlatformCloudWebsite
- The Options Clearing Corporation is seeking a leader for its Platform Engineering organization, focusing on governance, compliance, and cloud architecture. This executive role involves scaling SRE practices, improving platform health, and overseeing FinOps and SecOps domains...PlatformCloud
- ...Senior Associate, Data Strategy Enablement and Innovation... ...scalable, secure, and reliable systems; support the... ...-stack development and platform enablementBachelor's degree... ...solutions on cloud platforms, with required... ...of our KPMG US Careers site at Benefits & How We Work...PlatformCloudWebsiteH1bLocal area
- ...accepted only through our careers site by directly applying to the... ....What You'll Do:The Executive Director, Platform Architecture leads OCC's Platform... ...Architecture team, owning the technical strategy and standards for all infrastructure domains — cloud, onpremises, networking,...PlatformCloudWebsiteFull timeWork experience placementRemote work2 days per week
$165k - $225k
...bare-metal performance with cloud-native operational simplicity... ...workloads with enterprise-grade reliability and compliance. Your Role:... ..., network engineers, and platform engineering team, you'll architect... ...etcd management, and scaling strategies for high-performance compute...PlatformCloudWebsiteRemote workFlexible hours$195.42k - $370.53k
...currently seeking a Director in Deal Advisory and Strategy Analytics for... ...to build reliable and scalable data... ...Microsoft Azure Data Platform hands on... ...solutions using Cloud native technologies... ...teams and executives on data-related... ...KPMG US Careers site at Benefits & How...PlatformCloudWebsiteLocal area$158.5k - $172k
...restaurant technology, easy-to-use platforms, and an improved delivery... ...), engineering robust multi-cloud solutions across AWS and GCP... ...position driving continuous reliability, deep system optimization,... ...: help define architectural strategies, introducing modern platform...PlatformCloudWebsiteFull timeTemporary workWork at officeFlexible hours3 days per week- ...seeking a highly skilled Edge Site Reliability Engineer (Edge SRE) to lead... ...of Google Distributed Cloud Edge (GDCE) environments. This... ...deep expertise in cloud-native platforms, networking, and automation... ...steering, and edge routing strategies to reduce latency and maximize...PlatformCloudWebsiteFull timeContract work
$91.2k - $136.8k
Reliability Engineer - IE08GEWe’re determined to make a... ...performance of our systems in cloud and SAAS environments.... ...the technical strategy for the organization,... ...Infrastructure Engineering, Site Reliability... ...practices.Expertise in cloud platforms (AWS) and Kubernetes-based...PlatformCloudWebsiteFull timeTemporary workWork at office3 days per week$130k - $150k
...SolutionsInfrastructure, Networking and Cloud SolutionsInformation... ...Position OverviewThe Site Reliability Engineer (SRE) helps... ...infrastructure platforms and services,... ...monitoring/alerting strategy, and blameless... ...across stakeholders, execution of tabletop and technical...PlatformCloudWebsiteWork at officeWork from home3 days per week$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time... ...causes and develops preventive strategies.Drafts clear and insightful RCAs... ...incidents and reduce their impact on our platform.Partners with the larger Cloud Operations, SRE, Engineering...PlatformCloudWebsiteFull timeTemporary workWork experience placementFlexible hours- ...Zero Trust Exchange platform combined with... ...building a culture of execution centered on... ...are looking for a Site Reliability Engineer-SkillBridge... ...reporting to the Director, Site Reliability... ...world’s largest cloud security platform... ...a cloud-first strategy. What you’ll do (...PlatformCloudWebsiteInternshipWork at officeLocal areaRemote workWorldwide
$130k - $180k
...accomplishment.Being a Senior Site Reliability Engineer at iManage Means…... ...You’ll create middleware and platform guardrails that empower... ...SRE, you’ll help scale our cloud platform, collaborate across... ...CI/CD pipelines and rollout strategies. A bachelor’s degree (or equivalent...PlatformCloudWebsiteWork at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours$112.5k - $187.5k
...role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards... ...consequential work on the platform.As a Staff Site Reliability Engineer... ...5+ years of experience in Cloud Architecture, Site...PlatformCloudWebsiteFull timeTemporary workWork experience placementWork at officeFlexible hours2 days per week- ...Lenovo in the United States is seeking a Senior Site Reliability Engineer to strengthen reliability, observability, and operations for the Qira platform. You will work across device, edge, and cloud, collaborating with AI/ML, firmware, and platform teams to design scalable...PlatformCloudWebsiteRemote work
- ...Zero Trust Exchange platform combined with advanced... ...building a culture of execution centered on customer... ...We are looking for a Site Reliability Engineer-SkillBridge... ...the world’s largest cloud security platform, specifically... ...a cloud-first strategy. You'll bring your...PlatformCloudWebsiteInternshipWork at officeLocal areaRemote workNight shift
- .... Contractor will implement and maintain scalable and reliable infrastructure on Google Cloud Platform (“GCP”) for Snowflake data warehousing. Monitor, troubleshoot... ..., including vendor resources Willingness to work on-site at stated location in the job openingDepartment:...PlatformCloudWebsiteContract workFor contractorsWork experience placement
- ...a meaningful way. Tempus' proprietary platform connects an entire ecosystem of real-world... ...right patients, at the right time.The Site Reliability Engineering team works with all... ...and business units to provide dependable cloud infrastructure solutions, along with support...PlatformCloudWebsiteFull time
$130k - $170k
Senior Site Reliability Engineer About Us Founded in 2014,... ...industry’s first and only cloud‑based, fully‑... ...bridges engineering, platform, and security teams to... ...processes, and resilience strategies, shifting the... ...Core Values Form, execute, and communicate new...PlatformCloudWebsiteFull timeFlexible hoursShift work$194k - $267k
...with speed and urgency and execute with excellence.This is an opportunity... ...too, let's talk.The TeamThe Site Reliability team is dedicated to... ...tooling and CI/CD platforms that support Okta’s SRE ecosystem... ...Maintain a highly available cloud infrastructure edge for the...PlatformCloudWebsiteLocal areaWorldwideFlexible hours$118.3k - $219.8k
Are you excited to lead Site Reliability Engineering teams that keep mission-critical, 24/7 services running reliably and securely?Do you enjoy building automated cloud platforms, hardening security, and driving ongoing cost optimization through strong FinOps practices...PlatformCloudWebsiteFull timeLocal area$127k - $249k
...to support, maintain and grow the Atlas platform. As a senior SRE, you will be expected... ...Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure... ...to ops work”)Be familiar with a major cloud provider (AWS, Azure, or GCP) and...PlatformCloudWebsiteLocal areaRemote workWorldwideFlexible hours$80 - $90 per hour
...LaSalle Network is hiring for a Senior Site Reliability Engineer (Compute Platform) with a leading infrastructure and platform engineering firm known for... ..., imaging, and integrating Kubernetes into private clouds using Infrastructure-as-Code. This is an exciting project...PlatformCloudWebsiteHourly payContract workTemporary workRemote work- ...complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within... ...solutions. Through code and cloud infrastructure, you will configure, maintain... ...and scalability of your application or platform.Job ResponsibilitiesGuides and assists...PlatformCloudWebsite
- ...the fastest-growing fintech platforms in the U.S., supporting over... ...grade systems that prioritize reliability, security, and performance, and... ...the RoleAs a Senior/Staff Site Reliability Engineer, you will... ...-class SRE skill, along with cloud platform literacy, a strong dedication...PlatformCloudWebsiteWork at officeImmediate startRemote work
$117.63k - $176.44k
...company, provides comprehensive ad platforms for publishers, advertisers,... ...responsible for ensuring the reliability, scalability, and performance... ...Engineer.Experience with cloud platforms (e.g., AWS, GCP,... ...benefits summary on our careers site for more details....PlatformCloudWebsiteFull time$194k - $267k
...operate with speed and urgency and execute with excellence.This is an... ...StaffObservabilitySite Reliability Engineer with a specialty in... ...comprehensive, scalable Observability Platform that enables our SRE teams... ...scaling and managing Splunk Cloud at scale (1000+ SVCs),...PlatformCloudWebsitePermanent employmentWork at officeLocal areaWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Executive Director, Site Reliability and Cloud Platform Strategy. Be the first to apply!
- managing director of operations Chicago, IL
- chief Chicago, IL
- technology executive Chicago, IL
- chief development officer Chicago, IL
- chief executive officer Chicago, IL
- executive assistant to ceo Chicago, IL
- executive sales director Chicago, IL
- chief of psychiatry Chicago, IL
- chief officer Chicago, IL
- information technology executive Chicago, IL


