Manager Site Reliability Engineering
$125.83k - $221.28kThe Federal Home Loan Bank of Chicago
What you’ll do We are building a new Site Reliability Engineering function and seeking a leader who can establish SRE practices across the organization while developing a team of engineers new to the discipline. This is a unique opportunity to shape how reliability engineering is practiced at FHLBank Chicago from the ground up. The SRE team operates as a guiding and consultative partner to application and development teams rather than owning systems directly. Success in this role requires technical credibility, strong influencing skills, and the ability to drive change through collaboration and education rather than direct authority. This role is accountable for building and leading the SRE team, including hiring, performance management, coaching, and development of engineers transitioning into SRE practices. The manager establishes the SRE operating model (how SRE engages with application and development teams), ensures sustainable on-call and learning culture, and drives adoption of reliability standards through collaboration and influence. How you’ll make an impact Shape how Reliability Engineering is practiced and enforced across the Bank through collaboration Build deep relationships between IT and the greater organization in support of common goals Provide direction and development guidance to a team of 3-4 members in this new space Deepen further automation into application availability and reporting processes What you can expect Team Leadership and Development Build and develop a team of engineers transitioning from traditional operations and systems administration backgrounds into SRE practices Create psychological safety that enables learning, experimentation, and honest discussion of failures Establish career development paths and growth opportunities within the SRE discipline Foster a culture of blameless postmortems and continuous improvement Reliability Engineering Practice Define and implement Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budget policies across critical services Establish pager budgets and on‑call practices that are sustainable and effective Lead tuning and optimization of monitoring, alerting, and observability tooling Drive reduction of system disruptions through automation, tooling, and process improvement Develop and maintain incident management processes, including severity classification and escalation procedures Participate in FHLBank’s Disaster Recovery process and testing, coordinating and executing regularly scheduled DR exercises Consultation and Partnership Participate in troubleshooting meetings and production incidents, providing expert guidance and recommendations Partner with application owners, product owners, and development teams to improve system reliability Ensure deep technical analysis is performed for significant reliability issues, and provide escalation support as needed; coach the team in translating findings into actionable recommendations. Advocate for reliability investments and help teams prioritize reliability work against feature development Build relationships that enable SRE to influence architectural and operational decisions without direct ownership Organizational Leadership Secure and maintain executive sponsorship and governance mechanisms required for SLO and error budget practices (including defined decision rights when reliability thresholds are breached). Communicate the value and principles of SRE to leadership, helping secure sustained support and appropriate resource allocation Develop metrics and reporting that demonstrate SRE impact on business outcomes Navigate organizational dynamics to build credibility and trust for a new function Align SRE practices with existing compliance, risk management, and regulatory requirements What you’ll bring 2+ years of Site Reliability Engineering or directly-related primary function 5+ years of experience in infrastructure, operations, DevOps 2+ years of people management experience, with demonstrated ability to develop and grow technical talent Strong technical foundation in systems administration, networking, and infrastructure Proficiency in at least one programming or scripting language (.NET preferred, Python also valuable) Experience with monitoring, observability, and alerting tools and practices Proficiency with agentic AI tools such as Github Copilot, Claude Code, or Codex Demonstrated ability to influence outcomes without direct authority Strong written and verbal communication skills, including ability to explain technical concepts to non-technical stakeholders Experience conducting or leading incident response and postmortem processes Outstanding communication (verbal, written, and listening) skills Proven ability to consistently navigate crucial conversations Critical thinking - using logic and reasoning to identify the strengths and weaknesses of alternative solutions, conclusions or approaches to problems Systems thinking - approaching situations and scenarios understanding they are a complex web of interdependencies between items and other systems that are often initially unclear Ability to present ideas in business-friendly and user-friendly language Attention to detail Comfort with high levels of ambiguity and shared responsibility Pleasant demeanor with others with a good-natured, cooperative attitude Experience with Agile methods and concepts Knowledge of cloud computing principles, specifically related to Amazon Web Services Preferred Qualifications Experience implementing SRE practices in an organization new to the discipline Background in financial services or other regulated industries Experience defining and implementing SLIs, SLOs, and error budget policies Experience building or transforming teams through organizational change Knowledge of ITIL, DevOps, or related frameworks The Perks We offer a highly competitive compensation and bonus package, a retirement program with 401(k) and a pension plan, medical, dental and vision insurance, a Lifestyle Spending Account, a competitive PTO plan, 11 paid holidays per year and remote work flexibility. Salary Range $125,825.00 - $221,275.00 #J-18808-Ljbffr
- ...Senior Site Reliability Engineer We are looking for a Senior Reliability Engineer to join our Platform team. In this position, you will be... ...engineering organization to build automated processes and tools for managing application and service deployments Own and support...SuggestedTemporary workFlexible hours
- ...Site Reliability Engineer (SRE) Immediate need for a talented Site Reliability Engineer (SRE). This is a 12+ months contract opportunity with... ...issue/resolution status (written and verbal) to project team and management ~ Provide reactive, break-fix support...SuggestedContract workImmediate start
- ...with the investigation, Always validate if the team is following the SOPs or the process defined for an alerts/ issue. Contacting all the external vendors if in case their integrations fail, Measure the front-end metrics for the site with various tools available...Suggested
- ...Edward Jones Site Reliability Engineer 100% remote Initial contract is 6 months, but will be a multi year engagement. Position Overview... ...of our systems. You will be responsible for incident management, root cause analysis, and implementing postmortem processes...SuggestedContract workRemote work
- ...Site Reliability Engineer in Wealth Management Chicago (IL) / Tempe (AZ) Onsite Job ROLE: This role will be Responsible for application observability, maintenance, and support, identifying and implementing preventive measures proactively, evaluates and...SuggestedFlexible hours
$125.83k - $221.28k
...Site Reliability Engineering Manager At the Federal Home Loan Bank of Chicago, employees come first - that's why we offer a highly competitive compensation and bonus package, and access to a comprehensive benefits program designed to meet the needs of our employees...Work at officeRemote work$112.5k - $187.5k
...We Collect Your Privacy Choices Team Overview At TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work...Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$160k - $200k
...stage of our journey as we continue to grow. The Opportunity We, at Flywire, are looking for an experienced Manager II, Site Reliability Engineering to join our team. In this role, you’ll help drive reliability, automation and performance within our cloud-based...Full timeTemporary workLocal areaImmediate startRemote workShift work$150k - $200k
...@ Selby Jennings | Financial Technology We are seeking a Site Reliability Engineer to join our team and assist with the design, development,... .../written communication skills and documentation/knowledge management skills This role must sit in the firms Chicago office. Seniority...Full timeWork at office- ...Direct message the job poster from Algo Capital Group Senior Site Reliability Engineer - Observability and Automation A leading high-frequency... ...pipelines and a variety of languages Familiarity with AWS managed services (EKS, S3, EC2, etc.) A service-focused attitude: asking...Full timeWork at officeFlexible hours
$190.8k - $267.1k
...influential and trafficked corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team, you’ll use your... ...more. In this role, you will also take ownership of risk management, ensuring the reliability and performance of our systems....Work experience placementHome officeFlexible hours$250k - $350k
...where quantitative researchers, engineers, traders, and operational... ...boost stability, throughput, and reliability Qualifications Minimum of 3... ...in production support, site reliability, or infrastructure... ...and Bash Hands-on experience managing Kubernetes in a production setting...Full time$150k - $155k
Site Reliability Engineer Hybrid (3 days onsite, 2 days remote) full‑time. No visa sponsorship. Base pay: $150,000 - $155,000 per year, subject... ...large‑scale distributed systems Experience managing infrastructure in public cloud environments like AWS (preferred...Full timeWork experience placementRemote workVisa sponsorship$125.04k - $187.56k
...Digital and E-commerce, Technology and more. Overview The Site Reliability Engineer (SRE) III is responsible for ensuring the scalability, reliability... ...(SLOs) and service level indicators (SLIs). Build and manage microservices-based platforms leveraging Spring Boot, Java,...Full timeWork at officeRemote workFlexible hours$130k - $170k
Senior Site Reliability Engineer About Us Founded in 2014, we offer the industry’s first and only cloud‑based, fully‑customisable, end‑to‑end... .... By combining thought leadership in suitability and risk management with industry‑leading education and the latest technology,...Full timeFlexible hoursShift work- ...building and running systems that must perform reliably under real-time market conditions. The culture is highly collaborative, engineering-driven, and focused on continuous... ...a related field 3+ years of experience in site reliability, systems engineering, or technical...
$128.5k - $214.1k
We're looking for a Staff Site Reliability Engineer to join our team, focusing on the core systems that power global financial markets. This isn... ...with automation, CI/CD, orchestration, and configuration management.* **Observability knowledge**: Familiarity with logging...Work at officeWorldwide2 days per week$124k - $280k
...Data And Analytics Engineering Director At PwC, our people in data and analytics engineering focus on leveraging advanced technologies... ...data storage solutions using cloud services Designing and managing data warehouses and data lakes Implementing IAM roles and policies...$73.5k - $212.28k
...Requirements: Up to 60% At PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to... ...for coaching, leveraging team member's unique strengths, and managing performance to deliver on client expectations. With your...Full timeH1b- ...Instagram and YouTube . Job Description Opportunity at a Glance The Registrar Functional Tech position is responsible for managing and building Registrar and Degree Audit related configuration in Banner, and DegreeWorks. This involves coordinating with Business...Permanent employmentFull timeWork at officeRemote workFlexible hours
- ...Overview We are seeking an experienced Observability / Site Reliability Engineer (SRE) to design, scale, and maintain our enterprise monitoring... ...-native tools. Key Responsibilities GCP & Cloud Management: Architect, optimize, and maintain observability frameworks...Remote jobContract work
$110k - $120k
...the World’s leading AI-powered Quality Engineering Company? Ready to advance your career,... ...us at QualityAI! We are looking for a Site Reliability Engineer (SRE)) to join our growing... ...JIRA * Basic understanding of Release Management * Understanding of Agile methodologies...Full timeLocal area2 days per week3 days per week$165k - $225k
...demanding AI workloads with enterprise-grade reliability and compliance. Your Role: You will... ...core. Working closely with our systems engineers, network engineers, and platform... ...from the ground up (not just deploying in managed environments). You'll ensure enterprise-...Remote workFlexible hours$145k - $175k
...help you gain your full potential. Job Overview The Site Reliability Engineer supports deployments, cloud infrastructure, and monitoring... ...improve AWS infrastructure using Terraform and Atlantis. • Manage secrets and security operations with HashiCorp Vault. • Participate...Full timeTemporary workWork at officeLocal areaFlexible hours3 days per week- ...Qualifications: 8+ years of Software Engineering experience, or equivalent... ...and maintain scalable and reliable infrastructure on Google... ...effectively with the client, IT management and staff, and other groups... ...resources Willingness to work on-site at stated location in the job...Contract workFor contractorsWork experience placement
$130k - $180k
...belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a... ...rotations. You’ll be a key voice in observability, change management, and service scalability, providing guidance during...Work at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours$130k - $165k
...Job Title: Senior Software Engineer Company: Snapsheet Job Location: USA, Remote... ...Department: Technology Team : Site Reliability Engineering About Snapsheet:... ...virtual estimating and innovative claims management technology, transforming the end-to-...Full timeTemporary workLocal areaRemote workVisa sponsorshipWork visaFlexible hours$78.75k - $131.25k
...Information We Collect Your Privacy Choices Team Overview This role is for a Senior Developer Platform Engineering who will report into a Senior Manager for DevOps and operate as a member on one of the core teams responsible for enabling new features and...Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$194k - $267k
...talk. We are seeking a highly technical Observability Site Reliability Engineer with a specialty in Google Cloud, to own and expand our Observability... ...(The Essentials) GKE: Minimum 5+ Experience scaling and managing observability in a Google Cloud platform. Visualization:...Permanent employmentLocal areaWorldwideFlexible hours$194k - $267k
...We are seeking a highly technical Staff Observability Site Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem... .... Required Skills & Experience (The Essentials) Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud...Permanent employmentWork at officeLocal areaWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager Site Reliability Engineering. Be the first to apply!
- site reliability engineer sre Chicago, IL
- site reliability engineer Chicago, IL
- site safety Chicago, IL
- historic site Chicago, IL
- IT site lead Chicago, IL
- site leader Chicago, IL
- site recruiter Chicago, IL
- website coordinator Chicago, IL
- website content developer Chicago, IL
- site services specialist Chicago, IL



