Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Manager Site Reliability Engineering

$125.83k - $221.28k

The Federal Home Loan Bank of Chicago

What you’ll do We are building a new Site Reliability Engineering function and seeking a leader who can establish SRE practices across the organization while developing a team of engineers new to the discipline. This is a unique opportunity to shape how reliability engineering is practiced at FHLBank Chicago from the ground up. The SRE team operates as a guiding and consultative partner to application and development teams rather than owning systems directly. Success in this role requires technical credibility, strong influencing skills, and the ability to drive change through collaboration and education rather than direct authority. This role is accountable for building and leading the SRE team, including hiring, performance management, coaching, and development of engineers transitioning into SRE practices. The manager establishes the SRE operating model (how SRE engages with application and development teams), ensures sustainable on-call and learning culture, and drives adoption of reliability standards through collaboration and influence. How you’ll make an impact Shape how Reliability Engineering is practiced and enforced across the Bank through collaboration Build deep relationships between IT and the greater organization in support of common goals Provide direction and development guidance to a team of 3-4 members in this new space Deepen further automation into application availability and reporting processes What you can expect Team Leadership and Development Build and develop a team of engineers transitioning from traditional operations and systems administration backgrounds into SRE practices Create psychological safety that enables learning, experimentation, and honest discussion of failures Establish career development paths and growth opportunities within the SRE discipline Foster a culture of blameless postmortems and continuous improvement Reliability Engineering Practice Define and implement Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budget policies across critical services Establish pager budgets and on‑call practices that are sustainable and effective Lead tuning and optimization of monitoring, alerting, and observability tooling Drive reduction of system disruptions through automation, tooling, and process improvement Develop and maintain incident management processes, including severity classification and escalation procedures Participate in FHLBank’s Disaster Recovery process and testing, coordinating and executing regularly scheduled DR exercises Consultation and Partnership Participate in troubleshooting meetings and production incidents, providing expert guidance and recommendations Partner with application owners, product owners, and development teams to improve system reliability Ensure deep technical analysis is performed for significant reliability issues, and provide escalation support as needed; coach the team in translating findings into actionable recommendations. Advocate for reliability investments and help teams prioritize reliability work against feature development Build relationships that enable SRE to influence architectural and operational decisions without direct ownership Organizational Leadership Secure and maintain executive sponsorship and governance mechanisms required for SLO and error budget practices (including defined decision rights when reliability thresholds are breached). Communicate the value and principles of SRE to leadership, helping secure sustained support and appropriate resource allocation Develop metrics and reporting that demonstrate SRE impact on business outcomes Navigate organizational dynamics to build credibility and trust for a new function Align SRE practices with existing compliance, risk management, and regulatory requirements What you’ll bring 2+ years of Site Reliability Engineering or directly-related primary function 5+ years of experience in infrastructure, operations, DevOps 2+ years of people management experience, with demonstrated ability to develop and grow technical talent Strong technical foundation in systems administration, networking, and infrastructure Proficiency in at least one programming or scripting language (.NET preferred, Python also valuable) Experience with monitoring, observability, and alerting tools and practices Proficiency with agentic AI tools such as Github Copilot, Claude Code, or Codex Demonstrated ability to influence outcomes without direct authority Strong written and verbal communication skills, including ability to explain technical concepts to non-technical stakeholders Experience conducting or leading incident response and postmortem processes Outstanding communication (verbal, written, and listening) skills Proven ability to consistently navigate crucial conversations Critical thinking - using logic and reasoning to identify the strengths and weaknesses of alternative solutions, conclusions or approaches to problems Systems thinking - approaching situations and scenarios understanding they are a complex web of interdependencies between items and other systems that are often initially unclear Ability to present ideas in business-friendly and user-friendly language Attention to detail Comfort with high levels of ambiguity and shared responsibility Pleasant demeanor with others with a good-natured, cooperative attitude Experience with Agile methods and concepts Knowledge of cloud computing principles, specifically related to Amazon Web Services Preferred Qualifications Experience implementing SRE practices in an organization new to the discipline Background in financial services or other regulated industries Experience defining and implementing SLIs, SLOs, and error budget policies Experience building or transforming teams through organizational change Knowledge of ITIL, DevOps, or related frameworks The Perks We offer a highly competitive compensation and bonus package, a retirement program with 401(k) and a pension plan, medical, dental and vision insurance, a Lifestyle Spending Account, a competitive PTO plan, 11 paid holidays per year and remote work flexibility. Salary Range $125,825.00 - $221,275.00 #J-18808-Ljbffr

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Manager Site Reliability Engineering in Chicago, IL vacancy
  •  ...Senior Site Reliability Engineer We are looking for a Senior Reliability Engineer to join our Platform team. In this position, you will be...  ...engineering organization to build automated processes and tools for managing application and service deployments Own and support... 
    Suggested
    Temporary work
    Flexible hours

    1872 Consulting

    Chicago, IL
    2 days ago
  •  ...Site Reliability Engineer (SRE) Immediate need for a talented Site Reliability Engineer (SRE). This is a 12+ months contract opportunity with...  ...issue/resolution status (written and verbal) to project team and management ~ Provide reactive, break-fix support... 
    Suggested
    Contract work
    Immediate start

    Pyramid Consulting

    Chicago, IL
    2 days ago
  •  ...with the investigation, Always validate if the team is following the SOPs or the process defined for an alerts/ issue. Contacting all the external vendors if in case their integrations fail, Measure the front-end metrics for the site with various tools available... 
    Suggested

    Omni Inclusive

    Chicago, IL
    1 day ago
  •  ...Edward Jones Site Reliability Engineer 100% remote Initial contract is 6 months, but will be a multi year engagement. Position Overview...  ...of our systems. You will be responsible for incident management, root cause analysis, and implementing postmortem processes... 
    Suggested
    Contract work
    Remote work

    HCL Global Systems

    Chicago, IL
    1 day ago
  •  ...Site Reliability Engineer in Wealth Management Chicago (IL) / Tempe (AZ) Onsite Job ROLE: This role will be Responsible for application observability, maintenance, and support, identifying and implementing preventive measures proactively, evaluates and... 
    Suggested
    Flexible hours

    Info Way Solutions

    Chicago, IL
    3 days ago
  • $125.83k - $221.28k

     ...Site Reliability Engineering Manager At the Federal Home Loan Bank of Chicago, employees come first - that's why we offer a highly competitive compensation and bonus package, and access to a comprehensive benefits program designed to meet the needs of our employees... 
    Work at office
    Remote work

    FHLBank Chicago

    Chicago, IL
    5 days ago
  • $112.5k - $187.5k

     ...We Collect Your Privacy Choices Team Overview At TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work... 
    Full time
    Temporary work
    Work experience placement
    Work at office
    Flexible hours
    2 days per week

    TransUnion

    Chicago, IL
    1 day ago
  • $160k - $200k

     ...stage of our journey as we continue to grow. The Opportunity We, at Flywire, are looking for an experienced Manager II, Site Reliability Engineering to join our team. In this role, you’ll help drive reliability, automation and performance within our cloud-based... 
    Full time
    Temporary work
    Local area
    Immediate start
    Remote work
    Shift work

    Flywire

    Chicago, IL
    17 days ago
  • $150k - $200k

     ...@ Selby Jennings | Financial Technology We are seeking a Site Reliability Engineer to join our team and assist with the design, development,...  .../written communication skills and documentation/knowledge management skills This role must sit in the firms Chicago office. Seniority... 
    Full time
    Work at office

    Selby Jennings

    Chicago, IL
    4 days ago
  •  ...Direct message the job poster from Algo Capital Group Senior Site Reliability Engineer - Observability and Automation A leading high-frequency...  ...pipelines and a variety of languages Familiarity with AWS managed services (EKS, S3, EC2, etc.) A service-focused attitude: asking... 
    Full time
    Work at office
    Flexible hours

    Algo Capital Group

    Chicago, IL
    4 days ago
  • $190.8k - $267.1k

     ...influential and trafficked corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team, you’ll use your...  ...more. In this role, you will also take ownership of risk management, ensuring the reliability and performance of our systems.... 
    Work experience placement
    Home office
    Flexible hours

    Alien Blue

    Chicago, IL
    2 days ago
  • $250k - $350k

     ...where quantitative researchers, engineers, traders, and operational...  ...boost stability, throughput, and reliability Qualifications Minimum of 3...  ...in production support, site reliability, or infrastructure...  ...and Bash Hands-on experience managing Kubernetes in a production setting... 
    Full time

    Engtal

    Chicago, IL
    2 days ago
  • $150k - $155k

    Site Reliability Engineer Hybrid (3 days onsite, 2 days remote) full‑time. No visa sponsorship. Base pay: $150,000 - $155,000 per year, subject...  ...large‑scale distributed systems Experience managing infrastructure in public cloud environments like AWS (preferred... 
    Full time
    Work experience placement
    Remote work
    Visa sponsorship

    Request Technology, LLC

    Chicago, IL
    4 days ago
  • $125.04k - $187.56k

     ...Digital and E-commerce, Technology and more. Overview The Site Reliability Engineer (SRE) III is responsible for ensuring the scalability, reliability...  ...(SLOs) and service level indicators (SLIs). Build and manage microservices-based platforms leveraging Spring Boot, Java,... 
    Full time
    Work at office
    Remote work
    Flexible hours

    ViziRecruiter,LLC.

    Chicago, IL
    a month ago
  • $130k - $170k

    Senior Site Reliability Engineer About Us Founded in 2014, we offer the industry’s first and only cloud‑based, fully‑customisable, end‑to‑end...  .... By combining thought leadership in suitability and risk management with industry‑leading education and the latest technology,... 
    Full time
    Flexible hours
    Shift work

    Supernova Technology™

    Chicago, IL
    4 days ago
  •  ...building and running systems that must perform reliably under real-time market conditions. The culture is highly collaborative, engineering-driven, and focused on continuous...  ...a related field 3+ years of experience in site reliability, systems engineering, or technical... 

    Fintal Partners

    Chicago, IL
    4 days ago
  • $128.5k - $214.1k

    We're looking for a Staff Site Reliability Engineer to join our team, focusing on the core systems that power global financial markets. This isn...  ...with automation, CI/CD, orchestration, and configuration management.* **Observability knowledge**: Familiarity with logging... 
    Work at office
    Worldwide
    2 days per week

    CME Group Inc.

    Chicago, IL
    4 days ago
  • $124k - $280k

     ...Data And Analytics Engineering Director At PwC, our people in data and analytics engineering focus on leveraging advanced technologies...  ...data storage solutions using cloud services Designing and managing data warehouses and data lakes Implementing IAM roles and policies... 

    PwC (US)

    Chicago, IL
    2 days ago
  • $73.5k - $212.28k

     ...Requirements: Up to 60% At PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to...  ...for coaching, leveraging team member's unique strengths, and managing performance to deliver on client expectations. With your... 
    Full time
    H1b

    PwC

    Chicago, IL
    2 days ago
  •  ...Instagram and YouTube . Job Description Opportunity at a Glance The Registrar Functional Tech position is responsible for managing and building Registrar and Degree Audit related configuration in Banner, and DegreeWorks. This involves coordinating with Business... 
    Permanent employment
    Full time
    Work at office
    Remote work
    Flexible hours

    Covista

    La Grange, IL
    5 days ago
  •  ...Overview We are seeking an experienced Observability / Site Reliability Engineer (SRE) to design, scale, and maintain our enterprise monitoring...  ...-native tools. Key Responsibilities GCP & Cloud Management: Architect, optimize, and maintain observability frameworks... 
    Remote job
    Contract work

    Ontrac Solutions

    Chicago, IL
    23 days ago
  • $110k - $120k

     ...the World’s leading AI-powered Quality Engineering Company? Ready to advance your career,...  ...us at QualityAI! We are looking for a Site Reliability Engineer (SRE)) to join our growing...  ...JIRA * Basic understanding of Release Management * Understanding of Agile methodologies... 
    Full time
    Local area
    2 days per week
    3 days per week

    Qualitest Group

    Chicago, IL
    7 days ago
  • $165k - $225k

     ...demanding AI workloads with enterprise-grade reliability and compliance. Your Role: You will...  ...core. Working closely with our systems engineers, network engineers, and platform...  ...from the ground up (not just deploying in managed environments). You'll ensure enterprise-... 
    Remote work
    Flexible hours

    Moonlite

    Chicago, IL
    3 days ago
  • $145k - $175k

     ...help you gain your full potential. Job Overview The Site Reliability Engineer supports deployments, cloud infrastructure, and monitoring...  ...improve AWS infrastructure using Terraform and Atlantis. • Manage secrets and security operations with HashiCorp Vault. • Participate... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Rewards Network

    Chicago, IL
    15 days ago
  •  ...Qualifications: 8+ years of Software Engineering experience, or equivalent...  ...and maintain scalable and reliable infrastructure on Google...  ...effectively with the client, IT management and staff, and other groups...  ...resources Willingness to work on-site at stated location in the job... 
    Contract work
    For contractors
    Work experience placement

    CEDENT

    Chicago, IL
    more than 2 months ago
  • $130k - $180k

     ...belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means…  You are an engineer, a builder, and a...  ...rotations. You’ll be a key voice in observability, change management, and service scalability, providing guidance during... 
    Work at office
    Local area
    Remote work
    Worldwide
    Monday to Friday
    Flexible hours

    iManage

    Chicago, IL
    17 days ago
  • $130k - $165k

     ...Job Title: Senior Software Engineer Company: Snapsheet Job Location: USA, Remote...  ...Department: Technology  Team : Site Reliability Engineering   About Snapsheet:...  ...virtual estimating and innovative claims management technology, transforming the end-to-... 
    Full time
    Temporary work
    Local area
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    Snapsheet

    Chicago, IL
    24 days ago
  • $78.75k - $131.25k

     ...Information We Collect Your Privacy Choices Team Overview This role is for a Senior Developer Platform Engineering who will report into a Senior Manager for DevOps and operate as a member on one of the core teams responsible for enabling new features and... 
    Full time
    Temporary work
    Work experience placement
    Work at office
    Flexible hours
    2 days per week

    Transunion

    Chicago, IL
    6 days ago
  • $194k - $267k

     ...talk. We are seeking a highly technical Observability Site Reliability Engineer with a specialty in Google Cloud, to own and expand our Observability...  ...(The Essentials) GKE: Minimum 5+ Experience scaling and managing observability in a Google Cloud platform. Visualization:... 
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    Chicago, IL
    a month ago
  • $194k - $267k

     ...We are seeking a highly technical Staff Observability Site Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem...  .... Required Skills & Experience (The Essentials) Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Chicago, IL
    15 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Manager Site Reliability Engineering. Be the first to apply!