Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

$125.04k - $187.56k

ViziRecruiter

Introduction Ahold Delhaize USA, a division of global food retailer Ahold Delhaize, is part of the U.S. family of brands, which also includes five leading omnichannel grocery brands – Food Lion, Giant Food, The GIANT Company, Hannaford and Stop & Shop. Ahold Delhaize USA associates support the brands with a wide range of services, including Finance, Legal, Sustainability, Commercial, Digital and E-commerce, Technology and more. Overview The Site Reliability Engineer (SRE) III is responsible for ensuring the scalability, reliability, and performance of production systems through automation, observability, incident response, and infrastructure engineering. This role involves designing and implementing robust operational processes and tooling to support highly available, fault-tolerant systems in a cloud-native environment. The SRE III collaborates closely with engineering squads, product teams, and stakeholders to embed reliability best practices across the software delivery lifecycle. The role includes ownership of system uptime, service level objectives (SLOs), and operational excellence, along with mentoring junior engineers and leading cross-functional initiatives that improve system resilience. Applicants must be currently authorized to work in the United States on a full-time basis. Our flexible/hybrid work schedule includes 3 in-person days at our Chicago office and 2 remote days. Responsibilities Design and implement infrastructure solutions that ensure system availability, scalability, and reliability across cloud-native environments like AKS and Kubernetes. Develop automation for provisioning, deployment, configuration, monitoring, and incident remediation using tools such as Terraform, ArgoCD, and GitHub Actions. Collaborate with engineering teams to define and track service level objectives (SLOs) and service level indicators (SLIs). Build and manage microservices-based platforms leveraging Spring Boot, Java, Tomcat, and Redis. Monitor production environments using Datadog and proactively address performance and reliability issues. Perform root cause analysis and lead post-incident reviews to drive continual improvement. Manage CI/CD pipelines and deployment automation using GitHub, Docker, and container orchestration technologies. Create and maintain infrastructure as code (IaC) using Terraform, with deployment pipelines integrated into GitOps workflows. Lead and support operational readiness reviews, game days, chaos engineering practices, and failure mode analysis. Build scalable observability and alerting frameworks with Datadog. Implement resilient, asynchronous architectures using Kafka for event-driven services. Reduce operational toil through self-healing automation and proactive system tuning. Troubleshoot Linux-based environments such as Ubuntu and optimize them for reliability. Provide on-call support and ensure 24/7/365 system reliability for mission-critical applications. Collaborate with the security team to enforce secure operational practices and cloud compliance. Mentor junior engineers and contribute to documentation, technical design, and knowledge-sharing across the organization. Requirements Bachelor's Degree in Computer Science, Information Systems, or a related technical field; equivalent training, certifications, or experience will be considered. 5+ years of experience in a Site Reliability Engineering, or DevOps, or Java programming role. Experience managing production-grade systems and services on AKS/Kubernetes in distributed environments. Proficiency in programming and scripting languages including Python, Java, Bash, or Go. Proven experience with Spring Boot, Tomcat, Redis, and microservices architecture. Hands‑on experience in managing Linux environments, particularly Ubuntu. Proficiency with observability stacks and performance monitoring using Datadog, Prometheus, and ELK. Deep understanding of containerization and orchestration using Docker, Kubernetes, and ArgoCD. Experience managing event‑driven systems using Kafka. Expertise in IaC and automation using Terraform and GitHub Actions. Familiarity with networking concepts, DNS, load balancing, and cloud infrastructure (AWS, Azure, or GCP). Strong analytical, debugging, and problem‑solving skills. Excellent verbal and written communication skills and the ability to collaborate effectively across teams. Salary Range: $125,040 - $187,560 Actual compensation offered to a candidate may vary based on their unique qualifications and experience, internal equity, and market conditions. Final compensation decisions will be made in accordance with company policies and applicable laws. #J-18808-Ljbffr

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Chicago, IL vacancy
  • $158.5k - $172k

     ...the exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate,...  ...environment. This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire... 
    Senior
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    Chicago, IL
    3 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Senior
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Chicago, IL
    19 hours ago
  • $145k - $175k

     ...designed to help you gain your full potential. Job OverviewThe Site Reliability Engineer supports deployments, cloud infrastructure, and monitoring...  ...infrastructure improvements. You'll be joining a small, senior SRE team with broad ownership of the platforms and infrastructure... 
    Senior
    Full time
    Work at office
    Local area
    Flexible hours
    3 days per week

    Rewards Network

    Chicago, IL
    3 days ago
  • $130k - $180k

     ...best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform... 
    Senior
    Work at office
    Local area
    Remote work
    Worldwide
    Monday to Friday
    Flexible hours

    Imanage

    Chicago, IL
    19 hours ago
  • $127k - $249k

     ...Eastern or Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the...  ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background.... 
    Senior
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Chicago, IL
    1 day ago
  • $190.8k - $267.1k

     ...is a unique opportunity to leave your mark on one of the most influential and trafficked corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team, you’ll use your knowledge of distributed systems and architecture to improve the... 
    Senior
    Work experience placement
    Home office
    Flexible hours

    Alien Blue

    Chicago, IL
    1 day ago
  •  ...Get AI-powered advice on this job and more exclusive features. Direct message the job poster from Algo Capital Group Senior Site Reliability Engineer - Observability and Automation A leading high-frequency trading firm is seeking a mid to senior-level Site Reliability... 
    Senior
    Full time
    Work at office
    Flexible hours

    Algo Capital Group

    Chicago, IL
    1 day ago
  • $130k - $165k

     ...Job Title: Senior Software Engineer Company: Snapsheet Job Location: USA, Remote Job Type: Full-time, direct hire Job Department: Technology  Team : Site Reliability Engineering   About Snapsheet: Snapsheet exists to simplify claims. We leverage... 
    Senior
    Full time
    Temporary work
    Local area
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    Snapsheet

    Chicago, IL
    more than 2 months ago
  • $180k - $200k

     ...Company Name: tastytrade Role: Senior Site Reliability Engineer Location: Chicago, IL (Hybrid, 3 days/week in office) Role Summary Come join tastytrade, part of IG Group, as we build the reliability practice behind the brokerage platform that active options... 
    Senior
    Work at office
    3 days per week

    tastytrade

    Chicago, IL
    4 days ago
  • $80 - $90 per hour

     ...LaSalle Network is hiring for a Senior Site Reliability Engineer (Compute Platform) with a leading infrastructure and platform engineering firm known for innovation and cutting-edge technological solutions. This opportunity is a remote role focused on deep infrastructure... 
    Senior
    Hourly pay
    Contract work
    Temporary work
    Remote work

    LaSalle Network

    Chicago, IL
    11 days ago
  • $80 - $90 per hour

     ...LaSalle Network is hiring for a Senior Site Reliability Engineer (Storage Platforms) with a storage-focused, enterprise-leading company known for innovation and impactful cloud infrastructure. Join a dedicated team managing software-defined storage solutions and enterprise... 
    Senior
    Hourly pay
    Contract work
    Temporary work
    Remote work

    LaSalle Network

    Chicago, IL
    11 days ago
  • $130k - $170k

    Senior Site Reliability Engineer About Us Founded in 2014, we offer the industry’s first and only cloud‑based, fully‑customisable, end‑to‑end software solution to automate securities‑based lending from origination through the life of the loan. By combining thought leadership... 
    Senior
    Full time
    Flexible hours
    Shift work

    Supernova Technology™

    Chicago, IL
    3 days ago
  •  ...Hire Overview Our client is seeking a highly skilled Edge Site Reliability Engineer (Edge SRE) to lead the design, automation, and operations...  ..., leadership, and cross‑functional collaboration skills. Seniority level Not Applicable Employment type Full-time Job function... 
    Senior
    Full time
    Contract work

    CoSourcing Partners - Enterprise-AI and IT Services Company

    Chicago, IL
    1 day ago
  •  ...consumers and companies, alikeKlover’s engineering team powers one of the fastest-growing...  ...-grade systems that prioritize reliability, security, and performance, and that integrate...  ...the right candidateAbout the RoleAs a Senior/Staff Site Reliability Engineer, you will play a... 
    Senior
    Work at office
    Immediate start
    Remote work

    Attain Data

    Chicago, IL
    3 days ago
  • $140k - $170k

    We are looking for a Senior Site Reliability Engineer to work as part of a lean, product‑focused engineering organization. This role is about building and operating reliable cloud‑based systems by writing code, automating infrastructure and delivery workflows, and reducing... 
    Senior
    Full time
    Work experience placement
    Flexible hours

    SEI Investments Developments

    Chicago, IL
    4 days ago
  • $165k - $225k

     ...enterprises to deploy demanding AI workloads with enterprise-grade reliability and compliance. Your Role: You will be instrumental in...  ...expertise at its core. Working closely with our systems engineers, network engineers, and platform engineering team, you'll architect... 
    Senior
    Remote work
    Flexible hours

    Moonlite

    Chicago, IL
    more than 2 months ago
  • $117.63k - $176.44k

     ...you will be responsible for ensuring the reliability, scalability, and performance of our data systems. Working closely with data engineers and other operation sub-teams, you will manage...  ...and benefits summary on our careers site for more details.EducationBachelor's DegreeWhile... 
    Senior
    Full time

    Comcast

    Chicago, IL
    3 days ago
  • $138.1k - $198.2k

     ...more intuitive with technology that simply works.  The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments...  ...Our customers are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering... 
    Permanent employment
    Full time
    Temporary work
    Work experience placement
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    Chicago, IL
    2 days ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will solve complex and... 

    JP Morgan Chase

    Chicago, IL
    4 days ago
  • Play a key role in ensuring system reliability at one of the world’s most iconic and largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will use technology to solve... 

    JP Morgan Chase

    Chicago, IL
    4 days ago
  • $100k - $120k

    OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying... 
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Origami Risk

    Chicago, IL
    2 days ago
  • $130k - $225k

     ...expectations, integrity, innovation and a willingness to challenge consensus.The Algorithmic Trading Team is looking for a Site Reliability Engineer for our Chicago office. The SRE team is critical to the success of our trading - ensuring that our production trading... 
    Temporary work
    Work at office
    Flexible hours

    DRW

    Chicago, IL
    3 days ago
  • Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through...  ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform...  ...vendor resources Willingness to work on-site at stated location in the job openingDepartment... 
    Contract work
    For contractors
    Work experience placement

    Cedent Consulting

    Chicago, IL
    3 days ago
  •  ...to physicians, providing critical information about the right treatments for the right patients, at the right time.The Site Reliability Engineering team works with all departments and business units to provide dependable cloud infrastructure solutions, along with support... 
    Full time

    Tempus

    Chicago, IL
    19 hours ago
  • $130k - $150k

     ...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are...  ...career mentoring and performance coaching from an assigned senior colleague. Additional leadership and collaboration opportunities... 
    Work at office
    Work from home
    3 days per week

    CRA International

    Chicago, IL
    2 days ago
  • $100.7k - $167.8k

    Job SummaryThe Site Reliability Engineer III is a pivotal architect of stability for CME Clearing & Risk. You will engineer secure, scalable, and reliable technology solutions that safeguard the global marketplace. By bridging the gap between development and operations,... 
    Full time
    Worldwide

    CME- Group

    Chicago, IL
    3 days ago
  • $194k - $267k

     ...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Chicago, IL
    3 days ago
  • $194k - $267k

     ...do something more than once, automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Chicago, IL
    19 hours ago
  • $132.1k - $220.1k

    We're looking for a Staff Site Reliability Engineer to join our team, focusing on the core systems that power global financial markets. This isn't just about keeping the lights on; it's about pioneering the future of financial technology. As a member of our Clearing department... 
    Full time
    Work at office
    Worldwide
    2 days per week

    CME- Group

    Chicago, IL
    3 days ago
  • $112.5k - $187.5k

     ...TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering...  ...Site Reliability Engineer at TransUnion, you will serve as a senior technical leader and force multiplier on the SRE team.... 
    Full time
    Temporary work
    Work experience placement
    Work at office
    Flexible hours
    2 days per week

    TransUnion

    Chicago, IL
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!