Senior Site Reliability Engineer
$125.04k - $187.56kViziRecruiter
Introduction Ahold Delhaize USA, a division of global food retailer Ahold Delhaize, is part of the U.S. family of brands, which also includes five leading omnichannel grocery brands – Food Lion, Giant Food, The GIANT Company, Hannaford and Stop & Shop. Ahold Delhaize USA associates support the brands with a wide range of services, including Finance, Legal, Sustainability, Commercial, Digital and E-commerce, Technology and more. Overview The Site Reliability Engineer (SRE) III is responsible for ensuring the scalability, reliability, and performance of production systems through automation, observability, incident response, and infrastructure engineering. This role involves designing and implementing robust operational processes and tooling to support highly available, fault-tolerant systems in a cloud-native environment. The SRE III collaborates closely with engineering squads, product teams, and stakeholders to embed reliability best practices across the software delivery lifecycle. The role includes ownership of system uptime, service level objectives (SLOs), and operational excellence, along with mentoring junior engineers and leading cross-functional initiatives that improve system resilience. Applicants must be currently authorized to work in the United States on a full-time basis. Our flexible/hybrid work schedule includes 3 in-person days at our Chicago office and 2 remote days. Responsibilities Design and implement infrastructure solutions that ensure system availability, scalability, and reliability across cloud-native environments like AKS and Kubernetes. Develop automation for provisioning, deployment, configuration, monitoring, and incident remediation using tools such as Terraform, ArgoCD, and GitHub Actions. Collaborate with engineering teams to define and track service level objectives (SLOs) and service level indicators (SLIs). Build and manage microservices-based platforms leveraging Spring Boot, Java, Tomcat, and Redis. Monitor production environments using Datadog and proactively address performance and reliability issues. Perform root cause analysis and lead post-incident reviews to drive continual improvement. Manage CI/CD pipelines and deployment automation using GitHub, Docker, and container orchestration technologies. Create and maintain infrastructure as code (IaC) using Terraform, with deployment pipelines integrated into GitOps workflows. Lead and support operational readiness reviews, game days, chaos engineering practices, and failure mode analysis. Build scalable observability and alerting frameworks with Datadog. Implement resilient, asynchronous architectures using Kafka for event-driven services. Reduce operational toil through self-healing automation and proactive system tuning. Troubleshoot Linux-based environments such as Ubuntu and optimize them for reliability. Provide on-call support and ensure 24/7/365 system reliability for mission-critical applications. Collaborate with the security team to enforce secure operational practices and cloud compliance. Mentor junior engineers and contribute to documentation, technical design, and knowledge-sharing across the organization. Requirements Bachelor's Degree in Computer Science, Information Systems, or a related technical field; equivalent training, certifications, or experience will be considered. 5+ years of experience in a Site Reliability Engineering, or DevOps, or Java programming role. Experience managing production-grade systems and services on AKS/Kubernetes in distributed environments. Proficiency in programming and scripting languages including Python, Java, Bash, or Go. Proven experience with Spring Boot, Tomcat, Redis, and microservices architecture. Hands‑on experience in managing Linux environments, particularly Ubuntu. Proficiency with observability stacks and performance monitoring using Datadog, Prometheus, and ELK. Deep understanding of containerization and orchestration using Docker, Kubernetes, and ArgoCD. Experience managing event‑driven systems using Kafka. Expertise in IaC and automation using Terraform and GitHub Actions. Familiarity with networking concepts, DNS, load balancing, and cloud infrastructure (AWS, Azure, or GCP). Strong analytical, debugging, and problem‑solving skills. Excellent verbal and written communication skills and the ability to collaborate effectively across teams. Salary Range: $125,040 - $187,560 Actual compensation offered to a candidate may vary based on their unique qualifications and experience, internal equity, and market conditions. Final compensation decisions will be made in accordance with company policies and applicable laws. #J-18808-Ljbffr
- ...Overview: Senior Site Reliability Engineer (SRE) Location: Chicago, IL (Onsite) Type: Contract Role Overview: We are seeking a Senior Site Reliability Engineer (SRE) with strong expertise in AWS infrastructure, automation, observability, and production...SeniorContract work
- ...Get AI-powered advice on this job and more exclusive features. Direct message the job poster from Algo Capital Group Senior Site Reliability Engineer - Observability and Automation A leading high-frequency trading firm is seeking a mid to senior-level Site Reliability...SeniorFull timeWork at officeFlexible hours
$130k - $170k
...Senior Site Reliability Engineer About Us Founded in 2014, we offer the industry’s first and only cloud‑based, fully‑customisable, end‑to‑end software solution to automate securities‑based lending from origination through the life of the loan. By combining thought leadership...SeniorFull timeFlexible hoursShift work$190.8k - $267.1k
...is a unique opportunity to leave your mark on one of the most influential and trafficked corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team, you’ll use your knowledge of distributed systems and architecture to improve the...SeniorWork experience placementHome officeFlexible hours$127k - $249k
THE TEAM Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational... ..., alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$130k - $165k
...Job Title: Senior Software Engineer Company: Snapsheet Job Location: USA, Remote Job Type: Full-time, direct hire Job Department: Technology Team : Site Reliability Engineering About Snapsheet: Snapsheet exists to simplify claims. We leverage our expertise...SeniorFull timeTemporary workLocal areaRemote workVisa sponsorshipWork visaFlexible hours$130k - $180k
...of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform guardrails...SeniorFull timeWork at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours$145k - $175k
...help you gain your full potential. Job Overview The Site Reliability Engineer supports deployments, cloud infrastructure, and monitoring... ...and infrastructure improvements. You'll be joining a small, senior SRE team with broad ownership of the platforms and infrastructure...SeniorFull timeTemporary workWork at officeLocal areaFlexible hours3 days per week$250k - $350k
...where quantitative researchers, engineers, traders, and operational... ...boost stability, throughput, and reliability Qualifications Minimum of 3... ...in production support, site reliability, or infrastructure... ...in team-driven environments Seniority level Mid-Senior level Employment...Full time$150k - $200k
...@ Selby Jennings | Financial Technology We are seeking a Site Reliability Engineer to join our team and assist with the design, development,... ...management skills This role must sit in the firms Chicago office. Seniority level Seniority level Entry level Employment type...Full timeWork at office- ...Site Reliability Engineer (SRE) Immediate need for a talented Site Reliability Engineer (SRE). This is a 12+ months contract opportunity with long-term potential and is in Chicago, IL (Hybrid). Key Requirements and Technology Experience: ~ Must have skills:...Contract workLocal areaImmediate start
- ...SRE Engineer We are seeking a highly capable engineer to join our dynamic SRE team. This... .... You'll work closely with peers and senior engineers to manage and evolve our infrastructure... ...transfer teams to ensure secure and reliable operations. .NET Logging &...
- ...Site Reliability Engineer As a Site Reliability Engineer, you will build and secure infrastructure supporting our AI platform with special attention to safeguarding US customer data and supporting the Aerospace and Defense Industrial Base. You'll have strong ownership...
$152.5k - $219.2k
...global cloud platform. As a team of six engineers distributed across the US, Canada, and the... ...with a strong focus on automation, reliability, and operational excellence. We are one... ...Qualifications ~2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure...Permanent employmentFull timeTemporary workLocal areaWorldwideFlexible hours- ...building and running systems that must perform reliably under real-time market conditions. The culture is highly collaborative, engineering-driven, and focused on continuous... ...a related field 3+ years of experience in site reliability, systems engineering, or technical...
$86k - $105k
...generation of application infrastructure and to be responsible for reliability, automation and scalability using and the latest best... ...certifications. Minimum of 2 years prior DevOps, software engineering or related experience. Must be able to work different schedules...Hourly payWork at officeImmediate startVisa sponsorshipWork visaFlexible hours$150k - $155k
...Site Reliability Engineer Hybrid (3 days onsite, 2 days remote) full‑time. No visa sponsorship. Base pay: $150,000 – $155,000 per year, subject... ..., including scripting, incident report summarization, or AI workload maintenance Seniority Level Mid‑Senior #J-18808-Ljbffr...Full timeWork experience placementRemote workVisa sponsorship- ...Qualifications: 8+ years of software engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources. Willingness to work on-site at stated location in the job opening....For contractorsWork experience placement
- ...Job Title: Site Reliability Engineer Location: Chicago, IL FTE Only Job Description Must Have Technical/Functional Skills ~ We are looking for a Senior Site Reliability Engineer (SRE) with deep experience in AWS infrastructure...
$85k - $130k
...Site Reliability Engineer Passionate about precision medicine and advancing the healthcare industry? Recent advancements in underlying technology have finally made it possible for AI to impact clinical care in a meaningful way. Tempus' proprietary platform connects...$128.5k - $214.1k
...We're looking for a Staff Site Reliability Engineer to join our team, focusing on the core systems that power global financial markets. This isn't just about keeping the lights on; it's about pioneering the future of financial technology. As a member of our Clearing department...Work at officeWorldwide2 days per week$102.6k - $193.43k
...Software Engineer Chamberlain Group (CG) is a global leader in intelligent access and Blackstone portfolio company. Powered by our... ...and manage discussions with TECH/IT Executives, Architects and Senior Engineers Collaborate with security team on architecture design...Temporary workSecond jobWorldwide$112.5k - $187.5k
...TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering... ...Site Reliability Engineer at TransUnion, you will serve as a senior technical leader and force multiplier on the SRE team....Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week- This is an exciting opportunity for someone that wants to be apart of a team with some of the brightest developers in the logistics industry. You will have an opportunity to be apart of an aggressively growing logistics provider that has developed LoadRunner Technologies...SeniorFull time
$200k - $275k
...P2P is seeking a skilled developer to build reliable and scalable systems. The role includes working on both legacy code and new development projects and requires experience in Java, Python, and understanding micro-service architectures. A competitive salary ranging from...Senior- ...We’re expanding our forward-deployed engineering team in Chicago, a major hub for our customers... ...through. Travel 20–30% to customer sites nationwide to unblock deployments, map... ...judgment, especially when balancing speed, reliability, and real-world constraints. ~ Based...SeniorFull timeImmediate start
- ...potential. We are seeking an experienced and highly motivated Senior Software Engineer to contribute to the continuous improvement of platform... ...tests and integration tests to ensure code quality and reliability. Continuously improving the performance, reliability, and...SeniorJob sharingFull time
- ...Homeland Security, Clean Technology, Energy, B2B, Manufacturing, Engineering, LifeSciences (Pharmaceutical, Scientific, Medical Device,... ...A predominant information technology company is seeking a Senior Software Engineer with strong Clojure or Ruby development skills...SeniorFull time
- ...Brooksource is seeking a Senior AI Solutions Engineer to lead the organization's AI and intelligent automation initiatives in Waukegan, IL. This highly visible, hybrid role is ideal for someone who thrives on solving business problems, working directly with stakeholders...SeniorFlexible hours
$160k
...Description Eagle Seven is seeking a Senior Software Developer to design and build... ...systems and has a strong commitment to reliability and sound software design. In this role,... ...traders, strategy developers, and fellow engineers to develop infrastructure that is central...SeniorFull timeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre Chicago, IL
- site reliability engineer Chicago, IL
- senior vice president communications Chicago, IL
- senior manager quality engineering Chicago, IL
- senior device engineer Chicago, IL
- sr operations manager Chicago, IL
- senior supervisor Chicago, IL
- senior client services manager Chicago, IL
- senior recruitment consultant Chicago, IL
- sr. hr generalist Chicago, IL


