#23183 - Site Reliability Engineer
$110k - $120kQualiTest Group
Qualitest Group Are you interested in working with the World’s leading AI-powered Quality Engineering Company? Ready to advance your career, team up with global thought leaders across industries and make a difference every day? Join us at QualityAI! We are looking for a Site Reliability Engineer (SRE)) to join our growing team in the United States! Location: Riverwoods, IL (Hybrid – 2 to 3 days/week onsite) Position Overview We are seeking an experienced Site Reliability Engineer (SRE) with a strong background in AWS Cloud, monitoring/observability platforms, and automation. The ideal candidate will partner closely with application development teams to improve application reliability, resiliency, performance, and operational excellence across hybrid cloud environments. This role is ideal for engineers with hands-on experience in monitoring engineering using tools such as Datadog, Dynatrace, Grafana, Kibana, and expertise in AWS-based applications. Must-Have Skills
- 6–12 years of professional experience as a Site Reliability Engineer (SRE)
- Strong hands-on experience with AWS Cloud applications and services (mandatory)
- Experience supporting hybrid environments (AWS Cloud and on-premises deployments)
- Strong Linux/Unix administration and shell scripting experience
- Experience with Systems Observability and Application Performance Monitoring (APM) tools, preferably Datadog (Dynatrace experience is also valuable)
- Experience building dashboards using Grafana and Kibana
- Experience in performance testing and the ability to translate functional and non-functional requirements into automated non-functional testing (NFT) solutions
- Strong understanding of application integration, high availability, resilience, and observability
- Experience with DevOps practices and CI/CD pipelines
- Strong programming skills in one or more of the following:
- Python
- Java
- Shell Scripting (Unix/Linux)
- Strong understanding of Software Development Life Cycle (SDLC)
- Hands-on experience with ServiceNow (SNOW)
- Experience with container technologies such as Kubernetes and OpenShift
- Experience with Jenkins and CI/CD automation
- Experience with automation tools such as Ansible
- Strong working knowledge of JIRA
- Basic understanding of Release Management
- Understanding of Agile methodologies
- Experience with AWS Lambda services
- Knowledge of SQL, MySQL, and database concepts
- Partner with application development teams to improve application resiliency and reliability.
- Implement and maintain Service Level Objectives (SLOs), Service Level Indicators (SLIs), and operational best practices.
- Build end-to-end observability solutions using monitoring, logging, and tracing tools.
- Design and implement monitoring, alerting, dashboards, and health checks for production applications.
- Automate operational processes to reduce manual effort and improve system reliability.
- Develop and enhance capacity planning and performance management capabilities.
- Support disaster recovery (DR) planning and implementation for critical applications.
- Develop and support chaos engineering and resilience testing initiatives.
- Participate in production incident management and on-call support rotation.
- Collaborate with cross-functional teams to improve system availability, scalability, and operational efficiency.
- Cloud: AWS (mandatory), Hybrid Cloud, On-Prem, Linux/Unix, AWS Lambda
- Monitoring & Observability: Datadog (preferred), Dynatrace, Grafana, Kibana, ELK, APM tools
- Programming: Python, Java, Shell Scripting, Go (preferred), Ansible
- DevOps: Jenkins, CI/CD, Kubernetes, OpenShift
- Tools: JIRA, ServiceNow (SNOW), Agile, Release Management
- Database: SQL, MySQL
- Be a part of a company who strives to support for diversity and inclusion in the workplace - we are one, we are many at QualityAI. Celebrate culture, share knowledge with engineers from around the globe, and inspire each other through our differences.
- Local and global opportunities - we offer you internal rotation and international mobility opportunities to grow your career.
- Clear view of your career and progression with the company - QualityAI is growing massively (since Jan 2021 - added more than 2000 engineers) and giving you the opportunity to grow with us.
- Never stop experimenting and learning with QualityAI Tech academy: 3000+ training courses, mentorship programs, technical tribes, sponsored certifications, leadership programs and much more.
- Earn bonuses via our Client Referral and Employee Referral Program’s. Refer and earn - tap your network for net-worth.
- A Competitive pay, the salary range for the role is $110,000 - $120,000.
- Intrigued to find more about us?
- Visit our website at
- 6–12 years of professional experience as a Site Reliability Engineer (SRE)
- Strong hands-on experience with AWS Cloud applications and services (mandatory)
- Experience supporting hybrid environments (AWS Cloud and on-premises deployments)
- Strong Linux/Unix administration and shell scripting experience
- Experience with Systems Observability and Application Performance Monitoring (APM) tools, preferably Datadog (Dynatrace experience is also valuable)
- Experience building dashboards using Grafana and Kibana
- Experience in performance testing and the ability to translate functional and non-functional requirements into automated non-functional testing (NFT) solutions
- Strong understanding of application integration, high availability, resilience, and observability
- Experience with DevOps practices and CI/CD pipelines
- Strong programming skills in one or more of the following:
- Python
- Java
- Shell Scripting (Unix/Linux)
- Strong understanding of Software Development Life Cycle (SDLC)
- Key Responsibilities
- Partner with application development teams to improve application resiliency and reliability.
- Implement and maintain Service Level Objectives (SLOs), Service Level Indicators (SLIs), and operational best practices.
- Build end-to-end observability solutions using monitoring, logging, and tracing tools.
- Design and implement monitoring, alerting, dashboards, and health checks for production applications.
- Automate operational processes to reduce manual effort and improve system reliability.
- Develop and enhance capacity planning and performance management capabilities.
- Support disaster recovery (DR) planning and implementation for critical applications.
- Develop and support chaos engineering and resilience testing initiatives.
- Participate in production incident management and on-call support rotation.
- Collaborate with cross-functional teams to improve system availability, scalability, and operational efficiency.
- 6–12 years of professional experience as a Site Reliability Engineer (SRE)
- Strong hands-on experience with AWS Cloud applications and services (mandatory)
- Experience supporting hybrid environments (AWS Cloud and on-premises deployments)
- Strong Linux/Unix administration and shell scripting experience
- Experience with Systems Observability and Application Performance Monitoring (APM) tools, preferably Datadog (Dynatrace experience is also valuable)
- Experience building dashboards using Grafana and Kibana
- Experience in performance testing and the ability to translate functional and non-functional requirements into automated non-functional testing (NFT) solutions
- Strong understanding of application integration, high availability, resilience, and observability
- Experience with DevOps practices and CI/CD pipelines
- Strong programming skills in one or more of the following:
- Python
- Java
- Shell Scripting (Unix/Linux)
- Strong understanding of Software Development Life Cycle (SDLC)
- Hands-on experience with ServiceNow (SNOW)
- Experience with container technologies such as Kubernetes and OpenShift
- Experience with Jenkins and CI/CD automation
- Experience with automation tools such as Ansible
- Strong working knowledge of JIRA
- Basic understanding of Release Management
- Understanding of Agile methodologies
- Experience with AWS Lambda services
- Knowledge of SQL, MySQL, and database concepts
- Partner with application development teams to improve application resiliency and reliability.
- Implement and maintain Service Level Objectives (SLOs), Service Level Indicators (SLIs), and operational best practices.
- Build end-to-end observability solutions using monitoring, logging, and tracing tools.
- Design and implement monitoring, alerting, dashboards, and health checks for production applications.
- Automate operational processes to reduce manual effort and improve system reliability.
- Develop and enhance capacity planning and performance management capabilities.
- Support disaster recovery (DR) planning and implementation for critical applications.
- Develop and support chaos engineering and resilience testing initiatives.
- Participate in production incident management and on-call support rotation.
- Collaborate with cross-functional teams to improve system availability, scalability, and operational efficiency.
- Cloud: AWS (mandatory), Hybrid Cloud, On-Prem, Linux/Unix, AWS Lambda
- Monitoring & Observability: Datadog (preferred), Dynatrace, Grafana, Kibana, ELK, APM tools
- Programming: Python, Java, Shell Scripting, Go (preferred), Ansible
- DevOps: Jenkins, CI/CD, Kubernetes, OpenShift
- Tools: JIRA, ServiceNow (SNOW), Agile, Release Management
- Database: SQL, MySQL
- Be a part of a company who strives to support for diversity and inclusion in the workplace - we are one, we are many at QualityAI. Celebrate culture, share knowledge with engineers from around the globe, and inspire each other through our differences.
- Local and global opportunities - we offer you internal rotation and international mobility opportunities to grow your career.
- Clear view of your career and progression with the company - QualityAI is growing massively (since Jan 2021 - added more than 2000 engineers) and giving you the opportunity to grow with us.
- Never stop experimenting and learning with QualityAI Tech academy: 3000+ training courses, mentorship programs, technical tribes, sponsored certifications, leadership programs and much more.
- Earn bonuses via our Client Referral and Employee Referral Program’s. Refer and earn - tap your network for net-worth.
- A Competitive pay, the salary range for the role is $110,000 - $120,000.
- Intrigued to find more about us?
- Visit our website at
$112.5k - $187.5k
...Information We Collect Your Privacy Choices Team Overview At TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential...SuggestedFull timeTemporary workWork experience placementWork at officeFlexible hours2 days per week- ...Site Reliability EngineerRemote - United StatesJR013952 We are seeking an experienced Site Reliability Engineer (SRE) with expertise in Infrastructure as Code tools like Terraform, core CI/CD tools such as Azure DevOps, and monitoring tools including DataDog and AWS...SuggestedFull timeTemporary workRemote workWork from homeFlexible hours
- ...Job Title: Site Reliability Engineer Location: Chicago, IL FTE Only Job Description Must Have Technical/Functional Skills ~ We are looking for a Senior Site Reliability Engineer (SRE) with deep experience in AWS infrastructure...Suggested
- ...Qualifications: 8+ years of software engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources. Willingness to work on-site at stated location in the job opening....SuggestedFor contractorsWork experience placement
$86k - $105k
...generation of application infrastructure and to be responsible for reliability, automation and scalability using and the latest best... ...certifications. Minimum of 2 years prior DevOps, software engineering or related experience. Must be able to work different schedules...SuggestedHourly payWork at officeImmediate startVisa sponsorshipWork visaFlexible hours$93.9k - $156.5k
...work model, requiring 2 days per week on-site at our corporate office 20 S Wacker Dr,... ...low-latency performance and rock-solid reliability to seamlessly handle the world's busiest... ...successful candidate will work alongside senior engineers to learn how we observe, monitor,...Work at officeLocal areaWorldwide2 days per week$97.5k - $130k
...Overview: Site Reliability Engineer Salary: $97,500-$130,000 Role Summary The SRE & Cloud Engineer will support the client by accelerating remediation of security issues and contributing to cloud modernization initiatives. This role is hands-on and execution...- ...Sre Engineer We are seeking a highly capable engineer to join our dynamic SRE team. This role is ideal for someone with a strong background... ...Global Technology file transfer teams to ensure secure and reliable operations. .NET Logging & Monitoring Libraries Maintain and...
- ...Site Reliability Engineer (SRE) Immediate need for a talented Site Reliability Engineer (SRE). This is a 12+ months contract opportunity with long-term potential and is in Chicago, IL (Hybrid). Key Requirements and Technology Experience: ~ Must have skills:...Contract workLocal areaImmediate start
- ...Overview: Senior Site Reliability Engineer (SRE) Location: Chicago, IL (Onsite) Type: Contract Role Overview: We are seeking a Senior Site Reliability Engineer (SRE) with strong expertise in AWS infrastructure, automation, observability, and production...Contract work
- ...Site Reliability Engineer As a Site Reliability Engineer, you will build and secure infrastructure supporting our AI platform with special attention to safeguarding US customer data and supporting the Aerospace and Defense Industrial Base. You'll have strong ownership...
$57k - $113k
...Site Reliability Engineer Huntington will not sponsor applicants for this position for immigration benefits, including but not limited to assisting with obtaining work permission for F-1 students, H-1B professionals, O-1 workers, TN workers, E-3 workers, among other...Full timeH1bWork at officeRemote workWork from homeFlexible hours$85k - $130k
...Site Reliability Engineer Passionate about precision medicine and advancing the healthcare industry? Recent advancements in underlying technology have finally made it possible for AI to impact clinical care in a meaningful way. Tempus' proprietary platform connects...- ...Site Reliability Engineer We are looking for a Senior Reliability Engineer to join our Platform team. In this position, you will be responsible for maintaining, designing, implementing and upgrading our cloud infrastructure to support our microservices platforms. We...Temporary workFlexible hours
$127.33k - $159.17k
...Service Management. It's our goal to always provide an engaging, relevant, and simple experience for our customers. The Site Reliability Engineer (SRE) - Edge Platform is a key member of the Edge Operations and SRE team within Global Technology Infrastructure &...Local areaFlexible hoursShift work$130k - $170k
Senior Site Reliability Engineer About Us Founded in 2014, we offer the industry’s first and only cloud‑based, fully‑customisable, end‑to‑end software solution to automate securities‑based lending from origination through the life of the loan. By combining thought leadership...Full timeFlexible hoursShift work- Salary: Up to $500k split between base and bonus The company is seeking an experienced Site Reliability Engineer with strong Python and Linux skills to join a high-performance technology environment. The role is suited to someone who enjoys solving operational challenges...
$150k - $155k
Site Reliability Engineer Hybrid (3 days onsite, 2 days remote) full‑time. No visa sponsorship. Base pay: $150,000 - $155,000 per year, subject to skills and experience. A prestigious company seeks a Site Reliability Engineer focused on observation, logging, and capacity...Full timeWork experience placementRemote workVisa sponsorship- JPMorganChase is looking for a Technology Support III team member in Chicago to ensure operational stability of production application flows. You will diagnose complex issues and maintain high-volume payment systems. The ideal candidate will have at least 3 years of IT ...
- Early Warning is looking to fill the role of Staff Site Reliability Engineer in Chicago, IL. This position involves partnering with development teams to define and implement standards in availability and resiliency. Candidates must have over 8 years of experience in managing...
- As a Site Reliability Engineer, you will build and secure infrastructure supporting our AI platform with special attention to safeguarding US customer data and supporting the Aerospace and Defense Industrial Base. You'll have strong ownership of US operations while collaborating...Immediate start
$250k - $350k
...and India, where quantitative researchers, engineers, traders, and operational teams work together... ...that boost stability, throughput, and reliability Qualifications Minimum of 3 years’ experience in production support, site reliability, or infrastructure operations in...Full time$190.8k - $267.1k
...unique opportunity to leave your mark on one of the most influential and trafficked corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team, you’ll use your knowledge of distributed systems and architecture to improve the reliability...Work experience placementHome officeFlexible hours- Site Reliability Engineer at the organization. Key technologies: Kubernetes, Prometheus, Grafana. Key Responsibilities Define and track SLOs, SLIs and error budgets Design and implement observability stacks (metrics, logging, tracing) Automate toil and improve system...
$106k - $130k
...sponsorship. Overall Purpose To create and maintain the next generation of application infrastructure and to be responsible for reliability, automation and scalability using the latest best practices. Essential Functions Implement software and tools to improve performance...Hourly payWork experience placementWork at officeImmediate startVisa sponsorshipWork visaFlexible hours$150k - $200k
...message the job poster from Selby Jennings Recruitment Consultant @ Selby Jennings | Financial Technology We are seeking a Site Reliability Engineer to join our team and assist with the design, development, and administration of our trading and research systems. This...Full timeWork at office- No H1 or C2C. Must be Permanent Resident or US Citizen Senior Site Reliability Engineer Description and Requirements About Our Team We are building Quantum , a next‑generation hybrid AI platform that spans Windows, Android, and cloud. As part of this vision, we are expanding...Permanent employmentRemote work
$106.28k - $145k
CCC Information Services in Chicago is looking for a Senior Site Reliability Engineer to enhance and support their multi-cloud solutions. This hybrid position offers a salary range of $106,277.25 to $145,000.00, and candidates should have over two years of experience in...- ...building and running systems that must perform reliably under real-time market conditions. The culture is highly collaborative, engineering-driven, and focused on continuous... ...a related field 3+ years of experience in site reliability, systems engineering, or technical...
$145k - $175k
Job Overview The Site Reliability Engineer supports deployments, cloud infrastructure, and monitoring systems that power Rewards Network's applications and services. This hybrid role requires in‑office presence 3 days a week (Tuesday‑Thursday) in Chicago and is open to...Full timeTemporary workWork at officeFlexible hours3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to #23183 - Site Reliability Engineer. Be the first to apply!
- site reliability engineer Chicago, IL
- site reliability engineer remote Chicago, IL
- site reliability engineer sre Chicago, IL
- IT site lead Chicago, IL
- website coordinator Chicago, IL
- on site coordinator Chicago, IL
- junior website developer Chicago, IL
- site safety Chicago, IL
- site services specialist Chicago, IL
- site acquisition specialist Chicago, IL


