Site Reliability Engineer
$72.1k - $173.04kOak St. Health
Software Development Engineer, Site Reliability Engineering (SRE)
We're building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you'll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time.
Position Summary
We are seeking a highly skilled Software Development Engineer, Site Reliability Engineering (SRE), for Retail and Pharmacy platforms to drive reliability, scalability, and operational excellence. The ideal candidate will leverage AIOps practices and AI-powered tools to improve observability, automate incident response, reduce operational noise, and enable data-driven decision making across distributed systems.
Required Qualifications
- Strong experience in Software Engineering, SRE, or Platform Engineering roles.
- Experience in Edge/Distributed software Architecture model.
- Proficiency in one or more languages: Java, Python, Go, or similar.
- Hands-on experience with cloud platforms (AWS/Azure/GCP) and distributed systems.
- Experience in monitoring/observability tools (e.g., Prometheus, Grafana, Splunk, AppInsights, OpenTelemetry).
- Solid understanding of CI/CD, infrastructure as code, and automation practices.
Preferred Qualifications
- Experience with AIOps platforms (e.g. Dynatrace, Datadog Watchdog, Splunk ITSI, Azure Monitor with AI capabilities).
- Familiarity with AI/ML concepts applied to operations, including anomaly detection, predictive analytics, and event correlation.
- Hands-on experience using AI coding and productivity tools (e.g., GitHub Copilot, ChatGPT/Copilot, AI-assisted debugging tools) in daily development workflows.
- Experience building or integrating AI-driven automation (runbooks, bots, intelligent alerting systems).
- Knowledge of real user monitoring (RUM) with AI-driven insights and performance analytics.
- Exposure to AIOps-driven incident management and self-healing architectures.
- Strong analytical mindset with ability to leverage data insights and AI recommendations for operational decisions.
Education
Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field, or equivalent practical experience.
Anticipated Weekly Hours: 40
Time Type: Full time
Pay Range: $72,100.00 - $173,040.00
This pay range represents the base hourly rate or base annual full-time salary for all positions in the job grade within which this position falls. The actual base salary offer will depend on a variety of factors including experience, education, geography and other relevant factors. This position is eligible for a CVS Health bonus, commission or short-term incentive program in addition to the base pay range listed above.
Our people fuel our future. Our teams reflect the customers, patients, members and communities we serve and we are committed to fostering a workplace where every colleague feels valued and that they belong.
Great benefits for great people
We take pride in offering a comprehensive and competitive mix of pay and benefits that reflects our commitment to our colleagues and their families.
This full-time position is eligible for a comprehensive benefits package designed to support the physical, emotional, and financial well-being of colleagues and their families. The benefits for this position include medical, dental, and vision coverage, paid time off, retirement savings options, wellness programs, and other resources, based on eligibility.
$140k - $210.9k
...Senior Site Reliability Engineer Federal Reserve Financial Services (FRFS) delivers a suite of payments services to financial institutions via FedLine® Solutions, FedNowSM, Fedwire®, National Settlement Service (NSS), FedCash®, FedACH® (Automated Clearing House), and...SuggestedFull timeTemporary workPart timeWork at officeShift work$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...SuggestedPermanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$146.4k - $263.6k
...passionate about cutting edge technology? Do you enjoy working with a diverse multi-national team of engineering talents? Join our highly skilled Site Reliability team Our team designs, develops, and manages applications and infrastructure that support Akamai'...SuggestedWork experience placementWork at office$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...SuggestedWork experience placementWork at office$104.9k - $174.7k
...scale, 24x7, distributed and fault-tolerant systems within agreed reliability objectives, whilst enabling the fast flow of feature and... ...strong automation skills. About team; This diverse team of Engineers in assisting multiple product teams as we continue to innovate...SuggestedLocal areaImmediate startWorldwide$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...Work experience placementWork at office- ...Job Title: Site Reliability Engineer Location: Remote with Quarterly visits to Chennai, Tamil Nadu, India Duration: Full-Time bout BigRio: BigRio is a remote-based, technology consulting firm headquartered in Boston, MA. We deliver software solutions...Full timeRemote work
$146.4k - $263.6k
...passionate about cutting edge technology? Do you enjoy working with a diverse multi-national team of engineering talents? Join our highly skilled Site Reliability team Our team designs, develops, and manages applications and infrastructure that support Akamai's Compute...Work experience placementWork at office$150k - $190k
...Senior Site Reliability Engineer (SRE) This role is located in Somerville, MA - We are a hybrid work environment and are in the office 3+ days per week. Tulip, the leader in AI-native frontline operations, is helping companies around the world equip their workforce...Temporary workWork at officeFlexible hours3 days per week$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...Temporary workImmediate startFlexible hoursShift work- ...Site Reliability Engineer Cambridge, MA About Watershed Our vision is to become the leading biocomputing platform. The future of biology is in big data analysis, and we are on a mission to accelerate digital drug discovery with the Watershed platform. Watershed...
- ...Senior Site Reliability Engineer As a Senior Site Reliability Engineer at Blitzy's Cambridge headquarters, you will be the backbone of our platform's reliability, scalability, and operational excellence. You'll work at the intersection of software engineering and infrastructure...
- ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the... ...ideal candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based...Full time
- ...work and education) Education Desired: Bachelor of Computer Engineering Travel Percentage: 0% We are FIS. Our technology powers the... ...python, etc. A mindset/desire to improve application systems reliability and automate manual support tasks, to facilitate continuous...Full timeWork at officeRemote workWork from homeFlexible hours
$128k - $160k
...at a time. To those who see AI as a driver of progress, come build the future together. The Crown Is Yours As a Senior Site Reliability Engineer, you'll build and scale the critical infrastructure behind every product. In this role, you'll take on complex challenges...Full timeImmediate start$146.4k - $263.6k
...uses large datasets to analyze and measure the performance and reliability of our platform. We are networking data scientists: we... ...organizing project‑specific work. Driving partnership with Engineering, Operations and Product teams to help guide adoption of new features...Work at office$84.9k - $209.5k
...spirit that promotes an upbeat and creative environment. We are unencumbered and will need your contribution to make it a special engineering center with the focus on excellence. Health Data Intelligence Platform has a rare opportunity to play a critical role in how...Temporary workImmediate startFlexible hours$51.9 per hour
...Company: Allegheny Health Network Job Title: Site Reliability Engineering – Clinical & Facility Services General Overview This role ensures the reliability, availability, and performance of critical healthcare IT systems in the Environment of Care (EOC), supporting patients...Local area$160k - $225k
...tens of thousands of users across hundreds of organizations globally. About the Role Manifold is looking for a Staff Site Reliability Engineer (SRE) to work at the intersection of AI, data infrastructure, and life sciences. In this high-impact role, you will help...- ...Site Reliability Engineering (SRE) Team Lead The Site Reliability Engineering (SRE) team is foundational to the growth and scale of our platform. You and your team will help advance several initiatives tied to automation, SRE culture, and cloud architecture. You will...Shift work
$127k - $249k
The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper). As...Work at officeLocal areaRemote workWorldwideFlexible hours$126k - $248k
...As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB’s cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and...Local areaRemote workWorldwideFlexible hours- ...collaborating with Verisk to connect them with exceptional professionals for this role. Description We are hiring a Senior Software Engineer with deep expertise in AI/ML engineering and data-intensive systems to join our Catastrophic and Risk Solutions team. You will be...Work at officeFlexible hours
$205k - $260k
...software solutions. Solve business problems through innovation and engineering practices. Involved in all aspects of the Software Development... ...These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup...Full time$166k - $220k
...About The Team The Reliability Engineering team partners across Anduril’s engineering, manufacturing, and operations organizations to ensure our autonomous systems survive real‑world conditions and deliver consistent performance for our customers. The team’s mandate spans...Full time- ...supports the U.S. Air Force Cloud One Architecture and Common Shared Services contract and currently has an opening for a Reliability Engineer . The Reliability Engineer is responsible for ensuring the availability, performance, scalability, and resiliency of missioncritical...Contract workRemote work
$150k - $195k
...Senior Reliability Engineer At WHOOP, we're on a mission to unlock human performance and healthspan. WHOOP empowers members to perform at a higher level through a deeper understanding of their bodies and daily lives. WHOOP is seeking a Senior Reliability Engineer...Full timeWork at officeRelocation$103.71k - $138.28k
...demonstrated knowledge and experience in system architecture and engineering disciplines. Specific technical knowledge of enterprise level... ...Amazon Web Services. Supports due diligence activities including site surveys, design, design review, bill of materials creation,...Temporary workRemote work- hackajob is collaborating with Verisk to connect them with exceptional professionals for this role. Description Come be part of something new and exciting at a well-established analytical company. Help build scalable solutions for an industry leading catastrophe...
$204k - $348k
...Sr Principal/ Principal Software Engineer, AI Lab Execution System Cambridge, MA USA; San Francisco, CA USA Your Impact at LILA... ...user interfaces, services, high-performance APIs, databases, and reliability-critical systems that integrate advanced AI frameworks with...Full timeWork at officeLocal areaFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!


