Site Reliability Engineer
$72.1k - $173.04kOak St. Health
Software Development Engineer, Site Reliability Engineering (SRE)
We're building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you'll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time.
Position Summary
We are seeking a highly skilled Software Development Engineer, Site Reliability Engineering (SRE), for Retail and Pharmacy platforms to drive reliability, scalability, and operational excellence. The ideal candidate will leverage AIOps practices and AI-powered tools to improve observability, automate incident response, reduce operational noise, and enable data-driven decision making across distributed systems.
Required Qualifications
- Strong experience in Software Engineering, SRE, or Platform Engineering roles.
- Experience in Edge/Distributed software Architecture model.
- Proficiency in one or more languages: Java, Python, Go, or similar.
- Hands-on experience with cloud platforms (AWS/Azure/GCP) and distributed systems.
- Experience in monitoring/observability tools (e.g., Prometheus, Grafana, Splunk, AppInsights, OpenTelemetry).
- Solid understanding of CI/CD, infrastructure as code, and automation practices.
Preferred Qualifications
- Experience with AIOps platforms (e.g. Dynatrace, Datadog Watchdog, Splunk ITSI, Azure Monitor with AI capabilities).
- Familiarity with AI/ML concepts applied to operations, including anomaly detection, predictive analytics, and event correlation.
- Hands-on experience using AI coding and productivity tools (e.g., GitHub Copilot, ChatGPT/Copilot, AI-assisted debugging tools) in daily development workflows.
- Experience building or integrating AI-driven automation (runbooks, bots, intelligent alerting systems).
- Knowledge of real user monitoring (RUM) with AI-driven insights and performance analytics.
- Exposure to AIOps-driven incident management and self-healing architectures.
- Strong analytical mindset with ability to leverage data insights and AI recommendations for operational decisions.
Education
Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field, or equivalent practical experience.
Anticipated Weekly Hours: 40
Time Type: Full time
Pay Range: $72,100.00 - $173,040.00
This pay range represents the base hourly rate or base annual full-time salary for all positions in the job grade within which this position falls. The actual base salary offer will depend on a variety of factors including experience, education, geography and other relevant factors. This position is eligible for a CVS Health bonus, commission or short-term incentive program in addition to the base pay range listed above.
Our people fuel our future. Our teams reflect the customers, patients, members and communities we serve and we are committed to fostering a workplace where every colleague feels valued and that they belong.
Great benefits for great people
We take pride in offering a comprehensive and competitive mix of pay and benefits that reflects our commitment to our colleagues and their families.
This full-time position is eligible for a comprehensive benefits package designed to support the physical, emotional, and financial well-being of colleagues and their families. The benefits for this position include medical, dental, and vision coverage, paid time off, retirement savings options, wellness programs, and other resources, based on eligibility.
$160k - $200k
...data, ideally using promQLKey Responsibilities:Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coachPerform incident...SuggestedTemporary workWork at officeLocal areaFlexible hours3 days per week$134.25k - $214.8k
...matters at a company where you matter.Your ImpactAre you an engineer who gets excited about the challenge of making complex distributed... ...it.You will be part of the Observability team within Axon's Site Reliability organization — a focused team responsible for Axon's metrics,...SuggestedWork experience placementWork at officeRemote work$74.1k - $148.3k
...systems. Facilitate service capacity planning and demand forecasting, software performance analysis, and system tuning. As a Site Reliability Engineer, you will solve interesting technical challenges by defining, designing, deploying, and solving key Oracle Cloud services,...SuggestedTemporary workImmediate startFlexible hours$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...SuggestedWork experience placementWork at office- ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the... ...ideal candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based...SuggestedFull time
$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$160k - $200k
...Senior Site Reliability Engineer This role is located in Somerville, MA – We are a hybrid work environment and are in the office 3+ days/per week. Tulip, the leader in AI-native frontline operations, is helping companies around the world equip their workforce with...Temporary workWork at officeLocal areaFlexible hours3 days per week$134.25k - $214.8k
...change. Constantly grow as you work hard for a mission that matters at a company where you matter. Your Impact As a Senior Site Reliability Engineer within the APX SRE organization, you'll focus on delivering practical, scalable solutions to support the reliability and...Work experience placementWork at officeRemote workFlexible hours$130k - $150k
...systems and hybrid infrastructure, meaning experience with cloud technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are reliable, scalable, and performant across on-premises and cloud...Work at officeWork from home3 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....Work at officeLocal areaRemote workWorldwideFlexible hours$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office- ...Job Title: Site Reliability Engineer Location: Remote with Quarterly visits to Chennai, Tamil Nadu, India Duration: Full-Time bout BigRio: BigRio is a remote-based, technology consulting firm headquartered in Boston, MA. We deliver software solutions ranging...Full timeRemote work
$127k - $249k
...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas... ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This...Local areaRemote workWorldwideFlexible hours$160k - $200k
Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware...Local areaRemote work$166k - $220k
...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &... ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril...Full timeWork experience placementImmediate start$160k - $225k
...tens of thousands of users across hundreds of organizations globally. About the Role Manifold is looking for a Staff Site Reliability Engineer (SRE) to work at the intersection of AI, data infrastructure, and life sciences. In this high-impact role, you will help...- ...Information Technology group delivers secure, reliable technology solutions that enable DTCC to... ...This RoleAs a Senior Application Support Engineer, you will help power DTCC's global... ...trade processing and settlement.Leveraging Site Reliability Engineering (SRE) principles,...Remote workFlexible hours
- ...Site Reliability Engineering (SRE) Team Lead The Site Reliability Engineering (SRE) team is foundational to the growth and scale of our platform. You and your team will help advance several initiatives tied to automation, SRE culture, and cloud architecture. You will...Shift work
$130k - $140k
...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient...Ongoing contractFull timeTemporary workWork experience placement$151k - $297k
...As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB's cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and...Local areaRemote workWorldwideFlexible hours$166k - $220k
...way through to the field. When something breaks in a deployed environment, we fix it. About the Role We're looking for a Site Reliability Engineer to join the Imaging team. This is not a product development role, and it isn't a traditional cloud-SRE role either. You are...Full timeWork experience placementRemote workWeekend workDay shift$40 - $45.78 per hour
...Job Description Job Description Site Reliability Engineer 1 Job Details Site Reliability Engineer 1 (Contract) Location: Waltham, MA 02451 (Hybrid) Duration: 10/22/2025 to 4/03/2026 Team: Campaign Core RD US Key Responsibilities: Deploy and manage...Hourly payContract work- ...infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands,... ...and deployment workflows for accuracy and reliability. Work with AWS, Azure, GCP,... ...Azure DevOps Cloud Infrastructure Site Reliability Engineering (SRE) Platform...Remote jobFor contractors
$146k - $194k
...autonomy, AI, computer vision, sensor fusion, and networking technology to the military in months, not years.ABOUT THE TEAMThe Reliability Engineering team partners across Anduril's engineering, manufacturing, and operations organizations to ensure our autonomous systems...Full timeWork experience placementImmediate start$266.2k - $425.9k
...TeamHubSpot's Developer Acceleration group empowers over 2,000 engineers to build, test, and deploy their code at scale. The release... ...the tooling that underpins incident management and production reliability. This is not a team focused on maintaining existing systems. It...Live outWork at officeRemote workShift work$191k - $253k
...priorities, we want you to join Anduril’s Maritime Division and help us build the future of defense capability.About the JobSr. Software Engineers independently drive the delivery of a variety of software integrated in to our products. This includes autonomy, simulation, data...Full timeWork experience placementImmediate startRemote workFlexible hours$135k - $325k
Job OverviewWe are seeking an experienced Engineer to join our Trading Systems team within the Front Office Systems group. This role... ...engineering best practices.Identify opportunities to improve system reliability, performance, scalability, automation, and developer...Full timeLocal area$93.8k - $125.1k
...Home Secure. What You'll DoSimplisafe is looking for a Software Engineer II to join our User Systems team to develop and maintain our... ...User Systems team.Optimize our backend systems for performance, reliability and scalability.Ensure high quality standards by participating...Work experience placementWork at office$115k - $325k
...benchmarks, instruments, holdings and reference entities that feed our trading cycle. The ideal candidate is an accomplished and driven engineer with experience building complex systems using modern platforms. We value technologists who are passionate about innovation and...Full timeLocal area$52.7 - $62 per hour
...TitleReliability EngineerJob Description SummaryThis individual will provide support for Facility Operations/Engineering department sites as part of the Reliability Engineering team. The ideal candidate will be responsible for ensuring the reliability and performance of...Minimum wageFull timeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer remote Boston, MA
- site reliability engineer Boston, MA
- site reliability engineer sre Boston, MA
- junior website developer Boston, MA
- website content developer Boston, MA
- on site coordinator Boston, MA
- website coordinator Boston, MA
- site leader Boston, MA
- site recruiter Boston, MA
- historic site Boston, MA


