Site Reliability Engineer
ICONMA
Site Reliability Engineer
Our client, an IT Services and Consulting company, is looking for a Site Reliability Engineer for their Woonsocket, RI/ Hybrid location. Responsibilities include:
- Owning and driving the end-to-end reliability, availability, and performance of critical retail and pharmacy technology platforms across hybrid cloud and on-premises environments.
- Establishing and maintaining SLI/SLO health, alerting strategies, observability standards, and business-aligned monitoring for the assigned application domain.
- Leading production incident response as Incident Commander, driving root cause analysis, postmortems, and continuous reliability improvements.
- Partnering with engineering, product, and operations teams to embed reliability, resiliency, scalability, and operational readiness into system design and delivery.
- Building and optimizing automation, self-service capabilities, and operational tooling to eliminate toil, improve efficiency, and reduce manual intervention.
- Designing and executing proactive reliability initiatives, including production readiness reviews, dependency risk assessments, fault injection, and chaos engineering exercises.
- Mentoring engineers, championing SRE best practices, and enabling teams to independently detect, respond to, and learn from production issues with minimal SRE involvement. Influencing organizational adoption of SLO-driven engineering, observability, incident management, and reliability practices through collaboration, credibility, and measurable outcomes.
Requirements include:
- 8+ years of senior software engineering experience in SRE, DevOps, platform engineering, or related production-systems roles in distributed systems at production scale with active on-call responsibility.
- Demonstrated experience as an on-call Incident Commander (IC) for P1 or P2 incidents — structured leadership updates, not just participant involvement.
- Experience tuning and validating time-series anomaly detection models in a production observability context — this is a required qualification, not a preferred one; anomaly-based detection is a core function of this role.
- Strong programming proficiency in Python, React, and Java at production quality — capable of writing operational tooling that other engineers will rely on.
- Hands-on experience designing SLIs, SLOs, and managing error budgets for customer-facing or business-critical services.
- Deep observability platform experience: Prometheus, Grafana, Open Telemetry, and at least two of the log aggregation solutions (Loki, Splunk, Elasticsearch).
- Fleet-scale deployment awareness: familiarity with progressive rollout strategies, blast radius management, and configuration drift as a reliability risk in large unattended node deployments.
- Strong cloud platform expertise in Google Cloud Platform (GCP) and Rancher K3s.
- Advanced Kubernetes operational experience: debugging, resource management, networking policies, and workload failure modes. Experience with AI-assisted tooling and development.
- Experience diagnosing and resolving workflow orchestration issues, batch processing failures, scheduler performance problems, and building observability on data pipelines: Apache Airflow and Tidal.
- Experience owning Production Readiness Reviews or service launch gates.
- Strong proficiency in transforming large-scale operational and telemetry data into actionable business insights using SQL-based analytics and reporting frameworks: Google BigQuery, PostgreSQL.
- Hands-on chaos or fault injection experience.
- TIC (Technical Incident Commander) certification or equivalent structured incident command training.
- Experience operating distributed systems in retail, pharmacy, healthcare, or other operationally sensitive environments where failures have direct patient or customer impact.
- LLM integration for operational use cases (alert summarization, runbook suggestion, incident triage assistance) — design or implementation experience.
- Experience with streaming data platforms: Kafka.
- Experience with service mesh and traffic management: Istio, Envoy.
- Infrastructure-as-code proficiency at production scale: Terraform or Ansible.
- Years of Experience: 14.00 Years of Experience
Why Should You Apply?
- Health Benefits
- Referral Program
- Excellent growth and advancement opportunities
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Woonsocket, RI vacancy
$50 - $60 per hour
...Job Description A large healthcare client of ours is seeking a Site Reliability Engineer to join a fast-paced operations team responsible for maintaining platform stability, managing critical incidents, and supporting enterprise applications. This individual will play...Suggested$92.7k - $203.94k
...Senior Site Reliability Engineer We're building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you'll be surrounded by passionate colleagues who care deeply, innovate with purpose...SuggestedHourly payFull timeTemporary work$168k - $200k
...is passionate about creating transformative change in healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the forefront of building and operating a resilient, observable, and...SuggestedRemote work$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SuggestedTemporary workImmediate startFlexible hoursShift work$121.4k - $218.6k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...SuggestedWork experience placementWork at office$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office$140k - $210k
...Our Mission As the world’s number 1 job site*, our mission is to help people get jobs. We strive to cultivate an... ...Comscore, Total Visits, March 2026) Day to Day As an Engineering Manager in Site Reliability Engineering at Indeed, you will manage and grow a team that...Work experience placementLocal area$138.9k - $180.6k
...Description: Saab Inc.'s Autonomous and Undersea Systems (AUS) division is seeking an innovative and experienced Senior Software Engineer to participate on technical teams defining, architecting, implementing, integrating, verifying, delivering, and maintaining...Temporary workFor contractorsWork experience placementCasual workLocal areaRemote work- ...teams across manufacturing, quality, supply chain, finance, engineering, and other functions to reduce manual work, improve access to... ...help business users adopt new workflows successfully. · Track reliability, usage, time savings, cost reduction, quality improvement,...
- ...initiatives across global enterprises. Job Summary We are seeking an experienced Sr. Java Developer to join a high-performing engineering team supporting enterprise application development initiatives. The ideal candidate will have strong expertise in Java, Node.js,...Local area
- ...500 companies in the financial, banking, insurance and billing industries across the U.S. We are currently looking for a Software Engineer to join our Application Development team at one of our East coast facilities - Orlando, Florida or Lincoln, Rhode Island. This role...Work at officeRemote workMonday to Friday
$101.97k - $203.94k
...influence the overall technical strategy and innovation of digital engineering initiatives. Primary Responsibilities Design and Develop... ..., performs debugging, and resolves issues to guarantee the reliability, stability, and high quality of solutions. Creates and...Hourly payFull timeTemporary workWork at officeLocal area3 days per week$78.03k - $97.53k
Grand Junction, CO Mooresville, IN Wildwood, FL Santa Rosa, CA Fontana, CA Elgin, IL Del Rio, TX Green Bay, WI Tannersville, PA Katy, Texas Irving, TX Norcross, GA Kettleman City, CA Westfield, MA Lakeville, MN Justin, TX Byhalia, MS Benicia, CA Phoenixville, PA Milton...Full timeTemporary workWork experience placementLocal areaShift workDay shift- ...such as: Snowflake BigQuery Amazon Redshift Azure Synapse Databricks ~ Strong understanding of cloud-based data engineering and analytics architectures. SQL Strong SQL skills for data analysis, transformation, troubleshooting, and query...Contract work
$144.2k - $288.4k
...are looking for a hands-on, passionate engineering leader to join a high-energy, mission-driven... ...to ensure quality, efficiency, and reliable execution across teams ~ Organizational... ...engineering teams, including multi‑site and fully remote models Demonstrated...Hourly payFull timeTemporary workWork experience placementLocal areaRemote work$100k - $110k
...Job Description Job Description Sr React Software Engineer Branding Brand is searching for a Sr Software Engineer to help create mobile apps and sites for an international portfolio of high-profile brands in retail and hospitality. Ideal candidates are skilled...Work at officeWork from home- Must Haves: ~8+ years of experience as a Scrum Master or in a similar Agile leadership role. ~ Expertise managing and overseeing the release train and deployment schedules ~ Strong understanding of Agile methodologies (Scrum, Kanban) and frameworks. ...
- ...candidate will have strong experience in Java development, API engineering, cloud platforms, and modern integration technologies.... ...retrospectives. Optimize application performance, scalability, reliability, and security across cloud and hybrid environments. Support front...Remote job
- ...CVS Health in Rhode Island seeks a hands-on Lead Director of Software Engineering to guide the Health100 platform engineering roadmap and integration strategy. You will manage cross-functional teams, ensure high-quality, scalable software, and align with business goals...
- ...Databricks during the first phase of the engagement • Partner with the team to expand scope beyond file ingestion into broader data engineering and AI workflows • Help solve emerging problems in the audit and submission integrity space, where requirements are still taking...
- Job Title Manager is looking for a.NET Developer who is interested in CTH Must have experience with Vendor Applications and experience working with one vendor to another. Lot of coordination is needed in order to be successful in this role. Must have good multi-tasking ...2 days per week3 days per week
- Work Location & Reporting Address: Woonsocket, RI 02895 (hybrid and if required twice a week) Vendor Rate: $115/hr on C2C. $100/hr on W2 only GC/USC Minimum years of experience required: 12+ years of experience Certification needed: NA Must Have Skills: SAP EWM Module...Work experience placement
$175.1k - $334.75k
## Executive Director, Software Engineering AI SolutionsApply: Remote/Hybrid: RI - Woonsocket: Full time: Posted Today: End Date: October 12, 2026 (26 days left to apply): R0971791We’re building a world of health around every individual — shaping a more connected, convenient...Hourly payFull timeTemporary workRemote work- Senior Cloud Engineer Required Skills & Experience: 5-7 years of cloud or infrastructure engineering experience, including 3+ years of... ...capabilities that allow applications and systems to exchange data reliably, including Pub/Sub and the supporting networking, security,...
$200k
...About the Role Location: Providence, RI — on-site 1x per week, hybrid. We're hiring a Senior Software Engineer to take the reins on the software behind our... ...feature additions, debugging, performance and reliability improvements, and incremental refactoring Lead...Immediate start$58 - $59 per hour
job summary: Experience in developing pattern-based solutions and abstraction concepts (building libraries, interceptors, SDK etc.). Knowledge on data structure concepts such as Binary Tree, Binary Search Tree, Graphs and hands-on experience on Tree operations/ traversals...Hourly payContract workTemporary workWork experience placement$100k - $125k
...Overview Branding Brand is seeking a Senior / Lead Frontend Platform Engineer to architect and build a scalable frontend platform that... ...largest and fastest-growing provider of mobile commerce apps and sites for retailers. Branding Brand is a supporter of the...Work at officeImmediate startWork from home- ...CVS Health Corporation is seeking an Executive Director, Software Engineering – AI Solutions to lead the design, development, and delivery of AI-powered software at enterprise scale. Reporting to the VP of Engineering, you will build and scale a high-performing software...
$114.6k - $234.6k
...detection, and instrumentation Maintain documentation as a source of truth Cross-Functional Leadership Collaborate across OHD, engineering, operations, supply chain, and software teams Align with external partners (OEMs, silicon vendors) Drive technical decisions...Temporary workFlexible hours$218.6k - $286.1k
...Enterprise Premier Segment desires to bring onboard a Solutions Engineer (SE), that will be working as a trusted advisor and partner to top... ..., and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible to...Full timeTemporary workLocal areaImmediate startRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
Related searches
- on-site clinical research associate (traveling/remote) Woonsocket, RI
- construction site safety Woonsocket, RI
- junior site reliability engineer
- site reliability engineering manager
- site reliability engineer
- lead site reliability engineer
- site reliability engineer remote
- site reliability engineer sre
- website qa testing
- site buyer





