Site Reliability Engineer - Disaster Recovery & Business Continuity
$130k - $150kCRA International
About Charles River AssociatesFor over 50 years, Charles River Associates has been a premier consulting firm that offers employees a place to learn from a diverse group of consultants, industry experts, and academics. At CRA you will be exposed to leading minds who use economic, financial, and business analysis to solve complex world problems for an impressive roster of clients, including major law firms, Fortune 100 companies, and government agencies. Through a collegial environment, formal and informal training opportunities, and a broad array of professional development resources, your experience at CRA will open doors for you throughout your career.The Information Technology (ITS) department at Charles River Associates is currently a team of more than 40 professionals dedicated to enhancing, maintaining, and developing the firm's technology infrastructure and security. The team is comprised of four functions:Service Delivery & TelecomEnterprise Application SolutionsInfrastructure, Networking and Cloud SolutionsInformation SecurityInformation Technology staff are based in the Boston, Chicago, London, Munich, New York, Oakland, San Francisco, College Station and Washington, DC offices.Mainly a Microsoft house, CRA is looking to maximize the performance of our on-premise systems and hybrid infrastructure, meaning experience with cloud technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are reliable, scalable, and performant across on-premises and cloud environments. This role blends software engineering and operations practices to reduce manual toil through automation, improve service observability, and strengthen incident response. The SRE partners closely with infrastructure, security, application, and service delivery teams to define measurable reliability targets (SLIs/SLOs), implement resilient architectures, and drive continuous improvement through blameless post-incident learning.Key ResponsibilitiesHands-on System Engineering experience with core enterprise infrastructure platforms and services, including Windows Server, VMware vSphere, VMware Site Recovery Manager (SRM), SAN technologies, and the Rubrik ecosystem, with the ability to understand dependencies, recovery workflows, and failure modes across on-premises and cloud environmentsService Ownership & Reliability Targets: Partner with service owners to define and maintain service level indicators (SLIs) and service level objectives (SLOs) for availability, latency, and performance; track error budgets and reliability risk.Observability: Implement and continuously improve monitoring, logging, alerting, and dashboards to provide actionable, symptom-based signals and reduce mean time to detect/respond (MTTD/MTTR).Blameless Postmortems & Continuous Improvement: Facilitate post-incident reviews, identify root causes and contributing factors, and drive remediation items to completion; standardize learnings into runbooks and operational practices.DR Testing Program Build-Out: Design and launch a scalable DR testing program (scope, test types, cadence, success criteria, and evidence capture) in partnership with application, infrastructure, and security teams; maintain runbooks and lead regular tabletop and technical recovery exercises to validate RTO/RPO assumptions and improve recoverability.DR Readiness: Contribute to reliability architecture and disaster recovery readiness for key services, including dependency mapping, recovery testing inputs, and validation of recovery procedures.Cross-Functional Collaboration: Work day-to-day with infrastructure, network, cloud, security, and application teams to improve operational excellence, reliability culture, and shared ownership of production outcomes.Relevant Skills & ExperienceExperience operating and improving reliability of production services (on-prem and/or cloud), including incident response, operational readiness, and service ownershipWorking knowledge of SRE concepts and practices such as SLIs/SLOs, error budgets, monitoring/alerting strategy, and blameless postmortemsExperience with observability tooling and practices (logs, metrics, tracing, dashboards) and using data to drive reliability and performance improvementsExperience with disaster recovery orchestration and recovery testing using VMware Site Recovery Manager (SRM) and Azure Site Recovery (ASR) (or similar public cloud DR services)Proven experience building and operating a DR testing program, including dependency mapping, test planning, coordination across stakeholders, execution of tabletop and technical failover tests, documentation of results, and tracking remediation actions to closureStrong cross-functional communication and teamwork skills; comfortable partnering with engineering, security, and operations teams to drive shared outcomesAbility to document and standardize operational procedures (runbooks), participate in on-call rotations, and manage multiple priorities in a fast-moving environmentCareer Growth and Benefits CRA’s robust skills development programs, including a commitment to offering 100 hours of training annually through formal and informal programs, encourage you to thrive as an individual and team member. Beginning with research and analysis skill building, training continues with technical training, presentation skills, internal seminars, and career mentoring and performance coaching from an assigned senior colleague. Additional leadership and collaboration opportunities exist through internal firm development activities.We offer a comprehensive total rewards program including a superior benefits package, wellness programming to support physical, mental, emotional and financial well-being, and in-house immigration support for foreign nationals and international business travelers.Work Location FlexibilityCRA creates a work environment that enables our colleagues to benefit from being together in the office to best deliver on our promise of career growth, mentorship and inclusivity. At the same time, we recognize that individuals realize a range of benefits when working from home periodically. We currently expect that individuals spend at least 3 to 4 days a week working in the office (which may include traveling to another CRA office or to client meetings), with specific days determined in coordination with your practice or team.Our Commitment to Equal Employment OpportunityCharles River Associates is an equal opportunity employer (EOE). All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, age, disability, status as a protected veteran, or any other protected characteristic under applicable law.Salary and other compensationA good-faith estimate of the annual base salary range for this position is $130,000 - $150,000. Stating pay within this range may vary based on factors such as education level, experience, skills, geographic location, market conditions, and other qualifications of the successful candidate. This position may be eligible for additional bonus incentive compensation.CRA offers a comprehensive benefits package, subject to eligibility requirements, which may include: medical, dental, and vision insurance; 401(k) retirement plan with employer match; life and disability insurance; paid time off (vacation, sick leave, holidays); paid parental leave; wellness programs and employee assistance resources; and commuter benefits.
- ...delivers secure, reliable technology solutions... ...Support Engineer, you will help power... ...settlement.Leveraging Site Reliability... ...mission-critical business functions.Your Primary... ...application recovery, failover, disaster recovery (DR), and business continuity activities.Maintain...SuggestedRemote workFlexible hours
$125k - $145k
...lasting impact for our investors, teams, businesses, and the communities in which we... ....The Senior Application Support Engineer ensures stability, performance, and... ...(ServiceNow)Contribute to business continuity and disaster recovery planning and testingQualificationsRequired...SuggestedFull time- ...growth by fostering business development, infrastructure... ...the Director of Engineering Innovation,... ...delivery, and enhance the reliability, scalability, and security... ...capabilities. Drive continuous improvement... ...continuous improvement of disaster recovery, Continuity of...SuggestedFull timeContract workPart timeWork experience placementWork at officeWork from homeMonday to Friday
$145k - $160k
...experienced Cloud Platform Engineer to help design,... ...automation, and reliability across the... ...platforms that support business and application... ...availability, disaster recovery, security, and... ...recovery, and business continuity solutions.... ...vision plans, plus on-site gym access,...SuggestedFull timeSummer workImmediate startRemote workOverseas$134.25k - $214.8k
...where you matter.Your ImpactAre you an engineer who gets excited about the challenge of... ...of the Observability team within Axon's Site Reliability organization — a focused team... ...transferable skills, work experience, business needs, geographic market, and often a combination...SuggestedWork experience placementWork at officeRemote work$140k - $205k
Senior Technology Site Reliability EngineerCooley is seeking... ...Site Reliability Engineer to join the Infrastructure... ..., scaling, and recovery processes to reduce manual... ...post-mortems and continuous improvementDocument operational... ...with all levels of business professionals,...Full timeTemporary workWork at officeFlexible hoursWeekend work- ...where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’... ...by engineers to promote self-service.Continually seek improvement within the entire... ...supplemented at any time in accordance with business needs and conditions.Some roles may...Work at officeRemote workFlexible hours
$160k - $200k
...observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and... ...knowledge & skills, experience, business needs, geographical location, market... ...test as a condition of employment or continued employment. An employer who...Temporary workWork at officeLocal areaFlexible hours3 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range... ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager... ....To drive the personal growth and business impact of our employees, we’re committed...Work at officeLocal areaRemote workWorldwideFlexible hours$93.6k - $170.64k
...opportunity for a NOC AI-Ops Engineer to join our team located in... ...a Senior AIOps and Incident/Site Reliability Engineer to lead incident management... ...design to redefine how businesses run and succeed. Perficient... ...any time. #LI-BV1Incident & Recovery ManagementMonitor, document,...Work at officeLocal areaNight shift3 days per week$138.1k - $198.2k
..., patients, customers, and businesses. We’re making networking easier... ...simply works. The SRE Engineering Enablement Team supports our... ...at Cisco. Your Impact As a Site Reliability Engineer, you will be at... ...units, which vest following continued employment with Cisco for defined...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours- ...in major centers of business, finance, technology,... ...OverviewThe Cloud Solutions Engineer is responsible for... ..., and leads disaster recovery planning. Advanced expertise... ...using Azure Site Recovery, Azure Backup... ...for ongoing business continuity, professional development...Work at office
$127k - $249k
...are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain... ...Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong... ...resilient multi-cloud platform that hosts business critical applications for a wide & varied...Local areaRemote workWorldwideFlexible hours$166k - $220k
...expertise, technology, and business model of the 21st... ...services are reliable and maintainable. This... ...infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability... ...so that our systems continually improve. We actively... ...in health, recovery, and whatever comes...Full timeWork experience placementImmediate start$160k - $200k
Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware...Local areaRemote work$105k - $165k
...being able to help small business owners open their doors faster... ..., and assisting with disaster recovery . The work you do here every... ...Summary: As a Software Engineer III at OpenGov, you'll build... ...technical issues. Drive continuous improvement of development...Full timeTemporary workLocal area$165k - $190k
...RoleWe’re looking for a Senior DevOps Engineer to own and evolve the infrastructure powering... ..., ensuring our platform scales reliably as we onboard enterprise customers with... ...reliability practices: SLOs, capacity planning, disaster recovery, and runbook developmentOperationalize...$119k - $221k
...meet you. We are looking for a Sr. DevOps Engineer IIto help us build and scale the... ...foundation that enables Hi Marley to operate reliably at enterprise scale while deploying... ...drift detection and remediation Improve disaster recovery capabilities: documented and rehearsed...Work at officeLocal areaFlexible hours2 days per week3 days per week$120k - $225k
...Wellington Management you will continually learn, develop, and expand... ...experienced Lead Software Engineer to join a strong,... ..., and back-endPartner with business analysts, QA engineers, and... ...facilitate infrastructure upgrades, disaster recovery testing, and cross-...Full timeRemote workFlexible hours1 day per week$160k - $240k
...passionate about building unified IT solutions that simplify the way IT organizations work. We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the Platform Engineering organization and help us scale our products to millions of end-users. We...Permanent employmentFull timeRemote workWork from homeRelocationFlexible hours- As a Senior Platform Engineer , you’ll be at the heart of our infrastructure... ...are performant, stable, reliable, and secure. You’ll own... ...everything you can, and continuously improve how we deliver,... ..., and improve elasticity, disaster recovery, and high-availability...
- ...trust security. The world’s largest businesses, critical infrastructure organizations... ...cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or... ..., vulnerability management, and continuous monitoring. Proficiency in Linux administration...InternshipWork at officeLocal areaRemote workWorldwide
- ...our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team,... ..., and the implementation and driving of continuous improvement initiatives. You will work... ...someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and...Full time
$130k - $180k
...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and... ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to...Temporary workWork at officeImmediate startRemote workFlexible hours$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust...Work experience placementWork at office$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours- ...Senior Site Reliability Engineer As a Senior Site Reliability Engineer at Blitzy's Cambridge headquarters, you will be the backbone of our platform's reliability, scalability, and operational excellence. You'll work at the intersection of software engineering and infrastructure...
- Principal Solutions Support Engineer -- CRD, Boston, MA, Hybrid, 125k to 165k base, bonusThe... ...liaison with other technology areas and business teams to triage and resolve issues.... ...and on-going refinement, enhancement and continuous process improvement of Charles River based...Local area
$150k - $200k
...DevOps Principal Engineer, you will provide... ...faster and more reliably.This role is ideal... ...DevSecOps culture, continuously improving... ...DevOps solutions with business and product strategy... ...availability, disaster recovery, and business continuity... ....Exposure to Site Reliability...Full time- ...are the leading provider of comprehensive data backup, recovery and business continuity solutions with over five million customers and 8,000 partners... ...500 enterprises to small businesses. Some of our engineers use Mac OS X with Vim, some others use Linux with Sublime...Worldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer - Disaster Recovery & Business Continuity. Be the first to apply!
- site reliability engineer Boston, MA
- site reliability engineer sre Boston, MA
- site services specialist Boston, MA
- construction site safety Boston, MA
- site leader Boston, MA
- official site Boston, MA
- website content developer Boston, MA
- on site coordinator Boston, MA
- IT site lead Boston, MA
- site safety Boston, MA


