Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

$90k - $180k

Abbott

Job ID: 31156290Company: Abbott LaboratoriesLocation: United States - California - SunnyvaleJob Type: Full timeCategory: Research & DevelopmentIndustry: HealthcarePosted Date: 2026-07-30Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices, nutritionals and branded generic medicines. Our 115,000 colleagues serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We are seeking a highly skilled and mission-driven Senior Site Reliability Engineer (SRE) to join our DevOps team. In this critical role, you will be responsible for ensuring the reliability, scalability, performance, and operational excellence of Merlin.net — a remote monitoring platform designed to help doctors, cardiologists, and care teams automatically collect and review data from patients with implanted cardiac devices.This isn't just about keeping servers up; it's about building and maintaining the resilient backbone for systems where failure is not an option, and where our success directly impacts patient care around the world. You will embed within our DevOps team, acting as a bridge between development and operations.What You'll DoDesign, implement, and maintain highly available, fault-tolerant, and resilient systems that meet demanding uptime and safety requirements. Identify and eliminate performance bottlenecks in software and infrastructure, ensuring low-latency, high-throughput, and real-time responsiveness for customer-facing services. Define, monitor, and uphold Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets. Develop and implement comprehensive monitoring, logging, tracing, and alerting solutions to provide deep insights into system health and behavior at scale. Automate away manual operational tasks, from provisioning and deployment to testing and recovery. Develop and implement strategies for scaling our services and infrastructure to meet evolving business demands, including distributed systems and cloud deployments in Azure. Work closely with software engineering, security, quality, and compliance teams to integrate SRE best practices into our operational processes and infrastructure, ensuring the integrity, availability, and confidentiality of our systems that carry sensitive patient health data. Create clear, concise, and comprehensive documentation, runbooks, and playbooks for operational procedures. Lead blameless postmortem processes following incidents and drive systematic follow-through on action items. Work with a multi-disciplinary team on challenging problems in a fast-paced environment, contributing across architecture reviews, incident response, capacity planning, and reliability roadmap planning.Required QualificationsBachelor's in Computer Science, Software Engineering, Systems Engineering, or a related technical discipline; equivalent professional experience will be considered.Minimum 7 years of experience working in site reliability, software Engineering , Systems Engineering or a related technical disciplineExcellent communication skills with the demonstrated ability to work effectively in cross-functional teams, translate technical complexity for non-technical stakeholders, and collaborate with development, quality, security, marketing, and regulatory teams.Strong analytical, problem-solving, and debugging skills with a methodical and structured approach to diagnosing complex, distributed system issues under pressure — including production incidents with patient safety implications.Proficiency in at least one systems or automation programming language (e.g., Python, Go, Bash, PowerShell) for building tooling, automation, and operational systems.Demonstrated expertise with Microsoft Azure — including Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, and related managed services.Container orchestration expertise — hands-on production experience with Kubernetes and Docker at scale, including deployment strategies, resource management, and cluster operations.Observability platform experience with tools such as Prometheus, Grafana, the ELK/EFK stack, Datadog, Azure Monitor, or similar enterprise-grade monitoring and tracing platforms.Experience designing and operating CI/CD pipelines for continuous delivery of software in production environments, including safe deployment strategies such as blue/green, canary, and feature flag-gated rollouts.Deep understanding of distributed systems — including load balancing, service meshes, microservices, message queues, and fault-tolerant design patterns.Solid Linux & networking fundamentals — DNS, TCP/IP, TLS, load balancing, and networking in cloud environments.Incident management experience — including on-call rotation participation, structured incident response, root cause analysis (RCA), and systematic prevention of recurrence.Preferred QualificationsExperience working in a regulated healthcare, medical device, or life sciences environment, with familiarity in compliance frameworks such as HIPAA.Relevant professional certifications such as: Microsoft Certified Azure DevOps Engineer Expert, Azure Solutions Architect Expert, Certified Kubernetes Administrator (CKA), or Certified Kubernetes Security Specialist (CKS).Background in cost optimization for cloud-native architectures, including FinOps practices for Azure environments.Experience contributing to or driving disaster recovery (DR) design, business continuity planning, and tabletop exercises.The base pay for this position is $90,000.00 – $180,000.00. In specific locations, the pay range may vary from the range posted.Skillsazurekubernetespythongobashmicrosoft azure devopsprometheusgrafanaelk stackdatadogazure kubernetes serviceazure monitorazure policyservice level indicatorsservice level objectivescloud deploymentdistributed systemslinuxnetworking fundamentalsroot cause analysisproduction incident responsehealthcare compliancehipaa compliancecloud-native architecturecost optimizationfinopsdisaster recovery design

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Sunnyvale, CA vacancy
  • $174k - $253k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 
    Senior

    Google

    Sunnyvale, CA
    16 hours ago
  • LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is... 
    Senior
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    2 days ago
  • $148k - $235.75k

     ...see how you can make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable, job-centric insights and automation for GPU fleets. We’re hiring a DevOps Engineer... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $101k - $161k

     ...several prestigious awards, such as Best Engineering Team, Best Company for Diversity,...  ...DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s...  ...: EngineeringExperience level: Mid-Senior LevelIndustry: Computer Networking
    Senior

    Arista Networks

    Santa Clara, CA
    16 hours ago
  • $152k - $241.5k

     ...artificial intelligence.We’re looking for a Senior SRE to join our Compute Farm team and...  ...host lifecycle management, fleet reliability/auto-healing, E2E observability or data-...  ...Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $210.6k - $305.1k

     ...Minimum Qualifications:  You have led a distributed team of 5+ engineers, can demonstrate strong technical vision for your team, and ensure...  ..., and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible... 
    Senior
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    Los Altos, CA
    4 days ago
  •  ...A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure. The ideal candidate has over 8 years of experience in wireless network operations and a strong background in wireless... 
    Senior

    TechDigital Group

    Santa Clara, CA
    2 days ago
  • $145k - $165k

     ...A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key... 
    Senior

    Bolt Graphics, Inc.

    Sunnyvale, CA
    2 days ago
  • $262k - $365k

    Lead a team of Software/Systems Engineers on projects for users and be directly responsible...  ...in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and...  ...Software Engineer chose to join SRE.As the Senior Engineering Manager for Collaboration... 
    Senior

    Google

    Sunnyvale, CA
    4 days ago
  • Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with... 
    Senior
    Flexible hours

    Sumo Logic

    San Jose, CA
    16 hours ago
  •  ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and...  ...and networking teams to improve service reliability and deployment workflowsDeploy and...  ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago
  • $140k - $205k

    Senior Technology Site Reliability EngineerCooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operations team.Position summary: The Senior Technology Site Reliability Engineer (“SRE”) is responsible for ensuring the reliability... 
    Senior
    Full time
    Temporary work
    Work at office
    Flexible hours
    Weekend work

    Cooley

    Palo Alto, CA
    2 days ago
  • $159.2k - $301.6k

     ...for running Graphs on the cloud. In this reliability-focused role, you will own the...  ...infrastructure. You'll partner with the backend engineers building these APIs to make sure the system...  ...Science. 5-10 years of experience in site reliability engineering, infrastructure,... 
    Senior
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    3 days ago
  • $262k - $365k

     ...automation, and evolve systems by pushing for changes that improve reliability and velocity.Practice sustainable incident response and...  ...qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems... 
    Senior

    Google

    San Jose, CA
    16 hours ago
  • Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Infrastructure Platforms... 
    Senior

    JP Morgan Chase

    Palo Alto, CA
    2 days ago
  •  ...Platform powers compute provisioning and infrastructure orchestration across our physical data centers. We are looking for a Senior Site Reliability Engineer to improve the reliability, scalability, and operational maturity of these systems as Lambda’s fleet and customer base... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    1 day ago
  • $187.04k - $359.72k

     ...systems by pushing for changes that improve reliability and velocity. Qualifications Minimum...  ...degree in Computer Science, Electrical Engineering, Computer Engineering or related areas....  ...Product Ops, Corporate Functions and more. On-site presence across teams allows the company... 
    Senior
    Temporary work
    Local area
    Overseas
    Shift work

    Tik Tok

    San Jose, CA
    2 days ago
  • $217.57k - $260k

     ...job description explicitly states otherwise, all roles are on-site five days per week at one of our offices in McLean, VA;...  ...which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure... 
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours
    Shift work

    ID.me

    Mountain View, CA
    1 day ago
  •  ...professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of...  ...professionals and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the... 
    Senior

    J.P. Morgan

    Palo Alto, CA
    4 days ago
  •  ...Job Description Job Description About the Role Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform... 
    Senior

    Kody

    Sunnyvale, CA
    20 days ago
  •  ..., and the challenges of building in a high-growth startup, we’d love to talk. This is more than a job—it’s a journey. Site Reliability Engineers (SREs) are responsible for the overall performance and reliability of ASAPP's infrastructure and products. The team owns... 
    Senior
    Remote work

    ASAPP

    Mountain View, CA
    24 days ago
  • $230k - $250k

     ...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change...  ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"... 
    Night shift

    Forward Networks

    Santa Clara, CA
    2 days ago
  • $147k - $211k

     ...product or system development code.Review code developed by other engineers and provide feedback to ensure best practices (e.g., style...  ..., and troubleshooting large-scale distributed systems. Site Reliability Engineering (SRE) is what you get when you treat operations... 

    Google

    Sunnyvale, CA
    16 hours ago
  • $170k - $200k

    We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster, infrastructure, and monitoring systems to ensure high availability, performance,... 
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    4 days ago
  • $184k

     ...best work. Come join the team and see how you can make a lasting impact on the world. We are looking for a dedicated engineer for the Senior Systems Software Engineer role, focusing on GPU Performance at Scale. At NVIDIA, this role is uniquely positioned to drive... 
    Senior
    Full time

    NVIDIA

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

    At NVIDIA, Site Reliability Engineering provides a rare chance to define, develop, and support large-scale production systems with high efficiency and availability. This demanding position merges software and systems engineering efforts to guarantee flawless service operation... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

    NVIDIA is searching for a creative and highly motivated engineer with expertise in systems software to join the GPU Software team. You will design key aspects of our production GPU kernel drivers and embedded SW that impacts our products both in the datacenter and in gaming... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

    Our Autonomous Vehicles Platform team is searching for engineers to develop and bring NVIDIA's automotive platform out to the world. You will participate in a focused effort to develop and productize ground-breaking solutions that will revolutionize the world of transportation... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    16 hours ago
  • $152k - $241.5k

    The Autonomous Vehicles Platform team is seeking a Senior System Software Engineer to help bring NVIDIA's autonomous vehicle platform to new markets! This role involves developing and productizing innovative solutions that will transform transportation and the field of... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    The Autonomous Vehicles Platform team is looking for a hands-on System Software Engineer. As part of our team, you will work on our Autonomous Driving Platform software, implementing performant and scalable solutions for data collection and autonomous vehicle fleets. Our... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    16 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!