Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer II

$104.9k - $174.7k

RELX Group

About the Role:The SRE role is responsible for improving the reliability, availability, performance, and operational quality of production systems. This role provides technical input into project plans, schedules, methodologies, and operational strategies across multiple system environments.Job FunctionsLead and participate in incident response, postmortems, root-cause analysis, and gap assessments.Identify reliability, availability, performance, security, and operational risks across production environments.Develop, prioritize, and track corrective and preventive actions through completion.Follow up with engineering, development, security, support, and business stakeholders to ensure timely resolution of incidents and identified gaps.Respond to system-management alerts and operational exceptions within assigned enterprise systems and product offerings.Provide technical input into project plans, schedules, implementation methodologies, and operational readiness activities.Support the triage, planning, execution, documentation, and closure of changes, service requests, and operational tasks.Lead or contribute to Operations Team projects involving cloud, on-premises infrastructure, security, Kubernetes, automation, monitoring, and system modernization.Improve production quality and availability by creating new operational capabilities and remediating weaknesses in existing systems and processes.Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering, DevOps, Infrastructure Engineering, or a related field.Bachelor’s degree in Engineering, Computer Science, Information Technology, or equivalent professional experience.Demonstrated experience supporting highly available production systems.Experience leading incident reviews, postmortems, root-cause analysis, and remediation planning.Experience working across infrastructure, application, security, and operations teams.Strong problem-solving, analytical, organizational, and communication skills.Ability to manage multiple priorities and drive work to completion in a fast-paced operational environment.Technical SkillsStrong experience in Site Reliability Engineering (SRE),production operations, and IT service management processes including incident, problem, change, and service request management.Hands-on expertise with cloud and on-premises infrastructure, Kubernetes, containerized workloads, virtualization, and distributed systems.Advanced knowledge of Linux/UNIX and Windows environments, storage and file systems, including installation, configuration, troubleshooting, lifecycle management, backup, disaster recovery, and business continuity.Experience with monitoring, alerting, logging, observability, and performance analysis, including the ability to analyze system diagnostics, logs, traces, resource utilization, and operational metrics.Strong automation and infrastructure engineering skills, including Infrastructure as Code (IaC), configuration management, scripting (Python, Shell, PowerShell), system provisioning, deployments, remediation, and security risk mitigation.AccountabilitiesMonitor assigned environments, respond to alerts and incidents, diagnose system and performance issues, and coordinate escalation and recovery.Track remediation activities and stakeholder commitments through completion to improve production quality, reliability, and availability.Design and maintain automation, scripts, integrations, runbooks, and workflows for provisioning, health checks, deployments, remediation, and routine operations.Install, configure, troubleshoot, and support hardware, software, storage, network, cloud, Kubernetes, and other infrastructure services.Establish logging, monitoring, alerting, metrics, and tracing standards; improve alert quality by reducing noise and ensuring alerts are actionable.Build dashboards and visualizations that communicate system health, availability, performance, capacity, service-level objectives, and incident trends.Develop and maintain recovery procedures and participate in disaster-recovery, resilience, and business-continuity exercises.Partner with development, operations, security, support teams, vendors, and stakeholders to coordinate work, resolve issues, and meet delivery commitments.Lead or contribute to Operations Team projects from planning and implementation through documentation, transition to support, and closure.Plan, risk-assess, obtain approval for, implement, document, and close changes, service requests, and operational tasks.Review and improve technical procedures, scripts, automation, and operational documentation while providing guidance to less-experienced team members.Working for you:We know that your wellbeing and happiness are key to a long and successful career. These are some of the benefits we are delighted to offer:Health Benefits: Comprehensive, multi-carrier program for medical, dental and vision benefitsRetirement Benefits: 401(k) with match and an Employee Share Purchase PlanWellbeing: Wellness platform with incentives, Headspace app subscription, Employee Assistance and Time-off ProgramsShort-and-Long Term Disability, Life and Accidental Death Insurance, Critical Illness, and Hospital IndemnityFamily Benefits, including bonding and family care leaves, adoption and surrogacy benefitsHealth Savings, Health Care, Dependent Care and Commuter Spending AccountsIn addition to annual Paid Time Off, we offer up to two days of paid leave each to participate in Employee Resource Groups and to volunteer with your charity of choice U.S. National Base Pay Range: $104,900 - $174,700. Geographic differentials may apply in some locations to better reflect local market rates. This job is eligible for an annual incentive bonus. We know your well-being and happiness are key to a long and successful career. We are delighted to offer country specific benefits. Click here to access benefits specific to your location.We are committed to providing a fair and accessible hiring process. If you have a disability or other need that requires accommodation or adjustment, please let us know by completing our Applicant Request Support Form or please contact View phone number on click.appcast.io.Criminals may pose as recruiters asking for money or personal information. We never request money or banking details from job applicants. Learn more about spotting and avoiding scams here.Please read our Candidate Privacy Policy.We are an equal opportunity employer: qualified applicants are considered for and treated during employment without regard to race, color, creed, religion, sex, national origin, citizenship status, disability status, protected veteran status, age, marital status, sexual orientation, gender identity, genetic information, or any other characteristic protected by law.USA Job Seekers:EEO Know Your Rights.SummaryLocation: San Jose, CA; San Jose, CA (Santa Clara)Type: Full time

Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer II in San Jose, CA vacancy
  •  ...Job Description Job Description Site Reliability Engineer II Bay Area, offices in San Jose · Hybrid · 24/7 FedRAMP Operations · Rotational...  ...GovCloud experience on your CV, and a real stepping stone toward Senior SRE. ·        Real support.  From a delivery partner... 
    Suggested
    Hourly pay
    Contract work
    For contractors
    Shift work
    Night shift
    Weekend work

    C-Serv

    San Jose, CA
    26 days ago
  •  ...world running. Location: 5 On-Site Days a Week in Sunnyvale, CA Headquarters Our Engineering team is driven by a culture...  ...Your Impact As an SRE Engineer II, you will be responsible for...  ...will work on enhancing system reliability and scalability of Illumio SaaS... 
    Suggested
    Work experience placement
    Immediate start

    Illumio

    Sunnyvale, CA
    2 days ago
  •  ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering...  ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    a month ago
  •  ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and...  ...and networking teams to improve service reliability and deployment workflowsDeploy and...  ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    a month ago
  • $168k - $270.25k

     ...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $267k - $356k

     ...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-...  ...workloads in the industry, which means reliability and performance aren't just goals—they're...  ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc... 
    Senior
    Work experience placement
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    a month ago
  •  ...the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines...  ...the team for you. Your Impact You will be the most senior technical individual contributor on the team — setting the... 
    Senior

    Webex Events (formerly Socio)

    San Jose, CA
    1 day ago
  • $150.4k - $277.6k

     ...Senior Site Reliability Engineer The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples of Apple's long-held passion for combining art and technology. These are the people who power the App Store, Apple TV, Apple... 
    Senior
    Relocation
    Day shift

    Apple

    Cupertino, CA
    4 days ago
  •  ...Job Title: Senior Site Reliability Engineer Kubernetes Platform Location: San Jose, CA Full-Time Job Description Must Have Technical/Functional Skills: 10+ years of experience in SRE, DevOps, or infrastructure engineering Strong experience... 
    Senior
    Full time

    SFE

    San Jose, CA
    2 days ago
  •  ...Job Title: Mid-Senior Site Reliability Engineer Kubernetes Platform Location: San Jose, CA Full-Time Job Description Must Have Technical/Functional Skills: 8+ years of experience in SRE, DevOps, or platform engineering Hands-on experience... 
    Senior
    Full time

    SFE

    San Jose, CA
    2 days ago
  • $207k - $300k

    Manage a team of Software/Systems Engineers on projects for users and remain directly responsible...  ..., and establishing sustainable multi-site on-call rotations across distributed...  ...global hubs.Deep practical expertise in Site Reliability Engineering practices, including SLO/SLI... 

    Google

    San Jose, CA
    3 days ago
  • $182k - $242k

     ...) in March 2025. Learn more at  About the Role As a Senior Software Engineer II (IC4) on the AI Workload Orchestration Platform team, you...  ...will own meaningful components of the platform, drive reliability and performance improvements, and help scale the system as... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    26 days ago
  •  ..., GPU infrastructure, and Linux systems engineering. We partner closely with security, platform...  ...and resolving complex performance, reliability, or isolation issues across containers,...  ...defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    5 days ago
  • $182k - $242k

     ...regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    26 days ago
  • $133.92k

     ...motivated and skilled Software Development Engineer II to join our team. This role offers an...  ...Proactively collaborate with peers and senior engineers to solve technical challenges...  ...to the improvement of the performance, reliability, and success metrics of product... 
    Full time
    Internship
    Summer internship
    Local area
    Remote work

    F5

    San Jose, CA
    16 hours ago
  • $207k - $300k

     ....Drive incident response, maintain high reliability standards, and actively automate operational...  ...of experience managing and growing engineering teams, including performance management...  ...managing distributed teams across multiple sites or timezones.Experience developing long-... 

    Google

    San Jose, CA
    3 days ago
  •  ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems...  ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes... 
    Senior

    Saviynt

    Milpitas, CA
    a month ago
  • $212.5k - $250k

     ...every step of the software delivery lifecycle so Carta's engineers can do the best work of their careers. We own CI/CD infrastructure...  ...get to write it. The Problems You’ll Solve As a Senior Software Engineer II on DevEx, you will: Build the MCP servers, agents, and... 
    Senior
    Full time

    Carta

    Santa Clara, CA
    a month ago
  •  ...planet.What You Will Contribute:Software Engineer II (Full Stack)Are you excited about...  ...accelerate development while ensuring quality, reliability, and maintainabilityCollaborate with...  ...environment with at least three days per week on-site in Los Gatos, CAReceive competitive... 
    Full time
    Temporary work
    Local area
    Flexible hours
    3 days per week

    Badger Meter

    Los Gatos, CA
    3 days ago
  • $165.2k - $223.6k

     ...their information.As a Software Development Engineer II, you'll take ownership of production...  ...architectural strategies, seeking guidance from senior engineers when facing complex technical...  ...you investigate root causes to maintain reliable data deletion and opt-out mechanisms.... 
    Permanent employment
    Internship
    Local area
    Worldwide
    Flexible hours

    AmazonWebServices

    Santa Clara, CA
    a month ago
  •  ...software tasks of small to medium complexity. Working on a variety of technical problems of varying scope, the Software Development Engineer II implements high-quality code and deploys features or APIs under general supervision. Collaborates closely with team peers to... 
    Full time
    Work at office
    Local area
    Remote work
    Home office

    F5 Networks

    San Jose, CA
    13 days ago
  • $226.14k

     ...massive datasets in real time to building systems that operate reliably on a global scale. When you work here, your impact is...  ...meaningful challenges, we’d love to meet you. Job Title: Software Engineer II Location: 50 West San Fernando Street, Suite 1800, San Jose... 
    Full time
    Temporary work
    Work at office
    Remote work
    Worldwide

    The Trade Desk

    San Jose, CA
    5 days ago
  • $125.7k - $203.1k

    Software Engineer Embedded Systems II Join a vibrant community of passionate professionals from around...  ...platform interactions and ensure product reliability. Apply systems programming expertise...  .... Please see the Cisco careers site to discover more benefits and perks.... 
    Full time
    Temporary work
    Apprenticeship
    Work experience placement
    Local area
    Flexible hours

    Webex Events (formerly Socio)

    Milpitas, CA
    2 days ago
  • $75k - $150k

     ...We are seeking a Salesforce Developer II to design, develop, and deliver enterprise...  ...in Computer Science, Information Systems, Engineering, or a related field, or equivalent experience...  ..., applications, or resumes to this site or to any Columbia Bank employee and any... 

    Columbia Bank

    San Jose, CA
    16 hours ago
  • $87.8k - $131.7k

    Application Developer I/II/IIISalary Range: $87,798 - $131,697The expected pay range is based on many factors, such as experience, education, and the market. The range is subject to change.This posting is for one position and will be filled as an Application Developer... 
    Work at office
    Local area

    Santa Clara Family Health Plan

    San Jose, CA
    a month ago
  • $185k - $215k

     ...About the role Muon seeks a Senior Software Engineer to join our Information Systems team. The ideal candidate is a self-motivated and versatile...  ...a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (green card holder), (iii)... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Remote work
    Flexible hours

    Muon Space

    San Jose, CA
    more than 2 months ago
  • Software Engineer II TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider. We are a force...  ...Job Responsibilities Design, develop, and deploy scalable and reliable backend services and APIs. Build and maintain responsive and... 
    Remote work
    Monday to Friday

    TenEx

    San Jose, CA
    2 days ago
  • $184k - $208k

     ...About the role Muon seeks a Senior Software Engineer to join our Software Team. The ideal candidate is a brilliant generalist problem-solver...  ...a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (green card holder), (iii)... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Remote work
    Flexible hours

    Muon Space

    San Jose, CA
    more than 2 months ago
  • $105.5k - $213.5k

    Software Engineer II This role has been designed as 'Onsite' with an expectation that you will primarily work from an HPE office. Who We...  ...Collaborate with internal and external teams to deliver high-quality, reliable, and cost-effective software solutions. Design, develop and... 
    Work experience placement
    Work at office
    Local area
    Immediate start

    Hewlett Packard Enterprise

    Cupertino, CA
    2 days ago
  • $184k - $208k

     ...About the role Muon is looking for a Senior Software Engineer (Back-end) to join our Ground Software team. The ideal candidate is a self‑motivated...  ...a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (green card holder), (iii)... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Remote work

    muonspace

    San Jose, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer II. Be the first to apply!