Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer II

$104.9k - $174.7k

LexisNexis Risk Solutions Group

About the Role:The SRE role is responsible for improving the reliability, availability, performance, and operational quality of production systems. This role provides technical input into project plans, schedules, methodologies, and operational strategies across multiple system environments.Job FunctionsLead and participate in incident response, postmortems, root-cause analysis, and gap assessments.Identify reliability, availability, performance, security, and operational risks across production environments.Develop, prioritize, and track corrective and preventive actions through completion.Follow up with engineering, development, security, support, and business stakeholders to ensure timely resolution of incidents and identified gaps.Respond to system-management alerts and operational exceptions within assigned enterprise systems and product offerings.Provide technical input into project plans, schedules, implementation methodologies, and operational readiness activities.Support the triage, planning, execution, documentation, and closure of changes, service requests, and operational tasks.Lead or contribute to Operations Team projects involving cloud, on-premises infrastructure, security, Kubernetes, automation, monitoring, and system modernization.Improve production quality and availability by creating new operational capabilities and remediating weaknesses in existing systems and processes.Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering, DevOps, Infrastructure Engineering, or a related field.Bachelor’s degree in Engineering, Computer Science, Information Technology, or equivalent professional experience.Demonstrated experience supporting highly available production systems.Experience leading incident reviews, postmortems, root-cause analysis, and remediation planning.Experience working across infrastructure, application, security, and operations teams.Strong problem-solving, analytical, organizational, and communication skills.Ability to manage multiple priorities and drive work to completion in a fast-paced operational environment.Technical SkillsStrong experience in Site Reliability Engineering (SRE),production operations, and IT service management processes including incident, problem, change, and service request management.Hands-on expertise with cloud and on-premises infrastructure, Kubernetes, containerized workloads, virtualization, and distributed systems.Advanced knowledge of Linux/UNIX and Windows environments, storage and file systems, including installation, configuration, troubleshooting, lifecycle management, backup, disaster recovery, and business continuity.Experience with monitoring, alerting, logging, observability, and performance analysis, including the ability to analyze system diagnostics, logs, traces, resource utilization, and operational metrics.Strong automation and infrastructure engineering skills, including Infrastructure as Code (IaC), configuration management, scripting (Python, Shell, PowerShell), system provisioning, deployments, remediation, and security risk mitigation.AccountabilitiesMonitor assigned environments, respond to alerts and incidents, diagnose system and performance issues, and coordinate escalation and recovery.Track remediation activities and stakeholder commitments through completion to improve production quality, reliability, and availability.Design and maintain automation, scripts, integrations, runbooks, and workflows for provisioning, health checks, deployments, remediation, and routine operations.Install, configure, troubleshoot, and support hardware, software, storage, network, cloud, Kubernetes, and other infrastructure services.Establish logging, monitoring, alerting, metrics, and tracing standards; improve alert quality by reducing noise and ensuring alerts are actionable.Build dashboards and visualizations that communicate system health, availability, performance, capacity, service-level objectives, and incident trends.Develop and maintain recovery procedures and participate in disaster-recovery, resilience, and business-continuity exercises.Partner with development, operations, security, support teams, vendors, and stakeholders to coordinate work, resolve issues, and meet delivery commitments.Lead or contribute to Operations Team projects from planning and implementation through documentation, transition to support, and closure.Plan, risk-assess, obtain approval for, implement, document, and close changes, service requests, and operational tasks.Review and improve technical procedures, scripts, automation, and operational documentation while providing guidance to less-experienced team members.Working for you:We know that your wellbeing and happiness are key to a long and successful career. These are some of the benefits we are delighted to offer:Health Benefits: Comprehensive, multi-carrier program for medical, dental and vision benefitsRetirement Benefits: 401(k) with match and an Employee Share Purchase PlanWellbeing: Wellness platform with incentives, Headspace app subscription, Employee Assistance and Time-off ProgramsShort-and-Long Term Disability, Life and Accidental Death Insurance, Critical Illness, and Hospital IndemnityFamily Benefits, including bonding and family care leaves, adoption and surrogacy benefitsHealth Savings, Health Care, Dependent Care and Commuter Spending AccountsIn addition to annual Paid Time Off, we offer up to two days of paid leave each to participate in Employee Resource Groups and to volunteer with your charity of choice U.S. National Base Pay Range: $104,900 - $174,700. Geographic differentials may apply in some locations to better reflect local market rates. This job is eligible for an annual incentive bonus. We know your well-being and happiness are key to a long and successful career. We are delighted to offer country specific benefits. Click here to access benefits specific to your location.We are committed to providing a fair and accessible hiring process. If you have a disability or other need that requires accommodation or adjustment, please let us know by completing our Applicant Request Support Form or please contact View phone number on click.appcast.io.Criminals may pose as recruiters asking for money or personal information. We never request money or banking details from job applicants. Learn more about spotting and avoiding scams here.Please read our Candidate Privacy Policy.We are an equal opportunity employer: qualified applicants are considered for and treated during employment without regard to race, color, creed, religion, sex, national origin, citizenship status, disability status, protected veteran status, age, marital status, sexual orientation, gender identity, genetic information, or any other characteristic protected by law.USA Job Seekers:EEO Know Your Rights.SummaryLocation: San Jose, CA; San Jose, CA (Santa Clara)Type: Full time

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer II in San Jose, CA vacancy
  • $104.9k - $174.7k

     ...SRE role is responsible for improving the reliability, availability, performance, and...  ...actions through completion.Follow up with engineering, development, security, support, and business...  ...Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering... 
    Senior
    Full time
    Local area

    RELX Group

    San Jose, CA
    2 days ago
  •  ...Job Description Job Description Site Reliability Engineer II Bay Area, offices in San Jose · Hybrid · 24/7 FedRAMP Operations · Rotational...  ...GovCloud experience on your CV, and a real stepping stone toward Senior SRE. ·        Real support.  From a delivery partner... 
    Suggested
    Hourly pay
    Contract work
    For contractors
    Shift work
    Night shift
    Weekend work

    C-Serv

    San Jose, CA
    28 days ago
  •  ...world running. Location: 5 On-Site Days a Week in Sunnyvale, CA Headquarters Our Engineering team is driven by a culture...  ...Your Impact As an SRE Engineer II, you will be responsible for...  ...will work on enhancing system reliability and scalability of Illumio SaaS... 
    Suggested
    Work experience placement
    Immediate start

    Illumio

    Sunnyvale, CA
    4 days ago
  • $81.5k - $141.3k

     ...Site Reliability Engineer II Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices... 
    Suggested
    Remote work

    Abbott

    Sunnyvale, CA
    2 days ago
  • $132.6k - $214.5k

     ...you will collaborate closely with our engineering teams to develop innovative solutions that...  ...’ performance and health. As a Senior Staff SRE with the Cortex Observability...  ...operability of the product and ensure the reliability and availability of our services.... 
    Senior
    Full time
    Work at office
    Visa sponsorship
    Work visa

    Palo Alto Networks

    Santa Clara, CA
    2 days ago
  • $150.4k - $277.6k

     ...Services The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples of Apple’s long...  ...field with 4+ years experience At least 6 years in a Reliability Engineering, DevOps or infrastructure focused role Advanced... 
    Senior
    Relocation
    Day shift

    Apple

    Cupertino, CA
    2 days ago
  •  ...package to enable holistic well-being for you and your family.What you'll doYou will design, build, and operate scalable, secure, and reliable systems supporting SaaS applications and platforms. You will work in Agile, cross-functional teams alongside product managers,... 
    Senior
    Apprenticeship
    Work experience placement

    McKinsey & Company

    San Jose, CA
    8 hours ago
  •  ...the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines...  ...the team for you. Your Impact You will be the most senior technical individual contributor on the team — setting the... 
    Senior

    Webex Events (formerly Socio)

    San Jose, CA
    3 days ago
  • $182k - $242k

     ...regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    28 days ago
  • $212.5k - $250k

     ...every step of the software delivery lifecycle so Carta's engineers can do the best work of their careers. We own CI/CD infrastructure...  ...get to write it. The Problems You’ll Solve As a Senior Software Engineer II on DevEx, you will: Build the MCP servers, agents, and... 
    Senior
    Full time

    Carta

    Santa Clara, CA
    a month ago
  •  ...What You Will Contribute:### ## Software Engineer II (Full Stack) Are you excited about...  ...accelerate development while ensuring quality, reliability, and maintainability* Collaborate with...  ...environment with at least three days per week on-site in Los Gatos, CA* Receive competitive... 
    Temporary work
    Local area
    Flexible hours
    3 days per week

    Badger Meter

    Los Gatos, CA
    1 day ago
  • $75k - $150k

     ...We are seeking a Salesforce Developer II to design, develop, and deliver enterprise...  ...in Computer Science, Information Systems, Engineering, or a related field, or equivalent experience...  ..., applications, or resumes to this site or to any Columbia Bank employee and any... 

    Columbia Bank

    San Jose, CA
    2 days ago
  • $226.14k

     ...massive datasets in real time to building systems that operate reliably on a global scale. When you work here, your impact is...  ...meaningful challenges, we'd love to meet you. Job Title: Software Engineer II Location: 50 West San Fernando Street, Suite 1800, San Jose,... 
    Full time
    Temporary work
    Work at office
    Remote work
    Worldwide

    The Trade Desk

    San Jose, CA
    4 days ago
  • $115.2k - $172.8k

    Software Development Engineer II At F5, we strive to bring a better digital world to life. Our teams empower organizations across the globe to create, secure, and run applications that enhance how we experience our evolving digital world. We are passionate about cybersecurity... 
    Work at office
    Local area
    Remote work
    Home office

    F5

    San Jose, CA
    4 days ago
  • Software Engineer II TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider. We are a force...  ...Job Responsibilities Design, develop, and deploy scalable and reliable backend services and APIs. Build and maintain responsive and... 
    Remote work
    Monday to Friday

    TenEx

    San Jose, CA
    4 days ago
  • $262k - $364k

     ...infrastructure from SRE side, ensuring it is reliable, scalable, cost effective and performant, while working closely with senior technical leads in the development teams....  ...:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software... 
    Senior

    Google

    Sunnyvale, CA
    8 hours ago
  • $136.3k - $199.9k

     ...code with a strong emphasis on performance, scalability, and reliability across distributed or multi-threaded systems Identify, analyze...  ...closely with hardware, systems, and interdisciplinary engineering teams to deliver integrated solutions on complex equipment platforms... 
    Minimum wage
    Temporary work
    Work experience placement

    KLA

    Milpitas, CA
    4 hours ago
  • $152k - $241.5k

     ...infrastructure for AI workloads. We are looking for Software Engineers with SRE or Production Engineering experience who have worked...  ...initial provisioning through repair.Experience managing production reliability through on-call duties, incident response, observability, and... 
    Senior
    Permanent employment
    Full time

    Nvidia

    Santa Clara, CA
    8 hours ago
  • $184k - $208k

     ...About the role Muon is looking for a Senior Software Engineer (Back-end) to join our Ground Software team. The ideal candidate is a self‑motivated...  ...a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (green card holder), (iii)... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Remote work

    muonspace

    San Jose, CA
    5 days ago
  • $155k - $165k

     ...everyday workers. We are seeking a Senior AWS DevOps & Cloud Security Engineer to own and drive the security, production...  ...and secure Windows Server and IIS workloads hosted in AWS. Deploy and...  ...streamline operations and improve reliability. Participate in on-call rotations,... 
    Senior
    Work at office
    Local area

    Jobot

    Milpitas, CA
    2 days ago
  • $248k - $396.75k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused on designing, building, and operating large-scale production systems...  ..., analytics, and automated anomaly detection.Partner with senior leaders and engineers across Cloud, Platform, Security,... 
    Full time

    NVIDIA

    Santa Clara, CA
    2 days ago
  • $135.9k - $178.97k

     ...Senior Program Performance Auditor The Office of the City Auditor is seeking motivated...  ...Auditor or Program Performance Auditor I/II positions to lead or support City Council...  ...benefits above, there is an additional perks site to explore further benefits of working for... 
    Senior
    Full time
    Work at office

    City of San Jos

    San Jose, CA
    2 days ago
  • $151.6k - $245.3k

     ...Summary Your Career Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture... 
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

    NVIDIA is seeking an innovative and highly motivated engineer with deep expertise in systems software to join our GPU Software team. In...  ...environmentsCollaborate with globally distributed teams to deliver scalable, reliable, and high-impact GPU software solutionsWhat we need to see: BS... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    8 hours ago
  • $175k - $265k

     ...d-Matrix's SRE team owns the infrastructure layer that every engineering team and customer depends on — colocation facilities, on-premises...  .... This role is a core member of that team, responsible for reliability, automation, and observability across colo, on-premises lab,... 
    Senior
    Full time

    d-Matrix

    Santa Clara, CA
    2 days ago
  • $230k - $250k

     ...minds are shaping the future of network reliability, security, and AI‑ready operations. About...  ...you will be building the reliability engineering function at Forward — defining how we...  ...Looking For ~6+ years of experience in site reliability engineering, DevOps, or... 
    Night shift

    Forward

    Santa Clara, CA
    1 day ago
  •  ...Job Description Job Description Site Reliability Engineer Foxconn Industrial Internet (Fii), is a world leading professional design and manufacturing service provider of communication network equipment, cloud service equipment, precision tools and industrial robots... 
    Permanent employment
    Full time
    Work at office
    Local area

    Foxconn Industrial Internet - FII

    San Jose, CA
    a month ago
  •  ...Must Have Technical/Functional Skills: 2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a related role supporting cloud-based production environments. Practical vulnerability-management experience; familiarity with... 
    Full time
    Worldwide

    SFE

    San Jose, CA
    2 days ago
  •  ...of Huobi globe spanning infrastructure. •       Work with engineering teams to make sure new features and changes are deployed quickly...  .... •       Constantly improve our system performance and reliability through better tools, process and monitoring system. •... 
    Worldwide

    Cryptoware Technologies Inc

    Santa Clara, CA
    a month ago
  •  ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems...  ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes... 
    Senior

    Saviynt

    Milpitas, CA
    23 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer II. Be the first to apply!