Senior Site Reliability Engineer II
$104.9k - $174.7kLexisNexis Risk Solutions Group
About the Role:The SRE role is responsible for improving the reliability, availability, performance, and operational quality of production systems. This role provides technical input into project plans, schedules, methodologies, and operational strategies across multiple system environments.Job FunctionsLead and participate in incident response, postmortems, root-cause analysis, and gap assessments.Identify reliability, availability, performance, security, and operational risks across production environments.Develop, prioritize, and track corrective and preventive actions through completion.Follow up with engineering, development, security, support, and business stakeholders to ensure timely resolution of incidents and identified gaps.Respond to system-management alerts and operational exceptions within assigned enterprise systems and product offerings.Provide technical input into project plans, schedules, implementation methodologies, and operational readiness activities.Support the triage, planning, execution, documentation, and closure of changes, service requests, and operational tasks.Lead or contribute to Operations Team projects involving cloud, on-premises infrastructure, security, Kubernetes, automation, monitoring, and system modernization.Improve production quality and availability by creating new operational capabilities and remediating weaknesses in existing systems and processes.Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering, DevOps, Infrastructure Engineering, or a related field.Bachelor’s degree in Engineering, Computer Science, Information Technology, or equivalent professional experience.Demonstrated experience supporting highly available production systems.Experience leading incident reviews, postmortems, root-cause analysis, and remediation planning.Experience working across infrastructure, application, security, and operations teams.Strong problem-solving, analytical, organizational, and communication skills.Ability to manage multiple priorities and drive work to completion in a fast-paced operational environment.Technical SkillsStrong experience in Site Reliability Engineering (SRE),production operations, and IT service management processes including incident, problem, change, and service request management.Hands-on expertise with cloud and on-premises infrastructure, Kubernetes, containerized workloads, virtualization, and distributed systems.Advanced knowledge of Linux/UNIX and Windows environments, storage and file systems, including installation, configuration, troubleshooting, lifecycle management, backup, disaster recovery, and business continuity.Experience with monitoring, alerting, logging, observability, and performance analysis, including the ability to analyze system diagnostics, logs, traces, resource utilization, and operational metrics.Strong automation and infrastructure engineering skills, including Infrastructure as Code (IaC), configuration management, scripting (Python, Shell, PowerShell), system provisioning, deployments, remediation, and security risk mitigation.AccountabilitiesMonitor assigned environments, respond to alerts and incidents, diagnose system and performance issues, and coordinate escalation and recovery.Track remediation activities and stakeholder commitments through completion to improve production quality, reliability, and availability.Design and maintain automation, scripts, integrations, runbooks, and workflows for provisioning, health checks, deployments, remediation, and routine operations.Install, configure, troubleshoot, and support hardware, software, storage, network, cloud, Kubernetes, and other infrastructure services.Establish logging, monitoring, alerting, metrics, and tracing standards; improve alert quality by reducing noise and ensuring alerts are actionable.Build dashboards and visualizations that communicate system health, availability, performance, capacity, service-level objectives, and incident trends.Develop and maintain recovery procedures and participate in disaster-recovery, resilience, and business-continuity exercises.Partner with development, operations, security, support teams, vendors, and stakeholders to coordinate work, resolve issues, and meet delivery commitments.Lead or contribute to Operations Team projects from planning and implementation through documentation, transition to support, and closure.Plan, risk-assess, obtain approval for, implement, document, and close changes, service requests, and operational tasks.Review and improve technical procedures, scripts, automation, and operational documentation while providing guidance to less-experienced team members.Working for you:We know that your wellbeing and happiness are key to a long and successful career. These are some of the benefits we are delighted to offer:Health Benefits: Comprehensive, multi-carrier program for medical, dental and vision benefitsRetirement Benefits: 401(k) with match and an Employee Share Purchase PlanWellbeing: Wellness platform with incentives, Headspace app subscription, Employee Assistance and Time-off ProgramsShort-and-Long Term Disability, Life and Accidental Death Insurance, Critical Illness, and Hospital IndemnityFamily Benefits, including bonding and family care leaves, adoption and surrogacy benefitsHealth Savings, Health Care, Dependent Care and Commuter Spending AccountsIn addition to annual Paid Time Off, we offer up to two days of paid leave each to participate in Employee Resource Groups and to volunteer with your charity of choice U.S. National Base Pay Range: $104,900 - $174,700. Geographic differentials may apply in some locations to better reflect local market rates. This job is eligible for an annual incentive bonus. We know your well-being and happiness are key to a long and successful career. We are delighted to offer country specific benefits. Click here to access benefits specific to your location.We are committed to providing a fair and accessible hiring process. If you have a disability or other need that requires accommodation or adjustment, please let us know by completing our Applicant Request Support Form or please contact View phone number on click.appcast.io.Criminals may pose as recruiters asking for money or personal information. We never request money or banking details from job applicants. Learn more about spotting and avoiding scams here.Please read our Candidate Privacy Policy.We are an equal opportunity employer: qualified applicants are considered for and treated during employment without regard to race, color, creed, religion, sex, national origin, citizenship status, disability status, protected veteran status, age, marital status, sexual orientation, gender identity, genetic information, or any other characteristic protected by law.USA Job Seekers:EEO Know Your Rights.SummaryLocation: San Jose, CA; San Jose, CA (Santa Clara)Type: Full time
$104.9k - $174.7k
...SRE role is responsible for improving the reliability, availability, performance, and... ...actions through completion.Follow up with engineering, development, security, support, and business... ...Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering...SeniorFull timeLocal area- ...Job Description Job Description Site Reliability Engineer II Bay Area, offices in San Jose · Hybrid · 24/7 FedRAMP Operations · Rotational... ...GovCloud experience on your CV, and a real stepping stone toward Senior SRE. · Real support. From a delivery partner...SuggestedHourly payContract workFor contractorsShift workNight shiftWeekend work
- ...world running. Location: 5 On-Site Days a Week in Sunnyvale, CA Headquarters Our Engineering team is driven by a culture... ...Your Impact As an SRE Engineer II, you will be responsible for... ...will work on enhancing system reliability and scalability of Illumio SaaS...SuggestedWork experience placementImmediate start
$81.5k - $141.3k
...Site Reliability Engineer II Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices...SuggestedRemote work$132.6k - $214.5k
...you will collaborate closely with our engineering teams to develop innovative solutions that... ...’ performance and health. As a Senior Staff SRE with the Cortex Observability... ...operability of the product and ensure the reliability and availability of our services....SeniorFull timeWork at officeVisa sponsorshipWork visa$150.4k - $277.6k
...Services The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples of Apple’s long... ...field with 4+ years experience At least 6 years in a Reliability Engineering, DevOps or infrastructure focused role Advanced...SeniorRelocationDay shift- ...package to enable holistic well-being for you and your family.What you'll doYou will design, build, and operate scalable, secure, and reliable systems supporting SaaS applications and platforms. You will work in Agile, cross-functional teams alongside product managers,...SeniorApprenticeshipWork experience placement
- ...the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines... ...the team for you. Your Impact You will be the most senior technical individual contributor on the team — setting the...Senior
$182k - $242k
...regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible...SeniorPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$212.5k - $250k
...every step of the software delivery lifecycle so Carta's engineers can do the best work of their careers. We own CI/CD infrastructure... ...get to write it. The Problems You’ll Solve As a Senior Software Engineer II on DevEx, you will: Build the MCP servers, agents, and...SeniorFull time- ...What You Will Contribute:### ## Software Engineer II (Full Stack) Are you excited about... ...accelerate development while ensuring quality, reliability, and maintainability* Collaborate with... ...environment with at least three days per week on-site in Los Gatos, CA* Receive competitive...Temporary workLocal areaFlexible hours3 days per week
$75k - $150k
...We are seeking a Salesforce Developer II to design, develop, and deliver enterprise... ...in Computer Science, Information Systems, Engineering, or a related field, or equivalent experience... ..., applications, or resumes to this site or to any Columbia Bank employee and any...$226.14k
...massive datasets in real time to building systems that operate reliably on a global scale. When you work here, your impact is... ...meaningful challenges, we'd love to meet you. Job Title: Software Engineer II Location: 50 West San Fernando Street, Suite 1800, San Jose,...Full timeTemporary workWork at officeRemote workWorldwide$115.2k - $172.8k
Software Development Engineer II At F5, we strive to bring a better digital world to life. Our teams empower organizations across the globe to create, secure, and run applications that enhance how we experience our evolving digital world. We are passionate about cybersecurity...Work at officeLocal areaRemote workHome office- Software Engineer II TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider. We are a force... ...Job Responsibilities Design, develop, and deploy scalable and reliable backend services and APIs. Build and maintain responsive and...Remote workMonday to Friday
$262k - $364k
...infrastructure from SRE side, ensuring it is reliable, scalable, cost effective and performant, while working closely with senior technical leads in the development teams.... ...:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software...Senior$136.3k - $199.9k
...code with a strong emphasis on performance, scalability, and reliability across distributed or multi-threaded systems Identify, analyze... ...closely with hardware, systems, and interdisciplinary engineering teams to deliver integrated solutions on complex equipment platforms...Minimum wageTemporary workWork experience placement$152k - $241.5k
...infrastructure for AI workloads. We are looking for Software Engineers with SRE or Production Engineering experience who have worked... ...initial provisioning through repair.Experience managing production reliability through on-call duties, incident response, observability, and...SeniorPermanent employmentFull time$184k - $208k
...About the role Muon is looking for a Senior Software Engineer (Back-end) to join our Ground Software team. The ideal candidate is a self‑motivated... ...a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (green card holder), (iii)...SeniorPermanent employmentFull timeTemporary workRemote work$155k - $165k
...everyday workers. We are seeking a Senior AWS DevOps & Cloud Security Engineer to own and drive the security, production... ...and secure Windows Server and IIS workloads hosted in AWS. Deploy and... ...streamline operations and improve reliability. Participate in on-call rotations,...SeniorWork at officeLocal area$248k - $396.75k
Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused on designing, building, and operating large-scale production systems... ..., analytics, and automated anomaly detection.Partner with senior leaders and engineers across Cloud, Platform, Security,...Full time$135.9k - $178.97k
...Senior Program Performance Auditor The Office of the City Auditor is seeking motivated... ...Auditor or Program Performance Auditor I/II positions to lead or support City Council... ...benefits above, there is an additional perks site to explore further benefits of working for...SeniorFull timeWork at office$151.6k - $245.3k
...Summary Your Career Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture...Full timeWork at office$152k - $241.5k
NVIDIA is seeking an innovative and highly motivated engineer with deep expertise in systems software to join our GPU Software team. In... ...environmentsCollaborate with globally distributed teams to deliver scalable, reliable, and high-impact GPU software solutionsWhat we need to see: BS...SeniorFull time$175k - $265k
...d-Matrix's SRE team owns the infrastructure layer that every engineering team and customer depends on — colocation facilities, on-premises... .... This role is a core member of that team, responsible for reliability, automation, and observability across colo, on-premises lab,...SeniorFull time$230k - $250k
...minds are shaping the future of network reliability, security, and AI‑ready operations. About... ...you will be building the reliability engineering function at Forward — defining how we... ...Looking For ~6+ years of experience in site reliability engineering, DevOps, or...Night shift- ...Job Description Job Description Site Reliability Engineer Foxconn Industrial Internet (Fii), is a world leading professional design and manufacturing service provider of communication network equipment, cloud service equipment, precision tools and industrial robots...Permanent employmentFull timeWork at officeLocal area
- ...Must Have Technical/Functional Skills: 2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a related role supporting cloud-based production environments. Practical vulnerability-management experience; familiarity with...Full timeWorldwide
- ...of Huobi globe spanning infrastructure. • Work with engineering teams to make sure new features and changes are deployed quickly... .... • Constantly improve our system performance and reliability through better tools, process and monitoring system. •...Worldwide
- ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems... ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer II. Be the first to apply!
- site reliability engineer San Jose, CA
- site reliability engineer sre San Jose, CA
- senior manager customer operations San Jose, CA
- senior software engineer ruby on rails San Jose, CA
- sr finance manager San Jose, CA
- sr marketing manager San Jose, CA
- senior customer service San Jose, CA
- senior business manager San Jose, CA
- senior account executive San Jose, CA
- senior accounts receivable analyst San Jose, CA



