Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

$82.3k - $228.8k

GrabJobs

Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. We are seeking a highly experienced Senior Site Reliability Engineer – Compute Platforms to design, implement, and support Kubernetes on baremetal and hypervisor platforms in a private cloud environment. This role is responsible for the architecture, design, and standardization of enterprise compute and hypervisor environments spanning bare metal infrastructure, operating systems, hypervisors, private cloud orchestration, and Kubernetes using Infrastructure-as-Code and GitOps practices. This is a deeply technical role requiring expert-level understanding of compute hardware management, Kubernetes, OpenStack, hypervisors and extensive working knowledge on Linux Operating systems. You will also collaborate with platform and SRE teams to maintain secure, performant, and multi-tenant-isolated services that serve high-throughput, mission-critical applications. Key Responsibilities Lead the architecture and design of enterprise compute and hypervisor platform solutions across hardware, OS, virtualization, cloud orchestration, and container orchestration layers Define standards and automation frameworks for bare metal provisioning and lifecycle management Design and implement Bare Metal as a Service (BMaaS) capabilities for scalable infrastructure consumption Architect and design Kubernetes platforms on bare metal with QoS and Affinity (ArgoCD) Architect and validate automated deployments of operating systems and hypervisors including Ubuntu and Harvester Design and maintain PXE-based provisioning environments leveraging Redfish APIs for large-scale server deployments Develop Infrastructure-as-Code using Ansible, Terraform, Helm and Git, with Python/Bash automation. Implement CI/CD pipelines for infrastructure updates, patching, upgrades, testing, and rollback. Design automated workflows for server build, firmware lifecycle management, patching, and hardware validation Evaluate and standardize enterprise hardware platforms to meet performance, scalability, and reliability requirements Produce detailed high-level and low-level design documentation , build guides, and operational handoff materials Perform deep troubleshooting across storage, Kubernetes, hypervisors, networking, and Linux systems Partner with operations, network, storage, and platform teams to ensure designs are supportable and production-ready Participate in on-call escalation support for complex platform-related issues Collaborate globally on change management , documentation, and operational best practices Minimum Qualifications 6 + years of experience in infrastructure engineering, platform engineering, or DevOps with a strong focus on Compute system design Proven experience designing and automating bare metal compute environments at scale Strong hands-on experience with PXE boot, network-based OS provisioning, and automated server imaging Experience implementing or supporting Bare Metal as a Service (BMaaS) platforms Practical experience using Redfish APIs for hardware provisioning, power management, and remote lifecycle operations Deep expertise with Ubuntu Linux in enterprise environments Strong Hands-on experience with KVM hypervisors (Suse Harvester, OpenStack). Experience designing and deploying production-grade Kubernetes clusters Strong background with enterprise compute hardware platforms , including Cisco UCS, Dell PowerEdge, Supermicro systems & HPE Proficiency with Infrastructure as Code tools (e.g., Terraform, Ansible, or similar) Experience building or supporting CI/CD pipelines for infrastructure and platform automation Strong scripting skills in Python, Bash, or similar languages Demonstrated ability to produce clear, structured technical design documentation Excellent written and verbal communication skills Bachelor’s degree in computer science or equivalent professional experience Preferred Qualifications OpenStack, Ubuntu KVM administration. BareMetal as a Service (PXE, Redfish). Kubernetes on BareMetal CIS/NIST security and infrastructure lifecycle management. ITIL Foundation/advanced certifications in support of ITSM standard methodology. Background in telco, edge cloud, or large enterprise environments. Ubuntu Certifications, CNCF Certified Kubernetes Administrator (CKA), Certified Kubernetes Security Specialist (CKS) Master’s degree in computer science, IT, Engineering, or a related field preferred; equivalent experience and relevant industry certifications will also be considered What You’ll Get A collaborative team that’s deeply invested in infrastructure excellence. Complex technical challenges that require creative, scalable solutions. The opportunity to shape a next-generation private cloud platform-built reliability Access to the latest tools, frameworks, and upstream project developments Skills and Attributes: Analytical Thinking & Problem Solving: Demonstrated ability to translate complex, cross-domain requirements into scalable and resilient cloud infrastructure and automation solutions Collaboration & Teamwork: Strong interpersonal and communication skills with a proven track record of effective collaboration across multidisciplinary teams, including developers, operations, security, and product stakeholders Mentorship & Leadership: Passionate about knowledge-sharing and mentorship, with experience guiding junior engineers and fostering a team culture of continuous learning, innovation, and technical excellence in cloud engineering and DevOps practices Work Location: This role is fully remote for candidates who reside outside the 50 mile radius of our San Ramon office. For candidates who reside within 50 miles of our San Ramon location, this role is Hybrid and would require 3 days a week (M, W, TH) in our San Ramon office. As part of our continued commitment to diversity, equity, and inclusion, Five9 supports pay transparency during the entire recruitment process. Actual compensation packages are based on several factors that are unique to each candidate including, but not limited to: skill set, depth of experience, certifications, and specific work location. The range displayed reflects the minimum and maximum target for new hire salaries for the job across the United States. Your recruiter can share more about the specific compensation package during your hiring process. Additionally, the total compensation package for this position may also include an annual performance bonus, stock, and/or other applicable incentive compensation plans. Our total reward package also includes: Health, dental, and vision coverage, beginning on the first day of employment. Five9 covers 100% of the employee portion of the health, dental and vision coverage and shares a high portion of the dependent cost. We also offer Short & Long-Term Disability, Basic Life Insurance, and a 401k saving plan with employer matching. Access to an innovative mental health support platform that offers personalized care and resources in areas such as: therapy, coaching and self-guided mindfulness exercises for all covered employees and their covered dependents. Generous employee stock purchase plan. Paid Time Off, Company paid holidays, paid volunteer hours and 12 weeks paid parental leave. All compensation and benefits are subject to the requirements and restrictions set forth in the applicable plan documents and any written agreements between the parties. The US base salary range for this role is below. $82,300 - $228,800 USD Five9 embraces diversity and is committed to building a team that represents a variety of backgrounds, perspectives, and skills.  The more inclusive we are, the better we are.  Five9 is an equal opportunity employer. View our privacy policy, including our privacy notice to California residents here: . Note: Five9 will never request that an applicant send money as a prerequisite for commencing employment with Five9.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in San Jose, CA vacancy
  • Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with... 
    Senior
    Flexible hours

    Sumo Logic

    San Jose, CA
    2 days ago
  •  ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and...  ...and networking teams to improve service reliability and deployment workflowsDeploy and...  ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    1 day ago
  • $168k - $270.25k

    NVIDIA is looking for a Senior Site Reliability Engineer (SRE) to join its GeForce Now (GFN) team. SRE at NVIDIA ensures that our internal and external-facing GPU cloud gaming services have reliability and uptime as promised to the users and at the same time enables developers... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    4 hours ago
  • $168k - $270.25k

     ...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is... 
    Senior
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    4 days ago
  •  ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering...  ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    2 days ago
  • $267k - $356k

     ...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-...  ...workloads in the industry, which means reliability and performance aren't just goals—they're...  ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc... 
    Senior
    Work experience placement
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    1 day ago
  • $101k - $161k

     ...several prestigious awards, such as Best Engineering Team, Best Company for Diversity,...  ...DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s...  ...: EngineeringExperience level: Mid-Senior LevelIndustry: Computer Networking
    Senior

    Arista Networks

    Santa Clara, CA
    2 days ago
  • $148k - $235.75k

     ...see how you can make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable, job-centric insights and automation for GPU fleets. We’re hiring a DevOps Engineer... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

     ...artificial intelligence.We’re looking for a Senior SRE to join our Compute Farm team and...  ...host lifecycle management, fleet reliability/auto-healing, E2E observability or data-...  ...Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $167.7k - $245.2k

     ...requiring approximately 2 days per week on-site at Cisco offices in either San...  ...behave as intended, improving reliability and reducing risks. This unified approach...  ...enhanced observability and control.As a Senior Site Reliability Engineer (SRE), you will build, operate, and... 
    Senior
    Full time
    Temporary work
    Local area
    Flexible hours
    2 days per week

    CISCO Systems

    Milpitas, CA
    3 days ago
  • $262k - $364k

     ...automation, and evolve systems by pushing for changes that improve reliability and velocity.Practice sustainable incident response and...  ...qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems... 
    Senior

    Google

    San Jose, CA
    2 days ago
  • $210.6k - $305.1k

     ...Minimum Qualifications:  You have led a distributed team of 5+ engineers, can demonstrate strong technical vision for your team, and ensure...  ..., and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible... 
    Senior
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    1 day ago
  •  ...A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure. The ideal candidate has over 8 years of experience in wireless network operations and a strong background in wireless... 
    Senior

    TechDigital Group

    Santa Clara, CA
    4 days ago
  •  ...Platform powers compute provisioning and infrastructure orchestration across our physical data centers. We are looking for a Senior Site Reliability Engineer to improve the reliability, scalability, and operational maturity of these systems as Lambda’s fleet and customer base... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago
  • $101k - $161k

     ...Requirements: We require a BS or MS in Computer Science, or equivalent relevant experience. We look for 5+ years of software engineering experience. We need experience building or operating distributed database systems or scale-out applications in a SaaS... 
    Senior

    Arastra, Inc.

    Santa Clara, CA
    4 days ago
  •  ...Senior Site Reliability Engineer (Enterprise Platform) Location: Remote - US - Open to Europe if happy to overlap with EST Compensation: Competitive We are a high-growth software company supporting the development of a premier open-source, EVM-compatible public ledger... 
    Senior
    Contract work
    Currently hiring
    Remote work

    GrabJobs

    San Jose, CA
    1 day ago
  • $187.04k - $359.72k

     ...systems by pushing for changes that improve reliability and velocity. Qualifications Minimum...  ...degree in Computer Science, Electrical Engineering, Computer Engineering or related areas....  ...Product Ops, Corporate Functions and more. On-site presence across teams allows the company... 
    Senior
    Temporary work
    Local area
    Overseas
    Shift work

    Tik Tok

    San Jose, CA
    4 days ago
  •  ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems...  ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes... 
    Senior

    Saviynt

    Milpitas, CA
    a month ago
  • $90k - $180k

     ...generic medicines. Our 115,000 colleagues serve people in more than 160 countries.JOB DESCRIPTION:About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We are... 
    Senior
    Remote work
    Shift work

    Abbott

    Sunnyvale, CA
    4 hours ago
  • $145k - $165k

     ...A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key... 
    Senior

    Bolt Graphics, Inc.

    Sunnyvale, CA
    4 days ago
  •  ..., and the challenges of building in a high-growth startup, we’d love to talk. This is more than a job—it’s a journey. Site Reliability Engineers (SREs) are responsible for the overall performance and reliability of ASAPP's infrastructure and products. The team owns... 
    Senior
    Remote work

    ASAPP

    Mountain View, CA
    more than 2 months ago
  • $120k - $200k

    Sr Site Reliability Engineer (Prisma Access) 2 days ago Be among the first 25 applicants Job Description This role requires US Citizenship. Your Career Palo Alto Networks runs a large infrastructure and is one of the biggest GCP customers. As a Principal SRE, you'll be... 
    Senior
    Rotating shift

    Palo Alto Networks

    Santa Clara, CA
    4 days ago
  • $230k - $250k

     ...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change...  ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"... 
    Night shift

    Forward Networks

    Santa Clara, CA
    4 days ago
  • $128.6k - $184.9k

     ...global cloud platform. As a team of six engineers distributed across the US, Canada, and the...  ...with a strong focus on automation, reliability, and operational excellence. We are one...  ...Qualifications7+ years of experience in Site Reliability Engineering, DevOps, Infrastructure... 
    Permanent employment
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    CISCO Systems

    Santa Clara, CA
    3 days ago
  • $272k - $431.25k

    NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's (Infrastructure, Planning and Process) Cloud Infrastructure Team. IPP is a global organization within NVIDIA. This group works with various other groups within NVIDIA such as Graphics... 
    Full time
    Work experience placement
    Worldwide

    Nvidia

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

     ...the world.Join the Simulation Software team at NVIDIA as a Senior System Software Engineer! This role offers an outstanding opportunity to work on...  ...development by enabling Chips Simulation as a trusted and reliable virtual platform.What you will be doing:Drive early... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    NVIDIA is searching for a creative and highly motivated engineer with expertise in systems software to join the GPU Software team. You will design key aspects of our production GPU kernel drivers and embedded SW that impacts our products both in the datacenter and in gaming... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $122.5k - $175k

     ...we invite you to bring your talents to Zscaler and help shape the future of cybersecurity.RoleWe are looking for a Staff Site Reliability Engineer to join our team. This is a hybrid role going into the San Jose, CA office 3 days a week, reporting to the Chief Architect... 
    Full time
    Work at office
    Local area
    3 days per week

    Zscaler

    San Jose, CA
    4 hours ago
  • $207k - $300k

     ...implementation of solutions to enhance the reliability of systems that support F1.Scale systems...  ...for multiple teams.Engage in software engineering on services written in Java, C++, and Go...  ...related technical field.Experience in a Site Reliability Engineering role.Experience... 

    Google

    San Jose, CA
    4 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!