Senior Site Reliability Engineer
$82.3k - $228.8kGrabJobs
Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. We are seeking a highly experienced Senior Site Reliability Engineer – Compute Platforms to design, implement, and support Kubernetes on baremetal and hypervisor platforms in a private cloud environment. This role is responsible for the architecture, design, and standardization of enterprise compute and hypervisor environments spanning bare metal infrastructure, operating systems, hypervisors, private cloud orchestration, and Kubernetes using Infrastructure-as-Code and GitOps practices. This is a deeply technical role requiring expert-level understanding of compute hardware management, Kubernetes, OpenStack, hypervisors and extensive working knowledge on Linux Operating systems. You will also collaborate with platform and SRE teams to maintain secure, performant, and multi-tenant-isolated services that serve high-throughput, mission-critical applications. Key Responsibilities Lead the architecture and design of enterprise compute and hypervisor platform solutions across hardware, OS, virtualization, cloud orchestration, and container orchestration layers Define standards and automation frameworks for bare metal provisioning and lifecycle management Design and implement Bare Metal as a Service (BMaaS) capabilities for scalable infrastructure consumption Architect and design Kubernetes platforms on bare metal with QoS and Affinity (ArgoCD) Architect and validate automated deployments of operating systems and hypervisors including Ubuntu and Harvester Design and maintain PXE-based provisioning environments leveraging Redfish APIs for large-scale server deployments Develop Infrastructure-as-Code using Ansible, Terraform, Helm and Git, with Python/Bash automation. Implement CI/CD pipelines for infrastructure updates, patching, upgrades, testing, and rollback. Design automated workflows for server build, firmware lifecycle management, patching, and hardware validation Evaluate and standardize enterprise hardware platforms to meet performance, scalability, and reliability requirements Produce detailed high-level and low-level design documentation , build guides, and operational handoff materials Perform deep troubleshooting across storage, Kubernetes, hypervisors, networking, and Linux systems Partner with operations, network, storage, and platform teams to ensure designs are supportable and production-ready Participate in on-call escalation support for complex platform-related issues Collaborate globally on change management , documentation, and operational best practices Minimum Qualifications 6 + years of experience in infrastructure engineering, platform engineering, or DevOps with a strong focus on Compute system design Proven experience designing and automating bare metal compute environments at scale Strong hands-on experience with PXE boot, network-based OS provisioning, and automated server imaging Experience implementing or supporting Bare Metal as a Service (BMaaS) platforms Practical experience using Redfish APIs for hardware provisioning, power management, and remote lifecycle operations Deep expertise with Ubuntu Linux in enterprise environments Strong Hands-on experience with KVM hypervisors (Suse Harvester, OpenStack). Experience designing and deploying production-grade Kubernetes clusters Strong background with enterprise compute hardware platforms , including Cisco UCS, Dell PowerEdge, Supermicro systems & HPE Proficiency with Infrastructure as Code tools (e.g., Terraform, Ansible, or similar) Experience building or supporting CI/CD pipelines for infrastructure and platform automation Strong scripting skills in Python, Bash, or similar languages Demonstrated ability to produce clear, structured technical design documentation Excellent written and verbal communication skills Bachelor’s degree in computer science or equivalent professional experience Preferred Qualifications OpenStack, Ubuntu KVM administration. BareMetal as a Service (PXE, Redfish). Kubernetes on BareMetal CIS/NIST security and infrastructure lifecycle management. ITIL Foundation/advanced certifications in support of ITSM standard methodology. Background in telco, edge cloud, or large enterprise environments. Ubuntu Certifications, CNCF Certified Kubernetes Administrator (CKA), Certified Kubernetes Security Specialist (CKS) Master’s degree in computer science, IT, Engineering, or a related field preferred; equivalent experience and relevant industry certifications will also be considered What You’ll Get A collaborative team that’s deeply invested in infrastructure excellence. Complex technical challenges that require creative, scalable solutions. The opportunity to shape a next-generation private cloud platform-built reliability Access to the latest tools, frameworks, and upstream project developments Skills and Attributes: Analytical Thinking & Problem Solving: Demonstrated ability to translate complex, cross-domain requirements into scalable and resilient cloud infrastructure and automation solutions Collaboration & Teamwork: Strong interpersonal and communication skills with a proven track record of effective collaboration across multidisciplinary teams, including developers, operations, security, and product stakeholders Mentorship & Leadership: Passionate about knowledge-sharing and mentorship, with experience guiding junior engineers and fostering a team culture of continuous learning, innovation, and technical excellence in cloud engineering and DevOps practices Work Location: This role is fully remote for candidates who reside outside the 50 mile radius of our San Ramon office. For candidates who reside within 50 miles of our San Ramon location, this role is Hybrid and would require 3 days a week (M, W, TH) in our San Ramon office. As part of our continued commitment to diversity, equity, and inclusion, Five9 supports pay transparency during the entire recruitment process. Actual compensation packages are based on several factors that are unique to each candidate including, but not limited to: skill set, depth of experience, certifications, and specific work location. The range displayed reflects the minimum and maximum target for new hire salaries for the job across the United States. Your recruiter can share more about the specific compensation package during your hiring process. Additionally, the total compensation package for this position may also include an annual performance bonus, stock, and/or other applicable incentive compensation plans. Our total reward package also includes: Health, dental, and vision coverage, beginning on the first day of employment. Five9 covers 100% of the employee portion of the health, dental and vision coverage and shares a high portion of the dependent cost. We also offer Short & Long-Term Disability, Basic Life Insurance, and a 401k saving plan with employer matching. Access to an innovative mental health support platform that offers personalized care and resources in areas such as: therapy, coaching and self-guided mindfulness exercises for all covered employees and their covered dependents. Generous employee stock purchase plan. Paid Time Off, Company paid holidays, paid volunteer hours and 12 weeks paid parental leave. All compensation and benefits are subject to the requirements and restrictions set forth in the applicable plan documents and any written agreements between the parties. The US base salary range for this role is below. $82,300 - $228,800 USD Five9 embraces diversity and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. The more inclusive we are, the better we are. Five9 is an equal opportunity employer. View our privacy policy, including our privacy notice to California residents here: . Note: Five9 will never request that an applicant send money as a prerequisite for commencing employment with Five9.
$150k - $180k
...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business... ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or...SeniorPermanent employmentFull timeWork at officeLocal areaRemote workWorldwide$166k - $220k
...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the... ...-to-end solutions we ship.ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Irvine...SeniorFull timeWork experience placementImmediate start$207k - $284.9k
...We're all in on this mission. If you are too, let's talk.Senior Manager, Site Reliability EngineeringSecure Every Identity, from AI to... ...mission. If you are too, let's talk.The Federal Operations Engineering GroupOkta's Federal Operations team supports government...SeniorPermanent employmentLocal areaWorldwideFlexible hoursDay shift$166k - $220k
...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &... ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril...SeniorFull timeWork experience placementImmediate start- ...Azure, Oracle, Cassandra, SQL Server, My SQL and Mongo DB Seniority level Seniority level Mid-Senior level Employment type Employment... ...new job is posted. Sign in to set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,000.00-$259,000.00 1 day ago...SeniorContract workRemote work
$175k - $250k
...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or be willing to... ...ensuring scalability, performance, and reliability across environments. What You’ll Do Design...SeniorFull timeRemote workRelocationRelocation package$149.4k - $202k
...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing...SeniorRemote work- ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability Engineer (SRE) who will play a key role in building a sustainable platform by developing systems for analyzing environments, predicting, and resolving issues, and supporting the production environment...SeniorWork experience placement
$160k - $240k
...passionate about building unified IT solutions that simplify the way IT organizations work. We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the Platform Engineering organization and help us scale our products to millions of end-users....SeniorPermanent employmentFull timeRemote workWork from homeRelocationFlexible hours$140k - $205k
...Senior Technology Site Reliability Engineer Cooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operationsteam. Position summary: The Senior Technology Site Reliability Engineer("SRE") is responsible for ensuring the reliability...SeniorFull timeTemporary workWork at officeFlexible hoursWeekend work$185k - $230k
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services...SeniorFull timeLocal areaImmediate start$165k - $230k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts....SeniorPermanent employmentTemporary workImmediate startWeekend work- ...Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the guardian of our production ecosystems, ensuring that our complex, data-driven AI platforms remain resilient, scalable, and highly performant...SeniorLocal area
$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. The Senior Site Reliability Engineer Opportunity Reporting to the Manager, Site Reliability Engineering , this role will...SeniorPermanent employmentLocal areaWorldwideFlexible hours$115.5k - $164.8k
...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant... ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations...Work experience placementWork at officeRemote work$230k - $250k
GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is...Remote work$125k - $185k
...lifesaving drugs, forecast supply chain disruptions, locate missing children, and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-performance, scalable, and reliable services for our production...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package$125k - $185k
Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the world’s leading software for data-driven decisions... ...locate missing children, and more.The RoleWe’re looking for Site Reliability Engineers who can help us build, operate, and maintain high-...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package$112k - $179k
...delivery of system, network, software, and security solutions.About The RolePeraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems...Contract workWorldwideShift work- ...Candidates eligible for polygraph upgrade will be considered. Position Overview We are seeking an experienced Senior DevOps / Site Reliability Engineer (SRE) to support a mission-critical national security program. This position is ideal for a hands-on engineer with...SeniorVisa sponsorship
- ...Principal Site Reliability Engineer The Principal Site Reliability Engineer will be a critical technical leader responsible for driving the... ...for a key Randstad client in the Washington D.C. area. This senior role merges deep expertise in infrastructure automation (IaC...
$174k - $239k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Staff Site Reliability Engineer OpportunityOkta Federal, Inc. is looking for an experienced Staff TDI Site Reliability...Local areaWorldwideFlexible hours$174k - $238k
...work. We're all in on this mission. If you are too, let's talk.The Federal SRE TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the Emerging Products Group (EPG). Our mission is to build highly reliable, scalable,...Local areaWorldwideFlexible hours$135k - $145k
...and positive change. About the role: We're looking for a Senior Platform Engineer to build and operate the CI/CD pipelines, Infrastructure... ...Deliver platform improvements with measurable impact on reliability and developer productivity. Participate in two-week sprints...SeniorFull timeTemporary workLocal areaWeekend work$130k - $180k
...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and... ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to...Temporary workWork at officeImmediate startRemote workFlexible hours- ...Job Title: Site Reliability Engineer (SRE) Location: Washington, DC (Onsite) Clearance: TS/SCI Position Overview Seeking a highly motivated Site Reliability Engineer (SRE) to support mission-critical enterprise applications and infrastructure in...
$86.8k - $198k
Software Engineer, Senior The Opportunity: As a full stack developer, you can resolve a problem... ...APIs). You will work to ensure secure, reliable, RMF compliant application performance... ...the Resource page on our Careers site and reviewing Our Employee Benefits page...SeniorFull timeContract workPart timeWork at officeLocal areaRemote work- ...SRE Support Engineer - Observability While this position is not currently open, we are interviewing strong candidates for upcoming opportunities... ...support across Slack and tickets, improving monitoring reliability, and reducing incident impact through better triage,...Remote work
$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office- ...Site Reliability Engineer (SRE) Randstad is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our client in the Washington D.C. area, focusing on optimizing the availability, performance, and scalability of critical production services. The ideal...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Washington DC
- site reliability engineer sre Washington DC
- site reliability engineer remote Washington DC
- senior business analyst Washington DC
- senior risk manager Washington DC
- senior cost estimator Washington DC
- senior contract specialist Washington DC
- senior manager tax Washington DC
- senior automation engineer Washington DC
- senior devops Washington DC

