Staff Site Reliability Engineer
$122.5k - $175kZscaler
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform.We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler.RoleWe are looking for a Staff Site Reliability Engineer to join our team. This is a hybrid role going into the San Jose, CA office 3 days a week, reporting to the Chief Architect, REIS in the Cloud Infrastructure & Operations department. You are an SRE with proven experience in Linux/UNIX System Administration and hands-on expertise in building infrastructure and managing platforms like Kubernetes using automation tools and high security standards. In this role, you will troubleshoot complex Linux networking and security issues, apply a strong understanding of firewall technologies, and manage access across various systems, platforms, and applications.What you’ll do (Role Expectations)Create and maintain highly scalable solutions based on KVM Linux, Kubernetes, and Public Cloud ProvidersAnalyze and troubleshoot systems performance and complex issues across operating systems and applicationsMaintain platform security and observability through nftables and comprehensive monitoringManage and deploy systems and software in diverse environments, including Kubernetes and OpenstackWrite and maintain custom tools using Python, Golang, and BASHWho You Are (Success Profile)You thrive in ambiguity. You're comfortable building the path as you walk it. You thrive in a dynamic environment, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful.You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution.You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact.You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback—knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust.You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose.What We’re Looking for (Minimum Qualifications)Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain5+ years of experience in a Linux/UNIX System Administration or SysAdmin role with a strong understanding of web security, protocols, and SSHTechnical security aptitude including PGP, SSH, PKI, and Multi-factor authenticationMastery of network fundamentals including DHCP, ARP, subnetting, routing, NAT, firewalls, and IPv4/IPv6Experience in container orchestration services including Docker and Kubernetes, alongside automation tools like AnsibleWhat Will Make You Stand Out (Preferred Qualifications)Experience implementing AI-driven anomaly detection, intelligent log parsing, or AIOps tools to automate root-cause analysis and predictive auto-scaling across Linux and Kubernetes infrastructureStrong Openstack and CEPH experience paired with advanced Kubernetes expertiseExtensive professional history in dedicated Linux System Administration roles along with direct experience with Secrets Management Systems such as Hashicorp Vault or similar technologies#LI-Hybrid #LI-CM3Zscaler’s salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training.The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits.Base Pay Range$122,500—$175,000 USDAt Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure.Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including:Various health plansTime off plans for vacation and sick timeParental leave optionsRetirement optionsEducation reimbursementIn-office perks, and more!Learn more about Zscaler's hybrid working model and benefitshere.By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines.Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link.Pay TransparencyZscaler complies with all applicable federal, state, and local pay transparency rules.Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
- ...We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available... ...and data platforms. The role combines database engineering, site reliability engineering, Linux systems administration, and infrastructure...Suggested
- ...automated detection, drain/cordon/taint, workload rescheduling. Feed the AIOps substrate The remediation-actuator and workflow engine land here — you make the control plane safe for automated action. Your CRDs are the schema the platform's predictors and...SuggestedLocal area
- ...The RoleThis hybrid role combines the hands-on responsibilities of a Technical Support Engineer within a SaaS (Software as a Service) environment with a growing focus on Site Reliability Engineering (SRE).The ideal candidate has a strong technical foundation, thrives in a...SuggestedFull timeLocal area
- ...Oracle Cloud Infrastructure (OCI) seeks a Senior Principal Engineer to lead the design and implementation of reliability validation for OCI control plane services, focusing on a high-performance, low-level systems approach. You will mentor engineers, define validation...Suggested
- ...in Cupertino, California, invites an experienced CDN Solutions Engineer to join the Content Delivery Network Solutions team. You will... ...and collaborate with engineering groups across Apple to ensure reliable delivery at scale. The ideal candidate has 4+ years in CDNs and...Suggested
- ...ServiceNow in Santa Clara, CA, seeks a Staff Software Engineer – SRE & AIOps to drive infrastructure automation, resilience, and toil elimination... ...for global engineering teams. Embedded within the Site Reliability & Database Engineering organization, you will architect...
$187.04k - $359.72k
...systems by pushing for changes that improve reliability and velocity. Qualifications Minimum... ...degree in Computer Science, Electrical Engineering, Computer Engineering or related areas.... ...Product Ops, Corporate Functions and more. On-site presence across teams allows the company...Temporary workLocal areaOverseasShift work$248k - $396.75k
Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused on designing, building, and operating large-scale production systems with exceptional efficiency, resilience, and availability. It combines software and systems engineering practices with...Full time- ...Job Description Job Description Site Reliability Engineer Foxconn Industrial Internet (Fii), is a world leading professional design and manufacturing service provider of communication network equipment, cloud service equipment, precision tools and industrial robots...Permanent employmentFull timeWork at officeLocal area
$148k - $235.75k
...see how you can make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable, job-centric insights and automation for GPU fleets. We’re hiring a DevOps Engineer...Full time- ...Must Have Technical/Functional Skills: 2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a related role supporting cloud-based production environments. Practical vulnerability-management experience; familiarity with...Full timeWorldwide
- ...of Huobi globe spanning infrastructure. • Work with engineering teams to make sure new features and changes are deployed quickly... .... • Constantly improve our system performance and reliability through better tools, process and monitoring system. •...Worldwide
$267k - $356k
...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-... ...workloads in the industry, which means reliability and performance aren't just goals—they're... ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc...Work experience placementWork at officeLocal areaWork from homeFlexible hours$101k - $161k
...excellence has earned us several prestigious awards, such as Best Engineering Team, Best Company for Diversity, Compensation, and Work-... ...we do.Job DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s CloudVision-as-a-...- ...to join IBM in a full‑time role between December 2027 and August 2028 upon successful completion of their degree. As a Site Reliability Engineer, you will work in an agile, collaborative environment to build, deploy, configure, and maintain systems for the IBM client...Full timeContract workPart timeFixed term contractInternshipWorldwideFlexible hoursShift work
$207.4k - $259.2k
...built specifically for aviation. We’re seeking exceptional engineers, operators and builders to join us on our mission to build... ...members.We are seeking a highly experienced and passionate Sr. Staff Site Reliability Engineer (SRE) to join our growing team. In this critical...Permanent employmentLocal areaVisa sponsorshipNight shift- ...runs on complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems remain... ...engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on...
$207k - $300k
Manage a team of Software/Systems Engineers on projects for users and remain directly responsible... ..., and establishing sustainable multi-site on-call rotations across distributed... ...global hubs.Deep practical expertise in Site Reliability Engineering practices, including SLO/SLI...- ...Job Description Job Description We are hiring Site Reliability Engineer Manager- Hybrid for a Contract To Hire position in santa clara, CA The Role You will build and lead the Site Reliability Engineering team, owning the infrastructure that development, validation...Contract workRemote work
- ...Google is seeking a Software Engineering Manager II in Site Reliability Engineering, based in Sunnyvale, California. This onsite role leads a team to ensure reliability and performance of critical systems, partnering with product and engineering teams to deliver scalable...
- ...Overview Title: Site Reliability Engineer SRE – ML platform Location: Austin, TX or Sunnyvale, CA Employment type: Full-time • Seniority: Mid-Senior level • ONLY W2 Responsibilities Continuous Deployment using GitHub Actions, Flux, Kustomize Design and implement cloud...Full time
$276.1k - $311.4k
...Vehicle Software SRE team from the ground up — defining its charter, hiring its founding engineers, establishing the operating model, and creating the technical strategy that makes reliability a first-class property of the software running on our vehicles. You'll work in a...Permanent employmentFull timeWork at officeWork from home- ...function to support one of the world’s fastest-growing AI inference services, powered by the Wafer-Scale Engine (WSE). This team will help deliver world-class, ultra-reliable inference infrastructure for leading model builders such as OpenAI and other frontier labs.As a...Shift work
$150.4k - $277.6k
...Technical Operations & Site Reliability Engineer, Customer SystemsAt Apple, Customer Experience is at the forefront of everything we do. The Customer Systems Operations team is looking for a highly skilled and motivated TechOps Engineer (Technical Operations & Site Reliability...Work experience placementRelocation$170k - $200k
...We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster, infrastructure, and monitoring systems to ensure high availability, performance...Full timeWorldwide- ...Site Reliability Engineer, Data Platform - USDS Responsibilities Engage in and improve the whole lifecycle of service, from inception and design, through to deployment, operation and refinement. Ensure reliable, fault-tolerant, efficiently scalable and cost-effective data...
- ...design by customizing MES tool per business needs Education Requirements, Ideal Experience: Associate’s degree in Industrial Engineering or IT related field Minimum of 0-3 years’ relevant experience Experience in C#, Delphi desired Knowledge of the...Work at office
- Company Description Integrated Resources, Inc is a premier staffing firm recognized as one of the tri-states most well-respected professional specialty firms. IRI has built its reputation on excellent service and integrity since its inception in 1996. Our mission ...Full time
$272k - $431.25k
...Come join the team and see how you can make a lasting impact on the world! The Data Center MODS organization seeks a Principal Engineer to architect and scale next-generation L10 and L11 diagnostic systems for Cloud Service Providers (CSPs). In this high-impact role...Full timeRemote work$149k - $359k
...opportunities and leave your mark, come join us. THE ROLE Join our Core Engineering team to power the foundation of our Enterprise Data Cloud, delivering six 9s of system reliability to global enterprises. In this role, you will take end-to-end ownership of core...Full timeWork at officeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!
- engineering aide San Jose, CA
- technology administrator San Jose, CA
- staff data engineer San Jose, CA
- senior staff systems engineer San Jose, CA
- senior staff engineer San Jose, CA
- staff engineer San Jose, CA
- assistant engineer San Jose, CA
- staff design engineer San Jose, CA
- assistant electrical engineer San Jose, CA
- assistant mechanical engineer San Jose, CA





