Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Site Reliability Engineer

$151.6k - $245.3k

Palo Alto Networks

Our Mission

At Palo Alto Networks®, we’re united by a shared mission—to protect our digital way of life. We thrive at the intersection of innovation and impact, solving real-world problems with cutting-edge technology and bold thinking. Here, everyone has a voice, and every idea counts. If you’re ready to do the most meaningful work of your career alongside people who are just as passionate as you are, you’re in the right place.

Who We Are

In order to be the cybersecurity partner of choice, we must trailblaze the path and shape the future of our industry. This is something our employees work at each day and is defined by our values: Disruption, Collaboration, Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and use it to augment the impact every individual can have. If you are passionate about solving real-world problems and ideating beside the best and the brightest, we invite you to join us!

We believe collaboration thrives in person. That’s why most of our teams work from the office full time, with flexibility when it’s needed. This model supports real-time problem-solving, stronger relationships, and the kind of precision that drives great outcomes.

Job Summary

Your Career

Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture, performance, metrics, troubleshooting, security, and reliability.

Our stack includes Kubernetes, Docker, GCP, AWS, Ansible, Terraform, Vault, Gitlab, Spinnaker, Pub/sub, Bigtable, Memorystore, Bigquery, RabbitMq, Kafka, MySQL, Python, and Go. We don’t expect you to know all these, but we do expect you to learn the ones needed for this role.

Your Impact

  • Contribute to the success of SRE and DevOps

  • Develop expertise in new technologies

  • Work with developers, researchers, data scientists, and security experts

  • Design, build, and operate reliable, secure Cloud infrastructure

  • Ensure that applications are production-ready, scalable, and reliable

  • Develop tools and automation frameworks

  • Automate robust deployment of robust services

  • Orchestrate end-to-end monitoring and alerting

  • Participate with SRE and Dev teams in the on-call rotation

  • Lead root cause analysis of critical business and production issues

  • Mentor and champion SRE culture

  • Participate in design reviews

The Team

Wildfire is the industry's largest cloud-based malware protection engine that uses machine learning and crowdsourced intelligence to instantly prevent up to 95% of unknown malware variants inline without compromising business productivity. Wildfire infrastructure team supports the scalability and high availability of Wildfire clouds.

Qualifications

Required Experience

  • 8+ years of designing, building, and operating cloud infrastructure that enables reliable, rapid deployment of microservices with resilient operations and effective monitoring.

  • BS or MS in Computer Science, a related field, or equivalent professional experience or equivalent military experience

  • Expertise in configuration management with a framework such as Ansible, Terraform, Helm, Kubernetes

  • Proficient in Python and/or Go

  • Expertise in managing applications in the Kubenetes cluster with autoscaling enabled

  • Experience in Production Engineering, DevOps, or Site Reliability

  • Expertise in the public cloud (GCP or AWS), especially in GCP

  • Strong Linux administration, internals, and network troubleshooting

  • Proficiency with programming languages like Python, Golang, and shell scripting to automate tasks

  • Experience with CI/CD pipelines, GitLab, and GitHub preferred

  • Ability to diagnose and troubleshoot complex distributed systems handling high-volume transactions

  • Excellent written and verbal communication, able to collaborate and rally support

  • Self-disciplined, self-managed, self-motivated, and strong sense of ownership, urgency, and drive

  • Passion for infrastructure and monitoring as code

  • Ready to understand and dissect new technology stacks quickly

Compensation Disclosure

The compensation offered for this position will depend on qualifications, experience, and work location. For candidates who receive an offer at the posted level, the starting base salary (for non-sales roles) or base salary + commission target (for sales/com-missioned roles) is expected to be the annual range listed below. The offered compensation may also include restricted stock units and a bonus. A description of our employee benefits may be found here ( .

$151,600.00 - $245,300.00/yr

Our Commitment

We’re trailblazers that dream big, take risks, and challenge cybersecurity’s status quo. It’s simple: we can’t accomplish our mission without diverse teams innovating, together.

We are committed to providing reasonable accommodations for all qualified individuals with a disability. If you require assistance or accommodation due to a disability or special need, please contact us at View email address on click.appcast.io .

Palo Alto Networks is an equal opportunity employer. We celebrate diversity in our workplace, and all qualified applicants will receive consideration for employment without regard to age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or other legally protected characteristics.

All your information will be kept confidential according to EEO guidelines.

Is role eligible for Immigration Sponsorship?: Yes

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Principal Site Reliability Engineer in Santa Clara, CA vacancy
  • $248k - $396.75k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused on designing, building, and operating large-scale production...  ...strengthen the reliability of production environments.As a Principal SRE, you will shape the technical direction of NVIDIA’s... 
    Principal
    Full time

    NVIDIA

    Santa Clara, CA
    3 days ago
  • $260k - $275k

     ...used by global enterprises • Solve complex reliability challenges at scale • Influence architecture and engineering culture at a company level • Competitive compensation...  ...You Bring 1+ years of experience as a Principal SRE with a strong focus on building tools and... 
    Principal

    Saviynt

    Milpitas, CA
    1 day ago
  • $160k - $240k

     ...one another millions of times a day - quickly, reliably, and securely. Any time you swipe your credit...  ...come make a difference at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our global team in... 
    Suggested
    Full time

    Fiserv

    Sunnyvale, CA
    4 days ago
  • $174k - $252k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 
    Suggested

    Google

    Sunnyvale, CA
    4 hours ago
  • $104.9k - $174.7k

     ...SRE role is responsible for improving the reliability, availability, performance, and...  ...actions through completion.Follow up with engineering, development, security, support, and business...  ...Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering... 
    Suggested
    Full time
    Local area

    RELX Group

    San Jose, CA
    3 days ago
  • $262k - $364k

     ...and training AI infrastructure from SRE side, ensuring it is reliable, scalable, cost effective and performant, while working closely...  ...qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems engineering... 

    Google

    Sunnyvale, CA
    1 day ago
  •  ...Job Title : Senior Site Reliability Engineer Location : Santa Clara, CA Contract ENGAGEMENT SUMMARY The Candidate will provide SRE services for AI platforms and supporting infrastructure with emphasis on reliability engineering, incident response... 
    Contract work

    VDart

    Santa Clara, CA
    2 days ago
  • $132.6k - $214.5k

     ...As part of this role, you will collaborate closely with our engineering teams to develop innovative solutions that provide clear and...  ...team to influence the operability of the product and ensure the reliability and availability of our services. Qualifications... 
    Full time
    Work at office
    Visa sponsorship
    Work visa

    Palo Alto Networks

    Santa Clara, CA
    3 days ago
  • $150.4k - $277.6k

     ...Services The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples of Apple’s long...  ...field with 4+ years experience At least 6 years in a Reliability Engineering, DevOps or infrastructure focused role Advanced... 
    Relocation
    Day shift

    Apple

    Cupertino, CA
    3 days ago
  • $272k - $431.25k

     ...We are now looking for a Principal Software Engineer for LPX System Software! NVIDIA’s LPX System Software team builds the foundational software...  .... Demonstrated leadership driving triage of difficult reliability issues to clear, written root‑cause analysis. Low‑... 
    Principal
    Shift work

    NVIDIA Gruppe

    Santa Clara, CA
    4 days ago
  •  ...keep the world running. Location: 5 on-site days a week in Sunnyvale, CA Headquarters. Our Team's Vision: Our Engineering team is shaping the future of...  ...are looking for an experienced Senior Site Reliability Engineer (SRE) with a strong background in... 
    Work experience placement
    Immediate start

    Illumio

    Sunnyvale, CA
    5 days ago
  • $175k - $265k

     ...Overviewd-Matrix's SRE team owns the infrastructure layer that every engineering team and customer depends on — colocation facilities, on-...  .... This role is a core member of that team, responsible for reliability, automation, and observability across colo, on-premises lab,... 

    d-Matrix

    Santa Clara, CA
    4 hours ago
  •  ...Cadence Design Systems is seeking a talented Software Engineer for its R&D group to create advanced software for physical verification of semiconductor devices. The role emphasizes design, validation, and collaboration across a distributed team to advance technology at... 
    Principal

    Jobleads-US

    San Jose, CA
    2 days ago
  •  ...Cadence Design Systems seeks a Software Engineer for its Research & Development group, focusing on advanced software for physical verification at leading-edge semiconductor nodes. The role drives innovation in development and validation of software used to design and... 
    Principal

    Jobleads-US

    San Jose, CA
    2 days ago
  •  ...Oracle Cloud Infrastructure (OCI) seeks a Senior Principal Engineer to lead the design and implementation of reliability validation for OCI control plane services, focusing on a high-performance, low-level systems approach. You will mentor engineers, define validation... 
    Principal

    Jobleads-US

    Santa Clara, CA
    4 days ago
  • $170k - $200k

     ...We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster, infrastructure, and monitoring systems to ensure high availability, performance... 
    Full time

    Zoomcar

    Sunnyvale, CA
    2 days ago
  •  ...CloudOps— the team that keeps Splunk Cloud running for some of the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines at a scale very few teams ever get to operate at. When the... 

    Outshift by Cisco

    San Jose, CA
    4 days ago
  • $145k - $165k

     ...: Selflessly collaborate towards our shared purpose. About the role Bolt Graphics is seeking a highly experienced Site Reliability Engineer (SRE) to design, build, and operate highly reliable developer and production systems. This role is mission-critical to maintaining... 
    Work at office
    Immediate start

    Bolt Graphics, Inc.

    Sunnyvale, CA
    5 days ago
  • $230k - $250k

     ...minds are shaping the future of network reliability, security, and AI‑ready operations. About...  ...you will be building the reliability engineering function at Forward — defining how we...  ...Looking For ~6+ years of experience in site reliability engineering, DevOps, or... 
    Night shift

    Forward

    Santa Clara, CA
    2 days ago
  • $272k - $431.25k

     ...We are hiring senior engineers to work on the CUDA driver, a core component of our platform for accelerating general purpose computation on the GPU. Our team delivers features and improvements to better realize the potential of NVIDIA hardware for a growing range of computational... 
    Principal

    NVIDIA Gruppe

    Santa Clara, CA
    2 days ago
  •  ...Job Description Job Description Site Reliability Engineer Foxconn Industrial Internet (Fii), is a world leading professional design and manufacturing service provider of communication network equipment, cloud service equipment, precision tools and industrial robots... 
    Permanent employment
    Full time
    Work at office
    Local area

    Foxconn Industrial Internet - FII

    San Jose, CA
    a month ago
  •  ...Must Have Technical/Functional Skills: 2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a related role supporting cloud-based production environments. Practical vulnerability-management experience; familiarity with... 
    Full time
    Worldwide

    SFE

    San Jose, CA
    3 days ago
  •  ...design by customizing MES tool per business needs Education Requirements, Ideal Experience: Associate’s degree in Industrial Engineering or IT related field Minimum of 0-3 years’ relevant experience Experience in C#, Delphi desired Knowledge of the... 
    Work at office

    Foxconn Industrial Internet - FII

    Sunnyvale, CA
    a month ago
  •  ...Job Description Job Description Site Reliability Engineer II Bay Area, offices in San Jose · Hybrid · 24/7 FedRAMP Operations · Rotational Shift · Initial Contract till March 27. KEY REQUIREMENT This role requires US citizenship and residence on US soil.... 
    Hourly pay
    Contract work
    For contractors
    Shift work
    Night shift
    Weekend work

    C-Serv

    San Jose, CA
    29 days ago
  •  ...of Huobi globe spanning infrastructure. •       Work with engineering teams to make sure new features and changes are deployed quickly...  .... •       Constantly improve our system performance and reliability through better tools, process and monitoring system. •... 
    Worldwide

    Cryptoware Technologies Inc

    Santa Clara, CA
    a month ago
  •  ...that keep the world running. Location: 5 On-Site Days a Week in Sunnyvale, CA Headquarters Our Engineering team is driven by a culture that thrives on visionary...  ...to-day basis, you will work on enhancing system reliability and scalability of Illumio SaaS products, and... 
    Work experience placement
    Immediate start

    Illumio

    Sunnyvale, CA
    5 days ago
  • $60 - $62 per hour

     ...and improving existing processes to enhance overall system reliability. Key Responsibilities: Deploying software to cloud...  ...Computer Science or a related field. 3+ years of experience in Site Reliability Engineering. Proficiency with Kubernetes, Helm, Linux, AWS networking... 
    Hourly pay
    Contract work
    Remote work

    Akraya

    Santa Clara, CA
    5 days ago
  •  ...Position: Site Reliability Engineering (SRE) Location: Santa Clara, CA (Onsite) Duration: W2 / C2C Contract Experience: 10+ Years Job Description: • WS application and CI/CD pipelines, Microsoft Server admin and workload support (Data Center and AWS... 
    Contract work
    Immediate start

    Syntricate Technologies

    Santa Clara, CA
    2 days ago
  • Job Description : Need to have experience with ticket support, azure, Splunk, ServiceNow, and any Java experience is a plus. Ideally candidates that come from an Enterprise background Handling tickets for the Walmart environment. Splunk, Servicenow...

    3B Staffing LLC

    Sunnyvale, CA
    2 days ago
  • $149.8k - $224.6k

     ...This hybrid role combines the hands-on responsibilities of a Technical Support Engineer within a SaaS (Software as a Service) environment with a growing focus on Site Reliability Engineering (SRE). The ideal candidate has a strong technical foundation, thrives... 
    Local area

    F5

    San Jose, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!