Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Reliability Engineer

$166.5k - $291.4k

ServiceNow

Company DescriptionIt all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker, Phyllis. She cried tears of joy. That moment inspired Fred to build a company that could do that for everyone—freeing people from busywork so they could focus on meaningful work. Today, ServiceNow is the AI control tower for business reinvention. Our ServiceNow AI platform brings together any AI, any data, and any workflow— helping 85% of the Fortune 500 work smarter, faster, and better. We're building an AI-native culture where technology and talent are unstoppable together. And we're just getting started.Join us to put AI to work for people.Job DescriptionJoin us to build the next generation of cloud-native reliability, release, and test platforms that enable engineering excellence, developer productivity, and high-confidence ServiceNow releases through automation, observability, and AI-driven operations.What you get to do in this role:Design, build, and operate cloud-native engineering platforms for software validation, release validation, and production readinessDesign and maintain production-like release and test ServiceNow environments that improve release confidence and deployment readiness.Build and integrate automated test pipelines, observability, reliability signals, deployment intelligence, and quality gates into CI/CD workflows.Develop automation solutions that improve engineering productivity, streamline operations, and reduce manual toil through shift-left engineering practices.Build reusable frameworks, self-service engineering environments, test data management, mock services, and developer productivity tooling.Design and enhance Kubernetes-based platforms supporting scalable test infrastructure, release automation, cloud-native workloads, and developer self-service.Implement automated validation for failure detection, deployment verification, policy enforcement, security checks, resilience testing, and operational health assessments.Resolve complex platforms, infrastructure, and networking challenges through software engineering, systems design, and automation.Partner closely with engineering teams to improve platform reliability, release quality, cloud-native adoption, and engineering best practices.Participate in architecture reviews, technical design discussions, and implementation of scalable, automation-first engineering solutions.Influence technical decisions through strong engineering execution, collaboration, and delivery of high-quality platform capabilities.Mentor engineers through technical guidance, code reviews, knowledge sharing, and engineering best practices.Foster a culture of reliability, automation, operational excellence, continuous improvement, and customer-focused engineering.QualificationsTo be successful in this role you have:Experience in leveraging or critically thinking about how to integrate AI into work processes, decision-making, or problem-solving. This may include using AI-powered tools, automating workflows, analyzing AI-driven insights, or exploring AI's potential impact on the function or industry.8+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Software Engineering, or Infrastructure Engineering with a Bachelor's degree; or 6 years and a Master's degree; or a PhD with 3 years experience; or equivalent experience.Hands-on experience with Kubernetes across cluster operations, networking, storage, security, autoscaling, and multi-cluster environments.Experience building and operating cloud-native platforms supporting scalable, highly available services.Experience integrating Kubernetes with CI/CD, GitOps, automated test pipelines, deployment validation, and cloud-native deployment workflows.Experience designing and implementing automation to improve developer productivity, release quality, and operational efficiency.Experience with progressive delivery practices, including canary deployments, feature flags, automated rollback, and deployment verification.Experience with chaos engineering, resilience testing, disaster recovery, and reliability validation.Strong software engineering skills with hands-on experience designing, developing, testing, and debugging applications using Python, Go, Java, or Ruby.Experience leveraging AI-assisted engineering for intelligent testing, release risk analysis, incident diagnostics, or operational automation is a plus.Strong understanding of observability, monitoring, SLI/SLOs, incident management, and production operations for distributed systems.Demonstrated ability to solve complex technical problems, drive projects independently, and collaborate effectively across engineering teams.Thrives in fast-paced, ambiguous environments with a strong ownership mindset, bias for action, and a passion for continuous learning and automation.Low ego, intellectually curious, and an effective collaborator who enjoys partnering with globally distributed teams to deliver reliable engineering solutions.Good to have:Experience with observability and monitoring platforms for applications, services, and distributed systems at scale.Experience with DevOps automation, CI/CD pipelines, GitOps, and Agile development practices using tools such as GitLab CI/CD, Argo CD, or Flux.Experience building and maintaining enterprise-scale test automation frameworks using technologies such as Playwright, Selenium, Cypress, REST Assured, PyTest, JUnit/TestNG, or equivalent.Experience with test orchestration, intelligent regression testing, test impact analysis, flaky test detection, parallel execution, and test data management.Experience with service virtualization, contract testing, synthetic testing, and building developer self-service engineering platforms.Experience with Infrastructure as Code and configuration management tools such as Ansible, Terraform, or equivalent.Experience with the Kubernetes ecosystem, including Helm, Argo Workflows, Kustomize, Istio/Linkerd, Gateway API/Ingress, Prometheus, OpenTelemetry, and container runtime technologies.Experience operating Kubernetes platforms across public cloud providers, including AWS (EKS), Azure (AKS), and Google Cloud (GKE).Experience implementing progressive delivery practices, including canary deployments, feature flags, deployment verification, and automated rollback.Familiarity with AI-assisted engineering, intelligent testing, operational automation, or cloud-native engineering platforms.For positions in this location, we offer a base pay of $166,500 - $291,400, plus equity (when applicable), variable/incentive compensation and benefits. Sales positions generally offer a competitive On Target Earnings (OTE) incentive compensation structure. Please note that the base pay shown is a guideline, and individual total compensation will vary based on factors such as qualifications, skill level, competencies, and work location. We also offer health plans, including flexible spending accounts, a 401(k) Plan with company match, ESPP, matching donations, a flexible time away plan and family leave programs. Compensation is based on the geographic location in which the role is located and is subject to change based on work location.Additional InformationWork PersonasWe approach our distributed world of work with flexibility and trust. Work personas (flexible, remote, or required in office) are categories that are assigned to ServiceNow employees depending on the nature of their work and their assigned work location. Learn more here. To determine eligibility for a work persona, ServiceNow may confirm the distance between your primary residence and the closest ServiceNow office using a third-party service.Equal Opportunity EmployerServiceNow is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, national origin, age, disability, gender identity, veteran status, or any other category protected by law. In addition, all qualified applicants with arrest or conviction records will be considered for employment in accordance with legal requirements.  AccommodationsWe strive to create an accessible and inclusive experience for all candidates. If you require a reasonable accommodation to complete any part of the application process, or are unable to use this online application and need an alternative method to apply, please contact View email address on click.appcast.io for assistance. Export Control RegulationsFor positions requiring access to controlled technology subject to export control regulations, including the U.S. Export Administration Regulations (EAR), ServiceNow may be required to obtain export control approval from government authorities for certain individuals. All employment is contingent upon ServiceNow obtaining any export license or other approval that may be required by relevant export control authorities. From Fortune. 2026 Fortune Media IP Limited. All rights reserved. Used under license.SummaryType: Full-timeFunction: Information TechnologyExperience level: Not ApplicableIndustry: Information Technology And Services

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Staff Reliability Engineer in Santa Clara, CA vacancy
  • $146.2k - $255.9k

     ...applications, data, cloud environments, and AI agents. For engineers joining Veza today, this means the scale and resources of an...  ...industry. Job Description We are seeking an exceptional Staff Site Reliability Engineer to lead critical infrastructure initiatives and drive... 
    Suggested
    Full time
    Work at office
    Remote work
    Flexible hours

    ServiceNow

    Santa Clara, CA
    3 hours ago
  • At Bloom Energy, our vision for a world powered by clean, reliable, and affordable energy is more than just a dream—we’re making...  ...challenges of the 21st century.We are looking for a Staff Reliability Engineer to join our team in one of today’s most exciting technologies... 
    Suggested
    Full time
    Work experience placement
    Remote work
    Worldwide

    Bloom Energy

    San Jose, CA
    1 day ago
  • $207.4k - $259.2k

     ...differences, and supports and celebrates all of our team members.We are seeking a highly experienced and passionate Sr. Staff Site Reliability Engineer (SRE) to join our growing team. In this critical role, you will be responsible for the reliability, scalability,... 
    Suggested
    Permanent employment
    Local area

    Archer Aviation

    San Jose, CA
    1 day ago
  • $122.5k - $175k

     ...age, we invite you to bring your talents to Zscaler and help shape the future of cybersecurity.RoleWe are looking for a Staff Site Reliability Engineer to join our team. This is a hybrid role going into the San Jose, CA office 3 days a week, reporting to the Chief... 
    Suggested
    Full time
    Work at office
    Local area
    3 days per week

    Zscaler

    San Jose, CA
    3 days ago
  • $207k - $301k

    Develop strong, influential relationships with multiple stakeholders across the Site Reliability Engineering and Developer organizations.Serve as an expert on particular fields of knowledge related to rate limiting or sharding.Develop plans and lead projects on evolving... 
    Suggested

    Google

    Sunnyvale, CA
    7 hours ago
  • $207k - $301k

     ...design and lead the implementation of solutions to enhance the reliability of systems that support F1.Scale systems sustainably through...  ...improve reliability for multiple teams.Engage in software engineering on services written in Java, C++, and Go (including instrumenting... 

    Google

    San Jose, CA
    3 days ago
  • $217.57k - $260k

     ...AI can be appropriately used during your application and interviews which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure team responsible for owning environments with thousands... 
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours
    Shift work

    ID.me

    Mountain View, CA
    1 day ago
  • $165k - $192.5k

    Staff Reliability Engineer - Power Electronics Reliability Engineers at Lunar Energy will be responsible for ensuring product reliability throughout the entire lifecycle of our revolutionary home energy products. This includes providing input during the design phases,... 
    Full time

    Lunar Energy

    Mountain View, CA
    1 day ago
  • $220k - $255k

     ...and SUV miles with ones on vehicles that are more affordable, more enjoyable and 10-50x more efficient. ALSO is looking for a Reliability Engineer to play a key role in developing and leading the reliability of multiple electric mobility vehicles over the full... 
    Full time
    Local area
    Flexible hours

    ALSO

    Palo Alto, CA
    3 days ago
  • $181k - $262k

     ...get crucial goods where they need to go, and make mobility more efficient and accessible for all.We’re searching for a Staff Hardware Reliability Engineer.The Hardware Reliability team is dedicated to ensuring the robustness and dependability of hardware systems in the... 
    Contract work
    Work at office
    Local area
    3 days per week

    Aurora Innovation

    Mountain View, CA
    1 day ago
  • $181k - $262k

     ...get crucial goods where they need to go, and make mobility more efficient and accessible for all.We’re searching for a Staff Hardware Reliability Engineer who will join the Hardware Reliability team to ensure the robustness of the Aurora Driver by leading reliability... 
    Contract work
    Work at office
    Local area
    3 days per week

    Aurora Innovation

    Mountain View, CA
    2 days ago
  • $232k - $263k

     ...scaling rapidly toward long-term growth and IPO readiness.Sr. Staff Site Reliability EngineerAs a Sr. Staff SRE at Obsidian, you will define...  ...operate as a strategic partner to DevOps and Platform Engineering leadership, shaping a unified reliability strategy that scales... 
    Work from home

    Obsidian Security

    Palo Alto, CA
    1 day ago
  • $122.44k - $232.19k

     ...workloads. With a startup-like culture, we move quickly and give engineers the opportunity to drive significant technical and business...  ...chip development flow. Mission: Define and own the pod-level reliability specifications that ensure the availability, resilience, and... 
    Full time
    Local area
    Immediate start
    Shift work

    Intel Corporation

    Santa Clara, CA
    1 day ago
  • $168k - $264.5k

     ...human inventiveness and intelligence. Make the choice to join us today. We are seeking an outstanding candidate for Silicon Reliability Engineer to drive and utilize cutting edge technologies to deliver high performance products while ensuring world class reliability.What... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $86.48k - $118.91k

     ...software technologies into solutions that combat climate change, reliably connect humans and the world, and help drive advancements in...  ...your career and help innovate ahead of what’s possibleReliability Engineer will be responsible for:Project management of strategic... 
    Permanent employment
    Full time
    Work at office
    Day shift

    Analog Devices

    San Jose, CA
    2 days ago
  • $116k - $184k

     ...profoundly impacting society. Come join the team and help build the next era of computing!We're seeking an outstanding Senior HTOL Reliability Engineer to join our Santa Clara lab. This role requires deep device-circuitry knowledge and hands-on hardware development. You will... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $168k - $264.5k

     ...industry. As the computational power increases with every GPU generation, developing efficient and reliable systems is an imperative. We are looking for a System Reliability Engineer to join NVIDIA's existing Reliability Engineering team, involved in NVIDIA's diverse system... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...of the most challenging global health issues.Play a meaningful role in the development of new life science technology as a reliability engineer in the R&D Impact Engineering group. Provide hands-on, analytical, and reliability expertise to support engineering programs... 
    Full time

    BD, Becton and Dickinson

    San Jose, CA
    2 days ago
  • $135k - $185k

    Req ID: 138673 Region: Americas Country: USA State/Province: California City: San Jose SummaryOur Engineering Team is seeking a highly motivated and self-driven candidate qualified for the position of Mechanical Engineering to be a part of a market leading Hardware Design... 
    Work at office

    Celestica

    San Jose, CA
    1 day ago
  • $118.01k - $176.8k

     ...learn, and lead. Your Team, Your ImpactWe are seeking a Sr. Staff Thermo Mechanical Engineer to join our Advanced Technology Infrastructure team, focusing on mechanical simulation and structural reliability of semiconductor test and validation infrastructure. This role... 
    Permanent employment
    Full time
    Internship
    Work from home

    Marvell

    Santa Clara, CA
    1 day ago
  •  ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by...  ...Function of Position The Staff Mechanical Design Engineer will play a leading...  ...solutionsPerform design FMEADesign for reliability, cost and manufacturabilityExercise creativity... 
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    2 days ago
  • $209k - $219.45k

     ...detailed design of parts and subassemblies, technical drawing creation, and control of space allocation and routing envelopes. Conduct Engineering Analysis including thermal modeling, structural analysis, and aircraft level weight and balance modeling to validate packaging... 
    Local area

    Archer Aviation

    San Jose, CA
    7 hours ago
  • $163k - $253k

     ...and communities.Samsung Semiconductor is hiring for a Packaging Engineer, Mechanical Simulation role to lead the development and design...  ...and simulation to evaluate and improve the performance and reliability of cutting‑edge semiconductor packages and memory devices.Provide... 
    Flexible hours

    Samsung Semiconductor

    San Jose, CA
    1 day ago
  •  ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by...  ...purpose here.Job DescriptionManaging Staff, Systems Mechanical EngineeringJoin...  ...and through validation to ensure robust, reliable products.Lead cross-functional projects,... 
    Full time
    Local area
    Worldwide
    Relocation package
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    4 days ago
  •  ...people from diverse backgrounds and industries to create a safer, sustainable and more connected world. Job OverviewA Quality & Reliability Engineer Supervisor uses their engineering skills to assist in issues related to the Quality of the product. This job also involves... 
    Remote work

    TE connectivity

    San Jose, CA
    1 day ago
  •  ...we advance your career. THE ROLE:We are looking for a Quality Engineer II to support component- and board-level failure analysis for...  ...fault isolation, physical failure analysis, board-level analysis, reliability stress support, and clear technical reporting. This is a hands... 

    AMD

    San Jose, CA
    3 days ago
  • $147k - $202.5k

    Who We AreApplied Materials is a global leader in materials engineering solutions used to produce virtually every new chip and advanced display in the world. We design, build and service cutting-edge equipment that helps our customers manufacture display and semiconductor... 
    Full time

    Applied Materials

    Santa Clara, CA
    4 days ago
  • At Bloom Energy, our vision for a world powered by clean, reliable, and affordable energy is more than just a dream—we’re making it...  ...pressing challenges of the 21st century.We are looking for a Senior Staff Engineer, Mechanical Design, to join our team in one of today’s most... 
    Full time
    Work at office
    Worldwide
    Overseas

    Bloom Energy

    San Jose, CA
    7 hours ago
  • At Bloom Energy, our vision for a world powered by clean, reliable, and affordable energy is more than just a dream—we’re making...  ...of the 21st century.  We are looking for a  Sr. Staff Mechanical Engineer  to join our team in one of today’s most exciting technologies... 
    Full time
    Work at office
    Worldwide

    Bloom Energy

    San Jose, CA
    1 day ago
  • $217.63k - $228.51k

     ...ownership of hardware, from specification to design, prototyping, reliability validation and manufacturing. Requirements capture, topology...  ...Education Requirements: Master’s degree in Electrical Engineering or Power Electronics.Minimum Experience Requirements: 48 months... 
    Local area

    Archer Aviation

    San Jose, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Reliability Engineer. Be the first to apply!