Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Reliability Engineer

$166.5k - $291.4k

ServiceNow

Staff Reliability EngineerJoin us to build the next generation of cloud-native reliability, release, and test platforms that enable engineering excellence, developer productivity, and high-confidence ServiceNow releases through automation, observability, and AI-driven operations.What you get to do in this role:Design, build, and operate cloud-native engineering platforms for software validation, release validation, and production readinessDesign and maintain production-like release and test ServiceNow environments that improve release confidence and deployment readiness.Build and integrate automated test pipelines, observability, reliability signals, deployment intelligence, and quality gates into CI/CD workflows.Develop automation solutions that improve engineering productivity, streamline operations, and reduce manual toil through shift-left engineering practices.Build reusable frameworks, self-service engineering environments, test data management, mock services, and developer productivity tooling.Design and enhance Kubernetes-based platforms supporting scalable test infrastructure, release automation, cloud-native workloads, and developer self-service.Implement automated validation for failure detection, deployment verification, policy enforcement, security checks, resilience testing, and operational health assessments.Resolve complex platforms, infrastructure, and networking challenges through software engineering, systems design, and automation.Partner closely with engineering teams to improve platform reliability, release quality, cloud-native adoption, and engineering best practices.Participate in architecture reviews, technical design discussions, and implementation of scalable, automation-first engineering solutions.Influence technical decisions through strong engineering execution, collaboration, and delivery of high-quality platform capabilities.Mentor engineers through technical guidance, code reviews, knowledge sharing, and engineering best practices.Foster a culture of reliability, automation, operational excellence, continuous improvement, and customer-focused engineering.To be successful in this role you have:Experience in leveraging or critically thinking about how to integrate AI into work processes, decision-making, or problem-solving. This may include using AI-powered tools, automating workflows, analyzing AI-driven insights, or exploring AI's potential impact on the function or industry.8+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Software Engineering, or Infrastructure Engineering with a Bachelor's degree; or 6 years and a Master's degree; or a PhD with 3 years experience; or equivalent experience.Hands-on experience with Kubernetes across cluster operations, networking, storage, security, autoscaling, and multi-cluster environments.Experience building and operating cloud-native platforms supporting scalable, highly available services.Experience integrating Kubernetes with CI/CD, GitOps, automated test pipelines, deployment validation, and cloud-native deployment workflows.Experience designing and implementing automation to improve developer productivity, release quality, and operational efficiency.Experience with progressive delivery practices, including canary deployments, feature flags, automated rollback, and deployment verification.Experience with chaos engineering, resilience testing, disaster recovery, and reliability validation.Strong software engineering skills with hands-on experience designing, developing, testing, and debugging applications using Python, Go, Java, or Ruby.Experience leveraging AI-assisted engineering for intelligent testing, release risk analysis, incident diagnostics, or operational automation is a plus.Strong understanding of observability, monitoring, SLI/SLOs, incident management, and production operations for distributed systems.Demonstrated ability to solve complex technical problems, drive projects independently, and collaborate effectively across engineering teams.Thrives in fast-paced, ambiguous environments with a strong ownership mindset, bias for action, and a passion for continuous learning and automation.Low ego, intellectually curious, and an effective collaborator who enjoys partnering with globally distributed teams to deliver reliable engineering solutions.Good to have:Experience with observability and monitoring platforms for applications, services, and distributed systems at scale.Experience with DevOps automation, CI/CD pipelines, GitOps, and Agile development practices using tools such as GitLab CI/CD, Argo CD, or Flux.Experience building and maintaining enterprise-scale test automation frameworks using technologies such as Playwright, Selenium, Cypress, REST Assured, PyTest, JUnit/TestNG, or equivalent.Experience with test orchestration, intelligent regression testing, test impact analysis, flaky test detection, parallel execution, and test data management.Experience with service virtualization, contract testing, synthetic testing, and building developer self-service engineering platforms.Experience with Infrastructure as Code and configuration management tools such as Ansible, Terraform, or equivalent.Experience with the Kubernetes ecosystem, including Helm, Argo Workflows, Kustomize, Istio/Linkerd, Gateway API/Ingress, Prometheus, OpenTelemetry, and container runtime technologies.Experience operating Kubernetes platforms across public cloud providers, including AWS (EKS), Azure (AKS), and Google Cloud (GKE).Experience implementing progressive delivery practices, including canary deployments, feature flags, deployment verification, and automated rollback.Familiarity with AI-assisted engineering, intelligent testing, operational automation, or cloud-native engineering platforms.For positions in this location, we offer a base pay of $166,500 - $291,400, plus equity (when applicable), variable/incentive compensation and benefits. Sales positions generally offer a competitive On Target Earnings (OTE) incentive compensation structure. Please note that the base pay shown is a guideline, and individual total compensation will vary based on factors such as qualifications, skill level, competencies, and work location. We also offer health plans, including flexible spending accounts, a 401(k) Plan with company match, ESPP, matching donations, a flexible time away plan and family leave programs. Compensation is based on the geographic location in which the role is located and is subject to change based on work location.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Reliability Engineer in Santa Clara, CA vacancy
  • $130.17k - $188.49k

     ...software technologies into solutions that combat climate change, reliably connect humans and the world, and help drive advancements in...  ...on LinkedIn and X.Analog Devices, Inc. (ADI) is seeking a Staff Engineer to join our Reliability Hardware & Systems Development team in... 
    Suggested
    Permanent employment
    Full time
    Work at office
    Day shift

    Analog Devices

    San Jose, CA
    4 days ago
  • At Bloom Energy, our vision for a world powered by clean, reliable, and affordable energy is more than just a dream—we’re making...  ...challenges of the 21st century.We are looking for a Staff Reliability Engineer to join our team in one of today’s most exciting technologies... 
    Suggested
    Full time
    Work experience placement
    Remote work
    Worldwide

    Bloom Energy

    San Jose, CA
    6 days ago
  •  ...Staff Reliability EngineerTekfortune is a fast-growing consulting firm specialized in permanent, contract & project-based staffing services...  ...exposure.Job DescriptionAs a Staff Reliability Quality Engineer to ensure product reliability and quality within a complex medical... 
    Suggested
    Permanent employment
    Full time
    Contract work
    Remote work

    Tekfortune Inc

    San Jose, CA
    3 days ago
  • $188k - $274k

     ...needed.Minimum qualifications:Bachelor’s degree in Electrical Engineering, Computer Engineering, Computer Science, Physics, or a...  ...practical experience.6 years of experience working in a hardware reliability technical environment.3 years of experience in technical leadership... 
    Suggested
    Worldwide
    Flexible hours

    Google

    San Jose, CA
    4 days ago
  • $122.5k - $175k

     ...age, we invite you to bring your talents to Zscaler and help shape the future of cybersecurity.RoleWe are looking for a Staff Site Reliability Engineer to join our team. This is a hybrid role going into the San Jose, CA office 3 days a week, reporting to the Chief... 
    Suggested
    Full time
    Work at office
    Local area
    3 days per week

    Zscaler

    San Jose, CA
    3 days ago
  • $192.4k - $275.8k

     ...CloudOps— the team that keeps Splunk Cloud running for some of the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines at a scale very few teams ever get to operate at. When the... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    2 days ago
  • $207k - $300k

     ...areas within SU SRE, mentoring team members to enhance system reliability and efficiency.Initiate, own, and lead large-scale, complex projects...  ...techniques.3 years of experience as a Site Reliability Engineer.3 years of experience leading projects.3 years of experience designing... 

    Google

    San Jose, CA
    2 days ago
  • $200k - $322k

     ....We are seeking a highly skilled Senior Staff SRE to join our dynamic team. Our company...  ...includes building for performance and reliability at global scale, covering automation, monitoring...  ...with NVIDIA leadership, senior engineers, program managers, and product managers... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    6 days ago
  •  ...AI inference services, powered by the Wafer-Scale Engine (WSE). This team will help deliver world-class, ultra-reliable inference infrastructure for leading model builders such as OpenAI and other frontier labs.As a Staff SRE, you will lead the engineering effort to... 
    Shift work

    Cerebras Systems

    Sunnyvale, CA
    6 days ago
  • $220k - $255k

     ...and SUV miles with ones on vehicles that are more affordable, more enjoyable and 10-50x more efficient. ALSO is looking for a Reliability Engineer to play a key role in developing and leading the reliability of multiple electric mobility vehicles over the full... 
    Full time
    Local area
    Flexible hours

    ALSO

    Palo Alto, CA
    3 days ago
  • $181k - $262k

     ...get crucial goods where they need to go, and make mobility more efficient and accessible for all.We’re searching for a Staff Hardware Reliability Engineer.The Hardware Reliability team is dedicated to ensuring the robustness and dependability of hardware systems in the... 
    Contract work
    Work at office
    Local area
    3 days per week

    Aurora Innovation

    Mountain View, CA
    6 days ago
  • $165k - $192.5k

    Staff Reliability Engineer - Power Electronics Reliability Engineers at Lunar Energy will be responsible for ensuring product reliability throughout the entire lifecycle of our revolutionary home energy products. This includes providing input during the design phases,... 
    Full time

    Lunar Energy

    Mountain View, CA
    2 days ago
  • $262k - $364k

    Ensure that Google Ads services within the AViD ecosystem have reliability and uptime appropriate to users' needs with a fast rate of...  ...ever-watchful eye on capacity and performance.Build creative engineering solutions to operations and infrastructure problems, including... 

    Google

    Mountain View, CA
    6 days ago
  • $160k - $235k

     ...these cities preferred - no relocation) This role sits within our Low-Voltage GaN Marketing & Systems Group, acting as a Sr Staff System Engineer to define the future of our power semiconductor portfolio. In this strategic role, you will not simply support existing... 
    Full time
    Temporary work
    Local area
    Relocation
    Visa sponsorship
    Work visa
    Flexible hours
    Shift work

    Renesas Electronics

    San Jose, CA
    20 days ago
  • $189k - $301k

     ...solution to be the best in class. Systems team within the Location group is looking for an experienced Systems/DSP engineer at the level of Staff/Senior Staff. The candidate will work as part of the Location team responsible for receiver model development, receiver... 
    Flexible hours

    Samsung Semiconductor

    San Jose, CA
    23 days ago
  • $168k - $264.5k

     ...industry. As the computational power increases with every GPU generation, developing efficient and reliable systems is an imperative. We are looking for a System Reliability Engineer to join NVIDIA's existing Reliability Engineering team, involved in NVIDIA's diverse system... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $168k - $264.5k

     ...human inventiveness and intelligence. Make the choice to join us today. We are seeking an outstanding candidate for Silicon Reliability Engineer to drive and utilize cutting edge technologies to deliver high performance products while ensuring world class reliability.What... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $154k - $200k

     ...Job Description Antora Energy delivers affordable, reliable energy to industry, data centers, and the grid. Our thermal...  ...and power at scale. We are looking for a  Senior / Staff Turbomachinery Systems Engineer to lead the definition, integration, and execution of... 
    Contract work
    Flexible hours
    Shift work

    Antora Energy

    San Jose, CA
    22 days ago
  • $116k - $184k

     ...profoundly impacting society. Come join the team and help build the next era of computing!We're seeking an outstanding Senior HTOL Reliability Engineer to join our Santa Clara lab. This role requires deep device-circuitry knowledge and hands-on hardware development. You will... 
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • $109.94k - $151.17k

     ...software technologies into solutions that combat climate change, reliably connect humans and the world, and help drive advancements in...  ...Devices, Inc. (ADI) is seeking a talented and motivated Reliability Engineer to join our Product Reliability Engineering team in San Jose,... 
    Permanent employment
    Full time
    Temporary work
    Work at office
    Day shift

    Analog Devices

    San Jose, CA
    2 days ago
  • $100k - $136.5k

     ...AreApplied Materials is the global leader in materials science and engineering solutions that are at the foundation of virtually every new...  ...damage or injuryAssist in preparing & presenting safety and reliability reports;Assist with root cause analysis of field failures and... 
    Full time

    Applied Materials

    Santa Clara, CA
    2 days ago
  • At Bloom Energy, our vision for a world powered by clean, reliable, and affordable energy is more than just a dream—we’re making it...  ...pressing challenges of the 21st century.We are looking for a Senior Staff Engineer, Mechanical Design, to join our team in one of today’s most... 
    Full time
    Work at office
    Worldwide
    Overseas

    Bloom Energy

    San Jose, CA
    6 days ago
  •  ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by...  ...Function of Position The Staff Mechanical Design Engineer will play a leading...  ...solutionsPerform design FMEADesign for reliability, cost and manufacturabilityExercise creativity... 
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    3 days ago
  •  ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by...  ...of the PositionWe are seeking a Staff Electrical Engineer to contribute towards...  ...designs based on DFM/DFA/DFT, serviceability, reliability, longevity, and performance.Develop and... 
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    6 days ago
  • $209k - $219.45k

     ...detailed design of parts and subassemblies, technical drawing creation, and control of space allocation and routing envelopes. Conduct Engineering Analysis including thermal modeling, structural analysis, and aircraft level weight and balance modeling to validate packaging... 
    Local area

    Archer Aviation

    San Jose, CA
    5 days ago
  • At Bloom Energy, our vision for a world powered by clean, reliable, and affordable energy is more than just a dream—we’re making...  ...of the 21st century.  We are looking for a  Sr. Staff Mechanical Engineer  to join our team in one of today’s most exciting technologies... 
    Full time
    Work at office
    Worldwide

    Bloom Energy

    San Jose, CA
    6 days ago
  • $163k - $253k

     ...and communities.Samsung Semiconductor is hiring for a Packaging Engineer, Mechanical Simulation role to lead the development and design...  ...and simulation to evaluate and improve the performance and reliability of cutting‑edge semiconductor packages and memory devices.Provide... 
    Flexible hours

    Samsung Semiconductor

    San Jose, CA
    2 days ago
  •  ...we advance your career. THE ROLE:We are looking for a Quality Engineer II to support component- and board-level failure analysis for...  ...fault isolation, physical failure analysis, board-level analysis, reliability stress support, and clear technical reporting. This is a hands... 

    AMD

    San Jose, CA
    4 days ago
  • $147k - $202.5k

    Who We AreApplied Materials is a global leader in materials engineering solutions used to produce virtually every new chip and advanced display in the world. We design, build and service cutting-edge equipment that helps our customers manufacture display and semiconductor... 
    Full time

    Applied Materials

    Santa Clara, CA
    4 days ago
  • $189k - $301k

     ...building a better tomorrow for our employees, customers, partners, and communities. About the Role We are seeking a Senior Staff Engineer to build and optimize the EDA design environment and large-scale compute infrastructure that supports our semiconductor design... 
    Contract work
    Work at office
    Flexible hours

    Samsung Semiconductor

    San Jose, CA
    29 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Reliability Engineer. Be the first to apply!