Staff Reliability Engineer
$166.5k - $291.4kServiceNow
Company DescriptionIt all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker, Phyllis. She cried tears of joy. That moment inspired Fred to build a company that could do that for everyone—freeing people from busywork so they could focus on meaningful work. Today, ServiceNow is the AI control tower for business reinvention. Our ServiceNow AI platform brings together any AI, any data, and any workflow— helping 85% of the Fortune 500 work smarter, faster, and better. We're building an AI-native culture where technology and talent are unstoppable together. And we're just getting started.Join us to put AI to work for people.Job DescriptionJoin us to build the next generation of cloud-native reliability, release, and test platforms that enable engineering excellence, developer productivity, and high-confidence ServiceNow releases through automation, observability, and AI-driven operations.What you get to do in this role:Design, build, and operate cloud-native engineering platforms for software validation, release validation, and production readinessDesign and maintain production-like release and test ServiceNow environments that improve release confidence and deployment readiness.Build and integrate automated test pipelines, observability, reliability signals, deployment intelligence, and quality gates into CI/CD workflows.Develop automation solutions that improve engineering productivity, streamline operations, and reduce manual toil through shift-left engineering practices.Build reusable frameworks, self-service engineering environments, test data management, mock services, and developer productivity tooling.Design and enhance Kubernetes-based platforms supporting scalable test infrastructure, release automation, cloud-native workloads, and developer self-service.Implement automated validation for failure detection, deployment verification, policy enforcement, security checks, resilience testing, and operational health assessments.Resolve complex platforms, infrastructure, and networking challenges through software engineering, systems design, and automation.Partner closely with engineering teams to improve platform reliability, release quality, cloud-native adoption, and engineering best practices.Participate in architecture reviews, technical design discussions, and implementation of scalable, automation-first engineering solutions.Influence technical decisions through strong engineering execution, collaboration, and delivery of high-quality platform capabilities.Mentor engineers through technical guidance, code reviews, knowledge sharing, and engineering best practices.Foster a culture of reliability, automation, operational excellence, continuous improvement, and customer-focused engineering.QualificationsTo be successful in this role you have:Experience in leveraging or critically thinking about how to integrate AI into work processes, decision-making, or problem-solving. This may include using AI-powered tools, automating workflows, analyzing AI-driven insights, or exploring AI's potential impact on the function or industry.8+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Software Engineering, or Infrastructure Engineering with a Bachelor's degree; or 6 years and a Master's degree; or a PhD with 3 years experience; or equivalent experience.Hands-on experience with Kubernetes across cluster operations, networking, storage, security, autoscaling, and multi-cluster environments.Experience building and operating cloud-native platforms supporting scalable, highly available services.Experience integrating Kubernetes with CI/CD, GitOps, automated test pipelines, deployment validation, and cloud-native deployment workflows.Experience designing and implementing automation to improve developer productivity, release quality, and operational efficiency.Experience with progressive delivery practices, including canary deployments, feature flags, automated rollback, and deployment verification.Experience with chaos engineering, resilience testing, disaster recovery, and reliability validation.Strong software engineering skills with hands-on experience designing, developing, testing, and debugging applications using Python, Go, Java, or Ruby.Experience leveraging AI-assisted engineering for intelligent testing, release risk analysis, incident diagnostics, or operational automation is a plus.Strong understanding of observability, monitoring, SLI/SLOs, incident management, and production operations for distributed systems.Demonstrated ability to solve complex technical problems, drive projects independently, and collaborate effectively across engineering teams.Thrives in fast-paced, ambiguous environments with a strong ownership mindset, bias for action, and a passion for continuous learning and automation.Low ego, intellectually curious, and an effective collaborator who enjoys partnering with globally distributed teams to deliver reliable engineering solutions.Good to have:Experience with observability and monitoring platforms for applications, services, and distributed systems at scale.Experience with DevOps automation, CI/CD pipelines, GitOps, and Agile development practices using tools such as GitLab CI/CD, Argo CD, or Flux.Experience building and maintaining enterprise-scale test automation frameworks using technologies such as Playwright, Selenium, Cypress, REST Assured, PyTest, JUnit/TestNG, or equivalent.Experience with test orchestration, intelligent regression testing, test impact analysis, flaky test detection, parallel execution, and test data management.Experience with service virtualization, contract testing, synthetic testing, and building developer self-service engineering platforms.Experience with Infrastructure as Code and configuration management tools such as Ansible, Terraform, or equivalent.Experience with the Kubernetes ecosystem, including Helm, Argo Workflows, Kustomize, Istio/Linkerd, Gateway API/Ingress, Prometheus, OpenTelemetry, and container runtime technologies.Experience operating Kubernetes platforms across public cloud providers, including AWS (EKS), Azure (AKS), and Google Cloud (GKE).Experience implementing progressive delivery practices, including canary deployments, feature flags, deployment verification, and automated rollback.Familiarity with AI-assisted engineering, intelligent testing, operational automation, or cloud-native engineering platforms.For positions in this location, we offer a base pay of $166,500 - $291,400, plus equity (when applicable), variable/incentive compensation and benefits. Sales positions generally offer a competitive On Target Earnings (OTE) incentive compensation structure. Please note that the base pay shown is a guideline, and individual total compensation will vary based on factors such as qualifications, skill level, competencies, and work location. We also offer health plans, including flexible spending accounts, a 401(k) Plan with company match, ESPP, matching donations, a flexible time away plan and family leave programs. Compensation is based on the geographic location in which the role is located and is subject to change based on work location.Additional InformationWork PersonasWe approach our distributed world of work with flexibility and trust. Work personas (flexible, remote, or required in office) are categories that are assigned to ServiceNow employees depending on the nature of their work and their assigned work location. Learn more here. To determine eligibility for a work persona, ServiceNow may confirm the distance between your primary residence and the closest ServiceNow office using a third-party service.Equal Opportunity EmployerServiceNow is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, national origin, age, disability, gender identity, veteran status, or any other category protected by law. In addition, all qualified applicants with arrest or conviction records will be considered for employment in accordance with legal requirements. AccommodationsWe strive to create an accessible and inclusive experience for all candidates. If you require a reasonable accommodation to complete any part of the application process, or are unable to use this online application and need an alternative method to apply, please contact View email address on click.appcast.io for assistance. Export Control RegulationsFor positions requiring access to controlled technology subject to export control regulations, including the U.S. Export Administration Regulations (EAR), ServiceNow may be required to obtain export control approval from government authorities for certain individuals. All employment is contingent upon ServiceNow obtaining any export license or other approval that may be required by relevant export control authorities. From Fortune. 2026 Fortune Media IP Limited. All rights reserved. Used under license.SummaryType: Full-timeFunction: Information TechnologyExperience level: Not ApplicableIndustry: Information Technology And Services
$146.2k - $255.9k
...applications, data, cloud environments, and AI agents. For engineers joining Veza today, this means the scale and resources of an... ...industry. Job Description We are seeking an exceptional Staff Site Reliability Engineer to lead critical infrastructure initiatives and drive...SuggestedFull timeWork at officeRemote workFlexible hours- At Bloom Energy, our vision for a world powered by clean, reliable, and affordable energy is more than just a dream—we’re making... ...challenges of the 21st century.We are looking for a Staff Reliability Engineer to join our team in one of today’s most exciting technologies...SuggestedFull timeWork experience placementRemote workWorldwide
$207.4k - $259.2k
...differences, and supports and celebrates all of our team members.We are seeking a highly experienced and passionate Sr. Staff Site Reliability Engineer (SRE) to join our growing team. In this critical role, you will be responsible for the reliability, scalability,...SuggestedPermanent employmentLocal area$122.5k - $175k
...age, we invite you to bring your talents to Zscaler and help shape the future of cybersecurity.RoleWe are looking for a Staff Site Reliability Engineer to join our team. This is a hybrid role going into the San Jose, CA office 3 days a week, reporting to the Chief...SuggestedFull timeWork at officeLocal area3 days per week$207k - $301k
Develop strong, influential relationships with multiple stakeholders across the Site Reliability Engineering and Developer organizations.Serve as an expert on particular fields of knowledge related to rate limiting or sharding.Develop plans and lead projects on evolving...Suggested$207k - $301k
...design and lead the implementation of solutions to enhance the reliability of systems that support F1.Scale systems sustainably through... ...improve reliability for multiple teams.Engage in software engineering on services written in Java, C++, and Go (including instrumenting...$217.57k - $260k
...AI can be appropriately used during your application and interviews which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure team responsible for owning environments with thousands...Full timeTemporary workWork at officeRemote workFlexible hoursShift work$165k - $192.5k
Staff Reliability Engineer - Power Electronics Reliability Engineers at Lunar Energy will be responsible for ensuring product reliability throughout the entire lifecycle of our revolutionary home energy products. This includes providing input during the design phases,...Full time$220k - $255k
...and SUV miles with ones on vehicles that are more affordable, more enjoyable and 10-50x more efficient. ALSO is looking for a Reliability Engineer to play a key role in developing and leading the reliability of multiple electric mobility vehicles over the full...Full timeLocal areaFlexible hours$181k - $262k
...get crucial goods where they need to go, and make mobility more efficient and accessible for all.We’re searching for a Staff Hardware Reliability Engineer.The Hardware Reliability team is dedicated to ensuring the robustness and dependability of hardware systems in the...Contract workWork at officeLocal area3 days per week$181k - $262k
...get crucial goods where they need to go, and make mobility more efficient and accessible for all.We’re searching for a Staff Hardware Reliability Engineer who will join the Hardware Reliability team to ensure the robustness of the Aurora Driver by leading reliability...Contract workWork at officeLocal area3 days per week$232k - $263k
...scaling rapidly toward long-term growth and IPO readiness.Sr. Staff Site Reliability EngineerAs a Sr. Staff SRE at Obsidian, you will define... ...operate as a strategic partner to DevOps and Platform Engineering leadership, shaping a unified reliability strategy that scales...Work from home$122.44k - $232.19k
...workloads. With a startup-like culture, we move quickly and give engineers the opportunity to drive significant technical and business... ...chip development flow. Mission: Define and own the pod-level reliability specifications that ensure the availability, resilience, and...Full timeLocal areaImmediate startShift work$168k - $264.5k
...human inventiveness and intelligence. Make the choice to join us today. We are seeking an outstanding candidate for Silicon Reliability Engineer to drive and utilize cutting edge technologies to deliver high performance products while ensuring world class reliability.What...Full time$86.48k - $118.91k
...software technologies into solutions that combat climate change, reliably connect humans and the world, and help drive advancements in... ...your career and help innovate ahead of what’s possibleReliability Engineer will be responsible for:Project management of strategic...Permanent employmentFull timeWork at officeDay shift$116k - $184k
...profoundly impacting society. Come join the team and help build the next era of computing!We're seeking an outstanding Senior HTOL Reliability Engineer to join our Santa Clara lab. This role requires deep device-circuitry knowledge and hands-on hardware development. You will...Full time$168k - $264.5k
...industry. As the computational power increases with every GPU generation, developing efficient and reliable systems is an imperative. We are looking for a System Reliability Engineer to join NVIDIA's existing Reliability Engineering team, involved in NVIDIA's diverse system...Full time- ...of the most challenging global health issues.Play a meaningful role in the development of new life science technology as a reliability engineer in the R&D Impact Engineering group. Provide hands-on, analytical, and reliability expertise to support engineering programs...Full time
$135k - $185k
Req ID: 138673 Region: Americas Country: USA State/Province: California City: San Jose SummaryOur Engineering Team is seeking a highly motivated and self-driven candidate qualified for the position of Mechanical Engineering to be a part of a market leading Hardware Design...Work at office$118.01k - $176.8k
...learn, and lead. Your Team, Your ImpactWe are seeking a Sr. Staff Thermo Mechanical Engineer to join our Advanced Technology Infrastructure team, focusing on mechanical simulation and structural reliability of semiconductor test and validation infrastructure. This role...Permanent employmentFull timeInternshipWork from home- ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by... ...Function of Position The Staff Mechanical Design Engineer will play a leading... ...solutionsPerform design FMEADesign for reliability, cost and manufacturabilityExercise creativity...Local areaWorldwideFlexible hours
$209k - $219.45k
...detailed design of parts and subassemblies, technical drawing creation, and control of space allocation and routing envelopes. Conduct Engineering Analysis including thermal modeling, structural analysis, and aircraft level weight and balance modeling to validate packaging...Local area$163k - $253k
...and communities.Samsung Semiconductor is hiring for a Packaging Engineer, Mechanical Simulation role to lead the development and design... ...and simulation to evaluate and improve the performance and reliability of cutting‑edge semiconductor packages and memory devices.Provide...Flexible hours- ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by... ...purpose here.Job DescriptionManaging Staff, Systems Mechanical EngineeringJoin... ...and through validation to ensure robust, reliable products.Lead cross-functional projects,...Full timeLocal areaWorldwideRelocation packageFlexible hours
- ...people from diverse backgrounds and industries to create a safer, sustainable and more connected world. Job OverviewA Quality & Reliability Engineer Supervisor uses their engineering skills to assist in issues related to the Quality of the product. This job also involves...Remote work
- ...we advance your career. THE ROLE:We are looking for a Quality Engineer II to support component- and board-level failure analysis for... ...fault isolation, physical failure analysis, board-level analysis, reliability stress support, and clear technical reporting. This is a hands...
$147k - $202.5k
Who We AreApplied Materials is a global leader in materials engineering solutions used to produce virtually every new chip and advanced display in the world. We design, build and service cutting-edge equipment that helps our customers manufacture display and semiconductor...Full time- At Bloom Energy, our vision for a world powered by clean, reliable, and affordable energy is more than just a dream—we’re making it... ...pressing challenges of the 21st century.We are looking for a Senior Staff Engineer, Mechanical Design, to join our team in one of today’s most...Full timeWork at officeWorldwideOverseas
- At Bloom Energy, our vision for a world powered by clean, reliable, and affordable energy is more than just a dream—we’re making... ...of the 21st century. We are looking for a Sr. Staff Mechanical Engineer to join our team in one of today’s most exciting technologies...Full timeWork at officeWorldwide
$217.63k - $228.51k
...ownership of hardware, from specification to design, prototyping, reliability validation and manufacturing. Requirements capture, topology... ...Education Requirements: Master’s degree in Electrical Engineering or Power Electronics.Minimum Experience Requirements: 48 months...Local area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Reliability Engineer. Be the first to apply!
- engineering aide Santa Clara, CA
- technology administrator Santa Clara, CA
- senior staff engineer Santa Clara, CA
- staff engineer Santa Clara, CA
- senior staff systems engineer Santa Clara, CA
- assistant engineer Santa Clara, CA
- reliability engineer Santa Clara, CA
- staff devops engineer
- vessel assistant engineer
- staff automation engineer

