Staff Reliability Engineer
$166.5k - $291.4kServiceNow
Company Description It all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker, Phyllis. She cried tears of joy. That moment inspired Fred to build a company that could do that for everyone—freeing people from busywork so they could focus on meaningful work. Today, ServiceNow is the AI control tower for business reinvention. Our ServiceNow AI platform brings together any AI, any data, and any workflow— helping 85% of the Fortune 500® work smarter, faster, and better. We're building an AI-native culture where technology and talent are unstoppable together. And we're just getting started. Join us to put AI to work for people. Job Description Join us to build the next generation of cloud-native reliability, release, and test platforms that enable engineering excellence, developer productivity, and high-confidence ServiceNow releases through automation, observability, and AI-driven operations. What you get to do in this role: Design, build, and operate cloud-native engineering platforms for software validation, release validation, and production readiness Design and maintain production-like release and test ServiceNow environments that improve release confidence and deployment readiness. Build and integrate automated test pipelines, observability, reliability signals, deployment intelligence, and quality gates into CI/CD workflows. Develop automation solutions that improve engineering productivity, streamline operations, and reduce manual toil through shift-left engineering practices. Build reusable frameworks, self-service engineering environments, test data management, mock services, and developer productivity tooling. Design and enhance Kubernetes-based platforms supporting scalable test infrastructure, release automation, cloud-native workloads, and developer self-service. Implement automated validation for failure detection, deployment verification, policy enforcement, security checks, resilience testing, and operational health assessments. Resolve complex platforms, infrastructure, and networking challenges through software engineering, systems design, and automation. Partner closely with engineering teams to improve platform reliability, release quality, cloud-native adoption, and engineering best practices. Participate in architecture reviews, technical design discussions, and implementation of scalable, automation-first engineering solutions. Influence technical decisions through strong engineering execution, collaboration, and delivery of high-quality platform capabilities. Mentor engineers through technical guidance, code reviews, knowledge sharing, and engineering best practices. Foster a culture of reliability, automation, operational excellence, continuous improvement, and customer-focused engineering. Qualifications To be successful in this role you have: Experience in leveraging or critically thinking about how to integrate AI into work processes, decision-making, or problem-solving. This may include using AI-powered tools, automating workflows, analyzing AI-driven insights, or exploring AI's potential impact on the function or industry. 8+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Software Engineering, or Infrastructure Engineering with a Bachelor's degree; or 6 years and a Master's degree; or a PhD with 3 years experience; or equivalent experience. Hands-on experience with Kubernetes across cluster operations, networking, storage, security, autoscaling, and multi-cluster environments. Experience building and operating cloud-native platforms supporting scalable, highly available services. Experience integrating Kubernetes with CI/CD, GitOps, automated test pipelines, deployment validation, and cloud-native deployment workflows. Experience designing and implementing automation to improve developer productivity, release quality, and operational efficiency. Experience with progressive delivery practices, including canary deployments, feature flags, automated rollback, and deployment verification. Experience with chaos engineering, resilience testing, disaster recovery, and reliability validation. Strong software engineering skills with hands-on experience designing, developing, testing, and debugging applications using Python, Go, Java, or Ruby. Experience leveraging AI-assisted engineering for intelligent testing, release risk analysis, incident diagnostics, or operational automation is a plus. Strong understanding of observability, monitoring, SLI/SLOs, incident management, and production operations for distributed systems. Demonstrated ability to solve complex technical problems, drive projects independently, and collaborate effectively across engineering teams. Thrives in fast-paced, ambiguous environments with a strong ownership mindset, bias for action, and a passion for continuous learning and automation. Low ego, intellectually curious, and an effective collaborator who enjoys partnering with globally distributed teams to deliver reliable engineering solutions. Good to have: Experience with observability and monitoring platforms for applications, services, and distributed systems at scale. Experience with DevOps automation, CI/CD pipelines, GitOps, and Agile development practices using tools such as GitLab CI/CD, Argo CD, or Flux. Experience building and maintaining enterprise-scale test automation frameworks using technologies such as Playwright, Selenium, Cypress, REST Assured, PyTest, JUnit/TestNG, or equivalent. Experience with test orchestration, intelligent regression testing, test impact analysis, flaky test detection, parallel execution, and test data management. Experience with service virtualization, contract testing, synthetic testing, and building developer self-service engineering platforms. Experience with Infrastructure as Code and configuration management tools such as Ansible, Terraform, or equivalent. Experience with the Kubernetes ecosystem, including Helm, Argo Workflows, Kustomize, Istio/Linkerd, Gateway API/Ingress, Prometheus, OpenTelemetry, and container runtime technologies. Experience operating Kubernetes platforms across public cloud providers, including AWS (EKS), Azure (AKS), and Google Cloud (GKE). Experience implementing progressive delivery practices, including canary deployments, feature flags, deployment verification, and automated rollback. Familiarity with AI-assisted engineering, intelligent testing, operational automation, or cloud-native engineering platforms. For positions in this location, we offer a base pay of $166,500 - $291,400, plus equity (when applicable), variable/incentive compensation and benefits. Sales positions generally offer a competitive On Target Earnings (OTE) incentive compensation structure. Please note that the base pay shown is a guideline, and individual total compensation will vary based on factors such as qualifications, skill level, competencies, and work location. We also offer health plans, including flexible spending accounts, a 401(k) Plan with company match, ESPP, matching donations, a flexible time away plan and family leave programs. Compensation is based on the geographic location in which the role is located and is subject to change based on work location. Additional Information Work Personas We approach our distributed world of work with flexibility and trust. Work personas (flexible, remote, or required in office) are categories that are assigned to ServiceNow employees depending on the nature of their work and their assigned work location. Learn more here. To determine eligibility for a work persona, ServiceNow may confirm the distance between your primary residence and the closest ServiceNow office using a third-party service. Equal Opportunity Employer ServiceNow is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, national origin, age, disability, gender identity, veteran status, or any other category protected by law. In addition, all qualified applicants with arrest or conviction records will be considered for employment in accordance with legal requirements. Accommodations We strive to create an accessible and inclusive experience for all candidates. If you require a reasonable accommodation to complete any part of the application process, or are unable to use this online application and need an alternative method to apply, please contact [email protected] for assistance. Export Control Regulations For positions requiring access to controlled technology subject to export control regulations, including the U.S. Export Administration Regulations (EAR), ServiceNow may be required to obtain export control approval from government authorities for certain individuals. All employment is contingent upon ServiceNow obtaining any export license or other approval that may be required by relevant export control authorities. From Fortune. ©2026 Fortune Media IP Limited. All rights reserved. Used under license. Employee Type: Regular Region: AMS - North America and Canada Work Persona: Flexible or Remote
$235k - $275k
...Inc. as one of the most innovative and fastest-growing technology companies in the country. \n Role Summary As a Staff Site Reliability Engineer at Filevine, you are the senior technical authority on the SRE team and a strategic partner to engineering leadership....SuggestedPermanent employmentFull timeTemporary work- ...We're looking for a Staff Reliability Engineer for ICON. This role will serve as technical resource on reliability for our fleet of Vulcan and Titan print systems. This is a high-impact individual contributor role with org-wide reach: you'll diagnose failure patterns...SuggestedFor contractors
$220k - $255k
...and SUV miles with ones on vehicles that are more affordable, more enjoyable and 10-50x more efficient. ALSO is looking for a Reliability Engineer to play a key role in developing and leading the reliability of multiple electric mobility vehicles over the full...SuggestedFull timeLocal areaFlexible hours$165k - $218k
...computer vision, sensor fusion, and networking technology to the military in months, not years. ABOUT THE TEAM The Reliability Engineering team partners across Anduril's engineering, manufacturing, and operations organizations to ensure our autonomous systems survive...SuggestedFull timeWork experience placementImmediate start- ...technology designed to keep the electric grid secure and reliable, even during extended periods of stress. By strengthening... ...place. Role Description Form Energy is seeking a Staff Reliability Engineer to improve the reliability and lifetime of our iron-air battery...SuggestedFull timeRelocation package
- ...Senior Reliability Engineer The Senior Reliability Engineer position is responsible for improving equipment reliability, asset performance, and long-term equipment health across the facility. This role integrates engineering principles, data analytics, maintenance...
- ...Building cloud-native reliability and automation solutions, the full-time Staff Reliability Engineer will design and operate engineering platforms for software validation and release, while working flexibly or remotely to enhance developer productivity and deployment...Full timeRemote work
$148.9k - $260.6k
...applications, data, cloud environments, and AI agents. For engineers joining Veza today, this means the scale and resources of an... ...moment in the industry. We are seeking an exceptional Staff Site Reliability Engineer to lead critical infrastructure initiatives and drive...Work at officeRemote workFlexible hours- ...Arkema's Beaumont Plant is seeking a knowledgeable Staff Reliability Engineer to support rotating equipment performance, strengthen predictive maintenance programs, and contribute to long-term equipment reliability across the site. This role plays a key part in improving...For contractorsLocal area
$190k - $235k
...and highly collaborative team that is passionate about creating transformative change in healthcare. We are seeking a Staff Site Reliability Engineer to play a key role in designing, optimizing, and securing the underlying cloud infrastructure that powers our organization...Flexible hours$112.5k - $187.5k
...TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering... ...the most complex and consequential work on the platform. As a Staff Site Reliability Engineer at TransUnion, you will serve as a...Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$128.5k - $214.1k
...We're looking for a Staff Site Reliability Engineer to join our team, focusing on the core systems that power global financial markets. This isn't just about keeping the lights on; it's about pioneering the future of financial technology. As a member of our Clearing department...Work at officeWorldwide2 days per week$131.6k - $210.3k
...We are seeking a highly skilled and experiencedStaff Site Reliability Engineerto join our DevOps squad. This role will focus on leading... ...IaC), and cloud technologies, as well as the ability to mentor engineers and contribute to the overall stability and scalability of our...Work experience placementWork at officeLocal areaRemote work- ...Required U.S. Citizenship / No clearance needed / 100% remote within the US Staff Site Reliability Engineer / Cloud SME Location: 100% remote in the continental US Type: Long-term contract (3+ years) Role Summary As the Staff SRE/Cloud SME, you will be...Long term contractRemote work
$220k - $235k
...Staff/Senior Staff Site Reliability Engineer Ironclad is the leading AI contracting platform that transforms agreements into assets. Contracts move faster, insights surface instantly, and agents push work forward, all with you in control. Whether you're buying or selling...Full timeContract workWork at office$181k - $263k
...and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Staff Site Reliability Engineer who will set the technical direction for reliability engineering across LiveRamp's global infrastructure. This is a...Work from homeFlexible hoursNight shift$160k - $225k
...tens of thousands of users across hundreds of organizations globally. About the Role Manifold is looking for a Staff Site Reliability Engineer (SRE) to work at the intersection of AI, data infrastructure, and life sciences. In this high-impact role, you will help...$200k - $260k
...Join the Cloud Infrastructure Team as a technical leader driving reliability, automation, and scalability across the systems running Sight... ...and reliability practices across teams, mentor senior engineers, and be a primary escalation point for the org's hardest systems...Casual workWork at officeRemote workFlexible hours$170k - $210k
...businesses think about cybersecurity, digital experiences, and identity and access management. As a Ping Identity Site Reliability Engineer, you will be involved in every facet of our Cloud-based services. You will establish solutions for building, deploying, and...Local areaRemote workWorldwideFlexible hours$148k - $193k
...Staff Site Reliability Engineer Denver, CO, USA DAT is an award-winning employer of choice and a next-generation SaaS technology company that has been at the leading edge of innovation in transportation supply chain logistics for 45 years. We continue to transform...Temporary workWork experience placementWork at officeLocal areaImmediate startFlexible hours$131k - $164k
...Staff Site Reliability Engineer New York, New York, United States Position Overview We are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across VMware, Linux, and automation frameworks, to join our global Infrastructure...Work at officeLocal areaVisa sponsorshipFlexible hours- ...Series B and have grown 800% over the last 12 months. Engineering at Ivo Engineers at Ivo are inventors. Ivo was first-to... ...to hit our SLAs. What? We're looking for a Senior or Staff Site level Reliability Engineer as part of Infrastructure team to: Own uptime...Contract workWork at officeRemote workVisa sponsorshipRelocation packageFlexible hours
$194k - $267k
...more than once, automate it" and who can rapidly self-educate on new concepts and tools. Position Overview: The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$217.57k - $260k
...AI can be appropriately used during your application and interviews which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure team responsible for owning environments with thousands...Full timeTemporary workWork at officeRemote workFlexible hoursShift work- ...contributes. No one coasts. If you're driven by impact, pace, and raising the bar. This is the place. The Role As a Staff Site Reliability Engineer you'll play a lead role on the founding SRE team at our new NYC engineering hub. You'll own multi-team reliability and...Work at office
$102.6k - $193.43k
...Software Engineer Chamberlain Group (CG) is a global leader in intelligent access and Blackstone portfolio company. Powered by our... ...Establish working relationships with business members, technical staff, and subject matter experts to help create and execute the infrastructure...Temporary workSecond jobWorldwide- ...Job Description Job Description Staff Reliability Engineer Company Overview Embark on an enriching journey with PROCEPT BioRobotics, where our vision, mission, and values guide everything we do as a company. At PROCEPT, we put the patient first in everything we...Temporary workLocal areaFlexible hours
$180k - $230k
...Job Description Job Description Job Title: Staff Reliability Engineer Location: Burlingame, CA Department: ESS Engineering Reports To: Staff Reliability Engineer Position Type: Full-time About Peak Energy Peak Energy is the first American...Full timeImmediate startFlexible hours- ...worldwide. For more information, visit Follow Shield AI on LinkedIn, X, Instagram, and YouTube. Job Description: The Reliability and Maintainability Engineer is a senior individual contributor responsible for driving investigation, root cause analysis, corrective action,...Full timeTemporary workPart timeWorldwide
- ...Description Job Description Position Overview An innovative and growing medical technology company is seeking a Reliability Quality Engineer to ensure the product reliability and quality of a complex, next-generation medical device ecosystem. In this permanent...Permanent employmentFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Reliability Engineer. Be the first to apply!
- staff security engineer United States
- project engineer assistant project manager United States
- assistant chief engineer United States
- staff data engineer United States
- assistant building engineer United States
- senior staff engineer United States
- staff design engineer United States
- staff process engineer United States
- structural engineering assistant United States
- assistant mechanical engineer United States

