Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Reliability Engineer

$166.5k - $291.4k
Full-time

ServiceNow

Company Description It all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker, Phyllis. She cried tears of joy. That moment inspired Fred to build a company that could do that for everyone—freeing people from busywork so they could focus on meaningful work. Today, ServiceNow is the AI control tower for business reinvention. Our ServiceNow AI platform brings together any AI, any data, and any workflow— helping 85% of the Fortune 500® work smarter, faster, and better. We're building an AI-native culture where technology and talent are unstoppable together. And we're just getting started. Join us to put AI to work for people. Job Description Join us to build the next generation of cloud-native reliability, release, and test platforms that enable engineering excellence, developer productivity, and high-confidence ServiceNow releases through automation, observability, and AI-driven operations. What you get to do in this role: Design, build, and operate cloud-native engineering platforms for software validation, release validation, and production readiness Design and maintain production-like release and test ServiceNow environments that improve release confidence and deployment readiness. Build and integrate automated test pipelines, observability, reliability signals, deployment intelligence, and quality gates into CI/CD workflows. Develop automation solutions that improve engineering productivity, streamline operations, and reduce manual toil through shift-left engineering practices. Build reusable frameworks, self-service engineering environments, test data management, mock services, and developer productivity tooling. Design and enhance Kubernetes-based platforms supporting scalable test infrastructure, release automation, cloud-native workloads, and developer self-service. Implement automated validation for failure detection, deployment verification, policy enforcement, security checks, resilience testing, and operational health assessments. Resolve complex platforms, infrastructure, and networking challenges through software engineering, systems design, and automation. Partner closely with engineering teams to improve platform reliability, release quality, cloud-native adoption, and engineering best practices. Participate in architecture reviews, technical design discussions, and implementation of scalable, automation-first engineering solutions. Influence technical decisions through strong engineering execution, collaboration, and delivery of high-quality platform capabilities. Mentor engineers through technical guidance, code reviews, knowledge sharing, and engineering best practices. Foster a culture of reliability, automation, operational excellence, continuous improvement, and customer-focused engineering. Qualifications To be successful in this role you have: Experience in leveraging or critically thinking about how to integrate AI into work processes, decision-making, or problem-solving. This may include using AI-powered tools, automating workflows, analyzing AI-driven insights, or exploring AI's potential impact on the function or industry. 8+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Software Engineering, or Infrastructure Engineering with a Bachelor's degree; or 6 years and a Master's degree; or a PhD with 3 years experience; or equivalent experience. Hands-on experience with Kubernetes across cluster operations, networking, storage, security, autoscaling, and multi-cluster environments. Experience building and operating cloud-native platforms supporting scalable, highly available services. Experience integrating Kubernetes with CI/CD, GitOps, automated test pipelines, deployment validation, and cloud-native deployment workflows. Experience designing and implementing automation to improve developer productivity, release quality, and operational efficiency. Experience with progressive delivery practices, including canary deployments, feature flags, automated rollback, and deployment verification. Experience with chaos engineering, resilience testing, disaster recovery, and reliability validation. Strong software engineering skills with hands-on experience designing, developing, testing, and debugging applications using Python, Go, Java, or Ruby. Experience leveraging AI-assisted engineering for intelligent testing, release risk analysis, incident diagnostics, or operational automation is a plus. Strong understanding of observability, monitoring, SLI/SLOs, incident management, and production operations for distributed systems. Demonstrated ability to solve complex technical problems, drive projects independently, and collaborate effectively across engineering teams. Thrives in fast-paced, ambiguous environments with a strong ownership mindset, bias for action, and a passion for continuous learning and automation. Low ego, intellectually curious, and an effective collaborator who enjoys partnering with globally distributed teams to deliver reliable engineering solutions. Good to have: Experience with observability and monitoring platforms for applications, services, and distributed systems at scale. Experience with DevOps automation, CI/CD pipelines, GitOps, and Agile development practices using tools such as GitLab CI/CD, Argo CD, or Flux. Experience building and maintaining enterprise-scale test automation frameworks using technologies such as Playwright, Selenium, Cypress, REST Assured, PyTest, JUnit/TestNG, or equivalent. Experience with test orchestration, intelligent regression testing, test impact analysis, flaky test detection, parallel execution, and test data management. Experience with service virtualization, contract testing, synthetic testing, and building developer self-service engineering platforms. Experience with Infrastructure as Code and configuration management tools such as Ansible, Terraform, or equivalent. Experience with the Kubernetes ecosystem, including Helm, Argo Workflows, Kustomize, Istio/Linkerd, Gateway API/Ingress, Prometheus, OpenTelemetry, and container runtime technologies. Experience operating Kubernetes platforms across public cloud providers, including AWS (EKS), Azure (AKS), and Google Cloud (GKE). Experience implementing progressive delivery practices, including canary deployments, feature flags, deployment verification, and automated rollback. Familiarity with AI-assisted engineering, intelligent testing, operational automation, or cloud-native engineering platforms. For positions in this location, we offer a base pay of $166,500 - $291,400, plus equity (when applicable), variable/incentive compensation and benefits. Sales positions generally offer a competitive On Target Earnings (OTE) incentive compensation structure. Please note that the base pay shown is a guideline, and individual total compensation will vary based on factors such as qualifications, skill level, competencies, and work location. We also offer health plans, including flexible spending accounts, a 401(k) Plan with company match, ESPP, matching donations, a flexible time away plan and family leave programs. Compensation is based on the geographic location in which the role is located and is subject to change based on work location. Additional Information Work Personas We approach our distributed world of work with flexibility and trust. Work personas (flexible, remote, or required in office) are categories that are assigned to ServiceNow employees depending on the nature of their work and their assigned work location. Learn more here. To determine eligibility for a work persona, ServiceNow may confirm the distance between your primary residence and the closest ServiceNow office using a third-party service. Equal Opportunity Employer ServiceNow is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, national origin, age, disability, gender identity, veteran status, or any other category protected by law. In addition, all qualified applicants with arrest or conviction records will be considered for employment in accordance with legal requirements. Accommodations We strive to create an accessible and inclusive experience for all candidates. If you require a reasonable accommodation to complete any part of the application process, or are unable to use this online application and need an alternative method to apply, please contact [email protected] for assistance. Export Control Regulations For positions requiring access to controlled technology subject to export control regulations, including the U.S. Export Administration Regulations (EAR), ServiceNow may be required to obtain export control approval from government authorities for certain individuals. All employment is contingent upon ServiceNow obtaining any export license or other approval that may be required by relevant export control authorities. From Fortune. ©2026 Fortune Media IP Limited. All rights reserved. Used under license. Employee Type: Regular Region: AMS - North America and Canada Work Persona: Flexible or Remote

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Reliability Engineer in United States vacancy
  • $235k - $275k

     ...Inc. as one of the most innovative and fastest-growing technology companies in the country. \n Role Summary As a Staff Site Reliability Engineer at Filevine, you are the senior technical authority on the SRE team and a strategic partner to engineering leadership.... 
    Suggested
    Permanent employment
    Full time
    Temporary work

    Filevine

    United States
    1 day ago
  •  ...We're looking for a Staff Reliability Engineer for ICON. This role will serve as technical resource on reliability for our fleet of Vulcan and Titan print systems. This is a high-impact individual contributor role with org-wide reach: you'll diagnose failure patterns... 
    Suggested
    For contractors

    I-Con Technology

    Sunset Valley, TX
    3 days ago
  • $220k - $255k

     ...and SUV miles with ones on vehicles that are more affordable, more enjoyable and 10-50x more efficient. ALSO is looking for a Reliability Engineer to play a key role in developing and leading the reliability of multiple electric mobility vehicles over the full... 
    Suggested
    Full time
    Local area
    Flexible hours

    ALSO

    Palo Alto, CA
    21 hours ago
  • $165k - $218k

     ...computer vision, sensor fusion, and networking technology to the military in months, not years. ABOUT THE TEAM The Reliability Engineering team partners across Anduril's engineering, manufacturing, and operations organizations to ensure our autonomous systems survive... 
    Suggested
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Atlanta, GA
    1 day ago
  •  ...technology designed to keep the electric grid secure and reliable, even during extended periods of stress. By strengthening...  ...place. Role Description Form Energy is seeking a Staff Reliability Engineer to improve the reliability and lifetime of our iron-air battery... 
    Suggested
    Full time
    Relocation package

    Form Energy, Inc.

    Somerville, MA
    2 days ago
  •  ...Senior Reliability Engineer The Senior Reliability Engineer position is responsible for improving equipment reliability, asset performance, and long-term equipment health across the facility. This role integrates engineering principles, data analytics, maintenance... 

    Arconic

    Lancaster, PA
    21 hours ago
  •  ...Building cloud-native reliability and automation solutions, the full-time Staff Reliability Engineer will design and operate engineering platforms for software validation and release, while working flexibly or remotely to enhance developer productivity and deployment... 
    Full time
    Remote work

    Virtual Vocations Inc

    United States
    2 days ago
  • $148.9k - $260.6k

     ...applications, data, cloud environments, and AI agents. For engineers joining Veza today, this means the scale and resources of an...  ...moment in the industry. We are seeking an exceptional Staff Site Reliability Engineer to lead critical infrastructure initiatives and drive... 
    Work at office
    Remote work
    Flexible hours

    ServiceNow

    United States
    21 hours ago
  •  ...Arkema's Beaumont Plant is seeking a knowledgeable Staff Reliability Engineer to support rotating equipment performance, strengthen predictive maintenance programs, and contribute to long-term equipment reliability across the site. This role plays a key part in improving... 
    For contractors
    Local area

    Arkema

    Lumberton, TX
    1 day ago
  • $190k - $235k

     ...and highly collaborative team that is passionate about creating transformative change in healthcare. We are seeking a Staff Site Reliability Engineer to play a key role in designing, optimizing, and securing the underlying cloud infrastructure that powers our organization... 
    Flexible hours

    Datavant

    Augusta, ME
    3 days ago
  • $112.5k - $187.5k

     ...TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering...  ...the most complex and consequential work on the platform. As a Staff Site Reliability Engineer at TransUnion, you will serve as a... 
    Full time
    Temporary work
    Work experience placement
    Work at office
    Flexible hours
    2 days per week

    TransUnion

    Chicago, IL
    1 day ago
  • $128.5k - $214.1k

     ...We're looking for a Staff Site Reliability Engineer to join our team, focusing on the core systems that power global financial markets. This isn't just about keeping the lights on; it's about pioneering the future of financial technology. As a member of our Clearing department... 
    Work at office
    Worldwide
    2 days per week

    CME Group

    Chicago, IL
    1 day ago
  • $131.6k - $210.3k

     ...We are seeking a highly skilled and experiencedStaff Site Reliability Engineerto join our DevOps squad. This role will focus on leading...  ...IaC), and cloud technologies, as well as the ability to mentor engineers and contribute to the overall stability and scalability of our... 
    Work experience placement
    Work at office
    Local area
    Remote work

    Visa

    Austin, TX
    3 days ago
  •  ...Required U.S. Citizenship / No clearance needed / 100% remote within the US  Staff Site Reliability Engineer / Cloud SME Location: 100% remote in the continental US  Type: Long-term contract (3+ years) Role Summary As the Staff SRE/Cloud SME, you will be... 
    Long term contract
    Remote work

    ASCENDING LLC

    Fairfax, VA
    2 days ago
  • $220k - $235k

     ...Staff/Senior Staff Site Reliability Engineer Ironclad is the leading AI contracting platform that transforms agreements into assets. Contracts move faster, insights surface instantly, and agents push work forward, all with you in control. Whether you're buying or selling... 
    Full time
    Contract work
    Work at office

    Ironclad Inc

    San Francisco, CA
    21 hours ago
  • $181k - $263k

     ...and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Staff Site Reliability Engineer who will set the technical direction for reliability engineering across LiveRamp's global infrastructure. This is a... 
    Work from home
    Flexible hours
    Night shift

    LiveRamp

    San Francisco, CA
    21 hours ago
  • $160k - $225k

     ...tens of thousands of users across hundreds of organizations globally. About the Role Manifold is looking for a Staff Site Reliability Engineer (SRE) to work at the intersection of AI, data infrastructure, and life sciences. In this high-impact role, you will help... 

    Manifold AI

    Cambridge, MA
    21 hours ago
  • $200k - $260k

     ...Join the Cloud Infrastructure Team as a technical leader driving reliability, automation, and scalability across the systems running Sight...  ...and reliability practices across teams, mentor senior engineers, and be a primary escalation point for the org's hardest systems... 
    Casual work
    Work at office
    Remote work
    Flexible hours

    Sight Machine

    San Francisco, CA
    2 days ago
  • $170k - $210k

     ...businesses think about cybersecurity, digital experiences, and identity and access management.  As a Ping Identity Site Reliability Engineer, you will be involved in every facet of our Cloud-based services. You will establish solutions for building, deploying, and... 
    Local area
    Remote work
    Worldwide
    Flexible hours

    Ping Identity

    United States
    2 days ago
  • $148k - $193k

     ...Staff Site Reliability Engineer Denver, CO, USA DAT is an award-winning employer of choice and a next-generation SaaS technology company that has been at the leading edge of innovation in transportation supply chain logistics for 45 years. We continue to transform... 
    Temporary work
    Work experience placement
    Work at office
    Local area
    Immediate start
    Flexible hours

    DAT Freight & Analytics

    Denver, CO
    2 days ago
  • $131k - $164k

     ...Staff Site Reliability Engineer New York, New York, United States Position Overview We are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across VMware, Linux, and automation frameworks, to join our global Infrastructure... 
    Work at office
    Local area
    Visa sponsorship
    Flexible hours

    Diligent

    New York, NY
    21 hours ago
  •  ...Series B and have grown 800% over the last 12 months. Engineering at Ivo Engineers at Ivo are inventors. Ivo was first-to...  ...to hit our SLAs. What? We're looking for a Senior or Staff Site level Reliability Engineer as part of Infrastructure team to: Own uptime... 
    Contract work
    Work at office
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours

    IVO Inc

    San Francisco, CA
    2 days ago
  • $194k - $267k

     ...more than once, automate it" and who can rapidly self-educate on new concepts and tools. Position Overview: The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta, Inc.

    Bellevue, WA
    21 hours ago
  • $217.57k - $260k

     ...AI can be appropriately used during your application and interviews which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure team responsible for owning environments with thousands... 
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours
    Shift work

    ID.me

    Mountain View, CA
    3 days ago
  •  ...contributes. No one coasts. If you're driven by impact, pace, and raising the bar. This is the place. The Role As a Staff Site Reliability Engineer you'll play a lead role on the founding SRE team at our new NYC engineering hub. You'll own multi-team reliability and... 
    Work at office

    Legora

    New York, NY
    3 days ago
  • $102.6k - $193.43k

     ...Software Engineer Chamberlain Group (CG) is a global leader in intelligent access and Blackstone portfolio company. Powered by our...  ...Establish working relationships with business members, technical staff, and subject matter experts to help create and execute the infrastructure... 
    Temporary work
    Second job
    Worldwide

    Chamberlain Group

    Oak Brook, IL
    21 hours ago
  •  ...Job Description Job Description Staff Reliability Engineer Company Overview Embark on an enriching journey with PROCEPT BioRobotics, where our vision, mission, and values guide everything we do as a company. At PROCEPT, we put the patient first in everything we... 
    Temporary work
    Local area
    Flexible hours

    Gulf Coast Automation Group

    San Jose, CA
    24 days ago
  • $180k - $230k

     ...Job Description Job Description Job Title: Staff Reliability Engineer Location: Burlingame, CA Department: ESS Engineering Reports To: Staff Reliability Engineer Position Type: Full-time About Peak Energy Peak Energy is the first American... 
    Full time
    Immediate start
    Flexible hours

    Peak Energy

    San Francisco, CA
    23 days ago
  •  ...worldwide. For more information, visit  Follow Shield AI on LinkedIn, X, Instagram, and YouTube.  Job Description: The Reliability and Maintainability Engineer is a senior individual contributor responsible for driving investigation, root cause analysis, corrective action,... 
    Full time
    Temporary work
    Part time
    Worldwide

    Shield AI

    Dallas, TX
    26 days ago
  •  ...Description Job Description Position Overview An innovative and growing medical technology company is seeking a Reliability Quality Engineer to ensure the product reliability and quality of a complex, next-generation medical device ecosystem. In this permanent... 
    Permanent employment
    Full time

    Aviso, LLC

    Fremont, CA
    7 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Reliability Engineer. Be the first to apply!