Staff Software Engineer - SRE & AIOps
Moveworks
Company DescriptionIt all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker, Phyllis. She cried tears of joy. That moment inspired Fred to build a company that could do that for everyone—freeing people from busywork so they could focus on meaningful work. Today, ServiceNow is the AI control tower for business reinvention. Our ServiceNow AI platform brings together any AI, any data, and any workflow— helping 85% of the Fortune 500 work smarter, faster, and better. We're building an AI-native culture where technology and talent are unstoppable together. And we're just getting started.Join us to put AI to work for people.Job DescriptionAbout the role:ServiceNow is seeking a Staff Software Engineer - SRE & AIOps to contribute to infrastructure automation, operational resilience, and toil elimination across our hybrid cloud and data center operations. Embedded within the Site Reliability & Database Engineering organization, you will implement automation-first systems that reduce manual intervention, accelerate incident remediation, and enable our global engineering teams to operate reliably at scale. This role combines solid hands-on technical expertise in Kubernetes, cloud platforms, and DevOps practices with growing technical leadership capabilities. You will contribute to SRE tooling design, develop auto-remediation capabilities, and help establish patterns that maintain ServiceNow's cloud platform reliability while minimizing operational toil across follow-the-sun global teams.What you get to do in this role:Deploy, operate, and troubleshoot production Kubernetes clusters across hybrid and multi-cloud environments, maintaining operational standards and supporting high-velocity application deployments.Implement and maintain closed-loop auto-remediation systems that detect, classify, and resolve transient infrastructure failures, leveraging automation frameworks and machine learning insights to reduce MTTR and on-call burden.Contribute to the design and evolution of SRE tooling stack, including monitoring platforms, incident management systems, log aggregation, and observability integrations that support global on-call operations.Develop and maintain SLO frameworks, alerting policies, and automated runbooks that empower on-call engineers to resolve issues autonomously while managing alert fatigue.Build and maintain Infrastructure-as-Code frameworks and GitOps pipelines that enable reproducible infrastructure deployments across hybrid and multi-cloud environments with security and compliance guardrails.Support hybrid cloud and data center operations, including on-premises infrastructure, public cloud environments, and workload optimization across multi-region deployments.Contribute to adoption of containerization, microservices, and DevOps patterns across engineering teams, establishing CI/CD best practices and network security controls.Support on-call rotation operations and incident response processes across different time zones, helping develop runbooks and contributing to post-incident reviews that drive continuous improvement.Share knowledge and mentor junior SRE engineers on reliability patterns, incident investigation techniques, and automation best practices.Champion a culture of blameless incident analysis, data-driven decision-making, and continuous improvement through knowledge sharing and documentation.Identify and systematically automate repetitive operational tasks, from infrastructure provisioning to incident response, improving team efficiency and capacity.QualificationsTo be successful in this role you have:Kubernetes Proficiency: Solid hands-on experience operating production Kubernetes clusters, including deployment models, pod orchestration, resource management, network policies, and troubleshooting runtime issues.Incident Remediation Experience: Demonstrated experience designing and implementing automated remediation systems, including alert automation, runbook development, and self-healing mechanisms.Cloud Platform Knowledge: Strong hands-on experience with AWS (EKS, EC2, RDS) and/or Azure (AKS, VMs) or GCP (GKE), with understanding of core SRE-related services.DevOps & IaC Skills: Solid experience with Infrastructure-as-Code tools (Terraform, CloudFormation) and GitOps practices.SRE Tooling Familiarity: Working knowledge of observability platforms, incident management systems, and log aggregation tools.Distributed Systems Understanding: Understanding of distributed system challenges, fault tolerance, and resilience patterns.On-Call Operations: Experience participating in on-call rotations and understanding 24/7 operational models, runbook development, and escalation procedures.Cloud & Hybrid Operations: Hands-on experience working with cloud infrastructure and understanding hybrid cloud concepts.Systems Administration: Strong foundation in Linux system administration, performance troubleshooting, and scripting (Python, Go, or Bash).Collaborative Mindset: Ability to work effectively with infrastructure and application teams, contribute to technical discussions, and help drive reliability improvementsQualificationsExperience in leveraging or critically thinking about how to integrate AI into work processes, decision-making, or problem-solving. This may include using AI-powered tools, automating workflows, analyzing AI-driven insights, or exploring AI's potential impact on the function or industry.8+ years in software engineering or infrastructure operations, with 3+ years in SRE, DevOps, or cloud platform engineering roles with a Bachelor's degree; or 6 years and a Master's degree; or a PhD with 3 years of experience; or equivalent work experience.2+ years of hands-on experience working with production Kubernetes clusters.Proficiency in at least one Infrastructure-as-Code tool: Terraform, CloudFormation, or equivalent.Demonstrable hands-on experience with at least one major cloud platform: AWS, Azure, or GCP.Experience operating in on-call environments and participating in incident response.Experience implementing or improving automated remediation and alert systems.Strong foundation in Linux system administration, performance troubleshooting, and scripting (Python, Go, Bash).Demonstrated commitment to reliability engineering and continuous improvement through hands-on contributions.Bachelor's degree in computer science, Computer Engineering, or related field (or equivalent professional experience).Preferred:Kubernetes certification (CKA, CKAD, or equivalent).Experience with service mesh technologies or advanced Kubernetes networking.Background in cloud migration or infrastructure modernization projects.Experience with cost optimization in cloud environments.Track record of implementing automation solutions that significantly reduced operational toil.Why This Role?This role offers the opportunity to work with infrastructure automation and reliability engineering at scale. You will implement systems and practices that directly reduce operational burden across ServiceNow's global engineering teams. Your contributions will help establish reliable, automated infrastructure operations and provide a clear career path toward senior technical leadership. This is a role for an engineer who enjoys solving complex operational challenges, continuous learning, and working collaboratively to improve how systems operate.Additional InformationWork PersonasWe approach our distributed world of work with flexibility and trust. Work personas (flexible, remote, or required in office) are categories that are assigned to ServiceNow employees depending on the nature of their work and their assigned work location. Learn more here. To determine eligibility for a work persona, ServiceNow may confirm the distance between your primary residence and the closest ServiceNow office using a third-party service.Equal Opportunity EmployerServiceNow is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, national origin, age, disability, gender identity, veteran status, or any other category protected by law. In addition, all qualified applicants with arrest or conviction records will be considered for employment in accordance with legal requirements. AccommodationsWe strive to create an accessible and inclusive experience for all candidates. If you require a reasonable accommodation to complete any part of the application process, or are unable to use this online application and need an alternative method to apply, please contact View email address on click.appcast.io for assistance. Export Control RegulationsFor positions requiring access to controlled technology subject to export control regulations, including the U.S. Export Administration Regulations (EAR), ServiceNow may be required to obtain export control approval from government authorities for certain individuals. All employment is contingent upon ServiceNow obtaining any export license or other approval that may be required by relevant export control authorities. From Fortune. 2026 Fortune Media IP Limited. All rights reserved. Used under license.SummaryType: Full-timeFunction: Information TechnologyExperience level: Not ApplicableIndustry: Information Technology And Services
$129.5k - $186.1k
...matters—and so do you. About the Team: We are seeking a Staff Software Engineer with strong Go (Golang) skills to deliver and operate... ...with open-source cloud-native projects • Familiarity with SRE principles and DevOps practices • GitOps workflows and advanced...SuggestedFull time$139.73k - $209.59k
...business and society. The Challenge OneTrust is looking for a Staff Software Engineer to help lead technical direction for Developer Experience.... ...and Architecture before implementation starts.Work with QA, SRE, InfoSec, IAM, and R&D Operations when an initiative needs their...SuggestedWork experience placementWork at officeLocal areaImmediate startWorldwideFlexible hours3 days per week1 day per week$145.6k - $209.3k
...matters—and so do you.About the RoleWe're looking for a Sr. Staff Software Engineer to anchor the technical direction of our core platform and backend... ...-scale technical migrations. Exposure to observability and SRE practices in a high-availability production environment....SuggestedFixed term contract$180k - $247.5k
...operations space. You will define and drive the AIOps strategy for the organization, combining... ...operations, observability, automation, SRE practices, and AI/agentic solutions to... ...technical leader, you will work across Engineering, Cloud, Infrastructure, Security, and Operations...SuggestedFull timeWork at officeLocal areaVisa sponsorshipFlexible hours$139.73k - $209.59k
...future where trusted data becomes a transformative force for business and society. The Challenge We’re looking for a Staff Software Engineer with a passion for solving problems to join our agile AI Governance team at OneTrust. Staff Software Engineers are...SuggestedFull timeWork experience placementWork at officeLocal areaWorldwideFlexible hours3 days per week1 day per week$190k - $252k
...continuous improvement. That environment includes not just the software used to design, plan, and execute work, but also the... ...industrial base. ABOUT THE JOB: We are looking for a Staff Software Engineer to join our QualityOS team within Forge MES (Manufacturing...Full timeTemporary workWork experience placementImmediate start$163k - $204k
...interview process. We're looking for seasoned full-stack software engineers to join the teams behind Gusto's customer-facing products — the... ...owners and their employees rely on every day. As a Staff Software Engineer for the Time / Scheduling product team, you...Full timeWork at officeLocal areaRemote work2 days per week3 days per week- ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production support, incident management, monitoring, observability, troubleshooting, and improving system reliability and availability...Temporary work
$195k - $200k
...in-class experience for our members. Our engineering team builds and scales the platform that... .... We are looking for a Senior Software Engineer to join our Overall Game Experience... ...Summary We are seeking an experienced Staff Software Engineer with deep expertise in...Remote jobFull timeWork visaFlexible hours- ...Staff Software Engineer The CNN Growth team is hiring a Staff Software Engineer to help design, build, and evolve the core systems and experiences that support audience growth, engagement, and monetization across CNN's digital platforms. This role is ideal for a senior...
- ...Apply today to join Coreforce, where your Engineering expertise makes a real impact. Join Our Team as a Staff Software Engineer Company: Coreforce Location: Atlanta, GA Job Type: Full-time Salary: Based on Experience Company Overview:...Full timeFlexible hours
$260.1k
...together quarterly for intense in-person working sessions called “surges.”learn more about working at Coinbase. As a Senior Staff Software Engineer on thePlatform Payments team, you'll define the engineering vision for how fiat moves into and out of Coinbase across 50+...Local area$160.2k - $246.3k
...the accuracy, reliability, and efficiency of simulation tests used for autonomous vehicle software validation. Lead cross-functional initiatives with Autonomy, Systems Engineering, Simulation, and Data teams to tightly integrate team-ownedtest operations andevaluation...Full timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours$144.7k - $284.1k
...Simulation, and our partners Behaviors, Perception, and Safety Engineers. The specific duties may include ML/RL model development as... .... Work as part of an ML team and contribute strong software engineering (SWE) expertise. Support the ML team in accelerating...Full timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hoursShift work- ...rapid experimentation to scaled production deployment across diverse model types and use cases. Your New Role... As a Staff Software Engineer on ML Platform, you will work across teams to design, build, and operate the infrastructure foundations that power model...
$195k - $257.5k
...Staff Software Engineer Circle is one of the world's leading internet financial platform companies, building the foundation of a more open, global economy through digital assets, payment applications, and programmable blockchain infrastructure. Circle's platform includes...Contract workRemote workFlexible hours$161.91k - $181.91k
...: Design and deliver highly scalable multi-tiered distributed software applications; Build services and applications using code generation... ...architectural reviews, code quality standards, and best engineering practices; Drive strategic technology decisions. Service/Interface...InternshipRemote workWorldwide- ...CNN New Business Engineering Team Technical Leader The CNN New Business Engineering team is building a new portfolio of standalone consumer... ...for a seasoned technical leader with deep expertise in Web software and system architecture design, combined with hands-on...
$139.73k - $209.59k
...future where trusted data becomes a transformative force for business and society. The Challenge We're looking for a Staff Software Engineer with a passion for solving problems to join our agile AI Governance team at OneTrust. Staff Software Engineers are responsible...Work experience placementWork at officeLocal areaWorldwideFlexible hours3 days per week1 day per week$220k - $275k
...stepping onto a driven and highly collaborative team that is passionate about creating transformative change in healthcare. Staff Software Engineer The Role As a Staff Software Engineer, you will shape the technical direction and architecture of your business unit/...$148.8k - $204.6k
...Position OverviewThe Senior DevOps Engineer position drives reliability... ...What You’ll Do:Set enterprise SRE standards (SLO/SLI taxonomy,... ...evacuation exercises.Establish AIOps/event correlation to reduce... ...making and follow-through.Mentors staff and principal engineers,...Full timeWork at officeLocal areaVisa sponsorshipFlexible hours- ...for live event ticket inventory. We provide an end-to-end software platform for the live ticketing industry, managing thousands... ...fulfillment. THE POSITION Victory Live is looking for a Staff Software Engineer to own the technologies and processes that power our B2B...Full timeLocal area
$139.73k - $209.59k
...Strong professional experience building and operating production software systems as a highly autonomous individual contributor.... ...willingness to join an on-call rotation. Experience using AI engineering tools such as Devin, Claude, or similar systems to produce production...Full timeWork at officeFlexible hours3 days per week1 day per week- ...Engineer Lead (Staff Software Engineer Sr) Strategic Claims Pricing and Reimbursement Teams Location: Atlanta, GA Hybrid 2: This role requires associates to be in-office 3 days per week at the Atlanta, GA office, fostering collaboration and connectivity, while...Full timeTemporary workWork at officeLocal area3 days per week1 day per week
- ...networking equipment at AeroVect garage facilities. Manage software release delivery to autonomous tractors and on-premise devices... ...Bachelor’s or master’s degree in Computer Science, Electrical Engineering, or a related field. At least 7 years of experience managing...Full timeShift work
- ...About Us Dolby Cloud Solutions is a video streaming software company that helps enterprise and mid-market organizations deliver... ...work is simple: AI prepares, humans decide. This is a Staff-level engineering role at the center of that effort. You will own the...Local areaShift work
$196.5k - $294.75k
...We are looking for a Principal-level, US-based, customer-facing engineer who is deeply hands-on with large-scale, distributed data systems... ...of incidents in partnership with Product Engineering, SRE/CloudOps, and Support. Lead deep-dive technical sessions with...Full timeWork experience placementWork at officeLocal areaWorldwideFlexible hours3 days per week1 day per week- #CareersJC 1483593Qualifications· Strong experience supporting production systems hosted on AWS, including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS.· Hands-on experience with incident management and 24/7 production support models.· Proficiency with monitoring and observability...
- ...description:The Site Reliability Engineer role focuses on enhancing the... ...practices, mentoring SRE team members, and contributing... ...from time to time.1. Implements software architecture and engineering approaches... ...correlation leveraging AI and AIOps tools. Guide and enforce SLO/...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
$73.01k - $170.64k
...participate in all aspects of the software development lifecycle... ...leveraging modern observability, AIOps, and automation capabilities... ...Powered by our 7,000+ advisors, engineers, and designers, Perficient... ...technical and non-technical staff. Bachelor’s Degree in MIS, Computer...Full timeWork at officeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Software Engineer - SRE & AIOps. Be the first to apply!
- software sales representative Atlanta, GA
- embedded software Atlanta, GA
- software applications developer Atlanta, GA
- entry level software sales Atlanta, GA
- software technology Atlanta, GA
- software implementation project manager Atlanta, GA
- software support Atlanta, GA
- government software Atlanta, GA
- software engineer - cloud services Atlanta, GA
- software qa Atlanta, GA



