Senior Site Reliability Engineer - Unified Observability
NCR
About NCR VOYIXNCR Voyix Corporation (NYSE: VYX) is a global platform-powered leader in unified commerce for shopping and dining. Combining a flexible, intelligent platform with end-to-end payments capabilities and services developed through its deep industry experience, NCR Voyix empowers retailers and restaurants to accelerate new possibilities for their operations, experiences and business outcomes. NCR Voyix is headquartered in Atlanta, Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation Customer Unified Observability initiative. This strategic role will be responsible for building and evolving a unified enterprise observability platform that delivers end-to-end visibility across NCR Voyix Restaurants, Retail, and Payments environments.The ideal candidate will bring 10+ years of experience in Site Reliability Engineering, Platform Engineering, Cloud Operations, or related disciplines, with a proven track record of driving enterprise-scale observability, reliability, and operational excellence. This individual must be comfortable operating across organizational boundaries and partnering closely with Product Engineering, Infrastructure, Security, Operations, Architecture, and Executive Leadership teams to establish a comprehensive observability strategy and improve platform resilience.This role will serve as a key technical leader responsible for defining standards, influencing architecture decisions, and enabling proactive operations through unified monitoring, telemetry, automation, and AI-driven insights.Key ResponsibilitiesLead the architecture, design, implementation, and continuous improvement of enterprise observability solutions across Azure, Google Cloud Platform (GCP), Kubernetes, and hybrid environments.Establish and drive enterprise observability standards for monitoring, logging, distributed tracing, telemetry, and operational analytics.Develop and maintain executive, operational, and engineering dashboards that provide real-time visibility into infrastructure, applications, platform health, customer experience, and business transactions.Define, evangelize, and implement reliability frameworks including SLIs, SLOs, error budgets, operational KPIs, and service health metrics.Partner cross-functionally with Engineering, Infrastructure, Security, Product, and Operations teams to identify reliability risks and drive operational excellence initiatives.Lead efforts to improve incident prevention, detection, response, and recovery through intelligent alerting, automation, event correlation, and observability best practices.Integrate observability capabilities with ServiceNow, CI/CD pipelines, automation frameworks, and enterprise operational workflows.Influence technical strategy and roadmap decisions related to reliability engineering, platform observability, and operational readiness.Support and drive enterprise initiatives involving AI-driven observability, predictive analytics, anomaly detection, and event intelligence.Mentor engineers and serve as a subject matter expert for observability, reliability engineering, and cloud-native operations.Establish governance, adoption, and best practices across multiple product and engineering teams to ensure consistent observability standards enterprise-wide.Required QualificationsBachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience.10+ years of experience in Site Reliability Engineering, Cloud Engineering, Platform Engineering, DevOps, or related technical disciplines.Demonstrated success designing and operating observability platforms in large-scale enterprise environments.Deep expertise with Kubernetes platforms, including AKS and GKE.Strong experience with Azure and Google Cloud Platform services and architectures.Hands-on experience with enterprise observability tools such as Grafana, Datadog, Prometheus, OpenTelemetry, Dynatrace, New Relic, or similar platforms.Advanced knowledge of monitoring, logging, telemetry collection, distributed tracing, and observability engineering principles.Experience defining and operationalizing SLIs, SLOs, error budgets, reliability metrics, and service health frameworks.Strong automation and Infrastructure as Code expertise using Terraform and related tools.Proficiency developing automation solutions using Python, Go, PowerShell, or similar languages.Experience integrating observability solutions into CI/CD pipelines and modern DevOps workflows.Proven ability to influence technical direction and collaborate effectively with stakeholders across Engineering, Product, Infrastructure, Security, and Operations organizations.Strong communication, leadership, and stakeholder management skills with the ability to translate technical concepts for both technical and business audiences.Preferred QualificationsExperience leading enterprise observability transformations or platform modernization initiatives.Experience with AI Ops, event correlation, operational analytics, and predictive monitoring capabilities.Knowledge of ServiceNow integrations and ITSM/ITOM processes.Experience supporting highly available, customer-facing SaaS platforms at scale.One or more cloud certifications (Azure, Google Cloud, Kubernetes, or related technologies).Previous experience serving as a technical lead, mentor, or architect within a reliability engineering organization.Offers of employment are conditional upon passage of screening criteria applicable to the jobEEO StatementIntegrated into our shared values is NCR Voyix’s commitment to equal employment opportunity. All qualified applicants will receive consideration for employment without regard to sex, age, race, color, creed, religion, national origin, disability, sexual orientation, gender identity, veteran status, military service, genetic information, or any other characteristic or conduct protected by law. NCR Voyix is committed to being a globally inclusive company where all people are treated fairly, recognized for their individuality, promoted based on performance and encouraged to strive to reach their full potential. We believe in understanding and respecting differences among all people. Every individual at NCR Voyix has an ongoing responsibility to respect and support a globally diverse environment.Statement to Third Party AgenciesTo ALL recruitment agencies: NCR Voyix only accepts resumes from agencies on the preferred supplier list. Please do not forward resumes to our applicant tracking system, NCR Voyix employees, or any NCR Voyix facility. NCR Voyix is not responsible for any fees or charges associated with unsolicited resumes“When applying for a job, please make sure to only open emails that you will receive during your application process that come from a @ncrvoyix.com email domain.”SummaryLocation: ATLANTA, GA, USAType: Full time
- ...Senior Site Reliability Engineer Atlanta, Georgia Who We Are QGenda is redefining healthcare... ...strategic workforce decisions through our unified software platform. With more than 8... ...the organization by promoting observability, retrospectives, and continuous improvement...SeniorPermanent employmentFull timeWork at officeRemote workWork from homeWork visa
$104.9k - $174.7k
...the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the... ...designing infrastructure, writing Terraform, improving observability, and responding to real production incidents.If you live...SeniorFull timeWork at officeLocal areaRemote workWork from home$160k - $210k
...two coasts. Our culture thrives on cross-departmental collaboration and a unified sense of purpose, making teamwork a cornerstone of our success. We are looking for a Senior Site Reliability engineer to work on expanding our global footprint of datacenters and improve...SeniorWork at officeLocal areaImmediate startRemote work$218.11k - $274.19k
...Site Reliability Engineer We are Omnissa! Omnissa is the first AI-driven digital work platform,... ...industry-leading solutions—including Unified Endpoint Management, Virtual Apps and... ...solutions and practices to improve the observability and autoremediation/self-healing of our...SuggestedWork experience placementLocal areaRemote workVisa sponsorshipFlexible hours- ...The Home Depot is seeking a Senior Software Reliability Engineer to join the Platform Reliability Engineering team, ensuring the resilience, performance, and security of our enterprise Cloud Platform. You will mentor junior engineers, lead incident triage, root cause...Senior
$244k - $305k
...repositories into the monorepo. As a senior software engineer on the team, you will own... ...that are simpler and more reliable to use.Push performance... ...: Datadog is the leading observability and security platform for... ...providing businesses with unified visibility across the...SeniorWork at office$172.5k - $260.1k
...team in the Service Delivery Platform & Reliability group at Slack develops platforms and... ..., provide insights, and improves observability in Slack production services with a focus... ...and work closely with other teams in engineering, product development, and customer experience...SeniorFull time$116.48k - $174.71k
...Ready Governance Platform, unifies regulatory intelligence,... ...ChallengeWe're looking for a Senior Software Engineer that will report to the... ...and implement application observability and platform monitoring tools... ...budgets to balance system reliability with product feature...SeniorWork experience placementWork at officeLocal areaWorldwideFlexible hours3 days per week1 day per week- ...helping the world’s most important research sites do their best work. Our solutions are... ...to the Team:We are seeking a Site Reliability Engineer (SRE) to join one of our Scrum teams... ...while actively leveraging AI to improve observability, incident response, automation, and...Work at office
$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and... ...delivery.Cross trains colleagues on how to best leverage observability tools during incident and performance investigations....Full timeTemporary workWork experience placementFlexible hours$152.13k - $162.13k
...about what’s next. Join us.General Summary:Unum Group seeks Site Reliability Engineers in Atlanta, GA.Applicants who are interested in this... ...Ref #66753) for consideration.Design, build, and maintain observability, monitoring, and alerting capabilities across consumer and...Full timeTemporary workWork at officeRemote work$84.9k - $209.5k
...Job Description As a Principal Site Reliability Engineer (IC4), you will be responsible for designing... ...infrastructure, coding, automation, observability, incident management, and operational... ...incident response and act as a senior escalation point during major service...Temporary workFlexible hours- ...cloud-native systems. As a Staff Platform Engineer, you will play a critical role in... ...technical leadership role. You will own reliability for major platform domains, design scalable... ...Establish and enhance centralized Observability and Monitoring platforms and tools that...Senior
- ...The CNN Growth team is hiring a Senior Software Engineer to help build and evolve the systems... ...hands on across the stack, delivering reliable, performant features while contributing... ...AWS services, CI/CD pipelines, and observability tools. Requirements 4+ years of...SeniorFull time
$165k - $247.5k
...Governance Platform, unifies regulatory... ...ChallengeWe are hiring a Senior Staff DevOps Engineer to join our Detect &... ...CVEs; and ensuring the reliability, scalability, and security... ...expertise on-site.Container Hardening... ...code across the team.Observability & Platform Reliability...SeniorWork experience placementWork at officeLocal areaWorldwideFlexible hours3 days per week1 day per week- ...environments . Experienced in architectural design for reliability, scalability, and performance. Practical application of... ..., Docker, serverless computing). Proficient in observability solutions such as Dynatrace, Prometheus, Grafana, and ELK/EFK...
- ...Exchange (ICE) presents a unique opportunity for a full-time Engineer to join a team responsible for ICE OpenShift container platform... ...high availabilityExperience with platform and application observability (tracing, logging metrics)Top-tier analytics and problem solvingAbility...SeniorFull time
- ...apply for the Junior Software Developer - Observability role at Canonical 3 days ago Be among... ...as public cloud, data science, AI, engineering innovation, and IoT. Our customers include... ...give your application fair consideration. Seniority level Entry level Employment type Full‑...Full timeWork at officeRemote workWork from home
- ...technology and services Position: Senior AI/ML Engineer Location: Atlanta GA or Frisco TX... ...of customer records into accurate, unified profiles - and build the natural... ...prompt engineering, evaluation, and LLM observability - ensuring AI outputs meet the trust...SeniorTemporary workWork experience placement
$207.4k - $298.1k
...roadmap.About the role:We are hiring a Senior Principal Software Engineer to lead UKG Ready SMB —... ...container orchestration (Kubernetes), observability, capacity planning and SRE practices... ...cross-team integrations, operational reliability and cost-efficiency as Ready...SeniorWorldwide- ...Role: Site Reliability Engineering (SRE) Architect Location: Atlanta, GA (Hybrid on-site)... ...across the organization and much beyond observability pillar. Key Responsibilities:... ...Leadership & Consultation: Act as a senior technical advisor and subject matter...Contract workEarly shift
$120k - $150k
...following job description: The Senior ServiceNow Platform Engineer is responsible for the... ..., automation, reliability, and continuous improvement... ...Azure. • Experience with observability, monitoring, and operational... ...please visit our Benefits site. Depending on the position...SeniorPermanent employmentFull timePart timeWork experience placementH1bWork at officeLocal areaImmediate startWork visaShift workDay shift$228k - $342k
...We’re forming small, senior, cross-functional AI teams... ..., machine learning engineers, and full-stack builders... ...are scalable, observable, and enterprise-ready.... ...into systems that are reliable, explainable, and built... ...Careers. Please be aware of sites that may ask for you...SeniorFull timeWork at officeRemote workHome officeFlexible hours- ...A leading open-source technology firm is seeking a Junior Software Developer - Observability. In this role, you'll develop a cloud-native monitoring stack and work on exciting projects involving Linux and Kubernetes. Candidates should have skills in Python, a working knowledge...Remote work
- ...This role is pivotal to our platform engineering and site reliability initiatives, owning the Amazon EKS (... ...product teams build and run. The Senior Engineer, Platform & Site Reliability... ...the Istio service mesh, and modern observability to keep our systems secure, scalable...SeniorWork experience placement
- ...seamless efficiency. When operations are unified on a single platform, districts gain... ...to shape the future of education. Site Reliability Engineer (SRE) Overview: We are looking... .... You'll work with leading-edge observability and reliability tooling, and the calls...Full timeLive inWork at office
- ...including multi-threading, memory management and web services.History of building resilient, stateless, scalable, distributed, and observable systems.Experience in building REST services with high focus on performance.Familiarity with microservices and knowledge of...Senior
$112.8k - $257k
Observability and AI Ops Architect, SeniorThe Opportunity:We are seeking... ...with hands-on observability engineering to support government initiatives... ...skills for engaging senior leadershipGoogle Professional... ...Resource page on our Careers site and reviewing Our Employee Benefits...SeniorFull timeContract workPart timeWork at officeLocal areaRemote work$143k - $191k
...technology to the military in months, not years.ABOUT THE TEAMThe Reliability Engineering team partners across Anduril's engineering, manufacturing,... ...process and the security of our candidates. We've observed a rise in sophisticated phishing and fraudulent schemes where...SeniorFull timeWork experience placementImmediate start- ...candidate to join our talented Team.Job Title: Senior Software EngineerLocation(s): Atlanta,... ...seeking an experienced Senior Software Engineer to design, develop, and deliver scalable... ...using CloudWatch, X-Ray, and other observability tools.Optimize cloud performance, security...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer - Unified Observability. Be the first to apply!
- site reliability engineer Atlanta, GA
- site reliability engineer sre Atlanta, GA
- senior manufacturing manager Atlanta, GA
- senior business analyst Atlanta, GA
- senior risk manager Atlanta, GA
- senior cost estimator Atlanta, GA
- senior manager tax Atlanta, GA
- senior automation engineer Atlanta, GA
- senior devops Atlanta, GA
- senior recruiter Atlanta, GA

