Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer - Unified Observability

NCR

About NCR VOYIXNCR Voyix Corporation (NYSE: VYX) is a global platform-powered leader in unified commerce for shopping and dining. Combining a flexible, intelligent platform with end-to-end payments capabilities and services developed through its deep industry experience, NCR Voyix empowers retailers and restaurants to accelerate new possibilities for their operations, experiences and business outcomes. NCR Voyix is headquartered in Atlanta, Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation Customer Unified Observability initiative. This strategic role will be responsible for building and evolving a unified enterprise observability platform that delivers end-to-end visibility across NCR Voyix Restaurants, Retail, and Payments environments.The ideal candidate will bring 10+ years of experience in Site Reliability Engineering, Platform Engineering, Cloud Operations, or related disciplines, with a proven track record of driving enterprise-scale observability, reliability, and operational excellence. This individual must be comfortable operating across organizational boundaries and partnering closely with Product Engineering, Infrastructure, Security, Operations, Architecture, and Executive Leadership teams to establish a comprehensive observability strategy and improve platform resilience.This role will serve as a key technical leader responsible for defining standards, influencing architecture decisions, and enabling proactive operations through unified monitoring, telemetry, automation, and AI-driven insights.Key ResponsibilitiesLead the architecture, design, implementation, and continuous improvement of enterprise observability solutions across Azure, Google Cloud Platform (GCP), Kubernetes, and hybrid environments.Establish and drive enterprise observability standards for monitoring, logging, distributed tracing, telemetry, and operational analytics.Develop and maintain executive, operational, and engineering dashboards that provide real-time visibility into infrastructure, applications, platform health, customer experience, and business transactions.Define, evangelize, and implement reliability frameworks including SLIs, SLOs, error budgets, operational KPIs, and service health metrics.Partner cross-functionally with Engineering, Infrastructure, Security, Product, and Operations teams to identify reliability risks and drive operational excellence initiatives.Lead efforts to improve incident prevention, detection, response, and recovery through intelligent alerting, automation, event correlation, and observability best practices.Integrate observability capabilities with ServiceNow, CI/CD pipelines, automation frameworks, and enterprise operational workflows.Influence technical strategy and roadmap decisions related to reliability engineering, platform observability, and operational readiness.Support and drive enterprise initiatives involving AI-driven observability, predictive analytics, anomaly detection, and event intelligence.Mentor engineers and serve as a subject matter expert for observability, reliability engineering, and cloud-native operations.Establish governance, adoption, and best practices across multiple product and engineering teams to ensure consistent observability standards enterprise-wide.Required QualificationsBachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience.10+ years of experience in Site Reliability Engineering, Cloud Engineering, Platform Engineering, DevOps, or related technical disciplines.Demonstrated success designing and operating observability platforms in large-scale enterprise environments.Deep expertise with Kubernetes platforms, including AKS and GKE.Strong experience with Azure and Google Cloud Platform services and architectures.Hands-on experience with enterprise observability tools such as Grafana, Datadog, Prometheus, OpenTelemetry, Dynatrace, New Relic, or similar platforms.Advanced knowledge of monitoring, logging, telemetry collection, distributed tracing, and observability engineering principles.Experience defining and operationalizing SLIs, SLOs, error budgets, reliability metrics, and service health frameworks.Strong automation and Infrastructure as Code expertise using Terraform and related tools.Proficiency developing automation solutions using Python, Go, PowerShell, or similar languages.Experience integrating observability solutions into CI/CD pipelines and modern DevOps workflows.Proven ability to influence technical direction and collaborate effectively with stakeholders across Engineering, Product, Infrastructure, Security, and Operations organizations.Strong communication, leadership, and stakeholder management skills with the ability to translate technical concepts for both technical and business audiences.Preferred QualificationsExperience leading enterprise observability transformations or platform modernization initiatives.Experience with AI Ops, event correlation, operational analytics, and predictive monitoring capabilities.Knowledge of ServiceNow integrations and ITSM/ITOM processes.Experience supporting highly available, customer-facing SaaS platforms at scale.One or more cloud certifications (Azure, Google Cloud, Kubernetes, or related technologies).Previous experience serving as a technical lead, mentor, or architect within a reliability engineering organization.Offers of employment are conditional upon passage of screening criteria applicable to the jobEEO StatementIntegrated into our shared values is NCR Voyix’s commitment to equal employment opportunity. All qualified applicants will receive consideration for employment without regard to sex, age, race, color, creed, religion, national origin, disability, sexual orientation, gender identity, veteran status, military service, genetic information, or any other characteristic or conduct protected by law. NCR Voyix is committed to being a globally inclusive company where all people are treated fairly, recognized for their individuality, promoted based on performance and encouraged to strive to reach their full potential. We believe in understanding and respecting differences among all people. Every individual at NCR Voyix has an ongoing responsibility to respect and support a globally diverse environment.Statement to Third Party AgenciesTo ALL recruitment agencies: NCR Voyix only accepts resumes from agencies on the preferred supplier list. Please do not forward resumes to our applicant tracking system, NCR Voyix employees, or any NCR Voyix facility. NCR Voyix is not responsible for any fees or charges associated with unsolicited resumes“When applying for a job, please make sure to only open emails that you will receive during your application process that come from a @ncrvoyix.com email domain.”SummaryLocation: ATLANTA, GA, USAType: Full time

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer - Unified Observability in Atlanta, GA vacancy
  •  ...Senior Site Reliability Engineer Atlanta, Georgia Who We Are QGenda is redefining healthcare...  ...strategic workforce decisions through our unified software platform. With more than 8...  ...the organization by promoting observability, retrospectives, and continuous improvement... 
    Senior
    Permanent employment
    Full time
    Work at office
    Remote work
    Work from home
    Work visa

    QGenda

    Atlanta, GA
    4 days ago
  • $104.9k - $174.7k

     ...the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the...  ...designing infrastructure, writing Terraform, improving observability, and responding to real production incidents.If you live... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    RELX Group

    Atlanta, GA
    3 days ago
  • $160k - $210k

     ...two coasts. Our culture thrives on cross-departmental collaboration and a unified sense of purpose, making teamwork a cornerstone of our success. We are looking for a Senior Site Reliability engineer to work on expanding our global footprint of datacenters and improve... 
    Senior
    Work at office
    Local area
    Immediate start
    Remote work

    GrabJobs

    Atlanta, GA
    1 day ago
  • $218.11k - $274.19k

     ...Site Reliability Engineer We are Omnissa! Omnissa is the first AI-driven digital work platform,...  ...industry-leading solutions—including Unified Endpoint Management, Virtual Apps and...  ...solutions and practices to improve the observability and autoremediation/self-healing of our... 
    Suggested
    Work experience placement
    Local area
    Remote work
    Visa sponsorship
    Flexible hours

    Omnissa

    Atlanta, GA
    4 days ago
  •  ...The Home Depot is seeking a Senior Software Reliability Engineer to join the Platform Reliability Engineering team, ensuring the resilience, performance, and security of our enterprise Cloud Platform. You will mentor junior engineers, lead incident triage, root cause... 
    Senior

    Home Depot

    Atlanta, GA
    4 days ago
  • $244k - $305k

     ...repositories into the monorepo. As a senior software engineer on the team, you will own...  ...that are simpler and more reliable to use.Push performance...  ...: Datadog is the leading observability and security platform for...  ...providing businesses with unified visibility across the... 
    Senior
    Work at office

    Datadog

    Atlanta, GA
    3 days ago
  • $172.5k - $260.1k

     ...team in the Service Delivery Platform & Reliability group at Slack develops platforms and...  ..., provide insights, and improves observability in Slack production services with a focus...  ...and work closely with other teams in engineering, product development, and customer experience... 
    Senior
    Full time

    Salesforce

    Atlanta, GA
    12 hours ago
  • $116.48k - $174.71k

     ...Ready Governance Platform, unifies regulatory intelligence,...  ...ChallengeWe're looking for a Senior Software Engineer that will report to the...  ...and implement application observability and platform monitoring tools...  ...budgets to balance system reliability with product feature... 
    Senior
    Work experience placement
    Work at office
    Local area
    Worldwide
    Flexible hours
    3 days per week
    1 day per week

    OneTrust

    Atlanta, GA
    1 day ago
  •  ...helping the world’s most important research sites do their best work. Our solutions are...  ...to the Team:We are seeking a Site Reliability Engineer (SRE) to join one of our Scrum teams...  ...while actively leveraging AI to improve observability, incident response, automation, and... 
    Work at office

    Florence Healthcare

    Atlanta, GA
    1 day ago
  • $100k - $120k

    OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and...  ...delivery.Cross trains colleagues on how to best leverage observability tools during incident and performance investigations.... 
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Origami Risk

    Atlanta, GA
    2 days ago
  • $152.13k - $162.13k

     ...about what’s next. Join us.General Summary:Unum Group seeks Site Reliability Engineers in Atlanta, GA.Applicants who are interested in this...  ...Ref #66753) for consideration.Design, build, and maintain observability, monitoring, and alerting capabilities across consumer and... 
    Full time
    Temporary work
    Work at office
    Remote work

    Unum Group

    Atlanta, GA
    1 day ago
  • $84.9k - $209.5k

     ...Job Description As a Principal Site Reliability Engineer (IC4), you will be responsible for designing...  ...infrastructure, coding, automation, observability, incident management, and operational...  ...incident response and act as a senior escalation point during major service... 
    Temporary work
    Flexible hours

    Oracle

    Atlanta, GA
    4 days ago
  •  ...cloud-native systems. As a Staff Platform Engineer, you will play a critical role in...  ...technical leadership role. You will own reliability for major platform domains, design scalable...  ...Establish and enhance centralized Observability and Monitoring platforms and tools that... 
    Senior

    Saviynt

    Atlanta, GA
    29 days ago
  •  ...The CNN Growth team is hiring a Senior Software Engineer to help build and evolve the systems...  ...hands on across the stack, delivering reliable, performant features while contributing...  ...AWS services, CI/CD pipelines, and observability tools. Requirements  4+ years of... 
    Senior
    Full time

    Warner Bros. Discovery

    Atlanta, GA
    1 day ago
  • $165k - $247.5k

     ...Governance Platform, unifies regulatory...  ...ChallengeWe are hiring a Senior Staff DevOps Engineer to join our Detect &...  ...CVEs; and ensuring the reliability, scalability, and security...  ...expertise on-site.Container Hardening...  ...code across the team.Observability & Platform Reliability... 
    Senior
    Work experience placement
    Work at office
    Local area
    Worldwide
    Flexible hours
    3 days per week
    1 day per week

    OneTrust

    Atlanta, GA
    3 days ago
  •  ...environments . Experienced in architectural design for reliability, scalability, and performance. Practical application of...  ..., Docker, serverless computing). Proficient in observability solutions such as Dynatrace, Prometheus, Grafana, and ELK/EFK... 

    Purple Drive

    Atlanta, GA
    5 days ago
  •  ...Exchange (ICE) presents a unique opportunity for a full-time Engineer to join a team responsible for ICE OpenShift container platform...  ...high availabilityExperience with platform and application observability (tracing, logging metrics)Top-tier analytics and problem solvingAbility... 
    Senior
    Full time

    Intercontinental Exchange

    Atlanta, GA
    3 days ago
  •  ...apply for the Junior Software Developer - Observability role at Canonical 3 days ago Be among...  ...as public cloud, data science, AI, engineering innovation, and IoT. Our customers include...  ...give your application fair consideration. Seniority level Entry level Employment type Full‑... 
    Full time
    Work at office
    Remote work
    Work from home

    Canonical

    Atlanta, GA
    3 days ago
  •  ...technology and services Position: Senior AI/ML Engineer Location: Atlanta GA or Frisco TX...  ...of customer records into accurate, unified profiles - and build the natural...  ...prompt engineering, evaluation, and LLM observability - ensuring AI outputs meet the trust... 
    Senior
    Temporary work
    Work experience placement

    Tekwissen

    Atlanta, GA
    2 days ago
  • $207.4k - $298.1k

     ...roadmap.About the role:We are hiring a Senior Principal Software Engineer to lead UKG Ready SMB —...  ...container orchestration (Kubernetes), observability, capacity planning and SRE practices...  ...cross-team integrations, operational reliability and cost-efficiency as Ready... 
    Senior
    Worldwide

    Ultimate Software

    Atlanta, GA
    12 hours ago
  •  ...Role: Site Reliability Engineering (SRE) Architect Location: Atlanta, GA (Hybrid on-site)...  ...across the organization and much beyond observability pillar. Key Responsibilities:...  ...Leadership & Consultation: Act as a senior technical advisor and subject matter... 
    Contract work
    Early shift

    AceStack LLC

    Atlanta, GA
    3 days ago
  • $120k - $150k

     ...following job description: The Senior ServiceNow Platform Engineer is responsible for the...  ..., automation, reliability, and continuous improvement...  ...Azure. • Experience with observability, monitoring, and operational...  ...please visit our Benefits site. Depending on the position... 
    Senior
    Permanent employment
    Full time
    Part time
    Work experience placement
    H1b
    Work at office
    Local area
    Immediate start
    Work visa
    Shift work
    Day shift

    Truist

    Atlanta, GA
    1 day ago
  • $228k - $342k

     ...We’re forming small, senior, cross-functional AI teams...  ..., machine learning engineers, and full-stack builders...  ...are scalable, observable, and enterprise-ready....  ...into systems that are reliable, explainable, and built...  ...Careers. Please be aware of sites that may ask for you... 
    Senior
    Full time
    Work at office
    Remote work
    Home office
    Flexible hours

    Workday

    Atlanta, GA
    12 hours ago
  •  ...A leading open-source technology firm is seeking a Junior Software Developer - Observability. In this role, you'll develop a cloud-native monitoring stack and work on exciting projects involving Linux and Kubernetes. Candidates should have skills in Python, a working knowledge... 
    Remote work

    Canonical

    Atlanta, GA
    3 days ago
  •  ...This role is pivotal to our platform engineering and site reliability initiatives, owning the Amazon EKS (...  ...product teams build and run. The Senior Engineer, Platform & Site Reliability...  ...the Istio service mesh, and modern observability to keep our systems secure, scalable... 
    Senior
    Work experience placement

    Black Knight Financial Services

    Atlanta, GA
    3 days ago
  •  ...seamless efficiency. When operations are unified on a single platform, districts gain...  ...to shape the future of education. Site Reliability Engineer (SRE) Overview: We are looking...  .... You'll work with leading-edge observability and reliability tooling, and the calls... 
    Full time
    Live in
    Work at office

    Incident IQ

    Atlanta, GA
    12 days ago
  •  ...including multi-threading, memory management and web services.History of building resilient, stateless, scalable, distributed, and observable systems.Experience in building REST services with high focus on performance.Familiarity with microservices and knowledge of... 
    Senior

    Intercontinental Exchange

    Atlanta, GA
    1 day ago
  • $112.8k - $257k

    Observability and AI Ops Architect, SeniorThe Opportunity:We are seeking...  ...with hands-on observability engineering to support government initiatives...  ...skills for engaging senior leadershipGoogle Professional...  ...Resource page on our Careers site and reviewing Our Employee Benefits... 
    Senior
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Atlanta, GA
    3 days ago
  • $143k - $191k

     ...technology to the military in months, not years.ABOUT THE TEAMThe Reliability Engineering team partners across Anduril's engineering, manufacturing,...  ...process and the security of our candidates. We've observed a rise in sophisticated phishing and fraudulent schemes where... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Atlanta, GA
    12 hours ago
  •  ...candidate to join our talented Team.Job Title: Senior Software EngineerLocation(s): Atlanta,...  ...seeking an experienced Senior Software Engineer to design, develop, and deliver scalable...  ...using CloudWatch, X-Ray, and other observability tools.Optimize cloud performance, security... 
    Senior

    Ampcus

    Atlanta, GA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer - Unified Observability. Be the first to apply!