Senior Site Reliability Engineer - Unified Observability
NCR
About NCR VOYIXNCR Voyix Corporation (NYSE: VYX) is a global platform-powered leader in unified commerce for shopping and dining. Combining a flexible, intelligent platform with end-to-end payments capabilities and services developed through its deep industry experience, NCR Voyix empowers retailers and restaurants to accelerate new possibilities for their operations, experiences and business outcomes. NCR Voyix is headquartered in Atlanta, Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation Customer Unified Observability initiative. This strategic role will be responsible for building and evolving a unified enterprise observability platform that delivers end-to-end visibility across NCR Voyix Restaurants, Retail, and Payments environments.The ideal candidate will bring 10+ years of experience in Site Reliability Engineering, Platform Engineering, Cloud Operations, or related disciplines, with a proven track record of driving enterprise-scale observability, reliability, and operational excellence. This individual must be comfortable operating across organizational boundaries and partnering closely with Product Engineering, Infrastructure, Security, Operations, Architecture, and Executive Leadership teams to establish a comprehensive observability strategy and improve platform resilience.This role will serve as a key technical leader responsible for defining standards, influencing architecture decisions, and enabling proactive operations through unified monitoring, telemetry, automation, and AI-driven insights.Key ResponsibilitiesLead the architecture, design, implementation, and continuous improvement of enterprise observability solutions across Azure, Google Cloud Platform (GCP), Kubernetes, and hybrid environments.Establish and drive enterprise observability standards for monitoring, logging, distributed tracing, telemetry, and operational analytics.Develop and maintain executive, operational, and engineering dashboards that provide real-time visibility into infrastructure, applications, platform health, customer experience, and business transactions.Define, evangelize, and implement reliability frameworks including SLIs, SLOs, error budgets, operational KPIs, and service health metrics.Partner cross-functionally with Engineering, Infrastructure, Security, Product, and Operations teams to identify reliability risks and drive operational excellence initiatives.Lead efforts to improve incident prevention, detection, response, and recovery through intelligent alerting, automation, event correlation, and observability best practices.Integrate observability capabilities with ServiceNow, CI/CD pipelines, automation frameworks, and enterprise operational workflows.Influence technical strategy and roadmap decisions related to reliability engineering, platform observability, and operational readiness.Support and drive enterprise initiatives involving AI-driven observability, predictive analytics, anomaly detection, and event intelligence.Mentor engineers and serve as a subject matter expert for observability, reliability engineering, and cloud-native operations.Establish governance, adoption, and best practices across multiple product and engineering teams to ensure consistent observability standards enterprise-wide.Required QualificationsBachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience.10+ years of experience in Site Reliability Engineering, Cloud Engineering, Platform Engineering, DevOps, or related technical disciplines.Demonstrated success designing and operating observability platforms in large-scale enterprise environments.Deep expertise with Kubernetes platforms, including AKS and GKE.Strong experience with Azure and Google Cloud Platform services and architectures.Hands-on experience with enterprise observability tools such as Grafana, Datadog, Prometheus, OpenTelemetry, Dynatrace, New Relic, or similar platforms.Advanced knowledge of monitoring, logging, telemetry collection, distributed tracing, and observability engineering principles.Experience defining and operationalizing SLIs, SLOs, error budgets, reliability metrics, and service health frameworks.Strong automation and Infrastructure as Code expertise using Terraform and related tools.Proficiency developing automation solutions using Python, Go, PowerShell, or similar languages.Experience integrating observability solutions into CI/CD pipelines and modern DevOps workflows.Proven ability to influence technical direction and collaborate effectively with stakeholders across Engineering, Product, Infrastructure, Security, and Operations organizations.Strong communication, leadership, and stakeholder management skills with the ability to translate technical concepts for both technical and business audiences.Preferred QualificationsExperience leading enterprise observability transformations or platform modernization initiatives.Experience with AI Ops, event correlation, operational analytics, and predictive monitoring capabilities.Knowledge of ServiceNow integrations and ITSM/ITOM processes.Experience supporting highly available, customer-facing SaaS platforms at scale.One or more cloud certifications (Azure, Google Cloud, Kubernetes, or related technologies).Previous experience serving as a technical lead, mentor, or architect within a reliability engineering organization.Offers of employment are conditional upon passage of screening criteria applicable to the jobEEO StatementIntegrated into our shared values is NCR Voyix’s commitment to equal employment opportunity. All qualified applicants will receive consideration for employment without regard to sex, age, race, color, creed, religion, national origin, disability, sexual orientation, gender identity, veteran status, military service, genetic information, or any other characteristic or conduct protected by law. NCR Voyix is committed to being a globally inclusive company where all people are treated fairly, recognized for their individuality, promoted based on performance and encouraged to strive to reach their full potential. We believe in understanding and respecting differences among all people. Every individual at NCR Voyix has an ongoing responsibility to respect and support a globally diverse environment.Statement to Third Party AgenciesTo ALL recruitment agencies: NCR Voyix only accepts resumes from agencies on the preferred supplier list. Please do not forward resumes to our applicant tracking system, NCR Voyix employees, or any NCR Voyix facility. NCR Voyix is not responsible for any fees or charges associated with unsolicited resumes“When applying for a job, please make sure to only open emails that you will receive during your application process that come from a @ncrvoyix.com email domain.”SummaryLocation: ATLANTA, GA, USAType: Full time
$104.9k - $174.7k
...the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the... ...designing infrastructure, writing Terraform, improving observability, and responding to real production incidents.If you live...SeniorFull timeWork at officeLocal areaRemote workWork from home- ...software developers, platform engineers, and IT staff to improve... ...requirements, service quality, reliability, security, and compliance... ...: 8+ years of experience in Site Reliability Engineering, DevOps... ..., maintaining, and maturing observability tooling including monitoring...SeniorWork at officeRemote work
- ...experienced and innovative leaders and engineers in the field. Where we work Headquartered... ...multimodal architectures — fast, reliable, and effortless to deploy at massive scale... ...own the reliability, performance, and observability of the entire inference stack. Your day...Senior
- ...The Home Depot is seeking a Senior Software Reliability Engineer to join the Platform Reliability Engineering team, ensuring the resilience, performance, and security of our enterprise Cloud Platform. You will mentor junior engineers, lead incident triage, root cause...Senior
$244k - $305k
...repositories into the monorepo. As a senior software engineer on the team, you will own... ...that are simpler and more reliable to use.Push performance... ...: Datadog is the leading observability and security platform for... ...providing businesses with unified visibility across the...SeniorWork at office$172.5k - $260.1k
...team in the Service Delivery Platform & Reliability group at Slack develops platforms and... ..., provide insights, and improves observability in Slack production services with a focus... ...and work closely with other teams in engineering, product development, and customer experience...SeniorFull time$116.48k - $174.71k
...Ready Governance Platform, unifies regulatory intelligence,... ...ChallengeWe're looking for a Senior Software Engineer that will report to the... ...and implement application observability and platform monitoring tools... ...budgets to balance system reliability with product feature...SeniorWork experience placementWork at officeLocal areaWorldwideFlexible hours3 days per week1 day per week$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and... ...delivery.Cross trains colleagues on how to best leverage observability tools during incident and performance investigations....Full timeTemporary workWork experience placementFlexible hours- ...cloud-native systems. As a Staff Platform Engineer, you will play a critical role in... ...technical leadership role. You will own reliability for major platform domains, design scalable... ...Establish and enhance centralized Observability and Monitoring platforms and tools that...Senior
$165k - $247.5k
...Governance Platform, unifies regulatory... ...ChallengeWe are hiring a Senior Staff DevOps Engineer to join our Detect &... ...CVEs; and ensuring the reliability, scalability, and security... ...expertise on-site.Container Hardening... ...code across the team.Observability & Platform Reliability...SeniorWork experience placementWork at officeLocal areaWorldwideFlexible hours3 days per week1 day per week- ...Exchange (ICE) presents a unique opportunity for a full-time Engineer to join a team responsible for ICE OpenShift container platform... ...high availabilityExperience with platform and application observability (tracing, logging metrics)Top-tier analytics and problem solvingAbility...SeniorFull time
$207.4k - $298.1k
...roadmap.About the role:We are hiring a Senior Principal Software Engineer to lead UKG Ready SMB —... ...container orchestration (Kubernetes), observability, capacity planning and SRE practices... ...cross-team integrations, operational reliability and cost-efficiency as Ready...SeniorWorldwide$228k - $342k
...We’re forming small, senior, cross-functional AI teams... ..., machine learning engineers, and full-stack builders... ...are scalable, observable, and enterprise-ready.... ...into systems that are reliable, explainable, and built... ...Careers. Please be aware of sites that may ask for you...SeniorFull timeWork at officeRemote workHome officeFlexible hours- ...This role is pivotal to our platform engineering and site reliability initiatives, owning the Amazon EKS (... ...product teams build and run. The Senior Engineer, Platform & Site Reliability... ...the Istio service mesh, and modern observability to keep our systems secure, scalable...SeniorWork experience placement
- ...A leading open-source technology firm is seeking a Junior Software Developer - Observability. In this role, you'll develop a cloud-native monitoring stack and work on exciting projects involving Linux and Kubernetes. Candidates should have skills in Python, a working knowledge...Remote work
$143k - $191k
...technology to the military in months, not years.ABOUT THE TEAMThe Reliability Engineering team partners across Anduril's engineering, manufacturing,... ...process and the security of our candidates. We've observed a rise in sophisticated phishing and fraudulent schemes where...SeniorFull timeWork experience placementImmediate start$112.8k - $257k
Observability and AI Ops Architect, SeniorThe Opportunity:We are seeking... ...with hands-on observability engineering to support government initiatives... ...skills for engaging senior leadershipGoogle Professional... ...Resource page on our Careers site and reviewing Our Employee Benefits...SeniorFull timeContract workPart timeWork at officeLocal areaRemote work- ...seamless efficiency. When operations are unified on a single platform, districts gain... ...to shape the future of education. Site Reliability Engineer (SRE) Overview: We are looking... .... You'll work with leading-edge observability and reliability tooling, and the calls...Full timeLive inWork at office
- ...including multi-threading, memory management and web services.History of building resilient, stateless, scalable, distributed, and observable systems.Experience in building REST services with high focus on performance.Familiarity with microservices and knowledge of...Senior
- ...candidate to join our talented Team.Job Title: Senior Software EngineerLocation(s): Atlanta,... ...seeking an experienced Senior Software Engineer to design, develop, and deliver scalable... ...using CloudWatch, X-Ray, and other observability tools.Optimize cloud performance, security...Senior
$140.88k - $153.75k
General Information Job Title Senior Platform Engineer Job ID 105799 Work Areas Technology... ...services are secure-by-default, observable from day one, and operable under... ...development: testing, observability, security, reliability, and operational hygiene.Conduct...SeniorPermanent employmentFull timeContract workWork at officeLocal area$101.5k - $169.1k
...an incentive program.Job DescriptionCox Automotive’s Engineering & Technology organization is building a centralized Enterprise... ...Claude Code) to enterprise data sources in a secure, observable, and self-service way. This Senior Platform Engineer will own the technical...SeniorFull timeRemote workFlexible hours- ...of the largest, most reliable partnership platforms... ...relied on a trusted review site to make your decision?... ...location ()About CJ Engineering Are we a good fit for... ...CI/CD, automation, and observability to build reliable... ...to hear from you! As a Senior Software Engineer on the...SeniorTemporary workFreelanceWork at officeFlexible hours3 days per week
- ...apply for the Junior Software Developer - Observability role at Canonical 3 days ago Be among... ...as public cloud, data science, AI, engineering innovation, and IoT. Our customers include... ...your application fair consideration. Seniority level Entry level Employment type Full...Full timeWork at officeRemote workWork from home
$142k - $210k
Atlanta (Remote Friendly)Engineering /Full Time /RemoteGreenlight is... ...Greenlight is looking for a Senior Software Engineer, Full-Stack... ...user base.Improving the reliability and availability of the platform... ..., scalability, and observability when designing and building...SeniorFull timeWork at officeLocal areaRemote workWork from homeFlexible hoursDay shift- ...prioritizationPractical knowledge of software engineering best practices, including Agile... ...distributed systems using cloud-native observability tooling, experience with cloud-native architecture... ...you and your family.WHAT YOU'LL DOAs a Senior Software Engineer with the Life...SeniorApprenticeshipEasy work
$148.5k - $313.7k
...ExperienceNote: By applying to the Senior / Lead / Principal Software Engineer - Foundations Team posting,... ...cares deeply about performance, reliability, and engineering excellence — driving... ...robust CI/CD pipelines, advanced observability, and top-tier engineering tools....SeniorFull timeWork experience placementRemote work$183.8k - $263.6k
...integration. You will work closely with engineers across control plane, data plane, and platform... ...capabilities that support deployment, observability, and policy enforcement across... ...insurance. Please see the Cisco careers site to discover more benefits and perks. Employees...SeniorFull timeTemporary workLocal areaRemote workFlexible hours- ...global platform-powered leader in unified commerce for shopping and... ...worldwide.TITLE: SW Engineer IV LOCATION: Atlanta, GA About... ...industry. Our products are highly reliable, scalable, and configurable and... ...brands in the world. In this senior role, you will join a team building...SeniorFull timeWorldwideFlexible hours
$119.85k - $162.15k
...just a software provider; we are the unified subledger that eliminates time-intensive... ...Matter MostFinQuery is looking for a Senior Software Engineer to join our Engineering team. The Senior... ..., and guiding work through to reliable, production-quality outcomes. The position...SeniorFull timeContract workWork experience placementCasual workWork at officeImmediate startWork from homeWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer - Unified Observability. Be the first to apply!
- site reliability engineer Atlanta, GA
- senior business analyst Atlanta, GA
- senior cost estimator Atlanta, GA
- senior manager tax Atlanta, GA
- senior automation engineer Atlanta, GA
- senior devops Atlanta, GA
- senior recruiter Atlanta, GA
- senior property manager Atlanta, GA
- senior construction estimator Atlanta, GA
- senior paralegal Atlanta, GA

