Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Site Reliability Engineer

$184.3k - $264.95k

UKG, Inc.

Principal Site Reliability Engineer

At UKG, the work you do matters. The code you ship, the decisions you make, and the care you show a customer all add up to real impact. Today, tens of millions of workers start and end their days with our workforce operating platform. Helping people get paid, grow in their careers, and shape the future of their industries. That's what we do.

We never stop learning. We never stop challenging the norm. We push for better, and we celebrate the wins along the way. Here, you'll get flexibility that's real, benefits you can count on, and a team that succeeds together. Because at UKG, your work matters—and so do you.

Principal Site Reliability Engineers (SREs) at UKG are strategic technical leaders who play a critical role in shaping the reliability, scalability, and performance of our organization's services, at scale. They bring deep expertise across service delivery, infrastructure, and cloud architecture, and apply software engineering principles to solve complex operational challenges. Principal SREs drive technical vision and strategy across multiple teams and business units.

In this role, you will define the operational reliability strategy, architect large-scale solutions, and establish the standards and practices that enable teams across UKG to build and operate highly reliable services. This includes guiding the evolution of CI/CD ecosystems, designing automated testing frameworks, building capacity planning and performance analysis systems, establishing observability standards, and creating reliability automation strategies as self-healing and remediation at scale.

Principal SREs are passionate about driving business outcomes through technical excellence. They shape operational culture around reliability, mentor and develop SRE talent across teams, and relentlessly pursue the highest standards of customer experience through strategic automation and process innovation.

This is a principal-level individual contributor and technical leadership role, focused on operational reliability strategy, cross-functional influence, and enterprise-scale impact.

Define and evolve the organization's Site Reliability Engineering strategy, vision, and technical roadmap in partnership with engineering leadership and business stakeholders.

Guide the lifecycle of services from conception to end-of-life across the organization, including establishing service design review frameworks, capacity planning methodologies, and production readiness standards that scale across teams.

Establish and drive adoption of organization-wide standards and best practices related to system architecture, service delivery, reliability, automation, and operational excellence.

Influence technical decisions across multiple teams and business units.

Build and operate shared platforms, tooling, and frameworks that enable service and engineering teams across the organization to achieve high availability and improve incident detection and response.

Drive system performance, availability, and efficiency improvements across the organization through architectural guidance, automation strategy, process refinement, and deep analysis of operational patterns in incident reviews.

Partner closely with engineering and product leadership across the organization to establish reliability standards, shape technical direction, and deliver reliable services at scale.

Champion a culture of operational excellence by treating operational challenges such as software engineering problems and mentoring teams on reducing toil at scale.

Lead, mentor, and develop Site Reliability Engineering talent across multiple teams, establishing best practices and growing SRE capability within the organization.

Partner with executive stakeholders, product leadership, and business teams to align reliability investments with business priorities and drive technical strategy that supports business outcomes.

12+ years of hands-on experience in software engineering, systems engineering, and/or cloud-based environments.

10+ years of experience working with public cloud platforms (e.g., GCP, AWS, or Azure), including designing and operating large-scale systems.

5+ years of experience designing, operating, and maintaining applications and/or systems infrastructure in large-scale, customer-facing production environments.

Demonstrated experience architecting and influencing large-scale distributed systems and infrastructure solutions.

Proven track record of leading cross-functional technical initiatives and mentoring engineering teams.

Demonstrated understanding of observability best practices, including metric generation and collection, log aggregation pipelines, time-series databases, and distributed tracing.

Experience coding in one or more higher-level programming languages (e.g., Python, Java, C# or C++).

Strong working knowledge of Linux systems, including troubleshooting, performance analysis, and scripting in production environments.

Experience with GitHub Actions and modern CI/CD practices.

Exceptional communication and collaboration skills, with demonstrated ability to influence across teams, lead technical discussions with executives, and mentor engineers.

Experience defining and communicating technical vision and strategy across large organizations.

Deep expertise in distributed system design and architecture.

Hands-on experience with cloud-native applications and containerization technologies (e.g., Kubernetes, containers).

Experience designing and managing infrastructure-as-code and configuration management strategies across organizations (e.g., Terraform, Ansible).

Experience operating and optimizing production workloads at scale, including cost optimization and performance tuning.

Solid grounding in at least three of the following areas: Computer Science fundamentals, Cloud Architecture, Security or Network Design.

Experience designing observability strategies and building metrics pipeline, operational dashboards and alerts using observability tools such as Splunk or Grafana.

Experience partnering with business and product teams on technology decisions and translating technical recommendations into business value.

Experience driving operational change and building consensus around new technical practices and standards.

UKG is the Workforce Operating Platform that puts workforce understanding to work. With the world's largest collection of workforce insights, and people-first AI, our ability to reveal unseen ways to build trust, amplify productivity, and empower talent, is unmatched. It's this expertise that equips our customers with the intelligence to solve any challenge in any industry — because great organizations know their workforce is their competitive edge.

Equal Opportunity Employer

UKG is an equal opportunity employer. We evaluate qualified applicants without regard to race, color, disability, religion, sex, age, national origin, veteran status, genetic information, and other legally protected categories.

The pay range for this position is $184,300 to $264,950. The actual base pay offered may vary depending on skills, experience, job-related knowledge and work location. In addition to base pay, employees may be eligible to participate in a performance-based bonus plan and to receive restricted stock unit awards as part of total compensation.

Vacancy posted 13 hours ago
Similar jobs that could be interesting for youBased on the Principal Site Reliability Engineer in Seattle, WA vacancy
  •  ...customers depend on every day. We're hiring a senior, hands-on engineer to own the reliability, availability, security, and performance of that platform...  ...Who you are: ~5+ years of hands-on Cloud Operations and Site Reliability Engineering, operating production-scale SaaS (... 
    Suggested
    Full time

    MangoApps

    Seattle, WA
    2 days ago
  • $127k - $249k

    Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper). As... 
    Suggested
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Seattle, WA
    7 hours ago
  • $134.25k - $214.8k

     ...change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability and performance... 
    Suggested
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Axon

    Seattle, WA
    3 days ago
  • $134.25k - $214.8k

     ...upload. Every piece of digital evidence. Every chain of custody log that holds up in court. That's us.Axon's Platform team is the engine behind what hundreds of thousands of officers rely on every day. We're one of the world's largest blob storage customers, ingesting... 
    Suggested
    Work experience placement
    Work at office
    Remote work

    Axon

    Seattle, WA
    1 day ago
  • $55k - $151.47k

     ...ApplicableSpecialismIFS - Internal Firm Services - OtherManagement LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our... 
    Suggested
    Full time
    H1b

    PwC

    Seattle, WA
    2 days ago
  •  ...This is an engineering-first Senior SRE role. We’re looking for senior engineers who have: Built and shipped significant backend...  ...services end-to-end in production (design → launch → on-call → reliability improvements) Led incident response and driven durable... 

    Practice by Numbers

    Bellevue, WA
    3 days ago
  •  ...Senior Site Reliability Engineer (SRE) Location: Seattle, hybrid - 2 times a week in the office Job Type: Full-time, direct hire Industry: High-Growth Technology / SaaS About the Role We are seeking a highly skilled Senior Site Reliability Engineer to... 
    Full time
    Work at office

    TalentDome Staffing

    Seattle, WA
    2 days ago
  • Company DescriptionComtech LLC is a woman-owned small business focused on delivering end-to-end solutions and products. Since 1998, we have successfully serviced enterprises across the public and private sectors, and the Department of Defense. Our services span all aspects...

    Comtech

    Seattle, WA
    7 hours ago
  •  ...The RoleThis hybrid role combines the hands-on responsibilities of a Technical Support Engineer within a SaaS (Software as a Service) environment with a growing focus on Site Reliability Engineering (SRE).The ideal candidate has a strong technical foundation, thrives in a... 
    Full time
    Local area

    F5 Networks

    Seattle, WA
    1 day ago
  • $143k - $194k

     ...mission critical capabilities to our customers. System Deployment Engineers work in complex environments with shared environmental...  ...customersParticipate in customer demonstrations and exercisesWork with site reliability engineers to provide and refine requirements for tooling and... 
    Full time
    Temporary work
    Work experience placement
    Immediate start

    Anduril Industries

    Seattle, WA
    3 days ago
  • $151.2k - $204.6k

    Would you like to be an engineer who builds the systems that power advertising at scale,...  ...advertising queries every day, where latency, reliability, and quality translate directly into...  ...Development Engineer, operating as a Site Reliability Engineer, to raise the reliability... 
    Flexible hours

    Amazon

    Seattle, WA
    1 day ago
  •  ...system and process health, performance, and reliability, including alerting quality, dashboards,...  ...while aligning stakeholders across IT, engineering and partner teams. Proactively...  ...experience. ~5+ years of experience in site reliability engineering, infrastructure... 

    Blink Health

    Seattle, WA
    2 days ago
  • $170k - $219k

     ...Site Reliability Engineer RADAR runs data infrastructure across 1,600+ live retail stores, processing tens of billions of real-world events every day. We're hiring a Site Reliability Engineer to own the reliability of that system end to end — leading incident response... 
    Flexible hours
    Shift work
    Night shift

    Radar

    Seattle, WA
    2 days ago
  • $160k - $250k

     ...DevOps And Systems Engineer Hive is the leading provider of cloud-based AI solutions to understand, search, and generate content...  ...machine learning models, we also need to grow our DevOps and Site Reliability team to maintain the reliability of our enterprise SaaS offering... 

    Hive

    Seattle, WA
    2 days ago
  •  ...certification), ISO 27001:2005 Information Security Management System (ISMS), and CMMI-DEV Level 3. Job Description Sr. Site Reliability Engineer Location – Seattle, WA Duration – 12 months Interview – in-person if local or Phone + Skype Minimum... 
    Local area
    Worldwide

    Comtech LLC

    Seattle, WA
    2 days ago
  • $204k - $306k

     ...all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity, from...  ...in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions millions... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours
    2 days per week

    Okta

    Bellevue, WA
    3 days ago
  •  ...where you can push the limits of what's possible.As a Lead Software Engineer at JPMorganChase within the Enterprise Technology,...  .... These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare... 

    JP Morgan Chase

    Seattle, WA
    1 day ago
  • $194k - $267k

     ...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Bellevue, WA
    3 days ago
  • $194k - $267k

     ...do something more than once, automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Bellevue, WA
    4 days ago
  •  ...future together. We are responsible for the reliability of all the company's major data warehouse products, services, and query engines. We serve business needs across domains...  ...practices, and emerging technologies related to site reliability and infrastructure engineering.... 

    TikTok

    Seattle, WA
    7 hours ago
  • $163k - $244.6k

    We take play seriously. We’re looking for curious adventurers ready to find their party, fueled by imagination and drive to build what’s never been built before. At Hasbro and Wizards of the Coast, you’ll collaborate with passionate teams to reimagine our iconic brands ...
    Principal

    Hasbro

    Renton, WA
    4 days ago
  •  ...We're seeking an SRE to ensure the reliability and performance of our clients' critical systems. You'll work on observability, incident...  ...management Nice to have Experience with chaos engineering Knowledge of distributed systems Background in high-scale... 
    Remote work
    Flexible hours

    ACI Infotech

    Seattle, WA
    1 day ago
  •  ...Infrastructure SRE team is responsible for the reliability, scalability, and efficiency of the core...  ...not about building features, but about engineering the resilience and performance of the...  ...to maintain system stability.As a Site Reliability Engineer, you will be on the... 

    TikTok

    Seattle, WA
    7 hours ago
  • SingleStore engineers build the real-time data platform powering some of the world’s most...  ...Position SummaryWe are seeking a Senior/Principal Software Engineer to join the Engineering...  ...Demonstrated ability to design and build highly reliable, high-performance system software.... 
    Principal

    SingleStore

    Seattle, WA
    4 days ago
  •  ...on one unified cloud. One cloud for compute, inference, and agents. Role Overview We are seeking a skilled Site Reliability Engineer to join the GMI Global Infrastructure team. This role is hands-on and critical to ensuring the stability, efficiency, and... 

    GMI Cloud

    Seattle, WA
    14 hours ago
  • $95k - $134k

     ...disrupting the industry it helped build. Job Application Deadline: 10/31/2026 The Opportunity DAT is looking for a Site Reliability Engineer to join our SRE platform team. This position will work hybrid in Seattle, WA or Portland, OR Candidate profile DAT... 
    Temporary work
    For contractors
    Work experience placement
    Work at office
    Local area
    Immediate start
    Flexible hours

    DAT Freight & Analytics

    Seattle, WA
    3 days ago
  •  ...Engineering, Product, Design, and Marketing Engineering Compensation ~ Zone 1 Base Pay: $214K – $260K Superhuman offers...  ...role will be responsible for building software to ensure the reliability of our back-end systems, working with engineers who develop them... 
    Worldwide
    Home office
    Flexible hours

    Superhuman

    Seattle, WA
    3 days ago
  • $100k - $120k

     ...Senior Engineer Must Have Technical/Functional Skills ~8+ years of SRE, DevOps, or production operations experience supporting...  ...development and QA teams. Design, build, and maintain scalable, reliable, and high-performance systems and environments in Azure.... 
    Permanent employment

    Tata Consultancy Services

    Seattle, WA
    2 days ago
  • Job Title Required Skills: CHEF experience - Must have most critical Azure Cloud – experience - Must have most critical AKS- Azure Kubernetes services - Must have most critical Kubernetes - Must have most critical NoSQL DB – Cassandra / Mongo DB ...

    Syntricate Technologies

    Seattle, WA
    2 days ago
  •  ...Lululemon Site Reliability Engineering Engineer We are a yoga-inspired technical apparel company up to big things. The practice and philosophy of yoga informs our overall purpose to elevate the world through the power of practice. We are proud to be a growing global... 

    Samprasoft

    Seattle, WA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!