Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Site Reliability Engineer

LivePerson

Principal Site Reliability Engineer

LivePerson transforms customer care from voice calls to mobile messaging. Our cloud-based software platform, LiveEngage, allows brands with millions of customers and tens of thousands of care agents to deliver digital experiences at scale. As the market leader in real-time intelligent customer engagement, we are a B2B SaaS company with 20 years of experience and the heart of a startup.

The Cloud DevOps team at LivePerson is looking for a Principal Site Reliability Engineer (Principal SRE) to provide technical leadership across the organization and help shape the reliability, scalability, security, and operational excellence of our cloud platforms and services.

The ideal candidate is a highly experienced engineer who can solve complex technical problems, influence engineering teams without direct authority, and drive large-scale initiatives from strategy through implementation. This is a highly technical role focused on architecture, reliability engineering, automation, and improving the overall engineering maturity of the organization.

You will:

  • Provide technical leadership and direction for reliability and platform engineering across multiple teams.
  • Design and evolve highly available, scalable, secure, and resilient systems, with a strong focus on Google Cloud Platform (GCP).
  • Lead complex, cross-team initiatives across cloud infrastructure, Kubernetes, networking, observability, security, and software delivery.
  • Define and drive SRE practices including SLOs, SLIs, error budgets, reliability reviews, capacity planning, and operational readiness.
  • Lead technical response to complex production incidents and drive long-term corrective and preventative actions.
  • Develop automation and infrastructure-as-code solutions using Python, Terraform, Ansible, Bash, and other modern engineering tools.
  • Provide technical leadership for Kubernetes platforms and containerized workloads, including architecture, scalability, performance, and reliability.
  • Establish and evolve GitOps deployment practices using Kubernetes, Helm, and FluxCD. Define and improve CI/CD practices using GitLab CI/CD, focusing on reliability, security, scalability, and developer experience.
  • Drive observability improvements using metrics, logs, traces, dashboards, and actionable alerting. Identify systemic reliability risks, technical debt, and architectural weaknesses and drive sustainable solutions.
  • Create reusable platforms, tooling, and engineering patterns that enable teams to operate reliable services independently. Mentor Senior SREs, Team Leads, and other technical leaders while raising engineering standards across the organization. Influence architecture and technical decisions across teams and communicate complex technical concepts and trade-offs to engineering leadership.

You have:

  • Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
  • 7+ years of experience in SRE, DevOps, Cloud Infrastructure, Systems Engineering, or a related field.
  • Experience operating at a Principal, Staff+, Architect, or equivalent senior technical leadership level.
  • Extensive experience designing and operating highly available, scalable, secure, and distributed production systems.
  • Strong hands-on experience with Google Cloud Platform (GCP) and cloud architecture, including networking, IAM, compute, storage, and managed services.
  • Extensive experience with Kubernetes and containerization technologies such as Docker.
  • Hands-on experience with service meshes such as Istio.
  • Strong programming and scripting experience with Python, Bash, or similar languages, focused on automation and platform engineering.
  • Extensive experience with Terraform and automation/configuration management tools such as Ansible.
  • Strong experience with GitOps, Helm, and FluxCD.
  • Strong experience designing and maintaining CI/CD pipelines using GitLab CI/CD.
  • Deep understanding of Linux, networking, DNS, load balancing, TLS/SSL, authentication, and security fundamentals.
  • Strong experience with observability platforms such as Prometheus, Grafana, and Alertmanager.
  • Deep understanding of SRE principles, distributed systems, scalability, fault tolerance, and performance engineering.
  • Proven ability to influence architecture and technical decisions across multiple teams and organizational boundaries.
  • Excellent communication skills and the ability to mentor senior engineers and technical leaders.

Nice to Have:

  • Experience operating large-scale B2B SaaS or globally distributed systems.
  • Experience with secrets management and security platforms such as HashiCorp Vault.
  • Experience with PostgreSQL and other distributed data services.
  • Experience operating and modernizing hybrid cloud and on-premises infrastructure.
  • Experience leading large-scale cloud migrations or infrastructure modernization initiatives.
  • Experience designing internal developer platforms and self-service infrastructure capabilities.
  • Experience with disaster recovery, business continuity, and resilience engineering.
  • Experience defining organization-wide engineering standards and reliability frameworks.

Benefits:

  • Health: medical, dental, and vision
  • Time away: 28 vacation days
  • Development: Generous tuition reimbursement and access to internal professional development resources.
  • Additional: Food Vouchers.
  • #LI-Remote

Why you'll love working here:

As leaders in enterprise customer conversations, we celebrate diversity, empowering our team to forge impactful conversations globally. LivePerson is a place where uniqueness is embraced, growth is constant, and everyone is empowered to create their own success. And, we're very proud to have earned recognition from Fast Company, Newsweek, and BuiltIn for being a top innovative, beloved, and remote-friendly workplace.

Belonging at LivePerson:

We are proud to be an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable laws, regulations and ordinances. We also consider qualified applicants with criminal histories, consistent with applicable federal, state, and local law.

We are committed to the accessibility needs of applicants and employees. We provide reasonable accommodations to job applicants with physical or mental disabilities. Applicants with a disability who require reasonable accommodation for any part of the application or hiring process should inform their recruiting contact upon initial connection.

The talent acquisition team at LivePerson has recently been notified of a phishing scam targeting candidates applying for our open roles. Scammers have been posing as hiring managers and recruiters in an effort to access candidates' personal and financial information. This phishing scam is not isolated to only LivePerson and has been documented in news articles and media outlets. Please note that any communication from our hiring teams at LivePerson regarding a job opportunity will only be made by a LivePerson employee with an @ liveperson.com email address.

LivePerson does not ask for personal or financial information as part of our interview process, including but not limited to your social security number, online account passwords, credit card numbers, passport information and other related banking information. If you have any questions and or concerns, please feel free to contact View email address on click.appcast.io

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Principal Site Reliability Engineer in United States vacancy
  • $139.7k - $232.9k

     ...designing, implementing, and continuously improving highly reliable, scalable, and resilient platform solutions across the enterprise. Operates as a subject matter expert (SME) in Site Reliability Engineering, driving reliability engineering practices, operational excellence... 
    Principal
    Full time
    Work experience placement

    M&T Bank

    Buffalo, NY
    2 days ago
  • $142.8k - $274.8k

     ...yearEmployment type: Full-TimeWork site: 0 days / week in-office - remoteRole...  ...Software EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...world’s most demanding workloads. As a Principal Site Reliability Engineer, you will set technical and operational... 
    Principal
    Ongoing contract
    Work at office
    Local area

    Microsoft

    Redmond, WA
    3 days ago
  •  ...lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together.We are seeking a Principal Site Reliability Engineer (SRE) to define and scale reliability practices across large-scale cloud platforms.This is a senior individual contributor... 
    Principal
    Minimum wage
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work

    UnitedHealth Group

    Minnetonka, MN
    1 day ago
  • $134.6k - $230.8k

     ...Connecting. Growing together.Are you passionate about reimagining operations through AI? Optum Financial is seeking a Principal Site Reliability Engineer to lead the evolution of our reliability platform by combining modern SRE practices with AI-assisted operations. You'... 
    Principal
    Minimum wage
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work

    UnitedHealth Group

    Eden Prairie, MN
    8 hours ago
  • $248k - $396.75k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused on designing, building, and operating large-scale production...  ...strengthen the reliability of production environments.As a Principal SRE, you will shape the technical direction of NVIDIA’s... 
    Principal
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...Infrastructure Code. Builds reliability into the ecosystem by applying...  ...practices in resiliency engineering and observability by developing...  ...engineering techniques with site reliability engineering...  ...5) years of experience as a Principal Site Reliability Engineer (or... 
    Principal
    Full time

    Fidelity Investments

    Westlake, OH
    1 day ago
  •  ...Continuous Delivery (CI/CD) pipelines and Kubernetes. Supports Site Reliability Engineering (SRE) functions by establishing Service Level Objectives...  ...equivalent) and five (5) years of experience as a Principal Site Reliability Engineer (or closely related occupation)... 
    Principal
    Full time

    Fidelity Investments

    Durham, NC
    2 days ago
  •  ...To enhance the reliability and performance of production systems, the full-time Principal Site Reliability Engineer will lead incident management, implement SRE practices, and drive infrastructure automation in a fully remote environment. Key responsibilities: Manage... 
    Principal
    Full time
    Remote work

    Virtual Vocations Inc

    United States
    2 days ago
  • About this role:Shape the future of reliability, resilience, and customer experience at scale. Wells Fargo is seeking a highly influential Principal Engineer to lead the transformation of observability, automation, and Site Reliability Engineering (SRE) practices supporting... 
    Principal
    Full time
    Work experience placement
    Visa sponsorship
    3 days per week

    Wells Fargo

    Charlotte, NC
    1 day ago
  • $84.9k - $209.5k

    This role combines strategic architecture with practical systems engineering, deployment, automation, patching, troubleshooting, incident response, and compliance support. The Principal Site Reliability Engineer will work across Windows, Linux, Oracle Cloud Infrastructure... 
    Principal
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    1 day ago
  • $84.9k - $209.5k

     .... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers...  ...posted.Career Level - IC4Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and... 
    Principal
    Temporary work
    Monday to Friday
    Flexible hours
    Shift work
    Night shift

    Oracle Corporation

    Reston, VA
    8 hours ago
  • $159k - $272k

     ...generosity. Join us for the opportunity to grow and make a difference in ways that matter to you. Role SummaryIn this role as Principal Site Reliability Engineer, Infrastructure Observability you will help formulate, develop, and implement a team of Site Reliability Engineers (... 
    Principal
    Full time
    Private practice
    Local area
    Remote work
    Work from home
    3 days per week

    T. Rowe Price

    Owings Mills, MD
    1 day ago
  •  ...Principal Site Reliability Engineer The Principal Site Reliability Engineer will be a critical technical leader responsible for driving the operational excellence, resilience, and security of our core systems for a key Randstad client in the Washington D.C. area. This... 
    Principal

    Software Technology Inc

    Washington DC
    3 days ago
  • $96.3k - $264.1k

     ...infrastructure and service, ensuring alignment with reliability and functionality standards. Takes full...  ...tools and provides expertise in site reliability trends.Only Oracle brings...  ...LeadershipDefine and drive the site reliability engineering strategy for large-scale, distributed,... 
    Principal
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    6 days ago
  • $198.24k - $272.58k

    We’re looking for a Principal Site Reliability Engineer to join Procore’s Compute Division to work on our FedRAMP initiative. In this role, you’ll help build Procore’s next-generation construction compute platform for others to build upon, including Procore developers,... 
    Principal
    Full time
    Contract work
    Work at office
    Local area
    Immediate start

    Procore Technologies

    Austin, TX
    8 hours ago
  • $272k - $431.25k

    NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's (Infrastructure, Planning and Process) Cloud Infrastructure Team. IPP is a global organization within NVIDIA. This group works with various other groups within NVIDIA such as Graphics... 
    Principal
    Full time
    Work experience placement
    Worldwide

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...with software development teams to build reliable, scalable, secure, and cloud-native...  ...influence scalable architecture patterns across engineering teams, helping ensure systems are...  ...~8+ years of hands-on experience in Site Reliability Engineering, DevOps, cloud infrastructure... 
    Principal
    Remote work

    ABC Fitness Solutions, LLC

    United States
    5 days ago
  • $194k - $237k

    ## Principal Site Reliability EngineerApplylocations: Scottsdaletime type: Full timeposted on: Posted 5 Days Agojob requisition id: REQ2026426At...  ....**Overall Purpose**The Principal Site Reliability Engineer partners with development teams by designing availability... 
    Principal
    Hourly pay
    Work at office
    Immediate start
    Visa sponsorship
    Work visa
    Flexible hours

    Early Warning Services

    Scottsdale, AZ
    4 days ago
  • $240k - $250k

     ...MattersSaviynt’s platform is mission-critical for our customers. As we scale globally, reliability, availability, and performance are not optional—they are core product features.As a Principal Engineer, you will define and drive the reliability strategy for our SaaS platform.... 
    Principal
    Full time

    Saviynt

    Atlanta, TX
    2 days ago
  •  ...disrupt, and thrive! KēSTA I.T. is actively seeking a Principal Engineer for an immediate full-time opportunity with our industry creating...  ...An innovative technology company is seeking experienced Site Reliability Engineers to take ownership of building reliable, scalable... 
    Principal
    Permanent employment
    Full time
    Temporary work
    Immediate start

    KēSTA I.T.

    Beverly Hills, CA
    22 days ago
  •  ...Job Description Job Description Principal Site Reliability Engineer About ShipperHQ: ShipperHQ is a trusted leader in the e-commerce shipping space, with over 15 years of experience helping merchants deliver better checkout experiences. Founded in 2009, we power... 
    Principal
    Full time
    Work at office

    ShipperHQ

    Austin, TX
    21 days ago
  •  ...lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together.We are seeking a Principal Site Reliability Engineer (SRE) to define and scale reliability practices across large-scale cloud platforms.This is a senior individual contributor... 
    Principal
    Minimum wage
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work

    UnitedHealth Group

    Minnetonka, MN
    23 hours ago
  •  ...serve.The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted...  ...scalability, and performance of enterprise platforms.As a Principal Site Reliability Engineer (SRE), you will drive operational excellence across... 
    Principal
    Remote work
    Flexible hours

    DTCC- The Depository Trust & Clearing Corporation

    Jersey City, NJ
    1 day ago
  • $160k - $180k

     ...global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Principal Site Reliability Engineer to define the strategic vision and own the enterprise-wide reliability, scalability, and performance of our critical... 
    Principal
    Contract work
    Work from home
    Flexible hours

    Vertafore

    Denver, CO
    22 days ago
  • $129k - $231k

     ...where operational excellence, reliability, and quality are non-...  ...platforms. The Reliability Engineering team is the engineering-first...  ...Role Summary: As the Senior Principal SRE Engineering Lead, you are...  ...judgment of reliability at this site are yours. You are a senior... 
    Principal
    Contract work
    Flexible hours
    Shift work

    Eli Lilly and Company

    Indianapolis, IN
    1 day ago
  • $66k - $158.4k

     ...where operational excellence, reliability, and quality are non-...  ...platforms. The Reliability Engineering team is the engineering-first...  ...within the standards the Senior Principal SRE Engineer sets — SLOs, error...  ..., with meaningful time as a Site Reliability Engineer,... 
    Principal
    Full time
    H1b
    Visa sponsorship
    Work visa
    Flexible hours

    Eli Lilly and Company

    Indianapolis, IN
    1 day ago
  • $153k - $210k

     ...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating... 
    Full time

    Ridgeline

    New York, NY
    1 day ago
  •  ...to meet you. Our Enterprise Information Technology (EIT) organization is expanding, and we are seeking a Senior Site Reliability Engineer to help drive a major architectural modernization. In this role, you will move beyond traditional infrastructure maintenance... 
    Permanent employment
    Full time
    H1b
    Local area
    Remote work
    Shift work

    Jack Henry & Associates

    New York, NY
    1 day ago
  •  ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that...  ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to... 
    Full time

    Vanguard

    Wayne, PA
    2 days ago
  • About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure... 

    Alembic

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!