Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Site Reliability Engineer

$167.3k - $242.6k

Expel

Principal Site Reliability Engineer

Remote

Your passion for uptime was forged from experience in production and refined through incident response. You're an Expel Principal Site Reliability Engineer - a protector, champion, and leader of Expel's reputation for service reliability.

Innovation comes naturally to you, but you're also eager to help others. You understand that operational reliability is a shared mission across all of engineering, and that your role is to make it as easy as possible for Expel to achieve that mission. You spend your mornings collaborating with architects and product stakeholders to outline the next quarter's reliability initiatives, then the afternoon pair-programming with a junior SRE to mentor them in debugging a tricky Kubernetes deployment.

You apply your dedication to reliability and collaboration with the broader SRE community to ensure Expel maintains outstanding reliability standards within the cloud native ecosystem.

You take pride in all the nines of uptime you've achieved, but you know that an SRE's job is never done!

What Expel Can Do For You
  • Provide an opportunity to grow and maintain reliability-focused platform features within a cloud native engineering platform using modern infrastructure and tooling (Kubernetes/GKE/EKS runtime, Hashicorp toolset, etc)
  • Provide you a mission you can get behind: stopping evil hackers so our customers can focus on their business
  • Be included in a company focused on creating opportunities to do interesting work and creating space for employees to learn and grow
  • An opportunity to contribute to a best-in-class product
  • A leadership team that's embraced modern Site Reliability principles, as outlined by Google and other industry leaders.
What You Can Do For Expel
  • Lead project work to build and maintain platform features that cut across the Expel product's reliability, networking, and cloud infrastructure.
  • Contribute by pushing IaC commits daily, with occasional opportunities to write and test application code in Python, Golang, and Javascript
  • Mentor and motivate service owners on how to use the platform in order to deploy, measure, monitor, and operate their own services at scale.
  • Participate in a weekly support rotation that includes taking the on-call pager and providing nearly on-demand working-hours support to platform users.
  • Lead incident response, triage, and root cause analysis support
  • Poke fun at our leadership team in creative ways.
What You Should Bring With You
  • A passion for learning and improving your work product
  • Significant experience operating Kubernetes within highly distributed environments
  • Experience running systems in GCP or AWS
  • Exposure to monitoring and observability infrastructure and standard methodologies
  • An understanding of infrastructure-as-code practices, tools, and patterns
  • Some experience developing software in Linux environments, preferably with Python and/or Golang
  • A customer-minded approach that enables the success of platform users as well as building trust across the organization.
  • A collaborative disposition that allows you to work optimally on and across teams
  • Six years of systems experience either in operations or development
  • Missing some items on the list? That's ok! We still want to talk to you!
How Our Team Works Together

We build and run teams where everyone is pulling in the same direction and is learning from each other:

  • We work out of a shared backlog
  • We pair-program weekly, as it makes sense
  • We peer-review everything
  • We do weekly blame-free retros to reinforce what's going well, so we do more of it, and surface what's not going well, so we can do something about it. Same thing for projects and significant operational problems.
Our Hiring Process

We respect your time. You'll hear from us by the end of the next business day after completing an interview.

We also have a goal that all Expletives have a great manager and have a voice in how their team is run and who runs it. It's not the shortest process in the industry, but you'll get to meet nearly everyone you'll work with day-to-day and your Engineering leadership. New Expletives consistently say our interview process gave them an accurate picture of what it's like to work here. Here's our 3-stage process for this position (5.5 hours total interviewing time):

  • Chat with a recruiter (30 min)
  • Video interview with hiring manager (Engineering Manager) (60 minutes)
  • Pair programming interview (with two engineers) (60 minutes)
  • "Virtual onsite interview" (can be scheduled contiguous or broken up, 60 minutes each):
    • Engineering leadership (Engineering Director and Manager of Delivery Experience)
    • System design interview (with two engineers)
    • Technology and skills interview (with two engineers)
Additional Details

This role is remote within the United States.

The base salary range for this role is between $167,300 USD and $242,600 USD + bonus eligibility and equity. While the full salary band reflects our long-term compensation framework, we're primarily targeting candidates between $195,000 and $225,000 based on experience, skills, and market data.

We believe in paying transparently and equitably. Your salary will ultimately be based on factors such as your experience, skills, team equity, and market data. You'll also be eligible for unlimited PTO (which we model and encourage), work location flexibility, up to 24 weeks of parental leave, and really excellent health benefits.

We're only hiring those authorized to work in the United States. We do not currently sponsor immigration visas.

We're an Equal Opportunity Employer: You'll receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or on the basis of disability.

We'll ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please let us know if you need accommodation of any kind.

#LI-Remote

Salary Range

$167,300 - $242,600 USD

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Principal Site Reliability Engineer in United States vacancy
  • $139.7k - $232.9k

     ...designing, implementing, and continuously improving highly reliable, scalable, and resilient platform solutions across the enterprise. Operates as a subject matter expert (SME) in Site Reliability Engineering, driving reliability engineering practices, operational excellence... 
    Principal
    Full time
    Work experience placement

    M&T Bank

    Buffalo, NY
    21 hours ago
  •  ...future sponsorship.Maintain and enhance the reliability, availability, and performance of Navy...  ...page of the Navy Federal Career Site.Protect Yourself from Job Scams: Navy Federal...  ...Act.Master's degree in computer science, engineering, or the equivalent combination of education... 
    Principal
    Internship
    Monday to Friday

    Navy Federal Credit Union

    Vienna, VA
    21 hours ago
  • $142.8k - $274.8k

     ...yearEmployment type: Full-TimeWork site: 0 days / week in-office - remoteRole...  ...Software EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...world’s most demanding workloads. As a Principal Site Reliability Engineer, you will set technical and operational... 
    Principal
    Ongoing contract
    Work at office
    Local area

    Microsoft

    Redmond, WA
    2 days ago
  •  ...lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together.We are seeking a Principal Site Reliability Engineer (SRE) to define and scale reliability practices across large-scale cloud platforms.This is a senior individual contributor... 
    Principal
    Minimum wage
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work

    UnitedHealth Group

    Minnetonka, MN
    21 hours ago
  •  ...Working remotely within the United States, the full-time Principal Site Reliability Engineer will lead project work to enhance platform reliability, mentor junior engineers, and engage in incident response while collaborating closely with product stakeholders and architects... 
    Principal
    Full time
    Remote work

    Virtual Vocations Inc

    United States
    2 days ago
  •  ...Infrastructure Code. Builds reliability into the ecosystem by applying...  ...practices in resiliency engineering and observability by developing...  ...engineering techniques with site reliability engineering...  ...5) years of experience as a Principal Site Reliability Engineer (or... 
    Principal
    Full time

    Fidelity Investments

    Westlake, OH
    4 days ago
  • $175.5k - $235.4k

     ...MyDisneyExperience and Hey, Disney!This role sits in the Commerce Site Reliability Engineering (SRE) specifically supporting Ecommerce , Consumer...  ...Products Technology teams from across the company. The Principal of DXT SRE will report to the Director of DXT Commerce SREAbout... 
    Principal
    Worldwide

    Disney Interactive

    Orlando, FL
    1 day ago
  • Site Reliability Engineer - Equity Trading PlatformLocation: New York | Practice Area: Capital Markets - Technology & Engineering | Type: PermanentKeep critical equity trading platforms resilient, reliable, and ready for the markets.The RoleWe are seeking a highly motivated... 
    Principal
    Permanent employment
    Work at office
    Weekend work
    Afternoon shift

    Capco

    New York, NY
    2 days ago
  • $84.9k - $209.5k

    This role combines strategic architecture with practical systems engineering, deployment, automation, patching, troubleshooting, incident response, and compliance support. The Principal Site Reliability Engineer will work across Windows, Linux, Oracle Cloud Infrastructure... 
    Principal
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    21 hours ago
  • $163.62k - $212.71k

     ...maintaining the tools, platforms, and processes that improve our engineering teams' productivity and streamline the software...  ...Responsibilities:We are seeking a seasoned and strategic Lead/Principal Site Reliability Engineer to drive the reliability, scalability, and... 
    Principal
    Full time
    Part time
    Work experience placement
    Work at office
    Local area
    Immediate start
    Remote work
    Work from home
    Flexible hours
    Shift work
    3 days per week
    1 day per week

    iSpot.tv

    Bellevue, WA
    2 days ago
  •  ...Principal Site Reliability Engineer location- Washington, DC -Onsite Remote- No 6+ Months Job Summary At Amtrak, we're seeking a seasoned Principal Site Reliability Engineer with a strong focus on pipelines as code, CI/CD, and IaC. You... 
    Principal
    Remote work

    American IT Systems

    Washington DC
    1 day ago
  • $84.9k - $209.5k

     .... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers...  ...posted.Career Level - IC4Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and... 
    Principal
    Temporary work
    Monday to Friday
    Flexible hours
    Shift work
    Night shift

    Oracle Corporation

    Reston, VA
    4 days ago
  •  ...Senior Principal Site Reliability Engineer Hong Kong SAR About Us Established in 2018, Bybit is one of the world's leading cryptocurrency exchanges and digital financial platforms, serving over 80 million users across more than 200 countries and regions. Powered... 
    Principal
    Remote work

    Bybit

    United States
    4 days ago
  • $90k - $130k

     ...-on experience. We require 10+ years of experience in Site Reliability Engineering, Software Engineering, or Cloud Engineering. We need experience...  ...complex health care challenges. This is a senior, remote Principal Site Reliability Engineer role with preference for... 
    Principal
    Full time
    Remote work

    UnitedHealth Group

    Minnetonka, MN
    3 days ago
  •  ...Veracode is seeking an enthusiastic, motivated engineer with deep AWS knowledge and the ability to keep up with a high-performing team. This is a chance to be on the leading edge of our evolution to the cloud in a fast-paced environment. As a member of our SRE team... 
    Principal
    Work experience placement

    Veracode

    Burlington, MA
    5 days ago
  • $159k - $272k

     ...generosity. Join us for the opportunity to grow and make a difference in ways that matter to you. Role SummaryIn this role as Principal Site Reliability Engineer, Infrastructure Observability you will help formulate, develop, and implement a team of Site Reliability Engineers (... 
    Principal
    Full time
    Private practice
    Local area
    Remote work
    Work from home
    3 days per week

    T. Rowe Price

    Owings Mills, MD
    21 hours ago
  •  ...Site Reliability Engineering (SRE) Team Lead The Site Reliability Engineering (SRE) team is foundational to the growth and scale of our platform. You and your team will help advance several initiatives tied to automation, SRE culture, and cloud architecture. You will... 
    Principal
    Shift work

    Roberts Recruiting

    Boston, MA
    2 days ago
  • $96.3k - $264.1k

     ...infrastructure and service, ensuring alignment with reliability and functionality standards. Takes full...  ...tools and provides expertise in site reliability trends.Only Oracle brings...  ...LeadershipDefine and drive the site reliability engineering strategy for large-scale, distributed,... 
    Principal
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    21 hours ago
  •  ...disrupt, and thrive! KēSTA I.T. is actively seeking a Principal Engineer for an immediate full-time opportunity with our industry creating...  ...An innovative technology company is seeking experienced Site Reliability Engineers to take ownership of building reliable, scalable... 
    Principal
    Permanent employment
    Full time
    Temporary work
    Immediate start

    KēSTA I.T.

    Beverly Hills, CA
    18 days ago
  • $272k - $431.25k

    NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's (Infrastructure, Planning and Process) Cloud Infrastructure Team. IPP is a global organization within NVIDIA. This group works with various other groups within NVIDIA such as Graphics... 
    Principal
    Full time
    Work experience placement
    Worldwide

    Nvidia

    Santa Clara, CA
    21 hours ago
  •  ...Principal Site Reliability Engineer Deimos is a cloud-native developer and security operations technology services company. We help companies of all sizes adopt the cloud for improved service delivery to their clients. We're a fully remote African-based team of engineers... 
    Principal
    Currently hiring
    Remote work
    Work from home

    Deimos

    United States
    2 days ago
  • $200k - $250k

     ...shaping it, one bold step at a time. To those who see AI as a driver of progress, come build the future together. As a Principal Site Reliability Engineer, you'll shape the long-term strategy for the infrastructure behind one of the most demanding platforms in sports... 
    Principal
    Full time
    Immediate start
    Remote work

    DraftKings

    United States
    3 days ago
  •  ...Principal Site Reliability Engineer About ShipperHQ: ShipperHQ is a trusted leader in the e-commerce shipping space, with over 15 years of experience helping merchants deliver better checkout experiences. Founded in 2009, we power shipping logic and checkout optimization... 
    Principal
    Full time
    Work at office

    ShipperHQ

    Austin, TX
    a month ago
  • $240k - $250k

     ...MattersSaviynt’s platform is mission-critical for our customers. As we scale globally, reliability, availability, and performance are not optional—they are core product features.As a Principal Engineer, you will define and drive the reliability strategy for our SaaS platform.... 
    Principal
    Full time

    Saviynt

    Atlanta, TX
    1 day ago
  • $7,000 per month

     ...Principal Site Reliability Engineer Latin America The salary range for this role is negotiable, the range being $7000 - $12000 per month (Gross in USD) About Sezzle: With a mission to financially empower the next generation, Sezzle is revolutionizing the shopping... 
    Principal
    Remote work
    Flexible hours

    Sezzle

    United States
    4 days ago
  • $160k - $180k

     ...global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Principal Site Reliability Engineer to define the strategic vision and own the enterprise-wide reliability, scalability, and performance of our critical... 
    Principal
    Contract work
    Work from home
    Flexible hours

    Vertafore

    Denver, CO
    more than 2 months ago
  • $84.9k - $209.5k

     ...spirit that promotes an upbeat and creative environment. We are unencumbered and will need your contribution to make it a special engineering center with the focus on excellence. Health Data Intelligence Platform has a rare opportunity to play a critical role in how... 
    Principal
    Temporary work
    Immediate start
    Remote work
    Flexible hours

    Hackajob

    United States
    4 days ago
  •  ...with software development teams to build reliable, scalable, secure, and cloud-native...  ...influence scalable architecture patterns across engineering teams, helping ensure systems are...  ...~8+ years of hands-on experience in Site Reliability Engineering, DevOps, cloud infrastructure... 
    Principal
    Remote work

    ABC Fitness Solutions, LLC

    United States
    4 days ago
  • $151.6k - $245.3k

     ...Summary Your Career Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture... 
    Principal
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    1 day ago
  • Role Description Symmetrio is recruiting a Principal Site Reliability Engineer (SRE) for our customer, a rapidly growing healthcare technology organization focused on advanced healthcare technology solutions. This individual will play a critical role in ensuring the reliability... 
    Principal
    Full time

    Symmetrio

    Remote
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!