Principal Site Reliability Engineer
$167.3k - $242.6kExpel
Your passion for uptime was forged from experience in production and refined through incident response. You’re an Expel Principal Site Reliability Engineer - a protector, champion, and leader of Expel's reputation for service reliability.
Innovation comes naturally to you, but you're also eager to help others. You understand that operational reliability is a shared mission across all of engineering, and that your role is to make it as easy as possible for Expel to achieve that mission. You spend your mornings collaborating with architects and product stakeholders to outline the next quarter’s reliability initiatives, then the afternoon pair-programming with a junior SRE to mentor them in debugging a tricky Kubernetes deployment.
You apply your dedication to reliability and collaboration with the broader SRE community to ensure Expel maintains outstanding reliability standards within the cloud native ecosystem.
You take pride in all the nines of uptime you’ve achieved, but you know that an SRE’s job is never done!
What Expel can do for you
- Provide an opportunity to grow and maintain reliability-focused platform features within a cloud native engineering platform using modern infrastructure and tooling (Kubernetes/GKE/EKS runtime, Hashicorp toolset, etc)
- Provide you a mission you can get behind: stopping evil hackers so our customers can focus on their business
- Be included in a company focused on creating opportunities to do interesting work and creating space for employees to learn and grow
- An opportunity to contribute to a best-in-class product
- A leadership team that’s embraced modern Site Reliability principles, as outlined by Google and other industry leaders.
What you can do for Expel
- Lead project work to build and maintain platform features that cut across the Expel product’s reliability, networking, and cloud infrastructure.
- Contribute by pushing IaC commits daily, with occasional opportunities to write and test application code in Python, Golang, and Javascript
- Mentor and motivate service owners on how to use the platform in order to deploy, measure, monitor, and operate their own services at scale.
- Participate in a weekly support rotation that includes taking the on-call pager and providing nearly on-demand working-hours support to platform users.
- Lead incident response, triage, and root cause analysis support
- Poke fun at our leadership team in creative ways.
What you should bring with you
- A passion for learning and improving your work product
- Significant experience operating Kubernetes within highly distributed environments
- Experience running systems in GCP or AWS
- Exposure to monitoring and observability infrastructure and standard methodologies
- An understanding of infrastructure-as-code practices, tools, and patterns
- Some experience developing software in Linux environments, preferably with Python and/or Golang
- A customer-minded approach that enables the success of platform users as well as building trust across the organization.
- A collaborative disposition that allows you to work optimally on and across teams
- Six years of systems experience either in operations or development
- Missing some items on the list? That's ok! We still want to talk to you!
How our team works together
We build and run teams where everyone is pulling in the same direction and is learning from each other:
- We work out of a shared backlog
- We pair-program weekly, as it makes sense
- We peer-review everything
- We do weekly blame-free retros to reinforce what’s going well, so we do more of it, and surface what’s not going well, so we can do something about it. Same thing for projects and significant operational problems.
Our hiring process
We respect your time. You’ll hear from us by the end of the next business day after completing an interview.
We also have a goal that all Expletives have a great manager and have a voice in how their team is run and who runs it. It’s not the shortest process in the industry, but you’ll get to meet nearly everyone you’ll work with day-to-day and your Engineering leadership. New Expletives consistently say our interview process gave them an accurate picture of what it’s like to work here.
Here’s our 3-stage process for this position (5.5 hours total interviewing time):- Chat with a recruiter (30 min)
- Video interview with hiring manager (Engineering Manager) (60 minutes)
- Pair programming interview (with two engineers) (60 minutes)
- “Virtual onsite interview” (can be scheduled contiguous or broken up, 60 minutes each):
- Engineering leadership (Engineering Director and Manager of Delivery Experience)
- System design interview (with two engineers)
- Technology and skills interview (with two engineers)
Additional details
This role is remote within the United States.
The base salary range for this role is between $167,300 USD and $242,600 USD + bonus eligibility and equity. While the full salary band reflects our long-term compensation framework, we're primarily targeting candidates between $195,000 and $225,000 based on experience, skills, and market data.
We believe in paying transparently and equitably. Your salary will ultimately be based on factors such as your experience, skills, team equity, and market data. You’ll also be eligible for unlimited PTO (which we model and encourage), work location flexibility, up to 24 weeks of parental leave, and really excellent health benefits.
We’re only hiring those authorized to work in the United States. We do not currently sponsor immigration visas.
We're an Equal Opportunity Employer: You'll receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or on the basis of disability.
We’ll ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please let us know if you need accommodation of any kind.
#LI-Remote
Salary Range
$167,300—$242,600 USD
- ...future sponsorship.Maintain and enhance the reliability, availability, and performance of Navy... ...page of the Navy Federal Career Site.Protect Yourself from Job Scams: Navy Federal... ...Act.Master's degree in computer science, engineering, or the equivalent combination of education...PrincipalInternshipMonday to Friday
$142.8k - $274.8k
...yearEmployment type: Full-TimeWork site: 0 days / week in-office - remoteRole... ...Software EngineeringDiscipline: Site Reliability EngineeringCompany:... ...world’s most demanding workloads. As a Principal Site Reliability Engineer, you will set technical and operational...PrincipalOngoing contractWork at officeLocal area- ...lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together.We are seeking a Principal Site Reliability Engineer (SRE) to define and scale reliability practices across large-scale cloud platforms.This is a senior individual contributor...PrincipalMinimum wageFull timeWork experience placementWork at officeLocal areaRemote work
- ...Working remotely within the United States, the full-time Principal Site Reliability Engineer will lead project work to enhance platform reliability, mentor junior engineers, and engage in incident response while collaborating closely with product stakeholders and architects...PrincipalFull timeRemote work
$163.62k - $212.71k
...maintaining the tools, platforms, and processes that improve our engineering teams' productivity and streamline the software... ...Responsibilities:We are seeking a seasoned and strategic Lead/Principal Site Reliability Engineer to drive the reliability, scalability, and...PrincipalFull timePart timeWork experience placementWork at officeLocal areaImmediate startRemote workWork from homeFlexible hoursShift work3 days per week1 day per week- ...Principal Site Reliability Engineer location- Washington, DC -Onsite Remote- No 6+ Months Job Summary At Amtrak, we're seeking a seasoned Principal Site Reliability Engineer with a strong focus on pipelines as code, CI/CD, and IaC. You...PrincipalRemote work
- ...Senior Principal Site Reliability Engineer Hong Kong SAR About Us Established in 2018, Bybit is one of the world's leading cryptocurrency exchanges and digital financial platforms, serving over 80 million users across more than 200 countries and regions. Powered...PrincipalRemote work
$90k - $130k
...-on experience. We require 10+ years of experience in Site Reliability Engineering, Software Engineering, or Cloud Engineering. We need experience... ...complex health care challenges. This is a senior, remote Principal Site Reliability Engineer role with preference for...PrincipalFull timeRemote work$159k - $272k
...generosity. Join us for the opportunity to grow and make a difference in ways that matter to you. Role SummaryIn this role as Principal Site Reliability Engineer, Infrastructure Observability you will help formulate, develop, and implement a team of Site Reliability Engineers (...PrincipalFull timePrivate practiceLocal areaRemote workWork from home3 days per week$200k - $250k
...build the future together. The Crown Is Yours As a Principal Site Reliability Enginee r , you'll shape the long-term strategy for the... ...direction of our cloud and on-premise platforms, helping engineering teams build, deploy, and operate highly reliable systems...PrincipalFull timeImmediate startRemote work$7,000 per month
...Principal Site Reliability Engineer Latin America The salary range for this role is negotiable, the range being $7000 - $12000 per month (Gross in USD) About Sezzle: With a mission to financially empower the next generation, Sezzle is revolutionizing the shopping...PrincipalRemote workFlexible hours$160k - $180k
...global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Principal Site Reliability Engineer to define the strategic vision and own the enterprise-wide reliability, scalability, and performance of our critical...PrincipalContract workWork from homeFlexible hours- ...with software development teams to build reliable, scalable, secure, and cloud-native... ...influence scalable architecture patterns across engineering teams, helping ensure systems are... ...~8+ years of hands-on experience in Site Reliability Engineering, DevOps, cloud infrastructure...PrincipalRemote work
$142.8k - $274.8k
...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type:... ...Software EngineeringDiscipline: Site Reliability EngineeringCompany:... ...demanding workloads.We are seeking a Principal Site Reliability Engineering Manager to lead a team responsible...PrincipalOngoing contractTemporary workFixed term contractLocal areaImmediate start3 days per week- Role Description Symmetrio is recruiting a Principal Site Reliability Engineer (SRE) for our customer, a rapidly growing healthcare technology organization focused on advanced healthcare technology solutions. This individual will play a critical role in ensuring the reliability...PrincipalFull time
- ...serve.The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted... ...scalability, and performance of enterprise platforms.As a Principal Site Reliability Engineer (SRE), you will drive operational excellence across...PrincipalRemote workFlexible hours
- ...: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8 to... ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will have...Remote work
$65 - $75 per hour
DescriptionKforce has a client seeking a remote Senior Site Reliability Engineer to be a l be a leading member of the team working with a diverse range of technologies. You will enjoy working in a friendly environment and benefit from our investment in staff. The role also...Remote work- ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and... ...and networking teams to improve service reliability and deployment workflowsDeploy and... ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering...Work at officeLocal areaWork from homeFlexible hours
$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...Currently hiringRemote work- Recognized as the No. 1 site trusted by real estate professionals, Realtor.com has been at the forefront of online real estate... ...confidence through expert guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization,...Work at officeLocal area
$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...Full timeWork at officeLocal areaRemote workWork from home$90k - $180k
...nutritionals and branded generic medicines. Our 115,000 colleagues serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We are...Remote work$15k
...packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage engineering skills...Work at officeLocal areaRemote work$267k - $356k
...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-... ...workloads in the industry, which means reliability and performance aren't just goals—they're... ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc...Work experience placementWork at officeLocal areaWork from homeFlexible hours$158.5k - $172k
...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and... .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology...Full timeTemporary workWork at officeFlexible hours3 days per week$150k - $180k
...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business... ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or...Permanent employmentFull timeWork at officeLocal areaRemote workWorldwide$86.9k - $198k
Site Reliability Engineer, SeniorThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if...Full timeContract workPart timeWork at officeLocal areaRemote work$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....Work at officeLocal areaRemote workWorldwideFlexible hours$110k - $145k
...is global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is...Contract workWork at officeWork from homeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!
- principal network engineer Remote
- senior director engineering Remote
- principal developer Remote
- chief design engineer Remote
- principal engineer Remote
- director data engineering Remote
- principal infrastructure engineer Remote
- principal security engineer Remote
- senior civil engineer project manager Remote
- director software engineering Remote



