Principal Site Reliability Engineer
$173k - $230kEarly Warning Services, LLC
At Early Warning, we've powered and protected the U.S. financial system for over thirty years with cutting-edge solutions like Zelle , Paze , and so much more. As a trusted name in payments, we partner with thousands of institutions to increase access to financial services and protect transactions for hundreds of millions of consumers and small businesses. Positions located in Scottsdale, San Francisco, Chicago, or New York follow a hybrid work model to allow for a more collaborative working environment. Candidates responding to this posting must independently possess the eligibility to work in the United States, for any employer, at the date of hire. This position is ineligible for employment Visa sponsorship. Role Summary The Principal Site Reliability Engineer applies software engineering and systems engineering practices to improve the reliability, resilience, scalability, and operational health of production services. The role partners with Software Engineering and other technology teams to ensure reliability, observability, recoverability, performance, and operational readiness are engineered into systems throughout their lifecycle. The role operates at enterprise scope, establishing technical direction and applying evidence-driven engineering, technical rigor, sound judgment, automation, and broad systems expertise across organizational boundaries. Core Responsibilities
Phoenix, AZ in USD per year is: $173,000 - $230,000.
San Francisco, CA in USD per year is: $207,000 - $276,000.
Additionally, candidates are eligible for a discretionary incentive plan and benefits. This pay scale is subject to change and is not necessarily reflective of actual compensation that may be earned, nor a promise of any specific pay for any specific candidate, which is always dependent on legitimate factors considered at the time of job offer. Early Warning Services takes into consideration a variety of factors when determining a competitive salary offer, including, but not limited to, the job scope, market rates and geographic location of a position, candidate's education, experience, training, and specialized skills or certification(s) in relation to the job requirements and compared with internal equity (peers). The business actively supports and reviews wage equity to ensure that pay decisions are not based on gender, race, national origin, or any other protected classes. Physical Requirements Early Warning works together in a highly collaborative office environment.Working conditions consist of a normal office environment. Work is primarily sedentary and requires extensive use of a computer and involves sitting for periods of approximately four hours. Work may require occasional standing, walking, kneeling, and reaching. Must be able to lift 10 pounds occasionally and/or negligible amount of force frequently. Requires visual acuity and dexterity to view, prepare, and manipulate documents and office equipment including personal computers. Requires the ability to communicate with internal and/or external customers. Employee must be able to perform essential functions and physical requirements of position with or without reasonable accommodation. Candidates responding to this posting must independently possess the eligibility to work in the United States at the date of hire. Some of the Ways We Prioritize Your Health and Happiness
- Use software engineering, automation, and DevOps principles and practices to continually improve how services are built, tested, deployed, observed, operated, and recovered.
- Use data, evidence, experimentation, and rigorous engineering analysis appropriate to the level to identify reliability risks, test assumptions, and guide technical decisions.
- Define, implement, or improve SLIs, SLOs, error budgets, and other service-health measures appropriate to the scope of responsibility.
- Improve observability through metrics, logging, tracing, monitoring, alerting, dashboards, and service-health instrumentation.
- Drive continuous improvement across CI/CD, observability, deployment practices, Infrastructure as Code, automation, testing, incident response, capacity management, resilience, and operational readiness.
- Identify recurring or systemic production issues and translate operational experience into improvements in code, architecture, automation, tooling, and engineering practices.
- Partner with Software Engineering teams to incorporate reliability, resiliency, scalability, performance, observability, recoverability, and operational readiness throughout the development lifecycle.
- Participate in or lead incident response, troubleshooting, service restoration, and blameless post-incident learning appropriate to the level.
- Provides enterprise-level technical leadership for critical production incidents and establishes or influences engineering practices that improve incident response, escalation, service restoration and sustainable on-call operations across the organization.
- Reduce operational toil and unnecessary manual intervention through software, automation, reusable patterns, and better engineering practices.
- Acts as an enterprise force multiplier, raising the effectiveness and technical capability of engineers and teams across the organization while building sustainable organizational capability rather than individual dependency.
- Demonstrates software engineering, systems thinking, troubleshooting, and production reliability capabilities appropriate to the level.
- Applies evidence-driven reasoning and technical rigor to distinguish observed facts from assumptions and make defensible engineering recommendations.
- Shares knowledge and contributes to sustainable engineering capability rather than creating dependency on individual expertise.
- Operates with significant autonomy across the organization's most consequential reliability challenges.
- Establishes enterprise technical direction, develops senior technical leaders, and demonstrates impact well beyond systems personally touched.
- Typically 15+ years of relevant professional experience in Software Engineering, Site Reliability Engineering, Systems Engineering, Cloud/Platform Engineering, DevOps, Infrastructure Engineering, Architecture where applicable, or a comparable technical discipline.
- Experience with software development or scripting using one or more modern programming languages.
- Experience with software engineering principles, distributed systems, production troubleshooting, automation, and observability appropriate to the level.
- Experience with public cloud technologies and architectures, preferably AWS, along with infrastructure, networking, Linux/Unix, and modern application architectures appropriate to the level.
- Demonstrated analytical, problem-solving, communication, and collaboration skills appropriate to the scope of the role.
- Hands-on experience with AWS is preferred, or comparable experience with another major cloud platform such as Microsoft Azure, Google Cloud Platform (Google Cloud Platform), or Oracle Cloud Infrastructure (OCI).
- Experience developing, deploying, operating, or improving highly available production software or distributed systems.
- Experience with CI/CD, Infrastructure as Code, containers or orchestration, observability, monitoring, alerting, and software-delivery automation.
- Experience with SLIs, SLOs, error budgets, incident management, performance analysis, capacity management, resilience testing, disaster recovery, or operational readiness appropriate to the level.
- Experience creating reusable automation, tooling, platforms, patterns, or practices that improve engineering effectiveness.
- Bachelor's degree in Computer Science, Software Engineering, Computer Engineering, Information Systems, or a related technical field, or equivalent practical experience.
Phoenix, AZ in USD per year is: $173,000 - $230,000.
San Francisco, CA in USD per year is: $207,000 - $276,000.
Additionally, candidates are eligible for a discretionary incentive plan and benefits. This pay scale is subject to change and is not necessarily reflective of actual compensation that may be earned, nor a promise of any specific pay for any specific candidate, which is always dependent on legitimate factors considered at the time of job offer. Early Warning Services takes into consideration a variety of factors when determining a competitive salary offer, including, but not limited to, the job scope, market rates and geographic location of a position, candidate's education, experience, training, and specialized skills or certification(s) in relation to the job requirements and compared with internal equity (peers). The business actively supports and reviews wage equity to ensure that pay decisions are not based on gender, race, national origin, or any other protected classes. Physical Requirements Early Warning works together in a highly collaborative office environment.Working conditions consist of a normal office environment. Work is primarily sedentary and requires extensive use of a computer and involves sitting for periods of approximately four hours. Work may require occasional standing, walking, kneeling, and reaching. Must be able to lift 10 pounds occasionally and/or negligible amount of force frequently. Requires visual acuity and dexterity to view, prepare, and manipulate documents and office equipment including personal computers. Requires the ability to communicate with internal and/or external customers. Employee must be able to perform essential functions and physical requirements of position with or without reasonable accommodation. Candidates responding to this posting must independently possess the eligibility to work in the United States at the date of hire. Some of the Ways We Prioritize Your Health and Happiness
- Healthcare Coverage -Competitive medical (PPO/HDHP), dental, and vision plans as well as company contributions to your Health Savings Account (HSA) or pre-tax savings through flexible spending accounts (FSA) for commuting, health & dependent care expenses.
- 401(k) Retirement Plan -Featuring a 100% Company Safe Harbor Match on your first 6% deferral immediately upon eligibility.
- Paid Time Off - Flexible Time Off for Exempt (salaried) employees, as well as generous PTO for Non-Exempt (hourly) employees, plus 11 paid company holidays and a paid volunteer day.
- 12 weeks of Paid Parental Leave
- Maven Family Planning - provides support through your Parenting journey including egg freezing, fertility, adoption, surrogacy, pregnancy, postpartum, early pediatrics, and returning to work.
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Principal Site Reliability Engineer in San Francisco, CA vacancy
- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure...Suggested
$190.8k - $267.1k
...while helping Reddit grow its business. The reliability of our Ads systems directly impacts... ...Reliability team partners closely with Ads Engineering to improve reliability, scalability,... ...advertiser trust. We’re looking for a Senior Site Reliability Engineer to build, operate,...SuggestedFor contractorsWork experience placement$127k - $249k
The TeamPlatform Engineering sits within SRE and builds the core infrastructure powering MongoDB... ...plays a pivotal role in engineering the reliable, globally connected, multi-cloud network... ...are seeking a talented Senior Site Reliability Engineer (SRE) with a strong...SuggestedLocal areaRemote workWorldwideFlexible hours$152.5k - $205k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common libraries, and infrastructure that help Circle teams ship secure...SuggestedFlexible hours$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SuggestedWork at officeLocal areaRemote workWorldwideFlexible hours$117k - $209.33k
Job Requisition ID #26WD99273Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products.As part of a new SRE team supporting...Full timeFor contractors$139.76k - $287.75k
...their business.We are seeking a Senior Site ReliabilityEngineer to help operate, scale... ...will be instrumental in advancing the reliability, scalability, automation, observability,... ...The ideal candidate is a highly hands-on engineer with strong production experience and a...Work at officeLocal areaRelocationRelocation package$148.5k - $223.9k
...right place! Agentforce is the future of AI, and you are the future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations,...Full timeWorldwideWeekend work$165k - $227k
...opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.The Engineering OpportunityWe are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable...Local areaWorldwideFlexible hours$167.7k - $245.2k
...requiring approximately 2 days per week on-site at Cisco offices in either San Francisco... ...AI agents behave as intended, improving reliability and reducing risks. This unified... ...and control.As a Senior Site Reliability Engineer (SRE), you will build, operate, and continuously...Full timeTemporary workLocal areaFlexible hours2 days per week- ...let’s build what’s next.About the teamThe Engineering team at Airwallex is a diverse group of... ..., working together to build scalable, reliable, and secure products that empower businesses... ...services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work closely...Temporary workLocal areaWorldwide
- ...About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the...
$175k - $250k
...00/yr Job Title: Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On-Site only. Must live within commuting distance of... ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design, build...Full timeRemote workRelocationRelocation package- ...Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems safely... ...safe shipping the default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale...Permanent employmentWork experience placementWork at officeLocal area
$55k - $151.47k
...ApplicableSpecialismIFS - Internal Firm Services - OtherManagement LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our...Full timeH1b$113.4k - $162k
...break down barriers to communication and free the flow of conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd, reliability and everything in between!This role is about impact at...Temporary work- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Enterprise Technology, Infrastructure Platforms team, you will solve complex and broad business...
$153k - $191.3k
...hardware design, manufacturing, data processing, and software engineering, our office is a truly inspiring mix of experts from a... ...deployments across operating environments, to guarantee the reliability, scalability, and availability of our services. To do this, you...Full timeTemporary workFor contractorsWork at officeLocal areaRemote workHome office3 days per week- ...getting here.)About the RoleWe're building infrastructure that has to perform under real-world scale, reliability, and security demands — and we're looking for an engineer who wants to own the foundation it runs on. This isn't a traditional "keep the lights on" role.You'...
$167.7k - $245.2k
...within Cisco’s Networking, Security, Collaboration, and Observability portfolios.Your ImpactWe are seeking a skilled Senior Site Reliability Engineer (SRE) in Production Engineering with a strong background in SaaS and operations. You will design and manage large-scale,...Full timeTemporary workWork at officeLocal areaFlexible hours1 day per week- ...SingleStore engineers build the real-time data platform powering some of the world’s most... ...Position Summary We are seeking a Senior/Principal Software Engineer to join the... ...Demonstrated ability to design and build highly reliable, high-performance system software. ~ Experience...PrincipalFull time
$260k - $340k
...and be part of a high-performing team that believes in each other, come build with us at Crusoe.About This Role:As the Principal Systems Software Engineer, you will serve as the visionary lead for Crusoe’s next-generation AI infrastructure. This is a role for an industry-...PrincipalFull timeTemporary work$194k - $267k
...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$220k - $235k
...SRE to define the future of our cloud platform and champion engineering excellence across Ironclad. In this role, you will pair deep... ...Provide technical leadership and strategic direction for the Site Reliability Engineering team and our broader Cloud PlatformDefine and...Full timeContract workWork at office$217k - $303.9k
...information, visit .As Reddit continues to scale globally, reliability and performance are more critical than ever. The Site Experience SRE team sits at the intersection of infrastructure, product engineering, and user experience - ensuring that every interaction across...For contractorsWork experience placement$150k - $220k
...teams, and innovators in this way. The Role: As an engineering organization, we pride ourselves on engineering as a creative... ...can achieve autonomy, mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge’s SRE team responsible for keeping...Local area$194k - $267k
...do something more than once, automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and...Permanent employmentWork at officeLocal areaWorldwideFlexible hours- ...long-term maintainability. # Own core runtime foundations: distributed control, state management, fault handling, and reliability. # Drive engineering rigor: testability, code quality, review standards, performance regression prevention, and release processes. #...PrincipalFull timeRemote work
$204k - $306k
...all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity, from... ...in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions millions...Permanent employmentWork at officeLocal areaWorldwideFlexible hours2 days per week$195k - $257.5k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Staff Site Reliability Engineer on Circle’s Platform team, you’ll design, build, and operate the infrastructure that powers our blockchain platform at...Flexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!
Related searches
- senior chief engineer San Francisco, CA
- general engineer San Francisco, CA
- chief design engineer San Francisco, CA
- principal infrastructure engineer San Francisco, CA
- principal cloud engineer San Francisco, CA
- chief engineer San Francisco, CA
- principal developer San Francisco, CA
- senior principal engineer San Francisco, CA
- engineering director San Francisco, CA
- technical director engineering San Francisco, CA


