Staff Site Reliability Engineer
United IT Solutions
Job: Staff Site Reliability Engineer (SRE) Location: San Francisco, CA Job Responsibilities
As our Staff SRE, you'll be the primary expert responsible for our entire compute ecosystem. Your key responsibilities will include:
As a Staff SRE, you'll operate at the highest level of technical expertise and influence. You won't just solve problems; you'll prevent them at a fundamental level across organizational boundaries.
Required Qualifications
As our Staff SRE, you'll be the primary expert responsible for our entire compute ecosystem. Your key responsibilities will include:
As a Staff SRE, you'll operate at the highest level of technical expertise and influence. You won't just solve problems; you'll prevent them at a fundamental level across organizational boundaries.
- Design, implement, and lead large-scale, cross-functional projects to improve the reliability, performance, and efficiency of our core services and infrastructure (10× impact).
- Drive the reduction of toil by developing and deploying sophisticated automation tools and frameworks, championing the "everything as code" philosophy.
- Serve as a technical escalation point for critical incidents, perform deep-dive root cause analyses (RCAs), and implement robust corrective measures to prevent recurrence.
- Define and implement SLOs, SLIs, and Error Budgets for critical services. Enhance our monitoring, logging, and tracing systems to provide comprehensive visibility into system health.
- Set the technical direction and best practices for the entire SRE and engineering organization. Mentor mid-level and senior engineers on design patterns, operational rigor, and reliability principles.
Required Qualifications
- 8+ years of progressive experience in Site Reliability Engineering, Production Engineering, or a closely related role.
- Expert-level proficiency with AWS, including networking, compute, and storage.
- Deep expertise in Kubernetes and the cloud-native ecosystem.
- Fluency in at least one major scripting/programming language for automation and tooling (e.g., Python, Go, or Java).
- Solid experience with monitoring and logging solutions (Datadog)
- Proven ability to design and implement robust, highly available distributed systems.
- Demonstrated experience with Infrastructure as Code tools like Terraform.
- Exceptional communication skills, capable of explaining complex technical issues to both technical and non-technical audiences.
- Experience implementing Service Mesh technologies (e.g., Istio, Linkerd).
- A strong understanding of security principles and practices in a cloud environment.
- Certifications such as CKA (Certified Kubernetes Administrator) or CKAD (Certified Kubernetes Application Developer).
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Site Reliability Engineer in San Francisco, CA vacancy
$260k - $300k
...software agents. We're the makers of Devin, the first AI software engineer. Our team is extremely talent-dense. Among our founding... ...faster than anyone expects. You will own both the production reliability of our user-facing products and the platform engineering that...Suggested- ...The role We're looking for a world-class Site Reliability Engineer to ensure the reliability, performance, and scalability of our AI infrastructure platform. You'll be building and operating the core systems that power agentic AI at scale. Your mission: keep...Suggested
$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. The Senior Site Reliability Engineer Opportunity Reporting to the Manager, Site Reliability Engineering, this role will help...SuggestedPermanent employmentLocal areaWorldwideFlexible hours- ...Site Reliability Engineer Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling...SuggestedFlexible hours
- ...Senior Engineering Role at Salesforce Salesforce is the #1 AI CRM, where humans with agents drive customer success together. Here... ...Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with...SuggestedWorldwideWeekend work
$81.1k - $187k
...Site Reliability Engineer 3 We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving...Temporary workImmediate startFlexible hoursShift work- ...the globe. Join us on this journey to redefine resource management-and change lives along the way. The Role As a Site Reliability Engineer (SRE) at Air Apps, you will be responsible for ensuring the reliability, availability, and scalability of our systems. You...Temporary workWorldwide
- ...About the Role We're looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You'll partner with engineers and data scientists to build, automate, and maintain...
- ...an SRE to join our infrastructure team. This role will be responsible for building software to ensure the reliability of our back-end systems, working with engineers who develop them, and planning for our future growth. You will work with our existing production...WorldwideHome officeFlexible hours
- ...JOB DESCRIPTION Project Outline: We are looking for a Site Reliability Engineer with experience in incident response. In this role, you will help Shipt understand where we can improve stability and reliability. There will be a focus on the intersection of systems...
- ...Arena Intelligence Engineer Arena Intelligence is looking for an engineer to build the core infrastructure that sits beneath our online... ...foundational infrastructure for our users that scales, is reliable, and makes the complexities of operating this infrastructure at...Permanent employmentShift work
- ...Site Reliability Engineer Specter's mission is to help automate the physical world. Today, we build video sensors with state-of-the-art AI agents that answer any question, anywhere in their environments. Our systems can automatically detect and reason about any physical...Remote work
- ...come shape the future and be part of a truly unique global culture at OutSystems! Hybrid Onsite in Menlo Park, CA Site Reliability Engineering (SRE) is a discipline that incorporates aspects of software engineering and applies them to infrastructure and...Immediate startRemote workWorldwide
$170k - $220k
...Senior Site Reliability Engineer Supio is a trusted AI platform purpose-built for law firms, reshaping how data drives impactful outcomes. Our innovative approach blends technology with deep legal expertise, making us a leader in our field. We go beyond surface-level...Work at officeRemote workFlexible hours- ...enterprise that runs the real economy. Learn more about our vision in our manifesto. About the Role We're looking for a Site Reliability Engineer to take the lead on scaling our operational resilience as we grow. You'll own the stability, observability, and debugging...WorldwideShift work
$117k - $209.33k
...Full time 26WD99273 Job Requisition ID # 26WD99273 Position Overview Want to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud...Full timeFor contractors- ...Site Reliability Engineer Job Location: San Francisco, CA or Charlotte, NC. Job Type: Contract Work with local API development squads, platform teams, product owners, scrum masters, and architects. The SRE ensures that both our internally critical and our externally...Contract workLocal area
$230k - $310k
...millions of daily users while enabling our engineering teams to ship fast. You'll own the... ...building automation and tooling that improves reliability and partnering with engineering to... ...What you'll bring ~5+ years in site reliability engineering, DevOps, or systems...Full timeWork at officeWork from home$195k - $257.5k
...Staff Site Reliability Engineer Circle (NYSE: CRCL) is one of the world's leading internet financial platform companies, building the foundation of a more open, global economy through digital assets, payment applications, and programmable blockchain infrastructure....Flexible hours$194k - $267k
...something more than once, automate it" and who can rapidly self-educate on new concepts and tools. Position Overview: The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$120.6k - $150.9k
...Staff Site Reliability Engineer (SRE) We are looking for a highly motivated, high-potential Staff Site Reliability Engineer (SRE) to join our team as a technical leader and drive transformative impact across WEX's platform reliability and operational excellence. This...Flexible hours$150k - $220k
...teams, and innovators in this way. The Role: As an engineering organization, we pride ourselves on engineering as a creative... ...can achieve autonomy, mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge’s SRE team responsible for keeping...Local area- ...and our economy, and we're building a dynamic and diverse team for our future. We are seeking an experienced Lead Site Reliability Engineer to join our engineering team and drive the reliability, scalability, and performance of our critical systems. This role...Full timePart time
$181k - $263k
...and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Staff Site Reliability Engineer who will set the technical direction for reliability engineering across LiveRamp's global infrastructure. This is a...Full timeWork from homeWorldwideFlexible hoursNight shift$200k - $260k
...Join the Cloud Infrastructure Team as a technical leader driving reliability, automation, and scalability across the systems running Sight... ...and reliability practices across teams, mentor senior engineers, and be a primary escalation point for the org's hardest systems...Casual workWork at officeRemote workFlexible hours$221.2k - $300k
...Manager, Software Engineer, Site Reliability Engineering Share Manager, Software Engineer, Site Reliability Engineering ~ link Copy link corporate_fare Google place San Francisco, CA, USA Advanced Experience owning outcomes and decision making, solving ambiguous...Full timeWork at office$61k - $101k
...Salary: $61,000 - 101,000 per year Requirements: We expect formal training or certification in site reliability engineering, plus 3+ years of hands-on experience. We want strong familiarity with SRE culture and the practical application of reliability principles...Full time$250k
...Europe, while now significantly expanding its footprint in the United States. The company is looking for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and cloud environments powering GPU-intensive workloads. The role involves working...Full timeRemote work- ...design of information and operational support systems. Required Skills/Qualifications: BS/MS degree in Computer Science, Engineering, or a related subject. Equivalent experience accepted. Proven working experience in installing, configuring, and troubleshooting...Full timeWork experience placementRemote workFlexible hours
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Enterprise Technology, Infrastructure Platforms team, you will solve complex and broad...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!
Related searches
- engineering aide San Francisco, CA
- staff design engineer San Francisco, CA
- senior staff engineer San Francisco, CA
- assistant engineer San Francisco, CA
- software engineer staff San Francisco, CA
- staff engineer San Francisco, CA
- staff security engineer San Francisco, CA
- senior staff systems engineer San Francisco, CA
- staff data engineer San Francisco, CA
- technology administrator San Francisco, CA



