Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Sr. Staff Site Reliability Engineer

$232k - $263k

Obsidian Security

Job Description

Job Description

Obsidian Security is the leading SaaS security platform, trusted by global enterprises like Snowflake, T-Mobile, and Algolia. We protect 200+ organizations across North America, Europe, the Middle East, Southeast Asia, Australia, and New Zealand, including many of the world's largest Fortune 1000 and Global 2000 companies.

Founded in 2017 and backed by top investors like Greylock, Obsidian was built to close a critical gap: securing SaaS apps where business happens—Microsoft 365, Salesforce, and hundreds more. The company does this by offering a complete SaaS security platform to reduce risk, detect and respond to threats, and prevent breaches at the source. Obsidian was built by leaders who redefined endpoint and identity security at CrowdStrike, Okta, Cylance, and Carbon Black. Now, they're transforming how SaaS is secured.

With AI driving rapid SaaS growth and complexity, agentic AI tools gain privileged access to sensitive data through integrations, creating new risks most security tools miss. Obsidian uniquely detects anomalous OAuth token activity and manages integration risks. Major announcements are on the horizon. Recognizing that SaaS security needs to evolve, Obsidian enables growing organizations to start with a lightweight, prevention-focused browser extension and expand coverage over time.

With global momentum, a growing partner ecosystem including SentinelOne, Databricks, and Google Cloud, and a major fundraise ahead, Obsidian is scaling rapidly toward long-term growth and IPO readiness.

Sr. Staff Site Reliability Engineer

As a Sr. Staff SRE at Obsidian , you will define and drive the company-wide reliability vision for a complex, multi-tenant SaaS platform serving enterprise and financial customers. You will operate as a strategic partner to DevOps and Platform Engineering leadership, shaping a unified reliability strategy that scales across the organization.

Your core mandate: ensure Obsidian detects, diagnoses, and communicates system issues before customers are impacted—consistently and predictably.

This is a hands-on technical role that involves architecting and leading the implementation of systems that handle real-world complexity, including upstream SaaS dependencies, sparse and noisy signals, and mission-critical enterprise workloads.

Key Responsibilities

  • Reliability Strategy & Architecture - Define and lead long-term reliability strategy across services. Establish end-to-end system visibility frameworks and guide architecture for observability, detection, and resilience.
  • Cross-Org Leadership - Partner across teams to embed reliability, standardize SLI/SLOs, and serve as a technical escalation expert.
  • Detection & Observability - Build intelligent detection systems (anomaly detection, connector health models) and enable self-service observability.
  • Incident Management - Define and evolve a tiered incident communication strategy , improve response practices, and lead postmortems to strengthen reliability and customer trust.
  • Execution - Contribute hands-on to system design, monitoring, and debugging across distributed systems and data pipelines.

Required Qualifications

  • 5+ years in SRE, Production Engineering, or related roles
  • 3+ years operating at a senior or technical leadership level (Staff or equivalent scope)
  • Deep expertise in:
    • AWS and/or GCP
    • Kubernetes and Helm
    • Observability stacks (Prometheus, Grafana, or equivalent)
    • CI/CD systems (GitLab CI/CD, ArgoCD, etc.)
  • Proven experience designing and scaling reliability systems for multi-tenant SaaS platforms
  • Strong debugging and systems thinking across distributed microservices and legacy systems
  • Demonstrated ability to lead initiatives that improve incident detection, response, and system resilience
  • Hands-on engineering approach with a track record of building—not just configuring—reliability systems

Preferred Qualifications

  • Experience in B2B SaaS serving enterprise or financial customers
  • Familiarity with third-party SaaS connector architectures and ingestion patterns
  • Experience building anomaly detection or intelligent alerting systems
  • Experience designing customer-facing status pages and incident communication frameworks

Why This Role

  • Drive org-wide reliability strategy
  • Own and build new detection & observability systems
  • Tackle complex distributed systems challenges
  • Safeguard critical infrastructure for financial customers

What Success Looks Like

  • Issues caught and resolved before customer impact
  • Reliability is measurable and continuously improving
  • Teams self-serve observability with scalable tools
  • Clear, proactive incident communication builds trust
  • Reliability becomes a competitive advantage

Employee Benefits

Our competitive benefits packages are designed to support our employees' well-being, both at work and at home. Our US based employees enjoy:

  • Competitive compensation with equity and 401k
  • Comprehensive healthcare with dental and vision coverage
  • Flexible paid time off and paid holiday time off
  • 12 weeks of new parent or family leave
  • Personal and professional development resources

For more details on our US benefits, or for information on our international benefits, please see here.

Pay Transparancy

Please note that the base pay range is a guideline and for candidates who receive an offer, the base pay will vary based on factors such as work location, as well as the knowledge, skills and experience of the candidate. In addition to a competitive base salary, this position is eligible for equity awards and may be eligible for sales commission or incentive compensation based on the role or function within the company.

At Obsidian, we are proud to be an equal-opportunity employer. We value diversity and hire for talent, passion, and compassion. In compliance with federal law, all persons hired will be required to submit satisfactory proof of identity and legal authorization. If you have a need that requires accommodation, please contact View email address on ziprecruiter.com

Information collected and processed as part of any job applications you choose to submit is subject to Obsidian's Applicant Privacy Policy.

Base Salary Range

$232,000—$263,000 USD

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Sr. Staff Site Reliability Engineer in Palo Alto, CA vacancy
  •  ...exceptional professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of...  ...and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan Chase... 
    Senior

    J.P. Morgan

    Palo Alto, CA
    4 days ago
  • $217.57k - $260k

     ...description explicitly states otherwise, all roles are on-site five days per week at one of our offices in McLean, VA; Mountain...  ...which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure... 
    Suggested
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours
    Shift work

    ID.me

    Mountain View, CA
    23 hours ago
  • $165k - $280k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s... 
    Senior
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Palo Alto, CA
    2 days ago
  •  ...ActiveHours is looking for an experienced DevOps Engineer to enhance platform automation and site reliability in a collaborative environment. The role involves automating key systems, improving visibility through metrics, and troubleshooting critical problems. You'll... 
    Suggested

    ActiveHours

    Palo Alto, CA
    2 days ago
  •  ...Acryl Data seeks a Site Reliability Engineering (SRE) Tech Lead to enhance the reliability and scalability of its DataHub platform. The role involves leading infrastructure design, optimizing system performance, and driving continuous improvement across cloud deployments... 
    Suggested

    Acryl Data

    Palo Alto, CA
    2 days ago
  •  ...Site Reliability Engineer, Data Platform - USDS Responsibilities Engage in and improve the whole lifecycle of service, from inception and design, through to deployment, operation and refinement. Ensure reliable, fault-tolerant, efficiently scalable and cost-effective data... 

    Tik Tok

    Mountain View, CA
    2 days ago
  •  ..., Elise AI, IBM and Accern. Position Summary We are hiring for a hands‑on Head of SRE to establish, lead, and scale our Site Reliability Engineering function. This role combines strategic ownership with deep technical execution. You will be responsible for defining reliability... 
    Shift work

    Wand AI

    Palo Alto, CA
    2 days ago
  • $100k - $200k

     ...OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about... 
    Full time

    OPPO

    Palo Alto, CA
    3 days ago
  •  ...Job Description As a Senior DevOps Engineer, you will: Provide extended U.S. coverage for SEV-1 and SEV-2 production incidents...  ...Infrastructure as Code (Terraform) to ensure scalability, security, and reliability. Build and maintain monitoring and observability solutions... 
    Senior

    CultureFit

    Palo Alto, CA
    11 days ago
  •  ...A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure. The ideal candidate has over 8 years of experience in wireless network operations and a strong background in wireless... 
    Senior

    TechDigital Group

    Santa Clara, CA
    2 days ago
  • $145k - $165k

     ...A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key... 
    Senior

    Bolt Graphics, Inc.

    Sunnyvale, CA
    2 days ago
  • $174k - $253k

     ...MINIMUM QUALIFICATIONS: Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical experience. 5...  ...s degree in Computer Science or Engineering. ABOUT THE JOB: Site Reliability Engineering (SRE) is what you get when you treat operations... 
    Senior

    Socket

    Sunnyvale, CA
    1 day ago
  • SKU Testing Functional Testing L1 V2 Manual testing SKU Testing Functional Testing L1 V2 Manual testingSKU Testing Functional Testing L1 V2 Manual testing SKU Testing Functional Testing L1 V2 Manual testing
    Senior

    Texas State Library and Archives Commision

    Mountain View, CA
    3 days ago
  • $190.9k - $334.1k

     ...Company Description It all started when engineer Fred Luddy wrote code that automated a...  ...needs not yet anticipated. ~ Drive reliability and operability across the platform through...  ...-person team culture, collaborating on-site with colleagues every day. For... 
    Senior
    Full time
    Work at office
    Immediate start
    Remote work
    Flexible hours
    Shift work

    ServiceNow

    Mountain View, CA
    4 days ago
  • $129k - $212k

     ...select days, as determined by the business needs of the team.  Our mission is to provide a reliable, efficient and intuitive development ecosystem to empower our iOS engineers to be productive. We serve hundreds of iOS engineers who build Linkedin's iOS applications.... 
    Senior
    For contractors
    Work experience placement
    Work at office
    Flexible hours

    LinkedIn

    Mountain View, CA
    11 days ago
  • $167.2k - $316.6k

     ...components of the Android framework, enhancing the performance, reliability, and security of our IVI platform. The ideal candidate will...  ....Bachelor’s or Master’s degree in Computer Science, Software Engineering, or equivalent combination of relevant education and... 
    Senior
    Immediate start
    Visa sponsorship
    Flexible hours

    Ford

    Palo Alto, CA
    2 hours ago
  • $204k - $259k

     ...performance analysis, full-system debugging, system telemetry, and reliability. We work closely with the Hardware, Compute, Sensor,...  ...Lead Manager. You will: Work on a small team of Software Engineers to develop system software components from early prototyping to... 
    Senior
    Full time
    Work experience placement
    Remote work

    Waymo

    Mountain View, CA
    1 day ago
  • Job Recommendation Confirmation When you upload your resume, we provide job recommendations to you. Please confirm you have read and understand how your data may be processed pursuant to the Microsoft Data Privacy Notice and Transparency FAQ.
    Senior

    Microsoft Corporation

    Mountain View, CA
    12 hours ago
  • $193.93k - $291.15k

     ...remotely plays a key role in our business strategy. As a Software Engineer, Video Streaming you will work on our in-house Teleoperations...  ...-time communication systems. The team is expected to deliver reliable solutions and license to 3rd party teleoperation usages. About... 
    Senior
    Full time
    Remote work

    Nuro

    Mountain View, CA
    1 day ago
  • $284.9k - $427.3k

     ...architecture specification documents Maintain strong communication skills and work effectively in a dynamic environment with senior engineers and technologists Minimum Qualifications Bachelor’s degree in Electrical Engineering, Computer Engineering, Computer Science... 
    Senior
    Work from home

    Qualcomm

    Santa Clara, CA
    1 day ago
  •  ...can use deep data insights to improve their business. Our engineering teams build technical products that fulfill real, important needs...  ...all Databricks engineers and customers to monitor the reliability of our product. You will develop advanced workflows that accelerate... 
    Senior
    Full time

    Databricks

    Mountain View, CA
    1 day ago
  •  ...Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will... 
    Senior

    Kody

    Palo Alto, CA
    14 days ago
  • $180k - $260k

     ...management, and real-time communication. Designs low-latency, high-reliability protocols and safety-critical workflows. Collaborates across...  ...cross-functionally with teams in behavior planning, platform engineering, integration, safety, and operations to align technical... 
    Senior
    Odd job
    Full time
    Work at office
    Remote work

    Gatik Ai

    Mountain View, CA
    1 day ago
  • $193.3k - $261.5k

     ...Sr. Software Development Engineer - Amazon Redshift, Query Processing ID lavoro: 3119625 | Amazon Development Center U.S., Inc. Amazon Redshift is...  ...architecture of new and existing systems (design patterns, reliability, and scaling). Experience as a mentor, tech lead, or... 
    Senior

    Amazon

    Palo Alto, CA
    3 days ago
  •  ..., and the challenges of building in a high-growth startup, we’d love to talk. This is more than a job—it’s a journey. Site Reliability Engineers (SREs) are responsible for the overall performance and reliability of ASAPP's infrastructure and products. The team owns... 
    Senior
    Remote work

    ASAPP

    Mountain View, CA
    24 days ago
  • $204k - $259k

     ...Development: Work on a small team of System Software and Linux Kernel Engineers to design, develop, and deploy production-grade system...  ...robust, high-performance networking solutions, focusing on reliable data transfer over wifi and cellular modems. Infrastructure... 
    Senior
    Full time
    Remote work

    Waymo

    Mountain View, CA
    1 day ago
  • $210k - $250k

     ...Ad Experience Team Engineer The Ad Experience team builds the products that greet every...  ...peers and leaders across various Samsung Ad sites, driving cross-site efforts to improve...  ...browser compatibility. Passion for building reliable "done right the first time" applications.... 
    Senior
    Hourly pay
    Full time
    Worldwide

    Samsung

    Mountain View, CA
    5 days ago
  •  ...Job Title : Sr. Ios Developer Location : Mountain View, CA (hybrid) Job Description: Create impactful iOS...  ...Act as the technical subject matter expert, mentoring fellow engineers, leading a small team, and solving challenging programming and design... 
    Senior
    Work experience placement

    Ruri Software Technologies LLC

    Mountain View, CA
    4 days ago
  • $160k - $200k

     ...genetic datasets in the world. We're looking for a Senior Software Engineer to help build and operate the systems that make that data...  ...science, and engineering teams to translate requirements into reliable, scalable systems Build AI-powered data pipelines that deliver... 
    Senior
    Local area

    23andMe

    Palo Alto, CA
    4 days ago
  • $140k - $205k

    Senior Technology Site Reliability EngineerCooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operations team.Position summary: The Senior Technology Site Reliability Engineer (“SRE”) is responsible for ensuring the reliability... 
    Senior
    Full time
    Temporary work
    Work at office
    Flexible hours
    Weekend work

    Cooley

    Palo Alto, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Sr. Staff Site Reliability Engineer. Be the first to apply!