Staff Site Reliability Engineer - Observability GCP (San Francisco)
$194k - $267kOkta
Secure Every Identity, from AI to HumanIdentity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.We are seeking a highly technical ObservabilitySite Reliability Engineer with a specialty in Google Cloud, to own and expand our Observability ecosystem into GCP. In this role, you will move beyond simple monitoring to delivering a world class, comprehensive, scalable Observability Platform that enables our SRE teams and business partners. You will treat infrastructure as code—utilizing Terraform and strong coding proficiency in Go, Python, or Ruby—to automate the deployment of agents and collectors across complex distributed systems.Key ResponsibilitiesAutomated Infrastructure: Design, build, and maintain scalable observability infrastructure using tools like Terraform.GCP Observabilty Engineering: Optimize the collection, processing, and storage of Observabilty data to ensure high reliability and low latency of our Splunk and Grafana servicesIncident Response: Participate in on-call rotations and lead post-incident reviews to drive systemic improvements and observability-driven development.Automation: Eliminate toil by automating the deployment and scaling of observability agents and collectors.Required Skills & Experience (The Essentials)GKE: Minimum 5+ Experience scaling and managing observability in a Google Cloud platform. Visualization: Expertise in creating intuitive, actionable Splunk or Grafana dashboards that correlate data across multiple sources.SRE Mindset: Minimum 3+ years of experience in an SRE, DevOps, or Systems Engineering role with a focus on high-availability systems.Programming Proficiency: Strong coding skills in Python, Go for building internal tools and automating workflows.Distributed Systems: Deep understanding of Linux internals, networking (TCP/IP, DNS, Load Balancing), and container orchestration (Kubernetes/GKE).Problem Solving: A data-driven approach to debugging complex, cross-service performance bottlenecks.Bonus Skills (The Nice-to-Haves)Telemetry Standards: Hands-on experience with OpenTelemetry (OTel), Vector, or similar frameworks for instrumenting applications.Grafana Loki: Experience in migrating Splunk to Grafana LokiOther Cloud Platforms: Experience managing observability native tools within AWS.Additional requirements:This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.#LI-MM#LI-HybridP24517_3387022The annual base salary range for this position for candidates located in the San Francisco Bay area is between: $194,000—$267,000 USDBelow is the annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York and Washington. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: . The annual base salary range for this position for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York, and Washington is between:$174,000—$239,000 USDThe Okta ExperienceSupporting Your Well-BeingDriving Social Impact Developing Talent and Fostering Connection + CommunityWe are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use this Form to request an accommodation.Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please click here to view our full NYC AEDT Notice.
- ...Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with... ...solutions that blend automation, observability, and AI-powered platforms.... ...reliability/observability tooling.AWS/GCP professional-level...SuggestedPart timeWorldwideWeekend work
- ...company is headquartered in San Francisco with offices in New York,... ...and tooling that help engineering teams develop, deploy, and... ...every product team.As a Staff Site Reliability Engineer on Release Engineering... ...through expertise in observability, incident response, and...SuggestedPermanent employmentPart timeWork experience placementWork at officeLocal area
- ...looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence... ...: $210K - $240KLocationSan Francisco HQAddressSan Francisco, CaliforniaEmployment... ...with you. Most recent San Francisco benchmark data: $...SuggestedPart time
- ...About the teamThe Engineering team at Airwallex... ...to build scalable, reliable, and secure products... ...’ll doAs a Senior Site Reliability... ...incident response, observability, and automation across... ...cloud platforms (AWS/GCP), Kubernetes,... ...$250KLocationUS - San FranciscoEmployment...SuggestedTemporary workPart timeLocal areaWorldwide
$194k - $267k
...concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in... ...service communication, security, and observability within the Kubernetes clusters. Enable... ...person onboarding and travel to our San Francisco, CA HQ office or our Chicago office...SuggestedPermanent employmentPart timeWork at officeLocal areaWorldwideFlexible hours$227.2k - $324.5k
...About the Role:Site Reliability Engineering (SRE) at Tubi is not a traditional operations team. We... ...technical strategy and vision for Tubi’s observability, and automation platforms. Partner... ...to Los Angeles, New York City, and San Francisco$227,200—$324,500 USDTubi is a...Full timeContract workTemporary workPart timeLocal areaFlexible hours$204k - $281k
...too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity... ...2 days a week in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates,... ..., K8s platform, CI/CD, Observability, automation platform & tooling...Permanent employmentPart timeWork at officeLocal areaWorldwideFlexible hours2 days per week$215k - $275k
...Anyscale is looking for a Senior Site Reliability Engineer to join the Infrastructure... ..., scalability, and observability of Anyscale-managed Ray workloadsSupport... ...technologies (AWS, Azure, GCP) and Kubernetes-based... ...: $215K - $275KLocationSan Francisco; Palo AltoEmployment...Part timeWork at office- ...under real-world scale, reliability, and security demands... ...we're looking for an engineer who wants to own the foundation... ...systems, automation, observability, and reliability... ...0K - $240KLocationSan Francisco HQAddressSan Francisco... ...with you. Most recent San Francisco benchmark...Part time
$232k - $319k
...service with great people and reliable, cost-effective, and... ...networking, K8s platform, Observability, automation platform & tooling... ...velocity of SRE and product engineering by developing robust platforms... ...in California (excluding San Francisco Bay Area), Colorado, Illinois...Permanent employmentPart timeLocal areaWorldwideFlexible hours- ...managers, economists, and engineers from Google, Netflix,... ...partner, and ship reliable systems without... ...error handling, and observability Work with product... ...native environments (GCP preferred): you understand... ...distance of our offices in San Francisco, Seattle, and New...Full timeWork at officeWork from homeWorldwideFlexible hours
- ...San Francisco, California100% RemoteFull Time$188k - $248kA leading... ...is hiring a Senior Staff Systems Design Engineer to join a growing cloud and... ...Required Skills and Experience GCP expertise with production-... ...Optimize performance, cost, and reliability of systems Collaborate...Full timePart timeRemote work
$195.5k - $264.5k
...at the heart of Pave Engineering. Our customers are Pave... ...software fast and reliably.This is a high-impact... ...end, from CI/CD and observability standards to local development... ...cloud architecture (GCP or similar), with a... .... Headquartered in San Francisco's Financial District,...Part timeLocal area3 days per week$165k - $247k
...every Amplitude engineer relies on... ...partner with Staff engineers and... ...developer experience, reliability, or security,... ..., AWS, and GCP using Terraform... ...operate. Drive observability with Datadog... ...engineering, DevOps, or Site Reliability... ...Amplitude in San Francisco Bay Area of...Part timeShift work$159.2k - $301.6k
...on the cloud. In this reliability-focused role, you will... ...partner with the backend engineers building these APIs to... ...Build and maintain observability—metrics, logging, tracing... ...of experience in site reliability engineering... ...liability.SummaryLocation: San Jose; Seattle; San...Full timeTemporary workPart timeLocal areaWorldwide- ...building software to ensure the reliability of our back-end systems, working with engineers who develop them, and planning for... ...terraform.Have used AWS, Azure, or GCP.Java and Kubernetes skills... ...Range: $214K - $260KLocationHub - San FranciscoEmployment TypeFull timeLocation...Part timeWorldwideHome officeFlexible hours
- ...it culture, where engineers are responsible... ...performance, and reliability of what they ship... ...Member of Technical Staff (SMTS) to join... ...performance, and observability standards.Actively... ...(AWS preferred; GCP or Azure also considered... ...link: to the San Francisco Fair Chance...Part time
$165k - $247k
...OpportunityAs a Senior Software Engineer, Applied AI, you’ll... ...standards (SLOs, observability, incident response... ...ensure services scale reliably across the org.What You... ...cloud platforms (AWS/GCP/Azure).Familiarity with... ...candidates located in the San Francisco Bay area is between: $...Part timeLocal areaWorldwideFlexible hours$220k - $260k
...for an experienced Staff Software Engineer to help architect and... ...hybrid role based in our San Francisco office.You Will:... ...across AWS, GCP, and Azure, including... ...high standards for reliability, scalability, performance... ...healingMake resilience and observability integral to the...Part timeWork at office$245k - $290k
...looking for a Principal Software Engineer to help architect and build... ..., and a global policy and observability framework that enforces... ...a hybrid role based in our San Francisco office.You Will:Track emerging... ...infrastructure across AWS, GCP, or AzurePrior experience in...Part timeWork at officeRemote workShift work$172.5k - $260.1k
...Salesforce.Lead Software Engineer – Platform ServicesAt... ...design, testing, observability, scalability, and operational... ...building scalable, reliable, and secure customer-... ...microservices on AWS (preferred), GCP, or other public cloud... ...link: to the San Francisco Fair Chance Ordinance...Full timePart time$172.6k - $304.1k
...office hubs in San Francisco, New York City... ...software engineer to join our rapidly... ...that enable reliable agent workflows... ...SLAs, maturing observability, and driving... ...across Senior to Staff based on... ...experience on AWS, GCP, or Azure,... ...stipendCompany-wide off-sites and team off-...Full timePart timeWork at officeLocal areaFlexible hours$192k - $240k
...to grow your career.Engineering at BrexEngineering at... ...power Brex’s release, observability, and incident management... ...are safe, fast, and reliable, and that our... ...will be based in our San Francisco office. We are a hybrid... ...platforms (e.g., AWS, GCP, Azure)Deep familiarity...Part timeWork at officeRemote workWork from home$220k - $235k
...seeking a strategic, high-output Staff/Senior Staff SRE to define... ...platform and champion engineering excellence across Ironclad.... ...strategic direction for the Site Reliability Engineering team and our... ...this position based at our San Francisco headquarters. The actual base...Full timeContract workPart timeWork at office$174k - $239k
...we partner across functions to drive scale, reliability, and innovation through technology.The Staff Site Reliability Engineer OpportunityOkta Federal, Inc. is looking... ...candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York and...Part timeWork experience placementLocal areaWorldwideFlexible hours$194k - $267k
...you are too, let's talk.The TeamThe Site Reliability team is dedicated to architecting and... ...that maximize platform reliability and engineering velocity.The ideal candidate is... ...position for candidates located in the San Francisco Bay area is between: $194,000—$267,00...Part timeLocal areaWorldwideFlexible hours$230k - $390k
...person company based in San Francisco, with growing offices... ...’ll doAs a Software Engineer, Infrastructure at Sierra... ...secure, reliable, and scalable, enabling... ...maintainable ways.Enhance observability tooling (metrics, logging... ...cloud platforms (AWS, GCP, or Azure) and infrastructure...Full timePart timeFlexible hours- ...Fleet team is the engine behind how our platform... ...you come in.As a Staff Software Engineer,... ...are resilient, observable, and require... ...practices in CI/CD, reliability, container lifecycle... ...cloud providers (AWS, GCP, Azure) and/or... ...ON; Remote Poland; San Francisco, California, USType...Full timePart timeLocal areaRemote workWorldwideFlexible hours
$197.3k - $313.7k
...As the Principal Engineer focused on architecture... ..., performance, observability, and user... ...and role based sites such as the developer... ...operate accurately and reliably.Critically... ...AWS (preferred), GCP, or other public... ...following link: to the San Francisco Fair Chance...Full timePart time$237.5k - $255k
...operations.With offices in San Francisco and Bengaluru, Instabase... ...are seeking a visionary Staff Software Engineer to join our Product... ...maximize the velocity, reliability, and observability of all engineering teams... ...cloud infrastructure (AWS, GCP, or Azure). Deep expertise...Part timeWork at officeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer - Observability GCP (San Francisco). Be the first to apply!
- staff security engineer San Francisco, CA
- staff data engineer San Francisco, CA
- senior staff engineer San Francisco, CA
- engineering aide San Francisco, CA
- software engineer staff San Francisco, CA
- assistant engineer San Francisco, CA
- technology administrator San Francisco, CA
- senior staff systems engineer San Francisco, CA
- staff engineer San Francisco, CA
- site reliability engineer San Francisco, CA




















