Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

Full-time

Supabase

About Supabase: Supabase is a premier, internationally recognized open-source technology juggernaut, Postgres development platform titan, and Firebase alternative pioneer operating on an absolute mission to protect, optimize, and transform how developers manage backend infrastructure. Offering an all-in-one suite that includes deeply integrated Postgres Databases, Authentication grids, Edge Functions, Realtime channels, Storage vaults, and Vector Search clusters, Supabase serves an accelerating user base managing millions of database instances. Backed by $500M in venture funding and scaling a high-vibe global network of over 500,000+ community members, Supabase is a born-remote, open-source-first organization that builds in public, values technical excellence, and utilizes its own product stack in everyday internal operations. The company provides high-agency systems engineering leaders with an uncompromised remote canvas to leverage state-of-the-art cloud systems, manipulate multi-tenant data pipelines, and deploy robust, automation-driven SRE frameworks globally.

Position Overview

We are seeking a highly analytical, detail-obsessed, and systems-minded Site Reliability Engineer to join our core centralized Service Operations collective in a full-time remote capacity open to qualified infrastructure authorities resident anywhere across the globe. As we scale to support millions of concurrent Postgres nodes, we are concentrating our platform-wide availability initiatives into a dedicated SRE practice designed to tie our observability, release engineering, and incident pipelines together. Shifting completely away from routine manual system operations, reactive standalone alert logging, or acting as an isolated infrastructure cleanup crew, you will run an active reliability strategy and automation engineering laboratory—embedding alongside software feature teams to build the tools, runbooks, and feedback loops that allow them to own availability themselves. This position requires an infrastructure or developer-tooling veteran with 7+ years of craft depth who maps out scalable cloud patterns fluidly natively using DevOps mechanics, builds internal platform extensions or reliability dashboards cleanly natively leveraging Python or alternative software engineering code bases, and commands high-concurrency cloud deployments confidently under asynchronous, influence-driven distributed models.

Key Responsibilities

  • SRE Practice and Policy Architecture: Collaborate directly with distributed engineering units to formulate, document, and embed meaningful Service Level Indicators (SLIs) and Objectives (SLOs) tied to end-user experiences, enforcing code-driven error budgets cleanly natively utilizing DevOps methodologies.
  • Operational Readiness Governance (ORR): Own and evolve the systemic Operational Readiness Review (ORR) framework, conducting exhaustive architecture reviews, dependency mapping, capacity audits, and failure mode analyses for major platform updates.
  • Incident-to-Improvement Orchestration: Maximize the impact of our postmortem pipeline, facilitating deep root-cause investigations, identifying cross-platform failure signatures, and driving systemic code improvements to eliminate recurring operational risks.
  • Operational Toil Elimination: Identify, track, and quantify recurring administrative manual friction points across the engineering organization, writing automated developer-facing reliability tools cleanly natively leveraging Python or cloud-native script interfaces to replace them.
  • Sustainable On-Call Design: Help development teams engineer resilient on-call protocols, optimizing alert routing systems, minimizing warning noise, and ensuring absolute runbook documentation coverage.
  • Maturity and Resilience Tracking: Monitor and map organizational infrastructure maturity vectors, surfacing foundational design gaps and advising leadership blocks on systemic engineering remediation priorities.
  • Asynchronous Cloud Deployment: Write and optimize infrastructure-as-code definitions to manage complex multi-tenant system footprints inside Amazon Web Services (AWS) or alternative cloud topologies.

Required Skills & Qualifications

  • A minimum of 7 years of verified professional history running advanced Site Reliability Engineering (SRE), production software engineering, infrastructure architecture, or cloud-scale systems optimization.
  • Expert-tier capability automating infrastructure environments, managing multi-tenant networks, and deploying cloud systems cleanly natively utilizing DevOps parameters.
  • Practical operational familiarity developing testing runbooks, automating diagnostic loops, or parsing system logging outputs natively using Python or related software development runtimes.
  • Demonstrated software engineering mindset, showing a powerful track record of writing code, building customized reliability tools (such as SLO dashboards or ORR frameworks), and developing APIs rather than simply adjusting vendor configuration templates.
  • Hands-on experience operationalizing multi-tenant SLOs/SLIs at scale, including building out explicit error budget systems that actively directed high-level product engineering resource decisions.
  • Deep professional familiarity with distributed cloud infrastructure management (with an absolute preference for AWS) and programmatic Infrastructure-as-Code frameworks (with a preference for Pulumi, or advanced Terraform/AWS CDK models).
  • Outstanding written and scannable technical communication attributes in business-fluent English, enabling uncompromised capability to influence engineering structures without authority across an entirely distributed organization.
  • Location Context: Position open to qualified engineering craftspeople based anywhere globally to operate under a 100% remote work-from-home layout.

Preferred Strategic Indicators (Nice to Have)

  • Prior technical operations history managing large-scale distributed cloud database platforms, orchestrating cluster configurations, or handling Postgres engines at enterprise scale.
  • Direct hands-on experience structuring container operations inside Kubernetes-based production topologies.
  • Familiarity with cloud-native open-source observability ecosystems, including OpenTelemetry specifications, VictoriaMetrics datastores, or Grafana instrumentation.

What We Offer

  • Vetted Open-Source Sector Salaried Blueprint: A highly competitive, full-time global baseline annual corporate salary scale calibrated precisely to evaluate your SRE authority and systems craftsmanship, paired with immediate equity ownership through an impactful Employee Stock Ownership Plan (ESOP).
  • The spectacular professional canvas to claim absolute strategic ownership over the reliability systems protecting database instances for hundreds of thousands of developers worldwide.
  • Profound work-from-home remote parameters offering a 100% remote virtual layout anywhere on earth, complete scheduling trust, and zero physical geographic commuting friction, complemented by a global co-working allowance or WeWork membership.
  • Immediate access to top-tier health benefits, featuring 100% company-paid premium medical coverage for employees alongside an immediate 80% coverage match for dependents.
  • Access to elite lifestyle and wealth accumulation tracks, including a dedicated personal Tech Allowance budget to configure your ideal laptop, monitor, and accessory layout, an annual professional development education allowance, and highly flexible asynchronous work hours.
  • Direct company-funded access to our spectacular Annual Team Offsites, bringing the entire global team together in a new international city for a week of intense collaboration and connection.
Vacancy posted 15 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Remote vacancy
  •  ...ears, and hands on the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-critical platform running...  ...of deep technical ownership and customer-facing engineering: you'll define how we measure reliability, lead incident... 
    Suggested
    Full time
    Work at office
    Remote work
    Flexible hours

    Twenty

    Maryland
    22 days ago
  • Role Description We are looking for a Site Reliability Engineer (SRE) who is passionate about infrastructure reliability, automation, and building scalable production systems. ~Own and improve production infrastructure reliability and stability ~Prepare, execute, and... 
    Suggested
    Full time
    Remote work

    Social Discovery Group

    Remote
    7 days ago
  • Role Description En Experis Argentina nos encontramos en la búsqueda de nuestro Site Reliability Engineer (SRE) para importante Compañía del rubro de Telecomunicaciones. Condiciones de contratación: ~Jornada de trabajo: Full Time 9am – 6pm. ~Modalidad de trabajo:... 
    Suggested
    Full time
    Remote work

    ManpowerGroup Argentina

    Remote
    10 days ago
  • Role Description Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and services. Our SRE teams solve reliability, security, and usability at scale for... 
    Suggested
    Full time
    Work at office

    Akamai

    Remote
    5 days ago
  • $114k - $148k

    Role Description As a Site Reliability Engineer, you will focus on ensuring the platform and services customers rely on are reliable, performant, and highly available. If you enjoy staying at the forefront of technology and automating infrastructure deployments, then this... 
    Suggested
    Full time
    Temporary work
    Work experience placement

    OneStream Software

    Remote
    3 days ago
  • $110k - $137.49k

    Role Description The Sr Site Reliability Engineer, Release will prototype, write, maintain, and test code in multiple stages of the release process and in multiple environments in order to rapidly deliver automated solutions to our application releases. Assesses unusual... 
    Full time
    Work experience placement
    Remote work

    Alkami Technology

    Remote
    2 days ago
  • $54k - $150k

    Role Description As Senior Site Reliability Engineer for Remote Build, you'll own the operational excellence and infrastructure strategy that makes Build's platform reliable, performant, and safe for customers. You'll report to the Engineering Manager and work closely with... 
    Full time
    Local area
    Remote work
    Home office
    Flexible hours

    Remote

    Remote
    3 days ago
  • $190k - $240k

    Role Description As a Sr. Site Reliability Engineer (SRE) at ICD, you will play a critical role in ensuring the reliability and seamless operation of our global platform and AWS infrastructure to create scalable and highly reliable software systems. Job Responsibilities... 
    Full time
    Work at office
    Immediate start
    Flexible hours

    Tradeweb

    Remote
    2 days ago
  • Role Description Stack AV Site Reliability Engineers are responsible for enabling and ensuring our production systems meet their service-level objectives. Through the implementation of centralized observability and automation, the SRE team constantly ensures the health... 
    Full time

    Stack AV

    Remote
    5 days ago
  • Role Description We are seeking a Site Reliability Engineer (SRE) with deep expertise in monitoring, observability, and reliability engineering to support systems running across on-premises infrastructure and Google Cloud Platform (GCP). This role is primarily responsible... 
    Long term contract
    Full time
    Remote work
    Flexible hours

    Devsu

    Remote
    2 days ago
  • Role Description We’re looking for a Senior Platform Engineer to design, build, and operate the core services that power Optura’s AI...  ...systems end-to-end, from model and agent orchestration to routing, reliability, and observability. You will partner closely with product and... 
    Full time
    Remote work

    Optura

    Remote
    1 day ago
  •  ...through intelligent automation and modern engineering. We are seeking a Senior SRE Engineer...  ...efficient delivery, observability, and reliability across Sleek’s products and internal operations...  ...~6+ years of progressive experience in Site Reliability Engineering (SRE). ~6+... 
    Full time
    Remote work
    Flexible hours

    Sleek

    Remote
    5 days ago
  • $140k - $165k

     ...and optimizing development velocity without compromising on reliability. Qualifications ~Experience in an SRE role with an...  ...plan you will receive if you were to be hired as a Senior Site Reliability Engineer at Flock Safety. The First 30 Days ~Onboarding. Make a... 
    Full time
    Work at office
    Work from home
    Home office
    Flexible hours

    Flock

    Remote
    1 day ago
  • $54k - $150k

    Role Description As Senior Site Reliability Engineer for Remote Build, you'll own the operational excellence and infrastructure strategy that makes Build's platform reliable, performant, and safe for customers. You'll report to the Engineering Manager and work closely with... 
    Full time
    Local area
    Immediate start
    Remote work
    Home office
    Flexible hours

    Referral Board

    Remote
    2 hours ago
  • Role Description We're looking for an SRE to own the reliability, scalability, and operability of the SigNoz cloud platform. You'll keep...  ...Benefits ~Work on a globally used open-source project that engineers actually love. ~Huge scope and ownership — your work directly... 
    Full time
    Remote work

    SigNoz

    Remote
    5 days ago
  • Role Description Versant's Sports & Entertainment Digital Products division is seeking a Senior Site Reliability Engineer to help drive the reliability, scalability, and usability of internal developer platforms, tooling, and engineering workflows across a portfolio of... 
    Full time
    Local area
    Remote work
    Worldwide

    Versant

    Remote
    4 days ago
  • Role Description We are looking for a Site Reliability Engineer (SRE) to join our world-class team. This isn't just an operational "maintenance" role; you will be software-engineering the engine that powers tens of thousands of daily builds, ensuring our platform is as... 
    Full time

    source.dev

    Remote
    2 days ago
  • $130k - $160k

    Role Description We are seeking a Site Reliability Engineer to design, build, and maintain highly available systems and infrastructure. The SRE will work closely with software developers and operations teams to improve system reliability, automate processes, and minimize... 
    Full time

    BRG

    Remote
    2 days ago
  • Role Description We’re looking for a Senior Site Reliability Engineer who takes ownership seriously — someone who designs for reliability, ships the automation, and stands behind it in production. You’ll work across cloud-native infrastructure on systems that process millions... 
    Full time

    CertifyOS

    Remote
    1 day ago
  • Role Description As a Site Reliability Engineer on the Central AI team, you will help Health Catalyst engineer teams adopt AI responsibly and effectively. You bring deep experience solutioning and implementing AI systems, and you use that expertise to evaluate architectures... 
    Full time

    Health Catalyst

    Remote
    7 days ago
  • Role Description We are expanding our Site Reliability Engineering (SRE) team and seeking a highly skilled and passionate Senior SRE to join us. As a member of our growing SRE function, you will play a critical role in ensuring the reliability, scalability, and performance... 
    Full time
    Temporary work

    QAD, Inc.

    Remote
    1 day ago
  •  ...to grow, adapt and use your skills consistently. Our customers rely on us in the moments that matter. Engineering delivers on that promise. The Senior Site Reliability Engineer is responsible for ensuring our SaaS products are fast, stable and optimized for our... 
    Full time
    Work experience placement

    Donnelley Financial Solutions

    Remote
    7 hours ago
  • Role Description The Senior Site Reliability Engineer is a technical leader responsible for architecting the reliability strategy for large-scale, distributed government systems. You will lead the implementation of the SRE framework, driving the adoption of SLO-based management... 
    Contract work
    Remote work

    Arctiq

    Remote
    7 hours ago
  • $137.9k - $221.4k

     ...for someone to lead development aspects of the Infrastructure engineering team at ServiceTitan. You must have a strong background in...  ...leadership and strong architectural thought process. Our Site Reliability and Infrastructure Engineering team is an investment by Cloud... 
    Full time
    Immediate start
    Flexible hours

    ServiceTitan

    Remote
    7 days ago
  • $135k - $170k

    Role Description Climavision is seeking a Senior Site Reliability Engineer to contribute towards reliability, operational excellence, and production resilience for our customer-facing platform and weather data services. This role is focused on ensuring our systems consistently... 
    Full time
    Temporary work
    Flexible hours

    Climavision

    Remote
    5 days ago
  •  ...Istio) ~Defining and monitoring Service-Level Objectives (SLOs) and Service-Level Agreements (SLAs) to ensure that systems meet reliability and performance targets ~Monitoring Tools like New Relic, Prometheus, Grafana, and/or Datadog ~OpenTelemetry knowledge for... 
    Full time
    Remote work

    Shippo

    Remote
    1 day ago
  • Role Description We are looking for a talented and driven Sr. Site Reliability Engineering (SRE) to support our engineering team, which manages the infrastructure and services that power our Waystar products. This role is ideal for an experienced engineer who thrives in... 
    Full time
    Live out
    Flexible hours

    Waystar

    Remote
    1 day ago
  • Role Description As a Senior Site Reliability Engineer you will champion all things pertaining to reliability at Okta for Auth0. Working closely with the Product Engineers, Quality Engineers, Platform Engineers and Architecture teams, your primary focus will be on ensuring... 
    Full time
    Remote work

    Okta

    Remote
    2 hours ago
  • $95k - $171k

     .... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Columbia, SC
    1 day ago
  • $96k - $163k

     ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that... 
    Full time
    Part time
    Worldwide
    Flexible hours

    MasterCard

    O Fallon, MO
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!