Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

Full-time

Supabase

About Supabase: Supabase is a premier, internationally recognized open-source technology juggernaut, Postgres development platform titan, and Firebase alternative pioneer operating on an absolute mission to protect, optimize, and transform how developers manage backend infrastructure. Offering an all-in-one suite that includes deeply integrated Postgres Databases, Authentication grids, Edge Functions, Realtime channels, Storage vaults, and Vector Search clusters, Supabase serves an accelerating user base managing millions of database instances. Backed by $500M in venture funding and scaling a high-vibe global network of over 500,000+ community members, Supabase is a born-remote, open-source-first organization that builds in public, values technical excellence, and utilizes its own product stack in everyday internal operations. The company provides high-agency systems engineering leaders with an uncompromised remote canvas to leverage state-of-the-art cloud systems, manipulate multi-tenant data pipelines, and deploy robust, automation-driven SRE frameworks globally.

Position Overview

We are seeking a highly analytical, detail-obsessed, and systems-minded Site Reliability Engineer to join our core centralized Service Operations collective in a full-time remote capacity open to qualified infrastructure authorities resident anywhere across the globe. As we scale to support millions of concurrent Postgres nodes, we are concentrating our platform-wide availability initiatives into a dedicated SRE practice designed to tie our observability, release engineering, and incident pipelines together. Shifting completely away from routine manual system operations, reactive standalone alert logging, or acting as an isolated infrastructure cleanup crew, you will run an active reliability strategy and automation engineering laboratory—embedding alongside software feature teams to build the tools, runbooks, and feedback loops that allow them to own availability themselves. This position requires an infrastructure or developer-tooling veteran with 7+ years of craft depth who maps out scalable cloud patterns fluidly natively using DevOps mechanics, builds internal platform extensions or reliability dashboards cleanly natively leveraging Python or alternative software engineering code bases, and commands high-concurrency cloud deployments confidently under asynchronous, influence-driven distributed models.

Key Responsibilities

  • SRE Practice and Policy Architecture: Collaborate directly with distributed engineering units to formulate, document, and embed meaningful Service Level Indicators (SLIs) and Objectives (SLOs) tied to end-user experiences, enforcing code-driven error budgets cleanly natively utilizing DevOps methodologies.
  • Operational Readiness Governance (ORR): Own and evolve the systemic Operational Readiness Review (ORR) framework, conducting exhaustive architecture reviews, dependency mapping, capacity audits, and failure mode analyses for major platform updates.
  • Incident-to-Improvement Orchestration: Maximize the impact of our postmortem pipeline, facilitating deep root-cause investigations, identifying cross-platform failure signatures, and driving systemic code improvements to eliminate recurring operational risks.
  • Operational Toil Elimination: Identify, track, and quantify recurring administrative manual friction points across the engineering organization, writing automated developer-facing reliability tools cleanly natively leveraging Python or cloud-native script interfaces to replace them.
  • Sustainable On-Call Design: Help development teams engineer resilient on-call protocols, optimizing alert routing systems, minimizing warning noise, and ensuring absolute runbook documentation coverage.
  • Maturity and Resilience Tracking: Monitor and map organizational infrastructure maturity vectors, surfacing foundational design gaps and advising leadership blocks on systemic engineering remediation priorities.
  • Asynchronous Cloud Deployment: Write and optimize infrastructure-as-code definitions to manage complex multi-tenant system footprints inside Amazon Web Services (AWS) or alternative cloud topologies.

Required Skills & Qualifications

  • A minimum of 7 years of verified professional history running advanced Site Reliability Engineering (SRE), production software engineering, infrastructure architecture, or cloud-scale systems optimization.
  • Expert-tier capability automating infrastructure environments, managing multi-tenant networks, and deploying cloud systems cleanly natively utilizing DevOps parameters.
  • Practical operational familiarity developing testing runbooks, automating diagnostic loops, or parsing system logging outputs natively using Python or related software development runtimes.
  • Demonstrated software engineering mindset, showing a powerful track record of writing code, building customized reliability tools (such as SLO dashboards or ORR frameworks), and developing APIs rather than simply adjusting vendor configuration templates.
  • Hands-on experience operationalizing multi-tenant SLOs/SLIs at scale, including building out explicit error budget systems that actively directed high-level product engineering resource decisions.
  • Deep professional familiarity with distributed cloud infrastructure management (with an absolute preference for AWS) and programmatic Infrastructure-as-Code frameworks (with a preference for Pulumi, or advanced Terraform/AWS CDK models).
  • Outstanding written and scannable technical communication attributes in business-fluent English, enabling uncompromised capability to influence engineering structures without authority across an entirely distributed organization.
  • Location Context: Position open to qualified engineering craftspeople based anywhere globally to operate under a 100% remote work-from-home layout.

Preferred Strategic Indicators (Nice to Have)

  • Prior technical operations history managing large-scale distributed cloud database platforms, orchestrating cluster configurations, or handling Postgres engines at enterprise scale.
  • Direct hands-on experience structuring container operations inside Kubernetes-based production topologies.
  • Familiarity with cloud-native open-source observability ecosystems, including OpenTelemetry specifications, VictoriaMetrics datastores, or Grafana instrumentation.

What We Offer

  • Vetted Open-Source Sector Salaried Blueprint: A highly competitive, full-time global baseline annual corporate salary scale calibrated precisely to evaluate your SRE authority and systems craftsmanship, paired with immediate equity ownership through an impactful Employee Stock Ownership Plan (ESOP).
  • The spectacular professional canvas to claim absolute strategic ownership over the reliability systems protecting database instances for hundreds of thousands of developers worldwide.
  • Profound work-from-home remote parameters offering a 100% remote virtual layout anywhere on earth, complete scheduling trust, and zero physical geographic commuting friction, complemented by a global co-working allowance or WeWork membership.
  • Immediate access to top-tier health benefits, featuring 100% company-paid premium medical coverage for employees alongside an immediate 80% coverage match for dependents.
  • Access to elite lifestyle and wealth accumulation tracks, including a dedicated personal Tech Allowance budget to configure your ideal laptop, monitor, and accessory layout, an annual professional development education allowance, and highly flexible asynchronous work hours.
  • Direct company-funded access to our spectacular Annual Team Offsites, bringing the entire global team together in a new international city for a week of intense collaboration and connection.
Vacancy posted 15 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Remote vacancy
  •  ...ears, and hands on the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-critical platform running...  ...of deep technical ownership and customer-facing engineering: you'll define how we measure reliability, lead incident... 
    Suggested
    Full time
    Work at office
    Remote work
    Flexible hours

    Twenty

    Arlington, VA
    22 days ago
  • $150k - $175k

     ...Site Reliability Engineer At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed... 
    Suggested
    Remote work

    ASAPP

    Mountain View, CA
    3 days ago
  •  ...Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public Trust clearance. What You Will Do We are seeking a Site Reliability Engineer (SRE) to support the SBA Disaster Lending Platform modernization effort in a remote... 
    Suggested
    Full time
    Local area
    Remote work

    System One

    McLean, VA
    3 days ago
  • $47.52 - $55.52 per hour

     ...in Charlotte, NC. This is a 12+ month contract opportunity. This role is responsible for designing and leading the Site Reliability Engineering (SRE) strategy across banking and payments environments. The SRE Lead will establish SRE standards, drive automation, and... 
    Suggested
    Hourly pay
    Permanent employment
    Contract work
    Remote work
    Flexible hours

    Genesis10

    United States
    1 day ago
  • $81k - $142k

     ...As a Site Reliability Engineer reporting to Director, System Operations, you'll play a critical role in the delivery, integration, and support of complex, distributed, high-availability solutions for financial institutions and organizations around the globe. You'll... 
    Suggested
    Temporary work
    H1b
    Work at office
    Home office
    Flexible hours
    3 days per week

    Nasdaq

    Denver, CO
    1 day ago
  •  ...onsite for a final interview. Our client is seeking a Senior SRE with proven industry experience to join our remote-based Engineering team. Our teams are collaborative and forward thinking; the successful candidate will help shape the operations and support for... 
    Remote work

    Insight Global

    Boca Raton, FL
    4 days ago
  • $174.35k - $210k

     ...Site Reliability Engineer, IBM Corporation, Austin, TX (Up to 80% telecommuting permitted): Analyze business needs, determine problems, and advise on design and solutions. Design, build, test, deploy and maintain well-engineered information systems and ecosystems. Guide... 
    Remote work

    IBM

    Austin, TX
    15 hours ago
  • $166k - $220k

     ...Site Reliability Engineer (SRE) Anduril Industries is a defense technology company with a mission to transform U.S. and allied military capabilities with advanced technology. By bringing the expertise, technology, and business model of the 21st century's most innovative... 
    Full time
    Work experience placement
    Immediate start
    Remote work

    Colorwave Inc

    Costa Mesa, CA
    15 hours ago
  • $60 - $68 per hour

     ...Genesis10 is currently seeking a Site Reliability Engineer for a 12+ month contract position with a Global Financial Institution located in Plano, TX, Pennington, NJ or Charlotte, NC. This is an excellent opportunity to join a newly forming Site Reliability Engineering... 
    Hourly pay
    Permanent employment
    Contract work
    Remote work

    Genesis10

    United States
    2 days ago
  •  ...Senior Site Reliability Engineer Partner with software developers, platform engineers, and IT staff to improve system design, operability, deployment safety, and production support readiness. Define and maintain operational standards, runbooks, support procedures... 
    Work at office
    Remote work

    ARA Brand

    United States
    8 hours ago
  •  ...Boldly.; Deliver Results.; Hire and Develop the Best.; Be Curious and Learn.; Win as a Team. Job Summary We’re looking for a Site Reliability Engineer (SRE) who’s passionate about building resilient, high-performing systems that our customers can depend on every day. In... 
    Local area
    Remote work

    EBSCO Health

    Birmingham, AL
    2 days ago
  • $95k - $171k

     .... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Olympia, WA
    1 day ago
  •  ...is a recognized, award-winning leader in supply chain AI and a FedRAMP® authorized provider to the federal government. Site Reliability Engineer Location: U.S. (Hybrid) This role requires U.S. citizenship and eligibility for a U.S. security clearance. Role Summary... 
    Work at office
    Work from home
    Flexible hours

    Exiger

    Jersey City, NJ
    2 days ago
  • $67 - $70 per hour

     ...Site Reliability Engineer Location: Chandler, Arizona Hybrid Schedule (3 days onsite/2 days remote) Employment Type: 12 mo. Contract Pay Range: $67-70/hr. Role Overview We are seeking a Site Reliability Engineer to join a team responsible for the reliability... 
    Contract work
    Remote work
    Shift work
    Weekend work

    Apex Systems

    Chandler, AZ
    15 hours ago
  • $50 - $53 per hour

     ...Immediate need for a talented Site Reliability Engineer (SRE) This is a 12+ Months contract opportunity with long-term potential and is in Chicago, IL (Hybrid). Please review the job description below and contact me ASAP if you are interested. Job Diva ID... 
    Contract work
    Local area
    Immediate start
    Remote work

    Pyramid Consulting

    United States
    4 days ago
  • Location: Plano, TX (Hybrid)3 days onsite 2 days remote look for nearby Candidates Must have Skills: Need SRE mindset Preferred coming from development background AWS Splunk App Dynamics (good Monitoring background ) Job responsibilities ...
    Remote work

    Apex Informatics

    Plano, TX
    1 day ago
  • $150k - $250k

     ...Site Reliability Engineer role USC or GC only are considered at this time. San Francisco - Local to Bay area only but role is remote and occasion meeting required Latest update, 03/31/2026: The Site Reliability Engineer role is critical for... 
    Work experience placement
    Casual work
    Local area
    Immediate start
    Remote work

    3B Staffing LLC

    San Francisco, CA
    2 days ago
  • $111k - $160k

     ...Join Mizuho as a Site Reliability Engineer! In this role you will play a crucial role in maintaining the reliability, scalability, and overall performance of our production systems. This position collaborates closely with development, operations, and product teams to automate... 
    Work at office
    Local area
    Remote work

    Mizuho

    New York, NY
    2 days ago
  •  ...Job Title Design, automate, deploy, and operate highly reliable cloud systems supporting mission-critical workloads for U....  ...Government customers. This role is centered on DevSecOps and site reliability engineering, with a strong emphasis on deployment automation,... 
    Permanent employment
    Remote work

    Quindar

    United States
    2 days ago
  • $160k - $180k

     ...Senior Site Reliability Engineer Arkestro's Predictive Procurement Platform applies AI, game theory, and behavioral science to enterprise negotiations. It moves teams away from reactive supplier bidding toward data-driven offers that reduce friction and help both buyers... 
    Local area
    Remote work

    Arkestro

    United States
    2 days ago
  • $86.9k - $198k

     ...Site Reliability Engineer, Senior The Opportunity Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if... 
    Full time
    Contract work
    Part time
    Local area
    Remote work

    Booz Allen Hamilton

    East Aurora, NY
    8 hours ago
  • $170k - $290k

     ...Senior Site Reliability Engineer Luma's mission is to build multimodal AI to expand human imagination and capabilities. We believe that multimodality is critical for intelligence. This requires a massive, reliable, and performant GPU infrastructure that pushes the boundaries... 
    Work experience placement
    Remote work

    Luma AI

    United States
    3 days ago
  •  ...significantly reduces costs and improves the critically important 24x7 performance for building owners, developers and tenants. Site Reliability Engineer II The SRE II sits at the intersection of software engineering and platform operations. You will own the reliability,... 
    Remote work

    Kastle Systems

    New York, NY
    2 days ago
  •  ...Site Reliability Engineer II About PROS: PROS, Inc. is the leading offer management provider to the airline industry, helping airlines deliver seamless retail experiences designed to maximize revenue and margin growth. Powered by AI, the PROS Platform enables... 
    Remote work
    Flexible hours

    PROS Holdings, Inc.

    United States
    1 day ago
  •  ...generative AI and cloud-native platforms to advanced release engineering practices, our teams are redefining how financial technology...  ...AI-driven solutions that accelerate development and improve reliability. Your work will directly influence how GM Financial leverages... 
    Full time
    H1b
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours
    2 days per week
    3 days per week

    GMAC Financial Services

    Arlington, TX
    2 days ago
  • $189k - $283.6k

     ...the SRE team, you will proactively and reactively improve the reliability of Block's platform and critical infrastructure. You are metrics...  ...~ A strong desire to perform and grow as an engineer ~5+ years of software development experience Technologies... 
    Full time
    Local area
    Remote work
    Relocation package
    Flexible hours
    Shift work

    Block USA

    New York, NY
    2 days ago
  •  ...expertise standards and connect business - community through highly engaging hacking experiences. The Core Mission of the Site Reliability Engineer (SRE): As a Site Reliability Engineer at Hack The Box, your paramount mission is to empower our Content Engineering team by... 
    Work at office
    Remote work
    Work from home
    Flexible hours

    Hack The Box

    Greece, NY
    15 hours ago
  • $62k - $141k

     ...Job Number: R0238834 Site Reliability Engineer The Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Chantilly, Loudoun County, VA
    1 day ago
  • $76k - $127k

     ...Site Reliability Engineer II Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices... 
    Full time
    Part time
    Worldwide
    Flexible hours

    Dynamic Yield

    O Fallon, MO
    4 days ago
  • $150k - $200k

     ...Site Reliability Engineer at Triomics (W21) $150K - $200K AI Agents for Oncology EHRs Triomics is building the agentic AI layer for oncology electronic health records (EHRs). Cancer hospitals spend billions on highly trained staff manually reading unstructured patient... 
    Full time
    Work at office
    Remote work
    Day shift

    Triomics

    New York, NY
    15 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!