Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff AI Platform Engineer

Full-time

Supabase

About the Role

We are hiring a Staff AI Platform Engineer to build the execution layer for Supabase's internal AI systems.

Supabase is building an AI-native internal operating system: a common way of working across the company where AI carries a meaningful share of the operational load rather than sitting alongside it as an assistant. We are standing up a new central team to build those systems, enable the teams, and embed AI operations throughout the organization. You are the engineer on that team.

The execution layer is yours. You will build the platform that actually runs agents: an event-triggered queue, a headless model-agnostic runtime, durable state so work survives a restart, a human review gate, atomic rollback, and full logging of every prompt, tool call and decision so any run can be reconstructed. You will build the evaluation layer that makes any of it trustworthy, because an agent that cannot be measured cannot be trusted with anything beyond reading. And you will build the agents themselves, across everything from executive reporting down to a layer of agents that watch the platform and improve it.

This is a governance-heavy environment by design, and that is the interesting part of the problem. Agents are risk-tiered from read-only internal data through to external-facing output, with review depth, evaluation requirements and human approval scaling by tier. Some capabilities are permanently off limits: an agent may read and report, it may carry a human-authored update into a system of record once a human consents, and it may initiate contact within a strict budget, but it may never autonomously write a commitment (an owner, a due date, a status) into a shared work system. Your job is to make that structurally impossible rather than merely forbidden.

You will be the only engineer on this platform. You will close open architecture decisions yourself, own the infrastructure end to end, and instrument the system so it reports its own return.

What You'll Be Responsible For

In this role, you'll:

  • Ship the agent platform to production. An event-triggered queue, a headless model-agnostic runtime (choosing the runtime is an open decision you will close), durable state that survives a failed run, a human review gate, atomic rollback, and complete run logging in the warehouse.

  • Own the evaluation layer, and switch on the gate that depends on it. Golden suites with behavioral assertions rather than intuition, judge criteria with a written rubric, safety cases that must pass on every run, and a CI gate that blocks a regression from merging. This is the precondition for every agent that does anything beyond read internal data.

  • Build and register the agent portfolio. Reporting, drafting, linting, triage and question-answering agents across the executive, team-lead and individual-contributor layers, plus a meta layer that observes the platform and improves it.

  • Enforce governance in code. Risk tiers the pipeline actually enforces, least-privilege credentials per agent, tool-permission gates, a decision audit log, and autonomy classes where the dangerous class has no code path rather than a warning label. Nothing runs without a registered owner, tier, tool grant and human gate.

  • Design how the system contacts people. A hard interruption budget per person, message bundling instead of a stream of pings, and a structure that gives something useful before it asks for anything. Adoption depends on this more than on any other single design choice.

  • Own the platform tooling. The compiler and validator, inventory integrity, and the paths that distribute context and capabilities into the repositories and chat surfaces where work happens.

  • Compute the operating measures from production data. A pipeline from raw system activity through to a computed maturity grade per team, defensible enough that a team can dispute the result and be answered with the query rather than an opinion.

  • Instrument the platform's own return. A ledger that logs the work each agent absorbs and computes the monthly figure, so the value of the system is a measurement rather than a claim.

How You'll Think

Recursive Thinking

You build the generator, not the artifact. When you need thirty agents, you do not write thirty agents; you build the inventory, the compiler and the distribution path that makes the thirty-first cost an afternoon, then a meta layer whose job is to watch the platform, find its drift and file the fix.

The evaluation layer is the same move applied to trust. You are not checking whether one agent is correct today. You are building the machine that decides whether every future agent is allowed to ship, which means the suite has to be right in a way the agent does not, and the thing that grades has to be graded too.

Inversion Thinking

You start from the failure and work backward to the design. The rule is that an agent may never autonomously write a commitment into a shared work system. The weak implementation is an instruction in a prompt. The strong one is that the credential in the agent's tool grant physically cannot set an owner, a due date or a status, so no amount of clever input, prompt injection or model error produces the forbidden write. You reach for the second one first.

Same for evaluation. Before writing a suite you enumerate how the agent can be wrong: an update that invents progress that did not happen, one that quotes a private channel into a public digest, one that is accurate and reads as an accusation, one that credits the wrong person. Then the suite is that list, each case caught before the agent ships rather than after it embarrasses someone.

AI-Native Execution

You use agentic tools on real work, in files and repositories, with the same rigor you apply to anything else you ship. You have a setup of your own and can describe it mechanically: what triggers it, what it is allowed to touch, where the human approves, what it logs, and how you found out the one time it went wrong.

That experience is this job, generalized. You are building for async engineers across 40+ countries who will judge the platform by whether it saves them an hour or costs them one, so you are your own first user and your own harshest reviewer.

You Might Be a Good Fit If You

Must Have

  • Have shipped production LLM agent systems that other people depended on. Not demos, not internal showcases. Systems with operational history, real users and at least one incident you can talk through. Prompt engineering on its own does not clear this bar.

  • Design evaluations, not spot checks. You build golden sets, write behavioral assertions, define judge rubrics, set pass thresholds and gate CI on the result. You can explain why "we reviewed a bunch of outputs and they looked good" is not evaluation.

  • Have done deep API work against the systems work actually lives in , and have authored MCP servers. You know the specific failure modes of those APIs, not just that they exist.

  • Own infrastructure end to end in Python on GCP , with a cloud warehouse and infrastructure as code. You provision, deploy, monitor and roll back your own systems, and you close architecture decisions rather than routing them onward.

  • Have taste about how software contacts humans. You treat every notification as spending a limited amount of trust.

Strong Signal

  • Public work in this space. An open-source agent framework, an MCP server, an evaluation harness, or writing on agent reliability that other practitioners cite.

  • LLM observability and cost instrumentation. Tracing agent runs, attributing spend per run, and building the queries that turn raw logs into a report someone acts on.

  • You have built an internal platform that non-engineers adopted voluntarily , and can describe what you changed after watching them use it.

What Success Looks Like

  • The platform is boring. Agents run on a schedule and on events, state survives restarts, failed runs roll back cleanly, and every run can be reconstructed from its log.

  • Nothing ships unevaluated. Every registered agent has a real suite with safety cases, the gate blocks regressions, and when someone challenges an output the answer is a test case rather than an argument.

  • The dangerous action is impossible, not discouraged. An audit of any agent's credentials shows it cannot perform the writes it is not allowed to perform, and the audit log makes every consequential decision traceable.

  • Teams pull the platform instead of being pushed. Agents get adopted because the reports are useful and the contact is rare and well-timed, and each team ends up with at least one workflow that runs automatically.

  • The platform reports its own value. The work absorbed is measured and published, so the case for expanding it is made with data rather than enthusiasm.

Vacancy posted 25 days ago
Similar jobs that could be interesting for youBased on the Staff AI Platform Engineer in Remote vacancy
  • $129k - $210k

    Role Description The Staff Data and AI Platform Engineer at Workiva serves as the technical authority for the Enterprise data platform. You'll own design, reliability, security, and cost efficiency of account-level infrastructure while enabling domain teams to build and... 
    Suggested
    Permanent employment
    Full time
    Remote work

    Workiva Inc.

    Remote
    20 days ago
  • $117.2k - $175.8k

    Sr Data Engineer - GE07BEWe’re determined to make a difference and are proud to be an insurance company that goes well beyond coverages...  ...shape the future. The Hartford seeks energetic and passionate AI Platform Engineers to build AI Operations (AIOps, MLOps, FMOps, LLMOps)... 
    Suggested
    Full time
    Temporary work
    Work at office
    Remote work
    3 days per week

    The Hartford Financial Services Group

    Columbus, OH
    1 day ago
  •  ...want to be part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a Gen AI - Platform Cloud Engineer to join our team in Concord, California (US-CA), United States (US).In this role, you will: Lead and design the platform... 
    Suggested
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Concord, NH
    3 days ago
  • AI Platform EngineerHED is hiring an AI Platform Engineer to build durable, production-grade AI agent systems that integrate with governed data. About HEDWe are a team that is full of ideas, experience, creativity, passionate opinions, insatiable curiosity, uncompromising... 
    Suggested
    Work at office
    Work from home

    HED

    Boston, MA
    1 day ago
  • $99k - $225k

    AWS Agentic AI Platform EngineerThe Opportunity:As an experienced AI/ML engineer, you know that machine learning and generative AI are critical to understanding and processing large, complex datasets. Your ability to apply statistical learning, large language models, retrieval... 
    Suggested
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    San Antonio, TX
    1 day ago
  • Notable is the leading healthcare AI platform for transforming workforce productivity. Health...  ..., scaling growth without hiring more staff.We are on a mission to improve the lives...  ...for healthcare. As a Senior AI Platform Engineer, you will design, build, and maintain LLM... 
    Full time
    Temporary work
    Work at office
    Remote work
    3 days per week

    Notable

    San Mateo, CA
    3 days ago
  •  ...to entry, making the right connections, and building confidence through expert guidance.Realtor.com is looking for a Senior AI Platform Engineer to architect, build, and continuously improve the internal AI tooling foundation that powers safe, scalable AI adoption across... 
    Work at office
    Local area

    Realtor.com

    Austin, TX
    19 hours ago
  • We are seeking a Full Stack AI Platform Engineer to join our Data Engineering, AI & ML Platform team. This role is central to designing, building, and scaling the enterprise AI/ML platform that powers intelligent automation across a global portfolio.As a Full Stack AI... 
    Permanent employment
    Temporary work
    Local area
    Flexible hours

    Honeywell

    Atlanta, GA
    19 hours ago
  • $156.06k - $211.14k

    Afresh, the AI platform for grocery, began by tackling the most complex problem in the industry: fresh, and has evolved into the core AI...  ...commodity. The knowledge you feed them is not.As a Senior AI Platform Engineer, you build the AI and data platform that powers Afresh's... 
    Full time
    Live in
    Work at office
    Local area
    Remote work
    Work from home
    Home office
    Flexible hours
    3 days per week

    Afresh

    San Francisco, CA
    1 day ago
  • $102.68k - $175.13k

     ...to be part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a Cloud Agentic AI Platform Engineer / Architect to join our team in Charlotte, North Carolina (US-NC), United States (US).We are seeking a hands-on Cloud... 
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours

    NTT

    Charlotte, NC
    1 day ago
  • $204k - $337k

     ...products and services that help people, businesses and governments realize their greatest potential.Title and SummaryPrincipal AI Platform Engineer - AI Center of ExcellenceWho is Mastercard?Mastercard is a global technology company in the payments industry. Our mission... 
    Full time
    Part time
    Worldwide
    Flexible hours

    MasterCard

    San Carlos, CA
    19 hours ago
  •  ...want to be part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a AWS Data & AI Platform Engineer to join our team in Charlotte, North Carolina (US-NC), United States (US).We are seeking an experienced AI Platform Engineer... 
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Charlotte, NC
    3 days ago
  • $87.95k - $162.88k

     ...Senior AI DevOps Engineer (AI Ops / Platform Engineering)NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently... 
    Temporary work
    Work at office
    Remote work
    Flexible hours
    3 days per week

    NTT DATA

    Atlanta, GA
    2 days ago
  • $132.4k - $251.6k

     ...of aerospace and defense.Role SummaryThe AI Accelerator works across RTX to mature early...  ...patterns that strengthen the AI Factory platform.Qualifications You Must HaveTypically requires a degree in Science, Technology, Engineering or Mathematics (STEM) and a minimum of 10... 
    Permanent employment
    Temporary work
    Work experience placement
    Work at office
    Local area
    Remote work
    Flexible hours

    Raytheon

    Cambridge, MA
    3 days ago
  • $135k - $175k

     ...collegiality, and empowerment.Role Overviewbpx energy is building an enterprise AI capability that can scale safely and deliver real operational value. The Senior AI/ML Platform Engineer will help build and operate the technical foundation required to move AI/ML capabilities... 
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Worldwide
    Relocation
    Flexible hours

    BP

    Denver, CO
    1 day ago
  • $180k - $200k

     ...Senior AI Platform Engineer Remote, USA At Twin Health, we empower people to improve and prevent chronic metabolic diseases, like type 2 diabetes and obesity, with a new standard of care. Twin Health is the only company applying AI Digital Twin technology exclusively... 
    Remote work
    Flexible hours

    Twin Health

    United States
    2 days ago
  •  ...AI Platform Engineer This role involves building and operating the foundational platform for scalable, secure AI capabilities. Located remotely in India, this contract position focuses on developing services, pipelines, and SDKs for machine learning and large language... 
    Contract work
    Work experience placement
    Remote work

    Mitchell Martin

    United States
    4 days ago
  • $190k - $230k

     ...Continuous Threat Exposure Management (CTEM). The HackerOne Platform unites agentic AI solutions with the ingenuity of the world's largest...  ..., inclusion, respect, and accountability. Senior Software Engineer, AI Platform Location: Seattle, WA; Austin, TX; Boston, MA... 
    Apprenticeship
    Work at office
    Local area
    Remote work
    Flexible hours
    Shift work
    1 day per week

    HackerOne

    Washington DC
    19 hours ago
  •  ...To build the AI backbone of a fast-moving company, the full-time Senior AI Platform Engineer will design and implement agents, connectors, and automation systems that enhance productivity across the organization, working remotely and collaborating with various teams to... 
    Full time
    Remote work

    Virtual Vocations Inc

    United States
    19 hours ago
  •  ...AI Platform Engineer We are looking for a forward-thinking AI Platform Engineer to join our engineering team. In this role, you will bridge the gap between traditional infrastructure excellence and the next generation of AI-driven development. You will not only... 
    Remote work

    Localytics

    United States
    3 days ago
  •  ...Job Title: AI Platform Engineer Location: Remote Duration: Long Interview: Video The AI Platform Engineer translates AI reference architectures into secure, scalable cloud configurations that enable development teams to build and... 
    Work at office
    Remote work

    Anveta

    United States
    4 days ago
  •  ...AI Platform Engineer Tessera Labs is redefining how enterprises adopt and operationalize Artificial Intelligence. Backed by Foundation Capital and led by a world-class founding team, we build multi-agent AI systems that automate complex business workflows across platforms... 
    Remote work

    Tessera Labs

    United States
    3 days ago
  •  ...Senior Software Engineer - AI Platform NegotiateAI (Menlo Ventures-backed) About NegotiateAI NegotiateAI is building an enterprise AI platform to modernize procurement for manufacturing and industrial companies. Procurement today is fragmented, manual, and opaque... 
    Remote work

    NegotiateAI

    United States
    1 day ago
  •  ...AI & Platform Software EngineerLocation: Remote within the USA, or onsite in Buffalo, NY / Wilmington, DE (client preference for candidates...  ...Developer Experience (SDLC automation and tooling, quality engineering and test data management, and site reliability engineering).... 
    Contract work
    For contractors
    Work at office
    Remote work

    Glint Tech Solutions LLC

    Buffalo, NY
    1 day ago
  •  ...AI Platform Engineer Collinson is the global, privately-owned company dedicated to helping the world to travel with ease and confidence. The group offers a unique blend of industry and sector specialists who together provide market-leading airport experiences, loyalty... 
    Remote work
    Worldwide
    Sleeping nights

    COLLINSON, INC.

    United States
    19 hours ago
  • $500 per month

     ...Senior AI Platform Engineer Remote - North America (EST) Alpaca is a US-headquartered, global leader in agent-first brokerage infrastructure for stocks, ETFs, options, crypto, fixed income, 24/5 trading, and more. Amongst our subsidiaries, Alpaca is a licensed financial... 
    Remote work
    Home office

    Alpaca

    United States
    4 days ago
  •  ...One Platform, A Whole World Of Opportunity The best jobs have always clustered in a handful of the world's wealthiest...  ...to be based within UTC−6 to UTC+3. Oyster's Data Engineering team is building the foundational AI platform layer that will enable teams across Oyster... 
    Work at office
    Remote work
    Work from home
    Home office
    Flexible hours

    Oyster

    United States
    1 day ago
  •  ...AI Platform Engineer (Gen AI) This is a remote position. Job Title: AI Platform Engineer (Gen AI) Experience: 8 Years Location: US Job Summary: We are looking for an experienced Platform Engineer specializing in Generative AI to design, build, and operate scalable... 
    Remote work

    Veltris

    United States
    1 day ago
  •  ...AI Platform Engineer This is a remote position. Role Overview We are looking for an AI Platform Engineer with strong experience in Azure and Snowflake to design, implement, and govern our next-generation AI capabilities. You will be responsible for enabling and operating... 
    Remote work

    Kanini

    United States
    19 hours ago
  •  ...Senior AI Platform Engineer Brazil (Remote) Your wellbeing, our mission. Join a company shaping a healthier world. At Wellhub we're revolutionizing workplace wellness. Our platform connects employees worldwide to the best partners for fitness, mindfulness, therapy... 
    Part time
    Remote work
    Worldwide
    Home office
    Flexible hours

    Wellhub Inc.

    United States
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff AI Platform Engineer. Be the first to apply!