Staff AI Platform Engineer
Supabase
About the Role
We are hiring a Staff AI Platform Engineer to build the execution layer for Supabase's internal AI systems.
Supabase is building an AI-native internal operating system: a common way of working across the company where AI carries a meaningful share of the operational load rather than sitting alongside it as an assistant. We are standing up a new central team to build those systems, enable the teams, and embed AI operations throughout the organization. You are the engineer on that team.
The execution layer is yours. You will build the platform that actually runs agents: an event-triggered queue, a headless model-agnostic runtime, durable state so work survives a restart, a human review gate, atomic rollback, and full logging of every prompt, tool call and decision so any run can be reconstructed. You will build the evaluation layer that makes any of it trustworthy, because an agent that cannot be measured cannot be trusted with anything beyond reading. And you will build the agents themselves, across everything from executive reporting down to a layer of agents that watch the platform and improve it.
This is a governance-heavy environment by design, and that is the interesting part of the problem. Agents are risk-tiered from read-only internal data through to external-facing output, with review depth, evaluation requirements and human approval scaling by tier. Some capabilities are permanently off limits: an agent may read and report, it may carry a human-authored update into a system of record once a human consents, and it may initiate contact within a strict budget, but it may never autonomously write a commitment (an owner, a due date, a status) into a shared work system. Your job is to make that structurally impossible rather than merely forbidden.
You will be the only engineer on this platform. You will close open architecture decisions yourself, own the infrastructure end to end, and instrument the system so it reports its own return.
What You'll Be Responsible For
In this role, you'll:
Ship the agent platform to production. An event-triggered queue, a headless model-agnostic runtime (choosing the runtime is an open decision you will close), durable state that survives a failed run, a human review gate, atomic rollback, and complete run logging in the warehouse.
Own the evaluation layer, and switch on the gate that depends on it. Golden suites with behavioral assertions rather than intuition, judge criteria with a written rubric, safety cases that must pass on every run, and a CI gate that blocks a regression from merging. This is the precondition for every agent that does anything beyond read internal data.
Build and register the agent portfolio. Reporting, drafting, linting, triage and question-answering agents across the executive, team-lead and individual-contributor layers, plus a meta layer that observes the platform and improves it.
Enforce governance in code. Risk tiers the pipeline actually enforces, least-privilege credentials per agent, tool-permission gates, a decision audit log, and autonomy classes where the dangerous class has no code path rather than a warning label. Nothing runs without a registered owner, tier, tool grant and human gate.
Design how the system contacts people. A hard interruption budget per person, message bundling instead of a stream of pings, and a structure that gives something useful before it asks for anything. Adoption depends on this more than on any other single design choice.
Own the platform tooling. The compiler and validator, inventory integrity, and the paths that distribute context and capabilities into the repositories and chat surfaces where work happens.
Compute the operating measures from production data. A pipeline from raw system activity through to a computed maturity grade per team, defensible enough that a team can dispute the result and be answered with the query rather than an opinion.
Instrument the platform's own return. A ledger that logs the work each agent absorbs and computes the monthly figure, so the value of the system is a measurement rather than a claim.
How You'll Think
Recursive Thinking
You build the generator, not the artifact. When you need thirty agents, you do not write thirty agents; you build the inventory, the compiler and the distribution path that makes the thirty-first cost an afternoon, then a meta layer whose job is to watch the platform, find its drift and file the fix.
The evaluation layer is the same move applied to trust. You are not checking whether one agent is correct today. You are building the machine that decides whether every future agent is allowed to ship, which means the suite has to be right in a way the agent does not, and the thing that grades has to be graded too.
Inversion Thinking
You start from the failure and work backward to the design. The rule is that an agent may never autonomously write a commitment into a shared work system. The weak implementation is an instruction in a prompt. The strong one is that the credential in the agent's tool grant physically cannot set an owner, a due date or a status, so no amount of clever input, prompt injection or model error produces the forbidden write. You reach for the second one first.
Same for evaluation. Before writing a suite you enumerate how the agent can be wrong: an update that invents progress that did not happen, one that quotes a private channel into a public digest, one that is accurate and reads as an accusation, one that credits the wrong person. Then the suite is that list, each case caught before the agent ships rather than after it embarrasses someone.
AI-Native Execution
You use agentic tools on real work, in files and repositories, with the same rigor you apply to anything else you ship. You have a setup of your own and can describe it mechanically: what triggers it, what it is allowed to touch, where the human approves, what it logs, and how you found out the one time it went wrong.
That experience is this job, generalized. You are building for async engineers across 40+ countries who will judge the platform by whether it saves them an hour or costs them one, so you are your own first user and your own harshest reviewer.
You Might Be a Good Fit If You
Must Have
Have shipped production LLM agent systems that other people depended on. Not demos, not internal showcases. Systems with operational history, real users and at least one incident you can talk through. Prompt engineering on its own does not clear this bar.
Design evaluations, not spot checks. You build golden sets, write behavioral assertions, define judge rubrics, set pass thresholds and gate CI on the result. You can explain why "we reviewed a bunch of outputs and they looked good" is not evaluation.
Have done deep API work against the systems work actually lives in , and have authored MCP servers. You know the specific failure modes of those APIs, not just that they exist.
Own infrastructure end to end in Python on GCP , with a cloud warehouse and infrastructure as code. You provision, deploy, monitor and roll back your own systems, and you close architecture decisions rather than routing them onward.
Have taste about how software contacts humans. You treat every notification as spending a limited amount of trust.
Strong Signal
Public work in this space. An open-source agent framework, an MCP server, an evaluation harness, or writing on agent reliability that other practitioners cite.
LLM observability and cost instrumentation. Tracing agent runs, attributing spend per run, and building the queries that turn raw logs into a report someone acts on.
You have built an internal platform that non-engineers adopted voluntarily , and can describe what you changed after watching them use it.
What Success Looks Like
The platform is boring. Agents run on a schedule and on events, state survives restarts, failed runs roll back cleanly, and every run can be reconstructed from its log.
Nothing ships unevaluated. Every registered agent has a real suite with safety cases, the gate blocks regressions, and when someone challenges an output the answer is a test case rather than an argument.
The dangerous action is impossible, not discouraged. An audit of any agent's credentials shows it cannot perform the writes it is not allowed to perform, and the audit log makes every consequential decision traceable.
Teams pull the platform instead of being pushed. Agents get adopted because the reports are useful and the contact is rare and well-timed, and each team ends up with at least one workflow that runs automatically.
The platform reports its own value. The work absorbed is measured and published, so the case for expanding it is made with data rather than enthusiasm.
- ...Saviynt is a leader in identity security, delivering an AI-powered platform that governs and secures access to applications, data, and business... ...across the AI Platform. You define the standards ML engineers and scientists build on, and ensure every training signal is...SuggestedFull time
- ...AI Platform Engineer Cornelis Networks delivers high-performance scale-out networking solutions for AI and HPC datacenters. Our differentiated architecture integrates hardware, software, and system-level technologies to maximize the efficiency of GPU, CPU, and accelerator...SuggestedPermanent employmentFull timeRemote workFlexible hours
- ...and Business consulting services. We are in search of a highly motivated candidate to join our talented Team. Job Title: AI Platform Engineer - 50092-1 Location(s): 100% Remote Duration: 6 Months Job Summary: The AI Platform Engineer translates AI...SuggestedWork at officeRemote work
$100k - $150k
...consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This... ...offering tremendous career growth potential. Job Title: AI Platform Engineer Location: 100% Remote (Continental United States)...SuggestedFull timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...Vallen Distribution is building a governed, production-grade AI capability on an Azure-first, Databricks-centered architecture... ...position with broad ownership across the AI layer. The AI Platform Engineer will own the Databricks AI/ML platform layer, drive enterprise...SuggestedImmediate startRemote work
- ...Senior AI Platform Engineer Leidos is seeking a Senior AI Platform Engineer to join our DCSB Artificial Intelligence & Agent Development team supporting the U.S. Army C5ISR Center. In this role, you will lead the technical strategy for building and accrediting a...Remote work
- ...AI Platform Engineer About Tessera Labs Tessera Labs is redefining how enterprises adopt and operationalize Artificial Intelligence. Backed by Foundation Capital and led by a world-class founding team, we build multi-agent AI systems that automate complex business...Remote work
- ...AI Data Scientist Strategic Staffing Solutions (S3) has an opening! AI Data Scientist... ...applications, including prompt engineering, tool usage, and experience with frameworks... ...) before deploying the solution on cloud platform to save us the cloud cost during development...Contract workFor contractorsLocal areaRemote work
- ...Overview: Job Title: AI Platform Engineer Job Location: Hybrid (3 days onsite, 2 days remote), Chicago, IL Travel: Occasional international travel required Job Duration: Long Term Contract Job Summary: We are seeking an AI Platform Engineer to...Long term contractRemote work
$50 - $55 per hour
...Job #5159 Job Title: AI Platform Engineer Client: Direct---Fund Raising Domain Location: Memphis, TN preferred. Remote candidates may be considered if qualified local candidates are not identified. Duration: 18 Months + Extensions (Long...Long term contractPermanent employmentLocal areaImmediate startRemote work- ...Job Title: AI Platform Engineer II Preferred domain Fundraising Location: Memphis, TN (If Local Candidates (Memphis / Nashville Need onsite visit))or REMOTE Duration: 18 Months + Extensions (Long-term contract w/ no end date. Could possibly go perm...Long term contractPermanent employmentLocal areaRemote work
- ...AI Platform Engineer Job Level: Vice President Job Function: IT and Digital Development Location: Charlotte, NC, US Employment Type: Full Time Role Description As a Staff AI-Ops Engineer in the Platform Engineering team, you will play a pivotal role in operationalizing...Full timeWork at officeLocal areaWork from home
- ...role involves building and scaling a production multi-agent AI platform designed to serve thousands of internal users across various... ...Artificial Intelligence (AI), and Generative AI. Qualifications: ~4-6 years of experience in a relevant engineering role....
- ...To build and enhance a centralized AI platform, the full-time Senior AI Platform Engineer will design knowledge representation systems, develop platform services for knowledge management, and collaborate with domain experts to create structured knowledge for AI applications...Full timeRemote workFlexible hours
- ...Data & AI Platform Engineer Executes the implementation and integration of an enterprise data & AI platform (e.g., Palantir, C3 AI). Develops data governance models, ontologies, data pipelines, and applications to ensure interoperability across planning systems....Remote work
- ...Your career matters to us because your passion and excitement will help keep our company moving forward. Senior DevOps AI Platform Engineer Are you ready to take your career to the next level with a rapidly growing global company? As a Senior DevOps Platform Engineer...Local areaImmediate start
- ...scale creates a meaningful opportunity for practical enterprise AI. WAI is investing in AI, automation, data, and digital... ...customers more effectively. About the Role The Director, AI Platform Engineer is a senior technical leadership role responsible for building...For contractorsWork at officeRemote work
$100k
...About Givzey / Version2.ai Join the Future of Fundraising at... ...traditional AI tools that simply make staff more efficient, VEOs expand... ...In just three years, Givzey's platform has already helped... ...developer tooling to make sure engineers spend their time building product...Full timeLocal areaDay shift- ...About the Role Modern engineering teams invest a significant amount of time and money on simulation. Before anything gets built, teams... ...multiphysics simulations, scientific visualization, or engineering data platforms. You understand the kinds of problems engineers are trying to...Permanent employmentFull timeRemote work
$114.6k - $252.1k
Job Title: Staff Computer Vision AI/ML Engineer Job Category: Science Time Type: Full time Minimum Clearance Required to Start: TS/SCI Employee Type: Regular Percentage of Travel Required: None Type of Travel: None Anticipated Posting End: 12/31/20...Full timeContract workWork experience placementRemote workFlexible hours- ...LTS is seeking an AI Platform and Harness Engineer to develop and maintain the infrastructure, tooling, and evaluation frameworks that power enterprise AI solutions. This role is responsible for building the AI platform and reusable "AI harnesses" that enable Large...Full time
- ...building the telemetry infrastructure for the AI era. At Cribl, we partner with IT and... ...and infrastructure reality. As the AI Platform for Telemetry, we give customers the choice... ...Love This Role The Senior Legal AI Platform Engineer is the builder and architect inside that...Full timeContract workRemote work
- ...controls, and construction/facilities. We are a Woman-owned, SBA Certified 8(a), and HUBZone company. Job Overview: The AI Platform Engineer builds, hardens, and helps operate the enterprise GenAI/agentic platform that enables teams to safely develop and run AI...Full timeRemote workFlexible hours
- ...Job Description Job Description AI Platform Engineer HED is hiring an AI Platform Engineer to build durable, production-grade AI agent systems that integrate with governed data. About HED We are a team that is full of ideas, experience, creativity, passionate...Work at officeWork from home
- ...global team is a diverse group of experienced engineers, traders, and brokerage professionals who... ...apply. Your Role Alpaca is building an AI Enablement function to move the company... ...-wide productivity gains. As Senior AI Platform Engineer, you will own the technical capability...Full time
- ...technologies to unlock the power of data. We have an amazing team of 25,200 people in 32 countries. As a Senior Director of AI Platform Engineering, you will oversee a layer of our internal platform that makes AI real for an engineering organization of 5,000 developers....Full timeWork at officeRemote workFlexible hours
$221.2k - $387.1k
...It all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker,... ...focus on meaningful work. Today, ServiceNow is the AI control tower for business reinvention. Our ServiceNow AI platform brings together any AI, any data, and any workflow...Full timeWork at officeImmediate startRemote workFlexible hours- About the Role We are hiring a Staff AI Platform Engineer to build the execution layer for Supabase's internal AI systems. Supabase is building an AI-native internal operating system: a common way of working across the company where AI carries a meaningful share of the...Full time
- ...be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking a AWS Data & AI Platform Engineer to join our team in Bengaluru, Hyderabad, Chennai, Karnātaka (IN-KA), India (IN). We are seeking an experienced AI Platform...Work at officeRemote workFlexible hours
- ...assess, design, build, and migrate to the Microsoft Azure platform, and increasingly helps those same enterprises design, build, and run AI agent platforms on top of it. This role sits in the AIC Platform Engineering practice, alongside consultants who lead Azure...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff AI Platform Engineer. Be the first to apply!
- project engineer assistant project manager Remote
- engineering aide Remote
- staff design engineer Remote
- senior staff engineer Remote
- assistant engineer Remote
- software engineer staff Remote
- staff engineer Remote
- staff security engineer Remote
- assistant engineering manager Remote
- senior staff systems engineer Remote




