Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff AI Platform Engineer: Agent & Retrieval Infrastructure

$180k - $220k
Full-time

Bedrock Ocean Exploration

About Bedrock Ocean

Bedrock Ocean builds and operates autonomous underwater vehicles (AUVs) that collect georeferenced ocean-floor data at commercial scale. We deliver bathymetric and imagery data products to customers through our own platform, and we're scaling toward continuous, around-the-clock data collection campaigns spanning months at a time.

We are building AI agents on Amazon Bedrock to support our ocean data, internal operations, and customer platform. This role owns that architecture.

(One note on names. Amazon Bedrock is the AWS service. Bedrock Ocean is us. They are unrelated, and we are aware it is confusing.)

The Role

We are looking for a Staff Platform Engineer to lead our AI architecture. This role goes beyond building agents on existing platforms; you will create the infrastructure itself, including the orchestration layer, the data and retrieval pipeline, and the security model required to work with production data. You will also build the tools and abstractions that allow our engineering team to implement AI features independently.

This position combines software engineering, data engineering, and infrastructure operations. You will manage the full lifecycle of our Amazon Bedrock implementation, from initial data chunking to IAM access controls. While some of our data pipelines are already in place, they will require significant expansion, and others will need to be built from scratch.

Security is core to this role, not an afterthought. Because agents with tool access represent a new kind of system actor, you will define their operational boundaries, including what they can access, the actions they can perform autonomously, and the monitoring required to detect issues.

Our roadmap prioritizes internal engineering and operational systems first to ensure a fast feedback loop, followed by our ocean and survey data products. Customer-facing retrieval is the final, high-stakes phase. You will play a key role in defining this sequence. While you will not be responsible for the core data transport design (store-and-forward or hub-and-spoke), you will work closely with that team to ensure it meets our retrieval and data freshness requirements.

What You'll Do

  • Architect Agent Orchestration: Design the Amazon Bedrock integration, including agent and action group configuration, backend APIs, model access, throughput, and cross-environment deployment.

  • Manage Retrieval Data Plane: Own the end-to-end retrieval pipeline from ingestion and chunking to embedding and storage in Amazon OpenSearch Serverless. Focus on optimizing for index design, cost, and capacity.

  • Extend Data Pipelines: Adapt ingestion pipelines for internal knowledge, ocean data, and customer platforms, addressing challenges specific to geospatial and large-binary datasets.

  • Secure AI Infrastructure: Implement robust security including Bedrock Guardrails, VPC and PrivateLink network boundaries, least-privilege IAM, and audit trails to ensure data isolation.

  • Define Agent Governance: Build the mechanisms to enforce approval boundaries for autonomous actions, ensuring agents are safe and monitored.

  • Establish LLMOps & Observability: Implement comprehensive monitoring for tracing, tool calls, and retrieval performance, using CloudWatch and LLM-specific tools like Langfuse or Phoenix.

  • Build Evaluation Frameworks: Create the infrastructure to run automated evaluations, track results, and manage release gates for model accuracy.

  • Enable Engineering Productivity: Provide the team with abstraction layers, SDKs, and self-service environments that allow engineers to ship AI features independently.

  • Operational Excellence: Manage the environment as code across all stages, ensuring deployment safety and participating in incident reviews.

What We're Looking For

  • 8+ years in software and infrastructure engineering, including deep production backend experience (Python or TypeScript preferred, Go fine) and staff-level ownership of technical direction.

  • Hands-on experience standing up Amazon Bedrock in production: agents, knowledge bases, guardrails, model access, and the throughput and quota decisions that come with them.

  • Containerized service deployment on ECS, EKS, or Lambda, with CI/CD you have owned rather than inherited. The models are managed, but the backend APIs, tool endpoints, and ingestion jobs still run somewhere real.

  • Practical RAG and vector search experience: embeddings, chunking strategies, semantic search quality, and operating a managed vector database (OpenSearch Serverless, Pinecone, pgvector, or similar) at production scale and cost.

  • Real data engineering: you have built or substantially extended ingestion pipelines over messy, heterogeneous, unstructured sources, and you think about freshness and correctness as SLAs rather than afterthoughts.

  • Strong AWS ecosystem expertise: IAM roles and least privilege for machine identities, VPC networking and PrivateLink, Lambda, S3, KMS, CloudWatch, and provisioning safely through infrastructure as code (Terraform, CDK, or CloudFormation).

  • Production LLM exposure: you have moved LLM features or autonomous agents past the prototype stage into environments other people depend on.

  • A working point of view on securing agentic systems: scoping tool permissions, prompt injection and exfiltration risk, sensitive data handling in retrieval, and where a human belongs in the loop.

  • Experience designing developer-facing APIs, SDKs, or platform services with an API-first mindset. The interface is the product for the engineers who consume it.

  • Experience building and operating multi-tenant services, with isolation guarantees that hold when the data belongs to customers rather than to us.

  • Platform instinct: you build the abstraction other engineers stand on, and you measure yourself by what they ship rather than by what you ship directly.

  • Demonstrated technical leadership and system design judgment at staff level: you have driven an architectural direction across pods or teams you do not manage, and made it stick through influence rather than authority.

  • A pragmatic builder's bias. You reach for boring, fully managed infrastructure before complex self-hosted alternatives, and you can tell the difference between the two in an architecture review.

  • Comfort wearing several hats on a small team, and the discipline to write things down so the system runs without you.

Nice to Have

  • Experience deploying LLM evaluations to measure accuracy over time and treating eval results as a release gate.

  • Involvement in AI red-teaming or the AI security community.

  • Experience with GraphRAG or knowledge graphs.

  • Experience running retrieval over geospatial, scientific, or large-binary datasets.

  • Experience moving data across intermittent or unreliable links: store-and-forward, hub-and-spoke topologies, offload from disconnected or edge systems, and reconciliation once a link comes back.

  • Compliance experience such as SOC 2, or handling government or defense customer data.

  • Background supporting data platforms, autonomous systems, or field operations.

Not a Fit If

  • Your AI work has been prototypes and notebooks rather than systems other people depend on in production.

  • You want to be a model researcher, a prompt engineer, or to spend your time fine-tuning models. This role owns the platform underneath agents and partners closely with the people building them.

  • You treat security as a gate at the end of a project rather than something designed from the start.

  • You would rather self-host and build from scratch than adopt a managed service that already works. We are optimizing for a small team shipping, not for architectural purity.

  • You want a mature platform team and a narrow, well-bounded scope. This is an early build with a lot of surface area and few existing answers.

Why This Role Matters

At Bedrock Ocean, our mission is to make the ocean transparent. We are building more than just a survey service; we are creating a source of deep ocean intelligence that grows with every mission.

To realize this, we need to make our data accessible and actionable. This role is about building the infrastructure that allows our teams and eventually our customers to directly query and learn from our findings. Because this data is strategically sensitive, security is woven into the foundation of your work, not added on later.

Ultimately, an AI agent with access to our production data is a powerful tool, but it requires careful design to be a success. You will build a secure, reliable platform that ensures these systems empower our mission safely and effectively.

The base compensation for this role is expected to be $180,000- $220,000 annually plus equity.

Bedrock Ocean is an equal opportunity employer.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff AI Platform Engineer: Agent & Retrieval Infrastructure in Remote vacancy
  • $120k - $220k

     ...Content Intelligence platform shaping the future...  ...by advanced AI, recommendation systems...  ...: building the infrastructure layer for content...  ...We're building the agent platform that powers...  ...Agent Platform engineer to own this layer...  ...engineering layer — retrieval, ranking, compression... 
    Suggested
    Full time
    Local area
    Work from home

    News Break

    Remote
    1 day ago
  •  ...ShiftKey ShiftKey is a platform that is disrupting...  ...toward being an AI-native company. We are...  ...role exists. The Staff AI Engineer owns ShiftKey’s AI infrastructure layer: the organizational...  ...and operating AI agents in the cloud, and the retrieval layer that powers... 
    Suggested
    Full time
    Work experience placement
    Work at office
    Remote work
    Flexible hours
    Shift work

    Shift Key

    Remote
    1 day ago
  • $84.4k - $156.8k

     ...secure, and maintain AI application environments...  ...within Oracle Cloud Infrastructure (OCI).Adapt existing containerized...  .../Oracle Kubernetes Engine (OKE), OCI Container...  ....Integrate platform deployment with source...  ...AI, OCI Generative AI Agents, OCI Data Science, or... 
    Suggested
    Minimum wage
    Full time
    Contract work
    Flexible hours

    DXC Technology

    Arlington, VA
    2 days ago
  •  ...AI Infrastructure Engineer Redstone Arsenal/Huntsville, AL At IPT Associates...  ...solutions involving LLMs, retrieval-augmented generation, tool...  ...validation. Collaborate with platform engineers, product leads,...  ...evaluation frameworks, and agent testing harnesses. -... 
    Suggested
    Full time

    Interactive Process Technology Llc

    Remote
    1 day ago
  •  ...that. Our multimodal AI models understand...  ...a query, retrieves, segments, and reasons...  ...product. Built for agents, not just people....  ...production-grade infrastructure that scales to millions...  ...a senior AI engineer to own the integration...  ...that makes the platform trustworthy under... 
    Suggested
    Full time
    Live in
    Work at office
    Local area
    Remote work
    Visa sponsorship
    Flexible hours

    Twelve Labs

    Remote
    1 day ago
  • $157k - $235k

     ...other digital services.Snap Engineering teams build fun and...  ...Engineer to join the Content Retrieval Platform team, part of the Content ML...  ...’ll do:Design and optimize infrastructure systems for machine learning...  ...ensure fast and efficient AI model servingBuild infrastructure... 
    Full time
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    3 days ago
  •  ...the world's best data and AI infrastructure platform so our customers can use deep...  ...The Mission Databricks agents are only as good as the context they can retrieve. Whether an agent is...  ...We are hiring a  Senior Staff Applied AI Engineer to own context retrieval... 
    Full time
    Work at office
    Immediate start

    Databricks

    Remote
    1 day ago
  •  ...what is delivered. Our AI-native platform, Air Enterprise...  ...experienced Senior AI Engineer specializing in AI security...  ...adoption of our AI agent, Ace. As agents gain...  ...failures, and build the infrastructure necessary to deploy...  ..., tool use, retrieval, and multi-step agent... 
    Full time
    Work at office
    Remote work

    Air Company

    United States
    1 day ago
  • $140k - $165k

     ...electronic devices and IT infrastructure, enabling enhanced...  ...Build foundational AI infrastructure that powers...  ...seeking a hands-on AI Engineer to design, deploy, and...  ..., and architecting AI agents that operate...  ...including document chunking, retrieval strategies (hybrid, re... 
    Full time

    SK Hynix Memory Solutions America Inc.

    Remote
    1 day ago
  • $99k - $225k

    AWS Agentic AI Platform EngineerThe Opportunity:As an experienced AI/ML engineer, you know that machine learning...  ...language models, retrieval techniques, and agentic...  ...retrieval architectures, and agent orchestration...  ...containerization, and infrastructure as codeAbility to translate... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Austin, TX
    3 days ago
  • We are seeking a Full Stack AI Platform Engineer to join our Data...  ...pipelines and knowledge graph infrastructure, through LLM orchestration...  ...orchestration layers for multi-agent AI systems.Develop and integrate...  ...training or real-time retrieval (RAG) on edge devicesAbility... 
    Permanent employment
    Temporary work
    Local area
    Flexible hours

    Honeywell

    Atlanta, GA
    1 day ago
  • $180k - $225k

    AI Infrastructure Engineer - Agent Sandbox PlatformAs a Software Engineer on the AI Infrastructure team, you'll help build and evolve our agent sandboxing platform — the secure, high-performance code execution layer powering our agentic workflows, deployed across both internal... 
    Full time
    Immediate start
    Remote work

    Scale AI

    New York, NY
    1 day ago
  •  ...scaling reusable AI-driven workflows,...  ...models and autonomous agents. You will build...  ...Build shared APIs and platform components used broadly across engineering teams. Key...  ...systems or platform infrastructure. Practical experience...  ...Familiarity with retrieval systems,... 
    Full time

    Calliere

    Remote
    1 day ago
  • $207k - $301k

     ...coach a distributed team of engineers.Facilitate alignment and clarity...  ...and developing large-scale infrastructure, distributed systems or...  ...areas, including information retrieval, distributed computing, large...  ...the next generation of Google platforms, we make Google's product... 

    Google

    Sunnyvale, CA
    9 hours ago
  • $262k - $364k

     ...coach a distributed engineering team, fostering...  ...various NICs and platforms, tuning the...  ...RDMA, storage, and AI/ML, staying ahead...  ...Machine Learning Infrastructure.Google's software...  ...including information retrieval, distributed...  ...GDS).As a Senior Staff Software Engineer... 
    Remote work
    Worldwide

    Google

    Sunnyvale, CA
    9 hours ago
  •  ...Reference26-01572Principal AI Platform EngineerJob Number: 26-01...  ...a Principal AI Platform Engineer for our client in San...  ...and operates AWS Bedrock infrastructure including Agents, Knowledge Bases, Guardrails...  ...applying appropriate grounding, retrieval, and data access patterns... 
    Contract work
    Work experience placement

    Eclaro International

    San Diego, CA
    3 days ago
  • $200k - $275k

     ...Labs provides AI-powered intelligence...  ...crime. TRM's platforms enable...  ...secure. The AI Engineering Team is...  ...high-performance infrastructure, and operational...  ...in the LLM and agent space - including...  ...use, context retrieval, memory, and...  ...sharing. Staff Engineer: Drives... 
    Local area
    Remote work
    Worldwide

    TRM

    United States
    2 days ago
  •  ...AI Platform Engineer Cornelis Networks delivers high-performance scale-...  ...engineering, developer platforms, infrastructure, workflow automation, and...  ...a growing workforce of AI agents that automate meaningful...  ...assistant to generate text. Retrieval-Augmented Generation:... 
    Permanent employment
    Full time
    Remote work
    Flexible hours

    Cornelis Networks

    United States
    3 days ago
  •  ...AI Data Scientist Strategic Staffing...  ...search, and agent orchestration for...  ...including prompt engineering, tool usage, and...  ..., or AutoGen. Retrieval Quality, Evaluation...  ...Databases & Retrieval Infrastructure – Hands-on...  ...solution on cloud platform to save us the cloud... 
    Contract work
    For contractors
    Local area
    Remote work

    Leading Utilities Organization

    United States
    9 hours ago
  • POSITION SUMMARY:The Senior Platform & Infrastructure Engineer is the principal technical...  ...building, and operationalizing AI and agentic automation...  ...initiatives, including multi-agent AI architectures, LLM...  ...vector store design, RAG (Retrieval-Augmented Generation) pipelines... 
    Shift work

    Nexus Health Systems

    Houston, TX
    3 days ago
  •  ...AI Platform Engineer This role involves building and operating the foundational platform for...  ...requires experience with cloud services, infrastructure as code, and machine learning...  ...-ready pipelines, including document retrieval and vector databases. Develop orchestration... 
    Contract work
    Work experience placement
    Remote work

    Mitchell Martin

    United States
    5 days ago
  •  ...Job Title: AI Platform Engineer II Preferred domain Fundraising Location: Memphis, TN...  ...AI solutions, focusing on either AI infrastructure (LLM platform, backend systems, cloud...  ...development and maintenance of LLM gateways, retrieval-augmented generation platforms, or AI... 
    Long term contract
    Permanent employment
    Local area
    Remote work

    3B Staffing LLC

    Memphis, TN
    1 day ago
  •  ...AI Platform Engineer The AI Platform Engineer is responsible for designing, building, and operationalizing...  ...platform capabilities across cloud infrastructure, enterprise systems, and customer-...  ...AI APIs, orchestration services, retrieval-augmented generation (RAG) pipelines,... 
    Remote work

    USALCO

    United States
    4 days ago
  •  ...Role Description As a Staff AI-Ops Engineer in the Platform Engineering team, you...  ...models, pipelines, and agents that run on it. You will...  ..., Databricks, and other infrastructure providers to build and operate...  ...across data, prompts, retrievals, models, and responses... 
    Work at office
    Local area
    Work from home
    Worldwide

    Sumitomo Mitsui Financial Group, Inc.

    Charlotte, NC
    1 day ago
  •  ...Solutions is a Cloud and AI-driven IT company...  ..., cloud engineering, data and analytics...  ...intelligent agents using Microsoft...  ...them into AI-driven platforms. You'll work closely...  ...to RAG (Retrieval-Augmented Generation...  ...DevOps tools and infrastructure-as-code ( Terraform... 
    Full time

    Symphony Solutions

    Remote
    a month ago
  •  ...enterprises whether their AI agents can be trusted — and we...  ...and integrate it into the platform alongside our engineering and data science teams. This...  ...team Build the eval infrastructure that makes those...  ...what "good" means Own retrieval and context pipelines over... 
    Remote work

    Monte Carlo

    Remote
    19 days ago
  •  ...is the knowledge infrastructure for AI at scale. Its leading...  ...and knowledge engine, Pinecone Nexus,...  ...hiring a senior/staff software engineer...  ...generation knowledge retrieval system built for...  ...build scalable platform components leveraging...  ...and autonomous agents will love. -... 
    Full time
    Worldwide

    Pinecone

    Remote
    5 days ago
  • $100k

     ...Givzey / Version2.ai Join the Future of...  ...that simply make staff more efficient, VEOs...  ...years, Givzey's platform has already helped...  ...ll work across AWS infrastructure, Infrastructure as...  ...to make sure engineers spend their time building...  ...databases Retrieval-Augmented Generation... 
    Full time
    Local area
    Day shift

    Givzey

    Remote
    more than 2 months ago
  •  ...company. Job Overview: The AI Platform Engineer builds, hardens, and helps...  ...safely develop and run AI agents at scale. This role owns...  ..., vector/knowledge base infrastructure, CI/CD and IaC automation,...  ...Knowledge Bases/vector-backed retrieval components where applicable... 
    Full time
    Remote work
    Flexible hours

    Cinteot

    Newark, NJ
    15 days ago
  •  ...LTS is seeking an AI Platform and Harness Engineer to develop and maintain the infrastructure, tooling, and evaluation frameworks that power enterprise AI solutions...  ...enable Large Language Models (LLMs), AI agents, Retrieval-Augmented Generation (RAG), and Generative... 
    Full time

    LTS

    Remote
    a month ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff AI Platform Engineer: Agent & Retrieval Infrastructure. Be the first to apply!