Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Agent Engineer, Cloud Infrastructure

Orbis

Job Description

Job Description

Company Overview

Orbis is a technology company that delivers sovereign intelligence and decision capabilities to the United States Government and its allies. We convert complex information and operational systems into decisive outcomes, giving public institutions and critical enterprises the tools to move faster, see clearer, and act first in an increasingly contested world.

We are not a consulting firm and not a systems integrator. Every Orbis team is built around practitioners who have operated at the mission edge, and every engagement is judged on one thing: whether it produces working capability, not slide decks. Solutions forged from experience; that is how we build, and it is how we hire.

Job Summary

Orbis is building the agentic AI capability its mission support operations run on, and the AI Agent Engineer builds and operates it. Sixty percent of this role is common to every AI Agent Engineer at Orbis: designing, building, and deploying the multi-agent workflows behind triage, data retrieval, summarization, reporting, and escalation routing. The remaining forty percent is this role's focus — the cloud platform those workflows run on. This engineer owns the footprint across GCP, Cloudflare, and Azure, manages it as code, and holds its networking, security posture, and run cost to a standard that survives scrutiny as volume grows.

Travel: Occasional (up to 10%) to McLean, VA or team offsites as needed

Key Responsibilities — Common Core (60% of the Role)

Every AI Agent Engineer at Orbis, regardless of focus area, owns the following:

  • Design, build, and deploy agentic AI workflows. Architect and ship the multi-agent and automated workflows that run in production: triage, data retrieval, summarization, reporting, and escalation routing.
  • Translate in both directions. Explain to non-technical stakeholders what the systems do, where they are reliable, and where they are not, in terms those stakeholders can act on. Equally, take a process owner's plain-language description of how their work happens and turn it into a workflow specification that holds up.
  • Partner with process owners to build what they actually need. Work directly and iteratively with the colleagues who own a process so the resulting capability fits their operation rather than forcing them to fit it. Push back constructively when a stated requirement will not survive contact with production.
  • Keep the workflows reliable. Monitor what is running in production, diagnose failures across prompts, tools, data, and orchestration, and resolve them. Know which problems are yours to solve and when to bring in engineering.
  • Own the cost of execution. Track what your AI workloads cost to run. Identify where a tighter prompt, a caching layer, or a simple deterministic step delivers the same outcome, and keep those costs defensible as volume grows.
  • Set the quality bar for AI-assisted output. Define accuracy standards, human review checkpoints, and escalation criteria for the automated workflows, and build the checks that prove those standards are being met.
  • Document what you build. Author and maintain the runbooks and SOPs that allow others to operate and troubleshoot these workflows without you. Nothing lives only in one person's head.
  • Integrate across Orbis products. Work across Catalyst, Pulse, and Discovery, with engineering support where needed, so workflows can access and act on the data they require.
  • Deliver the reporting leadership and the team depend on. Build and maintain the outputs behind resolution rate, escalation rate, output accuracy, response time, and employee satisfaction, and turn the Senior Operations Manager's requirements into structured, recurring reports and data exports.
  • Close capability gaps and keep the practice current. Recognize when a need is not served by existing tooling, scope the gap clearly, and either build it or bring it to engineering and product with enough detail to act on. Track developments in agentic AI tooling and enterprise automation platforms, and form a defensible view on what is worth adopting, at what pace, and for which use cases.
Focus Area — Cloud Infrastructure (40% of the Role)

In addition to the common core above, this role owns the cloud infrastructure Orbis depends on — both the enterprise platform behind internal operations and the environments that support client delivery:

  • Design, deploy, and maintain the Orbis cloud footprint across GCP, AWS, Cloudflare, and Azure, spanning both enterprise workloads and client delivery environments.
  • Define and manage that footprint as infrastructure as code so environments are reproducible, reviewable, and recoverable.
  • Own networking, identity, and secrets handling for agent workloads, including the boundaries between environments and the audit of access.
  • Hold the security posture of the platform and keep it evidenced rather than asserted.
  • Track and control platform run cost across compute, inference, storage, and egress, and report it in terms leadership can act on.
  • Build and maintain the deployment and release path — CI/CD, environment promotion, rollback — so shipping an agent workflow is routine rather than an event.
Required Qualifications

Orbis weighs what a candidate has built over tenure or credentials. The following are true minimum requirements for this role:

  • Must have 5 years of proven experience translating user requirements into wireframes, workflows and such.
  • Must have 1-year proven experience utilizing Claude or similar commercial offering to build AI agents.
  • Demonstrated experience building a functioning multi-agent workflow, with the ability to walk through the design rationale, how the agents hand off work and share context, what broke, and how it was fixed
  • Practical command of prompt engineering, including retrieval-augmented generation (RAG) and tool or function calling
  • Proven ability to explain technical work to non-technical audiences — describing a system's behavior, limits, and failure modes to a process owner or executive without hedging or retreating into jargon
  • Demonstrated ability to elicit requirements from stakeholders who cannot hand over a specification, and to reach a working solution regardless
  • Ability to troubleshoot agent workflows built on platforms such as Claude, GPT-class models, LangChain, AutoGen, or CrewAI, and to isolate whether a failure originates in the prompt, the tool, the data, or the orchestration
  • Working proficiency with APIs and sufficient scripting ability to build and debug workflows independently; software engineering experience is not required, but a JSON payload or a Python script must not be a blocker
  • Sound judgment about model output, including the discipline to state plainly when an output should not be trusted
  • Strong written communication. This role produces documentation, runbooks, and reports that program leads and executives rely on
  • Hands-on production experience with at least one major cloud platform (GCP, Azure, or Cloudflare) and working familiarity with a second
  • Practical experience managing infrastructure as code using Terraform, Pulumi, or an equivalent
  • Working command of cloud networking, identity, and secrets management, and the ability to treat security posture and run cost as design constraints rather than afterthoughts
Desired Qualifications

Strong candidates will bring depth in one or more of the following. Do not screen yourself out for lacking items on this list.

  • 5-7 years experience in software development
  • Experience coding using Python
  • Cloud depth: Kubernetes or serverless runtimes at production scale; multi-cloud or cross-boundary deployment; cloud migration or consolidation work
  • Cost and governance: FinOps or formal cloud cost management practice; zero-trust architectures; federated access control models
  • Delivery and observability: owning CI/CD pipelines end to end; infrastructure monitoring, logging, and alerting for production workloads
  • Bridging technical and operational teams: prior work as a solutions consultant, technical program manager, forward-deployed engineer, or in a product-adjacent role
  • AI reliability practice: monitoring, evaluation, or observability tooling for AI systems in production; automated evaluations or regression checks for LLM-based workflows; human-in-the-loop review steps that maintain quality without creating bottlenecks
  • Process design: designing SOPs or escalation frameworks that hold up across multiple stakeholders
  • AI-assisted development tooling: working fluency with Claude, Copilot, or Cursor
  • Mission environment: operating within or supporting a government or defense program; OSINT, DIGINT, or intelligence analysis workflows and tradecraft; data governance concepts, federated access control models, or zero-trust architectures
  • Operating range: a track record of coming up to speed quickly on unfamiliar systems, and effectiveness in a distributed, multi-program environment with competing priorities
Physical Requirements
  • Prolonged periods of sitting at a desk and working on a computer
  • Participation in virtual and in-person meetings
  • Ability to attend planned meetings and/or work in classified spaces for extended periods within the specified work regions
  • Travel: Occasional (up to 10%) to McLean, VA or team offsites as needed

Orbis Benefits

Beyond the opportunity to work alongside practitioners who have operated at the mission edge, joining Orbis means real ownership over outcomes that matter; the chance to build technology that gives the United States Government and its allies a genuine decision advantage, not just another dashboard. From competitive retirement matching to a PTO policy that encourages real time off, we're committed to creating an environment where our team can do their best work and still have a life outside of it.

Our comprehensive benefits package is designed to meet the diverse needs of our employees and their families. A full list of benefits is shared with candidates following an initial conversation with our HR team, so you'll have a clear picture of the support and resources available to you as part of the Orbis team.

Orbis Locations

Orbis works in whatever arrangement the mission requires. Depending on the role, that might mean five days a week on a customer site, full-time in one of our offices, a hybrid schedule, or fully remote work; our team is spread across 25 states and four countries to match. We make sure every team member has the tools and access needed to do the work, wherever that work happens.

Travel is common across many Orbis roles, particularly those supporting customer missions in the field.

Orbis is headquartered in McLean, VA, with additional office locations in Taipei, Taiwan, and Canberra, Australia.

 

Orbis is an Equal Employment Opportunity employer. We are committed to providing a work environment free of discrimination and harassment, and we make employment decisions based on merit, 

qualifications, and business need. Orbis does not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, disability, protected veteran status, genetic information, or any other status protected by applicable law.

Orbis complies with all applicable equal opportunity and non-discrimination laws in every jurisdiction in which we operate. Reasonable accommodation is available to applicants and employees with disabilities upon request. 

 

Vacancy posted 6 days ago
Similar jobs that could be interesting for youBased on the AI Agent Engineer, Cloud Infrastructure in Washington DC vacancy
  •  ...Orbis is building the agentic AI capability its mission...  ...operations run on, and the AI Agent Engineer builds and operates it. Sixty...  ...is this role's focus — the cloud platform those workflows run...  ...cases. Focus Area — Cloud Infrastructure (40% of the Role) In addition... 
    Suggested
    Full time
    Work at office
    Remote work

    Jobleads-US

    Washington DC
    4 days ago
  • $84.4k - $156.8k

     ...public, private, and hybrid clouds. The world’s largest...  ..., secure, and maintain AI application environments within Oracle Cloud Infrastructure (OCI).Adapt existing...  ...Kubernetes/Oracle Kubernetes Engine (OKE), OCI Container...  ...AI, OCI Generative AI Agents, OCI Data Science, or... 
    Suggested
    Minimum wage
    Full time
    Contract work
    Flexible hours

    DXC Technology

    Arlington, VA
    1 day ago
  •  ...Orbis is building agentic AI capabilities for mission support operations. The AI Agent Engineer designs, builds, and deploys multi...  ..., reporting) and owns the cloud platform across GCP, AWS, Cloudflare...  ...Azure. You will manage infrastructure as code, ensure security... 
    Suggested

    Jobleads-US

    Washington DC
    4 days ago
  • $220k - $350k

     ...with the ultimate goal of enabling human life on Mars.SR. AI ENGINEER, PLATFORM INFRASTRUCTURE, SPECIAL PROGRAMSAs an AI Engineer, Platform...  ...pipeline.Implement profile-driven rendering for public cloud, enterprise on-prem, and classified air-gap targets based... 
    Suggested
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    4 days ago
  •  ...Mastercard is seeking a Principal AI Platform Engineer to build and scale next‑generation AI infrastructure across on‑premise private cloud, enabling high‑performance, low‑latency AI workloads at enterprise scale. You will own the AI infrastructure roadmap, partner... 
    Suggested

    Jobleads-US

    Arlington, VA
    4 days ago
  • $184k - $287.5k

    AI Infrastructure Engineers at NVIDIA build the systems, tooling, and data infrastructure that enable operation of our GPU cloud services. We are enabling engineering teams to innovate while proactively...  ...along with familiarity with AI agent frameworks or orchestration... 
    Full time
    Remote work

    Nvidia

    Washington DC
    2 days ago
  • $191k - $253k

     ...Senior AI Infrastructure Engineer Anduril Industries is a defense technology company with a mission to transform U.S. and allied military...  ...AI models (including LLMs, computer vision, and RL agents) across cloud environments and air-gapped, edge-deployed networks. Working... 

    anduril

    Washington DC
    2 days ago
  • $148.5k - $223.9k

     ...DetailsAbout SalesforceSalesforce is the #1 AI CRM, where humans with agents drive customer success together....  ...future of Salesforce.Our Public Cloud engineering teams are responsible for...  ...systems.Your Impact:Deliver cloud infrastructure automation tools, frameworks, workflows... 
    Full time

    Salesforce

    Washington DC
    3 days ago
  •  ...Job Details: Role: AI Infrastructure Engineer Location: Remote Duration: 12+ Months Contract...  ...by migrating systems to a secure cloud environment and modernizing software that...  ...internal data sources to Bedrock agents - not just configuring existing ones see... 
    Contract work
    Remote work

    Creative G C

    Washington DC
    3 days ago
  • $229.9k - $262.4k

    AI Engineer 5 (GenAI Platform, Agentic Infrastructure) At Capital One, we are creating responsible and reliable AI systems...  ...large language model inference, agents and multi-agent workflows,...  ...and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud... 
    Full time
    Part time
    Local area

    Capital One

    McLean, VA
    2 days ago
  •  ...United States Digital Space LLC in El Segundo, CA, is seeking a Senior Software Engineer for AI infrastructure to build and scale Starshield's GPU-centric platforms. You will design, deploy, and optimize on-prem systems and collaborate with AI engineers to deliver scalable... 

    Jobleads-US

    Washington DC
    3 days ago
  •  ...SpaceX is seeking a Senior Software Engineer for AI Infrastructure (Starshield) to design, operate, and scale GPU/CPU infrastructures supporting critical national security missions. You will deploy on-premise resources, build scalable software, and collaborate across... 

    Jobleads-US

    Washington DC
    4 days ago
  •  ...SpaceX is hiring a Software Engineer to design, operate and scale Starshield AI infrastructure. You will manage GPU/CPU infrastructure deployments in Top Secret datacenters and build automation for on-prem Kubernetes/AI clusters. You will collaborate with AI engineers... 

    Jobleads-US

    Washington DC
    3 days ago
  •  ...SpaceX is seeking a Software Engineer for AI Infrastructure (Starshield) to design, deploy, and scale AI GPU infrastructure and services. The role spans Site Reliability Engineering, DevOps, and GPU platforms, focusing on on‑premise compute resources, automation, and... 

    Jobleads-US

    Washington DC
    4 days ago
  • $197.3k - $225.1k

    AI Engineer 4 (Gen AI Platform Services: Agentic AI, Agent Guardrails, Agent Evaluation, Agent Memory) At Capital One, we...  ...Our investments in technology infrastructure and world-class talent — along...  ...and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud... 
    Full time
    Part time
    Local area

    Capital One

    McLean, VA
    1 day ago
  •  ...Cape is seeking a software engineer to help build privacy-focused telecom infrastructure. You’ll own the full lifecycle development and collaborate with world-class engineers across teams in a startup-like environment based in New York. This role emphasizes privacy, national... 
    Work at office
    Remote work

    Jobleads-US

    Arlington, VA
    2 days ago
  •  ...Piper Companies is seeking an Azure Infrastructure Engineer to join a growing financial services organization in Bethesda, MD. This hybrid role...  ...building and modernizing secure, scalable Azure infrastructure for cloud-native and serverless transitions. Responsibilities... 

    Jobleads-US

    Bethesda, MD
    3 days ago
  •  ...E-INFOSOL LLC is seeking an Infrastructure Operations Engineer to join a Washington, DC team. This full-time role ensures stability, security, and performance...  ...of SharePoint, Power BI RS, and SQL across on-prem and cloud environments. The candidate will manage Windows Server... 
    Full time

    Jobleads-US

    Washington DC
    4 days ago
  • $146k - $234k

    ResponsibilitiesThe Senior AI Platform Engineer will stand up and operate a bare-metal container...  ...technologies, and drive the transition towards infrastructure-as-code managed provisioning using...  ...of OpenShift versus Spectro Cloud, documenting licensing tradeoffs, operability... 
    Contract work
    Remote work
    Flexible hours
    Shift work

    Peraton Corporation

    Arlington, VA
    2 days ago
  • $107.9k - $195.05k

    Leidos is seeking a Senior AI Platform Engineer to join our DCSB Artificial Intelligence & Agent Development team supporting the...  ...(IL5). You will combine deep cloud and platform engineering expertise...  ...maintain CI/CD pipelines and Infrastructure as Code (IaC) using Platform... 
    Full time
    Remote work

    Leidos

    Adelphi, MD
    4 days ago
  • $114.6k - $252.1k

    Job Title: AI Cloud Platform EngineerJob Category: Information TechnologyTime...  ...are seeking an AI and Cloud Engineer to design, build, and scale...  ...and robust cloud infrastructure. You will be part of a team...  ...architecting complex multi-agent flows, implementing enterprise... 
    Contract work
    Work experience placement
    Flexible hours

    CACI International

    Washington DC
    2 days ago
  • $128.4k - $192.5k

     ...to build a better working world.Government and Infrastructure - Technology Consulting - AI & Data - AI Automation Engineer - Senior ConsultantFrom strategy to execution,...  ...measures and regression tests for prompts and agents — and track quality metrics over time.· Design... 
    For contractors
    Summer holiday
    Work at office
    Local area
    Immediate start
    Flexible hours

    EY (Ernst & Young)

    McLean, VA
    3 days ago
  • $160k - $230k

    MTSI is seeking a Senior Cloud Infrastructure Engineer who will oversee the performance of contracted support building out government cloud-based infrastructure...  ...pipelines, and collaboration services supporting multiple AI-enabled software programsYour essential job functions will... 
    Contract work

    Modern Technology Solutions

    Washington DC
    1 day ago
  • Primary Responsibilities:Infrastructure & Systems AdministrationAdminister and support Microsoft...  ...OperationsSupport and maintain AWS cloud infrastructure, ensuring high availability...  ...Technology, Computer Science, Engineering, or equivalent experience7+ years of hands... 
    Full time
    Work at office
    Monday to Friday

    Vanda Pharmaceuticals

    Washington DC
    4 days ago
  • $146k - $194k

     ...Anduril’s family of systems is powered by Lattice OS, an AI-powered operating system that turns thousands of data streams...  ...the military in months, not years.ABOUT THE JOBAnduril’s Cloud Infrastructure Engineering team is the foundation upon which our advanced defense... 
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    4 days ago
  •  ...Senior Cloud Engineer Remote Long Term The Cloud Engineer-Senior supports by designing, implementing, and operating secure, scalable cloud infrastructure for enterprise applications and services. This role is central to OIT's hybrid infrastructure strategy, ensuring... 
    Remote work

    Mastech

    Washington DC
    4 days ago
  • $103.2k - $203.4k

     ...Generative AI Applications Engineer (Agents & RAG) Washington, DC At Accenture Federal Services, nothing matters more than helping the US federal...  .... Nice to Have Integration with leading cloud AI services or on prem inference stacks Background in... 

    Accenture Federal Services

    Washington DC
    5 days ago
  •  ...Role: Senior AI/ML Cloud Engineer (TS/SCI Polygraph Required) Location: Remote w/ occasional travel D.C A high-growth AI infrastructure company is hiring a an AI/ML Cloud Engineer to support...  ...developer, administrator, and AI-agent workflow perspectives. Install,... 
    Full time
    Remote work

    Colossus Technologies Group

    Washington DC
    23 hours ago
  • $150.7k - $251.2k

     ...to build a better working world.Government and Infrastructure - Technology Consulting - AI & Data - AI Automation Engineer - ManagerFrom strategy to execution,...  ...measures and regression tests for prompts and agents — and track quality metrics over time.· Design... 
    For contractors
    Summer holiday
    Work at office
    Local area
    Immediate start
    Flexible hours

    EY (Ernst & Young)

    McLean, VA
    3 days ago
  •  ...We are seeking a Principal Cloud Infrastructure Architect to provide strategic and technical leadership...  ..., security, automation, platform engineering, and resilient infrastructure. The architect...  ...and architect infrastructure for AI/ML, Generative AI, and intelligent... 
    Contract work

    Techvilla Solutions

    Washington DC
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Agent Engineer, Cloud Infrastructure. Be the first to apply!