AI Agent Engineer, Cloud Infrastructure
Orbis
Job Description
Job Description
Company Overview
Orbis is a technology company that delivers sovereign intelligence and decision capabilities to the United States Government and its allies. We convert complex information and operational systems into decisive outcomes, giving public institutions and critical enterprises the tools to move faster, see clearer, and act first in an increasingly contested world.
We are not a consulting firm and not a systems integrator. Every Orbis team is built around practitioners who have operated at the mission edge, and every engagement is judged on one thing: whether it produces working capability, not slide decks. Solutions forged from experience; that is how we build, and it is how we hire.
Job Summary
Orbis is building the agentic AI capability its mission support operations run on, and the AI Agent Engineer builds and operates it. Sixty percent of this role is common to every AI Agent Engineer at Orbis: designing, building, and deploying the multi-agent workflows behind triage, data retrieval, summarization, reporting, and escalation routing. The remaining forty percent is this role's focus — the cloud platform those workflows run on. This engineer owns the footprint across GCP, Cloudflare, and Azure, manages it as code, and holds its networking, security posture, and run cost to a standard that survives scrutiny as volume grows.
Travel: Occasional (up to 10%) to McLean, VA or team offsites as needed
Key Responsibilities — Common Core (60% of the Role)
Every AI Agent Engineer at Orbis, regardless of focus area, owns the following:
- Design, build, and deploy agentic AI workflows. Architect and ship the multi-agent and automated workflows that run in production: triage, data retrieval, summarization, reporting, and escalation routing.
- Translate in both directions. Explain to non-technical stakeholders what the systems do, where they are reliable, and where they are not, in terms those stakeholders can act on. Equally, take a process owner's plain-language description of how their work happens and turn it into a workflow specification that holds up.
- Partner with process owners to build what they actually need. Work directly and iteratively with the colleagues who own a process so the resulting capability fits their operation rather than forcing them to fit it. Push back constructively when a stated requirement will not survive contact with production.
- Keep the workflows reliable. Monitor what is running in production, diagnose failures across prompts, tools, data, and orchestration, and resolve them. Know which problems are yours to solve and when to bring in engineering.
- Own the cost of execution. Track what your AI workloads cost to run. Identify where a tighter prompt, a caching layer, or a simple deterministic step delivers the same outcome, and keep those costs defensible as volume grows.
- Set the quality bar for AI-assisted output. Define accuracy standards, human review checkpoints, and escalation criteria for the automated workflows, and build the checks that prove those standards are being met.
- Document what you build. Author and maintain the runbooks and SOPs that allow others to operate and troubleshoot these workflows without you. Nothing lives only in one person's head.
- Integrate across Orbis products. Work across Catalyst, Pulse, and Discovery, with engineering support where needed, so workflows can access and act on the data they require.
- Deliver the reporting leadership and the team depend on. Build and maintain the outputs behind resolution rate, escalation rate, output accuracy, response time, and employee satisfaction, and turn the Senior Operations Manager's requirements into structured, recurring reports and data exports.
- Close capability gaps and keep the practice current. Recognize when a need is not served by existing tooling, scope the gap clearly, and either build it or bring it to engineering and product with enough detail to act on. Track developments in agentic AI tooling and enterprise automation platforms, and form a defensible view on what is worth adopting, at what pace, and for which use cases.
In addition to the common core above, this role owns the cloud infrastructure Orbis depends on — both the enterprise platform behind internal operations and the environments that support client delivery:
- Design, deploy, and maintain the Orbis cloud footprint across GCP, AWS, Cloudflare, and Azure, spanning both enterprise workloads and client delivery environments.
- Define and manage that footprint as infrastructure as code so environments are reproducible, reviewable, and recoverable.
- Own networking, identity, and secrets handling for agent workloads, including the boundaries between environments and the audit of access.
- Hold the security posture of the platform and keep it evidenced rather than asserted.
- Track and control platform run cost across compute, inference, storage, and egress, and report it in terms leadership can act on.
- Build and maintain the deployment and release path — CI/CD, environment promotion, rollback — so shipping an agent workflow is routine rather than an event.
Orbis weighs what a candidate has built over tenure or credentials. The following are true minimum requirements for this role:
- Must have 5 years of proven experience translating user requirements into wireframes, workflows and such.
- Must have 1-year proven experience utilizing Claude or similar commercial offering to build AI agents.
- Demonstrated experience building a functioning multi-agent workflow, with the ability to walk through the design rationale, how the agents hand off work and share context, what broke, and how it was fixed
- Practical command of prompt engineering, including retrieval-augmented generation (RAG) and tool or function calling
- Proven ability to explain technical work to non-technical audiences — describing a system's behavior, limits, and failure modes to a process owner or executive without hedging or retreating into jargon
- Demonstrated ability to elicit requirements from stakeholders who cannot hand over a specification, and to reach a working solution regardless
- Ability to troubleshoot agent workflows built on platforms such as Claude, GPT-class models, LangChain, AutoGen, or CrewAI, and to isolate whether a failure originates in the prompt, the tool, the data, or the orchestration
- Working proficiency with APIs and sufficient scripting ability to build and debug workflows independently; software engineering experience is not required, but a JSON payload or a Python script must not be a blocker
- Sound judgment about model output, including the discipline to state plainly when an output should not be trusted
- Strong written communication. This role produces documentation, runbooks, and reports that program leads and executives rely on
- Hands-on production experience with at least one major cloud platform (GCP, Azure, or Cloudflare) and working familiarity with a second
- Practical experience managing infrastructure as code using Terraform, Pulumi, or an equivalent
- Working command of cloud networking, identity, and secrets management, and the ability to treat security posture and run cost as design constraints rather than afterthoughts
Strong candidates will bring depth in one or more of the following. Do not screen yourself out for lacking items on this list.
- 5-7 years experience in software development
- Experience coding using Python
- Cloud depth: Kubernetes or serverless runtimes at production scale; multi-cloud or cross-boundary deployment; cloud migration or consolidation work
- Cost and governance: FinOps or formal cloud cost management practice; zero-trust architectures; federated access control models
- Delivery and observability: owning CI/CD pipelines end to end; infrastructure monitoring, logging, and alerting for production workloads
- Bridging technical and operational teams: prior work as a solutions consultant, technical program manager, forward-deployed engineer, or in a product-adjacent role
- AI reliability practice: monitoring, evaluation, or observability tooling for AI systems in production; automated evaluations or regression checks for LLM-based workflows; human-in-the-loop review steps that maintain quality without creating bottlenecks
- Process design: designing SOPs or escalation frameworks that hold up across multiple stakeholders
- AI-assisted development tooling: working fluency with Claude, Copilot, or Cursor
- Mission environment: operating within or supporting a government or defense program; OSINT, DIGINT, or intelligence analysis workflows and tradecraft; data governance concepts, federated access control models, or zero-trust architectures
- Operating range: a track record of coming up to speed quickly on unfamiliar systems, and effectiveness in a distributed, multi-program environment with competing priorities
- Prolonged periods of sitting at a desk and working on a computer
- Participation in virtual and in-person meetings
- Ability to attend planned meetings and/or work in classified spaces for extended periods within the specified work regions
- Travel: Occasional (up to 10%) to McLean, VA or team offsites as needed
Orbis Benefits
Beyond the opportunity to work alongside practitioners who have operated at the mission edge, joining Orbis means real ownership over outcomes that matter; the chance to build technology that gives the United States Government and its allies a genuine decision advantage, not just another dashboard. From competitive retirement matching to a PTO policy that encourages real time off, we're committed to creating an environment where our team can do their best work and still have a life outside of it.
Our comprehensive benefits package is designed to meet the diverse needs of our employees and their families. A full list of benefits is shared with candidates following an initial conversation with our HR team, so you'll have a clear picture of the support and resources available to you as part of the Orbis team.
Orbis Locations
Orbis works in whatever arrangement the mission requires. Depending on the role, that might mean five days a week on a customer site, full-time in one of our offices, a hybrid schedule, or fully remote work; our team is spread across 25 states and four countries to match. We make sure every team member has the tools and access needed to do the work, wherever that work happens.
Travel is common across many Orbis roles, particularly those supporting customer missions in the field.
Orbis is headquartered in McLean, VA, with additional office locations in Taipei, Taiwan, and Canberra, Australia.
Orbis is an Equal Employment Opportunity employer. We are committed to providing a work environment free of discrimination and harassment, and we make employment decisions based on merit,
qualifications, and business need. Orbis does not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, disability, protected veteran status, genetic information, or any other status protected by applicable law.
Orbis complies with all applicable equal opportunity and non-discrimination laws in every jurisdiction in which we operate. Reasonable accommodation is available to applicants and employees with disabilities upon request.
- ...Orbis is building the agentic AI capability its mission... ...operations run on, and the AI Agent Engineer builds and operates it. Sixty... ...is this role's focus — the cloud platform those workflows run... ...cases. Focus Area — Cloud Infrastructure (40% of the Role) In addition...SuggestedFull timeWork at officeRemote work
$84.4k - $156.8k
...public, private, and hybrid clouds. The world’s largest... ..., secure, and maintain AI application environments within Oracle Cloud Infrastructure (OCI).Adapt existing... ...Kubernetes/Oracle Kubernetes Engine (OKE), OCI Container... ...AI, OCI Generative AI Agents, OCI Data Science, or...SuggestedMinimum wageFull timeContract workFlexible hours- ...Orbis is building agentic AI capabilities for mission support operations. The AI Agent Engineer designs, builds, and deploys multi... ..., reporting) and owns the cloud platform across GCP, AWS, Cloudflare... ...Azure. You will manage infrastructure as code, ensure security...Suggested
$220k - $350k
...with the ultimate goal of enabling human life on Mars.SR. AI ENGINEER, PLATFORM INFRASTRUCTURE, SPECIAL PROGRAMSAs an AI Engineer, Platform... ...pipeline.Implement profile-driven rendering for public cloud, enterprise on-prem, and classified air-gap targets based...SuggestedPermanent employmentTemporary workImmediate startWeekend work- ...Mastercard is seeking a Principal AI Platform Engineer to build and scale next‑generation AI infrastructure across on‑premise private cloud, enabling high‑performance, low‑latency AI workloads at enterprise scale. You will own the AI infrastructure roadmap, partner...Suggested
$184k - $287.5k
AI Infrastructure Engineers at NVIDIA build the systems, tooling, and data infrastructure that enable operation of our GPU cloud services. We are enabling engineering teams to innovate while proactively... ...along with familiarity with AI agent frameworks or orchestration...Full timeRemote work$191k - $253k
...Senior AI Infrastructure Engineer Anduril Industries is a defense technology company with a mission to transform U.S. and allied military... ...AI models (including LLMs, computer vision, and RL agents) across cloud environments and air-gapped, edge-deployed networks. Working...$148.5k - $223.9k
...DetailsAbout SalesforceSalesforce is the #1 AI CRM, where humans with agents drive customer success together.... ...future of Salesforce.Our Public Cloud engineering teams are responsible for... ...systems.Your Impact:Deliver cloud infrastructure automation tools, frameworks, workflows...Full time- ...Job Details: Role: AI Infrastructure Engineer Location: Remote Duration: 12+ Months Contract... ...by migrating systems to a secure cloud environment and modernizing software that... ...internal data sources to Bedrock agents - not just configuring existing ones see...Contract workRemote work
$229.9k - $262.4k
AI Engineer 5 (GenAI Platform, Agentic Infrastructure) At Capital One, we are creating responsible and reliable AI systems... ...large language model inference, agents and multi-agent workflows,... ...and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud...Full timePart timeLocal area- ...United States Digital Space LLC in El Segundo, CA, is seeking a Senior Software Engineer for AI infrastructure to build and scale Starshield's GPU-centric platforms. You will design, deploy, and optimize on-prem systems and collaborate with AI engineers to deliver scalable...
- ...SpaceX is seeking a Senior Software Engineer for AI Infrastructure (Starshield) to design, operate, and scale GPU/CPU infrastructures supporting critical national security missions. You will deploy on-premise resources, build scalable software, and collaborate across...
- ...SpaceX is hiring a Software Engineer to design, operate and scale Starshield AI infrastructure. You will manage GPU/CPU infrastructure deployments in Top Secret datacenters and build automation for on-prem Kubernetes/AI clusters. You will collaborate with AI engineers...
- ...SpaceX is seeking a Software Engineer for AI Infrastructure (Starshield) to design, deploy, and scale AI GPU infrastructure and services. The role spans Site Reliability Engineering, DevOps, and GPU platforms, focusing on on‑premise compute resources, automation, and...
$197.3k - $225.1k
AI Engineer 4 (Gen AI Platform Services: Agentic AI, Agent Guardrails, Agent Evaluation, Agent Memory) At Capital One, we... ...Our investments in technology infrastructure and world-class talent — along... ...and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud...Full timePart timeLocal area- ...Cape is seeking a software engineer to help build privacy-focused telecom infrastructure. You’ll own the full lifecycle development and collaborate with world-class engineers across teams in a startup-like environment based in New York. This role emphasizes privacy, national...Work at officeRemote work
- ...Piper Companies is seeking an Azure Infrastructure Engineer to join a growing financial services organization in Bethesda, MD. This hybrid role... ...building and modernizing secure, scalable Azure infrastructure for cloud-native and serverless transitions. Responsibilities...
- ...E-INFOSOL LLC is seeking an Infrastructure Operations Engineer to join a Washington, DC team. This full-time role ensures stability, security, and performance... ...of SharePoint, Power BI RS, and SQL across on-prem and cloud environments. The candidate will manage Windows Server...Full time
$146k - $234k
ResponsibilitiesThe Senior AI Platform Engineer will stand up and operate a bare-metal container... ...technologies, and drive the transition towards infrastructure-as-code managed provisioning using... ...of OpenShift versus Spectro Cloud, documenting licensing tradeoffs, operability...Contract workRemote workFlexible hoursShift work$107.9k - $195.05k
Leidos is seeking a Senior AI Platform Engineer to join our DCSB Artificial Intelligence & Agent Development team supporting the... ...(IL5). You will combine deep cloud and platform engineering expertise... ...maintain CI/CD pipelines and Infrastructure as Code (IaC) using Platform...Full timeRemote work$114.6k - $252.1k
Job Title: AI Cloud Platform EngineerJob Category: Information TechnologyTime... ...are seeking an AI and Cloud Engineer to design, build, and scale... ...and robust cloud infrastructure. You will be part of a team... ...architecting complex multi-agent flows, implementing enterprise...Contract workWork experience placementFlexible hours$128.4k - $192.5k
...to build a better working world.Government and Infrastructure - Technology Consulting - AI & Data - AI Automation Engineer - Senior ConsultantFrom strategy to execution,... ...measures and regression tests for prompts and agents — and track quality metrics over time.· Design...For contractorsSummer holidayWork at officeLocal areaImmediate startFlexible hours$160k - $230k
MTSI is seeking a Senior Cloud Infrastructure Engineer who will oversee the performance of contracted support building out government cloud-based infrastructure... ...pipelines, and collaboration services supporting multiple AI-enabled software programsYour essential job functions will...Contract work- Primary Responsibilities:Infrastructure & Systems AdministrationAdminister and support Microsoft... ...OperationsSupport and maintain AWS cloud infrastructure, ensuring high availability... ...Technology, Computer Science, Engineering, or equivalent experience7+ years of hands...Full timeWork at officeMonday to Friday
$146k - $194k
...Anduril’s family of systems is powered by Lattice OS, an AI-powered operating system that turns thousands of data streams... ...the military in months, not years.ABOUT THE JOBAnduril’s Cloud Infrastructure Engineering team is the foundation upon which our advanced defense...Full timeWork experience placementImmediate start- ...Senior Cloud Engineer Remote Long Term The Cloud Engineer-Senior supports by designing, implementing, and operating secure, scalable cloud infrastructure for enterprise applications and services. This role is central to OIT's hybrid infrastructure strategy, ensuring...Remote work
$103.2k - $203.4k
...Generative AI Applications Engineer (Agents & RAG) Washington, DC At Accenture Federal Services, nothing matters more than helping the US federal... .... Nice to Have Integration with leading cloud AI services or on prem inference stacks Background in...- ...Role: Senior AI/ML Cloud Engineer (TS/SCI Polygraph Required) Location: Remote w/ occasional travel D.C A high-growth AI infrastructure company is hiring a an AI/ML Cloud Engineer to support... ...developer, administrator, and AI-agent workflow perspectives. Install,...Full timeRemote work
$150.7k - $251.2k
...to build a better working world.Government and Infrastructure - Technology Consulting - AI & Data - AI Automation Engineer - ManagerFrom strategy to execution,... ...measures and regression tests for prompts and agents — and track quality metrics over time.· Design...For contractorsSummer holidayWork at officeLocal areaImmediate startFlexible hours- ...We are seeking a Principal Cloud Infrastructure Architect to provide strategic and technical leadership... ..., security, automation, platform engineering, and resilient infrastructure. The architect... ...and architect infrastructure for AI/ML, Generative AI, and intelligent...Contract work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Agent Engineer, Cloud Infrastructure. Be the first to apply!
- import export agent Washington DC
- operations agent Washington DC
- freight agent no experience Washington DC
- cruise agent Washington DC
- commissioning agent Washington DC
- tsa agent Washington DC
- registered agent Washington DC
- executive protection agent Washington DC
- work from home chat agent Washington DC
- special agent Washington DC






