Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Engineer 4 (LLM Gateway, FM Hosting)

$197.3k - $225.1k

Capital One National Association

AI Engineer 4 (LLM Gateway, FM Hosting)

At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine learning — position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build.

Team Description:

The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact.

What You’ll Do:

  • Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One.
  • Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.
  • Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more.
  • Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance — scalability, cost, latency, throughput — of large scale production AI systems.
  • Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
  • Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment
  • Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift
  • Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines
  • Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met
  • Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation

The Ideal Candidate:

  • You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good
  • Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production
  • You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven
  • You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss
  • You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown

Basic Qualifications:

  • Bachelor's Degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies
  • At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java

Preferred Qualifications:

  • Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy
  • 6+ years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud)
  • Experience designing, developing, delivering, and supporting AI services
  • Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang
  • Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost
  • Experience in building agentic AI systems and agentic workflows
  • Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production
  • Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale
  • Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules
  • Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms

Capital One will consider sponsoring a new qualified applicant for employment authorization for this position.

Cambridge, MA: $197,300 - $225,100 for AI Engineer 4McLean, VA: $197,300 - $225,100 for AI Engineer 4New York, NY: $215,200 - $245,600 for AI Engineer 4San Jose, CA: $215,200 - $245,600 for AI Engineer 4

Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate’s offer letter.

This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan.

Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website. Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level.

This role is expected to accept applications for a minimum of 5 business days.

Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections 4901-4920; New York City’s Fair Chance Act; Philadelphia’s Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries.

View email address on click.appcast.io

View email address on click.appcast.io

Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site.

Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).

#J-18808-Ljbffr

Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the AI Engineer 4 (LLM Gateway, FM Hosting) in McLean, VA vacancy
  • $229.9k - $262.4k

     ...AI Engineer 5 (FM Hosting, LLM Inference) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years...  ...Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or... 
    Suggested
    Full time
    Part time
    Local area

    Capital One National Association

    McLean, VA
    10 hours ago
  • $229.9k - $262.4k

     ...creating responsible and reliable AI systems, changing banking for...  ...‑class applied science and engineering teams to deliver our industry leading...  ...related fields plus at least 4 years of experience developing...  ...or technologies (e.g. LLM Inference, Similarity Search and... 
    Suggested
    Full time
    Part time
    Local area

    Capital One National Association

    McLean, VA
    10 hours ago
  • $229.9k - $262.4k

    ## Senior Lead AI Engineer (FM Hosting, LLM Inference)Applylocations: McLean, VA: New York, NY: San Jose, CAtime type: Full timeposted on: Posted Todayjob...  ..., Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or... 
    Suggested
    Full time
    Part time
    Local area

    Capital One Group

    McLean, VA
    1 day ago
  • $229.9k - $262.4k

    Senior Lead AI Engineer (FM Hosting, LLM Inference) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking...  ...Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or... 
    Suggested
    Full time
    Part time
    Local area

    Capital One

    McLean, VA
    2 days ago
  •  ...Capital One is seeking an AI Engineer 5 to advance foundational AI systems within the IFX group in McLean, VA. You will design, implement, and optimize AI components including model training, inference, and orchestration, partnering with engineers, researchers, and PMs... 
    Suggested

    Jobleads-US

    McLean, VA
    9 hours ago
  • $229.9k - $262.4k

     ...AI Engineer 5 (AI Foundations, LLM Core and Agentic AI) At Capital One, we are creating responsible and reliable AI systems, changing banking for good...  ..., Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies... 
    Full time
    Part time
    Local area

    Capital One National Association

    McLean, VA
    10 hours ago
  •  ...Capital One is seeking an AI Engineer 4 focused on AI Foundations, LLM customization and finetuning to advance responsible AI in banking. You will design and deliver AI powered products with cross-functional teams. You will train models, optimize inference pipelines... 

    Jobleads-US

    McLean, VA
    9 hours ago
  •  ...defense technology company, is seeking an AI Implementation Engineer to be part of our Warfare Systems...  ...available AI tooling/chat, hosting LLM servers and has demonstrated ability...  ...NVIDIA preferred) Experience with AI Gateways Experience with LLM servers and model... 
    Full time
    Work at office
    Local area
    Immediate start
    Shift work

    Innovative Defense Technologies

    Arlington, VA
    24 days ago
  •  ...Capital One is seeking an AI Engineer 5 to help build and scale responsible AI systems. You will collaborate with cross-functional teams on foundation models, LLM inference, and multi-model orchestration within a scalable AI infrastructure. The role emphasizes performance... 

    Jobleads-US

    McLean, VA
    9 hours ago
  •  ...Capital One invites applications for an AI Engineer 5 focused on AI Foundations and VLM customization. This role partners with cross‑functional...  ...deploy AI software components, including foundation models, LLM inference, and multi‑model workflows. The ideal candidate... 

    Jobleads-US

    McLean, VA
    9 hours ago
  •  ...Capital One is seeking an AI Engineer 5 to build and optimize foundational AI systems within the Intelligent Foundations and Experiences...  ...support AI components, including foundation model training and LLM inference, across a broad stack of open-source and SaaS technologies... 

    Jobleads-US

    McLean, VA
    9 hours ago
  •  ...Capital One is seeking an AI Engineer 4 to help build responsible AI systems and scalable AI infrastructure. You will design and deploy AI components, training pipelines, inference engines, and governance mechanisms while partnering with cross-functional teams across... 

    Jobleads-US

    McLean, VA
    9 hours ago
  • $135k - $150k

     ...Generative AI Application Engineer Location: United States (Remote...  ...large language model (LLM) APIs — along with the...  ...Work with models hosted on Hugging Face and with...  ...production At least 4 years of hands-on Python...  ....cpp, or similar). Gateway/routing layers like LiteLLM... 
    Full time
    Temporary work
    Immediate start
    Remote work
    Flexible hours
    Weekend work

    Purple Squirrel Enterprises

    Washington DC
    2 days ago
  •  ...We build AI agents that actually work in enterprise...  ...prototypes, not demos. We need engineers who can own the entire...  ...that scales, and LLM integrations that are model...  .../Fargate, Bedrock, API Gateway), Oracle OCI (OKE, Functions...  ...on our engagements. ~4+ years of software... 
    Temporary work

    Trilagen

    Bethesda, MD
    3 days ago
  • $150k - $200k

     ...Description Job Description Forward Deployed AI Engineer Location: Washington, DC (McLean, VA –...  ...Deployed AI Engineer with approximately 4–6 years of experience to architect and...  ..., tool-calling agents, or related LLM orchestration technologies • Experience... 
    Full time
    For contractors
    Local area
    Immediate start
    Night shift

    Fuze HR Solutions Inc.

    McLean, VA
    a month ago
  • $75 - $80 per hour

     ...hr Direct message the job poster from Matlen Silver Award Winning Senior Technical Recruiter at Matlen Silver Job Title: Senior AI/LLM Engineer Duration: 12+ Months Location: Washington, DC Required Pay Scale: $75-$80/hour W2 ***Due to client requirements this role is... 
    Full time

    Matlen Silver

    Washington DC
    20 hours ago
  • $229.9k - $262.4k

     ...creating responsible and reliable AI systems, changing banking for...  ...-class applied science and engineering teams to deliver our industry leading...  ...related fields plus at least 4 years of experience developing...  ...or technologies (e.g. LLM Inference, Similarity Search and... 
    Full time
    Part time
    Local area

    Capital One National Association

    McLean, VA
    10 hours ago
  • ID.me is hiring a Staff Software Development Engineer to lead the Tools Team in McLean, VA. You will champion AI-native engineering, architect production-grade solutions...  ...in full-stack engineering, including hands-on LLM integration experience. This on-site role fosters... 

    ID.me

    Mc Lean, VA
    1 day ago
  • $150k - $275k

     ...accomplishing hard things, together.Lead AI Security Engineer (Agentic SOC)Location: McLean/Tysons, VA...  ...incident logs into agent prompts.LLM Performance Engineering: Continuously evaluate...  ...Engineering, or related fields plus at least 4 years of experience developing AI and... 
    Work experience placement
    Local area
    Flexible hours

    Appian

    McLean, VA
    3 days ago
  •  ...Solutions Inc. is seeking a Lead AI Engineer / Architect to serve as the...  ...Generative AI, Large Language Model (LLM), Retrieval-Augmented...  ...integrated, including version, hosting approach, FedRAMP status, benchmarks...  ...(HSA) Retirement Plan with 4% match and discretionary... 
    Permanent employment
    Temporary work
    Interim role
    Work at office
    Flexible hours

    PRECISE SOFTWARE SOLUTIONS INCORPORATED

    Washington DC
    10 hours ago
  • $179.4k - $204.7k

     ...Data Engineer 4 (Python, AWS, SQL, GenAI) (Enterprise Platforms Technology) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative, inclusive and iterative delivery environment? At Capital... 
    Internship
    H1b
    Local area

    Capital One National Association

    McLean, VA
    10 hours ago
  • $145k - $205k

     ...computer science, software engineering, artificial...  ...integrating generative AI or LLMs into production...  ...orchestration frameworks, model gateways, MCP-compatible...  ...open-source, or locally hosted LLMs in cloud, on-premises...  ...actions. Implement LLM techniques including prompt... 
    Full time

    Vantor

    Herndon, VA
    2 days ago
  • $145k - $205k

     ...We are seeking a Senior AI/ML Software Engineer to integrate secure chatbot...  ...engineering skills, practical LLM experience, and an understanding...  ...frameworks, model gateways, MCP-compatible services, or...  ...commercial, open-source, or locally hosted LLMs in cloud, on-premises,... 

    Vantor

    Herndon, VA
    2 days ago
  • $179.4k - $204.7k

     ...Full-stack Engineer 4, Agentic Technology Do you love building and pioneering in the...  ...orchestration runtime that powers Capital One's AI content generation platform. You will...  ...agentic pipeline with the enterprise LLM model gateway — manage API calls, handle rate... 
    Full time
    Part time
    Internship
    H1b
    Local area

    Capital One National Association

    McLean, VA
    10 hours ago
  • $99k - $225k

    Agentic AI Forward-Deployed EngineerThe Opportunity:As an AI-forward engineer, you know that intelligence has become a commodity. Every organization has access to the same...  ...multi-agent orchestrationExperience deploying an LLM-powered system or AI agent into a production... 
    Full time
    Contract work
    Part time
    Local area
    Remote work

    Booz Allen Hamilton

    Arlington, VA
    3 days ago
  •  ...architecture, design patterns, and engineering standards for how we build and deploy AI agents and ML systems across cloud...  ...with large language models. Self-hosting open-weight models, fine-tuning, quantization...  ..., smaller models) in addition to LLM systems Ontology / semantic... 
    Full time

    Stellent IT LLC

    Arlington, VA
    2 days ago
  • $197.3k - $225.1k

     ...Machine Learning Engineer 4 (Python, AWS, SQL, GenAI) (Enterprise Platforms Technology) Do you love building and pioneering in the AI and technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative, inclusive, and iterative delivery... 
    Full time
    Part time
    H1b
    Local area

    Capital One National Association

    McLean, VA
    10 hours ago
  •  ...Capital One is hiring a Data Engineer 4 in McLean, VA to help drive a major transformation by delivering cloud-first data solutions. You will work with cross-functional teams to design and implement data pipelines, leverage Python and Spark, and engage with lakehouse architectures... 

    Jobleads-US

    McLean, VA
    9 hours ago
  • $197.3k - $225.1k

     ...Overview Full-stack Engineer 4 (Python, AWS) Do you love building and pioneering in the technology space? Do you enjoy solving complex...  ...within cloud environments ~ Experience leveraging modern AI-based coding tools and IDE copilots  ~1+ year of experience mentoring... 
    Full time
    Part time
    Internship
    H1b
    Local area

    Capital One

    McLean, VA
    10 days ago
  • $99k - $225k

     ...AI Engineer The Opportunity: As an AI engineer, you'll design and deliver production-grade AI systems. You'll own RAG pipelines, evaluation...  ..., including AI / ML applications ~ Experience building LLM systems using RAG frameworks such as LangChain and LlamaIndex,... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    BOOZ, ALLEN & HAMILTON, INC.

    Arlington, VA
    20 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Engineer 4 (LLM Gateway, FM Hosting). Be the first to apply!