AI Engineer 5 (LLM Gateway, FM Hosting)
$250.8k - $286.2kCapital One Bank
At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine learning — position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build.
Team Description:
The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact.
What You’ll Do:
Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One.
Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.
Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more.
Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance — scalability, cost, latency, throughput — of large scale production AI systems.
Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems
Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency
Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards
Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity
The Ideal Candidate:
You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good
Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production
You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven
You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss
You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown
Basic Qualifications:
Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies
At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java
Preferred Qualifications:
Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy
7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud)
Experience designing, developing, delivering, and supporting complex AI systems
Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang
Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost
Experience in building agentic AI systems and agentic workflows
Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production
Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers
Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines
Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes
Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression
Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs)
Capital One will consider sponsoring a new qualified applicant for employment authorization for this position.
The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked.
Cambridge, MA: $229,900 - $262,400 for AI Engineer 5McLean, VA: $229,900 - $262,400 for AI Engineer 5New York, NY: $250,800 - $286,200 for AI Engineer 5San Jose, CA: $250,800 - $286,200 for AI Engineer 5Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate’s offer letter.
This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan.Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website. Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level.
This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections 4901-4920; New York City’s Fair Chance Act; Philadelphia’s Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries.If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at View phone number on aiapply.co or via email at View email address on aiapply.co. All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations.
For technical support or questions about Capital One's recruiting process, please send an email to View email address on aiapply.co
Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site.
Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
$229.9k - $262.4k
...Senior Lead AI Engineer (LLM Gateway, FM Hosting) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good... ...role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal...SuggestedFull timePart timeLocal area$229.9k - $286.2k
AI Engineer 5 (FM Hosting, LLM Inference) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences...SuggestedFull timePart timeLocal area$229.9k - $262.4k
...Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform) Overview: At Capital One, we are creating responsible... ...and introduce state-of-the-art LLM optimization techniques to improve the... ...applications for a minimum of 5 business days.No agencies please. Capital...SuggestedFull timePart timeLocal area$250.8k - $286.2k
AI Engineer 5 ((AI Foundations, LLM Core and Agentic AI) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized...SuggestedFull timePart timeLocal area$112.8k - $257k
AI Gateway Security Engineer, LeadThe Opportunity:Designs, deploys, integrates, and maintains enterprise AI gateway capabilities across complex, multi... ..., and machine-to-machine authentication supporting LLM inference and agent-based architectures. Works without considerable...SuggestedFull timeContract workPart timeWork at officeLocal areaRemote work$250.8k - $286.2k
AI Engineer 5 (AI Foundations, VLM Customization) At Capital One, we are creating responsible and reliable AI systems, changing banking for... ...Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory)...Full timePart timeLocal area$114k - $231k
...technology company, is seeking an AI Implementation Engineer to be part of our Warfare... ...available AI tooling/chat, hosting LLM servers and has... ...technical field highly desired 5-10+ years of professional... ...preferred) Experience with AI Gateways Experience with LLM servers...Full timeWork at officeLocal areaImmediate startShift work$229.9k - $286.2k
AI Engineer 5 (Gen AI Platform Services - Agentic AI) Overview: At Capital One, we are creating responsible and reliable AI systems, changing... ...developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory)...Full timePart timeLocal area- ...technology company, is seeking an AI-Native Software Engineer to be part of our Warfare... ...Do:Use generative AI tools (LLM copilots, autonomous coding... ...professional experience.5-7 years of full-time professional... ...deploying or operating self-hosted LLMs in secure...Full timeWork at officeImmediate start
- ...defense technology company, is seeking an AI Implementation Engineer to be part of our Warfare Systems team... ...available AI tooling/chat, hosting LLM servers and has demonstrated ability to... ...other technical field highly desired 5-10+ years of professional experience...Full timeWork at officeLocal areaImmediate startShift work
- Position: Consultant, AI Software Engineering Location:... ...quantitative or technical field.2-5 years of experience in... ...and workflows.Experience with LLM platforms and services such as... ...services such as Lambda, API Gateway, S3, databases, IAM, CloudWatch...Permanent employmentTemporary workH1bRelocation
$71.5k - $190k
AI Engineering And Delivery Lead Systems Planning and Analysis... ...Foundry/OpenAI), self-hosted inference on AKS, and local... ..., DLP policies, gateway policies, and alignment... ...) and OWASP Top 10 for LLM Applications. Infrastructure... ..., IT, or related field 5-7 years in enterprise...Contract workWork at officeLocal areaRemote work$84.4k - $156.8k
...configure, secure, and maintain AI application environments... .../Oracle Kubernetes Engine (OKE), OCI Container... ...application access to approved LLM services, APIs,... ...experience may be considered.5+ years of hands-on... ...OCI Functions, and API Gateway.Experience integrating containerized...Minimum wageFull timeContract workFlexible hours- ...what is delivered. Our AI-native platform, Air Enterprise... ...experienced Senior AI Engineer specializing in AI... ...Required Skills: ~5+ years of experience building... ...understanding of how modern LLM and agentic systems... ...systems, vector stores, model gateways, or other components of...Full timeWork at officeRemote work
$75 - $80 per hour
...hr Direct message the job poster from Matlen Silver Award Winning Senior Technical Recruiter at Matlen Silver Job Title: Senior AI/LLM Engineer Duration: 12+ Months Location: Washington, DC Required Pay Scale: $75-$80/hour W2 ***Due to client requirements this role is...Full time- ...AI Engineer Location: McLean, VA Job Type: Full-Time Experience: 5+ Years Job Summary We are looking for an experienced AI Engineer to... ...AI solutions . Build and integrate LLM-based applications, RAG pipelines, embeddings...Full time
- ...AI Red Team Engineer SEI conducts research and development in software engineering, systems engineering... ...contexts. But this isn't a "make the LLM say the bad thing" type of AI red team.... ...technical/engineering field with five (5) years of experience; PhD in computer...Full timePart timeWork experience placementWork at office
- ...citizen services. We’re a nonprofit engineering, applied research, and advanced... ...and help shape the future.Our AI-enhanced Discovery and... ...areas: large language models (LLM), generative and agentic AI, decision... ...Qualifications:Requires a minimum of 5 years of related experience...InternshipLocal areaImmediate start3 days per week
$150k - $180k
...contact us at .Job DescriptionThe AI & Digital Engineering Software Engineer III... ...DMS, Kinesis, EventBridge, API Gateway, and Lambda.Parse, normalize,... ...methodology (golden datasets, LLM-as-judge, red-teaming, hallucination... ...office schedule are 8:00am-5:00pm ET, Mon-Fri...Full timeWork at office$220k - $245k
...contact us at .Job DescriptionThe AI & Digital Engineering Engineer contributes to X-... ..., Kinesis, EventBridge, API Gateway, and Lambda.Parse, normalize,... ...(golden datasets, LLM-as-judge, red-teaming, hallucination... ...Standard office schedule are 8:00am-5:00pm ET, Mon-Fri...Full timeWork at office$150k - $200k
...Description Job Description Forward Deployed AI Engineer Location: Washington, DC (McLean, VA – on-site 5 days/week) Compensation: $150,000 – $200,000 +... ...agent frameworks, tool-calling agents, or related LLM orchestration technologies • Experience building...Full timeFor contractorsLocal areaImmediate startNight shift$180k - $205k
...contact us at .Job DescriptionThe AI & Digital Engineering Software Engineer IV... ...DMS, Kinesis, EventBridge, API Gateway, and Lambda.Parse, normalize,... ...methodology (golden datasets, LLM-as-judge, red-teaming, hallucination... ...office schedule are 8:00am-5:00pm ET, Mon-Fri...Full timeWork experience placementWork at office$99k - $225k
Enterprise Agentic AI EngineerThe Opportunity:As an experienced engineer, you know that agentic AI is transforming how organizations... ...us. The world can’t wait.You Have:5+ years of experience supporting... ...interaction patternsKnowledge of LLM platforms and providers, including...Full timeContract workPart timeWork at officeLocal areaRemote workShift work$170k - $265k
...enabling human life on Mars.SR. AI SECURITY SOFTWARE ENGINEER (STARSHIELD)Starshield... ...observation, communications, and hosted payloads. The Starshield... ...another STEM discipline AND 5+ years of professional experience... ...solutions with LLMs, LLM-based agents, or other AI topicsPREFERRED...Permanent employmentTemporary workImmediate startFlexible hoursWeekend work$100k - $200k
...dedicated to accomplishing hard things, together.AI Security Engineer (Agentic SOC)Location: McLean, VA | On-site (5 days/week)Are you ready to build the security systems... ...and historical incident logs into agent prompts.LLM Performance Engineering: Continuously evaluate,...Work experience placementLocal areaFlexible hours- ...needs and what is delivered. Our AI-native platform, Air Enterprise... ...an experienced Senior AI Engineer to join our Agentic AI team as... ...required Required Skills: ~5+ years of experience building... ...Strong understanding of modern LLM systems, including model inference...Full timeWork at officeRemote work
$150k - $190k
...innovation meets mission. Our AI, cloud, cyber, and modernization... ...highly skilled Senior Google AI Engineer. We are growing our Google... ..., Model Monitoring) and Gemini/LLM services-optimized for performance... ...substantial handson experience (5+ years) in AI/ML solution development...Temporary workImmediate startWorldwide$176k
...Current PhD, AI Engineering Internship Program - Summer 2027 Key Role Details This is a full-... ...work with AI/ML domains and systems (e.g., LLM Inference, Search and Retrieval, NLP,... ...to accept applications for a minimum of 5 business days. No agencies please. Capital...Full timePart timeSummer workInternshipLocal area$209k - $238.5k
...Sr. Lead AI Engineer Overview: At Capital One, we are creating responsible and reliable AI systems... .... Invent and introduce state-of-the-art LLM optimization techniques to improve the performance... ...to accept applications for a minimum of 5 business days.No agencies please. Capital...Full timePart timeLocal area$150k - $300k
...accomplishing hard things, together.Lead Applied AI EngineerGrow your engineering career building the AI systems that... ...are expected to be in the office 5 days a week to foster that culture... ...intelligence and retrieval, and rigorous LLM evaluation. These skills are shared...Permanent employmentWork experience placementWork at officeLocal areaFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Engineer 5 (LLM Gateway, FM Hosting). Be the first to apply!



