AI Engineer 4 (LLM Gateway, FM Hosting)
$197.3k - $225.1kCapital One National Association
AI Engineer 4 (LLM Gateway, FM Hosting)
At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine learning — position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build.
Team Description:
The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact.
What You’ll Do:
- Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One.
- Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.
- Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more.
- Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance — scalability, cost, latency, throughput — of large scale production AI systems.
- Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
- Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment
- Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift
- Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines
- Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met
- Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation
The Ideal Candidate:
- You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good
- Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production
- You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven
- You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss
- You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown
Basic Qualifications:
- Bachelor's Degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies
- At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java
Preferred Qualifications:
- Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy
- 6+ years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud)
- Experience designing, developing, delivering, and supporting AI services
- Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang
- Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost
- Experience in building agentic AI systems and agentic workflows
- Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production
- Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale
- Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules
- Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms
Capital One will consider sponsoring a new qualified applicant for employment authorization for this position.
Cambridge, MA: $197,300 - $225,100 for AI Engineer 4McLean, VA: $197,300 - $225,100 for AI Engineer 4New York, NY: $215,200 - $245,600 for AI Engineer 4San Jose, CA: $215,200 - $245,600 for AI Engineer 4
Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate’s offer letter.
This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan.
Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website. Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level.
This role is expected to accept applications for a minimum of 5 business days.
Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections 4901-4920; New York City’s Fair Chance Act; Philadelphia’s Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries.
View email address on click.appcast.io
View email address on click.appcast.io
Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site.
Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
#J-18808-Ljbffr$229.9k - $262.4k
...AI Engineer 5 (FM Hosting, LLM Inference) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years... ...Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or...SuggestedFull timePart timeLocal area$229.9k - $262.4k
...creating responsible and reliable AI systems, changing banking for... ...‑class applied science and engineering teams to deliver our industry leading... ...related fields plus at least 4 years of experience developing... ...or technologies (e.g. LLM Inference, Similarity Search and...SuggestedFull timePart timeLocal area$229.9k - $262.4k
## Senior Lead AI Engineer (FM Hosting, LLM Inference)Applylocations: McLean, VA: New York, NY: San Jose, CAtime type: Full timeposted on: Posted Todayjob... ..., Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or...SuggestedFull timePart timeLocal area$229.9k - $262.4k
Senior Lead AI Engineer (FM Hosting, LLM Inference) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking... ...Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or...SuggestedFull timePart timeLocal area- ...Capital One is seeking an AI Engineer 5 to advance foundational AI systems within the IFX group in McLean, VA. You will design, implement, and optimize AI components including model training, inference, and orchestration, partnering with engineers, researchers, and PMs...Suggested
$229.9k - $262.4k
...AI Engineer 5 (AI Foundations, LLM Core and Agentic AI) At Capital One, we are creating responsible and reliable AI systems, changing banking for good... ..., Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies...Full timePart timeLocal area- ...Capital One is seeking an AI Engineer 4 focused on AI Foundations, LLM customization and finetuning to advance responsible AI in banking. You will design and deliver AI powered products with cross-functional teams. You will train models, optimize inference pipelines...
- ...defense technology company, is seeking an AI Implementation Engineer to be part of our Warfare Systems... ...available AI tooling/chat, hosting LLM servers and has demonstrated ability... ...NVIDIA preferred) Experience with AI Gateways Experience with LLM servers and model...Full timeWork at officeLocal areaImmediate startShift work
- ...Capital One is seeking an AI Engineer 5 to help build and scale responsible AI systems. You will collaborate with cross-functional teams on foundation models, LLM inference, and multi-model orchestration within a scalable AI infrastructure. The role emphasizes performance...
- ...Capital One invites applications for an AI Engineer 5 focused on AI Foundations and VLM customization. This role partners with cross‑functional... ...deploy AI software components, including foundation models, LLM inference, and multi‑model workflows. The ideal candidate...
- ...Capital One is seeking an AI Engineer 5 to build and optimize foundational AI systems within the Intelligent Foundations and Experiences... ...support AI components, including foundation model training and LLM inference, across a broad stack of open-source and SaaS technologies...
- ...Capital One is seeking an AI Engineer 4 to help build responsible AI systems and scalable AI infrastructure. You will design and deploy AI components, training pipelines, inference engines, and governance mechanisms while partnering with cross-functional teams across...
$135k - $150k
...Generative AI Application Engineer Location: United States (Remote... ...large language model (LLM) APIs — along with the... ...Work with models hosted on Hugging Face and with... ...production At least 4 years of hands-on Python... ....cpp, or similar). Gateway/routing layers like LiteLLM...Full timeTemporary workImmediate startRemote workFlexible hoursWeekend work- ...We build AI agents that actually work in enterprise... ...prototypes, not demos. We need engineers who can own the entire... ...that scales, and LLM integrations that are model... .../Fargate, Bedrock, API Gateway), Oracle OCI (OKE, Functions... ...on our engagements. ~4+ years of software...Temporary work
$150k - $200k
...Description Job Description Forward Deployed AI Engineer Location: Washington, DC (McLean, VA –... ...Deployed AI Engineer with approximately 4–6 years of experience to architect and... ..., tool-calling agents, or related LLM orchestration technologies • Experience...Full timeFor contractorsLocal areaImmediate startNight shift$75 - $80 per hour
...hr Direct message the job poster from Matlen Silver Award Winning Senior Technical Recruiter at Matlen Silver Job Title: Senior AI/LLM Engineer Duration: 12+ Months Location: Washington, DC Required Pay Scale: $75-$80/hour W2 ***Due to client requirements this role is...Full time$229.9k - $262.4k
...creating responsible and reliable AI systems, changing banking for... ...-class applied science and engineering teams to deliver our industry leading... ...related fields plus at least 4 years of experience developing... ...or technologies (e.g. LLM Inference, Similarity Search and...Full timePart timeLocal area- ID.me is hiring a Staff Software Development Engineer to lead the Tools Team in McLean, VA. You will champion AI-native engineering, architect production-grade solutions... ...in full-stack engineering, including hands-on LLM integration experience. This on-site role fosters...
$150k - $275k
...accomplishing hard things, together.Lead AI Security Engineer (Agentic SOC)Location: McLean/Tysons, VA... ...incident logs into agent prompts.LLM Performance Engineering: Continuously evaluate... ...Engineering, or related fields plus at least 4 years of experience developing AI and...Work experience placementLocal areaFlexible hours- ...Solutions Inc. is seeking a Lead AI Engineer / Architect to serve as the... ...Generative AI, Large Language Model (LLM), Retrieval-Augmented... ...integrated, including version, hosting approach, FedRAMP status, benchmarks... ...(HSA) Retirement Plan with 4% match and discretionary...Permanent employmentTemporary workInterim roleWork at officeFlexible hours
$179.4k - $204.7k
...Data Engineer 4 (Python, AWS, SQL, GenAI) (Enterprise Platforms Technology) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative, inclusive and iterative delivery environment? At Capital...InternshipH1bLocal area$145k - $205k
...computer science, software engineering, artificial... ...integrating generative AI or LLMs into production... ...orchestration frameworks, model gateways, MCP-compatible... ...open-source, or locally hosted LLMs in cloud, on-premises... ...actions. Implement LLM techniques including prompt...Full time$145k - $205k
...We are seeking a Senior AI/ML Software Engineer to integrate secure chatbot... ...engineering skills, practical LLM experience, and an understanding... ...frameworks, model gateways, MCP-compatible services, or... ...commercial, open-source, or locally hosted LLMs in cloud, on-premises,...$179.4k - $204.7k
...Full-stack Engineer 4, Agentic Technology Do you love building and pioneering in the... ...orchestration runtime that powers Capital One's AI content generation platform. You will... ...agentic pipeline with the enterprise LLM model gateway — manage API calls, handle rate...Full timePart timeInternshipH1bLocal area$99k - $225k
Agentic AI Forward-Deployed EngineerThe Opportunity:As an AI-forward engineer, you know that intelligence has become a commodity. Every organization has access to the same... ...multi-agent orchestrationExperience deploying an LLM-powered system or AI agent into a production...Full timeContract workPart timeLocal areaRemote work- ...architecture, design patterns, and engineering standards for how we build and deploy AI agents and ML systems across cloud... ...with large language models. Self-hosting open-weight models, fine-tuning, quantization... ..., smaller models) in addition to LLM systems Ontology / semantic...Full time
$197.3k - $225.1k
...Machine Learning Engineer 4 (Python, AWS, SQL, GenAI) (Enterprise Platforms Technology) Do you love building and pioneering in the AI and technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative, inclusive, and iterative delivery...Full timePart timeH1bLocal area- ...Capital One is hiring a Data Engineer 4 in McLean, VA to help drive a major transformation by delivering cloud-first data solutions. You will work with cross-functional teams to design and implement data pipelines, leverage Python and Spark, and engage with lakehouse architectures...
$197.3k - $225.1k
...Overview Full-stack Engineer 4 (Python, AWS) Do you love building and pioneering in the technology space? Do you enjoy solving complex... ...within cloud environments ~ Experience leveraging modern AI-based coding tools and IDE copilots ~1+ year of experience mentoring...Full timePart timeInternshipH1bLocal area$99k - $225k
...AI Engineer The Opportunity: As an AI engineer, you'll design and deliver production-grade AI systems. You'll own RAG pipelines, evaluation... ..., including AI / ML applications ~ Experience building LLM systems using RAG frameworks such as LangChain and LlamaIndex,...Full timeContract workPart timeWork at officeLocal areaRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Engineer 4 (LLM Gateway, FM Hosting). Be the first to apply!




