AI Engineer 4 (LLM Gateway, FM Hosting)
$197.3k - $225.1kMilitary Friendly Company
AI Engineer 4 (LLM Gateway, FM Hosting)
At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine learning — position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build.
Team Description:
The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact.
What You’ll Do:
Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One.
Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.
Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more.
Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance — scalability, cost, latency, throughput — of large scale production AI systems.
Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment
Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift
Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines
Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met
Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation
The Ideal Candidate:
You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good
Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production
You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven
You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss
You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown
Basic Qualifications:
Bachelor's Degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies
At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java
Preferred Qualifications:
Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy
6+ years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud)
Experience designing, developing, delivering, and supporting AI services
Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang
Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost
Experience in building agentic AI systems and agentic workflows
Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production
Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale
Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules
Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms
Capital One will consider sponsoring a new qualified applicant for employment authorization for this position.
The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked.
Cambridge, MA: $197,300 - $225,100 for AI Engineer 4
McLean, VA: $197,300 - $225,100 for AI Engineer 4
New York, NY: $215,200 - $245,600 for AI Engineer 4
San Jose, CA: $215,200 - $245,600 for AI Engineer 4
Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate’s offer letter.
This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan.
Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website ( . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level.
This role is expected to accept applications for a minimum of 5 business days.
No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections 4901-4920; New York City’s Fair Chance Act; Philadelphia’s Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries.
If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at View phone number on click.appcast.io or via email at View email address on click.appcast.io . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations.
For technical support or questions about Capital One's recruiting process, please send an email to View email address on click.appcast.io
Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site.
Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
#J-18808-Ljbffr$197.3k - $225.1k
Lead AI Engineer (AI Foundations, LLM Core and Agentic AI) Overview At Capital One, we are creating responsible and reliable AI systems, changing banking... ..., Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or...SuggestedFull timePart timeLocal area- ...through all the related job information below. AI Engineer | Data Science & AI Engineering Full-Time... ...learning, generative and agentic AI, and LLM-based techniques to real-world challenges,... ...colleagues. The Minimum Qualifications 4+ years of experience in data science, machine...SuggestedFull timeWork at office3 days per week
$121.6k - $194.5k
...The Role: Moderna is seeking an AI Engineer in Cambridge, MA to provide deep technical leadership... ...experience will be considered. ~4+ years of software, data, or AI engineering... ...experience architecting and operating LLM/GenAI systems, including RAG, agents/tool...SuggestedPermanent employmentWork from home$107k - $222.6k
...: Annually Role Join Suffolk's AI Studio in Boston as a core engineer transforming how AI powers construction... ...and templates. Qualifications ~4-6 years of professional software... ...systems using AWS Lambda, ECS/EKS, API Gateway, Step Functions, S3, CloudFront, and...SuggestedFull timeTemporary workFor contractorsWork at office$121.6k - $194.5k
...The Role Moderna is seeking an AI Engineer in Cambridge, MA to provide deep technical leadership... ...technical experience will be considered. 4+ years of software, data, or AI... ...Demonstrated experience architecting and operating LLM/GenAI systems, including RAG, agents/tool...SuggestedPermanent employmentWork from home- ...partnership with Nebius AI. Our team is fully remote... ...program for experienced engineers transitioning into... ...it to life for students: hosts the live sessions, reviews... ...Distributed Systems Module 4 — Cloud Infrastructure... ...Design & Integration (LLM integration, RAG, ML...Work at officeRemote workFlexible hours
$120k - $149k
...in which we practice.About the RoleThe AI Defense Engineer is a technical contributor who helps secure... ...across Copilot and other commercial LLM platforms.Adversarial Testing & Red... ...stacks, including SIEM, SOAR, EDR, WAF, API gateways, and identity platforms.Incident...Work experience placement- ...environments were never built to be AI-ready. They were built to... ...hit the accuracy cliff with an LLM and built around it instead of... ..., recurring patterns — back to engineering Develop implementation playbooks... ...'re Looking For Experience ~4–8 years combining hands-on...Full time
$197.3k - $225.1k
...Data Engineer 4 (Python/AWS) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative, inclusive and iterative delivery environment? At Capital One, you'll be part of a big group of makers...Full timePart timeInternshipH1bLocal area$160k - $220k
...builds the best-in-class Voice AI models powering the next... ...hiring our first Senior GTM AI Engineer . This is a software engineering... ...without you. What You'll Need ~4+ years as a software engineer... ...text | Speech Understanding | LLM Gateway Try the Playground Our $...Day shift$120k - $202.5k
...Analytics (CyberDNA) team is seeking a Senior AI Security Automation Engineer to help shape the next generation of... ...AI and Large Language Model (LLM) based solutions, including agentic workflows... ...in-office presence requirement of 2-4 days per week, consistent with the...Full timeTemporary workWork at officeLocal areaFlexible hoursShift work2 days per week$90.16k - $167.44k
AI EngineerJob Posting DescriptionThe AI Engineer will join the AI team supporting the Long-Term Care program in Manulife/John Hancock. This role will help design... ..., and anomaly detection.Implement GenAI and LLM-based solutions, including retrieval-augmented generation...Full timeTemporary workLocal area- ...That work changes the operating model an engineering organization runs on, the ways of working... ...sets the target. What we design from it is AI-native by construction. We redesign... ...enterprise use cases across one or more LLM platforms (e.g., OpenAI, Anthropic, or similar...Full timeWork experience placementLive inWork at officeLocal area
- ...matters at a company where you matter. AI Infrastructure Engineer, Corporate AI Team Team & Role Overview... ...deployment patterns, with a focus on Vercel-hosted applications and tools deployed across... ...evolving requirements. What You Bring 4+ years of experience in platform...
$135k - $169k
...Role The Artificial Intelligence Engineer is responsible for developing and implementing cutting-edge AI solutions and models to solve... ...effectiveness of the Large Language Models (LLM) solutions. What You Will Be... ...language models such as GPT-3/4, BERT, Dolly, LLaMA, PaLM2, etc...Full timeWork at officeShift work- AI/ML Engineer**** Please note: This role is not eligible for 100% remote work. Employees must live... ...and improve how people work* Develop LLM-powered applications using RAG, agents, prompt... ...learning algorithms into production.* 2-4 years of professional Managed or IT...Temporary workWork at officeLocal area3 days per week
$130k - $190k
...onsite / 2 days remote Role Summary As the technical engineer driving the company’s enterprise AI programs, you will design, build, and support AI... ...security, and compliance requirements. Qualifications ~4+ years in AI engineering, data science, or ML‑focused software...Full timeRemote work$150k - $240k
...TetraScience is the Scientific Data and AI Company building Tetra OS, the operating system... ...large continuous datasets and developing LLM agent harnesses, governance, and evaluation... ...engage directly with customers onsite up to 4-5 days per week in the Boston region. ~...Immediate startVisa sponsorship$175k - $225k
...work that defines Acadian. Position Overview: The Senior AI Engineer is a senior technical leader responsible for defining,... ..., agentic deep research, retrieval-augmented generation (RAG), LLM training, inference optimization, model integration, evaluations...Casual workWork at officeWorldwideFlexible hours3 days per week$101k - $194k
...? Join the #VTeamLife.What you'll be doing…As a Senior Agentic AI Engineer within our team, you will be at the forefront of Verizon's transformation... ...science and machine learning, and you've extended that into LLM orchestration, multi-agent architecture, and production agentic...Full timeTemporary workPart timeWork experience placementWork at officeWork from homeShift work3 days per week$137.35k - $206.09k
...Design, build, and operate the agentic and LLM-powered systems that biologics scientists... ...the interfaces and the supporting engineering (authentication, logging, evaluation, error... ...end GPU compute (including AZ’s sovereign AI compute platform built on NVIDIA DGX SuperPOD...Full timeTemporary workWork at officeRemote work$197.3k - $225.1k
...Full-stack Engineer 4 (Python, Typescript, React, AWS) Do you love building and pioneering in the technology space? Do you enjoy solving... ...issues within cloud environments ~ Experience leveraging modern AI-based coding tools and IDE copilots ~1+ year of experience...Full timePart timeInternshipH1bLocal area- ...Subject Matter Expert (SME)/Technical Lead Engineer (TLE) Level Test & Verification Engineers... ...initiative involving CI/CD pipelines and AI tools. \tThis position will be physically... ...for avionics platforms. \tSenior (Level 4) \tLeads the setup and configuration of...Full timeWork experience placementRemote workRelocation
$200k - $325k
Job Description We are building an AI Engineering function to enable productivity and agentic... ...the core AI platform, including managed LLM inference services (Amazon Bedrock and related... ...including MCP servers, an MCP registry/gateway, and authorization services that connect...Work at officeLocal area$229.9k - $262.4k
...creating responsible and reliable AI systems, changing banking for... ...-class applied science and engineering teams to deliver our industry leading... ...related fields plus at least 4 years of experience developing... ...or technologies (e.g. LLM Inference, Similarity Search and...Full timePart timeLocal area$170k - $240k
...clinical trial expertise with cutting-edge AI, we connect sponsors' scientific... ...that sits at the intersection of product, engineering, and operations. You will partner closely... ...workflows and agentic systems leveraging modern LLM platforms, orchestration frameworks, and...Temporary workWork at officeImmediate start2 days per week$175k - $210k
Role OverviewQuEra is standing up a new AI Engineering team to help every group in the company put AI to work. We are hiring senior founding... ...-wide AI workshop into real, deployable tools, and build LLM- and agent-powered software that makes our slow, expensive, and...$15 - $20 per hour
Job Description Insight Global is sourcing for a Python/AI Engineer to join a global consulting firm, supporting various internal IT & Operational... ...(Multi agentic architecture, Agents orchestration, RAG, LLM evaluations, APIs development etc.). Excellent communication...Remote work$191k - $305k
About this role:Wells Fargo is seeking a Principal AI Engineer to join the CCIBT Gen AI team, which is responsible for building AI frameworks... ...technologies Collaborate with enterprise teams to integrate LLM models with the existing CCIBT products and systemsLead the design...Full timeWork experience placement$115k - $159k
...by forward-thinking peers.The AI Developer is a member of the Digital... ...first and foremost a software engineering role, requiring solid... ...education and experience.* Minimum 2-4 years of professional software... ...automation technologies.* Exposure to LLM APIs or AI SDKs. Experience...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Engineer 4 (LLM Gateway, FM Hosting). Be the first to apply!




