Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Engineer 5 (LLM Gateway, FM Hosting)

$229.9k - $262.4k

Capital One

San Jose, CA

Overview

AI Engineer 5 (LLM Gateway, FM Hosting)

At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine learning — position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build.

Team Description:

The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers.  Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact.

What You’ll Do: 

  • Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One.

  • Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.

  • Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more.

  • Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance — scalability, cost, latency, throughput — of large scale production AI systems.

  • Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.

  • Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems

  • Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency

  • Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards

  • Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity

The Ideal Candidate: 

  • You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good

  • Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production

  • You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven

  • You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss

  • You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown

Basic Qualifications: 

  • Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies

  • At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java

Preferred Qualifications: 

  • Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy

  • 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud)

  • Experience designing, developing, delivering, and supporting complex AI systems

  • Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang

  • Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost

  • Experience in building agentic AI systems and agentic workflows

  • Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production

  • Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers

  • Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines

  • Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes

  • Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression

  • Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs)

Capital One will consider sponsoring a new qualified applicant for employment authorization for this position.

The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked.

Cambridge, MA: $229,900 - $262,400 for AI Engineer 5

McLean, VA: $229,900 - $262,400 for AI Engineer 5

New York, NY: $250,800 - $286,200 for AI Engineer 5

San Jose, CA: $250,800 - $286,200 for AI Engineer 5

Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate’s offer letter.

This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan.

Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level.

This role is expected to accept applications for a minimum of 5 business days.

No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections 4901-4920; New York City’s Fair Chance Act; Philadelphia’s Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries.

If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at Show phone number or via email at View email address on capitalonecareers.com . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations.

For technical support or questions about Capital One's recruiting process, please send an email to View email address on capitalonecareers.com

Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site.

Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).

Vacancy posted 7 days ago
Similar jobs that could be interesting for youBased on the AI Engineer 5 (LLM Gateway, FM Hosting) in San Jose, CA vacancy
  • $229.9k - $286.2k

    AI Engineer 5 (FM Hosting, LLM Inference) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences... 
    Suggested
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    1 day ago
  • $229.9k - $262.4k

    Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform) Overview: At Capital One, we are creating responsible...  ...and introduce state-of-the-art LLM optimization techniques to improve the...  ...applications for a minimum of 5 business days.No agencies please. Capital... 
    Suggested
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    1 day ago
  • $197.3k - $225.1k

    Lead AI Engineer (AI Foundations, LLM Core and Agentic AI) Overview At Capital One, we are creating responsible and reliable AI systems, changing banking...  ...role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal... 
    Suggested
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    2 days ago
  • $209k - $286.2k

    AI Engineer 5 (Gen AI Platform Services) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for...  ...developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory)... 
    Suggested
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    2 days ago
  • $184k - $287.5k

     ...unlimited potential of AI to define the next era of...  ...hiring senior software engineers in its Infrastructure, Planning...  ...Large Language Mode (LLM), Machine Learning (ML),...  .../NoSQL), and deployment/hosting (e.g., AWS, Azure, GCP)....  ...- 356,500 USD for Level 5.You will also be... 
    Suggested
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    2 days ago
  • $131.54k - $197k

     ...Cybersecurity Engineer For AI Security Marvell's semiconductor solutions are the essential building...  ...What We're Looking For ~5 to 8 years in cybersecurity engineering...  ...Familiarity with AI platform components including LLM gateways, MCP frameworks, RAG architecture, and... 
    Permanent employment
    Internship
    Work from home

    Marvell

    Santa Clara, CA
    1 day ago
  • $229.9k - $262.4k

     ...Overview AI Engineer 5 (MLX, Agentic AI, Gen AI platform Services) At Capital One, we are creating responsible and reliable AI systems...  ...developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory)... 
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    7 days ago
  • $229.9k - $262.4k

     ...Overview AI Engineer 5 (AI Foundations) At Capital One, we are creating responsible and reliable AI systems, changing banking for good...  ...developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory)... 
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    7 days ago
  • $151.8k - $332.2k

    What you can expect We are looking for an AI Inference Engineer with a solid background in speech...  ...experience in speech recognition, speech-llm or AI model inference.Display knowledge...  ...locationsAt Zoom, we offer a window of at least 5 days for you to apply because we... 
    Full time
    Work at office
    Remote work

    Zoom

    San Jose, CA
    1 day ago
  •  ...Description In this AI/ML ASIC Performance Engineer position, you will develop AI...  ...CPUs on the device side and Host CPUs. Author architecture...  ...during the execution phase. LLM Workload analysis and...  ...; (4) geographic location; (5) shift; (6) internal and external... 
    Full time
    Temporary work
    Remote work
    Flexible hours
    Shift work
    Night shift

    Sandisk

    Milpitas, CA
    14 days ago
  • $216.3k - $280.8k

     ...at the cutting edge of AI, functioning within the...  ...broader Customer Experience Engineering organization as a...  ...hybrid cloud environments, LLM orchestration, and inference...  ...knowledge of AI API gateways and proxying solutions like...  ...up to 50% of quota;1.5% of incentive target for... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    3 days ago
  • $137.7k - $275.4k

    What you can expect​As a Video AI Engineer, you’ll enhance video codecs, video generation, and...  ...application layers for our distributed, cloud-hosted backend. Working alongside leading...  ...locationsAt Zoom, we offer a window of at least 5 days for you to apply because we believe... 
    Full time
    Work at office
    Remote work

    Zoom

    San Jose, CA
    4 days ago
  • $152k - $241.5k

    We are seeking a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters at NVIDIA to join our AI Efficiency...  ...across Robotics, Autonomous vehicles, LLM’s, Videos and moreCollaborate across...  ...area (or equivalent experience) Minimum 5+ years of experience designing and... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $388k

     ...Software Engineer 5 - Ads Signals Conversion API New York, New York, United States of America • Seattle, Washington, United States of...  ...flexible time off. Inclusion is a Netflix value and we strive to host a meaningful interview experience for all candidates. If you... 
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    2 days ago
  • $125k - $175k

     ...Overview Blue Acorn iCi is seeking an AI Engineer to design and ship AI-powered solutions...  ...team. Required Skills & Experience ~5+ years of software engineering experience...  ...experiences. ~ Practical experience with LLM APIs such as Anthropic, OpenAI, or equivalent... 
    Full time
    Temporary work
    H1b
    3 days per week

    Blue Acorn iCi

    San Jose, CA
    a month ago
  • $272k - $431.25k

     ...framework for serving generative AI and reasoning models across multi-...  ...resilient deployment of cutting-edge LLM workloads.We are seeking a Principal Systems Engineer to define the vision and roadmap...  ...layer that spans GPU memory, pinned host memory, RDMA-accessible memory,... 
    Full time
    Local area
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $136k - $218.5k

     ...into the unlimited potential of AI to define the next era of...  ...looking for dedicated Software Engineers to work on developing and deploying...  ...needed infrastructure to deploy LLM-powered solutions for engineering...  ...or equivalent experience.5+ years of proven industry experience... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $250.6k - $362.6k

     ...be providedMeet the TeamThe CX AI Incubation and Agent team builds...  ...customers. We partner across engineering, product, and design to develop...  ...architecture and build core platform to host Agentic solutions.Lead...  ...Master’s degree with 12+ years).5+ years of engineering management... 
    Full time
    Temporary work
    Local area
    Relocation
    Flexible hours

    CISCO Systems

    San Jose, CA
    4 days ago
  • $168k - $264.5k

     ...Design (SOCD) team is looking for an Applied AI Engineer who is passionate about eliminating...  ..., from RAG-grounded knowledge systems and LLM-powered assistants to multi-step agents that...  ..., and 196,000 USD - 310,500 USD for Level 5.You will also be eligible for equity and... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $186.2k - $316.5k

     ...Our expert teams of physicists, engineers, data scientists and problem-...  ...Division is seeking a builder-first AI Engineering Architect to define...  ...with agentic AI, tool-using LLM workflows, multi-agent systems,...  ...and related work experience of 5 years; Master's Level Degree... 
    Minimum wage
    Full time
    Work experience placement
    Flexible hours

    KLA-Tencor

    Milpitas, CA
    4 days ago
  • $184k - $287.5k

     ...performance computing, gaming and AI. Our GPUs and SOCs give...  ...generation alone! Now we're hiring the engineer who will lead the rebuild of...  ...pipelines, including at least one LLM-backed system that SMEs depend...  ...00 USD - 356,500 USD for Level 5.You will also be eligible for... 
    Full time
    Immediate start

    Nvidia

    Santa Clara, CA
    1 day ago
  • $128.6k - $184.9k

     ...the TeamJoin Cisco’s Enterprise AI team, the core group enabling...  ...and security —partnering across engineering, security, compliance, and product...  ...AI systems, including LLM-based applications, RAG pipelines...  ...attainment up to 50% of quota;1.5% of incentive target for each... 
    Full time
    Temporary work
    Work experience placement
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    17 hours ago
  • $109.7k - $175.5k

     ...AI Software Engineer Broadcom is seeking an experienced AI Software Engineer to integrate commercial...  ...for use with Large Language Model (LLM) agents and the MCP protocol Qualifications...  ...working on-site at the Broadcom office, 5 days a week. This is not a remote-work... 
    Work at office
    Local area
    Worldwide

    Broadcom Corporation

    San Jose, CA
    3 days ago
  • $152k - $241.5k

    We are looking for a software engineer with a strong background in parallel...  ...at the intersection of AI, high-performance computing, and...  ...algorithms and software design.5+ years of relevant work or research...  ...with TensorRT, TensorRT-LLM, and cuTile.Experience parallelizing... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $249k

     ...for travelers everywhere.Principal Data & AI Engineer, Reporting and InsightsIntroduction to the...  ...reporting solutions. Design and implement LLM-powered reporting workflows, including prompt...  ...platform. Preferred qualifications: 5+ years of experience leading cross-functional... 
    Full time
    Work at office

    Expedia

    San Jose, CA
    4 days ago
  • $197.3k - $225.1k

    Lead AI Engineer At Capital One, we are creating responsible and reliable AI systems, changing...  ...Invent and introduce state-of-the-art LLM optimization techniques to improve the performance...  ...to accept applications for a minimum of 5 business days.No agencies please. Capital... 
    Full time
    Part time
    Local area

    Capital One Financial Corp

    San Jose, CA
    5 days ago
  • $152k - $241.5k

     ...NVIDIA is looking for a Senior Applied AI Engineer to help build intelligent software systems...  ...continuous improvement. What we need to see: ~5+ years of proven experience or related...  ...automation workflows. ~ Experience with LLM-based applications, retrieval-augmented... 
    Full time

    NVIDIA

    Santa Clara, CA
    3 days ago
  • $193.13k - $257.5k

    About Eightfold.ai:Eightfold is a global leader in AI-native enterprise...  ..., and high standards. Our engineers, product leaders, and go-to-...  ...equivalent years of experience.Min 5-7+ years of relevant work...  ...inference optimization (vLLM, TensorRT-LLM).Desired Skills & Experience:... 
    Work experience placement
    Work at office
    Remote work
    Flexible hours
    3 days per week

    Eightfold

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...product that powers innovative AI research and developers. We focus...  ...an AI infrastructure software engineer to join our team. You'll be...  ...and tools for large-scale AI, LLM, and GenAI infrastructure.Develop...  ...00 USD - 356,500 USD for Level 5.You will also be eligible for... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $190k - $237k

     ...our team members.What You’ll DoAs a Staff AI Systems Engineer, you will architect, deploy, and manage...  ...compute scheduling alongside advanced LLM serving engines.Cross-Functional Collaboration...  ..., or a related field.Experience: 5+ years of professional software engineering... 
    Local area

    Archer Aviation

    San Jose, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Engineer 5 (LLM Gateway, FM Hosting). Be the first to apply!