Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Engineer 5 (LLM Gateway, FM Hosting)

$229.9k - $262.4k

Jobleads-US

At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine learning — position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world‑class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build.

Team Description:

The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact.

What You’ll Do:

  • Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One.

  • Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.

  • Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more.

  • Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance — scalability, cost, latency, throughput — of large scale production AI systems.

  • Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.

  • Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems

  • Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency

  • Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards

  • Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity

The Ideal Candidate:

  • You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good

  • Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production

  • You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven

  • You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss

  • You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown

Basic Qualifications:

  • Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies

  • At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java

Preferred Qualifications:

  • Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy

  • 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud)

  • Experience designing, developing, delivering, and supporting complex AI systems

  • Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang

  • Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost

  • Experience in building agentic AI systems and agentic workflows

  • Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production

  • Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers

  • Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines

  • Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes

  • Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression

  • Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs)

Capital One will consider sponsoring a new qualified applicant for employment authorization for this position.

The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked.

Cambridge, MA: $229,900 - $262,400 for AI Engineer 5McLean, VA: $229,900 - $262,400 for AI Engineer 5New York, NY: $250,800 - $286,200 for AI Engineer 5San Jose, CA: $250,800 - $286,200 for AI Engineer 5

Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate’s offer letter.

This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan.

Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website. Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level.

This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug‑free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections 4901-4920; New York City’s Fair Chance Act; Philadelphia’s Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries.

If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at View phone number on click.appcast.io or via email at View email address on click.appcast.io. All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations.

For technical support or questions about Capital One's recruiting process, please send an email to View email address on click.appcast.io.

Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site.

Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).

#J-18808-Ljbffr Jobleads-US
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the AI Engineer 5 (LLM Gateway, FM Hosting) in San Jose, CA vacancy
  • $197.3k - $225.1k

     ...Overview AI Engineer 4 (LLM Gateway, FM Hosting) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For...  ...role is expected to accept applications for a minimum of 5 business days. No agencies please. Capital One is an... 
    Suggested
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    1 day ago
  • $229.9k - $262.4k

     ...Overview AI Engineer 5 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital...  ...developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory)... 
    Suggested
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    9 days ago
  • $184k - $287.5k

     ...unlimited potential of AI to define the next era of...  ...hiring senior software engineers in its Infrastructure, Planning...  ...Large Language Mode (LLM), Machine Learning (ML),...  .../NoSQL), and deployment/hosting (e.g., AWS, Azure, GCP)....  ...- 356,500 USD for Level 5.You will also be... 
    Suggested
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...NVIDIA Corporation’s Local AI team is building software that optimizes LLM inference on edge AI hardware. You’ll evaluate open-source frameworks, map models to GPU architecture, and drive performance characterization across multi-node configurations. You will own validation... 
    Suggested
    Local area

    Jobleads-US

    Santa Clara, CA
    1 day ago
  • $229.9k - $262.4k

     ...Overview AI Engineer 5 (MLX, Agentic AI, Gen AI platform Services) At Capital One, we are creating responsible and reliable AI systems...  ...developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory)... 
    Suggested
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    22 days ago
  • $229.9k - $262.4k

     ...Overview AI Engineer 5 (AI Foundations, VLM Customization) At Capital One, we are creating responsible and reliable AI systems, changing...  ...developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory)... 
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    22 days ago
  • $151.8k - $332.2k

    What you can expectAs an AI Engineer specializing in Agentic AI, you will develop intelligent agents...  ...and runtime operations.Building scalable LLM-powered systems including RAG pipelines,...  ...(or a Master's and 8+ years, or a PhD + 5 years).Show an understanding of machine learning... 
    Full time
    Work at office
    Remote work

    Zoom

    San Jose, CA
    2 days ago
  • $223k - $306.5k

     ...and Inclusion. We weave AI into the fabric of everything...  ...As a Sr Principal AI Engineer, you will join a dynamic...  ...of an "AI Gateway" intermediary layer to ensure...  ...complexities of individual LLM providers, providing centralized...  ...a Master's degree, or 5+ years with a PhD. ~... 
    Full time
    Work at office

    Jobleads-US

    Santa Clara, CA
    2 days ago
  • $216.3k - $280.8k

     ...at the cutting edge of AI, functioning within the...  ...broader Customer Experience Engineering organization as a...  ...hybrid cloud environments, LLM orchestration, and inference...  ...knowledge of AI API gateways and proxying solutions like...  ...up to 50% of quota;1.5% of incentive target for... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    3 days ago
  •  ...ServiceNow's Applied AI Forward Deployed Engineering (FDE) team builds production-grade, AI-native software for enterprise customers, delivering scalable LLM-enabled applications across backend, data orchestration, and frontend UI. You’ll work in the field with customers... 

    Jobleads-US

    Santa Clara, CA
    1 day ago
  •  ...Clara, CA seeks a deeply technical Senior Full-Stack Software Engineer to build next-gen AI platforms and products enhancing business efficiency. You...  ...role demands 8+ years in distributed systems, expertise in LLM-powered architectures, and hands-on Kubernetes/Docker... 

    Jobleads-US

    Santa Clara, CA
    2 days ago
  • $137.7k - $275.4k

    What you can expect​As a Video AI Engineer, you’ll enhance video codecs, video generation, and...  ...application layers for our distributed, cloud-hosted backend. Working alongside leading...  ...locationsAt Zoom, we offer a window of at least 5 days for you to apply because we believe... 
    Full time
    Work at office
    Remote work

    Zoom

    San Jose, CA
    4 days ago
  • $152k - $241.5k

    We are seeking a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters at NVIDIA to join our AI Efficiency...  ...across Robotics, Autonomous vehicles, LLM’s, Videos and moreCollaborate across...  ...area (or equivalent experience) Minimum 5+ years of experience designing and... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $272k - $431.25k

     ...framework for serving generative AI and reasoning models across multi-...  ...resilient deployment of cutting-edge LLM workloads.We are seeking a Principal Systems Engineer to define the vision and roadmap...  ...layer that spans GPU memory, pinned host memory, RDMA-accessible memory,... 
    Full time
    Local area
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $388k

     ...platforms, frameworks, or SDKs consumed by other engineering teams - not just point solutions. This...  ...with evaluation frameworks for LLM/agent quality.Generally, our compensation...  ...Inclusion is a Netflix value and we strive to host a meaningful interview experience for all... 
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    1 day ago
  • $125k - $175k

     ...Overview Blue Acorn iCi is seeking an AI Engineer to design and ship AI-powered solutions...  ...team. Required Skills & Experience ~5+ years of software engineering experience...  ...experiences. ~ Practical experience with LLM APIs such as Anthropic, OpenAI, or equivalent... 
    Full time
    Temporary work
    H1b
    3 days per week

    Blue Acorn iCi

    San Jose, CA
    a month ago
  • $250.6k - $362.6k

     ...be providedMeet the TeamThe CX AI Incubation and Agent team builds...  ...customers. We partner across engineering, product, and design to develop...  ...architecture and build core platform to host Agentic solutions.Lead...  ...Master’s degree with 12+ years).5+ years of engineering management... 
    Full time
    Temporary work
    Local area
    Relocation
    Flexible hours

    CISCO Systems

    San Jose, CA
    22 hours ago
  • $152k - $241.5k

    NVIDIA is looking for a Senior Applied AI Engineer to help build intelligent software systems that...  ...improvement.What we need to see:5+ years of proven experience or related field...  ...intelligent automation workflows.Experience with LLM-based applications, retrieval-augmented... 
    Full time

    Nvidia

    Santa Clara, CA
    15 hours ago
  • $128.6k - $184.9k

     ...the TeamJoin Cisco’s Enterprise AI team, the core group enabling...  ...and security —partnering across engineering, security, compliance, and product...  ...AI systems, including LLM-based applications, RAG pipelines...  ...attainment up to 50% of quota;1.5% of incentive target for each... 
    Full time
    Temporary work
    Work experience placement
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    1 day ago
  • $152k - $241.5k

    We are looking for a software engineer with a strong background in parallel...  ...at the intersection of AI, high-performance computing, and...  ...algorithms and software design.5+ years of relevant work or research...  ...with TensorRT, TensorRT-LLM, and cuTile.Experience parallelizing... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $136k - $218.5k

     ...into the unlimited potential of AI to define the next era of...  ...looking for dedicated Software Engineers to work on developing and deploying...  ...needed infrastructure to deploy LLM-powered solutions for engineering...  ...or equivalent experience.5+ years of proven industry experience... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $186.2k - $316.5k

     ...Our expert teams of physicists, engineers, data scientists and problem-...  ...Division is seeking a builder-first AI Engineering Architect to define...  ...with agentic AI, tool-using LLM workflows, multi-agent systems,...  ...and related work experience of 5 years; Master's Level Degree... 
    Minimum wage
    Full time
    Work experience placement
    Flexible hours

    KLA-Tencor

    Milpitas, CA
    4 days ago
  • $168k - $264.5k

     ...Design (SOCD) team is looking for an Applied AI Engineer who is passionate about eliminating...  ..., from RAG-grounded knowledge systems and LLM-powered assistants to multi-step agents that...  ..., and 196,000 USD - 310,500 USD for Level 5.You will also be eligible for equity and... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $109.7k - $175.5k

     ...AI Software Engineer Broadcom is seeking an experienced AI Software Engineer to integrate commercial...  ...for use with Large Language Model (LLM) agents and the MCP protocol Qualifications...  ...working on-site at the Broadcom office, 5 days a week. This is not a remote-work... 
    Work at office
    Local area
    Worldwide

    Broadcom Corporation

    San Jose, CA
    3 days ago
  • $194.3k - $243k

     ...CompanyWe are looking to hire a Sr. Staff AI Security Engineer to lead the secure deployment and...  ...Your mission: provide a secure foundation—gateways, orchestration layers, and tool ecosystems...  ...and operationalize the OWASP LLM Top 10 framework to categorize and mitigate... 
    Temporary work
    Work at office
    Remote work
    Flexible hours

    Bill.Com

    San Jose, CA
    1 day ago
  • $184k - $287.5k

     ...into the unlimited potential of AI to define the next era of...  ...looking for talented Software Engineers to work on developing and deploying...  ...decision systems (e.g., RL, planning, LLM-based agents) to optimize...  ...00 USD - 356,500 USD for Level 5.You will also be eligible for... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $114.1k - $214.95k

     ...Opportunity We are looking for a hands‑on AI Agent Engineer to develop, build, and maintain...  ...integrations. What you need to succeed ~3–5+ years of experience in AI/ML...  ...backend development ~ Prior exposure to LLM APIs or AI‑powered services in production... 
    Temporary work
    Local area
    Worldwide

    Adobe

    San Jose, CA
    1 day ago
  • $128.6k - $184.9k

     ...the TeamJoin Cisco’s Enterprise AI team, the core group enabling...  ...and security —partnering across engineering, security, compliance, and product...  ...AI systems, including LLM-based applications, RAG pipelines...  ...attainment up to 50% of quota;1.5% of incentive target for each... 
    Full time
    Temporary work
    Work experience placement
    Local area
    Flexible hours

    Cisco

    San Jose, CA
    2 days ago
  • $193.13k - $257.5k

    About Eightfold.ai:Eightfold is a global leader in AI-native enterprise...  ..., and high standards. Our engineers, product leaders, and go-to-...  ...equivalent years of experience.Min 5-7+ years of relevant work...  ...inference optimization (vLLM, TensorRT-LLM).Desired Skills & Experience:... 
    Work experience placement
    Work at office
    Remote work
    Flexible hours
    3 days per week

    Eightfold

    Santa Clara, CA
    3 days ago
  • $175.8k - $293k

     ...advantage. We’re looking for a Principal AI Engineer to architect, build, and harden the...  ...agent orchestration, RAG and grounding, gateways and routing) up to the agents and workflows...  ...on production traffic, calibrated LLM-as-a-judge graders, and A/B experiments... 

    Jobleads-US

    Santa Clara, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Engineer 5 (LLM Gateway, FM Hosting). Be the first to apply!