Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform)

$229.9k - $262.4k

Capital One

Sr. Lead AI Engineer (Inference Optimization, FM Hosting, AI Platform)At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine learning — position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build.The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact.In this role, you will:Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One.Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.Invent and introduce state-of-the-art LLM optimization techniques to improve the performance — scalability, cost, latency, throughput — of large scale production AI systems.Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.The Ideal Candidate:You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good.Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production.You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven.You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enable you to see and exploit optimization opportunities that others miss.You are a resilient trail blazer who can forge new paths to achieve business goals when the route is unknown.Basic Qualifications:Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologiesAt least 6 years of experience programming with Python, Go, Scala, or JavaPreferred Qualifications:7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud)Experience designing, developing, integrating, delivering, and supporting complex AI systemsDemonstrated ability to lead and mentor an engineering team and influence cross-functional stakeholdersExperience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or GolangExperience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and costPassion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in productionExcellent communication and presentation skills, with the ability to articulate complex AI concepts to peersCapital One will consider sponsoring a new qualified applicant for employment authorization for this position.The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked.Cambridge, MA: $229,900 - $262,400 for Sr. Lead AI EngineerMcLean, VA: $229,900 - $262,400 for Sr. Lead AI EngineerNew York, NY: $250,800 - $286,200 for Sr. Lead AI EngineerSan Francisco, CA: $250,800 - $286,200 for Sr. Lead AI EngineerSan Jose, CA: $250,800 - $286,200 for Sr. Lead AI EngineerCandidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter.This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan.Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level.

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform) in San Jose, CA vacancy
  • $229.9k - $262.4k

    Senior Lead AI Engineer (FM Hosting, LLM Inference) Overview: At Capital One, we are creating...  ...customers. Our AI models and platforms empower teams across...  ...introduce state-of-the-art LLM optimization techniques to improve the...  ...$229,900 - $262,400 for Sr. Lead AI Engineer New... 
    Platform
    Senior
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    5 days ago
  • $229.9k - $262.4k

     ...Overview AI Engineer 5 (FM Hosting, LLM Inference) At Capital One, we are creating...  ...to deliver our industry leading capabilities with breakthrough...  ....  Our AI models and platforms empower teams across...  ...the-art foundation model optimization techniques to improve the... 
    Platform
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    13 days ago
  • $171.9k - $300.8k

     ...what data.Veza's Access Graph platform maps an organization's...  ...combination brings together Veza's AI-native Access Graph with...  ...environments, and AI agents. For engineers joining Veza today, this...  ...Exposure to LLM fine-tuning or inference optimization in production.Why join us... 
    Platform
    Senior
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    ServiceNow

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    NVIDIA is the platform upon which every new AI-powered application is built...  ...a Senior Software Engineer - AI Inference Performance to...  ...What you'll be doing:Lead end-to-end analysis...  ...decode workloads. Optimize time to first token...  ...Eliminate bottlenecks in host code, CUDA kernels,... 
    Platform
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 hours ago
  • $274k - $300k

     ...Description Saviynt's AI-powered identity platform manages and governs human...  ...and empower the world’s leading brands, Fortune 500 companies...  ...visit  AI Platform Engineer – Training & Inference Saviynt's AI-...  ...aware fallback between self-hosted SLMs and cloud LLMs •... 
    Platform

    Saviynt

    Milpitas, CA
    20 days ago
  •  ...discovery to powering AI and the technologies...  ...breakthroughs, or bringing leading edge products to...  ...motivated AI Model Optimization & Software Engineer Interns/Co-op to join...  ...training, fine-tuning, and inference across CPU, GPU, and accelerator platforms. Profile AI... 
    Platform
    Full time
    Summer work
    Internship
    Summer internship
    Worldwide

    AMD

    San Jose, CA
    3 days ago
  • $184k - $287.5k

     ...NVIDIA's DGX Cloud AI Efficiency Team...  ...developing tools for optimizing efficiency and...  ...training, post-training, inference. Our objective is...  ...software engineer to join our team....  ...underpinning NVIDIA's AI platforms.Define meaningful...  ...JAX, and RayNVIDIA leads the way in groundbreaking... 
    Platform
    Senior
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 hours ago
  • $255.65k

     ...can expect: As a Senior AI Software Engineer, you will collaborate to design, implement, and optimize AI algorithms and software applications...  ...will ensure AI training, inference, deployment, and operation...  ...the best collaboration platform for the enterprise, and today... 
    Platform
    Senior
    Work at office
    Remote work

    Pantera Capital

    San Jose, CA
    3 days ago
  • $160k - $198k

     ...advanced air mobility platform that delivers air taxis...  ...artificial intelligence ("AI") solutions, and other...  ...As a Senior AI Systems Engineer, you will architect,...  ...AI model training and inference. You will ensure our machine...  ...-to streamline and optimize the AI development lifecycle... 
    Platform
    Senior
    Local area
    Worldwide
    Visa sponsorship

    Archer

    San Jose, CA
    4 days ago
  •  ...quality of life. As an AI / Embedded Engineer, you will be...  ...model development, optimization, and deployment on embedded...  ...model size and inference latency ◦ Use frameworks...  ...and cloud or edge-hosted LLM components ◦...  ...Experience with RTOS platforms such as FreeRTOS or... 
    Platform
    Full time
    Work at office
    Immediate start
    Visa sponsorship
    Night shift

    E-Space

    Saratoga, CA
    a month ago
  • $229.9k - $262.4k

     ...Overview Sr. Manager, AI Engineer (IFX) At Capital One, we...  ...deliver our industry leading capabilities with breakthrough...  ...Our AI models and platforms empower teams across...  ...language model inference, similarity search,...  ...foundation model optimization techniques to improve... 
    Platform
    Senior
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    9 days ago
  • $234k - $286k

     ...leader in next-generation AI infrastructure, delivering a full-stack inference platform for customers...  ...Principal AI Solutions Engineer: a hands-on technical...  ...knowledge workers. You will lead hands-on engagements...  ...driven prompt and program optimization (e.g. DSPy-style),... 
    Platform
    Senior
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    SambaNova

    San Jose, CA
    2 days ago
  • $100k

     ...Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance...  ...models, compilers, platforms, networking, and...  ...-generation RISC-V CPUs optimized for AI, high-performance...  ...HPC & Agentic Software Engineering Lead, you will operate at... 
    Platform
    Full time
    Remote work

    Tenstorrent

    Santa Clara, CA
    21 days ago
  • $224k - $356.5k

     ...for a Senior Software Engineer to join our team! We...  ...by combining the best AI and graphics...  ...open-source modding platform for remastering classic...  ...content and controls.Optimize models and inference for latency, token throughput...  ...alongside rendering.Lead technical decisions,... 
    Platform
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...looking for a GenAI Engineer with strong...  ...high-performance inference services. The ideal...  ...enterprise GenAI platforms across GPU infrastructure...  ...Deploy, host, and manage Large...  ...and Ray Serve. Optimize model serving for...  ...efficiency. Develop AI platform services... 
    Platform
    Contract work

    2T Consulting

    Santa Clara, CA
    a month ago
  •  ...enterprises move beyond AI experimentation and...  ...through enterprise AI platforms, products, and services. We combine deep engineering expertise with AI innovation...  ...business outcomes.As a Lead Generative AI Engineer,...  ..., deployment, and optimization.Partner with technical... 
    Platform
    Senior

    Publicis Media

    San Jose, CA
    1 day ago
  • $184k - $287.5k

     ...unlimited potential of AI to define the next era...  ...of data center platform designs. From single node...  ...Grace CPUs, and a fully optimized NVIDIA AI and HPC software...  ...a highly motivated engineer to lead performance benchmarking...  ...-world AI training, inference, and HPC workloads at... 
    Platform
    Senior
    Full time
    Remote work

    NVIDIA

    Santa Clara, CA
    2 days ago
  • $242.2k

     ...watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and...  ...role Roku is seeking a Senior AI Transformation Engineer to join the AI Enablement Team and...  ...sources, pipelines, and metric layers Lead onboarding, enablement, and change... 
    Platform
    Senior
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    2 days ago
  • $200k - $270k

     ...Description Samsung SDS America AI Team is researching the...  ...partners, we use their platforms to develop complete embodied...  ...looking for a Senior Physical AI Engineer to join the team developing...  ..., or similar platforms Optimize low-latency inference and control systems for... 
    Platform
    Senior
    Worldwide
    Flexible hours

    Samsung SDS America

    Mountain View, CA
    a month ago
  • $180k - $250k

     ...a Cyber Security Engineer, with hands-on experience...  ...API, MosaicML Platform -Experience using...  ...-Experience using optimization tools such as...  ...Distillation, DeepSpeed-Inference /...  ...Global is hiring a AI Cyber Security Engineer...  ...Engineer Technical Lead, Identity Sunnyvale... 
    Platform
    Full time

    Insight Global

    San Jose, CA
    1 day ago
  • $123.24k - $200k

    Overview Of Role As a Sr./Principal AI Engineer within TSMC's Artificial Intelligence...  ...learning for manufacturing optimization. This role is for a builder...  .... Responsibilities Lead System Architecture: Own the...  ...new AI‑native products and platforms, from initial concept and data... 
    Platform
    Senior
    Work at office

    TSMC

    San Jose, CA
    6 days ago
  • $229.9k - $286.2k

    AI Engineer 5 (MLXT) At Capital One, we are creating responsible...  ...deliver our industry leading capabilities with...  .... Our AI models and platforms empower teams across...  ...large language model inference, agents and multi-...  ...art foundation model optimization techniques to improve... 
    Platform
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    3 days ago
  • $40 - $85 per hour

     ...is one of the world's leading entertainment...  ...and languages. The AI Platform team builds the infrastructure...  ...to GPU-optimized inference and serving. You'll...  ...Machine Learning, Computer Engineering, or a related field...  ...value and we strive to host a meaningful interview... 
    Platform
    Hourly pay
    Full time
    Internship
    Immediate start
    Remote work
    Flexible hours

    Netflix

    Los Gatos, CA
    6 hours ago
  •  ...computing experiences—from AI and data centers, to PCs, gaming...  ...: We are seeking a DevOps / Platform Engineer to join our team building...  ...communicate effectively and work optimally with their peers within our...  ...pods, CI pipelines, inference services, benchmarking jobs)... 
    Platform

    AMD

    San Jose, CA
    6 days ago
  • $2,500 per month

     ...heavily focused on inference . Backed by hundreds...  ...investors and staffed by leading engineers, Etched is redefining...  ...We are using AI to build AI chips. AI...  ...developer-facing internal platforms, CI/CD at scale, or infrastructure...  ...AlphaEvolve-style optimization loops that propose... 
    Platform
    Work at office
    Relocation package
    Night shift

    Etched

    San Jose, CA
    13 days ago
  • $229.9k - $262.4k

     ...Overview AI Engineer 5 (GenAI Platform, Agentic Infrastructure) At Capital...  ...teams to deliver our industry leading capabilities with...  ...training, large language model inference, agents and multi-agent workflows...  ...-the-art foundation model optimization techniques to improve the... 
    Platform
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    9 days ago
  •  ...Bitdeer is a world-leading technology company for AI and Bitcoin mining...  ...Cloud Senior DevOps Engineer to join our AI Cloud...  ...products and platforms are delivered with...  ...driving automation, optimizing cloud-native architectures...  ...vLLM, TGI, Triton Inference Server), and... 
    Platform
    Senior
    Remote job
    Full time
    Local area

    Bitdeer Technologies Group

    San Jose, CA
    a month ago
  • $124.36k - $146.3k

     ...Description Job Summary The Senior Engineer (Generative AI) is responsible for designing,...  ...deployment Performance tuning and optimization Support secure deployment, horizontal...  ...production environments 3. Cloud, Platform & Scalability Engineering Develop... 
    Platform
    Senior
    Full time
    Temporary work
    Work experience placement
    Local area
    3 days per week

    U.S. Bank

    Cupertino, CA
    11 hours ago
  • $2,500 per month

     ...heavily focused on inference . Backed by hundreds...  ...investors and staffed by leading engineers, Etched is redefining...  ...We are using AI to build and ship AI...  ...You will embed with Platform, Production, Supply Chain...  ...frontier models and system optimization methods. You may... 
    Platform
    Work at office
    Relocation package

    Etched

    San Jose, CA
    13 days ago
  • $248k - $391k

     ...unlimited potential of AI to define the next era...  ...Principal AI Workflow Engineer to join the Enterprise...  ..., adaptive, and self-optimizing integration strategies...  ...real-time integration platforms using event-driven, asynchronous...  ...as one of the world’s leading technology companies,... 
    Platform
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform). Be the first to apply!