Lead AI Engineer (FM Hosting, LLM Inference)
$197.3k - $225.1kCapital One Financial Corp
Lead AI Engineer (FM Hosting, LLM Inference)
Overview At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. In this role, you will:- Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One.
- Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.
- Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
- Invent and introduce state-of-the-art LLM optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems.
- Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
- You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good.
- Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production.
- You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven.
- You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enable you to see and exploit optimization opportunities that others miss.
- You are a resilient trail blazer who can forge new paths to achieve business goals when the route is unknown.
- Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies
- At least 4 years of experience programming with Python, Go, Scala, or Java
- 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud)
- Experience designing, developing, delivering, and supporting AI services
- Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang
- Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost
- Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Lead AI Engineer (FM Hosting, LLM Inference) in San Francisco, CA vacancy
$300k - $400k
...reinvent how designers work in the AI era. We’re backed by top... ...the Role We’re hiring an Lead AI Engineer to own and scale our AI infrastructure... ...and ship production-grade, LLM-powered features, define... ...Do Own the training-to-inference pipeline for large code...SuggestedFull time$190k - $270k
...We AreHP IQ is HP’s new AI innovation lab.... ...diverse, world-class team—engineers, designers, researchers... .... We are looking for a Lead Software Engineer to design... ...solutions for real-time AI inference and processing.... ...or Python.Proficient in LLM integration into multi-...SuggestedFull timeTemporary workLocal areaFlexible hours$197.3k - $225.1k
...Overview Lead AI Engineer (MLX, Agentic AI, Gen AI platform Services) At Capital One, we... ...foundation model training, large language model inference, similarity search, guardrails, model... ...Invent and introduce state-of-the-art LLM optimization techniques to improve the...SuggestedFull timePart timeLocal area$215k - $260k
...the only vertically integrated AI infrastructure company built... ...production. That means owning the inference stack end to end: profiling... ...also work directly with customer engineering teams to tailor deployments to... ...inference.Comfort with modern LLM serving frameworks such as...SuggestedTemporary work$190k - $270k
...HP IQ is HP's new AI innovation lab. Combining... ...diverse, world-class team-engineers, designers, researchers... .... We are looking for a Lead Software Engineer to design... ...for real-time AI inference and processing. Implement... .... ~ Proficient in LLM integration into multi-...SuggestedFull timeTemporary workLocal areaFlexible hours$250k
...career? Join a rapidly growing AI cloud infrastructure provider... ...large-scale AI training and inference workloads. With expanding GPU... ...As a Senior ML Infrastructure Engineer, the successful candidate will... ...using vLLM, SGLang, TensorRT‑LLM, or Triton Knowledge of GPU...$229.9k - $262.4k
...Senior Lead AI Engineer (Gen AI Platform Services, Agentic AI) Overview At Capital One, we... ...foundation model training, large language model inference, similarity search, guardrails, model... ..., and more. Invent state‑of‑the‑art LLM optimization techniques to improve...Local area$166.5k - $266.2k
...something unprecedented — an AI foundation that will push the... ...areas.The Forward Deployed AI Engineer is the connective tissue between... ...petabytes of multi-omics data, leading end-to-end deployments from... ...reliability, and on-call readinessApply LLM, retrieval-augmented...Flexible hours$203.5k
General Information Job Title Lead, AI Engineering Job ID 102641 Work Areas Analytics, Data & Research, Management Consulting... ...experimentation frameworks, and data labeling strategies for LLM applicationsExperience with:RAG architectures (vector-based retrieval...Permanent employmentFull timeApprenticeshipWork at officeLocal areaWork from homeHome office3 days per week- ...people interact with the web by building AI agents that can reliably do everyday digital... ...models) Scale infra for agentic inference (throughput and latency of perception‑planning... ...web‑agent Work closely with product engineers to translate cutting‑edge AI capabilities...Work at officeRelocationVisa sponsorship
$190k - $265k
...about enabling data and AI teams to solve the... ...their business. Founded by engineers — and customer-obsessed... ...across real-time and batch inference, powering model... ...impact you will have:Build LLM infrastructure powering... ...Anthropic, Gemini) and self-hosted models (Qwen, GPT-OSS,...Local areaWorldwide$145.5k - $249.5k
...Join us to start Caring. Connecting. Growing together. As a Lead AI/ML Engineer on the AI Platform Team within Optum AI, you will drive the design... ...with the challenges of delivering and monitoring LLM-based cloud systems in a production environmentProficiency with...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work- ...Firmable is the market-leading B2B sales... ...served to humans and AI agents alike. This... ...As an Applied AI Engineer on this team, you... ...regression suites, LLM-as-judge where its... ...negotiable. You have run inference over billions of... ...at runtime Self-hosted inference: Hugging...Remote workShift work
$269.1k - $307.2k
Distinguished AI Engineer (Agentic AI Platform) At Capital... ...teams to deliver our industry leading capabilities with... ...will coach and evangelize - hosting architecture office hours, mentoring... ...algorithms or technologies (e.g. LLM Inference, Similarity Search and...Full timePart timeWork at officeLocal area$200k
A top global hedge fund is looking for an AI Platform Engineer to lead the development and management of cutting-edge AI infrastructure for our complex... ...-as-code tools. Experience with AI agent frameworks, LLM integration, or Model Context Protocol (MCP) implementations...$229.9k - $262.4k
...Overview Senior Lead Software Engineer (Golang + EKS, Kubernetes, LLM's + Agentic flows + control/data planes) Do you love building and pioneering in the... ...committed to pioneering and responsibly implementing AI/ML across Capital One . We achieve this by building...Full timePart timeInternshipLocal area- ...Overview Join a boutique quantamental hedge fund as our Lead AI Platform Engineer. Spearhead the buildout of a new internal data lake and platform... ...models, and feature engineering Demonstrated proficiency in LLM fine-tuning, system prompting, and multi-agent frameworks (e...Full timeImmediate start
- ...Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor,... ...help build the platform engineers turn to to ship AI... ...RESPONSIBILITIES: Own and lead Voice AI product areas end... ...performance profiling across host-device boundaries (e.g. PyTorch...Full timeFlexible hours
- ...powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion... ...help build the platform engineers turn to to ship AI... ...that powers large-scale LLM inference across our platform... ...LLM models with industry-leading performance, scalability,...Full timeFlexible hours
- ...About the Team OpenAI’s Inference team powers the deployment of our... ...a small, fast-moving team of engineers focused on delivering a world-... ...pushing the boundaries of what AI can do. We’re expanding into... ...inference tooling like vLLM, TensorRT-LLM, or custom model parallel...Full time
$250k - $300k
...You.com, we are building the AI Search Infrastructure that powers... ...vertical indexes with LLM-optimized retrieval systems to... ...useful. Our team includes engineers, researchers, product builders... ...improving agentic results, cutting inference cost and token usage, and getting...Full timeImmediate startRemote workWork from homeFlexible hours$200k - $350k
...infrastructure that powers high-volume, real-time business operations across multiple systems and platforms. We are seeking a Lead AI Engineer to lead the company's AI engineering organization and establish the technical strategy for production AI systems. What You...Remote jobFull timeImmediate start$115k - $175k
Gong harnesses the power of AI to transform how revenue teams win... ...visibility.As a GTM AI Engineer, you'll design, build, and ship... ...data foundation, and AI tooling/hosting infrastructure this work builds... ...teamExperience with modern AI/LLM tooling and integration patterns...Remote workWork from homeFlexible hours$175k - $215k
...are experimenting with AI in their GTM motion. We... ...this is the person who leads that build.Samba TV's core... ...architecture, prompt engineering standards, and... ...others depend onHands-on LLM production experience:... ...scale, drift monitoring, inference cost managementPrior experience...Full time- ...Description:This is a fullstack engineering role with a strong frontend... ...to infrastructure supporting AI models. Location: New York, NY... ...that support scalable AI model inference and data processingRequirements... ...AWSExperience with Large Language Models (LLM)1+ year of experience at a...3 days per week
$110.7k - $372.9k
...need. Deloitte has a new AI-first effort, backed by... ...lab. As an Agentic AI Engineer, you will design, build... ...and operationalize the LLM- and SLM-powered... ...and/or open-weight/self-hosted models (e.g., Llama via... ...salary is benchmarked to leading technology companies rather...Local areaVisa sponsorship- ...the Role Roger is an AI platform that frees home... ...their homes. Backed by leading healthcare investors like... ...records requires LLM systems that extract structure... ...a Senior Applied AI Engineer to build the intelligence... ..., cost-efficient inference infrastructure with great...Full timeRemote workWork from home
$200k - $350k
...platforms. We are seeking a Staff AI Engineer to lead the design and deployment of production... ...production AI systems. Design model serving, inference, evaluation, and optimization pipelines... ...expertise. ~ Strong production LLM experience. ~ Strong distributed...Remote jobFull timeImmediate start$229.9k - $262.4k
...Overview Senior Lead AI Engineering (GenAI Platform Services: Agentic Systems) Overview:... ...foundation model training, large language model inference, similarity search, guardrails, model... ...Invent and introduce state-of-the-art LLM optimization techniques to improve the...Full timePart timeLocal area$147k - $202k
At Prologis, we don’t just lead the industry—we define it with a 1.3 billion square foot... ...but building what comes next.Job Title:Lead AI EngineerCompany:PrologisTitle: Lead AI... ...be considered.A day in the lifeThe Lead AI Engineer builds production AI platform capabilities...Full timeWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lead AI Engineer (FM Hosting, LLM Inference). Be the first to apply!
Related searches
- lead network engineer San Francisco, CA
- lead engineer San Francisco, CA
- lead algorithm engineer San Francisco, CA
- lead operating engineer San Francisco, CA
- lead web developer San Francisco, CA
- lead infrastructure engineer San Francisco, CA
- machine learning ai engineer San Francisco, CA
- ai ml engineer San Francisco, CA
- ai prompt engineer San Francisco, CA
- ai engineer San Francisco, CA



