Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform)
$229.9k - $262.4kCapital One
Sr. Lead AI Engineer (Inference Optimization, FM Hosting, AI Platform)At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine learning — position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build.The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact.In this role, you will:Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One.Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.Invent and introduce state-of-the-art LLM optimization techniques to improve the performance — scalability, cost, latency, throughput — of large scale production AI systems.Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.The Ideal Candidate:You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good.Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production.You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven.You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enable you to see and exploit optimization opportunities that others miss.You are a resilient trail blazer who can forge new paths to achieve business goals when the route is unknown.Basic Qualifications:Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologiesAt least 6 years of experience programming with Python, Go, Scala, or JavaPreferred Qualifications:7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud)Experience designing, developing, integrating, delivering, and supporting complex AI systemsDemonstrated ability to lead and mentor an engineering team and influence cross-functional stakeholdersExperience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or GolangExperience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and costPassion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in productionExcellent communication and presentation skills, with the ability to articulate complex AI concepts to peersCapital One will consider sponsoring a new qualified applicant for employment authorization for this position.The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked.Cambridge, MA: $229,900 - $262,400 for Sr. Lead AI EngineerMcLean, VA: $229,900 - $262,400 for Sr. Lead AI EngineerNew York, NY: $250,800 - $286,200 for Sr. Lead AI EngineerSan Francisco, CA: $250,800 - $286,200 for Sr. Lead AI EngineerSan Jose, CA: $250,800 - $286,200 for Sr. Lead AI EngineerCandidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter.This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan.Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level.
$229.9k - $262.4k
Senior Lead AI Engineer (FM Hosting, LLM Inference) Overview: At Capital One, we are creating... ...customers. Our AI models and platforms empower teams across... ...introduce state-of-the-art LLM optimization techniques to improve the... ...$229,900 - $262,400 for Sr. Lead AI Engineer New...PlatformSeniorFull timePart timeLocal area$229.9k - $262.4k
...Overview AI Engineer 5 (FM Hosting, LLM Inference) At Capital One, we are creating... ...to deliver our industry leading capabilities with breakthrough... .... Our AI models and platforms empower teams across... ...the-art foundation model optimization techniques to improve the...PlatformFull timePart timeLocal area$171.9k - $300.8k
...what data.Veza's Access Graph platform maps an organization's... ...combination brings together Veza's AI-native Access Graph with... ...environments, and AI agents. For engineers joining Veza today, this... ...Exposure to LLM fine-tuning or inference optimization in production.Why join us...PlatformSeniorWork experience placementWork at officeRemote workFlexible hours$184k - $287.5k
NVIDIA is the platform upon which every new AI-powered application is built... ...a Senior Software Engineer - AI Inference Performance to... ...What you'll be doing:Lead end-to-end analysis... ...decode workloads. Optimize time to first token... ...Eliminate bottlenecks in host code, CUDA kernels,...PlatformSeniorFull time$274k - $300k
...Description Saviynt's AI-powered identity platform manages and governs human... ...and empower the world’s leading brands, Fortune 500 companies... ...visit AI Platform Engineer – Training & Inference Saviynt's AI-... ...aware fallback between self-hosted SLMs and cloud LLMs •...Platform- ...discovery to powering AI and the technologies... ...breakthroughs, or bringing leading edge products to... ...motivated AI Model Optimization & Software Engineer Interns/Co-op to join... ...training, fine-tuning, and inference across CPU, GPU, and accelerator platforms. Profile AI...PlatformFull timeSummer workInternshipSummer internshipWorldwide
$184k - $287.5k
...NVIDIA's DGX Cloud AI Efficiency Team... ...developing tools for optimizing efficiency and... ...training, post-training, inference. Our objective is... ...software engineer to join our team.... ...underpinning NVIDIA's AI platforms.Define meaningful... ...JAX, and RayNVIDIA leads the way in groundbreaking...PlatformSeniorFull timeRemote work$255.65k
...can expect: As a Senior AI Software Engineer, you will collaborate to design, implement, and optimize AI algorithms and software applications... ...will ensure AI training, inference, deployment, and operation... ...the best collaboration platform for the enterprise, and today...PlatformSeniorWork at officeRemote work$160k - $198k
...advanced air mobility platform that delivers air taxis... ...artificial intelligence ("AI") solutions, and other... ...As a Senior AI Systems Engineer, you will architect,... ...AI model training and inference. You will ensure our machine... ...-to streamline and optimize the AI development lifecycle...PlatformSeniorLocal areaWorldwideVisa sponsorship- ...quality of life. As an AI / Embedded Engineer, you will be... ...model development, optimization, and deployment on embedded... ...model size and inference latency ◦ Use frameworks... ...and cloud or edge-hosted LLM components ◦... ...Experience with RTOS platforms such as FreeRTOS or...PlatformFull timeWork at officeImmediate startVisa sponsorshipNight shift
$229.9k - $262.4k
...Overview Sr. Manager, AI Engineer (IFX) At Capital One, we... ...deliver our industry leading capabilities with breakthrough... ...Our AI models and platforms empower teams across... ...language model inference, similarity search,... ...foundation model optimization techniques to improve...PlatformSeniorFull timePart timeLocal area$234k - $286k
...leader in next-generation AI infrastructure, delivering a full-stack inference platform for customers... ...Principal AI Solutions Engineer: a hands-on technical... ...knowledge workers. You will lead hands-on engagements... ...driven prompt and program optimization (e.g. DSPy-style),...PlatformSeniorFull timeTemporary workLocal areaWorldwideFlexible hours$100k
...Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance... ...models, compilers, platforms, networking, and... ...-generation RISC-V CPUs optimized for AI, high-performance... ...HPC & Agentic Software Engineering Lead, you will operate at...PlatformFull timeRemote work$224k - $356.5k
...for a Senior Software Engineer to join our team! We... ...by combining the best AI and graphics... ...open-source modding platform for remastering classic... ...content and controls.Optimize models and inference for latency, token throughput... ...alongside rendering.Lead technical decisions,...PlatformFull timeRemote work- ...looking for a GenAI Engineer with strong... ...high-performance inference services. The ideal... ...enterprise GenAI platforms across GPU infrastructure... ...Deploy, host, and manage Large... ...and Ray Serve. Optimize model serving for... ...efficiency. Develop AI platform services...PlatformContract work
- ...enterprises move beyond AI experimentation and... ...through enterprise AI platforms, products, and services. We combine deep engineering expertise with AI innovation... ...business outcomes.As a Lead Generative AI Engineer,... ..., deployment, and optimization.Partner with technical...PlatformSenior
$184k - $287.5k
...unlimited potential of AI to define the next era... ...of data center platform designs. From single node... ...Grace CPUs, and a fully optimized NVIDIA AI and HPC software... ...a highly motivated engineer to lead performance benchmarking... ...-world AI training, inference, and HPC workloads at...PlatformSeniorFull timeRemote work$242.2k
...watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and... ...role Roku is seeking a Senior AI Transformation Engineer to join the AI Enablement Team and... ...sources, pipelines, and metric layers Lead onboarding, enablement, and change...PlatformSeniorWork at officeLocal areaRemote workMonday to ThursdayFlexible hours$200k - $270k
...Description Samsung SDS America AI Team is researching the... ...partners, we use their platforms to develop complete embodied... ...looking for a Senior Physical AI Engineer to join the team developing... ..., or similar platforms Optimize low-latency inference and control systems for...PlatformSeniorWorldwideFlexible hours$180k - $250k
...a Cyber Security Engineer, with hands-on experience... ...API, MosaicML Platform -Experience using... ...-Experience using optimization tools such as... ...Distillation, DeepSpeed-Inference /... ...Global is hiring a AI Cyber Security Engineer... ...Engineer Technical Lead, Identity Sunnyvale...PlatformFull time$123.24k - $200k
Overview Of Role As a Sr./Principal AI Engineer within TSMC's Artificial Intelligence... ...learning for manufacturing optimization. This role is for a builder... .... Responsibilities Lead System Architecture: Own the... ...new AI‑native products and platforms, from initial concept and data...PlatformSeniorWork at office$229.9k - $286.2k
AI Engineer 5 (MLXT) At Capital One, we are creating responsible... ...deliver our industry leading capabilities with... .... Our AI models and platforms empower teams across... ...large language model inference, agents and multi-... ...art foundation model optimization techniques to improve...PlatformFull timePart timeLocal area$40 - $85 per hour
...is one of the world's leading entertainment... ...and languages. The AI Platform team builds the infrastructure... ...to GPU-optimized inference and serving. You'll... ...Machine Learning, Computer Engineering, or a related field... ...value and we strive to host a meaningful interview...PlatformHourly payFull timeInternshipImmediate startRemote workFlexible hours- ...computing experiences—from AI and data centers, to PCs, gaming... ...: We are seeking a DevOps / Platform Engineer to join our team building... ...communicate effectively and work optimally with their peers within our... ...pods, CI pipelines, inference services, benchmarking jobs)...Platform
$2,500 per month
...heavily focused on inference . Backed by hundreds... ...investors and staffed by leading engineers, Etched is redefining... ...We are using AI to build AI chips. AI... ...developer-facing internal platforms, CI/CD at scale, or infrastructure... ...AlphaEvolve-style optimization loops that propose...PlatformWork at officeRelocation packageNight shift$229.9k - $262.4k
...Overview AI Engineer 5 (GenAI Platform, Agentic Infrastructure) At Capital... ...teams to deliver our industry leading capabilities with... ...training, large language model inference, agents and multi-agent workflows... ...-the-art foundation model optimization techniques to improve the...PlatformFull timePart timeLocal area- ...Bitdeer is a world-leading technology company for AI and Bitcoin mining... ...Cloud Senior DevOps Engineer to join our AI Cloud... ...products and platforms are delivered with... ...driving automation, optimizing cloud-native architectures... ...vLLM, TGI, Triton Inference Server), and...PlatformSeniorRemote jobFull timeLocal area
$124.36k - $146.3k
...Description Job Summary The Senior Engineer (Generative AI) is responsible for designing,... ...deployment Performance tuning and optimization Support secure deployment, horizontal... ...production environments 3. Cloud, Platform & Scalability Engineering Develop...PlatformSeniorFull timeTemporary workWork experience placementLocal area3 days per week$2,500 per month
...heavily focused on inference . Backed by hundreds... ...investors and staffed by leading engineers, Etched is redefining... ...We are using AI to build and ship AI... ...You will embed with Platform, Production, Supply Chain... ...frontier models and system optimization methods. You may...PlatformWork at officeRelocation package$248k - $391k
...unlimited potential of AI to define the next era... ...Principal AI Workflow Engineer to join the Enterprise... ..., adaptive, and self-optimizing integration strategies... ...real-time integration platforms using event-driven, asynchronous... ...as one of the world’s leading technology companies,...PlatformFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform). Be the first to apply!
- lead operating engineer San Jose, CA
- lead engineer San Jose, CA
- senior ai engineer San Jose, CA
- ai developer San Jose, CA
- ai prompt engineer San Jose, CA
- ai engineer San Jose, CA
- ai engineer remote San Jose, CA
- senior manager customer operations San Jose, CA
- senior software engineer ruby on rails San Jose, CA
- sr finance manager San Jose, CA





