AI Engineer LLM Infra
Yutori
Yutori is reimagining how people interact with the web by building AI agents that can reliably do everyday digital tasks. We are building the entire stack to be agent‑first, from training our own models to generative product interfaces. Towards this goal, we are looking for a member of the AI technical staff to join the founding team. Someone technically strong, and excited about building superhuman AI agents that take actions on the web. Our founders — Devi Parikh, Abhishek Das, Dhruv Batra — have decades of experience in AI research and product spanning generative, multimodal and embodied AI at Meta. Our team combines AI experience with design‑minded product thinking to build and deliver on Yutori’s mission. Yutori is backed by a stellar set of visionary investors — Elad Gil, Sarah Guo, Jeff Dean, Fei‑Fei Li, Amjad Masad, Guillermo Rauch, Akshay Kothari, Soleio, Oliver Cameron, Julien Chaumond, Logan Kilpatrick, Bryan McCann, Vladlen Koltun, Jamie Cuffe, Michele Catasta, etc. Responsibilities: Scale infra for post-training of multimodal LLMs (CPT, SFT, RL, search, reward models) Scale infra for agentic inference (throughput and latency of perception‑planning‑action loops) Build the foundations of a superhuman generalist web‑agent Work closely with product engineers to translate cutting‑edge AI capabilities into reliable product experiences. What we’re looking for: Experience with ML infrastructure (GPU clusters) and supporting networking (NCCL) Experience optimizing post‑training and inference performance of multimodal LLMs (data/tensor/pipeline/context/expert parallelism, optimizing MFU, throughput, latency) Low level systems experience (Triton, CUDA) High IQ, high EQ, high agency, high craftsmanship, low ego. Proactive, clear communication. Benefits and perks: Competitive salary and equity Visa sponsorship and relocation stipend to bring you to SF Generous health, dental, vision insurance for you and your dependents 20 days of paid time off per year Work laptop and budget to set up your work office Daily team lunches Commuter benefits Small, focused team of high‑potential individuals. In‑person in SF. #J-18808-Ljbffr Yutori
$197.3k - $225.1k
Lead AI Engineer (FM Hosting, LLM Inference) Overview At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer...SuggestedFull timePart timeLocal area- ...revolutionizing software development with AI-powered formal verification. We... ...Join our team as an AI Engineer and help us push the boundaries... ...LLMs Optimizing and scaling LLM pipelines Adjust frameworks... ...of production experience in ML Infra, DataOps, distributed training....SuggestedFull timeContract work
$250k
...in your career? Join a rapidly growing AI cloud infrastructure provider building high... ...limitations. As a Senior ML Infrastructure Engineer, the successful candidate will help build... ...workloads using vLLM, SGLang, TensorRT-LLM, or Triton Knowledge of GPU networking...SuggestedFull time$182.52k - $297k
...Plaid Machine Learning Infrastructure (ML Infra) team is responsible for creating and... ...for Plaid use cases to streamline feature engineering for both batch and real-time streaming... ...also help pioneer on early foundation of LLM AI platform for Pl. Working with MLEs and product...SuggestedFull timeWork experience placement$225k - $255k
...or sales motions change. Our AI-native platform equips companies... ...of 8 (pricing experts and engineers) to deliver that capability. Our... ...that are still manual. Own LLM infrastructure: routing across... ...quality tradeoffs. Maintain infra and data residency boundaries...SuggestedWork experience placementImmediate startVisa sponsorship- ...partnering with a fast-moving technology company to find an AI Engineer to build and ship LLM-powered applications. Our client is looking for someone... ...Able to work across the stack (frontend, backend, infra) to problem solve and ship features. Applied-AI literacy...Full timeWork experience placementWork at officeRemote work
$250k - $400k
...hire. Title of Role: Senior Applied AI Engineer Location: San Francisco, CA (On-site,... ...systems (planning, memory, tool use) LLM applications in production environments... ...production AI transitions Strong backend + infra fundamentals Strong Signals...H1bWork at officeRemote workVisa sponsorship$180k - $240k
...About Sauna by Wordware Most AI products hand you a coworker on... ...the Role: As an Applied AI Engineer, you’ll be responsible for building... ...uptime. You’ll work across infra, frontend, and product to make... ...agent-like systems: multi-step LLM pipelines, tool-using bots,...Full timeLive inWork at officeRelocationVisa sponsorshipFlexible hours$150k - $200k
...Applied AI Engineer Soulside AI • US On-Site • Reports to the CTO About Soulside Soulside... ...or prompt change. Optimize the full LLM pipeline-prompting, retrieval, structured... ...AI (or comparable inference/training infra). ~ Demonstrated ability to build...H1bImmediate startRemote workVisa sponsorshipFlexible hours$269.1k - $307.2k
Distinguished AI Engineer (Agentic AI Platform) At Capital One, we are creating responsible... ...without wrestling with model minutiae or infra plumbing. You will design the agentic... ...and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and...Full timePart timeWork at officeLocal area- ...the data and infrastructure layer for taste. Our goal is to end AI slop. To make AI feel right, not just be correct. We raised $18.5... ...Build agent harnesses and context layers Work on scalable RL infra Work with internal research teams on our training pipelines...Full time
- ...Founding AI Engineer We founded Bild AI to tackle the mess that is blueprint reading, cost estimation, and permit applications in construction... ...of the application-layer Apply SotA computer vision, LLM, and multimodal AI approaches to messy, real-world documents...Full timeWork at officeRelocation
$300 per month
...us Edison Scientific builds and deploys AI scientist agents to accelerate science and... ...are an ambitious team run by scientists and engineers from leading institutions across biology,... ...Hands-on experience building and deploying LLM-powered or agentic AI systems in...Full timeWork at officeRemote work$180k - $300k
...About The Role You'll own the core AI systems that power Gamma: the models, prompts... ...our AI stack. You'll work closely with engineering and product to ship improvements that millions... .... What You'll Do Own Gamma's LLM and image prompts, measuring and...Full timeWork at officeImmediate startWork from home- ...Everyone's building AI agents, but almost nobody gets them to production. Building... ...authorization, governance, and trust become the real engineering challenge. Arcade is the MCP runtime... ...database team at Redis, shipped 100+ LLM applications, and is a contributor to...Full timeShift work
- ...Role: ML/AI Engineers (This role is open to US Citizens, Green Card holders, GC-EAD only. We do not sponsor visas.) Summary: Adidev... ...the major clouds. Hands-on experience with Deep Learning, LLM, Python, TensorFlow, PyTorch and other AI frameworks Experience...Full timeRemote workVisa sponsorshipRelocation package
- AI Engineer Location: SF/ NYC/ Hybrid Type: Full-Time | Founding Team Opportunity About Navigate AI Navigate is building frontier... ..., video, and intuitive UX Experiment with AI prompting or LLM tools to layer intelligence over video data Help shape the...Full timeLive inRemote work
- ...AI Engineer As an AI Engineer on our research team, you'll work on our hardest research and AI infra problems. You'll both help prototype new products and also work on turning prototypes... ...Main focus is on improving models, LLM orchestration, prompt optimizations, evaluation...Work at officeFlexible hours
- ...About the role We’re seeking an experienced engineer to deploy enterprise-grade AI solutions, focusing on Retrieval-Augmented Generation (RAG) pipelines and large language model (LLM) workflows. This role is vital to expanding our reach with Fortune 500 and enterprise...Full time
- ...LiteLLM is the world's most popular AI Gateway, trusted by companies like Adobe, Netflix, and NASA. We're building... .... What you will be working on Skills: Python, LLM APIs, FastAPI As the Backend LLM Engineer, you'll be responsible for ensuring LiteLLM unifies the...Full timeImmediate start
$150k - $250k
...Description Max AI – Stripe for Healthcare Max AI is the World’s first human-free... ...for over 10 years. And our Head of Engineering was one of the earliest engineers at Figma... ...transformer-based models, etc) Experience with LLM, including supervised fine-tuning, usage...Full time- ...exceptional and thrive at work through human, AI, and software-based coaching. We’re on a... ...for a product-minded senior software engineer with experience working with data to help... ...figured out how to productively integrate LLM-based coding tools (things like cursor, cody...Remote jobFull timeHome office
$150k - $200k
...Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place... ...for their life’s work. About the Role: As a Software Engineer on Collections Infra, you’ll help scale the infrastructure behind Notion’s database...Full timeLocal area- ...final run, and (3) building the research and infra for horizontal integrations, such as... ...reliable, and unblocked. You will work across engineering and infrastructure problems as they... ...teams. About OpenAI OpenAI is an AI research and deployment company dedicated...Full time
$350k
...for understanding and debugging AI systems. We build world-class,... ...for an exceptional AI systems engineer to lead the design and development... ...building and path-set on what infra we should build Help other team... ...and scale) Bonus: can set up LLM pipelines, e.g. multiple specialized...Visa sponsorshipFlexible hours$120k - $200k
...$120,000 - $200,000 + 0.5‑2% Equity + Flexible PTO Are you an AI engineer who wants to build and own the core agentic systems powering an... ...with React/TypeScript Experience building agentic systems with LLM platforms (Anthropic, OpenAI, or similar) Deep hands‑on experience...Full timeImmediate startFlexible hours- ...At Middleware — one of the fastest-growing AI observability platforms — we're looking... ...a talented and experienced AI/ML founding engineer to join our team. You'll play a key role in... ..., anomaly detection, root cause analysis, LLM-based insights) to find what works best....
- ...Responsibilities Design, develop, and deploy AI software components including foundation... ...systems Implement state-of-the-art LLM optimization techniques to improve... ...~ Strong foundation in mathematics and engineering principles for AI system optimization ~...
- ...intelligence layer that compounds. The Platform engineer creates the tools, you create the... ...judgment under noisy real-world data. Strong LLM orchestration or agent-system experience... ..., or automated recommendation loops. AI-first development workflow and ability to...Full time
- ...AI Engineer Opportunity at Goodfin Goodfin is an AI-native investment platform giving accredited investors access to pre-IPO and alternative... ...across data, tools, and actions. Design and implement LLM-powered systems that go beyond single prompts (multi-step planning...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Engineer LLM Infra. Be the first to apply!
- senior ai engineer San Francisco, CA
- ai ml engineer San Francisco, CA
- ai engineer remote San Francisco, CA
- ai developer San Francisco, CA
- ai prompt engineer San Francisco, CA
- machine learning ai engineer San Francisco, CA
- ai engineer San Francisco, CA
- ai engineer contract
- ai agent engineer
- ai research engineer



