AI Infra Engineer - Universal Inference (On-site SF)
Touring Capital
Infinity Artificial Intelligence Institute in San Francisco Bay Area seeks an AI Infrastructure Engineer to advance an inference stack that spans multiple chips and models. You will build optimization kernels, a scalable library generator, and a benchmark-driven pipeline that continuously evolves with new papers and hardware. You should be fluent in Python, have experience with system languages (Rust/C/C++), and be comfortable reading research papers to implement novel ideas inside #J-18808-Ljbffr Touring Capital
- ...Notice: This is an on-site role based in San Francisco... ...extremely talented software engineers across the stack. In... ...that power Pylon's AI features - prompt executions, search infra, and more! Improve LLM... ...your own work You're in SF or you're willing to relocate...WebsiteFull timeWork at officeRelocation
- ...company in San Francisco is seeking a Senior Cloud Infrastructure Engineer to design and manage large-scale distributed systems. The ideal... ..., and Terraform, and thrives in a fast-paced startup environment that values autonomy. This is an on-site position. #J-18808-LjbffrWebsite
- ...with the web by building AI agents that can... ...Responsibilities: Scale infra for post-training of multimodal... ...infra for agentic inference (throughput and latency... ...closely with product engineers to translate cutting‑edge... ...to bring you to SF ~ Generous health, dental...SuggestedWork at officeRelocationVisa sponsorship
- Aware Health is hiring a Director of Engineering in San Francisco for a full-time, on-site/ hybrid schedule (3 days in SF). You will lead and scale our engineering team, shaping AI-driven healthcare solutions and platforms to advance orthopedic care. Ideal candidates bring...WebsiteFull time
$260k - $300k
...healthcare organizations to deploy robust AI infrastructure that directly serves... ...About this role As a Staff Backend/Infra Engineer at Amigo, you'll build the core services... ...yourself to a high bar You can work on site in New York City or San Francisco Nice...WebsiteFull timeFlexible hours- ...technology company to find an AI Engineer to build and ship LLM-... ...Build and optimize inference pipelines and real-... ...from the client's SF office. Nice to Haves... ...degree from a four-year university. Hard skills... ...stack (frontend, backend, infra) to problem solve and...Full timeWork experience placementWork at officeRemote work
- ...looking for exceptional Full-Stack Engineers who thrive in fast-moving... ...record from a Top 20 university or equivalent globally recognized... ...modern web technologies, and AI-powered applications. ✅ Builders... ...underlying models—slashing inference costs by ~50%, dropping...
$269.1k - $307.2k
Distinguished AI Engineer (Agentic AI Platform) At Capital One... ...with model minutiae or infra plumbing. You will design... ...algorithms or technologies (e.g. LLM Inference, Similarity Search and... ...information available through this site. Capital One...WebsiteFull timePart timeWork at officeLocal area$175k - $300k
...the intersection of platform engineering, site reliability, and applied ML systems... ..., and operability of Meshy's AI model serving stack, along... ...core capabilities for the AI inference platform, including key... ...such as AI infrastructure (AI Infra), inference systems, and AI agent...WebsiteWork at officeRemote workFlexible hours$180k - $250k
...San Francisco, CA · On-site (5 days/week) · Full-time... ...Our client builds AI that operates computers... ...document understanding, and inference optimization — making... ...PyTorch Applied ML/AI engineering experience at a strong... ...: a six-person team in SF; this hire owns the...WebsiteFull timeH1bRelocationVisa sponsorship$160k - $300k
...San Francisco, is looking for a Founding Engineer to lead the development of their cutting-edge... ...software solutions. In this full-time on-site role, you will collaborate with the team... ...infrastructure and optimize the performance of AI agents. Ideal candidates will have strong...WebsiteFull time$150k - $200k
..., CA (North Beach) · On-site · Full-time Compensation... ...into action by giving AI agents a dependable way... ...paired with a founding engineer from a major search team... ...thinker ~ On-site in SF (North Beach), 5 days a... ...generalist (not a 10-year infra specialist) ~ Substantial...WebsiteFull timeH1bVisa sponsorship- ...We're a team of ex-Google engineers who built some of the largest defensive platforms on the planet - Safe Browsing and reCAPTCHA . Now... ...an even bigger challenge: stopping the new wave of adversarial AI attacks already hitting organizations today. We're still in...WebsiteFull timeFlexible hours
- AI Engineer Location: SF/ NYC/ Hybrid Type: Full-Time | Founding Team Opportunity About Navigate AI Navigate is building frontier... ...video-based intelligence to life, using real footage from job sites, spatial data, and user context. You'll shape the...WebsiteFull timeLive inRemote work
$314.8k - $359.3k
{"description": "Senior Distinguished AI Engineer At Capital One, we are creating responsible... ...model training, large language model inference, similarity search, guardrails, model... ...information available through this site. Capital One Financial is made...WebsiteFull timePart timeLocal area- Join to apply for the AI Applications Engineer role at Quadric Join to apply for... ...to run neural network (NN) inference workloads in a wide variety... ...required to customer sites Requirements Bachelor’s or... ...Design Engineer, Platforms, University Graduate Sunnyvale, CA $105...WebsiteFull timeTemporary workWorldwide
$286.2k - $326.7k
Sr. Distinguished AI Engineer (Remote Eligible) At Capital One, we are creating responsible... ...model training, large language model inference, similarity search, guardrails, model... ...information available through this site. Capital One Financial is made...WebsiteFull timePart timeLocal areaRemote work$250k - $300k
...intelligence. As the only vertically integrated AI infrastructure company built from the... ...in production. That means owning the inference stack end to end: profiling where time... ...you will also work directly with customer engineering teams to tailor deployments to their...Temporary work- ...deploy, monitor, and scale autonomous AI agents with full visibility and... ...are looking for a Senior Applied AI Engineer to join our on-site Research & Intelligence team in San Francisco... ...to fine-tuning, PEFT, or cost-aware inference strategies. Experience working in...WebsiteFull time
- ...the role Our client is a well-funded AI startup building production-grade ML... .... They are looking for a Senior AI/ML Engineer to own model training pipelines, evaluation systems, and inference serving at scale. Full-time, on-site in San Francisco. What you will do...WebsiteFull time
$188k - $275k
...Description CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers,... ...'ll Do Description of the team: The Inference team is responsible for delivering high-performance... ...: We are looking for an Applied AI Engineer to help us understand, measure, and...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$240k - $270k
...At Sigma, we’re not just adding AI—we’re building the future of... ...where you come in. As an AI/ML Engineer, you’ll join a growing team focused... ...in all our offices in SF, NYC, London and Sydney. Our... ...submit a job application on this site, Sigma processes your personal...WebsiteFull timeWork at officeFlexible hours$250k
...in your career? Join a rapidly growing AI cloud infrastructure provider building high... ...for large-scale AI training and inference workloads. With expanding GPU infrastructure... ...limitations. As a Senior ML Infrastructure Engineer, the successful candidate will help build...Full time$220k
Perplexity is looking for an engineer to join their team in San Francisco. You will work on building and operating the inference engine, supporting new models, migrating GPU kernels, and developing a Rust-based serving runtime. The ideal candidate has 3+ years of experience...$150k - $200k
...consumer social tech startup — using AI to combat loneliness among Gen... ...is looking for a Founding AI Engineer to join their core team in... ...expanding across hundreds of university campuses in the US. What... ...Location This is a full-time, on-site role based in San Francisco,...WebsiteFull timeRemote work$180k - $300k
...across the United States to help them hire. AI Engineer (Full-stack) Location: San Francisco, CA (On-site) Company Stage of Funding: Seed ($18.5M raised... ...scraping systems for visual datasets Build inference serving systems Develop APIs powering client-...WebsiteH1bWork at officeRemote workVisa sponsorship$200k
...lifetime. Convex has assembled a team of engineers who have built and designed some of the... ...to work in-person at Convex's office in SF. ~ Ability to write high quality code (... ...understand the demands of a user-facing live-site service? We generally weigh experience...WebsiteFull timeWork at officeNight shift$240k - $270k
...RoleAt Sigma, we’re not just adding AI—we’re building the future of... ...where you come in. As an AI/ML Engineer, you’ll join a growing team... ...environment in all our offices in SF, NYC, London and Sydney.Our... ...submit a job application on this site, Sigma processes your personal...WebsiteFull timeWork at office- ...the world’s leading generative AI studio behind niji・journey ,... ...for an AI Infrastructure Engineer to join us in building out end... ...implement and run our next-generation inference architecture for running all... ...-person teams, and prefer on-site collaboration in either our...WebsiteWork experience placementWork at officeVisa sponsorship
- Sail builds the world’s most efficient software for inference and agent hosting. In this role, you’ll own token processing down to the lowest layers of the stack, optimize kernel performance, develop new request scheduling and parallelism strategies, and help us use a heterogeneous...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Infra Engineer - Universal Inference (On-site SF). Be the first to apply!
- ai engineer San Francisco, CA
- ai research engineer San Francisco, CA
- machine learning ai engineer San Francisco, CA
- senior ai engineer San Francisco, CA
- ai ml engineer San Francisco, CA
- ai prompt engineer San Francisco, CA
- ai engineer remote San Francisco, CA
- ai developer San Francisco, CA
- on-site clinical research associate (traveling/remote) San Francisco, CA
- website coordinator San Francisco, CA





