Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Infra Engineer - Universal Inference (On-site SF)

Touring Capital

Infinity Artificial Intelligence Institute in San Francisco Bay Area seeks an AI Infrastructure Engineer to advance an inference stack that spans multiple chips and models. You will build optimization kernels, a scalable library generator, and a benchmark-driven pipeline that continuously evolves with new papers and hardware. You should be fluent in Python, have experience with system languages (Rust/C/C++), and be comfortable reading research papers to implement novel ideas inside #J-18808-Ljbffr Touring Capital

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the AI Infra Engineer - Universal Inference (On-site SF) in San Francisco, CA vacancy
  •  ...Notice: This is an on-site role based in San Francisco...  ...extremely talented software engineers across the stack. In...  ...that power Pylon's AI features - prompt executions, search infra, and more! Improve LLM...  ...your own work You're in SF or you're willing to relocate... 
    Website
    Full time
    Work at office
    Relocation

    Pylon

    San Francisco, CA
    8 hours ago
  •  ...company in San Francisco is seeking a Senior Cloud Infrastructure Engineer to design and manage large-scale distributed systems. The ideal...  ..., and Terraform, and thrives in a fast-paced startup environment that values autonomy. This is an on-site position. #J-18808-Ljbffr
    Website

    The Recruiting Guy

    San Francisco, CA
    1 day ago
  •  ...with the web by building AI agents that can...  ...Responsibilities: Scale infra for post-training of multimodal...  ...infra for agentic inference (throughput and latency...  ...closely with product engineers to translate cutting‑edge...  ...to bring you to SF ~ Generous health, dental... 
    Suggested
    Work at office
    Relocation
    Visa sponsorship

    Yutori

    San Francisco, CA
    1 day ago
  • Aware Health is hiring a Director of Engineering in San Francisco for a full-time, on-site/ hybrid schedule (3 days in SF). You will lead and scale our engineering team, shaping AI-driven healthcare solutions and platforms to advance orthopedic care. Ideal candidates bring... 
    Website
    Full time

    ApplyMint

    San Francisco, CA
    5 days ago
  • $260k - $300k

     ...healthcare organizations to deploy robust AI infrastructure that directly serves...  ...About this role As a Staff Backend/Infra Engineer at Amigo, you'll build the core services...  ...yourself to a high bar You can work on site in New York City or San Francisco Nice... 
    Website
    Full time
    Flexible hours

    Amigo

    San Francisco, CA
    16 days ago
  •  ...technology company to find an AI Engineer to build and ship LLM-...  ...Build and optimize inference pipelines and real-...  ...from the client's SF office. Nice to Haves...  ...degree from a four-year university. Hard skills...  ...stack (frontend, backend, infra) to problem solve and... 
    Full time
    Work experience placement
    Work at office
    Remote work

    twenty80.io

    San Francisco, CA
    a month ago
  •  ...looking for exceptional Full-Stack Engineers who thrive in fast-moving...  ...record from a Top 20 university or equivalent globally recognized...  ...modern web technologies, and AI-powered applications. ✅ Builders...  ...underlying models—slashing inference costs by ~50%, dropping... 

    Mintstage

    San Francisco, CA
    8 days ago
  • $269.1k - $307.2k

    Distinguished AI Engineer (Agentic AI Platform) At Capital One...  ...with model minutiae or infra plumbing. You will design...  ...algorithms or technologies (e.g. LLM Inference, Similarity Search and...  ...information available through this site. Capital One... 
    Website
    Full time
    Part time
    Work at office
    Local area

    Capital One Financial Corporation

    San Francisco, CA
    8 hours ago
  • $175k - $300k

     ...the intersection of platform engineering, site reliability, and applied ML systems...  ..., and operability of Meshy's AI model serving stack, along...  ...core capabilities for the AI inference platform, including key...  ...such as AI infrastructure (AI Infra), inference systems, and AI agent... 
    Website
    Work at office
    Remote work
    Flexible hours

    Meshy

    San Francisco, CA
    5 days ago
  • $180k - $250k

     ...San Francisco, CA · On-site (5 days/week) · Full-time...  ...Our client builds AI that operates computers...  ...document understanding, and inference optimization — making...  ...PyTorch Applied ML/AI engineering experience at a strong...  ...: a six-person team in SF; this hire owns the... 
    Website
    Full time
    H1b
    Relocation
    Visa sponsorship

    David Joseph & Company

    San Francisco, CA
    22 days ago
  • $160k - $300k

     ...San Francisco, is looking for a Founding Engineer to lead the development of their cutting-edge...  ...software solutions. In this full-time on-site role, you will collaborate with the team...  ...infrastructure and optimize the performance of AI agents. Ideal candidates will have strong... 
    Website
    Full time

    Lance , Inc.

    San Francisco, CA
    22 hours ago
  • $150k - $200k

     ..., CA (North Beach) · On-site · Full-time Compensation...  ...into action by giving AI agents a dependable way...  ...paired with a founding engineer from a major search team...  ...thinker ~ On-site in SF (North Beach), 5 days a...  ...generalist (not a 10-year infra specialist) ~ Substantial... 
    Website
    Full time
    H1b
    Visa sponsorship

    David Joseph & Company

    San Francisco, CA
    21 days ago
  •  ...We're a team of ex-Google engineers who built some of the largest defensive platforms on the planet - Safe Browsing and reCAPTCHA . Now...  ...an even bigger challenge: stopping the new wave of adversarial AI attacks already hitting organizations today. We're still in... 
    Website
    Full time
    Flexible hours

    Aegis Ai

    San Francisco, CA
    8 hours ago
  • AI Engineer Location: SF/ NYC/ Hybrid Type: Full-Time | Founding Team Opportunity About Navigate AI Navigate is building frontier...  ...video-based intelligence to life, using real footage from job sites, spatial data, and user context. You'll shape the... 
    Website
    Full time
    Live in
    Remote work

    Navigate Ai

    San Francisco, CA
    8 hours ago
  • $314.8k - $359.3k

    {"description": "Senior Distinguished AI Engineer At Capital One, we are creating responsible...  ...model training, large language model inference, similarity search, guardrails, model...  ...information available through this site. Capital One Financial is made... 
    Website
    Full time
    Part time
    Local area

    Capital One Financial Corporation

    San Francisco, CA
    8 hours ago
  • Join to apply for the AI Applications Engineer role at Quadric Join to apply for...  ...to run neural network (NN) inference workloads in a wide variety...  ...required to customer sites Requirements Bachelor’s or...  ...Design Engineer, Platforms, University Graduate Sunnyvale, CA $105... 
    Website
    Full time
    Temporary work
    Worldwide

    Quadric

    San Francisco, CA
    4 days ago
  • $286.2k - $326.7k

    Sr. Distinguished AI Engineer (Remote Eligible) At Capital One, we are creating responsible...  ...model training, large language model inference, similarity search, guardrails, model...  ...information available through this site. Capital One Financial is made... 
    Website
    Full time
    Part time
    Local area
    Remote work

    Capital One Financial Corporation

    San Francisco, CA
    8 hours ago
  • $250k - $300k

     ...intelligence. As the only vertically integrated AI infrastructure company built from the...  ...in production. That means owning the inference stack end to end: profiling where time...  ...you will also work directly with customer engineering teams to tailor deployments to their... 
    Temporary work

    Crusoe

    San Francisco, CA
    4 days ago
  •  ...deploy, monitor, and scale autonomous AI agents with full visibility and...  ...are looking for a Senior Applied AI Engineer to join our on-site Research & Intelligence team in San Francisco...  ...to fine-tuning, PEFT, or cost-aware inference strategies. Experience working in... 
    Website
    Full time

    Alterion, Inc.

    San Francisco, CA
    8 hours ago
  •  ...the role Our client is a well-funded AI startup building production-grade ML...  .... They are looking for a Senior AI/ML Engineer to own model training pipelines, evaluation systems, and inference serving at scale. Full-time, on-site in San Francisco. What you will do... 
    Website
    Full time

    Clera

    San Francisco, CA
    8 hours ago
  • $188k - $275k

     ...Description CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers,...  ...'ll Do Description of the team: The Inference team is responsible for delivering high-performance...  ...: We are looking for an Applied AI Engineer to help us understand, measure, and... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    San Francisco, CA
    13 days ago
  • $240k - $270k

     ...At Sigma, we’re not just adding AI—we’re building the future of...  ...where you come in. As an AI/ML Engineer, you’ll join a growing team focused...  ...in all our offices in SF, NYC, London and Sydney. Our...  ...submit a job application on this site, Sigma processes your personal... 
    Website
    Full time
    Work at office
    Flexible hours

    Sigma Computing

    San Francisco, CA
    8 hours ago
  • $250k

     ...in your career? Join a rapidly growing AI cloud infrastructure provider building high...  ...for large-scale AI training and inference workloads. With expanding GPU infrastructure...  ...limitations. As a Senior ML Infrastructure Engineer, the successful candidate will help build... 
    Full time
    San Francisco, CA
    a month ago
  • $220k

    Perplexity is looking for an engineer to join their team in San Francisco. You will work on building and operating the inference engine, supporting new models, migrating GPU kernels, and developing a Rust-based serving runtime. The ideal candidate has 3+ years of experience... 

    Perplexity

    San Francisco, CA
    2 days ago
  • $150k - $200k

     ...consumer social tech startup — using AI to combat loneliness among Gen...  ...is looking for a Founding AI Engineer to join their core team in...  ...expanding across hundreds of university campuses in the US. What...  ...Location This is a full-time, on-site role based in San Francisco,... 
    Website
    Full time
    Remote work

    Clera

    San Francisco, CA
    4 days ago
  • $180k - $300k

     ...across the United States to help them hire. AI Engineer (Full-stack) Location: San Francisco, CA (On-site) Company Stage of Funding: Seed ($18.5M raised...  ...scraping systems for visual datasets Build inference serving systems Develop APIs powering client-... 
    Website
    H1b
    Work at office
    Remote work
    Visa sponsorship

    Recruiting from Scratch

    San Francisco, CA
    2 days ago
  • $200k

     ...lifetime. Convex has assembled a team of engineers who have built and designed some of the...  ...to work in-person at Convex's office in SF. ~ Ability to write high quality code (...  ...understand the demands of a user-facing live-site service? We generally weigh experience... 
    Website
    Full time
    Work at office
    Night shift

    Convex

    San Francisco, CA
    8 hours ago
  • $240k - $270k

     ...RoleAt Sigma, we’re not just adding AI—we’re building the future of...  ...where you come in. As an AI/ML Engineer, you’ll join a growing team...  ...environment in all our offices in SF, NYC, London and Sydney.Our...  ...submit a job application on this site, Sigma processes your personal... 
    Website
    Full time
    Work at office

    Sigma Computing

    San Francisco, CA
    1 day ago
  •  ...the world’s leading generative AI studio behind niji・journey ,...  ...for an AI Infrastructure Engineer to join us in building out end...  ...implement and run our next-generation inference architecture for running all...  ...-person teams, and prefer on-site collaboration in either our... 
    Website
    Work experience placement
    Work at office
    Visa sponsorship

    Spellbrush

    San Francisco, CA
    more than 2 months ago
  • Sail builds the world’s most efficient software for inference and agent hosting. In this role, you’ll own token processing down to the lowest layers of the stack, optimize kernel performance, develop new request scheduling and parallelism strategies, and help us use a heterogeneous... 

    SAIL

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Infra Engineer - Universal Inference (On-site SF). Be the first to apply!