Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Machine Learning Inference Engineer

$401k

Oscar

An AI Unicorn startup is hiring a Senior Machine Learning Inference Engineer for a full-time role. You will be responsible for improving efficiency for AI-native infrastructure powered by generative and multimodal models. The ideal candidate has over 3 years of professional experience and a strong understanding of GPU infrastructure, Python, and PyTorch. This is a highly autonomous role with significant ownership across inference systems and model performance in production. This role is hybrid in San Francisco Bay Area and offers full benefits and equity. Experience: Building AI applications at scale from the ground up Strong understanding of GPU infrastructure including Triton, TensorRT, or vLLM frameworks Hands-on experience with Python and PyTorch Building model-serving Microservices Diffusion and Multimodal model experience is a plus Equity $401k matching Medical coverage #J-18808-Ljbffr Oscar

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Machine Learning Inference Engineer in San Francisco, CA vacancy
  • $160k - $230k

     ...About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models models and... 
    Suggested
    Full time

    Together Ai

    San Francisco, CA
    1 day ago
  • $209k - $313k

     ...themselves, live in the moment, learn about the world, and have fun...  ...other digital services.Snap Engineering teams build fun and...  ...forefront.We’re looking for a Machine Learning Engineer to join Snap...  ...Strong understanding of causal inference and modern approaches to estimating... 
    Suggested
    Full time
    Live in
    Work at office
    Local area

    Snap

    San Francisco, CA
    2 days ago
  • $180k - $270k

     ...highest standards of data security and privacy protection. To learn more about Plaud, please visit and follow along on...  ...experience building and deploying high-throughput, ultra-low-latency inference engines for large language models or foundational speech models.... 
    Suggested
    Full time
    Work at office
    Worldwide

    Plaud

    San Francisco, CA
    1 day ago
  • $203.5k - $299.3k

     ...creates a causal question.About the RoleWe are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash...  ...because you have…Deep practical experience with causal inference, econometrics, experimentation, or causal ML.Experience... 
    Suggested
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    3 days ago
  • $155k - $180k

     ...half the Fortune 100, use Roboflow’s machine learning open source and hosted tools. That includes...  ...on all roles (not only product and engineering), so Roboflow employs developers...  ...At the center of all of this is inference — one of our most important open source... 
    Suggested
    Full time
    Second job
    Remote work
    Work from home
    Relocation package
    Flexible hours
    Night shift

    Roboflow

    San Francisco, CA
    1 day ago
  • $203.5k - $299.3k

     ...a causal question. About the Role We are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash...  ...you because you have… Deep practical experience with causal inference, econometrics, experimentation, or causal ML . Experience... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Visa Hunt

    San Francisco, CA
    2 days ago
  •  ...a causal question. About the Role We are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash...  ...re excited about you Deep practical experience with causal inference, econometrics, experimentation, or causal ML. Experience shipping... 
    Hourly pay
    Work at office
    Local area
    Flexible hours

    DoorDash, Inc.

    San Francisco, CA
    1 day ago
  •  ...requires expertise in deploying GPU systems for high-throughput inference and model performance optimization. The ideal candidate will...  ...inference frameworks and a solid understanding of reinforcement learning technologies. Comprehensive healthcare benefits, parental... 

    Reflection AI

    San Francisco, CA
    1 day ago
  • A media technology company in San Francisco is seeking a Founding Engineer specializing in ML Inference. This highly technical role requires expertise in the ML infrastructure stack and aims to optimize generative media performance. The ideal candidate will drive innovations... 
    Relocation package

    Reactor.am

    San Francisco, CA
    1 day ago
  • Oscar is hiring a Senior Machine Learning Inference Engineer for a full-time role in the San Francisco Bay Area. You will focus on improving efficiency for AI-native infrastructure powering generative and multimodal models, with significant ownership over production inference... 
    Full time

    Oscar

    San Francisco, CA
    5 days ago
  • Reactor in San Francisco is seeking an ML Inference Engineer to maximize performance of generative media models and push ultra-low-latency, high-throughput inference. You will craft an in-house runtime, implement optimizations with PyTorch tools, and collaborate with partner... 

    Reactor

    San Francisco, CA
    3 days ago
  •  ...our growing team. About the Role We're looking for a Machine Learning Engineer to design, build, and deploy production-grade ML systems...  ...scalable ML pipelines for training, evaluation, monitoring, and inference Build intelligent services using modern NLP, LLM,... 
    Full time
    Work at office
    Remote work
    Flexible hours
    2 days per week

    Plenful

    San Francisco, CA
    1 day ago
  •  ...that runs the real economy. Learn more about our vision in our...  ...Collaborate with product and engineering teams to integrate and deploy...  ...Have Strong experience in machine learning, deep learning, and...  ...generative AI, or real-time inference systems. Hands-on experience... 
    Full time
    Worldwide
    Shift work

    HappyRobot

    San Francisco, CA
    1 day ago
  •  ...We are a small, fast-growing team of engineers in San Francisco powering Fortune 100 enterprises...  ...our San Francisco office ~ Eager to learn and adapt quickly ~ Prior startup or...  ...active learning pipelines Optimize inference, batching, and quantization on GPU Productionize... 
    Full time
    Work at office
    Visa sponsorship
    Relocation package

    The Pulse

    San Francisco, CA
    1 day ago
  •  ...about the fruit they are seeing. We are looking for a Machine Learning Engineer to build creative, practical, and robust solutions to ML/...  ...monitor infrastructure for model training, evaluation, and inference, both in the cloud and on edge devices. Design and... 
    Full time
    Work at office
    Flexible hours
    Weekend work

    Orchard Robotics

    San Francisco, CA
    1 day ago
  • $150k - $200k

     ...reliable, high-speed robot autonomy software stack optimized for inference performance ● Advance SOTA dexterous manipulation...  ...Required Qualifications ● PhD or MS degree in Computer Science, Machine Learning, Robotics, or equivalent technical discipline ● Deep... 
    Full time

    Deft Ai, Inc.

    San Francisco, CA
    1 day ago
  • $200k - $400k

     ...top interpretability researchers and engineers from organizations like OpenAI and DeepMind...  ...About the role We’re looking for Machine Learning Engineers to help build our platform...  ...model interpretability, training, and inference. Integrate new machine learning... 
    Full time

    Goodfire

    San Francisco, CA
    1 day ago
  •  ...About the Role We're looking for founding Machine Learning Engineers (MLEs) to own and improve our core action models end-to-end - the intelligence...  ...platform. You'll work at the intersection of LLM inference, browser understanding, and low-latency systems, shipping... 
    Full time
    Sleeping nights

    Composite

    San Francisco, CA
    1 day ago
  •  ...is to reinvent the way people learn, starting with language....  ...role We’re hiring an ML Engineer, Assessments to help build best...  ...(Content/Learning Design) , Machine Learning, Product, and Engineering...  ...→ model training → inference → feedback generation) Own... 
    Full time
    Live in
    Immediate start

    Speak

    San Francisco, CA
    1 day ago
  • $150k - $200k

     ...instructional materials with dynamic digital learning. Through unparalleled curriculum...  ...other departments, including Product, Engineering, Machine Learning and Analytics, to understand...  ...execution of code, especially for model inference and lightweight processing tasks.  ~... 
    Permanent employment
    Full time
    Work at office
    Local area
    Remote work

    Kiddom

    San Francisco, CA
    1 day ago
  •  ...from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the...  ...scaling , post-training and Reinforcement Learning , sandbox environments for evaluation...  ...agentic learning , and deployment + inference optimization . You’ll build and iterate... 
    Full time
    Flexible hours

    Sciforium

    San Francisco, CA
    1 day ago
  •  ...connect and drive people forward. We are looking for a Machine Learning Engineer to join the growing AI and Machine Learning team at Strava...  ...to shipping production code to scaling and optimizing inference and deployment Shape AI at Strava : Bring your voice and... 
    Full time
    Work at office
    Worldwide
    Flexible hours
    3 days per week

    Strava

    San Francisco, CA
    1 day ago
  •  ...Francisco, NYC, or London offices. About the Role As a Machine Learning Engineer on the Marketplace team, you will build the models and...  ...not just top-of-funnel engagement • Real-time and batch inference systems embedded in product-critical workflows Example... 
    Full time
    Work at office
    Relocation package

    Mercor

    San Francisco, CA
    1 day ago
  • $150k - $190k

     ...-driven simulation software stack for engineering and manufacturing across advanced industries...  ..., multi-physics simulation through AI inference across the entire engineering...  ...goals. Who We're Looking For As a Machine Learning Engineer in Delivery, you are a... 
    Remote job
    Full time
    Flexible hours

    Physicsx

    San Francisco, CA
    1 day ago
  • $165k - $230k

     ...like. About the role We're looking for exceptional Machine Learning Engineers focused on Ads to help take Higgsfield's advertising...  ...reliably at significant scale, from experimentation through inference and serving. Work closely with Product, Research, Engineering... 
    Full time
    Work at office
    Remote work
    Worldwide
    3 days per week

    Higgsfield

    San Francisco, CA
    1 day ago
  •  ...is the place for you. The Role We’re looking for a Machine Learning Engineer who loves getting close to the metal. This is a hands-on...  ...scheduling, and squeezing performance out of complex training and inference workloads. They should be just as comfortable optimizing... 
    Full time
    Work at office

    Relace

    San Francisco, CA
    1 day ago
  •  ...looking for a Member of Technical Staff focused on ML systems and inference in San Francisco. You will design and build inference systems...  .... Candidates should have strong foundations in software engineering, experience with ML inference systems, and performance tuning... 

    Gimlet Labs, Inc.

    San Francisco, CA
    1 day ago
  • MakerMaker.AI is looking for a Senior Machine Learning Systems Engineer in San Francisco. In this role, you will build and operate production inference systems, optimizing for performance and reliability. The ideal candidate will have 3+ years of experience in production... 

    MakerMaker.AI

    San Francisco, CA
    3 days ago
  •  ...is seeking a Member of Technical Staff to design and optimize inference systems. The role involves managing KV cache allocation and improving...  ...components. Ideal candidates should have strong software engineering skills and experience with ML inference systems, particularly... 

    Gimlet Labs

    San Francisco, CA
    4 days ago
  • OpenAI in San Francisco seeks an experienced Software Engineer to help bring inference workloads to AWS Trainium and build the software stack to run frontier models efficiently on the platform. This deeply technical, cross‑stack role covers kernels, compilers, and model... 

    Slope

    San Francisco, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Machine Learning Inference Engineer. Be the first to apply!