Machine Learning Inference Engineer
$401kOscar
An AI Unicorn startup is hiring a Senior Machine Learning Inference Engineer for a full-time role. You will be responsible for improving efficiency for AI-native infrastructure powered by generative and multimodal models. The ideal candidate has over 3 years of professional experience and a strong understanding of GPU infrastructure, Python, and PyTorch. This is a highly autonomous role with significant ownership across inference systems and model performance in production. This role is hybrid in San Francisco Bay Area and offers full benefits and equity. Experience: Building AI applications at scale from the ground up Strong understanding of GPU infrastructure including Triton, TensorRT, or vLLM frameworks Hands-on experience with Python and PyTorch Building model-serving Microservices Diffusion and Multimodal model experience is a plus Equity $401k matching Medical coverage #J-18808-Ljbffr Oscar
$160k - $230k
...About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models models and...SuggestedFull time$209k - $313k
...themselves, live in the moment, learn about the world, and have fun... ...other digital services.Snap Engineering teams build fun and... ...forefront.We’re looking for a Machine Learning Engineer to join Snap... ...Strong understanding of causal inference and modern approaches to estimating...SuggestedFull timeLive inWork at officeLocal area$180k - $270k
...highest standards of data security and privacy protection. To learn more about Plaud, please visit and follow along on... ...experience building and deploying high-throughput, ultra-low-latency inference engines for large language models or foundational speech models....SuggestedFull timeWork at officeWorldwide$203.5k - $299.3k
...creates a causal question.About the RoleWe are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash... ...because you have…Deep practical experience with causal inference, econometrics, experimentation, or causal ML.Experience...SuggestedHourly payWork at officeLocal areaRemote workFlexible hours$155k - $180k
...half the Fortune 100, use Roboflow’s machine learning open source and hosted tools. That includes... ...on all roles (not only product and engineering), so Roboflow employs developers... ...At the center of all of this is inference — one of our most important open source...SuggestedFull timeSecond jobRemote workWork from homeRelocation packageFlexible hoursNight shift$203.5k - $299.3k
...a causal question. About the Role We are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash... ...you because you have… Deep practical experience with causal inference, econometrics, experimentation, or causal ML . Experience...Hourly payWork at officeLocal areaRemote workFlexible hours- ...a causal question. About the Role We are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash... ...re excited about you Deep practical experience with causal inference, econometrics, experimentation, or causal ML. Experience shipping...Hourly payWork at officeLocal areaFlexible hours
- ...requires expertise in deploying GPU systems for high-throughput inference and model performance optimization. The ideal candidate will... ...inference frameworks and a solid understanding of reinforcement learning technologies. Comprehensive healthcare benefits, parental...
- A media technology company in San Francisco is seeking a Founding Engineer specializing in ML Inference. This highly technical role requires expertise in the ML infrastructure stack and aims to optimize generative media performance. The ideal candidate will drive innovations...Relocation package
- Oscar is hiring a Senior Machine Learning Inference Engineer for a full-time role in the San Francisco Bay Area. You will focus on improving efficiency for AI-native infrastructure powering generative and multimodal models, with significant ownership over production inference...Full time
- Reactor in San Francisco is seeking an ML Inference Engineer to maximize performance of generative media models and push ultra-low-latency, high-throughput inference. You will craft an in-house runtime, implement optimizations with PyTorch tools, and collaborate with partner...
- ...our growing team. About the Role We're looking for a Machine Learning Engineer to design, build, and deploy production-grade ML systems... ...scalable ML pipelines for training, evaluation, monitoring, and inference Build intelligent services using modern NLP, LLM,...Full timeWork at officeRemote workFlexible hours2 days per week
- ...that runs the real economy. Learn more about our vision in our... ...Collaborate with product and engineering teams to integrate and deploy... ...Have Strong experience in machine learning, deep learning, and... ...generative AI, or real-time inference systems. Hands-on experience...Full timeWorldwideShift work
- ...We are a small, fast-growing team of engineers in San Francisco powering Fortune 100 enterprises... ...our San Francisco office ~ Eager to learn and adapt quickly ~ Prior startup or... ...active learning pipelines Optimize inference, batching, and quantization on GPU Productionize...Full timeWork at officeVisa sponsorshipRelocation package
- ...about the fruit they are seeing. We are looking for a Machine Learning Engineer to build creative, practical, and robust solutions to ML/... ...monitor infrastructure for model training, evaluation, and inference, both in the cloud and on edge devices. Design and...Full timeWork at officeFlexible hoursWeekend work
$150k - $200k
...reliable, high-speed robot autonomy software stack optimized for inference performance ● Advance SOTA dexterous manipulation... ...Required Qualifications ● PhD or MS degree in Computer Science, Machine Learning, Robotics, or equivalent technical discipline ● Deep...Full time$200k - $400k
...top interpretability researchers and engineers from organizations like OpenAI and DeepMind... ...About the role We’re looking for Machine Learning Engineers to help build our platform... ...model interpretability, training, and inference. Integrate new machine learning...Full time- ...About the Role We're looking for founding Machine Learning Engineers (MLEs) to own and improve our core action models end-to-end - the intelligence... ...platform. You'll work at the intersection of LLM inference, browser understanding, and low-latency systems, shipping...Full timeSleeping nights
- ...is to reinvent the way people learn, starting with language.... ...role We’re hiring an ML Engineer, Assessments to help build best... ...(Content/Learning Design) , Machine Learning, Product, and Engineering... ...→ model training → inference → feedback generation) Own...Full timeLive inImmediate start
$150k - $200k
...instructional materials with dynamic digital learning. Through unparalleled curriculum... ...other departments, including Product, Engineering, Machine Learning and Analytics, to understand... ...execution of code, especially for model inference and lightweight processing tasks. ~...Permanent employmentFull timeWork at officeLocal areaRemote work- ...from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the... ...scaling , post-training and Reinforcement Learning , sandbox environments for evaluation... ...agentic learning , and deployment + inference optimization . You’ll build and iterate...Full timeFlexible hours
- ...connect and drive people forward. We are looking for a Machine Learning Engineer to join the growing AI and Machine Learning team at Strava... ...to shipping production code to scaling and optimizing inference and deployment Shape AI at Strava : Bring your voice and...Full timeWork at officeWorldwideFlexible hours3 days per week
- ...Francisco, NYC, or London offices. About the Role As a Machine Learning Engineer on the Marketplace team, you will build the models and... ...not just top-of-funnel engagement • Real-time and batch inference systems embedded in product-critical workflows Example...Full timeWork at officeRelocation package
$150k - $190k
...-driven simulation software stack for engineering and manufacturing across advanced industries... ..., multi-physics simulation through AI inference across the entire engineering... ...goals. Who We're Looking For As a Machine Learning Engineer in Delivery, you are a...Remote jobFull timeFlexible hours$165k - $230k
...like. About the role We're looking for exceptional Machine Learning Engineers focused on Ads to help take Higgsfield's advertising... ...reliably at significant scale, from experimentation through inference and serving. Work closely with Product, Research, Engineering...Full timeWork at officeRemote workWorldwide3 days per week- ...is the place for you. The Role We’re looking for a Machine Learning Engineer who loves getting close to the metal. This is a hands-on... ...scheduling, and squeezing performance out of complex training and inference workloads. They should be just as comfortable optimizing...Full timeWork at office
- ...looking for a Member of Technical Staff focused on ML systems and inference in San Francisco. You will design and build inference systems... .... Candidates should have strong foundations in software engineering, experience with ML inference systems, and performance tuning...
- MakerMaker.AI is looking for a Senior Machine Learning Systems Engineer in San Francisco. In this role, you will build and operate production inference systems, optimizing for performance and reliability. The ideal candidate will have 3+ years of experience in production...
- ...is seeking a Member of Technical Staff to design and optimize inference systems. The role involves managing KV cache allocation and improving... ...components. Ideal candidates should have strong software engineering skills and experience with ML inference systems, particularly...
- OpenAI in San Francisco seeks an experienced Software Engineer to help bring inference workloads to AWS Trainium and build the software stack to run frontier models efficiently on the platform. This deeply technical, cross‑stack role covers kernels, compilers, and model...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Machine Learning Inference Engineer. Be the first to apply!
- data scientist machine learning engineer San Francisco, CA
- machine learning ai engineer San Francisco, CA
- computer vision machine learning engineer San Francisco, CA
- machine learning engineer San Francisco, CA
- ai ml engineer San Francisco, CA
- graduate machine learning engineer San Francisco, CA
- machine learning software engineer San Francisco, CA
- junior machine learning research engineer San Francisco, CA
- senior ml engineer San Francisco, CA
- internship machine learning San Francisco, CA

