Machine Learning Engineer, Runtime & Optimization
$213k - $263kWaymo
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver-The World's Most Experienced Driver-to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo's fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states. The ML Platform team at Waymo provides a set of tools to support and automate the lifecycle of the machine learning workflow, including feature and experiment management, model development, optimization and monitoring. These efforts have resulted in making machine learning more accessible to teams at Waymo, including Perception, Planner, Research and Simulation. We are looking for engineers with ML software or ML systems expertise to help us improve compute performance on both cloud and car. You'll work across the entire ML stack from the system perspective, from efficient deep learning models, model compression, ML software (e.g. JAX, XLA, Triton, and CUDA), to . You will be pleasantly challenged with deploying Waymo ML models on limited computation resources. In this hybrid role, you will report to the Senior Manager of Runtime and Optimization. You will: Lead the collaboration with the world-class Waymo ML scientists in perception, planner, research and simulation. Identify opportunities in both systems and models to make ML workloads faster. Lead projects from proposals through execution by developing junior engineers. Analyze and improve ML system workloads on both cloud and self-driving cars . Apply model optimization, efficient deep learning techniques and ML software improvements to Waymo's ML systems. You have: M.S. in CS, EE, Deep Learning or a related field 2+ years of experience as a technical lead, including writing project plans, engaging with customer teams, mentoring, responsible for goals & execution, reporting status. 5+ years of experience developing solutions in ML systems or ML software stack (Pytorch/JAX/TF, runtime libraries, ML compiler). Deep understanding of ML system architecture, performance analysis and tools. Strong Python or C++ programming skills We prefer you have one or more of the following: PhD in CS, EE, Deep Learning or a related field. Familiarity with the HW architecture of ML hardware accelerators (e.g., GPU/TPU). Deep knowledge of model optimization or efficient deep learning techniques for foundation models or LLM. Experience with GPU HW or TPU HW and related system software. #LI-Hybrid The expected base salary range for this full-time position across US locations is listed below. Actual starting pay will be based on job-related factors, including exact work location, experience, relevant training and education, and skill level. Your recruiter can share more about the specific salary range for the role location or, if the role can be performed remote, the specific salary range for your preferred location, during the hiring process. Waymo employees are also eligible to participate in Waymo's discretionary annual bonus program, equity incentive plan, and generous Company benefits program, subject to eligibility requirements. Salary Range $213,000—$263,000 USD
$159.05k - $199.3k
...the role We are looking for a software engineer with deep experience in optimizing ML models and deploying them on production‑grade embedded runtime environments. You’ll work across the... ...Experience in working with deep learning frameworks (e.g., PyTorch, JAX, ONNX,...SuggestedFull timeFor contractorsFor subcontractor$204k - $259k
An autonomous driving technology company is looking for an experienced engineer to improve compute performance in machine learning systems. This hybrid role involves collaboration with a world-class ML team and requires strong expertise in ML software or systems. The ideal...Suggested$250k - $350k
...are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's... ...experiences at scale.You will design and optimize inference pipelines, implement state-of-... ..., attention acceleration, and deep learning compiler stacks.GPU & Parallelism: Deep...SuggestedWork at office3 days per week$128.7k - $261.3k
...development, and performance engineering so that every cycle on our... ...Deployments, AI Solutions, Runtime, and AI Kernels teams to co... ...and turns them into highly optimized inference artifacts running... ...developing and deploying machine learning modelsCompensation: The compensation...SuggestedFlexible hours- ...workflows, and continuously learn and adapt.Moveworks is... ...Moveworks’ Reasoning Engine and natural language... ...'re building the runtime infrastructure that powers... ...engine — A state machine that manages long-running... ...bot scoping, batch read optimization, and hot-reload...SuggestedWork at officeRemote workFlexible hours
- ...the company. The Role You will build and optimize production infrastructure for large-... ...Distributed model serving and scheduling Runtime and systems optimization Latency, throughput... ...is an opportunity to join before the engineering organization scales significantly. You...Work at office
- ...quickly and reliably. About this role We are hiring a Machine Learning Engineer, LLM Optimization to build a world-leading inference optimization team... ...strategies for large-scale LLM serving across model execution, runtime systems, and production inference platforms. Advance...Worldwide
$260k - $330k
...Must have: Experience building large-scale prediction or optimization systemsPubMatic is the leading AI-powered ad tech company... ...environments.About the Role:We are looking for a Senior Principal Machine Learning Engineer to help build the next generation of performance...Work at officeRemote work- ...patients worldwide.We’re a team of engineers, clinicians, and innovators... ...DescriptionAbout This Opportunity:As a Staff Machine Learning Engineer, you will be responsible... ...-time inference, GPU/throughput optimization (e.g., TensorRT, ONNX Runtime, mixed precision), and building...Work at officeLocal areaWorldwideFlexible hours
- Applied Intuition Inc. is seeking a software engineer with deep expertise in optimizing ML models for production-grade embedded runtime environments. You will work across the ML framework stack and target a range of embedded compute platforms used in on- and off-road ADAS...
- ...games run inside a modern, browser-native runtime (built on technologies such as WebGPU... ...entirely within that runtime. As our Principal Engineer for On-Device AI Inference & Systems,... ...leaves research, through export, optimization, and kernel-level tuning, to a shipped feature...
$140k - $230k
...tackles the core challenges of machine learning for 3D perception, sensor... ...driving data, solving optimization problems in computer vision... ...a skilled Machine Learning Engineer to help advance a cutting-edge... ...~ Deep understanding of runtime complexity, distributed/cloud...Temporary workWork at officeFlexible hours3 days per week$215k - $285k
...intelligence to every moving machine on the planet. Applied... ...; Seoul; and Tokyo. Learn more at applied.co. We... ...for a performance engineer who specializes in... ..., and evaluation. The optimization target here is not tail... ...Server, TensorRT, ONNX Runtime, Ray, or similar) Fluency...Full timeFor contractorsFor subcontractorCasual workWork at officeRemote workDay shift$195.2k - $361.2k
## Sr. Inference Optimization Engineer (local / edge runtime)Applylocations: US, California, Santa Clara: US, Oregon... ...models run directly on the user's machine (AI PC, edge, on-prem, and beyond),... ...where it helps us# What you’ll learn / grow into*Curiosity is required....InternshipLocal areaShift work$200k - $300k
Machine Learning Systems Engineer Location - Palo Alto, CA (On-site) - Five days per week in-office in the... ..., experimentation, performance optimization, and production-scale AI infrastructure... ...such as vLLM, TensorRT, ONNX Runtime, and SGLang Develop and optimize GPU...H1bWork at office$262k - $364k
...partnership with DeepMind, Research, and Ads Machine Learning teams.Design, prototype, and scale high... ..., including AI Overviews and AI Mode.Engineer mathematical loss functions and... ...learning workflows to automate and accelerate optimal model architecture and feature space...$196k - $221k
...responsible for ML and work alongside industry-veteran scientists and engineers. As a Machine Learning Engineer, you’ll bring your strong software engineering mindset to machine learning in order to scale and optimize our ML systems—creating and transforming innovative research...Permanent employment$209k - $313k
...themselves, live in the moment, learn about the world, and have fun... ...other digital services.Snap Engineering teams build fun and... ...forefront.We’re looking for a Machine Learning Engineer to join Snap... ...that quantify causal impact, optimize decision-making, and drive value...Full timeLive inWork at officeLocal area$132k - $189k
Optimize image quality across the hardware and software stack, ensuring hardware capabilities... ....Build software tools, utilizing machine learning (ML) automation to streamline image... ...qualifications:Bachelor’s degree in Electrical Engineering, Computer Science, Imaging Science,...- ...researchers, data scientists, and engineers, tackling the most... ...performance computing in deep learning, driving impactful discoveries... ...Horovod) • Implement distributed optimizers from mathematical specs •... ...Experience with large-scale machine learning workloads (strong ML...Flexible hours
$160.5k - $240.7k
Company Qualcomm Technologies, Inc. Job Area Engineering Group, Engineering Group > Machine Learning Engineering About the Role Join the Qualcomm AI Hub team... ...this role you will develop tools to help developers optimize and deploy machine learning models on edge and mobile...Work experience placementImmediate startWork from home- A leading tech company based in California is seeking a Research Software Engineer to work at the intersection of computer vision and machine learning. The role involves optimizing network performance and developing technologies for next-generation GPU platforms. Candidates...
- Rivian is seeking a Staff Software Engineer for ML Optimization and Hardware Acceleration to lead optimization of ML models for autonomous driving. You will bridge ML research and embedded deployment, focusing on Transformers, LLMs, VLMs and LDMs across diverse embedded...
$170k - $216k
...team builds the system which learns the spatial-temporal representation... ...with downstream teams on the optimization and integration into the... ...set of sensors, enabling engineers like you to (1) develop methods... ...Lead Manager. You will: Apply machine learning techniques to build...Full timeRemote work$184.7k - $324.8k
...Senior Machine Learning Software Engineer - Computer VisionApple is where individual imaginations converge, united by values that drive exceptional... ...novel hardware for innovative applications.Implementing and optimizing algorithms to run efficiently on mobile devices.Building...Work experience placementRelocation$193.3k - $261.5k
...resourceful Senior Software Development Engineer to bring diverse perspectives, ideas,... .../offline evaluation, and serving optimization. - Partnering with Scientists, Product... ...technologies in distributed systems and machine learning infrastructure. You will foster a culture...InternshipLocal areaWorldwideFlexible hours$251k - $310k
...mission is to build the ultimate cognitive engine for autonomous driving. We are... ...art fine-tuning (SFT) and Reinforcement Learning (RLHF/RLAIF, DPO/GRPO/PPO) pipelines. Drastically... ...: Conduct rigorous ablation studies to optimize model architectures, token budgets, and...Full timeRemote work$251k - $310k
...Staff Machine Learning Engineer, Tech Lead, Labeling AutomationWaymo is an autonomous driving technology company with the mission to be the world... ...labels.Train & Deploy SOTA Computer Vision Models: Train, optimize, and push into production advanced 2D and 3D computer...Full timeRemote work- ...Intuition's Axion Data team builds the data engine to train perception models, facilitating... ...model iteration. The engineer will optimize pipelines, scale data labels, and implement... ...of ML data workflows, with ongoing learning and product impact. #J-18808-Ljbffr...
- ...Applied Intuition, Inc. is seeking a performance engineer to accelerate large-scale ML workloads in the data center. You will own profiling, optimization, and cost-efficiency for distributed training and large offline inferences. You will work across accelerators, ML...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Machine Learning Engineer, Runtime & Optimization. Be the first to apply!
- ai ml engineer Mountain View, CA
- senior ml engineer Mountain View, CA
- machine learning ai engineer Mountain View, CA
- computer vision machine learning engineer Mountain View, CA
- machine learning engineer Mountain View, CA
- machine learning software engineer Mountain View, CA
- artificial intelligence - machine learning intern Mountain View, CA
- internship machine learning Mountain View, CA
- machine learning researcher Mountain View, CA
- machine learning Mountain View, CA



