Machine Learning Engineer, Runtime & Optimization
$213k - $263kWaymo
The ML Platform team at Waymo provides a set of tools to support and automate the lifecycle of the machine learning workflow, including feature and experiment management, model development, optimization and monitoring. These efforts have resulted in making machine learning more accessible to teams at Waymo, including Perception, Planner, Research and Simulation.
We are looking for engineers with ML software or ML systems expertise to help us improve compute performance on both cloud and car. You'll work across the entire ML stack from the system perspective, from efficient deep learning models, model compression, ML software (e.g. JAX, XLA, Triton, and CUDA), to . You will be pleasantly challenged with deploying Waymo ML models on limited computation resources. In this hybrid role, you will report to the Senior Manager of Runtime and Optimization. You Will- Lead the collaboration with the world-class Waymo ML scientists in perception, planner, research and simulation. Identify opportunities in both systems and models to make ML workloads faster.
- Lead projects from proposals through execution by developing junior engineers.
- Analyze and improve ML system workloads on both cloud and self-driving cars .
- Apply model optimization, efficient deep learning techniques and ML software improvements to Waymo's ML systems.
- M.S. in CS, EE, Deep Learning or a related field
- 2+ years of experience as a technical lead, including writing project plans, engaging with customer teams, mentoring, responsible for goals & execution, reporting status.
- 5+ years of experience developing solutions in ML systems or ML software stack (Pytorch/JAX/TF, runtime libraries, ML compiler).
- Deep understanding of ML system architecture, performance analysis and tools.
- Strong Python or C++ programming skills
- PhD in CS, EE, Deep Learning or a related field.
- Familiarity with the HW architecture of ML hardware accelerators (e.g., GPU/TPU).
- Deep knowledge of model optimization or efficient deep learning techniques for foundation models or LLM.
- Experience with GPU HW or TPU HW and related system software.
$159.05k - $199.3k
...intelligence to every moving machine on the planet. Applied... ...; Seoul; and Tokyo. Learn more at applied.co. We... ...are looking for a software engineer with deep experience in optimizing ML models and deploying them... ...-grade embedded runtime environments. You’ll work...SuggestedFull timeFor contractorsFor subcontractorCasual workWork at officeRemote workDay shift$137.1k - $201.6k
...The mission of the Marketplace Optimization team is to ensure we maintain a healthy... ...artificial intelligence and advanced ML, deep learning techniques to power decision-making... ...the Role We’re looking for a Machine Learning Engineer to help design, build, optimize and scale...SuggestedHourly payWork at officeLocal areaRemote workFlexible hours- ...the TeamThe mission of the Marketplace Optimization team is to ensure we maintain a... ...artificial intelligence and advanced ML, deep learning techniques to power decision-making... .... About the RoleWe’re looking for a Machine Learning Engineer to help design, build, optimize and...SuggestedHourly payWork at officeLocal areaRemote workFlexible hours
$170k - $216k
...team builds the system which learns the spatial-temporal representation... ...with downstream teams on the optimization and integration into the... ...set of sensors, enabling engineers like you to (1) develop methods... ...experience ~3+ years experience in Machine Learning and/or Computer...SuggestedFull timeRemote work$260k - $330k
...Must have: Experience building large-scale prediction or optimization systemsPubMatic is the leading AI-powered ad tech company... ...environments.About the Role:We are looking for a Senior Principal Machine Learning Engineer to help build the next generation of performance...SuggestedWork at officeRemote work$298k - $368k
...team builds the system which learns the spatial-temporal representation... ...with downstream teams on the optimization and integration into the... ...set of sensors, enabling engineers like you to (1) develop methods... ...~7+ years of experience in Machine Learning, with a focus on large...Full timeRemote work$216.7k - $303.4k
...Description This role sits in the Ads Optimization organizations, which are responsible for... .... You’ll join a set of tight-knit engineers working on high-impact, internet-scale... ...Ads. Role Description We are hiring Machine Learning Engineers (IC4) to build and evolve the...For contractorsWork experience placementFlexible hours$160k - $230k
...About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI... ...performance at scale. Develop and optimize runtime inference services for large-scale AI applications...Full time$150k - $200k
...collision and force guards, smoothing, and runtime monitors for safety-critical deployment... ...-speed robot autonomy software stack optimized for inference performance ●... ...PhD or MS degree in Computer Science, Machine Learning, Robotics, or equivalent technical discipline...Full time- ...with hands-on support from AMD engineers the team is scaling rapidly... ...-training and Reinforcement Learning , sandbox environments for... ...and deployment + inference optimization . You’ll build and iterate... ...and optimize GPU kernels, runtime bottlenecks, and memory behavior...Full timeFlexible hours
$250k - $350k
...are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's... ...experiences at scale.You will design and optimize inference pipelines, implement state-of-... ..., attention acceleration, and deep learning compiler stacks.GPU & Parallelism: Deep...Work at office3 days per week$120k - $160k
...are unavailable or unreliable. As a Machine Learning Engineer, you will own and scale the training,... ...validate to redeploy loops.Deploy and optimize models for real-time edge inference on... ...(quantization/pruning, TensorRT/ONNX Runtime); profile CPU/GPU and hit tight...Work experience placementWork at officeLocal areaShift work$204k - $259k
An autonomous driving technology company is looking for an experienced engineer to improve compute performance in machine learning systems. This hybrid role involves collaboration with a world-class ML team and requires strong expertise in ML software or systems. The ideal...$195.2k - $262.2k
...infrastructure. Built by engineers, for engineers. From large-... ...orchestration to inference optimization, we own the hard problems across... ...across model code, kernels, runtime, scheduler, gateway, and... ...compensation Career growth and learning opportunities...Full timeTemporary workImmediate startRemote work$213k - $263k
...across 15+ U.S. states. The ML Optimization team at Waymo provides a... ...the lifecycle of the machine learning workflow, including feature... ...Simulation. We are looking for engineers with ML software & systems... ...to the Senior Manager of Runtime and Optimization. You...Full timeRemote work- Apple is seeking an ML Infrastructure Engineer focused on ML user experience APIs and integration for on-device AI. You... ...to integrate these APIs into internal and external model repositories, optimize pipelines from authoring to runtime, and #J-18808-Ljbffr Socket.dev
$180k - $220k
...build sensors and tools for engineers, roboticists, and researchers... ...for a highly technical Machine Learning Engineer to lead our efforts... ...research papers into code, and optimizing these systems for real-time,... ...Experience with TensorRT, ONNX Runtime, or edge-specific hardware (...Full timeWork experience placementLocal area$174.72k - $295.68k
...through cutting-edge R&D in AI, machine learning, and smart connectivity.Our... ...programming and software engineering skills.Ability to work... ...Experience with one or more LLM runtimes, such as TensorRT-LLM, vLLM... ....Contributions to model optimization, inference, compiler, or serving...Full time$195.2k - $361.2k
...run directly on the user's machine (AI PC, edge, on-prem, and beyond... ...are seeking a **Machine Learning Engineer / Data Scientist** to join our... ...-intensive tasks.Debug and optimize training runs — Profile... ...model choices interact with runtime constraints on edge hardwareIMPORTANT...Full timeInternshipLocal areaImmediate startShift work$119.25k - $150.85k
...Inference Solutions team deploys machine learning models from training... ...fast and predictable, and to optimize models so they meet the real... ...Escalade IQ, and we’re hiring engineers to help deliver the next generation... ...for compiler, kernel, runtime, and parity bugs....Full timeInternshipLocal areaWork from homeRelocation packageFlexible hours$151.8k - $265.35k
...adjacent verticals. We are hiring a Senior Machine Learning Engineer to build the pipelines and services... ...for self-serve fine-tuning, optimized VLM deployments for media intelligence... ...quantization, batching, and serving runtimes; custom CUDA a plus. Strong, data-driven...Full timeTemporary workLocal areaWorldwide$193.3k - $261.5k
The Product: AWS Machine Learning accelerators are at the forefront of AWS... ...includes an ML compiler, runtime and natively integrates into... ...disciplines including silicon engineering, hardware design and verification... ...AWS Neuron team works to optimize the performance of complex...InternshipLocal areaWork from homeRelocationFlexible hours$193.3k - $261.5k
...AWS our vision is to make deep learning pervasive for everyday... .... AWS Neuron is the SDK that optimizes the performance of complex ML... ...role is for a senior software engineer in the Compiler team for AWS... ...by side with chip architects, runtime/OS engineers, scientists and...Local areaFlexible hours$160.5k - $240.7k
Company: Qualcomm Technologies, Inc. Job Area: Engineering Group, Engineering Group > Machine Learning Engineering General Summary: About the Role Join the... .... You will develop and supportcutting-edgemodel optimization workflows — pushing the boundary ofwhat'spossible...Work experience placementImmediate startWork from home$195.2k - $361.2k
...efficient models run directly on the user's machine (AI PC, edge, on-prem, and beyond),... ...the hardware people actually own. You optimize inference engines (llama.cpp, vLLM) for constrained... ...engines where it helps usWhat you’ll learn / grow intoCuriosity is required. You...Full timeInternshipLocal areaImmediate startShift work- ...run inside a modern, browser-native runtime (built on technologies such as WebGPU... ...entirely within that runtime. As a Senior Machine Learning Engineer for On-Device & Mobile AI, you will... ...hands-on role. You will own the optimization and deployment of significant parts of...Full timeWork at officeRemote workWorldwide
$206.4k - $379.1k
...Premiere. We are hiring a Principal Machine Learning Engineer to serve as the technical lead for our... ...our generative models are architected, optimized, and served at enterprise scale. You... ...and MLOps technologies — serving runtimes, quantization, GPU scheduling — to improve...Full timeTemporary workLocal areaWorldwide- ...advance your career. THE ROLE:We are looking for a Principal Machine Learning Engineer to join our Models and Applications team. If you are... ...scale.Improve the end-to-end training pipeline performance.Optimize the distributed training pipeline and algorithm to scale out...
$179.4k - $303.6k
...through cutting-edge R&D in AI, machine learning, and smart connectivity.... ...a strong Machine Learning Engineer / Computer Vision Engineer... ...model training, evaluation, optimization, quantization, and... ...quantization accuracy drop, or runtime performance bottlenecks.Familiarity...Full time- ...patients worldwide.We’re a team of engineers, clinicians, and innovators... ...DescriptionAbout This Opportunity:As a Staff Machine Learning Engineer, you will be responsible... ...-time inference, GPU/throughput optimization (e.g., TensorRT, ONNX Runtime, mixed precision), and building...Work at officeLocal areaWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Machine Learning Engineer, Runtime & Optimization. Be the first to apply!
- computer vision machine learning engineer California
- machine learning engineer California
- machine learning scientist California
- machine learning part time California
- machine learning intern California
- machine learning remote California
- machine learning California
- machine learning research scientist California
- internship machine learning California
- artificial intelligence - machine learning intern California


