Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Machine Learning Engineer, Runtime & Optimization

$213k - $263k
Full-time

Waymo

Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.

The ML Platform team at Waymo provides a set of tools to support and automate the lifecycle of the machine learning workflow, including feature and experiment management, model development, optimization and monitoring. These efforts have resulted in making machine learning more accessible to teams at Waymo, including Perception, Planner, Research and Simulation.

We are looking for engineers with ML software or ML systems expertise to help us improve compute performance on both cloud and car. You'll work across the entire ML stack from the system perspective, from efficient deep learning models, model compression, ML software (e.g. JAX, XLA, Triton, and CUDA), to . You will be pleasantly challenged with deploying Waymo ML models on limited computation resources. In this hybrid role, you will report to the Senior Manager of Runtime and Optimization.

You Will

  • Lead the collaboration with the world-class Waymo ML scientists in perception, planner, research and simulation. Identify opportunities in both systems and models to make ML workloads faster.
  • Lead projects from proposals through execution by developing junior engineers.
  • Analyze and improve ML system workloads on both cloud and self-driving cars .
  • Apply model optimization, efficient deep learning techniques and ML software improvements to Waymo's ML systems.

You Have

  • M.S. in CS, EE, Deep Learning or a related field
  • 2+ years of experience as a technical lead, including writing project plans, engaging with customer teams, mentoring, responsible for goals & execution, reporting status.
  • 5+ years of experience developing solutions in ML systems or ML software stack (Pytorch/JAX/TF, runtime libraries, ML compiler).
  • Deep understanding of ML system architecture, performance analysis and tools.
  • Strong Python or C++ programming skills

We prefer you have one or more of the following:

  • PhD in CS, EE, Deep Learning or a related field.
  • Familiarity with the HW architecture of ML hardware accelerators (e.g., GPU/TPU).
  • Deep knowledge of model optimization or efficient deep learning techniques for foundation models or LLM.
  • Experience with GPU HW or TPU HW and related system software.

The expected base salary range for this full-time position across US locations is listed below. Actual starting pay will be based on job-related factors, including exact work location, experience, relevant training and education, and skill level. Your recruiter can share more about the specific salary range for the role location or, if the role can be performed remote, the specific salary range for your preferred location, during the hiring process.

Waymo employees are also eligible to participate in Waymo’s discretionary annual bonus program, equity incentive plan, and generous Company benefits program, subject to eligibility requirements.

Salary Range

$213,000—$263,000 USD

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the Machine Learning Engineer, Runtime & Optimization in California vacancy
  • $159.05k - $199.3k

     ...intelligence to every moving machine on the planet. Applied...  ...; Seoul; and Tokyo. Learn more at applied.co. We...  ...are looking for a software engineer with deep experience in optimizing ML models and deploying them...  ...-grade embedded runtime environments. You’ll work... 
    Suggested
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Remote work
    Day shift

    Applied Intuition

    California
    16 days ago
  • $137.1k - $201.6k

     ...The mission of the Marketplace Optimization team is to ensure we maintain a healthy...  ...artificial intelligence and advanced ML, deep learning techniques to power decision-making...  ...the Role We’re looking for a Machine Learning Engineer to help design, build, optimize and scale... 
    Suggested
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash Usa

    San Francisco, CA
    1 day ago
  •  ...the TeamThe mission of the Marketplace Optimization team is to ensure we maintain a...  ...artificial intelligence and advanced ML, deep learning techniques to power decision-making...  .... About the RoleWe’re looking for a Machine Learning Engineer to help design, build, optimize and... 
    Suggested
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    3 days ago
  • $170k - $216k

     ...team builds the system which learns the spatial-temporal representation...  ...with downstream teams on the optimization and integration into the...  ...set of sensors, enabling engineers like you to (1) develop methods...  ...experience ~3+ years experience in Machine Learning and/or Computer... 
    Suggested
    Full time
    Remote work

    Waymo

    San Francisco, CA
    1 day ago
  • $260k - $330k

     ...Must have: Experience building large-scale prediction or optimization systemsPubMatic is the leading AI-powered ad tech company...  ...environments.About the Role:We are looking for a Senior Principal Machine Learning Engineer to help build the next generation of performance... 
    Suggested
    Work at office
    Remote work

    PubMatic

    Redwood City, CA
    4 days ago
  • $298k - $368k

     ...team builds the system which learns the spatial-temporal representation...  ...with downstream teams on the optimization and integration into the...  ...set of sensors, enabling engineers like you to (1) develop methods...  ...~7+ years of experience in Machine Learning, with a focus on large... 
    Full time
    Remote work

    Waymo

    San Francisco, CA
    1 day ago
  • $216.7k - $303.4k

     ...Description This role sits in the Ads Optimization organizations, which are responsible for...  .... You’ll join a set of tight-knit engineers working on high-impact, internet-scale...  ...Ads. Role Description We are hiring Machine Learning Engineers (IC4) to build and evolve the... 
    For contractors
    Work experience placement
    Flexible hours

    Reddit, Inc.

    San Francisco, CA
    5 days ago
  • $160k - $230k

     ...About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI...  ...performance at scale. Develop and optimize runtime inference services for large-scale AI applications... 
    Full time

    Together Ai

    San Francisco, CA
    1 day ago
  • $150k - $200k

     ...collision and force guards, smoothing, and runtime monitors for safety-critical deployment...  ...-speed robot autonomy software stack optimized for inference performance ●...  ...PhD or MS degree in Computer Science, Machine Learning, Robotics, or equivalent technical discipline... 
    Full time

    Deft Ai, Inc.

    San Francisco, CA
    1 day ago
  •  ...with hands-on support from AMD engineers the team is scaling rapidly...  ...-training and Reinforcement Learning , sandbox environments for...  ...and deployment + inference optimization . You’ll build and iterate...  ...and optimize GPU kernels, runtime bottlenecks, and memory behavior... 
    Full time
    Flexible hours

    Sciforium

    San Francisco, CA
    1 day ago
  • $250k - $350k

     ...are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's...  ...experiences at scale.You will design and optimize inference pipelines, implement state-of-...  ..., attention acceleration, and deep learning compiler stacks.GPU & Parallelism: Deep... 
    Work at office
    3 days per week

    Pika

    Palo Alto, CA
    4 days ago
  • $120k - $160k

     ...are unavailable or unreliable. As a Machine Learning Engineer, you will own and scale the training,...  ...validate to redeploy loops.Deploy and optimize models for real-time edge inference on...  ...(quantization/pruning, TensorRT/ONNX Runtime); profile CPU/GPU and hit tight... 
    Work experience placement
    Work at office
    Local area
    Shift work

    Mach Industries

    Huntington Beach, CA
    4 days ago
  • $204k - $259k

    An autonomous driving technology company is looking for an experienced engineer to improve compute performance in machine learning systems. This hybrid role involves collaboration with a world-class ML team and requires strong expertise in ML software or systems. The ideal... 

    Waymo

    Mountain View, CA
    1 day ago
  • $195.2k - $262.2k

     ...infrastructure. Built by engineers, for engineers. From large-...  ...orchestration to inference optimization, we own the hard problems across...  ...across model code, kernels, runtime, scheduler, gateway, and...  ...compensation Career growth and learning opportunities... 
    Full time
    Temporary work
    Immediate start
    Remote work

    Nebius

    Palo Alto, CA
    2 days ago
  • $213k - $263k

     ...across 15+ U.S. states. The ML Optimization team at Waymo provides a...  ...the lifecycle of the machine learning workflow, including feature...  ...Simulation. We are looking for engineers with ML software & systems...  ...to the Senior Manager of Runtime and Optimization. You... 
    Full time
    Remote work

    Waymo

    Mountain View, CA
    3 days ago
  • Apple is seeking an ML Infrastructure Engineer focused on ML user experience APIs and integration for on-device AI. You...  ...to integrate these APIs into internal and external model repositories, optimize pipelines from authoring to runtime, and #J-18808-Ljbffr Socket.dev

    Socket.dev

    Cupertino, CA
    1 day ago
  • $180k - $220k

     ...build sensors and tools for engineers, roboticists, and researchers...  ...for a highly technical Machine Learning Engineer to lead our efforts...  ...research papers into code, and optimizing these systems for real-time,...  ...Experience with TensorRT, ONNX Runtime, or edge-specific hardware (... 
    Full time
    Work experience placement
    Local area

    Ouster

    San Francisco, CA
    1 day ago
  • $174.72k - $295.68k

     ...through cutting-edge R&D in AI, machine learning, and smart connectivity.Our...  ...programming and software engineering skills.Ability to work...  ...Experience with one or more LLM runtimes, such as TensorRT-LLM, vLLM...  ....Contributions to model optimization, inference, compiler, or serving... 
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $195.2k - $361.2k

     ...run directly on the user's machine (AI PC, edge, on-prem, and beyond...  ...are seeking a **Machine Learning Engineer / Data Scientist** to join our...  ...-intensive tasks.Debug and optimize training runs — Profile...  ...model choices interact with runtime constraints on edge hardwareIMPORTANT... 
    Full time
    Internship
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    3 days ago
  • $119.25k - $150.85k

     ...Inference Solutions team deploys machine learning models from training...  ...fast and predictable, and to optimize models so they meet the real...  ...Escalade IQ, and we’re hiring engineers to help deliver the next generation...  ...for compiler, kernel, runtime, and parity bugs.... 
    Full time
    Internship
    Local area
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • $151.8k - $265.35k

     ...adjacent verticals. We are hiring a Senior Machine Learning Engineer to build the pipelines and services...  ...for self-serve fine-tuning, optimized VLM deployments for media intelligence...  ...quantization, batching, and serving runtimes; custom CUDA a plus. Strong, data-driven... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    2 days ago
  • $193.3k - $261.5k

    The Product: AWS Machine Learning accelerators are at the forefront of AWS...  ...includes an ML compiler, runtime and natively integrates into...  ...disciplines including silicon engineering, hardware design and verification...  ...AWS Neuron team works to optimize the performance of complex... 
    Internship
    Local area
    Work from home
    Relocation
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $193.3k - $261.5k

     ...AWS our vision is to make deep learning pervasive for everyday...  .... AWS Neuron is the SDK that optimizes the performance of complex ML...  ...role is for a senior software engineer in the Compiler team for AWS...  ...by side with chip architects, runtime/OS engineers, scientists and... 
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $160.5k - $240.7k

    Company: Qualcomm Technologies, Inc. Job Area: Engineering Group, Engineering Group > Machine Learning Engineering General Summary: About the Role Join the...  .... You will develop and supportcutting-edgemodel optimization workflows — pushing the boundary ofwhat'spossible... 
    Work experience placement
    Immediate start
    Work from home

    Socket.dev

    Santa Clara, CA
    4 days ago
  • $195.2k - $361.2k

     ...efficient models run directly on the user's machine (AI PC, edge, on-prem, and beyond),...  ...the hardware people actually own. You optimize inference engines (llama.cpp, vLLM) for constrained...  ...engines where it helps usWhat you’ll learn / grow intoCuriosity is required. You... 
    Full time
    Internship
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    3 days ago
  •  ...run inside a modern, browser-native runtime (built on technologies such as WebGPU...  ...entirely within that runtime. As a Senior Machine Learning Engineer for On-Device & Mobile AI, you will...  ...hands-on role. You will own the optimization and deployment of significant parts of... 
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    4 days ago
  • $206.4k - $379.1k

     ...Premiere. We are hiring a Principal Machine Learning Engineer to serve as the technical lead for our...  ...our generative models are architected, optimized, and served at enterprise scale. You...  ...and MLOps technologies — serving runtimes, quantization, GPU scheduling — to improve... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    3 days ago
  •  ...advance your career. THE ROLE:We are looking for a Principal Machine Learning Engineer to join our Models and Applications team. If you are...  ...scale.Improve the end-to-end training pipeline performance.Optimize the distributed training pipeline and algorithm to scale out... 

    AMD

    San Jose, CA
    9 hours ago
  • $179.4k - $303.6k

     ...through cutting-edge R&D in AI, machine learning, and smart connectivity....  ...a strong Machine Learning Engineer / Computer Vision Engineer...  ...model training, evaluation, optimization, quantization, and...  ...quantization accuracy drop, or runtime performance bottlenecks.Familiarity... 
    Full time

    XPENG Motors

    Santa Clara, CA
    3 days ago
  •  ...patients worldwide.We’re a team of engineers, clinicians, and innovators...  ...DescriptionAbout This Opportunity:As a Staff Machine Learning Engineer, you will be responsible...  ...-time inference, GPU/throughput optimization (e.g., TensorRT, ONNX Runtime, mixed precision), and building... 
    Work at office
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    9 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Machine Learning Engineer, Runtime & Optimization. Be the first to apply!