Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff ML Engineer, Generative Model Performance & Efficiency

$251k - $310k
Full-time

Waymo

Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.

Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.

The Simulator Team at Waymo builds state-of-the-art simulations of realistic environments for testing, training, and validation of the Waymo Driver. Our team is a diverse, and collaborative group of machine learning (ML) engineers, software engineers, and ML research engineers. We develop industry-leading simulation solutions using advanced generative and reconstructive ML algorithms, to model the real world, encompassing realistic agents, roads, traffic systems, weather, and the full sensor suite (Camera, Lidar, Radar).

To accelerate the fidelity, scalability, controllability, and richness of our simulations, we are pushing the frontiers of 3D world modeling. We leverage state-of-the-art ML technologies trained on large-scale datasets to create dynamic and semantically rich virtual worlds, directly impacting the development and validation of the Waymo Driver.

In this role, you will report to a Senior Staff Engineering Manager

 

You will:


  • Analyze model architectures and identify bottlenecks in training and inference performance (e.g., memory bandwidth, compute, communication).

  • Apply and develop techniques such as quantization (e.g., FP8, INT4), pruning, knowledge distillation, and efficient attention mechanisms.

  • Optimize model code for specific hardware accelerators (TPUs, GPUs), leveraging compiler features and low-level libraries (e.g., XLA).

  • Experiment with different model partitioning and sharding strategies (e.g., data, tensor, pipeline parallelism, expert parallelism) to improve scalability and efficiency.

  • Design and implement low-latency, high-throughput serving solutions for generative models and optimize training pipelines to reduce training time.

  • Build and maintain tools for performance analysis, profiling (e.g., xprof), and debugging of ML models.

 

You have:


  • MS or PhD in Computer Science, Machine Learning, Robotics, or a related field.

  • 5+ years of experience with deep learning architectures (especially Transformers, Diffusion Models, MoEs), algorithms, and optimization techniques.

  • Proficiency in JAX, Flax, and potentially TensorFlow/PyTorch.

  • Expertise in using profiling tools (e.g., XProf, Perfetto, NVIDIA Nsight) to diagnose performance issues in ML workloads.

  • Hands-on experience with quantization, pruning, distillation, and other model compression methods.

  • Strong programming skills in Python and potentially C++, with experience in software development best practices.

 

We prefer:


  • Knowledge of TPU and GPU architectures and how to optimize code for them.

  • Familiarity with ML compilers like XLA and an understanding of how they translate high-level code to efficient hardware instructions.

  • Understanding of concepts related to training and serving models across multiple devices and machines.

  • Experience contributing to frameworks and libraries that improve training speed and scalability (e.g., JAX, Gemax, XManager).

The expected base salary range for this full-time position across US locations is listed below. Actual starting pay will be based on job-related factors, including exact work location, experience, relevant training and education, and skill level. Your recruiter can share more about the specific salary range for the role location or, if the role can be performed remote, the specific salary range for your preferred location, during the hiring process. 

Waymo employees are also eligible to participate in Waymo’s discretionary annual bonus program, equity incentive plan, and generous Company benefits program, subject to eligibility requirements. 

Salary Range

$251,000 - $310,000 USD

Vacancy posted 12 hours ago
Similar jobs that could be interesting for youBased on the Staff ML Engineer, Generative Model Performance & Efficiency in Remote vacancy
  •  ...space industry delivering next-generation solutions to support the...  ...Position Overview: The Staff AI/ML Engineer (LLMs) will lead the...  ...Adapt and fine-tune foundation models for specialized use cases...  ...based on company and employee performance ~ Company paid life insurance... 
    Performance
    Full time
    Temporary work
    Work at office
    Visa sponsorship
    Relocation package
    Flexible hours

    Arka Group, L.p.

    Remote
    12 hours ago
  • $298k - $368k

     ...diverse set of sensors, enabling engineers like you to (1) develop methods for efficiently and continuously learning from...  ...real-world data, to (2) develop models and model training at scale, to...  ...architectures. Optimize model performance for on-device use cases (memory,... 
    Performance
    Full time
    Remote work

    Waymo

    San Francisco, CA
    12 hours ago
  •  ...safety, reliability, and efficiency of modern operations. Stack...  ...environment. We're looking for a Staff or Senior Engineer to lead the development,...  ...the required safety and performance standards are met or...  ...fusion of existing detection models to support improving performance... 
    Performance
    Remote job
    Full time

    Stack Av

    Remote
    12 hours ago
  • $251k - $310k

     ...group of software engineers, machine learning (ML) engineers, and...  ...and enhance the performance of the Waymo...  ...machine learning, we model the real world,...  ...effort is the generation of foundational...  ...build scalable and efficient systems to...  ...report to a Sr Staff TLM. You will... 
    Performance
    Full time
    Remote work

    Waymo

    Remote
    12 hours ago
  • $170k - $230k

     ...to unlock human performance. WHOOP empowers...  ...expert knowledge to generate features that...  ...members.  As a Staff Machine Learning Engineer on our Clinical...  ...and productionize ML systems that...  ...support robust model performance....  ...latency, and cost efficiency.  Partner with... 
    Performance
    Full time
    Work at office
    Relocation

    Whoop

    Remote
    12 hours ago
  • $242k - $290k

     ...We enable safe and efficient navigation in...  ...systems. As an engineer on the Localization...  ...on improving the ML systems that utilize...  ...integrate computer vision models into the online...  ...ML models to perform vision based localization...  ...provide the next generation of mobility-as-a-... 
    Performance
    Full time
    Temporary work
    Relocation package

    Zoox

    Remote
    12 hours ago
  • $242k - $333k

     ...Machine Learning Engineer to join our team to...  ...in improving the efficiency and scalability of...  ...the intersection of ML and data science....  ...space used by our models to better identify...  .... Integrate AV Performance Data: You'll incorporate...  ...provide the next generation of mobility-as-a-... 
    Performance
    Full time
    Temporary work
    Relocation package

    Zoox

    Remote
    12 hours ago
  • $210k - $260k

     ...Machine Learning Engineers at Rocket Money further...  ...support strategy with ML and AI powered user experiences...  ...on end users. At the Staff level, Machine...  ...Guide others to help generate impact through effective...  ...systematic assessment of model performance, supporting rapid iteration... 
    Performance
    Full time
    Work at office

    Rocket Money

    Remote
    12 hours ago
  •  ...building the next generation of AI-powered equity...  ...-edge LLMs / ML, large-scale data...  ...team of exceptional engineers, analysts, and investors...  ...you think about modeling, signals, and...  ...re looking for a Staff ML Engineer to lead...  ...systems, including performance regressions, bias,... 
    Performance
    Full time
    Work at office
    Local area

    Versant Limited

    Remote
    12 hours ago
  • $220k - $247k

     ...’ll Do As a  Senior Staff Machine Learning Engineer , you will operate at the...  ...and shaping the future of generative AI at Typeface.   You will...  ...the design of large-scale ML systems and shared...  ...safety, reliability, and performance at scale   Identify and... 
    Performance
    Full time
    Work at office
    Immediate start
    Flexible hours
    3 days per week

    Typeface

    Remote
    12 hours ago
  •  ...Marketing we rely on ML to ensure that...  ...initiatives by adopting the Generative AI technologies to...  ...various AI models, ML services and...  ...integrations, and performance optimizations to solve...  ...design, and other engineering counterparts to design and build efficient AI solutions for... 
    Performance
    Remote job
    Full time
    Casual work
    Live in
    Work at office

    Airbnb, Inc.

    San Francisco, CA
    12 hours ago
  • $281k - $356k

     ...advanced machine learning models to deliver training...  ...and software engineers who are passionate about...  ...driver to improve the performance of our technology stack...  ...strategy for our next generation of machine learning-based...  ...in Python and standard ML frameworks (e.g., JAX,... 
    Performance
    Full time

    Waymo

    Remote
    12 hours ago
  • $151k - $177.5k

     ...Us We are Sila, a next-generation battery materials company. Our...  ...the inside out today. We engineer and manufacture ground-breaking...  ...execution speed Develop Models That Matter Build and...  ...product quality and performance outcomes Build feedforward... 
    Performance
    Full time

    Sila

    Remote
    12 hours ago
  • $141k - $249k

     ...with autonomy and algorithm engineers to scale safe self-driving systems...  ...approach. Expand the model deployment pipeline to new...  ...embedded systems for the next generation of our onboard compute...  ...runtime and memory to pinpoint performance bottlenecks. Qualifications... 
    Performance
    Work at office
    Work from home
    Flexible hours

    Waabi

    Pittsburgh, PA
    2 days ago
  •  ...The Role We're hiring a Staff AI/ML Engineer to help build the AI at the...  ...prediction and optimization models that power smarter...  ...working fluency across both: Generative / agentic AI — multi-agent...  ...salary, stock options, and performance-based bonuses Fully remote... 
    Performance
    Remote job
    Full time

    Burq

    United States
    12 hours ago
  • $225k

     ...seeking an experienced Staff Machine Learning Engineer to join our...  ...build and deploy AI/ML capabilities that enhance...  ...Build and optimize ML model architectures for...  ...and optimize ML model performance in production environments...  ...retrieval-augmented generation (RAG), or advanced... 
    Performance
    Remote job
    Full time
    Local area
    Immediate start

    Dragos

    United States
    12 hours ago
  •  ..., our proprietary models, and a sophisticated...  ...’ Reasoning Engine and natural language...  ...optimizing a cutting-edge Generative AI product that...  ...into the performance of our conversational...  ...automated metrics for efficient monitoring and...  ...Collaborate closely with ML engineers,... 
    Performance
    Full time
    Work at office
    Remote work
    Flexible hours

    Servicenow

    Remote
    12 hours ago
  • $253.3k - $354.6k

     ...machine learning models that power...  ...Impact As a Staff Machine Learning Engineer , you will own...  ...and high-impact ML systems. You will...  ...approaches, and ensure efficient, reliable...  ...development of next-generation, large-scale...  ...teams to build high-performance, distributed... 
    Performance
    Remote job
    Full time
    For contractors
    Work experience placement
    Flexible hours

    Reddit

    United States
    12 hours ago
  • $140k - $230k

     ...capacity. Xometry is seeking a Staff Machine Learning Engineer to lead our Generative AI efforts. This is a rare...  ...marketplace. Build cutting-edge models – Develop and deploy large language...  ...grow – Guide teammates on advanced ML methods, model architecture, and best... 
    Full time
    Immediate start

    Xometry

    Remote
    12 hours ago
  •  ...LLMs, our proprietary models, and a sophisticated Agentic...  ...Moveworks’ Reasoning Engine and natural language...  ...technologies, particularly Generative AI, for business...  ...customer perceptions of ML/GAI's business impact....  ...success. We seek high performance, a passion for enhancing... 
    Performance
    Full time
    Work at office
    Remote work
    Flexible hours

    Servicenow

    Remote
    12 hours ago
  •  ...troubleshoot campaign performance, and drive advertiser growth. As a Staff Machine Learning Engineer focused on Agentic AI...  ..., you will lead the ML strategy and execution...  ..., feature pipelines, model development, and...  ...architectures such as candidate generation, retrieval, ranking,... 
    Performance
    Full time
    Work at office
    Remote work
    Relocation
    Relocation package
    1 day per week

    Pinterest

    San Francisco, CA
    12 hours ago
  • $251k - $310k

     ...set of sensors, enabling engineers like you to (1) develop methods for efficiently and continuously...  ...world data, to (2) develop models and model training at scale...  ...Own tasks in the ML Driver, take responsibility...  ...task scaling and task performance, create ML methods and... 
    Performance
    Full time
    Temporary work
    Remote work

    Waymo

    San Francisco, CA
    12 hours ago
  •  ...auctions. In this Sr. Staff MLE role, the...  ...strategy for next-generation ranking systems, applying...  ...to accelerate modeling innovation,...  ...impact advertiser performance, Pinner experience...  ...architecture across the ads ML stack, and...  ...partners across product, engineering and research.... 
    Performance
    Full time
    Work at office
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    12 hours ago
  • $200k - $235k

     ...experienced Senior Staff Machine Learning Engineer to join our dynamic...  ...of machine learning models in a production environment...  ...deliver value and performance at scale. You will...  ...-scale datasets efficiently. Stay up-to-date...  ...and experience with ML libraries like TensorFlow... 
    Performance
    Full time
    Local area
    Remote work
    Relocation package
    Flexible hours
    2 days per week
    3 days per week

    Flex

    United States
    12 hours ago
  • $215.14k - $307.34k

     ...Spotify is looking for a Staff ML Engineer to join the Native Ads product area in the Music Mission...  ...improve overall supply and campaign performance. You will build ML-driven solutions...  ...in emerging agent technologies and generative recommender systems You care about... 
    Performance
    Remote job
    Full time
    Flexible hours

    Spotify

    Remote
    12 hours ago
  • $242k - $333k

     ...As a machine learning engineer within the Attributes team in the...  ...enhancing sophisticated behavioral models for various road users,...  ...perception attribute models that generate critical signals for our...  ...domain knowledge, and interview performance. The salary range listed in this... 
    Performance
    Full time
    Temporary work
    Relocation package

    Zoox

    Remote
    12 hours ago
  • $159k - $208.95k

     ...our organization. As a Staff Machine Learning Engineer at FanDuel, you will...  ...critical production ML systems across real-time...  ...to help organize, model, and present our data...  ...employees may be required to perform other such duties as...  ..., and cost-efficiency An inclusive culture... 
    Performance
    Full time
    Temporary work
    Local area
    Worldwide

    Fanduel

    Remote
    12 hours ago
  •  ...helping improve the safety, efficiency and sustainability of the physical...  ...customers as well as core ML infrastructure for Samsara. As a Staff Machine Learning Engineer, you will be working with...  ...and optimize the ML model performance on edge devices. Stay connected... 
    Performance
    Full time
    Remote work
    Relocation package

    Samsara

    Remote
    12 hours ago
  • $200k

     ...relationship. We’re doubling down on ML as the future of Grindr, and...  ...to serve millions, balancing performance and innovation. Leverage...  ...cross-functionally with engineering, data science and product teams...  ...machine/deep learning models with at least one common framework... 
    Performance
    Full time
    Casual work
    Work at office
    Immediate start
    Worldwide
    Flexible hours

    Grindr

    Remote
    12 hours ago
  • $266k - $372.4k

     ...We’re looking for a Senior Staff Machine Learning Engineer to lead Reddit’s next-generation user understanding initiative...  ...deep expertise in mainstream ML user modeling approaches (e.g., large-scale...  ...balancing latency, cost, and performance. Reimagine user understanding... 
    Performance
    Full time
    For contractors
    Work experience placement
    Immediate start
    Remote work
    Flexible hours
    Shift work

    Reddit

    United States
    12 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff ML Engineer, Generative Model Performance & Efficiency. Be the first to apply!