Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff ML Engineer, Generative Model Performance & Efficiency

$251k - $310k
Full-time

Waymo

Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.

Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.

The Simulator Team at Waymo builds state-of-the-art simulations of realistic environments for testing, training, and validation of the Waymo Driver. Our team is a diverse, and collaborative group of machine learning (ML) engineers, software engineers, and ML research engineers. We develop industry-leading simulation solutions using advanced generative and reconstructive ML algorithms, to model the real world, encompassing realistic agents, roads, traffic systems, weather, and the full sensor suite (Camera, Lidar, Radar).

To accelerate the fidelity, scalability, controllability, and richness of our simulations, we are pushing the frontiers of 3D world modeling. We leverage state-of-the-art ML technologies trained on large-scale datasets to create dynamic and semantically rich virtual worlds, directly impacting the development and validation of the Waymo Driver.

In this role, you will report to a Senior Staff Engineering Manager

 

You will:


  • Analyze model architectures and identify bottlenecks in training and inference performance (e.g., memory bandwidth, compute, communication).

  • Apply and develop techniques such as quantization (e.g., FP8, INT4), pruning, knowledge distillation, and efficient attention mechanisms.

  • Optimize model code for specific hardware accelerators (TPUs, GPUs), leveraging compiler features and low-level libraries (e.g., XLA).

  • Experiment with different model partitioning and sharding strategies (e.g., data, tensor, pipeline parallelism, expert parallelism) to improve scalability and efficiency.

  • Design and implement low-latency, high-throughput serving solutions for generative models and optimize training pipelines to reduce training time.

  • Build and maintain tools for performance analysis, profiling (e.g., xprof), and debugging of ML models.

 

You have:


  • MS or PhD in Computer Science, Machine Learning, Robotics, or a related field.

  • 5+ years of experience with deep learning architectures (especially Transformers, Diffusion Models, MoEs), algorithms, and optimization techniques.

  • Proficiency in JAX, Flax, and potentially TensorFlow/PyTorch.

  • Expertise in using profiling tools (e.g., XProf, Perfetto, NVIDIA Nsight) to diagnose performance issues in ML workloads.

  • Hands-on experience with quantization, pruning, distillation, and other model compression methods.

  • Strong programming skills in Python and potentially C++, with experience in software development best practices.

 

We prefer:


  • Knowledge of TPU and GPU architectures and how to optimize code for them.

  • Familiarity with ML compilers like XLA and an understanding of how they translate high-level code to efficient hardware instructions.

  • Understanding of concepts related to training and serving models across multiple devices and machines.

  • Experience contributing to frameworks and libraries that improve training speed and scalability (e.g., JAX, Gemax, XManager).

The expected base salary range for this full-time position across US locations is listed below. Actual starting pay will be based on job-related factors, including exact work location, experience, relevant training and education, and skill level. Your recruiter can share more about the specific salary range for the role location or, if the role can be performed remote, the specific salary range for your preferred location, during the hiring process. 

Waymo employees are also eligible to participate in Waymo’s discretionary annual bonus program, equity incentive plan, and generous Company benefits program, subject to eligibility requirements. 

Salary Range

$251,000 - $310,000 USD

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Staff ML Engineer, Generative Model Performance & Efficiency in Remote vacancy
  • $207k - $300k

    Build advanced AI-generated content (AIGC)...  ...advanced context engineering and agentic feedback...  ...to improve the performance of the AIGC stack...  ...content to optimize efficiency for low-latency...  ...for Large Language Models (e.g., LLM-as-a-...  ...interests. As a Staff Machine Learning... 
    Performance

    Google

    Mountain View, CA
    4 days ago
  •  ...space industry delivering next-generation solutions to support the...  ...Position Overview: The Staff AI/ML Engineer (LLMs) will lead the...  ...Adapt and fine-tune foundation models for specialized use cases...  ...based on company and employee performance ~ Company paid life insurance... 
    Performance
    Full time
    Temporary work
    Work at office
    Visa sponsorship
    Relocation package
    Flexible hours

    Arka Group, L.p.

    Remote
    1 day ago
  • $298k - $368k

     ...diverse set of sensors, enabling engineers like you to (1) develop methods for efficiently and continuously learning from...  ...real-world data, to (2) develop models and model training at scale, to...  ...architectures. Optimize model performance for on-device use cases (memory,... 
    Performance
    Full time
    Remote work

    Waymo

    San Francisco, CA
    1 day ago
  • $249.6k - $299.5k

     ...AV 3.0 strategy. The Scene Generation team builds sensor simulation...  ...Neural Rendering and generative models for Torc's data-driven...  ...across the complete AV stack. As Staff ML Engineer, you will lead this team,...  ...Proficiency with CUDA programming for efficient rendering of large-scale... 
    Suggested
    Full time
    Work experience placement
    Immediate start
    Relocation

    Torc Robotics

    Remote
    1 day ago
  • $251k - $310k

     ...from demonstration, generative modeling, Bayesian inference,...  ...and World models to perform 3D Perception using sensor...  ...effectively with engineering and research teams across...  ..., and implement efficient workflows for model development...  ...), and debugging of ML models. You have:... 
    Performance
    Full time
    Temporary work
    Remote work

    Waymo

    New York, NY
    1 day ago
  • $189.3k - $320.7k

     ...intelligent software, and next-generation safety and...  ...world scenarios.As a Staff ML Engineer on the Prometheus team...  ...teams, including onboard model teams, simulation,...  ....Design and build efficient infrastructure, pipelines...  ...based on company performance, job level, and individual... 
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    4 days ago
  • $185.1k - $335.3k

     ...software that can run efficiently and reliably on...  ...approaches to model export, kernel development, and performance engineering so that every cycle...  ...GM’s next‑generation autonomous and assisted...  .... The RoleAs a Staff Compiler Engineer...  ...and effortless for ML engineers across... 
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Austin, TX
    5 days ago
  • $189k - $300k

     ...software, and next-generation safety and...  ...AV Foundation model development and...  ...on and delivers ML models to the product...  ...AV product performance through smart...  ...team of AI/ML engineers, data scientists...  ...vehicles. As a Staff AI/ML Engineer...  ...Design and build efficient infrastructure,... 
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    4 days ago
  • $250k - $280k

     ...powers breakthrough AI models at leading research...  ...experts for next-generation AI models #...  ...both halves — the engineering throughput of a strong...  ...distributed systems, ML infrastructure, or...  ...designed for scalability, performance, and developer efficiency: Frontend:... 
    Performance
    Full time
    Work at office
    Flexible hours
    3 days per week

    Labelbox

    Remote
    1 day ago
  • $189.3k - $290.7k

     ...software, and next-generation safety and entertainment...  ...scenarios.As a Staff ML Infra Engineer, you will drive the...  ...Autonomous Driving models. From enabling large...  ...pipelines that are performant, easy to use, and exceptionally...  ..., and cost-efficient systems on modern cloud... 
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    3 days ago
  • $161k - $221.5k

     ...partner is looking for a Staff ML Engineer (ML/AI) based in...  ...machine learning and generative AI systems. You will...  ..., reliably, and efficiently. The role combines...  ...from data lineage and model development through production...  ...for high-performance production systems.... 
    Performance
    Full time
    Remote work

    jobgether

    United States
    4 days ago
  • $117.7k - $221.4k

     ..., practical, and cost efficient for embodied AI systems...  .... We believe the next generation of autonomy and...  ...not only on stronger models, but also on better infrastructure...  ...approach that first performs the cheapest reusable...  ...reflects how Cola engineers think: build durable intermediate... 
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    5 days ago
  • $189.3k - $320.7k

     ...software, and next-generation safety and entertainment...  ...scenarios.As a Staff AI/ML Future Sensing Engineer in the Embodied AI...  ...autonomous driving performance, with emphasis on future...  ...perception models and pipelines for detection...  ...Design and build efficient infrastructure,... 
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    5 days ago
  •  ...safety, reliability, and efficiency of modern operations. Stack...  ...environment. We're looking for a Staff or Senior Engineer to lead the development,...  ...the required safety and performance standards are met or...  ...fusion of existing detection models to support improving performance... 
    Performance
    Remote job
    Full time

    Stack Av

    Remote
    1 day ago
  • $170k - $230k

     ...to unlock human performance. WHOOP empowers...  ...expert knowledge to generate features that...  ...members.  As a Staff Machine Learning Engineer on our Clinical...  ...and productionize ML systems that...  ...support robust model performance....  ...latency, and cost efficiency.  Partner with... 
    Performance
    Full time
    Work at office
    Relocation

    Whoop

    Remote
    1 day ago
  • $251k - $310k

     ...group of software engineers, machine learning (ML) engineers, and...  ...and enhance the performance of the Waymo...  ...machine learning, we model the real world,...  ...effort is the generation of foundational...  ...build scalable and efficient systems to...  ...report to a Sr Staff TLM. You will... 
    Performance
    Full time
    Remote work

    Waymo

    Remote
    1 day ago
  • $242k - $290k

     ...We enable safe and efficient navigation in...  ...systems. As an engineer on the Localization...  ...on improving the ML systems that utilize...  ...integrate computer vision models into the online...  ...ML models to perform vision based localization...  ...provide the next generation of mobility-as-a-... 
    Performance
    Full time
    Temporary work
    Relocation package

    Zoox

    Remote
    1 day ago
  • $242k - $333k

     ...Machine Learning Engineer to join our team to...  ...in improving the efficiency and scalability of...  ...the intersection of ML and data science....  ...space used by our models to better identify...  .... Integrate AV Performance Data: You'll incorporate...  ...provide the next generation of mobility-as-a-... 
    Performance
    Full time
    Temporary work
    Relocation package

    Zoox

    Remote
    1 day ago
  • $220k - $247k

     ...’ll Do As a  Senior Staff Machine Learning Engineer , you will operate at the...  ...and shaping the future of generative AI at Typeface.   You will...  ...the design of large-scale ML systems and shared...  ...safety, reliability, and performance at scale   Identify and... 
    Performance
    Full time
    Work at office
    Immediate start
    Flexible hours
    3 days per week

    Typeface

    Remote
    1 day ago
  •  ...building the next generation of AI-powered equity...  ...-edge LLMs / ML, large-scale data...  ...team of exceptional engineers, analysts, and investors...  ...you think about modeling, signals, and...  ...re looking for a Staff ML Engineer to lead...  ...systems, including performance regressions, bias,... 
    Performance
    Full time
    Work at office
    Local area

    Versant Limited

    Remote
    1 day ago
  •  ...Marketing we rely on ML to ensure that...  ...initiatives by adopting the Generative AI technologies to...  ...various AI models, ML services and...  ...integrations, and performance optimizations to solve...  ...design, and other engineering counterparts to design and build efficient AI solutions for... 
    Performance
    Remote job
    Full time
    Casual work
    Live in
    Work at office

    Airbnb, Inc.

    San Francisco, CA
    1 day ago
  •  ...world.Role OverviewAs our Staff Software Engineer, ML infra Engineer for Search...  ...and implementing durable & efficient software solutions that handle...  ...to train complex ML models and efficiently serve them...  ...data structures, algorithms, performance complexity, and... 
    Performance
    Temporary work

    Coupang

    Mountain View, CA
    5 days ago
  • $281k - $356k

     ...advanced machine learning models to deliver training...  ...and software engineers who are passionate about...  ...driver to improve the performance of our technology stack...  ...strategy for our next generation of machine learning-based...  ...in Python and standard ML frameworks (e.g., JAX,... 
    Performance
    Full time

    Waymo

    Remote
    1 day ago
  •  ...building the next generation of AI-driven game...  ...running generative models on-device, right where...  ...Machine Learning Engineer for On-Device &...  ...bars.Do low-level performance work: write and tune...  ...level.Apply efficiency techniques — dynamic...  ...integration between the ML runtime and the... 
    Performance
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    1 day ago
  • $151k - $177.5k

     ...Us We are Sila, a next-generation battery materials company. Our...  ...the inside out today. We engineer and manufacture ground-breaking...  ...execution speed Develop Models That Matter Build and...  ...product quality and performance outcomes Build feedforward... 
    Performance
    Full time

    Sila

    Remote
    1 day ago
  •  ...build cutting-edge foundation AI models and end-to-end products that...  ...is a team of researchers, engineers, designers, and more, who are...  ...still the bottleneck. The Model Efficiency team is responsible for...  ...our preferred locations.As a Staff Research Engineer, you will develop... 
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    5 days ago
  • $230k - $300k

     ...Machine Learning Engineers at Cresta work across...  ...and build next-generation agentic AI systems...  ...requires strong pre-LLM ML foundations, deep...  ..., robustness, and performance of LLM‑powered...  ...pipelines that ground models in enterprise data...  ..., and cost‑efficient LLM‑powered systems... 
    Performance
    Work at office
    Remote work
    Home office
    Flexible hours

    Cresta

    United States
    2 days ago
  • $141k - $249k

     ...with autonomy and algorithm engineers to scale safe self-driving systems...  ...approach. Expand the model deployment pipeline to new...  ...embedded systems for the next generation of our onboard compute...  ...runtime and memory to pinpoint performance bottlenecks. Qualifications... 
    Performance
    Work at office
    Work from home
    Flexible hours

    Waabi

    Pittsburgh, PA
    2 days ago
  •  ...Deep Learning Engineer/Scientist Terra AI is building...  ..., probabilistic modeling, and deep geoscience...  ...confidence and capital efficiency. The company's...  ...In the same way image generators have shown the remarkable...  ...data to improve model performance Adapt diffusion modeling... 
    Performance
    Remote work

    terra.ai Inc.

    United States
    4 days ago
  • $248k - $310k

     ...Marketing we rely on ML to ensure that...  ...by adopting the Generative AI technologies...  ...various AI models, ML services and...  ...As a senior staff machine learning engineer, you will be responsible...  ...models for high-performance deployment on...  ...performance and efficiency. Hands-on prototype... 
    Performance
    Full time
    Work experience placement
    Casual work
    Live in
    Work at office
    Remote work

    Airbnb

    Remote
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff ML Engineer, Generative Model Performance & Efficiency. Be the first to apply!