Staff ML Engineer, Generative Model Performance & Efficiency
$251k - $310kWaymo
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.
The Simulator Team at Waymo builds state-of-the-art simulations of realistic environments for testing, training, and validation of the Waymo Driver. Our team is a diverse, and collaborative group of machine learning (ML) engineers, software engineers, and ML research engineers. We develop industry-leading simulation solutions using advanced generative and reconstructive ML algorithms, to model the real world, encompassing realistic agents, roads, traffic systems, weather, and the full sensor suite (Camera, Lidar, Radar).
To accelerate the fidelity, scalability, controllability, and richness of our simulations, we are pushing the frontiers of 3D world modeling. We leverage state-of-the-art ML technologies trained on large-scale datasets to create dynamic and semantically rich virtual worlds, directly impacting the development and validation of the Waymo Driver.
In this role, you will report to a Senior Staff Engineering Manager
You will:
- Analyze model architectures and identify bottlenecks in training and inference performance (e.g., memory bandwidth, compute, communication).
- Apply and develop techniques such as quantization (e.g., FP8, INT4), pruning, knowledge distillation, and efficient attention mechanisms.
- Optimize model code for specific hardware accelerators (TPUs, GPUs), leveraging compiler features and low-level libraries (e.g., XLA).
- Experiment with different model partitioning and sharding strategies (e.g., data, tensor, pipeline parallelism, expert parallelism) to improve scalability and efficiency.
- Design and implement low-latency, high-throughput serving solutions for generative models and optimize training pipelines to reduce training time.
- Build and maintain tools for performance analysis, profiling (e.g., xprof), and debugging of ML models.
You have:
- MS or PhD in Computer Science, Machine Learning, Robotics, or a related field.
- 5+ years of experience with deep learning architectures (especially Transformers, Diffusion Models, MoEs), algorithms, and optimization techniques.
- Proficiency in JAX, Flax, and potentially TensorFlow/PyTorch.
- Expertise in using profiling tools (e.g., XProf, Perfetto, NVIDIA Nsight) to diagnose performance issues in ML workloads.
- Hands-on experience with quantization, pruning, distillation, and other model compression methods.
- Strong programming skills in Python and potentially C++, with experience in software development best practices.
We prefer:
- Knowledge of TPU and GPU architectures and how to optimize code for them.
- Familiarity with ML compilers like XLA and an understanding of how they translate high-level code to efficient hardware instructions.
- Understanding of concepts related to training and serving models across multiple devices and machines.
- Experience contributing to frameworks and libraries that improve training speed and scalability (e.g., JAX, Gemax, XManager).
The expected base salary range for this full-time position across US locations is listed below. Actual starting pay will be based on job-related factors, including exact work location, experience, relevant training and education, and skill level. Your recruiter can share more about the specific salary range for the role location or, if the role can be performed remote, the specific salary range for your preferred location, during the hiring process.
Waymo employees are also eligible to participate in Waymo’s discretionary annual bonus program, equity incentive plan, and generous Company benefits program, subject to eligibility requirements.
Salary Range
$251,000 - $310,000 USD
- ...space industry delivering next-generation solutions to support the... ...Position Overview: The Staff AI/ML Engineer (LLMs) will lead the... ...Adapt and fine-tune foundation models for specialized use cases... ...based on company and employee performance ~ Company paid life insurance...PerformanceFull timeTemporary workWork at officeVisa sponsorshipRelocation packageFlexible hours
$298k - $368k
...diverse set of sensors, enabling engineers like you to (1) develop methods for efficiently and continuously learning from... ...real-world data, to (2) develop models and model training at scale, to... ...architectures. Optimize model performance for on-device use cases (memory,...PerformanceFull timeRemote work- ...safety, reliability, and efficiency of modern operations. Stack... ...environment. We're looking for a Staff or Senior Engineer to lead the development,... ...the required safety and performance standards are met or... ...fusion of existing detection models to support improving performance...PerformanceRemote jobFull time
$251k - $310k
...group of software engineers, machine learning (ML) engineers, and... ...and enhance the performance of the Waymo... ...machine learning, we model the real world,... ...effort is the generation of foundational... ...build scalable and efficient systems to... ...report to a Sr Staff TLM. You will...PerformanceFull timeRemote work$170k - $230k
...to unlock human performance. WHOOP empowers... ...expert knowledge to generate features that... ...members. As a Staff Machine Learning Engineer on our Clinical... ...and productionize ML systems that... ...support robust model performance.... ...latency, and cost efficiency. Partner with...PerformanceFull timeWork at officeRelocation$242k - $290k
...We enable safe and efficient navigation in... ...systems. As an engineer on the Localization... ...on improving the ML systems that utilize... ...integrate computer vision models into the online... ...ML models to perform vision based localization... ...provide the next generation of mobility-as-a-...PerformanceFull timeTemporary workRelocation package$242k - $333k
...Machine Learning Engineer to join our team to... ...in improving the efficiency and scalability of... ...the intersection of ML and data science.... ...space used by our models to better identify... .... Integrate AV Performance Data: You'll incorporate... ...provide the next generation of mobility-as-a-...PerformanceFull timeTemporary workRelocation package$210k - $260k
...Machine Learning Engineers at Rocket Money further... ...support strategy with ML and AI powered user experiences... ...on end users. At the Staff level, Machine... ...Guide others to help generate impact through effective... ...systematic assessment of model performance, supporting rapid iteration...PerformanceFull timeWork at office- ...building the next generation of AI-powered equity... ...-edge LLMs / ML, large-scale data... ...team of exceptional engineers, analysts, and investors... ...you think about modeling, signals, and... ...re looking for a Staff ML Engineer to lead... ...systems, including performance regressions, bias,...PerformanceFull timeWork at officeLocal area
$220k - $247k
...’ll Do As a Senior Staff Machine Learning Engineer , you will operate at the... ...and shaping the future of generative AI at Typeface. You will... ...the design of large-scale ML systems and shared... ...safety, reliability, and performance at scale Identify and...PerformanceFull timeWork at officeImmediate startFlexible hours3 days per week- ...Marketing we rely on ML to ensure that... ...initiatives by adopting the Generative AI technologies to... ...various AI models, ML services and... ...integrations, and performance optimizations to solve... ...design, and other engineering counterparts to design and build efficient AI solutions for...PerformanceRemote jobFull timeCasual workLive inWork at office
$281k - $356k
...advanced machine learning models to deliver training... ...and software engineers who are passionate about... ...driver to improve the performance of our technology stack... ...strategy for our next generation of machine learning-based... ...in Python and standard ML frameworks (e.g., JAX,...PerformanceFull time$151k - $177.5k
...Us We are Sila, a next-generation battery materials company. Our... ...the inside out today. We engineer and manufacture ground-breaking... ...execution speed Develop Models That Matter Build and... ...product quality and performance outcomes Build feedforward...PerformanceFull time$141k - $249k
...with autonomy and algorithm engineers to scale safe self-driving systems... ...approach. Expand the model deployment pipeline to new... ...embedded systems for the next generation of our onboard compute... ...runtime and memory to pinpoint performance bottlenecks. Qualifications...PerformanceWork at officeWork from homeFlexible hours- ...The Role We're hiring a Staff AI/ML Engineer to help build the AI at the... ...prediction and optimization models that power smarter... ...working fluency across both: Generative / agentic AI — multi-agent... ...salary, stock options, and performance-based bonuses Fully remote...PerformanceRemote jobFull time
$225k
...seeking an experienced Staff Machine Learning Engineer to join our... ...build and deploy AI/ML capabilities that enhance... ...Build and optimize ML model architectures for... ...and optimize ML model performance in production environments... ...retrieval-augmented generation (RAG), or advanced...PerformanceRemote jobFull timeLocal areaImmediate start- ..., our proprietary models, and a sophisticated... ...’ Reasoning Engine and natural language... ...optimizing a cutting-edge Generative AI product that... ...into the performance of our conversational... ...automated metrics for efficient monitoring and... ...Collaborate closely with ML engineers,...PerformanceFull timeWork at officeRemote workFlexible hours
$253.3k - $354.6k
...machine learning models that power... ...Impact As a Staff Machine Learning Engineer , you will own... ...and high-impact ML systems. You will... ...approaches, and ensure efficient, reliable... ...development of next-generation, large-scale... ...teams to build high-performance, distributed...PerformanceRemote jobFull timeFor contractorsWork experience placementFlexible hours$140k - $230k
...capacity. Xometry is seeking a Staff Machine Learning Engineer to lead our Generative AI efforts. This is a rare... ...marketplace. Build cutting-edge models – Develop and deploy large language... ...grow – Guide teammates on advanced ML methods, model architecture, and best...Full timeImmediate start- ...LLMs, our proprietary models, and a sophisticated Agentic... ...Moveworks’ Reasoning Engine and natural language... ...technologies, particularly Generative AI, for business... ...customer perceptions of ML/GAI's business impact.... ...success. We seek high performance, a passion for enhancing...PerformanceFull timeWork at officeRemote workFlexible hours
- ...troubleshoot campaign performance, and drive advertiser growth. As a Staff Machine Learning Engineer focused on Agentic AI... ..., you will lead the ML strategy and execution... ..., feature pipelines, model development, and... ...architectures such as candidate generation, retrieval, ranking,...PerformanceFull timeWork at officeRemote workRelocationRelocation package1 day per week
$251k - $310k
...set of sensors, enabling engineers like you to (1) develop methods for efficiently and continuously... ...world data, to (2) develop models and model training at scale... ...Own tasks in the ML Driver, take responsibility... ...task scaling and task performance, create ML methods and...PerformanceFull timeTemporary workRemote work- ...auctions. In this Sr. Staff MLE role, the... ...strategy for next-generation ranking systems, applying... ...to accelerate modeling innovation,... ...impact advertiser performance, Pinner experience... ...architecture across the ads ML stack, and... ...partners across product, engineering and research....PerformanceFull timeWork at officeRelocationRelocation package
$200k - $235k
...experienced Senior Staff Machine Learning Engineer to join our dynamic... ...of machine learning models in a production environment... ...deliver value and performance at scale. You will... ...-scale datasets efficiently. Stay up-to-date... ...and experience with ML libraries like TensorFlow...PerformanceFull timeLocal areaRemote workRelocation packageFlexible hours2 days per week3 days per week$215.14k - $307.34k
...Spotify is looking for a Staff ML Engineer to join the Native Ads product area in the Music Mission... ...improve overall supply and campaign performance. You will build ML-driven solutions... ...in emerging agent technologies and generative recommender systems You care about...PerformanceRemote jobFull timeFlexible hours$242k - $333k
...As a machine learning engineer within the Attributes team in the... ...enhancing sophisticated behavioral models for various road users,... ...perception attribute models that generate critical signals for our... ...domain knowledge, and interview performance. The salary range listed in this...PerformanceFull timeTemporary workRelocation package$159k - $208.95k
...our organization. As a Staff Machine Learning Engineer at FanDuel, you will... ...critical production ML systems across real-time... ...to help organize, model, and present our data... ...employees may be required to perform other such duties as... ..., and cost-efficiency An inclusive culture...PerformanceFull timeTemporary workLocal areaWorldwide- ...helping improve the safety, efficiency and sustainability of the physical... ...customers as well as core ML infrastructure for Samsara. As a Staff Machine Learning Engineer, you will be working with... ...and optimize the ML model performance on edge devices. Stay connected...PerformanceFull timeRemote workRelocation package
$200k
...relationship. We’re doubling down on ML as the future of Grindr, and... ...to serve millions, balancing performance and innovation. Leverage... ...cross-functionally with engineering, data science and product teams... ...machine/deep learning models with at least one common framework...PerformanceFull timeCasual workWork at officeImmediate startWorldwideFlexible hours$266k - $372.4k
...We’re looking for a Senior Staff Machine Learning Engineer to lead Reddit’s next-generation user understanding initiative... ...deep expertise in mainstream ML user modeling approaches (e.g., large-scale... ...balancing latency, cost, and performance. Reimagine user understanding...PerformanceFull timeFor contractorsWork experience placementImmediate startRemote workFlexible hoursShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff ML Engineer, Generative Model Performance & Efficiency. Be the first to apply!
- assistant engineer Remote
- staff design engineer Remote
- assistant engineering manager Remote
- senior staff systems engineer Remote
- assistant chief engineer Remote
- staff data engineer Remote
- project engineer assistant project manager Remote
- engineering aide Remote
- senior staff engineer Remote
- staff security engineer Remote


