Staff ML Engineer, Generative Model Performance & Efficiency
$251k - $310kWaymo
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.
The Simulator Team at Waymo builds state-of-the-art simulations of realistic environments for testing, training, and validation of the Waymo Driver. Our team is a diverse, and collaborative group of machine learning (ML) engineers, software engineers, and ML research engineers. We develop industry-leading simulation solutions using advanced generative and reconstructive ML algorithms, to model the real world, encompassing realistic agents, roads, traffic systems, weather, and the full sensor suite (Camera, Lidar, Radar).
To accelerate the fidelity, scalability, controllability, and richness of our simulations, we are pushing the frontiers of 3D world modeling. We leverage state-of-the-art ML technologies trained on large-scale datasets to create dynamic and semantically rich virtual worlds, directly impacting the development and validation of the Waymo Driver.
In this role, you will report to a Senior Staff Engineering Manager
You will:
- Analyze model architectures and identify bottlenecks in training and inference performance (e.g., memory bandwidth, compute, communication).
- Apply and develop techniques such as quantization (e.g., FP8, INT4), pruning, knowledge distillation, and efficient attention mechanisms.
- Optimize model code for specific hardware accelerators (TPUs, GPUs), leveraging compiler features and low-level libraries (e.g., XLA).
- Experiment with different model partitioning and sharding strategies (e.g., data, tensor, pipeline parallelism, expert parallelism) to improve scalability and efficiency.
- Design and implement low-latency, high-throughput serving solutions for generative models and optimize training pipelines to reduce training time.
- Build and maintain tools for performance analysis, profiling (e.g., xprof), and debugging of ML models.
You have:
- MS or PhD in Computer Science, Machine Learning, Robotics, or a related field.
- 5+ years of experience with deep learning architectures (especially Transformers, Diffusion Models, MoEs), algorithms, and optimization techniques.
- Proficiency in JAX, Flax, and potentially TensorFlow/PyTorch.
- Expertise in using profiling tools (e.g., XProf, Perfetto, NVIDIA Nsight) to diagnose performance issues in ML workloads.
- Hands-on experience with quantization, pruning, distillation, and other model compression methods.
- Strong programming skills in Python and potentially C++, with experience in software development best practices.
We prefer:
- Knowledge of TPU and GPU architectures and how to optimize code for them.
- Familiarity with ML compilers like XLA and an understanding of how they translate high-level code to efficient hardware instructions.
- Understanding of concepts related to training and serving models across multiple devices and machines.
- Experience contributing to frameworks and libraries that improve training speed and scalability (e.g., JAX, Gemax, XManager).
The expected base salary range for this full-time position across US locations is listed below. Actual starting pay will be based on job-related factors, including exact work location, experience, relevant training and education, and skill level. Your recruiter can share more about the specific salary range for the role location or, if the role can be performed remote, the specific salary range for your preferred location, during the hiring process.
Waymo employees are also eligible to participate in Waymo’s discretionary annual bonus program, equity incentive plan, and generous Company benefits program, subject to eligibility requirements.
Salary Range
$251,000 - $310,000 USD
$207k - $300k
Build advanced AI-generated content (AIGC)... ...advanced context engineering and agentic feedback... ...to improve the performance of the AIGC stack... ...content to optimize efficiency for low-latency... ...for Large Language Models (e.g., LLM-as-a-... ...interests. As a Staff Machine Learning...Performance- ...space industry delivering next-generation solutions to support the... ...Position Overview: The Staff AI/ML Engineer (LLMs) will lead the... ...Adapt and fine-tune foundation models for specialized use cases... ...based on company and employee performance ~ Company paid life insurance...PerformanceFull timeTemporary workWork at officeVisa sponsorshipRelocation packageFlexible hours
$298k - $368k
...diverse set of sensors, enabling engineers like you to (1) develop methods for efficiently and continuously learning from... ...real-world data, to (2) develop models and model training at scale, to... ...architectures. Optimize model performance for on-device use cases (memory,...PerformanceFull timeRemote work$249.6k - $299.5k
...AV 3.0 strategy. The Scene Generation team builds sensor simulation... ...Neural Rendering and generative models for Torc's data-driven... ...across the complete AV stack. As Staff ML Engineer, you will lead this team,... ...Proficiency with CUDA programming for efficient rendering of large-scale...SuggestedFull timeWork experience placementImmediate startRelocation$251k - $310k
...from demonstration, generative modeling, Bayesian inference,... ...and World models to perform 3D Perception using sensor... ...effectively with engineering and research teams across... ..., and implement efficient workflows for model development... ...), and debugging of ML models. You have:...PerformanceFull timeTemporary workRemote work$189.3k - $320.7k
...intelligent software, and next-generation safety and... ...world scenarios.As a Staff ML Engineer on the Prometheus team... ...teams, including onboard model teams, simulation,... ....Design and build efficient infrastructure, pipelines... ...based on company performance, job level, and individual...PerformanceFull timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours$185.1k - $335.3k
...software that can run efficiently and reliably on... ...approaches to model export, kernel development, and performance engineering so that every cycle... ...GM’s next‑generation autonomous and assisted... .... The RoleAs a Staff Compiler Engineer... ...and effortless for ML engineers across...PerformanceFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours$189k - $300k
...software, and next-generation safety and... ...AV Foundation model development and... ...on and delivers ML models to the product... ...AV product performance through smart... ...team of AI/ML engineers, data scientists... ...vehicles. As a Staff AI/ML Engineer... ...Design and build efficient infrastructure,...PerformanceFull timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours$250k - $280k
...powers breakthrough AI models at leading research... ...experts for next-generation AI models #... ...both halves — the engineering throughput of a strong... ...distributed systems, ML infrastructure, or... ...designed for scalability, performance, and developer efficiency: Frontend:...PerformanceFull timeWork at officeFlexible hours3 days per week$189.3k - $290.7k
...software, and next-generation safety and entertainment... ...scenarios.As a Staff ML Infra Engineer, you will drive the... ...Autonomous Driving models. From enabling large... ...pipelines that are performant, easy to use, and exceptionally... ..., and cost-efficient systems on modern cloud...PerformanceFull timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours$161k - $221.5k
...partner is looking for a Staff ML Engineer (ML/AI) based in... ...machine learning and generative AI systems. You will... ..., reliably, and efficiently. The role combines... ...from data lineage and model development through production... ...for high-performance production systems....PerformanceFull timeRemote work$117.7k - $221.4k
..., practical, and cost efficient for embodied AI systems... .... We believe the next generation of autonomy and... ...not only on stronger models, but also on better infrastructure... ...approach that first performs the cheapest reusable... ...reflects how Cola engineers think: build durable intermediate...PerformanceFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours$189.3k - $320.7k
...software, and next-generation safety and entertainment... ...scenarios.As a Staff AI/ML Future Sensing Engineer in the Embodied AI... ...autonomous driving performance, with emphasis on future... ...perception models and pipelines for detection... ...Design and build efficient infrastructure,...PerformanceFull timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours- ...safety, reliability, and efficiency of modern operations. Stack... ...environment. We're looking for a Staff or Senior Engineer to lead the development,... ...the required safety and performance standards are met or... ...fusion of existing detection models to support improving performance...PerformanceRemote jobFull time
$170k - $230k
...to unlock human performance. WHOOP empowers... ...expert knowledge to generate features that... ...members. As a Staff Machine Learning Engineer on our Clinical... ...and productionize ML systems that... ...support robust model performance.... ...latency, and cost efficiency. Partner with...PerformanceFull timeWork at officeRelocation$251k - $310k
...group of software engineers, machine learning (ML) engineers, and... ...and enhance the performance of the Waymo... ...machine learning, we model the real world,... ...effort is the generation of foundational... ...build scalable and efficient systems to... ...report to a Sr Staff TLM. You will...PerformanceFull timeRemote work$242k - $290k
...We enable safe and efficient navigation in... ...systems. As an engineer on the Localization... ...on improving the ML systems that utilize... ...integrate computer vision models into the online... ...ML models to perform vision based localization... ...provide the next generation of mobility-as-a-...PerformanceFull timeTemporary workRelocation package$242k - $333k
...Machine Learning Engineer to join our team to... ...in improving the efficiency and scalability of... ...the intersection of ML and data science.... ...space used by our models to better identify... .... Integrate AV Performance Data: You'll incorporate... ...provide the next generation of mobility-as-a-...PerformanceFull timeTemporary workRelocation package$220k - $247k
...’ll Do As a Senior Staff Machine Learning Engineer , you will operate at the... ...and shaping the future of generative AI at Typeface. You will... ...the design of large-scale ML systems and shared... ...safety, reliability, and performance at scale Identify and...PerformanceFull timeWork at officeImmediate startFlexible hours3 days per week- ...building the next generation of AI-powered equity... ...-edge LLMs / ML, large-scale data... ...team of exceptional engineers, analysts, and investors... ...you think about modeling, signals, and... ...re looking for a Staff ML Engineer to lead... ...systems, including performance regressions, bias,...PerformanceFull timeWork at officeLocal area
- ...Marketing we rely on ML to ensure that... ...initiatives by adopting the Generative AI technologies to... ...various AI models, ML services and... ...integrations, and performance optimizations to solve... ...design, and other engineering counterparts to design and build efficient AI solutions for...PerformanceRemote jobFull timeCasual workLive inWork at office
- ...world.Role OverviewAs our Staff Software Engineer, ML infra Engineer for Search... ...and implementing durable & efficient software solutions that handle... ...to train complex ML models and efficiently serve them... ...data structures, algorithms, performance complexity, and...PerformanceTemporary work
$281k - $356k
...advanced machine learning models to deliver training... ...and software engineers who are passionate about... ...driver to improve the performance of our technology stack... ...strategy for our next generation of machine learning-based... ...in Python and standard ML frameworks (e.g., JAX,...PerformanceFull time- ...building the next generation of AI-driven game... ...running generative models on-device, right where... ...Machine Learning Engineer for On-Device &... ...bars.Do low-level performance work: write and tune... ...level.Apply efficiency techniques — dynamic... ...integration between the ML runtime and the...PerformanceFull timeWork at officeRemote workWorldwide
$151k - $177.5k
...Us We are Sila, a next-generation battery materials company. Our... ...the inside out today. We engineer and manufacture ground-breaking... ...execution speed Develop Models That Matter Build and... ...product quality and performance outcomes Build feedforward...PerformanceFull time- ...build cutting-edge foundation AI models and end-to-end products that... ...is a team of researchers, engineers, designers, and more, who are... ...still the bottleneck. The Model Efficiency team is responsible for... ...our preferred locations.As a Staff Research Engineer, you will develop...Work at officeLocal areaRemote workHome office
$230k - $300k
...Machine Learning Engineers at Cresta work across... ...and build next-generation agentic AI systems... ...requires strong pre-LLM ML foundations, deep... ..., robustness, and performance of LLM‑powered... ...pipelines that ground models in enterprise data... ..., and cost‑efficient LLM‑powered systems...PerformanceWork at officeRemote workHome officeFlexible hours$141k - $249k
...with autonomy and algorithm engineers to scale safe self-driving systems... ...approach. Expand the model deployment pipeline to new... ...embedded systems for the next generation of our onboard compute... ...runtime and memory to pinpoint performance bottlenecks. Qualifications...PerformanceWork at officeWork from homeFlexible hours- ...Deep Learning Engineer/Scientist Terra AI is building... ..., probabilistic modeling, and deep geoscience... ...confidence and capital efficiency. The company's... ...In the same way image generators have shown the remarkable... ...data to improve model performance Adapt diffusion modeling...PerformanceRemote work
$248k - $310k
...Marketing we rely on ML to ensure that... ...by adopting the Generative AI technologies... ...various AI models, ML services and... ...As a senior staff machine learning engineer, you will be responsible... ...models for high-performance deployment on... ...performance and efficiency. Hands-on prototype...PerformanceFull timeWork experience placementCasual workLive inWork at officeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff ML Engineer, Generative Model Performance & Efficiency. Be the first to apply!
- software engineer staff Remote
- assistant engineer Remote
- engineering aide Remote
- staff engineer Remote
- staff security engineer Remote
- assistant engineering manager Remote
- senior staff systems engineer Remote
- technology administrator Remote
- project engineer assistant project manager Remote
- senior staff engineer Remote



