Senior ML Inference Engineer - Platform
$129k - $261kGeneral Motors Proving Ground
Hybrid - Sunnyvale, California, United States of America May 6, 2026
$129K - $261K
Job Description The Model Deployment & Inference Solutions team in GM AV deploys machine learning models from training frameworks (e.g. PyTorch) onto autonomous vehicle hardware. Our mission is two‑fold: build the ML deployment platform that makes model rollouts fast and predictable, and optimize models so they meet the real‑time latency and memory budgets required to run on‑vehicle. Our work is on the critical path of GM's publicly committed launch of eyes‑off (hands‑free, eyes‑free) autonomous driving in 2028, debuting on the Cadillac Escalade IQ, building on Super Cruise's billion‑plus hands‑free miles. About the Team The Model Deployment & Inference Solutions team in GM AV deploys machine learning models from training frameworks (e.g. PyTorch) onto autonomous vehicle hardware. Our mission is two‑fold: build the ML deployment platform that makes model rollouts fast and predictable, and optimize models so they meet the real‑time latency and memory budgets required to run on‑vehicle. Our work is on the critical path of GM's publicly committed launch of eyes‑off (hands‑free, eyes‑free) autonomous driving in 2028, debuting on the Cadillac Escalade IQ, building on Super Cruise's billion‑plus hands‑free miles. About the Role This role sits in the team's Platform pillar. We own the unified ML deployment platform that automates the path from a trained model to inference on the vehicle, along with the developer‑experience and agentic‑tooling layer that makes deployment self‑serve for every ML model development team at GM. What you’ll be doing (Responsibilities) Design, build, and operate the ML deployment platform that automates the path from trained model to on‑vehicle inference. Drive cross‑organization model deployments to the autonomous vehicle stack, partnering with model development teams to take high‑value models from training to production on‑vehicle. Build agentic tools that diagnose and fix deployment‑blocking issues, automating workflows currently performed manually by engineers. Build the developer experience that ML model development teams use day to day: tooling, dashboards, automation, and observability. Drive shift‑left validation that surfaces deployment risk (compile, runtime, parity, latency) early in the model development cycle. Build platform tools that integrate the work of our sister teams (kernels, compiler, reduced‑precision and parity) so their optimization wins land directly in the deployment workflow. Partner with the team's Performance pillar and model development teams across the AV organization. Your Skills & Abilities (Required Qualifications) BS, MS, or PhD in Computer Science or a related technical field. 3+ years of relevant industry experience. Strong fundamentals and excellent coding ability in Python. Experience building or operating production platform or infrastructure systems where reliability, observability, and extensibility matter. Experience with ML model deployment, inference integration, model optimization workflows, or model serving infrastructure, with at least one prior context where you owned the path from a trained model to a running inference workload. Experience using coding agents (Cursor, Claude Code, GitHub Copilot, or equivalent) as part of your engineering workflow. Experience designing clean, well‑tested software with clear interfaces and good abstractions. Strong cross‑team collaboration skills. What will give you a competitive edge (Preferred Qualifications) Experience building agentic or LLM‑powered developer tooling. Experience with ML or workflow orchestration frameworks (Airflow, Temporal, Flyte, Ray, Kubeflow, or equivalent). Familiarity with the NVIDIA GPU stack at the integration level (CUDA‑aware Python, TensorRT, Triton inference server, torch.compile, ONNX). Experience with inference‑serving frameworks (Triton, TorchServe, Ray Serve, vLLM) or edge‑deployment toolchains. Experience with low‑latency or real‑time systems. Experience in autonomous vehicles, robotics, or other safety‑critical ML deployment domains. Open‑source contributions to PyTorch, Ray, Airflow, Temporal, vLLM, TensorRT, or related projects. 3+ years of relevant industry experience. Compensation The compensation information is a good faith estimate only. It is based on what a successful applicant might be paid in accordance with applicable state laws. The compensation may not be representative for positions located outside of New York, Colorado, California, or Washington. Bonus Potential An incentive pay program offers payouts based on company performance, job level, and individual performance. Benefits GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays, tuition assistance programs, employee assistance program, GM vehicle discounts and more. Non-Discrimination and Equal Employment Opportunities (U.S.) General Motors is committed to being a workplace that is not only free of unlawful discrimination, but one that genuinely fosters inclusion and belonging. We strongly believe that providing an inclusive workplace creates an environment in which our employees can thrive and develop better products for our customers. All employment decisions are made on a non‑discriminatory basis without regard to sex, race, color, national origin, citizenship status, religion, age, disability, pregnancy or maternity status, sexual orientation, gender identity, status as a veteran or protected veteran, or any other similarly protected status in accordance with federal, state and local laws. Accommodations General Motors offers opportunities to all job seekers including individuals with disabilities. If you need a reasonable accommodation to assist with your job search or application for employment, email us or call us at View phone number on click.appcast.io. In your email, please include a description of the specific accommodation you are requesting as well as the job title and requisition number of the position for which you are applying. Apply For This Role Company General Motors Location Hybrid - Sunnyvale, California, United States of America #J-18808-Ljbffr General MotorsVacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior ML Inference Engineer - Platform in Sunnyvale, CA vacancy
- General Motors is seeking a senior ML deployment professional to own the platform that turns trained models into on-vehicle inferences. You will drive cross‑organization model deployments for the autonomous vehicle stack, partnering with model development teams to move...Senior
- NVIDIA is seeking a Senior Machine Learning Applications and Compiler Engineer to craft high-performance runtime and compiler components for large-scale inference workloads. You will map neural network graphs onto NVIDIA platforms, extending the CUDA ecosystem with efficient...Senior
- Cloud Hybrid Technologies, LLC is seeking a developer with hands-on experience in AI and ML systems. The role focuses on building AI/ML algorithms and technologies such as LLM inference, similarity search, and vector databases, with guardrails and LangChain integration....Senior
$165k - $242k
Dormont Manufacturing Co is seeking a Senior Engineer to lead designs and improve engineering standards. The role focuses on evolving our Kubernetes-native inference platform and ensuring reliability across multiple services. Qualified candidates should have 5-8 years...Senior- Unity Technologies is seeking a Senior Machine Learning Infrastructure Engineer to join the Vector Ads team. You will build and operate... ...powering Unity's global advertising platform, ensuring low-latency, high-throughput ML model serving at scale. You will collaborate...SeniorRemote work
- A leading AI company in California seeks a skilled engineer to develop foundational infrastructure for large-scale multimodal AI models. You will architect model serving pipelines, manage scheduling systems for GPU resources, and own CI/CD pipelines for checkpoints. Ideal...Senior
- ...located in Mountain View, California, is at the forefront of autonomous driving technology and is looking for a skilled ML Infrastructure Engineer. In this role, you will enhance the infrastructure used in machine learning projects, collaborating with various teams to...Senior
- ...is looking for a talented individual to join our Conversation Engine team. You'll be at the forefront of scaling and optimizing our... ...in software engineering will directly impact how efficient our platform functions. The ideal candidate will have a strong background in...Senior
$193.93k - $291.15k
Nuro in Mountain View is seeking an experienced engineer to work on its Mapping team, focusing on building and improving HD map generation and release pipelines. You will develop scalable workflows and manage APIs while engaging with other mapping teams. Required qualifications...Senior$202k - $224k
A leading technology company in Sunnyvale is seeking an ML engineer for its AV Labs to advance autonomous vehicle systems. The ideal candidate should have a PhD or MS in a related field and at least 3 years of experience in autonomous vehicles or computer vision. Responsibilities...Senior- Gravity Engineering Services Pvt Ltd. is looking for a skilled AI/ML Developer in Mountain View, California. The ideal candidate will design, develop, and maintain AI/ML systems for core product features and collaborate with cross-functional teams to implement these solutions...Senior
$165k - $242k
A cloud service provider is seeking a Senior Software Engineer II for their Inference team in Sunnyvale, California. In this role, you'll lead design reviews, implement optimizations, and improve service reliability. The ideal candidate has extensive experience with distributed...Senior$139k - $204k
CoreWeave is seeking a Senior Engineer to lead designs and enhance engineering standards within their Kubernetes-native inference platform. Responsibilities include driving architecture, defining SLIs/SLOs, and mentoring engineers, with 3-8 years of experience preferred...SeniorRemote jobFlexible hours- Gravity Engineering Services Pvt Ltd. is seeking a Senior Engineer to develop a next-generation inference platform for embedding models in MongoDB Atlas. You'll collaborate with AI researchers to enhance model inference, focusing on performance and reliability. The ideal...Senior
- United States Digital Space LLC in Palo Alto is looking for a Senior Engineer to develop a next-generation inference platform integrated with Atlas. This role involves building scalable infrastructure and collaborating with teams to enhance AI capabilities. Ideal candidates...Senior
- About the Role We’re looking for a Senior Engineer to help build the next-generation inference platform that supports embedding models used for semantic search, retrieval... ...AI Platform organization and collaborate with ML researchers and engineers from our Voyage.ai acquisition...Senior
$126k - $248k
About the Role We’re looking for a Senior Engineer to help build the next-generation inference platform that supports embedding models used for semantic search, retrieval... ...AI Platform organization and collaborate with ML researchers and engineers from our Voyage.ai acquisition...SeniorLocal areaFlexible hours- Crusoe Energy Systems seeks a Senior Software Engineer to join their Cloud Managed AI team in California. This on-site role involves leading the design and implementation of AI inference platform infrastructure, ensuring high performance and availability. The ideal candidate...Senior
- ...robotics a reality. We're looking for an Inference Optimization MLE to help build and... ...versions Collaborate closely with research engineers to translate model innovations into optimized... ...of experience in inference optimization, ML systems, or a closely related field Deep...
$230k - $260k
Typeface is hiring a Principal Machine Learning Engineer to define technical strategy and drive architectural decisions. This hybrid role based in Palo Alto involves leading design of large-scale ML systems while collaborating closely in office 3 days a week. The ideal...SeniorWork at officeFlexible hours3 days per week$175k - $215k
Waymo is looking for a skilled engineer to implement and scale the economic engine for its ride-hailing services. In this role, you will develop high-level infrastructure code for ML models that facilitate real-time decision-making for pricing and vehicle matching. The...SeniorFull time- A leading AI-powered fraud detection platform in Mountain View is seeking experienced platform engineers to design and build advanced machine learning systems. You will engage in improving core detection algorithms, using unsupervised and supervised machine learning, and...Senior
- ...background, hands-on software engineering experience, and a knack for... ...you'll do: Own development of ML models end-to-end from data strategy... ..., optimization, production platform validation, and fine-tuning... ..., high-throughput cloud inference pipeline for evaluation and KPI...SeniorWork experience placementWork at office
- ...oversee development of high scale, reliable data platform to manage, visualize and serve large-scale datasets for ML model training and validation. Build up the... ...Bachelor's degree or higher in Computer Science, Engineering, Robotics, or a similar technical field. Minimum...SeniorWork experience placement
$213k - $263k
Waymo is seeking experienced engineers with ML software and systems expertise to develop the next generation of its onboard ML inference engine. The role involves architecting high-performance ML systems for autonomous vehicles, requiring over 5 years of software engineering...SeniorFull time- A leading security technology company based in Santa Clara is seeking a Senior IT AI/ML Engineer to design and implement AI solutions across various business functions. You will collaborate with cross-functional teams and leverage deep technical expertise to deliver impactful...Senior
- Apple Inc. in Cupertino is searching for a passionate Software Engineer to enhance Siri's capabilities on innovative devices. The role focuses on developing natural interaction platforms, requiring deep experience in product development and a strong technical background...Senior
- Quanata, LLC is seeking a Senior Data Engineer specializing in MLOps to lead model development and automation across the ML lifecycle. You will partner with data engineers and data scientists to build a scalable platform that accelerates time-to-market for data science...SeniorRemote job
$193.93k - $291.15k
A leading robotics company in Mountain View is seeking a Senior Perception ML Data Infrastructure Engineer to optimize the core data platform. The ideal candidate has over 4 years of software engineering experience and must be fluent in C++ with a strong grasp of Python...Senior$232.2k - $283.8k
Senior Machine Learning Engineer: ML Platforms Mountain View, US ABOUT EARNIN As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck...SeniorFull timeWork at officeLocal area2 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior ML Inference Engineer - Platform. Be the first to apply!
Related searches
- machine learning engineer Sunnyvale, CA
- ai ml engineer Sunnyvale, CA
- machine learning software engineer Sunnyvale, CA
- senior ml engineer Sunnyvale, CA
- machine learning ai engineer Sunnyvale, CA
- computer vision machine learning engineer Sunnyvale, CA
- data platform engineer Sunnyvale, CA
- client platform engineer Sunnyvale, CA
- platform engineer Sunnyvale, CA
- platform developer Sunnyvale, CA
