AI/ML Engineer - Model Inference
$117.7k - $221.4kGeneral Motors
Description
The Team
Cola is part of GM’s autonomous vehicle effort, focused on helping teams discover, understand, and curate high-value data from large-scale real-world sensor streams. The team sits at the intersection of machine learning, data infrastructure, and developer productivity, building systems that make it easier to search for important scenarios, prepare training-ready data, and support fast iteration across perception and evaluation workflows.
Our goal is to make world understanding scalable, practical, and cost efficient for embodied AI systems. We believe the next generation of autonomy and robotics depends not only on stronger models, but also on better infrastructure for turning massive volumes of multimodal data into reusable signals, searchable artifacts, and high-quality evaluation loops. That means building systems that can operate at industrial scale while preserving the flexibility to adapt quickly to new questions, new edge cases, and new model capabilities.
A core idea behind how we work is EMWU, or Efficient Multi-Tier World Understanding. At a high level, EMWU is a cost-aware approach that first performs the cheapest reusable work, such as detection, featurization, and retrieval, and then applies deeper reasoning only where it adds meaningful value. This operating model reflects how Cola engineers think: build durable intermediate artifacts, design for scale from the start, and balance quality, speed, and cost instead of optimizing any one of them in isolation.
The Role
We are looking for a hands-on machine learning engineer to help build the data processing, featurization, and inference foundations that power scalable world understanding. This role is ideal for someone who is equally comfortable working on machine learning systems, production infrastructure, and evaluation loops, and who enjoys turning ambiguous problems into practical, reliable solutions.
What You’ll Do
Design, build, and productionize data processing and featurization pipelines for large-scale multimodal data
Improve inference frameworks for computer vision and multimodal models, with a focus on reliability, extensibility, and operational simplicity
Drive scalability and cost efficiency across the end-to-end pipeline, including compute utilization, throughput, storage, and query performance
Work closely with partners across machine learning, infrastructure, and evaluation to deliver systems that support both rapid experimentation and production use
Develop and refine evaluation methods for model quality, retrieval quality, and system-level performance
Help shape technical direction through strong execution, thoughtful tradeoff analysis, and clear engineering judgment
Take ownership of ambiguous problem spaces, define practical paths forward, and move quickly from prototype to production
Operate with urgency and a strong bias toward execution velocity while maintaining a high bar for engineering quality
Your Skills & Abilities
BS, MS, or PhD in Computer Science, Electrical Engineering, Robotics, or a related technical field, or equivalent practical experience
Experience building production data processing or machine learning pipelines at scale
Experience with featurization, embedding, inference, or retrieval systems for vision or multimodal workloads
Strong understanding of computer vision models and the practical challenges of deploying them in production environments
Experience evaluating machine learning systems using clear metrics, experiments, and regression safeguards
Proven ability to work hands-on in fast-moving environments with incomplete information
Strong ownership mindset, sound technical judgment, and the ability to drive execution through ambiguity
What Will Give You A Competitive Edge
Experience with world models or large-scale world understanding systems
Experience with simulation workflows or synthetic data systems
Experience with vector search, approximate nearest neighbor retrieval, or large-scale embedding infrastructure
Experience working on embodied AI, autonomous systems, or safety-critical machine learning applications
Compensation
The salary range for this role is $117,700 and $221,400. The actual base salary a successful candidate will be offered within this range will vary based on factors relevant to the position (along with level.)
Bonus Potential: An incentive pay program offers payouts based on company performance, job level, and individual performance.
Benefits:
GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays, tuition assistance programs, employee assistance program, GM vehicle discounts and more.
This role is based remotely, but if the selected candidate lives within a specific mile radius of a GM hub, they will be expected to report to the location three times a week {or other frequency dictated by your manager}.
Relocation benefits are available for candidates who qualify under company policy.
About GM
Our vision is a world with Zero Crashes, Zero Emissions and Zero Congestion and we embrace the responsibility to lead the change that will make our world better, safer and more equitable for all.
Why Join Us
We believe we all must make a choice every day – individually and collectively – to drive meaningful change through our words, our deeds and our culture. Every day, we want every employee to feel they belong to one General Motors team.
Total Rewards | Benefits Overview
From day one, we're looking out for your well-being–at work and at home–so you can focus on realizing your ambitions. Learn how GM supports a rewarding career that rewards you personally by visiting Total Rewards resources.
Non-Discrimination and Equal Employment Opportunities (U.S.)
General Motors is committed to being a workplace that is not only free of unlawful discrimination, but one that genuinely fosters inclusion and belonging. We strongly believe that providing an inclusive workplace creates an environment in which our employees can thrive and develop better products for our customers.
All employment decisions are made on a non-discriminatory basis without regard to sex, race, color, national origin, citizenship status, religion, age, disability, pregnancy or maternity status, sexual orientation, gender identity, status as a veteran or protected veteran, or any other similarly protected status in accordance with federal, state and local laws.
We encourage interested candidates to review the key responsibilities and qualifications for each role and apply for any positions that match their skills and capabilities. Applicants in the recruitment process may be required, where applicable, to successfully complete a role-related assessment(s) and/or a pre-employment screening prior to beginning employment. To learn more, visit How we Hire.
Accommodations
General Motors offers opportunities to all job seekers including individuals with disabilities. If you need a reasonable accommodation to assist with your job search or application for employment, email us Show email or call us at Show phone number. In your email, please include a description of the specific accommodation you are requesting as well as the job title and requisition number of the position for which you are applying.
$189.3k - $290.7k
.... The Role General Motors is bringing multimodal AI into the vehicle, and we are looking for a Staff AI/ML Software Engineer to lead the adaptation, fine-tuning, and distillation of foundation models for the automotive edge. You will build models that understand...SuggestedFull timeWork at officeLocal areaWork from homeRelocation package3 days per week- ...Systems builds the world's largest AI chip, 56 times larger than... ...-leading training and inference speeds; over 10 times faster... ...Cerebras works with the leading model labs, global enterprises, and... ...loop." You'll sit between engineering, product, and customer-facing...SuggestedFull time
$165.2k - $223.6k
...Inferentia and Trainium ML accelerators. This... ...enabling unparalleled ML inference and training performance... ...running a wide range of models and supporting novel architecture... ...software boundary, our engineers build systematic... ...of what's possible in AI acceleration.As part of...SuggestedWork experience placementInternshipLocal areaFlexible hours$170.6k - $261.3k
...Autonomous Vehicle organization makes aggressive model optimization safe enough to ship... ...looking for a mathematically rigorous engineer to own model numerics: how we measure, bound... ...quantization, compiler toolchains, or inference-time numerical parity. Experience...SuggestedFull timeLocal areaWork from homeRelocation packageFlexible hours$193.3k - $261.5k
...custom cloud-scale machine learning accelerators. Join us to optimize the latest models to run really fast on the Trainium hardware.As a Sr. Software Development Engineer on the Inference Model Enablement team, you will onboard and optimize state-of-the-art open-source...SuggestedInternshipLocal areaFlexible hours$144.7k - $261.3k
...environments, cloud infrastructure, and ML/AI GPU platforms for AV research... ...for a Senior Performance Engineer to join the AV Capacity and... ...large-scale ML training and inference environments. Your s kills &... ...H100, B200, and GB200 . Model Deployment: Experience deploying...Full timeWork at officeLocal areaRemote workWork from homeFlexible hours3 days per week$150.4k - $277.6k
AI/ML Software Engineer The Video Computer Vision organization is working on exciting technologies for... ...understanding of machine learning inference and performance oriented development is... ...calling components Integrate multimodal models with audio, image and video modalities...Relocation$275.8k - $340.5k
.... About the team: The AV ML Infra team at GM builds ML... ...meet the unique demands of AI and ML innovation, supporting... ...the productivity of ML engineers, and drive the adoption of... ...includes: AI Validation & Inference: Ensures robust model performance by running large...Full timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours- ...Company Overview: At Viven, we're building AI-powered Digital Twins for businesses. A... ...Overview: We are looking for an AI/ML Engineer with hands-on experience in machine learning... ...language processing, and large language models. This AI/ML Engineer will play a critical...Full time
- ...computing experiences—from AI and data centers, to PCs, gaming... ...THE ROLE:We are hiring AI / ML Platform Engineers to build the platform layer... ..., distributed training and inference, experiment tracking,... ...profiling, kernel benchmarking, model serving, or distributed training...
- ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by one... ...next-generation robotic platforms. As a Senior AI/ML Research Engineer, you will develop and fine-tune the foundation models—VFMs, VLMs, and VLA models—that let our Embodied...Local areaWorldwideFlexible hours
$150k - $225k
...quality of life. As a Senior AI / Embedded Engineer, you will be responsible for... ...includes data ingestion, model development, optimization,... ...reliable, low-power, real-time ML systems that operate at the... ...to reduce model size and inference latency ◦ Use frameworks including...Full timeWork at officeImmediate startVisa sponsorshipNight shift- ...headquartered in Santa Clara, CA, seeks a Principal System Software Engineer for AI Inference Execution. You will join the software team to productize the... ..., developing deployment software and collaborating with ML, compiler, and hardware experts. Required: strong...
- ...is seeking a Machine Learning Engineer to support advanced technical... ...on developing and applying AI and machine learning methods... ...transparency, and clinical utility of ML/AI models intended for translational... ...training, evaluation, and inference that can support a range of...
- ...to take your software engineering career to the next level... ...enterprise-authorized AI coding assist tools within... ..., with emphasis on ML systems. Hands-on experience... ...Exposure to ML model serving frameworks (e.g... ...TensorFlow Serving, Triton Inference Server) Familiarity...
$140k - $224.25k
We are looking for a highly motivated AI/ML Software Engineer to join the Enterprise Agentic AI Platform team within IT. You will work closely... ...develop, and deploy Agentic AI applications using Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), and AI...Full time- ...and scientific discovery to powering AI and the technologies people rely on... ...We are seeking highly motivated AI Model Optimization & Software Engineer Interns/Co-op to join our teams. We... ...software for training, fine-tuning, and inference across CPU, GPU, and accelerator...Full timeSummer workInternshipSummer internshipWorldwide
$119.25k - $150.85k
...unprecedented scale. Within GM AV, the Model Deployment & Inference Solutions team deploys... ...mission is to build the ML deployment platform that... ...Escalade IQ, and we’re hiring engineers to help deliver the next generation... ...Hands-on experience in AI/ML (e.g., machine learning,...Full timeInternshipLocal areaWork from homeRelocation packageFlexible hours- AI/ML Engineer Location: Sunnyvale, CA, USA Duration: 6-12+ Months Experience: 9-12 Years Preferred Background Apple, PayPal, Netflix, Meta, AWS, Product Companies, Startups Avoid candidates from Banking, Finance, Telecom, Healthcare, Insurance, Wells Fargo, CVS, and similar...
- Job Title: Senior AI/ML Engineer Work Location with ZIP: Sunnyvale, CA 94085 (Hybrid) Technical Hiring Criteria (Must Haves) Top 3 Required... ...with Generative AI / LLMs (OpenAI, Azure OpenAI, open'source models) - Experience with Prompt Engineering, RAG, LangChain / Semantic...
$190k - $260k
...developed an artificial intelligence (AI) powered technology stack... ...it. Every improvement to our models – from GigaFusionNet to large-... .... We are looking for engineers who make model training fast:... ...streamable formats Partner with ML teams to scale new architectures...Temporary workWork at officeVisa sponsorshipFlexible hours$178.42k - $230.5k
...successful candidate is in the Seattle, Washington area. About Us The AI Cloud and Developer Infrastructure organization is responsible for delivering and maintaining the tools and services engineers here at GM use every day to do their best work and drive our cars...Full timeWork experience placementWork at officeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours3 days per week- Lead AI/ML Engineer We are seeking a Lead AI/ML Engineer to architect, build, and operationalize... ...from UAT logs and utilize frontier models to synthetically generate variations (typos... ...pipelines supporting parallel inference execution across 1,000+ trajectory datasets...Local area
$274k - $304k
...Saviynt's AI-powered identity platform manages and governs human and... ...information, please visit AI Platform Engineer - Training & Inference Saviynt's AI-powered... ..., evaluates, and serves every AI model at Saviynt. We need an ML Platform Engineer to own distributed...$296.3k - $423.9k
...Generation team within the Embodied AI organization. This role... ...will shape the algorithms and ML systems that translate sensor... ...lead a high-performing team of engineers building ML-driven trajectory... ...the roadmap that takes these models from research to production deployment...Full timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours$207k - $300k
...month Forward Deployed Engineer (FDE) embeds and 2-4-week... ...and Python code for model compression, speculative... ...sequential decision making), ML infrastructure, or... ...integrating generative AI tools or LLM interfaces... ...pipelines, optimize inference serving engines, and establish...$200k - $400k
...intelligent machines at scale. At Scout AI, we’re developing Fury, the first robotic foundation model for defense, to give U.S.... ...for a Senior or Staff AI Engineer to join the Fury Orchestration... ...multi-agent coordination, and edge inference optimization. Expect to rapidly...Full timeRelocation package$50k - $120k
...Mission Altimate AI, founded in 2022 in San... ...combines multiple language models and a custom-built... ...of the AI-powered data engineering revolution. You can read... ...~5+ years of hands-on ML/AI experience with at least... ...large-scale training, inference, and multi-agent...Full timeWorldwide$169k - $338k
...WalmartBusiness Segment: Home OfficeAI/ML & Agentic Systems Technical... ...and develop advanced agentic AI systems that can autonomously handle complex reliability engineering workflows, predictive failure... ...that enable continuous learning, model deployment, monitoring, and...Full timeTemporary workPart time$215k - $260k
...vertically integrated AI infrastructure company... ...making large language models run faster, cheaper, and... ...That means owning the inference stack end to end: profiling... ...directly with customer engineering teams to tailor... ...methods across many kinds of ML models, with an emphasis...Temporary work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI/ML Engineer - Model Inference. Be the first to apply!
- machine learning ai engineer Sunnyvale, CA
- ai ml engineer Sunnyvale, CA
- ai prompt engineer Sunnyvale, CA
- ai engineer Sunnyvale, CA
- ai engineer remote Sunnyvale, CA
- ai developer Sunnyvale, CA
- senior ai engineer Sunnyvale, CA
- machine learning engineer Sunnyvale, CA
- senior ml engineer Sunnyvale, CA
- machine learning software engineer Sunnyvale, CA





