Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Machine Learning Engineer, AI Inference Solutions (Early in Career)

$119.25k - $150.85k
Full-time

General Motors

Description

General Motors is a global leader in advanced driver assistance, with Super Cruise hands-free technology in more than 500,000 equipped vehicles on the road and over 700 million hands-free miles driven—demonstrating that automation can be trusted, intuitive, and helpful while reaching everyday drivers at unprecedented scale. Within GM AV, the Model Deployment & Inference Solutions team deploys machine learning models from training frameworks (e.g., PyTorch) onto autonomous-vehicle hardware; our two-fold mission is to build the ML deployment platform that makes model rollouts fast and predictable, and to optimize models so they meet the real-time latency and memory budgets required to run on-vehicle. Our work sits on the critical path for GM’s publicly committed launch of eyes-off (hands-free, eyes-free) autonomous driving in 2028 on the Cadillac Escalade IQ, and we’re hiring engineers to help deliver the next generation of safe, delightful personal autonomous-vehicle experiences.

About the Role

As an early career Engineer on the Model Deployment & Inference Solutions team, you’ll contribute across both sides of our mission: building the ML deployment platform and optimizing models for on-vehicle inference. You’ll work with and learn from senior engineers on real production deployments, platform features, and model-optimization workflows that ship to GM’s Super Cruise fleet at large scale, with structured mentorship and a clear onboarding plan. You’ll also collaborate closely with our sister teams (kernels, compiler, reduced precision, and parity) on the end-to-end path that takes trained models from research frameworks to ultra-efficient, safety-critical inference on the car. This is an early-career / new graduate role designed for candidates who have recently or will be completing their degree by August 2026.

What You’ll Do (Responsibilities)

Contribute production code across the ML deployment platform, model-optimization workflows, and inference benchmarking/profiling infrastructure.

Pair with senior engineers on deployment workflows, performance investigations, model-optimization experiments (e.g., quantization, pruning, distillation), and platform tooling.

Build, test, and maintain platform tools (e.g., validators, performance probes, parity and sensitivity analyzers, agentic specialists) with technical guidance and code review support.

Investigate and help root-cause production deployment or performance issues; learn and apply the diagnostic playbook for compiler, kernel, runtime, and parity bugs.

Collaborate with cross-functional teams across the AV organization; including kernels, compiler, reduced-precision, parity, and model-development groups—to plan and execute model deployments to the AV stack, working under the guidance of senior engineers

Participate in code reviews, design discussions, and technical documentation to ensure reliability, correctness, and clear abstractions in a large-scale codebase.

Learn and follow secure coding, safety, and compliance practices required for on-vehicle autonomous driving software.

Your Skills & Abilities (Required Qualifications)

Recently completed or completing a Bachelor’s or Master’s degree by Spring 2026 in Computer Science, ECE, or a related technical field. (Degree must be completed before your start date.)

Strong computer science fundamentals (e.g., data structures, algorithms, operating systems, computer architecture) and solid coding skills in Python and/or C++, demonstrated through coursework, internships, or substantial projects.

Hands-on experience in AI/ML (e.g., machine learning, deep learning, computer vision, NLP, or ML systems) via classes, research, internships, or personal projects.

Depth in at least one of: computer architecture, operating systems, distributed systems, or compilers.

Demonstrated software-engineering experience (internships, coursework, open-source, research code, or competitions) showing good judgment around r eliability, correctness, and clean abstractions.

Experience with—or strong interest in—using coding assistants/agents (e.g., Cursor, Claude Code, GitHub Copilot) as part of your workflow.

Ability to work effectively in collaborative, cross-functional teams and communicate clearly—both in writing and verbally—including explaining technical work partners

What Will Give You a Competitive Edge (Preferred Qualifications)

Internship, research, or advanced coursework in ML systems, ML compilers, GPU programming (CUDA, OpenAI Triton), inference optimization, or distributed training/serving infrastructure.

Familiarity with PyTorch and modern ML compiler/runtime stacks (e.g., torch.compile, TensorRT, ONNX, Triton Inference Server, vLLM, or equivalent).

Exposure to model optimization (quantization, pruning, distillation) or GPU profiling tools (Nsight Systems, Nsight Compute, PyTorch Profiler).

Familiarity with workflow/ML platforms such as Airflow, Temporal, Flyte, Ray, or Kubeflow.

Experience building agentic or LLM-powered tools or workflows.

Open-source contributions related to PyTorch, TensorRT, vLLM, OpenAI Triton, or similar projects.

Coursework, projects, or publications touching ML systems (e.g., MLSys, OSDI, ASPLOS, HPCA, NeurIPS systems track).

Familiarity with a systems language (e.g., C++) and development in a Linux environment.

Location

Sunnyvale, CA

This role is categorized as hybrid. This means the selected candidate is expected to report to a specific location at least 3 times a week.

This job may be eligible for relocation benefits

Compensation

The compensation information is a good faith estimate only. It is based on what a successful applicant might be paid in accordance with applicable state laws. The compensation may not be representative for positions located outside of New York, Colorado, California, or Washington.

The salary range for this role is $119,250 to $150,850. The actual base salary a successful candidate will be offered within this range will vary based on factors relevant to the position.

Bonus Potential : An incentive pay program offers payouts based on company performance, job level, and individual performance.

Benefits: GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays, tuition assistance programs, employee assistance program, GM vehicle discounts and more.

About GM

Our vision is a world with Zero Crashes, Zero Emissions and Zero Congestion and we embrace the responsibility to lead the change that will make our world better, safer and more equitable for all.

Why Join Us 

We believe we all must make a choice every day – individually and collectively – to drive meaningful change through our words, our deeds and our culture. Every day, we want every employee to feel they belong to one General Motors team.

Total Rewards | Benefits Overview

From day one, we're looking out for your well-being–at work and at home–so you can focus on realizing your ambitions. Learn how GM supports a rewarding career that rewards you personally by visiting Total Rewards resources. 

Non-Discrimination and Equal Employment Opportunities (U.S.)

General Motors is committed to being a workplace that is not only free of unlawful discrimination, but one that genuinely fosters inclusion and belonging. We strongly believe that providing an inclusive workplace creates an environment in which our employees can thrive and develop better products for our customers.

All employment decisions are made on a non-discriminatory basis without regard to sex, race, color, national origin, citizenship status, religion, age, disability, pregnancy or maternity status, sexual orientation, gender identity, status as a veteran or protected veteran, or any other similarly protected status in accordance with federal, state and local laws. 

We encourage interested candidates to review the key responsibilities and qualifications for each role and apply for any positions that match their skills and capabilities. Applicants in the recruitment process may be required, where applicable, to successfully complete a role-related assessment(s) and/or a pre-employment screening prior to beginning employment. To learn more, visit How we Hire.

Accommodations

General Motors offers opportunities to all job seekers including individuals with disabilities. If you need a reasonable accommodation to assist with your job search or application for employment, email us View email address on us.fitly.work or call us at View phone number on us.fitly.work. In your email, please include a description of the specific accommodation you are requesting as well as the job title and requisition number of the position for which you are applying.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Machine Learning Engineer, AI Inference Solutions (Early in Career) in Sunnyvale, CA vacancy
  • $278.1k - $347.6k

     ...are building the next generation of AI-driven game experiences, running generative...  ...that runtime. As our Principal Engineer for On-Device AI Inference & Systems, you will be the foremost engineering...  ..., op-support gaps, and cost models early so model design and deployment... 
    Suggested
    Work at office
    Worldwide
    Relocation package

    Unity

    Mountain View, CA
    5 days ago
  •  ...About the job Machine Learning Engineer The Role We are...  ...production-ready ML solutions. This role is well...  ...centric, and embodied AI problems.....  ...collection, training, inference, and deployment. Background...  ...A clear path for career growth in technical leadership... 
    Suggested

    Glint Tech Solutions LLC

    Santa Clara, CA
    5 days ago
  • $174.72k - $295.68k

     ...Senior Machine Learning Engineer - AI Foundation Santa Clara, CA XPENG is a leading smart technology...  ...and accelerating model training/inference. Our mission is to solve the autonomous...  ...will enable the next-generation E2E solution of autonomous driving. Job... 
    Suggested
    Full time

    XPENG

    Santa Clara, CA
    1 day ago
  •  ...500 company and a leading AI platform for managing people...  ...Whether you're building smarter solutions, supporting customers, or...  ...About the Role As a Machine Learning Engineer on the AI Platform team, you...  ...developing ETL pipelines and inference services that use large... 
    Suggested
    Full time

    Workday

    Santa Clara, CA
    1 day ago
  •  ...how people, data, and AI agents connect across...  ...suite of cloud-based solutions, Proofpoint helps...  ...Proofpoint. Job Title: Machine Learning Engineer Location :...  ...latency, efficiency, inference cost, and operational...  ...that an exceptional career experience includes a... 
    Suggested
    Full time
    Remote work
    Flexible hours

    Proofpoint

    Sunnyvale, CA
    2 days ago
  • $209k - $313k

     ..., live in the moment, learn about the world, and have...  ...services. Snap Engineering ( teams build fun and...  ...We're looking for a Machine Learning Engineer to join...  ...causal machine learning solutions (e.g., uplift modeling...  ...of causal inference and modern approaches... 
    Full time
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    4 days ago
  • $197.5k - $272k

     ...driving the transformation to AI-enabled software-defined...  ...intelligent AI-defined vehicles. Our solutions are already on the road...  ...looking for a great Staff Machine Learning Engineer to join our seasoned AI...  ...modern C++ (C++14/17 for inference). ~ Deep proficiency with... 
    Work at office
    Worldwide
    Flexible hours
    Shift work
    3 days per week

    Sonatus

    Sunnyvale, CA
    5 days ago
  • $162.51k - $342.75k

     ...Staff Machine Learning Engineer We are Omnissa! Omnissa is the first AI-driven digital work platform, built to support flexible...  ...We integrate industry-leading solutions—including Unified Endpoint...  ...evaluation, and real time/batch inference. Optimize ML models and pipelines... 
    Work experience placement
    Local area
    Flexible hours

    Omnissa

    Mountain View, CA
    1 day ago
  •  ...next generation of AI builders, and...  ...data scientists, and engineers, tackling the most...  ...groundbreaking AI solutions that have the potential...  ...computing in deep learning, driving impactful...  ...for the machine learning software...  ...especially at training and inference, and support the... 
    Work experience placement
    Visa sponsorship

    Institute of Foundation Models

    Sunnyvale, CA
    16 days ago
  •  ...next generation of AI builders, and...  ...data scientists, and engineers, tackling the most...  ...groundbreaking AI solutions that have the potential...  ...computing in deep learning, driving impactful...  .../or distributed inference optimization team...  ...with large-scale machine learning workloads... 
    Flexible hours

    Institute of Foundation Models

    Sunnyvale, CA
    16 days ago
  • $160k - $200k

     ...PlusAI is a Physical AI company pioneering...  ...ML Infrastructure Engineer at Plus, you will...  ...both training and inference phases. You will...  ...state-of-the-art deep learning frameworks like...  ...'s possible in machine learning infrastructure...  ...to cutting-edge solutions, this position is... 

    PlusAI

    Santa Clara, CA
    16 days ago
  • $166k - $244k

     ...working to build the AI-powered electric grid...  ...energy, AI, software engineering, and products to build...  ...at global scale. Learn more about our team and...  ...We're looking for an early career Machine Learning Engineer to...  ...with purpose: We build solutions that solve real problems... 
    Entry level
    Full time
    Flexible hours

    X Company

    Mountain View, CA
    6 days ago
  •  ...development, and deployment of advanced AI agents and agentic systems....  ..., UX designers, and other engineers to define requirements and deliver impactful solutions. Diagnose and troubleshoot...  ...execution. Knowledge and passion in machine learning algorithms, GenAI, LLMs, and... 
    Full time
    Work experience placement

    Eightfold

    Santa Clara, CA
    14 hours ago
  •  ...Job Description: We are looking for a Machine Learning Engineer to join our core research and...  ...scale training, high-throughput video inference, and reliable production pipelines over...  ...demonstrated impact in applied ML or AI systems. What We Offer ~ Competitive... 
    Full time

    Maxinsights Corporation

    Santa Clara, CA
    14 hours ago
  • $150k

     ...the next generation of AI builders, and drive transformative...  ..., data scientists, and engineers, tackling the most...  ...of groundbreaking AI solutions that have the potential...  ...computing in deep learning, driving impactful discoveries...  .... The Role As a Machine Learning Engineer at... 
    Full time
    Worldwide
    Visa sponsorship

    Institute Of Foundation Models

    Sunnyvale, CA
    14 hours ago
  • $125k - $165k

     ...Data Labeling Engineering team designs, builds...  ...operates hybrid human/machine data labeling...  ...vehicle machine learning models within General...  ..., and  AI/ML , defining the...  ...led training data solutions at  foundation...  ...Role  As an early-career Software Engineer... 
    Entry level
    Full time
    Internship
    Work at office
    Local area
    Work from home
    Relocation package

    General Motors

    Sunnyvale, CA
    4 days ago
  •  ...Machine Learning Software Engineer The Omnichannel Core Platform team is looking for a Machine Learning Software...  ...tasks and idea to a implementable solution. Work with minimum hand holding,...  ...in building NLP, Computer vision AI models ~ Hands on expertise and contributions... 
    Work experience placement

    Omega Solutions

    Santa Clara, CA
    2 days ago
  •  ...About the job ML Engineer Our Client...  ...teams work. Their AI technology captures...  ...how businesses learn from and optimize...  ...vertical in Applied AI, Machine Learning, and Data...  ...AI-powered solutions enabling natural speech...  ..., deployment, inference, and monitoring in... 
    Full time

    Catalyst Labs, LLC

    Mountain View, CA
    2 days ago
  • $174.3k - $200k

     ...Machine Learning Engineer UnitX builds the world's leading physical AI systems to automate repetitive visual tasks in factories....  ...level precision and real-time inference. Build Automated Pipelines...  ...continuously optimize deployed solutions. Deploy & Maintain... 

    UnitX

    Milpitas, CA
    4 days ago
  •  ...are seeking a highly skilled Machine Learning Engineer to design and build a low-...  ...a focus on sub-second inference, CPU-based execution, and scalable...  ...without reliance on hosted AI services. Design and...  ...workflows into scalable ML solutions. Deliver a working... 
    Local area

    Sparktek

    San Jose, CA
    5 days ago
  • $196k - $221k

     ...build and deploy cutting-edge AI technology to help people...  ...-veteran scientists and engineers. As a Machine Learning Engineer, you'll bring your...  ...tuning, post-training, and inference strategies for large language...  .... The company is backed by early investors in Google, DeepMind... 
    Permanent employment

    Otter.ai

    Mountain View, CA
    5 days ago
  •  ...Moveworks is the Agentic AI Assistant platform that...  ..., and continuously learn and adapt. Moveworks...  ...with Moveworks’ Reasoning Engine and natural language...  ...demands of the enterprise solution space. Take a look at...  ...collaborate closely with machine learning experts and... 
    Work at office
    Remote work
    Flexible hours

    ServiceNow

    Mountain View, CA
    4 days ago
  • $123.75k - $185k

     ...Eightfold is a global leader in AI-native enterprise talent...  ...collaboration, and high standards. Our engineers, product leaders, and go-to-...  ...and deliver impactful solutions. Diagnose and...  ...Knowledge and passion in machine learning algorithms, Gen AI, LLMs, and... 
    Work experience placement
    Work at office
    3 days per week

    Eightfold LLC

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

     ...36 NVIDIA is looking for a talented Machine Learning Engineer to drive the development, evaluation,...  ...end-to-end lifecycle management of our AI-powered systems. This role bridges advanced...  ...for high-throughput, low-latency AI inference workflows. GitLab CI/CD & Security... 
    Full time
    Remote work
    Flexible hours

    NVIDIA

    Santa Clara, CA
    4 days ago
  • $210.3k - $273.4k

     ...systems. Our next-generation agentic AI - goal-directed, adaptive, and capable...  ...to life. We're looking for a Senior Machine Learning Engineer to lead the development of these foundational...  ...behaviors, AI planning and goal inference frameworks for NPCs and simulations,... 
    Full time
    Work at office
    Worldwide

    Unity Technologies

    Mountain View, CA
    2 days ago
  • $190k - $250k

     ...Atoms is building the machines that power the next era...  ...Atoms builds Physical AI- real-world robots for...  ...environments, operate them, learn from them, and improve...  ...We are roboticists, engineers, operators, and builders...  ...applied machine learning solutions in a production... 
    Full time
    Temporary work
    Work at office
    Flexible hours

    ATOMS Careers page

    Mountain View, CA
    5 days ago
  • $229.5k - $360k

     ...leveraging state-of-the-art machine learning. Our mission is to...  ...blends innovation, engineering excellence, and a...  ...How will I use AI at Roku? At Roku...  ...feature store, real-time inference services, vector DBs,...  ...experience building software solutions to concrete problems... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    2 days ago
  • $100k

     ...focus on getting you a tech Job we make careers. In this Job market also, our...  ..., Data analysts/ Data Scientists, and Machine Learning engineers for full-time positions with clients....  ...cycle Knowledge of Statistics, Gen AI, LLM, Python, Computer Vision, data visualization... 
    Entry level
    Full time
    H1b

    SynergisticIT

    Cupertino, CA
    1 day ago
  • $246.5k

     ...and Roku. The systems and solutions span multiple...  ...with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization...  ...our Machine Learning and Inference Platform that powers the...  ...someone excited to mentor engineers, innovate at scale, and... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    4 days ago
  • $182k - $242k

     ...Senior Applied ML Engineer Bellevue, WA...  ...Cloud for AI™. Built for pioneers...  ...to help agents learn from experience...  ...a proven solution. However, there...  ...fraction of the total inference market, which...  ...Science, Machine Learning, Robotics...  ...please contact: careers@coreweave.com.... 
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Machine Learning Engineer, AI Inference Solutions (Early in Career). Be the first to apply!