Machine Learning Engineer, AI Inference Solutions (University Grad)
$119.25k - $150.85kGeneral Motors Proving Ground
Job DescriptionGeneral Motors is a global leader in advanced driver assistance, with Super Cruise hands-free technology in more than 500,000 equipped vehicles on the road and over 700 million hands-free miles driven—demonstrating that automation can be trusted, intuitive, and helpful while reaching everyday drivers at unprecedented scale. Within GM AV, the Model Deployment & Inference Solutions team deploys machine learning models from training frameworks (e.g., PyTorch) onto autonomous-vehicle hardware; our two-fold mission is to build the ML deployment platform that makes model rollouts fast and predictable, and to optimize models so they meet the real-time latency and memory budgets required to run on-vehicle. Our work sits on the critical path for GM’s publicly committed launch of eyes-off (hands-free, eyes-free) autonomous driving in 2028 on the Cadillac Escalade IQ, and we’re hiring engineers to help deliver the next generation of safe, delightful personal autonomous-vehicle experiences. About the Role As an early career Engineer on the Model Deployment & Inference Solutions team, you’ll contribute across both sides of our mission: building the ML deployment platform and optimizing models for on-vehicle inference. You’ll work with and learn from senior engineers on real production deployments, platform features, and model-optimization workflows that ship to GM’s Super Cruise fleet at large scale, with structured mentorship and a clear onboarding plan. You’ll also collaborate closely with our sister teams (kernels, compiler, reduced precision, and parity) on the end-to-end path that takes trained models from research frameworks to ultra-efficient, safety-critical inference on the car. This is an early-career / new graduate role designed for candidates who have recently or will be completing their degree by June 2026. What You’ll Do (Responsibilities) Contribute production code across the ML deployment platform, model-optimization workflows, and inference benchmarking/profiling infrastructure. Pair with senior engineers on deployment workflows, performance investigations, model-optimization experiments (e.g., quantization, pruning, distillation), and platform tooling. Build, test, and maintain platform tools (e.g., validators, performance probes, parity and sensitivity analyzers, agentic specialists) with technical guidance and code review support. Investigate and help root-cause production deployment or performance issues; learn and apply the diagnostic playbook for compiler, kernel, runtime, and parity bugs. Collaborate with cross-functional teams across the AV organization; including kernels, compiler, reduced-precision, parity, and model-development groups—to plan and execute model deployments to the AV stack, working under the guidance of senior engineers Participate in code reviews, design discussions, and technical documentation to ensure reliability, correctness, and clear abstractions in a large-scale codebase. Learn and follow secure coding, safety, and compliance practices required for on-vehicle autonomous driving software. Your Skills & Abilities (Required Qualifications) Recently completed or completing a Bachelor’s or Master’s degree by Spring 2026 in Computer Science, ECE, or a related technical field. (Degree must be completed before your start date.) Strong computer science fundamentals (e.g., data structures, algorithms, operating systems, computer architecture) and solid coding skills in Python and/or C++, demonstrated through coursework, internships, or substantial projects. Hands-on experience in AI/ML (e.g., machine learning, deep learning, computer vision, NLP, or ML systems) via classes, research, internships, or personal projects. Depth in at least one of: computer architecture, operating systems, distributed systems, or compilers. Demonstrated software-engineering experience (internships, coursework, open-source, research code, or competitions) showing good judgment around reliability, correctness, and clean abstractions. Experience with—or strong interest in—using coding assistants/agents (e.g., Cursor, Claude Code, GitHub Copilot) as part of your workflow. Ability to work effectively in collaborative, cross-functional teams and communicate clearly—both in writing and verbally—including explaining technical work partners What Will Give You a Competitive Edge (Preferred Qualifications) Internship, research, or advanced coursework in ML systems, ML compilers, GPU programming (CUDA, OpenAI Triton), inference optimization, or distributed training/serving infrastructure. Familiarity with PyTorch and modern ML compiler/runtime stacks (e.g., torch.compile, TensorRT, ONNX, Triton Inference Server, vLLM, or equivalent). Exposure to model optimization (quantization, pruning, distillation) or GPU profiling tools (Nsight Systems, Nsight Compute, PyTorch Profiler). Familiarity with workflow/ML platforms such as Airflow, Temporal, Flyte, Ray, or Kubeflow. Experience building agentic or LLM-powered tools or workflows. Open-source contributions related to PyTorch, TensorRT, vLLM, OpenAI Triton, or similar projects. Coursework, projects, or publications touching ML systems (e.g., MLSys, OSDI, ASPLOS, HPCA, NeurIPS systems track). Familiarity with a systems language (e.g., C++) and development in a Linux environment. Location Sunnyvale, CA This role is categorized as hybrid. This means the selected candidate is expected to report to a specific location at least 3 times a week. This job may be eligible for relocation benefits Compensation The compensation information is a good faith estimate only. It is based on what a successful applicant might be paid in accordance with applicable state laws. The compensation may not be representative for positions located outside of New York, Colorado, California, or Washington. The salary range for this roleis $119,250 to $150,850. The actual base salary a successful candidate will be offered within this range will vary based on factors relevant to the position. Bonus Potential: An incentive pay program offers payouts based on company performance, job level, and individual performance. Benefits: GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays, tuition assistance programs, employee assistance program, GM vehicle discounts and more. About GMOur vision is a world with Zero Crashes, Zero Emissions and Zero Congestion and we embrace the responsibility to lead the change that will make our world better, safer and more equitable for all.Why Join UsWe believe we all must make a choice every day – individually and collectively – to drive meaningful change through our words, our deeds and our culture. Every day, we want every employee to feel they belong to one General Motors team.Benefits OverviewFrom day one, we're looking out for your well-being–at work and at home–so you can focus on realizing your ambitions. Learn how GM supports a rewarding career that rewards you personally by visiting Total Rewards resources.Non-Discrimination and Equal Employment Opportunities (U.S.)General Motors is committed to being a workplace that is not only free of unlawful discrimination, but one that genuinely fosters inclusion and belonging. We strongly believe that providing an inclusive workplace creates an environment in which our employees can thrive and develop better products for our customers.All employment decisions are made on a non-discriminatory basis without regard to sex, race, color, national origin, citizenship status, religion, age, disability, pregnancy or maternity status, sexual orientation, gender identity, status as a veteran or protected veteran, or any other similarly protected status in accordance with federal, state and local laws. We encourage interested candidates to review the key responsibilities and qualifications for each role and apply for any positions that match their skills and capabilities. Applicants in the recruitment process may be required, where applicable, to successfully complete a role-related assessment(s) and/or a pre-employment screening prior to beginning employment. To learn more, visit How we Hire.AccommodationsGeneral Motors offers opportunities to all job seekers including individuals with disabilities. If you need a reasonable accommodation to assist with your job search or application for employment, emailus or call us at View phone number on click.appcast.io. In your email, please include a description of the specific accommodation you are requesting as well as the job title and requisition number of the position for which you are applying.SummaryLocation: Sunnyvale, California, United States of AmericaType: Full time
- ...future of physical AI. Founded in 2017... ...intelligence to every moving machine on the planet.... ...the company’s solutions to deliver... ...Seoul; and Tokyo. Learn more at applied.co... ...for a performance engineer who specializes in... ...-throughput batch inference sweeping petabytes...SuggestedFull timeFor contractorsFor subcontractorCasual workWork at officeRemote workDay shift
$209k - $313k
...live in the moment, learn about the world,... ...digital services.Snap Engineering teams build fun... ...’re looking for a Machine Learning Engineer... ...machine learning solutions (e.g., uplift modeling... ...of causal inference and modern approaches... ...+ 4+ year of post-grad machine learning experience...SuggestedFull timeLive inWork at officeLocal area$160k - $200k
Santa Clara, CAData Engineering - Machine Learning and Data Engineer /Full... ...is a Physical AI company pioneering AI... ...for both training and inference phases. You will build... ...contribute to cutting-edge solutions, this position is an... ...skillsPhd new grad or Masters with 3+ years...SuggestedFull time$116k - $189.75k
...focused on visual and AI computing. For two decades... ..., running deep learning algorithms and acting... ...; architect and build solutions. Test and release models... ...with existing machine learning, design automation... ...Electrical or Computer Engineering, Computer Science, or...Suggested$112.7k - $169.1k
...The opportunity Unity's Vector AI team builds the machine learning systems that decide which ads reach... ...monthly users on the world's leading game engine. Recommendation and ranking systems... ...rigorous experiments using causal inference, A/B testing, and offline evaluation...SuggestedInternshipWork at officeWorldwideRelocation packageShift work$246.5k
...and Roku. The systems and solutions span multiple... ...with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization... ...our Machine Learning and Inference Platform that powers the... ...someone excited to mentor engineers, innovate at scale, and...Work at officeLocal areaRemote workMonday to ThursdayFlexible hours$150k
...next generation of AI builders, and... ...data scientists, and engineers, tackling the most... ...groundbreaking AI solutions that have the potential... ...computing in deep learning, driving impactful... ...for the machine learning software... ...especially at training and inference, and support the...Full timeWork experience placementVisa sponsorship- ...developer of Embodied AI technology. Our... ...groundbreaking solutions. We aim high and stay... ..., constantly learning and evolving as we... ...As an ML Engineer within the Application... ...for success as a Machine Learning Engineer... ...or driver intent inference. Experience integrating...Full timeWork at officeWork from home
$117.7k - $221.4k
...at the intersection of machine learning, data infrastructure,... ...efficient for embodied AI systems. We believe... ...model reflects how Cola engineers think: build durable intermediate... ..., featurization, and inference foundations that power... ...practical, reliable solutions.What You’ll DoDesign,...Full timeLocal areaRemote workWork from homeRelocation packageFlexible hours$128k - $260.5k
...for the cloud and AI era. We secure and... ...platform, its Zero Trust Engine, and the powerful... ...at Netskope to learn more. Follow us on... ...(AI) and machine learning (ML) to protect... ...enterprise-scale AI solutions. Working closely... ...bleeding edge of LLM inference optimization,...$174.72k - $295.68k
...integrating advanced AI and autonomous driving... ...cutting-edge R&D in AI, machine learning, and smart... ...time Machine Learning Engineer - AI Foundation, with... ...accelerating model training/inference. Our mission is to solve... ...the next-generation E2E solution of autonomous driving....Full time$250k - $350k
...seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In... ..., and deep learning compiler stacks.GPU &... ...Mindset: Self-driven, solutions-oriented, and capable... ...seamless, intuitive, and universally accessible. Our mission...Work at office3 days per week$184k - $287.5k
Intelligent machines powered by Artificial Intelligence... ...that can learn, reason, and... ...vehicle powered by AI can meander through... ...Machine Learning Engineers with a background... ...ideas into reliable solutions for autonomous driving... ...for real-time inference on embedded or automotive...Full timeWorldwideNight shift$300k
...next generation of AI builders, and... ...data scientists, and engineers, tackling the most... ...groundbreaking AI solutions that have the potential... ...computing in deep learning, driving impactful... .../or distributed inference optimization team... ...with large-scale machine learning workloads...Full timeFlexible hours- ...worldwide.We’re a team of engineers, clinicians, and... ...Opportunity:As a Staff Machine Learning Engineer, you will be... ...novel machine learning solutions for pathology image analysis... ..., and implement AI/ML approaches to extract... ...for real-time inference, GPU/throughput optimization...Work at officeLocal areaWorldwideFlexible hours
$232k - $310k
...is a global leader in AI-native enterprise talent... ...high standards. Our engineers, product leaders, and... ...Eightfold's cutting-edge solutions. We're a group of... ...boundaries of applied machine learning. We work with massive... ...strategies (QLORA, DPO) and inference optimization (vLLM,...Work experience placementWork at officeRemote workFlexible hours3 days per week$272k - $431.25k
...NVIDIA is looking for a Machine Learning (ML) Engineer to join the GPU accelerated Apache... ...ML/DL model training and inference pipelines, spanning many... ...You will apply the latest ML/AI methods to empower... ...implement machine learning solutions for performance prediction...Full time- Overview As a Principal Machine Learning Engineer, you will drive the development... ...and analytics teams to build AI functionalities into Atlassian... ...offline training and online inference at scaleOversee end-to-end deployment of ML solutions into production, ensuring continuous...Local areaRemote work
- ...resilience. Powered by the Illumio AI Security Graph, our breach... ...Our Team's Vision: Our Engineering team is shaping the future of... ...simultaneously shipping the solution on-premises. Our guiding philosophy... ...managing proprietary model inference endpoints. This position...Immediate start
$160k - $225k
...Machine Learning Engineer Location: Mountain View, CA Company Stage of Funding: Seed Stage AI Startup ($25M Raised) Office Type: Onsite (Hybrid... ...support model training and inference. Build customer-facing... ...needs into scalable AI solutions. Help shape the technical...H1bWork at officeVisa sponsorship$140k - $160k
...the only premium engineering and scientific consulting... ...clients with solutions that create a... ...programs at over 500 universities and extensive... ...and a culture of learning. Thanks for your... ...currently seeking a Machine Learning Engineer... ..., and foundation AI modelsProficiency...Full timeWork at officeFlexible hours$111k - $231.25k
Machine Learning Engineer II (Finance) Yahoo Mail is the ultimate consumer inbox... ...systems and state-of-the-art AI - including NLP, Generative... ...mechanisms in AI/ML solutions. Responsibilities Drive Personalization... ...vLLM, TensorRT-LLM, Triton Inference Server, or ONNX Runtime....Work at officeFlexible hours- ...financial wellness solutions to nearly 500... ...impact talent who see AI as a teammate - leveraging... ...and continuous learning, and we seek out... .... We build machine learning solutions... ...Machine Learning Engineer I to build models,... ...modeling, causal inference, contextual bandits...Flexible hours
$162.51k - $342.75k
Staff Machine Learning Engineer (Finance) Job Description: We are Omnissa ! Omnissa is the first AI-driven digital work platform, built to support... ...industry-leading solutions-including Unified Endpoint... ...evaluation, and real time/batch inference. Optimize ML models and...Work experience placementLocal areaFlexible hours$195k - $343k
..., live in the moment, learn about the world, and have... ...builds cutting-edge AI technologies that... ...device and server-side inference. Our team creates intuitive... ...We’re looking for a Machine Learning Engineer to join Snap Inc! What... ...+ 7+ years of post-grad ML or related experience...Full timeLive inWork at officeLocal areaWorldwide$235.03k - $352.29k
...profound opportunity for AI to drive positive... ...why we’re building a universal autonomy platform: self... ...looking for a Staff Machine Learning Engineer to be a technical... ...to develop holistic solutions to top autonomy challenges... ..., training, onboard inference, closed-loop and open...Immediate startFlexible hours- ...At Mistral AI, we believe in the... ...time, and enhance learning and creativity. Our... ...models, products and solutions. Our comprehensive... ...seeking a Applied AI Engineer to facilitate the... ...for tasks such as inference and fine-tuning.... ...algorithms underlying machine learning and LLMs...Full timeWork at officeVisa sponsorship
- ...development, and deployment of advanced AI agents and agentic systems.... ..., UX designers, and other engineers to define requirements and deliver impactful solutions. Diagnose and troubleshoot... ...execution. Knowledge and passion in machine learning algorithms, Gen AI, LLMs, and...Full timeWork experience placement
$175k - $275k
...About Abaka AI Abaka AI is built on... ...reliable, and scalable data solutions. Our offerings... ..., or full-cycle data engineering, Abaka AI provides the... ...’re hiring our first Machine Learning Engineer in the United... ...training and inference systems. ~ Familiarity...Full timeImmediate startFlexible hours- ...Job Description: We are looking for a Machine Learning Engineer to join our core research and... ...scale training, high-throughput video inference, and reliable production pipelines over... ...demonstrated impact in applied ML or AI systems. What We Offer ~ Competitive...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Machine Learning Engineer, AI Inference Solutions (University Grad). Be the first to apply!
- senior ml engineer Sunnyvale, CA
- machine learning engineer Sunnyvale, CA
- machine learning ai engineer Sunnyvale, CA
- machine learning software engineer Sunnyvale, CA
- ai ml engineer Sunnyvale, CA
- computer vision machine learning engineer Sunnyvale, CA
- machine learning researcher Sunnyvale, CA
- machine learning part time Sunnyvale, CA
- machine learning research scientist Sunnyvale, CA
- machine learning Sunnyvale, CA



