Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI/ML Engineer - Model Inference

$117.7k - $221.4k

General Motors Proving Ground

Job DescriptionThe TeamCola is part of GM’s autonomous vehicle effort, focused on helping teams discover, understand, and curate high-value data from large-scale real-world sensor streams. The team sits at the intersection of machine learning, data infrastructure, and developer productivity, building systems that make it easier to search for important scenarios, prepare training-ready data, and support fast iteration across perception and evaluation workflows.Our goal is to make world understanding scalable, practical, and cost efficient for embodied AI systems. We believe the next generation of autonomy and robotics depends not only on stronger models, but also on better infrastructure for turning massive volumes of multimodal data into reusable signals, searchable artifacts, and high-quality evaluation loops. That means building systems that can operate at industrial scale while preserving the flexibility to adapt quickly to new questions, new edge cases, and new model capabilities.A core idea behind how we work is EMWU, or Efficient Multi-Tier World Understanding. At a high level, EMWU is a cost-aware approach that first performs the cheapest reusable work, such as detection, featurization, and retrieval, and then applies deeper reasoning only where it adds meaningful value. This operating model reflects how Cola engineers think: build durable intermediate artifacts, design for scale from the start, and balance quality, speed, and cost instead of optimizing any one of them in isolation.The RoleWe are looking for a hands-on machine learning engineer to help build the data processing, featurization, and inference foundations that power scalable world understanding. This role is ideal for someone who is equally comfortable working on machine learning systems, production infrastructure, and evaluation loops, and who enjoys turning ambiguous problems into practical, reliable solutions.What You’ll DoDesign, build, and productionize data processing and featurization pipelines for large-scale multimodal dataImprove inference frameworks for computer vision and multimodal models, with a focus on reliability, extensibility, and operational simplicityDrive scalability and cost efficiency across the end-to-end pipeline, including compute utilization, throughput, storage, and query performanceWork closely with partners across machine learning, infrastructure, and evaluation to deliver systems that support both rapid experimentation and production useDevelop and refine evaluation methods for model quality, retrieval quality, and system-level performanceHelp shape technical direction through strong execution, thoughtful tradeoff analysis, and clear engineering judgmentTake ownership of ambiguous problem spaces, define practical paths forward, and move quickly from prototype to productionOperate with urgency and a strong bias toward execution velocity while maintaining a high bar for engineering qualityYour Skills & AbilitiesBS, MS, or PhD in Computer Science, Electrical Engineering, Robotics, or a related technical field, or equivalent practical experienceExperience building production data processing or machine learning pipelines at scaleExperience with featurization, embedding, inference, or retrieval systems for vision or multimodal workloadsStrong understanding of computer vision models and the practical challenges of deploying them in production environmentsExperience evaluating machine learning systems using clear metrics, experiments, and regression safeguardsProven ability to work hands-on in fast-moving environments with incomplete informationStrong ownership mindset, sound technical judgment, and the ability to drive execution through ambiguityWhat Will Give You A Competitive EdgeExperience with world models or large-scale world understanding systemsExperience with simulation workflows or synthetic data systemsExperience with vector search, approximate nearest neighbor retrieval, or large-scale embedding infrastructureExperience working on embodied AI, autonomous systems, or safety-critical machine learning applicationsCompensationThe salary range for this role is $117,700 and $221,400. The actual base salary a successful candidate will be offered within this range will vary based on factors relevant to the position (along with level.) Bonus Potential: An incentive pay program offers payouts based on company performance, job level, and individual performance. Benefits:GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays, tuition assistance programs, employee assistance program, GM vehicle discounts and more. This role is based remotely, but if the selected candidate lives within a specific mile radius of a GM hub, they will be expected to report to the location three times a week {or other frequency dictated by your manager}. This job may be eligible for relocation benefits. About GMOur vision is a world with Zero Crashes, Zero Emissions and Zero Congestion and we embrace the responsibility to lead the change that will make our world better, safer and more equitable for all.Why Join UsWe believe we all must make a choice every day – individually and collectively – to drive meaningful change through our words, our deeds and our culture. Every day, we want every employee to feel they belong to one General Motors team.Benefits OverviewFrom day one, we're looking out for your well-being–at work and at home–so you can focus on realizing your ambitions. Learn how GM supports a rewarding career that rewards you personally by visiting Total Rewards resources.Non-Discrimination and Equal Employment Opportunities (U.S.)General Motors is committed to being a workplace that is not only free of unlawful discrimination, but one that genuinely fosters inclusion and belonging. We strongly believe that providing an inclusive workplace creates an environment in which our employees can thrive and develop better products for our customers.All employment decisions are made on a non-discriminatory basis without regard to sex, race, color, national origin, citizenship status, religion, age, disability, pregnancy or maternity status, sexual orientation, gender identity, status as a veteran or protected veteran, or any other similarly protected status in accordance with federal, state and local laws. We encourage interested candidates to review the key responsibilities and qualifications for each role and apply for any positions that match their skills and capabilities. Applicants in the recruitment process may be required, where applicable, to successfully complete a role-related assessment(s) and/or a pre-employment screening prior to beginning employment. To learn more, visit How we Hire.AccommodationsGeneral Motors offers opportunities to all job seekers including individuals with disabilities. If you need a reasonable accommodation to assist with your job search or application for employment, emailus or call us at View phone number on click.appcast.io. In your email, please include a description of the specific accommodation you are requesting as well as the job title and requisition number of the position for which you are applying.SummaryLocation: Sunnyvale, California, United States of AmericaType: Full time

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the AI/ML Engineer - Model Inference in Sunnyvale, CA vacancy
  •  ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by one...  ...next-generation robotic platforms. As a Senior AI/ML Research Engineer, you will develop and fine-tune the foundation models—VFMs, VLMs, and VLA models—that let our Embodied... 
    Suggested
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    8 hours ago
  •  ...Systems builds the world's largest AI chip, 56 times larger than...  ...-leading training and inference speeds; over 10 times faster...  ...Cerebras works with the leading model labs, global enterprises, and...  ...on a loop."You'll sit between engineering, product, and customer-facing... 
    Suggested

    Cerebras Systems

    Sunnyvale, CA
    2 days ago
  • $193.3k - $261.5k

     ...Inferentia and Trainium ML accelerators. This...  ...enabling unparalleled ML inference and training performance...  ...running a wide range of models and supporting novel architecture...  ...software boundary, our engineers build systematic...  ...of what's possible in AI acceleration.As part of... 
    Suggested
    Work experience placement
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $193.3k - $261.5k

     ...s custom cloud-scale machine learning accelerators. Join us to optimize the latest models to run really fast on the Trainium hardware.As a Software Development Engineer on the Inference Model Enablement team, you will onboard and optimize state-of-the-art open-source and... 
    Suggested
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $170k - $190k

     ...simple: deliver the best AI-powered customer...  ...on outcomes. ASAPP’s AI Engineering team is seeking an enterprising...  ...experienced Lead AI/ML Engineer to join our...  ...expertise in foundational models and enterprise AI...  ...interruption handling, streaming inference, and audio quality, and... 
    Suggested
    Full time

    Asapp

    Mountain View, CA
    12 hours ago
  • $170.6k - $261.3k

     ...us. About the team: The AV ML Infra team at GM builds end...  ...meet the unique demands of AI and ML innovation, supporting...  ...the productivity of ML engineers, and drive the adoption of...  ...includes: AI Validation & Inference: Ensures robust model performance by running large... 
    Full time
    Local area
    Work from home
    Flexible hours

    General Motors

    Sunnyvale, CA
    1 day ago
  •  ...computing experiences—from AI and data centers, to PCs, gaming...  ...THE ROLEWe are hiring AI / ML Platform Engineers to build the platform layer...  ..., distributed training and inference, experiment tracking, benchmark...  ..., kernel benchmarking, model serving, or distributed training... 

    AMD

    Santa Clara, CA
    3 days ago
  • $275.8k - $340.5k

     ...About the team: The AV ML Infra team at GM builds ML...  ...meet the unique demands of AI and ML innovation, supporting...  ...enhance the productivity of ML engineers, and drive the adoption of...  ...: AI Validation & Inference: Ensures robust model performance by running large... 
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    3 days ago
  • $184k - $287.5k

     ...NVIDIA, we aren't just powering the AI revolution—we're accelerating it. The TensorRT inference platform is the backbone of...  ...of cutting-edge deep learning models on every NVIDIA GPU. With demand...  ...a highly skilled and driven Engineering Manager to take the lead in developing... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $105k - $115k

     ...Quest Global delivers world-class end-to-end engineering solutions by leveraging our deep industry...  ...struggles with by utilizing Gen AI or other machine learning techniques Experience...  ...APIsExperience in cloud platformsAI Engineer/ ML Engineer with python or typescript. They... 
    Temporary work

    Quest Global Services

    Sunnyvale, CA
    4 days ago
  • $144.7k - $261.3k

     ...environments, cloud infrastructure, and ML/AI GPU platforms for AV research and development...  ...: GM is looking for a Senior Capacity Engineer to join the AV Capacity and Performance Engineering...  ...) Develop and refine technical models for capacity planning for GM’s AV infrastructure... 
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Work from home
    Flexible hours
    3 days per week

    General Motors

    Sunnyvale, CA
    1 day ago
  • $182.4k - $250.6k

     ...autonomous driving? Join the Embodied AI team at General Motors. Our team is developing...  ...real-world scenarios.As a Senior AI/ML Future Sensing Engineer in the Embodied AI organization, you...  ...and improving ML and perception models that support safe and reliable vehicle... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  •  ...generation computing experiences—from AI and data centers, to PCs,...  ...is seeking an experienced AI/ML Engineer to join our Applied Research...  ...optimize state-of-the-art AI models for image generation,...  ....Build scalable training and inference pipelines leveraging high-performance... 

    AMD

    San Jose, CA
    2 days ago
  • $151.8k - $332.2k

    What you can expect We are looking for an AI Inference Engineer with a solid background in speech recognition and model inference. In this role, you will develop state-of-the...  ...Python, shell scripts, C/C++; familiarity with ML frameworks such as PyTorch and TensorFlow.... 
    Full time
    Work at office
    Remote work

    Zoom

    San Jose, CA
    2 days ago
  •  ...GorusuCompany: SRI Tech SolutionsJob Title: AI / ML engineerLocation: Sunnyvale, CANOTE:...  ...and experienced Senior Machine Learning Engineer / AI Solutions Architect with over 7 years...  ...developing, and deploying machine learning models and AI solutions. The ideal candidate will... 
    Full time

    SRI Tech

    Sunnyvale, CA
    2 days ago
  • $212.7k - $287.7k

     ...accelerators. Join us to optimize LLMs to run really fast on the Trainium hardware.As an SDM for the LLM Inference Model Enablement team, you will lead a team of expert AI/ML engineers to onboard and optimize state-of-the-art open-source and customer LLMs, both dense and MoE, for... 
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $250k - $350k

    About the RoleWe are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you...  ...inference acceleration, GPU parallelism, advanced model deployment, and video generation technologies. Your... 
    Work at office
    3 days per week

    Pika

    Palo Alto, CA
    3 days ago
  • $190k - $260k

     ...developed an artificial intelligence (AI) powered technology stack...  ...it. Every improvement to our models - from GigaFusionNet to large-...  .... We are looking for engineers who make model training fast:...  ...streamable formatsPartner with ML teams to scale new architectures... 
    Temporary work
    Work at office
    Visa sponsorship

    Kodiak Robotics

    Mountain View, CA
    2 days ago
  • $296.3k - $423.9k

     ...Generation team within the Embodied AI organization. This role...  ...will shape the algorithms and ML systems that translate sensor...  ...lead a high-performing team of engineers building ML-driven trajectory...  ...the roadmap that takes these models from research to production deployment... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • $207k - $300k

     ...generation infrastructure that powers our AI/ML applications. You will operate at the...  ...infrastructure, including creative data engines, generation and rendering pipelines, and...  ...and learning infrastructure to accelerate model training, evaluation, and rapid deployment... 

    Google

    Mountain View, CA
    1 day ago
  • $307k - $422k

     ...About the teamX’s central Applied AI team is looking for a leader...  ...and develop a diverse team of ML/AI experts, as well as guide...  ...techniques with a focus on foundation models and transformer architectures....  ...).Provide research and engineering leadership on X early... 
    Full time

    X Company

    Mountain View, CA
    2 days ago
  • $189.3k - $320.7k

     ...autonomous driving? Join the Embodied AI team at General Motors. Our team is developing...  ...real-world scenarios.As a Staff AI/ML Future Sensing Engineer in the Embodied AI organization, you...  ...Architect and evaluate perception models and pipelines for detection, reconstruction... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • $129k - $198.4k

    Job DescriptionRole: As an AI/ML Engineer on the Metrics Frameworks team, part of the Simulation, Evaluation, and Data organization, you will...  ...vehicle engineering to enable rapid development and model feedback. Maintain a high technical standard for code quality... 
    Full time
    Local area
    Work from home

    General Motors

    Sunnyvale, CA
    1 day ago
  • $148.7k - $297.3k

     ...female executives, and scientists.THE OPPORTUNITYThis Principal AI/ML Engineer position can work out of our Santa Clara, CA location.The...  ...science and engineering. Implement robust CI/CD workflows for ML models, including testing, rollout, rollback strategies, and compliance... 

    Abbott

    Santa Clara, CA
    4 days ago
  • $169k - $338k

     ...OfficePosition Summary...What you'll do...As a Distinguished AI/ML Engineer within Walmart Global Tech’s Reliability Engineering Organization...  ...for Reliability Engineering, including:Large language models for automated incident responseReinforcement learning agents for... 
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    8 hours ago
  • $120.8k - $193.3k

    DescriptionJob Title: Sr. Engineer, AI/ML Apps Engineering Job Location: San Jose, CA (This position...  ...use cases.Help customers integrate AI models, perception pipelines, AI agents, and...  ...preferred.Strong understanding of AI inference optimization, model deployment, and... 
    Full time
    Work at office

    SiMa Technologies

    San Jose, CA
    2 days ago
  • $165.2k - $223.6k

     ...introduction deployment. This involves building agents, working with AI/ML such as VLM/VLA/LLMs that drive robotic automation in...  ...incorporating AI/ML trends.A day in the lifeAs a Software Development Engineer, you will be responsible for driving world-class manufacturing... 
    Contract work
    Internship
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    3 days ago
  •  ...Job Title: Senior AI/ML Engineer Work Location with ZIP: Sunnyvale, CA 94085 (Hybrid) Technical Hiring Criteria (Must Haves) Top 3...  ...with Generative AI / LLMs (OpenAI, Azure OpenAI, open'source models) - Experience with Prompt Engineering, RAG, LangChain / Semantic... 

    eTeam

    Sunnyvale, CA
    3 days ago
  •  ...Job Title: AI/ML Engineer Location: Cupertino, CA Onsite/ Remote: Onsite JD: Build and optimize RAG pipelines, vector databases, embeddings, and document-processing workflows Design agentic AI systems - including tool calling, orchestration, reasoning loops... 
    Remote work

    Yantran LLC

    Cupertino, CA
    12 hours ago
  • $150k - $225k

     ...quality of life. As a Senior AI / Embedded Engineer, you will be responsible for...  ...includes data ingestion, model development, optimization,...  ...reliable, low-power, real-time ML systems that operate at the...  ...to reduce model size and inference latency ◦ Use frameworks including... 
    Full time
    Work at office
    Immediate start
    Visa sponsorship
    Night shift

    E-Space

    Saratoga, CA
    12 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI/ML Engineer - Model Inference. Be the first to apply!