Senior ML Accelerator Engineer - GPU
$170.1k - $258.3kGeneral Motors
Job DescriptionAbout the MissionGM’s vision of Zero Crashes, Zero Emissions, and Zero Congestion guides everything we do in autonomous and assisted driving. The AV organization is building advanced automated driving technologies, including Level 4–capable fully self-driving systems, to move us toward safer, more sustainable, and more accessible mobility. For the AI Kernels & Compilers team, that mission shows up in the details: turning cutting‑edge perception, prediction, and planning research into production‑grade software that can run efficiently and reliably on real vehicles at scale. We pioneer new approaches to model export, kernel development, and performance engineering so that every cycle on our accelerators translates into better situational awareness, faster reaction times, and more robust behavior on the road. If you want your compiler and kernels work to directly influence how automated vehicles understand and react to the world — while operating at the safety, reliability and scale of a company like GM — this is where that impact becomes real.About the TeamThe AI Kernels team builds high‑performance GPU kernels and custom libraries that sit at the heart of our on‑vehicle ML inference for ADAS and autonomous driving. We own making core AI workloads faster, more reliable, and easier to maintain and deploy on real cars, under real‑world constraints.That means: Designing and implementing custom operators when vendor libraries hit their limitsIntegrating those kernels deep into our ML runtime stackDebugging and tuning GPU performance across the AV software stack, often on hardware‑in‑the‑loopWe partner closely with AI Solutions, AI Compilers, AI Architecture, and AI Tooling to ensure models deploy efficiently to the car while consistently meeting strict latency, throughput, and reliability targets. If you enjoy pushing GPUs to their limits and seeing your work directly impact how autonomous vehicles perceive and act in the world, this is the team for you.What you’ll be doing (Responsibilities) Design, implement, benchmark, and iterate on CUDA-based kernels and custom operators to squeeze every last drop of performance out of on-vehicle inference workloads.Build and improve tooling and infrastructure that make it easier to profile, debug, and validate CUDA kernels and accelerator-backend code across the AV stack.Partner with AI Solutions, Compilers, and Architecture to translate model and system requirements into concrete kernel roadmaps, priorities, and project plans.Collaborate with cross-functional teams (compiler, performance tooling, runtime, deployment solutions) to deliver reusable, reliable, high-performance libraries into production.Maintain high technology standards, methodologies, processes, and guidelines for GPU kernel development and performance engineering through code review.Manage relationships with internal customers to ensure our kernels and libraries meet real-world needsYour Skills & Abilities (Required Qualifications) Minimum 2+ years of relevant industry experience or equivalent experienceBS, MS or PhD in CS, or related technical fieldExcellent GPU programming skills in CUDA, with a thorough understanding of parallel programming patterns and GPU architecture.Hands-on experience benchmarking, profiling, debugging and optimizing accelerator libraries and kernels to extract optimal performance using the NSight suite of tools or similar.Strong background in software architecture, library design, and design patterns.Strong C++ programming skills with the ability to feel comfortable in large codebases.Solid background in system performance, high performance computing and/or architecture-aware optimizations. Strong communication skills and the ability to work collaboratively within a teamExcellent analytical and problem-solving skillsWhat Will Give You A Competitive Edge (Preferred Qualifications) 2+ years of relevant industry experience or equivalent experienceExperience with tensor core programming, CUTLASS and/or CuTeExperience with ML model architectures, in particular transformer-basedExperience with low latency or real time systemsExperience with lower levels of an accelerator software stack (i.e. drivers, runtimes, and compilers)Compensation: The compensation information is a good faith estimate only. It is based on what a successful applicant might be paid in accordance with applicable state laws. The compensation may not be representative for positions located outside of New York, Colorado, California, or Washington. The salary range for this role: is $170,100 to $258,300. The actual base salary a successful candidate will be offered within this range will vary based on factors relevant to the position. Bonus Potential: An incentive pay program offers payouts based on company performance, job level, and individual performance. Benefits: GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays, tuition assistance programs, employee assistance program, GM vehicle discounts and more. This role is categorized as hybrid. This means the selected candidate is expected to report to a specific location at least 3 times a week {or other frequency dictated by their manager}. The selected candidate will be required to travel <25% for this role. This job may be eligible for relocation benefits. About GMOur vision is a world with Zero Crashes, Zero Emissions and Zero Congestion and we embrace the responsibility to lead the change that will make our world better, safer and more equitable for all.Why Join UsWe believe we all must make a choice every day – individually and collectively – to drive meaningful change through our words, our deeds and our culture. Every day, we want every employee to feel they belong to one General Motors team.Benefits OverviewFrom day one, we're looking out for your well-being–at work and at home–so you can focus on realizing your ambitions. Learn how GM supports a rewarding career that rewards you personally by visiting Total Rewards resources.Non-Discrimination and Equal Employment Opportunities (U.S.)General Motors is committed to being a workplace that is not only free of unlawful discrimination, but one that genuinely fosters inclusion and belonging. We strongly believe that providing an inclusive workplace creates an environment in which our employees can thrive and develop better products for our customers.All employment decisions are made on a non-discriminatory basis without regard to sex, race, color, national origin, citizenship status, religion, age, disability, pregnancy or maternity status, sexual orientation, gender identity, status as a veteran or protected veteran, or any other similarly protected status in accordance with federal, state and local laws. We encourage interested candidates to review the key responsibilities and qualifications for each role and apply for any positions that match their skills and capabilities. Applicants in the recruitment process may be required, where applicable, to successfully complete a role-related assessment(s) and/or a pre-employment screening prior to beginning employment. To learn more, visit How we Hire.AccommodationsGeneral Motors offers opportunities to all job seekers including individuals with disabilities. If you need a reasonable accommodation to assist with your job search or application for employment, emailus or call us at View phone number on click.appcast.io. In your email, please include a description of the specific accommodation you are requesting as well as the job title and requisition number of the position for which you are applying.SummaryLocation: Sunnyvale, California, United States of America; Remote - Washington; Austin, Texas, United States of America; San Francisco, California, United States of America; Warren, Michigan, United States of AmericaType: Full time
- A cutting-edge AI technology company based in San Francisco is seeking a specialist to design and operate large-scale GPU infrastructure. This role requires expertise in deploying GPU systems for high-throughput inference and model performance optimization. The ideal candidate...Senior
- A leading AI infrastructure company is seeking a Senior ML Performance Engineer to design a comprehensive performance testing platform for large language... ...in performance engineering and strong experience with GPU programming and ML inference workloads. Candidates should...Senior
- ...company in San Francisco is looking for a Senior Software Engineer to build scalable infrastructure for... ...distributed training systems and optimize GPU utilization while collaborating with... ...candidates have over 5 years of experience in ML infrastructure and a strong background...Senior
$128.7k - $261.3k
...development, and performance engineering so that every cycle on our accelerators translates into better... ...compiler, systems, and GPU engineers who enjoy working... ...driving. The RoleAs a Senior Compiler Engineer on the... ...reliable, and effortless for ML engineers across the AV...SeniorFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours- MakerMaker.AI is looking for a Senior Machine Learning Systems Engineer in San Francisco. In this role, you will build and operate production inference... ..., be fluent in Python, and have strong knowledge in GPU-accelerated inference. Excellent communication skills are...Senior
- A technology startup is seeking a Founding Engineer (Systems + ML) to develop GPU-accelerated engines and build end-to-end pipelines. The ideal candidate has proven experience in C++ and GPU code, alongside a deep understanding of systems performance. You'll join as the...Full time
$200k - $260k
...and reliability. We're looking for a Senior ML Engineer to drive the model serving layer for... ...throughput to the frontier. You'll profile GPU utilization, design batching strategies... ...Work directly with state-of-the-art accelerators (H100s, H200s, B200s) to optimize voice...SeniorFull time$148.5k - $223.9k
...and you are the future of Salesforce.This role is for a Senior Machine Learning Engineer within the Trust Intelligence Platform team who will architect... ...you to find balance and be your best, and our AI agents accelerate your impact so you can do your best. Together, we’ll...SeniorFull time$151.8k - $265.35k
...verticals. We are hiring a Senior Machine Learning Engineer to build the pipelines and services... ...reaching production. Own GPU capacity and cost -... ...efficiency, and right-sizing acceleration fleets against latency SLAs. Run production ML operationally - on-call, incident...SeniorFull timeTemporary workLocal areaWorldwide$148.5k - $223.9k
...duplicating efforts. Job Category Software Engineering Job Details About Salesforce... ...future of Salesforce. This role is for a Senior Machine Learning Engineer within the... ...and be your best , and our AI agents accelerate your impact so you can do your best ....Senior$179k - $218k
Crusoe is on a mission to accelerate the abundance of energy and intelligence.... ...must be bridged.We are seeking a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture to be the definitive... ...& Telemetry: Leverage AI/ML methodologies to analyze fleet-wide...SeniorTemporary work- ...throughput and stability while advancing production-ready training pipelines. This role emphasizes collaboration with ML researchers, debugging across GPU stacks, and improving communication and memory efficiency in large-scale training environments. #J-18808-Ljbffr...Senior
- Oscar is hiring a Senior Machine Learning Inference Engineer for a full-time role in the San Francisco Bay Area. You will focus on improving efficiency for... ...has 3+ years of professional experience, deep GPU infrastructure knowledge, and strong Python and PyTorch...SeniorFull time
- Responsibilities Design, deploy, and maintain large distributed ML training and inference clusters Develop efficient, scalable end-... ...across different model scales Analyze, profile and debug low-level GPU operations to optimize performance Stay up-to-date on research...Senior
- ...funded AI startup building production-grade ML infrastructure used by enterprise customers. They are looking for a Senior AI/ML Engineer to own model training pipelines, evaluation... ...~ Experience with distributed training, GPU optimization, or inference serving ~ Pragmatic...SeniorFull time
- Crusoe is on a mission to accelerate the abundance of energy and intelligence. We are seeking a Software Engineer to join Crusoe’s Data Center Infrastructure Engineering team,... ...focusing on software for managing a fleet of GPU servers and the data centers that house them...Senior
- Position: Senior ML Performance Engineer Location: SF Bay Area (US) or Toronto (Canada) - Hybrid Employment Type: Full-Time Industry: AI Infrastructure... ...(LLMs) before and after compiler optimization on modern GPU architectures. This role sits at the intersection of ML...SeniorFull time
- Oscar Associates Limited (US) in San Francisco offers a hybrid Senior ML Inference Engineer role. You will enhance AI-native infrastructure powering... ..., strong Python and PyTorch skills, and familiarity with GPU tooling like Triton, TensorRT, and vLLM. Equity and full benefits...Senior
- Genesis AI in San Francisco is seeking a senior ML infrastructure engineer to design and optimize distributed training systems and performance-critical... ...ensure efficient hardware utilization across multi‑node GPU clusters. Join a team focused on scalable AI foundations,...Senior
- ...governance, maintain auditability, and deliver reliable outcomes at scale. About the Role We are looking for a visionary Senior ML Engineer who will bridge the gap between high-level architecture and hands-on execution, specifically focusing on simplifying...SeniorFull timeShift work
- Senior ML Systems Engineer, Frameworks & Tooling at Cohere Our mission is to scale intelligence to serve humanity. We’re training and deploying... ...models. You’ll build tools and systems that directly accelerate research and model quality. Sample Projects Build a high...SeniorFull timeWork at officeRemote workFlexible hours
- ...that will be used by 1000s of developers and enterprise users.ML performance, quality, and systems acumen-ship: Experience in tuning... ...$149,400 - $195,050 Qualifications 7+ years of software engineering experience building and operating enterprise systems, APIs, or...SeniorWork at office
$200k - $235k
...managers, data scientists, software engineers, fraud intelligence, and... ..., you'll design and build ML solutions that have direct, meaningful... ...You Will Make: As a Senior Machine Learning Engineer on... ...use LLMs and AI agents to accelerate how we build models. Write...SeniorWork experience placementCasual workLive inWork at officeRemote work$180k - $336k
...Applied AI org and seeking talented Senior/Staff Machine Learning Engineers with expertise in LLM training, evaluation, and production-oriented ML systems. You’ll work on improving Lila... ...scientific method autonomously, accelerating discovery at unprecedented speed, scale...SeniorFull timeWork at officeLocal areaFlexible hours$160k - $225k
Cacheflow is seeking a Senior Software Engineer for AI Runtime at Databricks, located in San Francisco... ...and scaling systems for large-scale GPU training, ensuring high throughput and... ...in training across expansive fleets of accelerators. The role demands 5+ years of experience...Senior$298k - $368k
...miles of driving data from a diverse set of sensors, enabling engineers like you to (1) develop methods for efficiently and... ...device inference techniques and a deep understanding of hardware acceleration. ~ Extensive experience with deep learning frameworks (e.g....SeniorFull timeRemote work$77k - $202k
...generation analytics platform tailored for healthcare payers. As a Senior Associate, you will analyze complex problems, mentor junior... ...Bachelor's Degree- At least 4 years of experience in software engineering or data engineeringWhat Sets You Apart- Master's Degree in...SeniorFull timeH1b$166k - $210.25k
RDQ127R59SummaryAs a Senior Applied ML Engineer on the Applied AI team at Databricks, you will use machine learning, scheduling, and optimization algorithms to maximize the efficiency and performance of our infrastructure. Your work will span the entire stack—from cluster...SeniorLocal areaWorldwide- ...Senior Client Engineer SAN FRANCISCO, CA ENGINEERING FULL-TIME What will you be doing? Training machine learning models over billions of data points. Quantifying predictive uncertainty using probabilistic and Bayesian methods. Creating models that quickly generalize...SeniorFull timeWork experience placement
$218.4k - $273k
...research in Robotics and developing ML pipelines for training and fine-tuning... ...exceptionally motivated and experienced Senior Machine Learning Engineer, Computer Vision to drive cutting-... ...Force. We are expanding our team to accelerate the development of AI applications....SeniorFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior ML Accelerator Engineer - GPU. Be the first to apply!
- data scientist machine learning engineer San Francisco, CA
- machine learning ai engineer San Francisco, CA
- computer vision machine learning engineer San Francisco, CA
- machine learning engineer San Francisco, CA
- ai ml engineer San Francisco, CA
- graduate machine learning engineer San Francisco, CA
- machine learning software engineer San Francisco, CA
- junior machine learning research engineer San Francisco, CA
- senior ml engineer San Francisco, CA
- srs distribution San Francisco, CA


