AI Engineer, Model Distillation, Tesla AI
$124kFull-time
Tesla
What To Expect
At Tesla AI, we are solving real-world autonomy at a scale that does not exist anywhere else. We are building the brains for millions of robots and to achieve this, we train frontier-scale foundation models on the world’s richest, most diverse dataset of multimodal telemetry, vision, language, and robotic action. You will have access to unparalleled resources that set us apart from other companies in the AI industry. You will work with one of the richest real-world driving, robotics and human-interaction datasets available, spanning vision, language, action, and on-robot telemetry providing a unique environment to study how capability transfers from large teacher models into compact students that must run under strict latency, power, and reliability constraints. Tesla also offers one of the highest GPU resources per engineer in the industry, giving you the computational budget to train frontier-scale teachers, generate high-quality distillation corpora at volume, and iterate on student architectures far beyond typical research environments. These resources enable you to run distillation experiments across language, multimodal, and policy models at a fidelity and scale unmatched elsewhere, and to close the loop by deploying distilled models onto FSD / Optimus in the real world. What You'll Do Design and execute state-of-the-art distillation pipelines that transfer complex reasoning, vision-language understanding, and action policies from massive teacher models to ultra-efficient student models for FSD (Full Self-Driving), Optimus, and Digital Optimus Develop and iterate on advanced distillation methodologies, including logit-level, sequence-level, and trajectory-level distillation. You will pioneer novel recipes combining synthetic data generation, soft-label training, supervised fine-tuning (SFT), and Reinforcement Learning (RL)Conduct experiments to uncover the scaling behaviors of distillation: teacher size, student capacity, student architecture design, data mixture, compute allocation, and how each factor affects capability retention, latency, and on-robot performance Build and maintain infrastructure for efficient large-scale teacher inference, distillation data curation, and distributed student training, resolving compute and memory bottlenecks end-to-end Evaluate distilled models against teacher fidelity and product metrics like task success, robustness, and real-world behaviors and close gaps where students diverge from teachers Work closely with cross-functional teams to ship distilled models to production, meeting stringent performance, safety, and reliability standards Contribute tools and frameworks that make distillation reproducible, measurable, and reusable across Tesla AI model families What You'll Bring Deep, proven expertise in deep learning fundamentals, particularly in training, compressing, or distilling large-scale language, vision, or multimodal models Strong grasp of teacher–student interaction, synthetic data, and the tradeoffs between capability, latency, and compute In-depth knowledge of loss design (e.g., KL / soft-target objectives), and modern neural architectures (mixture of experts, hybrid attention, and beyond) Hands-on familiarity with post-training methods such as distillation, supervised fine-tuning, and policy optimization / RL Strong expertise in distributed computing and large-scale training or inference pipelines Proficiency in Python and a deep understanding of software engineering best practices Experience with deep learning frameworks such as PyTorch, TensorFlow, or JAX Demonstrated ability to work collaboratively in a cross-functional team environment Strong problem-solving skills and the ability to troubleshoot complex system-level issues across data, training, and deployment Benefits
Compensation and Benefits
Along with competitive pay, as a full-time Tesla employee, you are eligible for the following benefits at day 1 of hire:
Medical plans > plan options with $0 payroll deduction Family-building, fertility, adoption and surrogacy benefits Dental (including orthodontic coverage) and vision plans, both have options with a $0 paycheck contribution Company Paid (Health Savings Accounts) HSA Contribution when enrolled in the High-Deductible medical plan with HSA Healthcare and Dependent Care Flexible Spending Accounts (FSA) 401(k) with employer match, Employee Stock Purchase Plans, and other financial benefits Company paid Basic Life, AD&D Short-term and long-term disability insurance (90 day waiting period) Employee Assistance Program Sick and Vacation time (Flex time for salary positions, Accrued hours for Hourly positions), and Paid Holidays Back-up childcare and parenting support resources Voluntary benefits to include: critical illness, hospital indemnity, accident insurance, theft & legal services, and pet insurance Weight Loss and Tobacco Cessation Programs Tesla Babies program Commuter benefits Employee discounts and perks program Expected Compensation $124,000 - $558,000/annual salary + cash and stock awards + benefits Pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. The total compensation package for this position may also include other elements dependent on the position offered. Details of participation in these benefit plans will be provided if an employee receives an offer of employment. , Tesla
At Tesla AI, we are solving real-world autonomy at a scale that does not exist anywhere else. We are building the brains for millions of robots and to achieve this, we train frontier-scale foundation models on the world’s richest, most diverse dataset of multimodal telemetry, vision, language, and robotic action. You will have access to unparalleled resources that set us apart from other companies in the AI industry. You will work with one of the richest real-world driving, robotics and human-interaction datasets available, spanning vision, language, action, and on-robot telemetry providing a unique environment to study how capability transfers from large teacher models into compact students that must run under strict latency, power, and reliability constraints. Tesla also offers one of the highest GPU resources per engineer in the industry, giving you the computational budget to train frontier-scale teachers, generate high-quality distillation corpora at volume, and iterate on student architectures far beyond typical research environments. These resources enable you to run distillation experiments across language, multimodal, and policy models at a fidelity and scale unmatched elsewhere, and to close the loop by deploying distilled models onto FSD / Optimus in the real world. What You'll Do Design and execute state-of-the-art distillation pipelines that transfer complex reasoning, vision-language understanding, and action policies from massive teacher models to ultra-efficient student models for FSD (Full Self-Driving), Optimus, and Digital Optimus Develop and iterate on advanced distillation methodologies, including logit-level, sequence-level, and trajectory-level distillation. You will pioneer novel recipes combining synthetic data generation, soft-label training, supervised fine-tuning (SFT), and Reinforcement Learning (RL)Conduct experiments to uncover the scaling behaviors of distillation: teacher size, student capacity, student architecture design, data mixture, compute allocation, and how each factor affects capability retention, latency, and on-robot performance Build and maintain infrastructure for efficient large-scale teacher inference, distillation data curation, and distributed student training, resolving compute and memory bottlenecks end-to-end Evaluate distilled models against teacher fidelity and product metrics like task success, robustness, and real-world behaviors and close gaps where students diverge from teachers Work closely with cross-functional teams to ship distilled models to production, meeting stringent performance, safety, and reliability standards Contribute tools and frameworks that make distillation reproducible, measurable, and reusable across Tesla AI model families What You'll Bring Deep, proven expertise in deep learning fundamentals, particularly in training, compressing, or distilling large-scale language, vision, or multimodal models Strong grasp of teacher–student interaction, synthetic data, and the tradeoffs between capability, latency, and compute In-depth knowledge of loss design (e.g., KL / soft-target objectives), and modern neural architectures (mixture of experts, hybrid attention, and beyond) Hands-on familiarity with post-training methods such as distillation, supervised fine-tuning, and policy optimization / RL Strong expertise in distributed computing and large-scale training or inference pipelines Proficiency in Python and a deep understanding of software engineering best practices Experience with deep learning frameworks such as PyTorch, TensorFlow, or JAX Demonstrated ability to work collaboratively in a cross-functional team environment Strong problem-solving skills and the ability to troubleshoot complex system-level issues across data, training, and deployment Benefits
Compensation and Benefits
Along with competitive pay, as a full-time Tesla employee, you are eligible for the following benefits at day 1 of hire:
Medical plans > plan options with $0 payroll deduction Family-building, fertility, adoption and surrogacy benefits Dental (including orthodontic coverage) and vision plans, both have options with a $0 paycheck contribution Company Paid (Health Savings Accounts) HSA Contribution when enrolled in the High-Deductible medical plan with HSA Healthcare and Dependent Care Flexible Spending Accounts (FSA) 401(k) with employer match, Employee Stock Purchase Plans, and other financial benefits Company paid Basic Life, AD&D Short-term and long-term disability insurance (90 day waiting period) Employee Assistance Program Sick and Vacation time (Flex time for salary positions, Accrued hours for Hourly positions), and Paid Holidays Back-up childcare and parenting support resources Voluntary benefits to include: critical illness, hospital indemnity, accident insurance, theft & legal services, and pet insurance Weight Loss and Tobacco Cessation Programs Tesla Babies program Commuter benefits Employee discounts and perks program Expected Compensation $124,000 - $558,000/annual salary + cash and stock awards + benefits Pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. The total compensation package for this position may also include other elements dependent on the position offered. Details of participation in these benefit plans will be provided if an employee receives an offer of employment. , Tesla
Vacancy posted 14 days ago
Similar jobs that could be interesting for youBased on the AI Engineer, Model Distillation, Tesla AI in California vacancy
- ...Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This... ...computation. Cerebras works with the leading model labs, global enterprises, and cutting-... ...do this on a loop." You'll sit between engineering, product, and customer-facing teams....SuggestedFull time
$190k - $260k
...has developed an artificial intelligence (AI) powered technology stack purpose-built... ...we can train it. Every improvement to our models - from GigaFusionNet to large-scale world... ...training throughput. We are looking for engineers who make model training fast: streaming massive...SuggestedTemporary workWork at officeVisa sponsorship$117.7k - $221.4k
...practical, and cost efficient for embodied AI systems. We believe the next generation... ...and robotics depends not only on stronger models, but also on better infrastructure for... .... This operating model reflects how Cola engineers think: build durable intermediate artifacts...SuggestedFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours$180k - $225k
As a Software Engineer on the ML Infrastructure team, you will design and build platforms for... ...and engineers to integrate and optimize models for production and research use cases.... ...Scale, our mission is to develop reliable AI systems for the world's most important decisions...SuggestedFull time- ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by one... ...next-generation robotic platforms. As a Senior AI/ML Research Engineer, you will develop and fine-tune the foundation models—VFMs, VLMs, and VLA models—that let our Embodied...SuggestedLocal areaWorldwideFlexible hours
$150k - $250k
...About Distyl AI Distyl is an applied AI technology company... ...What We Are Looking For AI Engineers build and operate production... ...compound AI workflows that combine models, prompts, agents, tools,... ...Engineers, AI Strategists, and other Distillers to make pragmatic system...Full timeWork at officeFlexible hours3 days per week$50k - $120k
...Mission Altimate AI, founded in 2022 in San Francisco... ...that combines multiple language models and a custom-built knowledge... ...forefront of the AI-powered data engineering revolution. You can read more... ...in model compression, distillation, and deployment at scale Model...Full timeWorldwide- ...patients worldwide.We’re a team of engineers, clinicians, and innovators... ...of Position:The Finance AI Engineer designs, builds, and... ...tools think in terms of financial models and business outcomes, not technical... ...skills, with the ability to distill complex AI and data topics for...Work at officeLocal areaWorldwideFlexible hours
$150k - $250k
...will build state-of-the-art AI capabilities for Cylake's next... ...train deep learning and LLM models that enable security teams to... ...security researchers and software engineers to develop, productize, and... ...optimization, and model distillation. ~ Strong understanding of...Full time$300k - $400k
...reinvent how designers work in the AI era. We’re backed by top... ...We’re hiring an Lead AI Engineer to own and scale our AI infrastructure... ...intersection of research, model training, and product—you’ll... ...inference with quantization, distillation, and caching techniques....Full time- ...The role: SoFi’s Staff AI Engineer is a hands-on AI engineering role in SoFi’s growing... ...distributed nodes and regional failovers Deep Model Optimization: Pioneer and... ...Coordinate with cross-functional teams to distill specific requirements, project roadmaps,...Full time
- ...that sit at the intersection of AI, biology, chemistry, and large-scale engineering. Our goal is to translate complex... ...of large-scale machine learning models that form the core of Absentia Labs... ...Familiarity with model compression, distillation, or inference optimization....Full timeRemote workFlexible hours
$70k - $200k
...Who are We? Field AI is transforming how robots interact with... ...results and rapidly improving models through real-field applications... ...empower our customers and engineers to deliver value. Field-insight... ...Boston Dynamics, NVIDIA, Amazon, Tesla Autopilot, Cruise Self-Driving...Full timeRemote work$70k - $300k
...Field AI is transforming how robots interact with the real world.... ...world results and rapidly improving models through real-field applications. Robotics AI Engineer – Calibration, Localization, and... ...Dynamics, NVIDIA, Amazon, Tesla Autopilot, Cruise Self-Driving,...Full timeWork at officeRemote workFlexible hours$197.5k - $272k
...driving the transformation to AI-enabled software-defined... ...looking for a great Staff AI Engineer to join our seasoned AI team... ...you will build and deploy AI models that analyze continuous data... ...Apply quantization, pruning, distillation, and memory optimization to ensure...Work at officeWorldwideFlexible hoursShift work3 days per week$176.6k - $239k
...think big about how Physical AI, robotics, and simulation technologies... ...Prototyping and AI Customer Engineering (PACE) team transforms... ...synthetic data pipelines, and robotic model training. You'll leverage AWS... ...across the AWS ecosystem to distill customer needs and influence...Local areaWorldwideFlexible hoursDay shift- ...join us in our journey? Most engineers sit behind a desk. This one doesn't. As our Field AI Engineer , you'll be embedded... ...something useful. This role is modeled on the Forward Deployed... ...noisy week of conversations and distill it into three things that actually...Full timeImmediate start
- ...SignalFire’s Talent Network for Principal AI/ML Engineer Roles at VC-Backed Startups This is... ...edge machine learning and deep learning models ✔ Experienced in architecting... ...ONNX, TensorRT, pruning, quantization, distillation What Happens Next? # Submit your application...Full time
$152k - $241.5k
...seeking a Senior GenAI Algorithms Engineer to advance the state of the art in foundation model development, training, and... ...VLMs, model efficiency, multimodal AI, and open-source AI infrastructure... ...4, INT4), pruning, knowledge distillation, neural architecture search, and...Full timeRemote work$251.1k - $385k
...It all started when engineer Fred Luddy wrote code that automated... ...work. Today, ServiceNow is the AI control tower for business reinvention... ...team, building the operating model, technical foundation, and... ...orchestration, fine-tuning, distillation, function calling, and LLMOps...Full timeWork at officeImmediate startRemote workFlexible hours$224k - $280k
...a Job Info Back to Free Job Search AI Engineer, Intern Postman ~ Berkeley, CA, US... ...coursework-style projects. What You'll Do Model Development Partner with product... ...work (quantization, batching, distillation) under supervision -this is a learning...Full timeFreelanceInternshipWork at officeRemote workFlexible hours3 days per week$90k - $115k
...Generative AI Engineer – Remote Bright Vision Technologies is a technology consulting... ...fine-tuning workflows for large language models across supervised, preference-based, and... ...synthetic data generation and dataset distillation. Open-source contributions to LLM training...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship$187k - $292k
...differentiated by its expertise in imagining, engineering, and delivering robots with advanced... ...in simulation. About the Role The AI Controls team builds high-rate learned... ...tactile sensing. Familiarity with policy distillation (e.g. teacher–student) for transferring...Full timeTemporary workRelocation packageFlexible hours$220k - $320k
...and hosts specialized language models for companies that need frontier-quality AI at a fraction of the cost. The models... ...handles everything end-to-end: distillation, training, evaluation, and... ...well-funded ten-person team of engineers who work in-person in downtown San...Full timeWork at office$150k - $225k
...overall quality of life. As a Senior AI / Embedded Engineer, you will be responsible for the full... .... This includes data ingestion, model development, optimization, and deployment... ...quantization, pruning, and knowledge distillation to reduce model size and inference latency...Full timeWork at officeImmediate startVisa sponsorshipNight shift$190k - $240k
...members.What You’ll Do:As a Staff AI and ML fundamentalist who is... ...algorithms.Iterate on AI model development; starting from the... ...with other researcher engineers to prototype and validate complex... ...policies, fine-tuning, model distillation, mixture of experts, and other...Local area$180k - $250k
...-8 years experience as a Cyber Security Engineer, with hands-on experience in security research... ...such as Quantization, FlashAttention, Distillation, DeepSpeed-Inference / FasterTransformer... ...Description Insight Global is hiring a AI Cyber Security Engineer join the team,...Full time- ...through an intelligent, auditable AI platform that spans the... ...seeking a Principal Generative AI Engineer to serve as the technical... ...the engineering standards for model evaluation, system... ...engineering (quantization, pruning, distillation, tensor/model parallelism) for...Full time
$148.19k - $231.98k
...SalesforceSalesforce is the #1 AI CRM, where humans with agents... ...foundation, a clear operating model, and an adoption path that scales... ..., and close to Product, Engineering, Support, and delivery teams.... ...idea to execution.The ability to distill and communicate complex technical...Full timeWork experience placementImmediate startRemote work$190k - $250k
...has developed an artificial intelligence (AI) powered technology stack purpose-built... ...developing large-scale generative world models that learn to predict realistic, physically... ...and long-tail scenario generation, and is distilled into the onboard models that drive our...Temporary workWork at officeVisa sponsorship
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Engineer, Model Distillation, Tesla AI. Be the first to apply!


