Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Resident - Learning From Videos (LFV)

$45 - $60 per hour

Toyota Research Institute

At Toyota Research Institute (TRI), we’re on a mission to improve the quality of human life. We’re developing new tools and capabilities to amplify the human experience. To lead this transformative shift in mobility, we’ve built a world‑class team advancing the state of the art in AI, robotics, driving, and material sciences. The Team The Learning From Videos (LFV) team in the Robotics division focuses on the development of foundation models capable of leveraging large‑scale multi‑modal (RGB, depth, flow, semantics, bounding boxes, tactile, audio, etc.) data from multiple domains (driving, robotics, indoors, outdoors, etc.) to improve downstream task performance. Our approach emphasizes training scalability: by learning from multiple modalities, models can develop useful data‑driven priors about 3D geometry, physics, and dynamics for world understanding. Our research interests include, but are not limited to: Video Generation World Models 4D Reconstruction Multi‑Modal Models Multi‑View Geometry Data Augmentation Video‑Language‑Action Models We focus primarily on embodied applications and aim to tackle some of the hardest scientific challenges in spatio‑temporal reasoning, enabling autonomous agents to operate in real‑world, unstructured environments. The AI Resident This year‑long AI Residency is a research‑focused position designed for early‑career researchers and engineers who are excited to work on ambitious problems in embodied AI. The resident will be deeply integrated into the LFV team, contributing to both ongoing and new research efforts in areas including: 4D World Models Physical and Embodied Intelligence Multi‑Modal Learning As an AI Resident, you will collaborate closely with researchers and engineers at TRI on high‑risk, high‑impact research that pushes our understanding of spatio‑temporal reasoning and zero‑shot generalization. This is a research‑focused position, targeting the development of methods and techniques that can solve real‑world problems. We welcome you to join a positive, friendly, and enthusiastic team of researchers, where you will contribute to helping people gain and maintain independence, access, and mobility. We work closely with other Toyota affiliates, and actively collaborate towards research publications and the productization of our developed technologies. Responsibilities Develop, integrate, and deploy algorithms for Multi‑Modal and 4D reasoning targeting physical applications. Handle the ingestion of large‑scale datasets for training, including streaming, online, and continual learning. Contribute innovative solutions at the intersection of machine learning, computer vision, and robotics to improve real‑world task performance. Work closely with robotics and machine learning researchers and engineers to understand theoretical and practical needs. Follow best practices producing maintainable code, both for internal use as well as for open‑sourcing to the scientific community. Contribute to research publications and technical reports. Qualifications Bachelor’s or Master’s degree in Computer Science, Electrical Engineering, Robotics, or a related technical field. Exceptional candidates with equivalent research experience (e.g., strong publication record, open‑source contributions, or industry research experience) are encouraged to apply. Strong background in computer vision and its applications to robotics and embodied systems. Demonstrated research experience through publications, technical projects, or open‑source contributions. Strong communication skills and a collaborative mindset, with the ability to learn quickly and contribute to team research efforts. Passionate about assisting and amplifying older adults and those in need through dexterous manipulation, human‑robot collaboration, and physical assistance innovation. Bonus Qualifications Spatio‑temporal (4D) computer vision, including multi‑view geometry, 3D/4D reconstruction, video generation, self‑supervised learning, occlusion reasoning, etc. Large‑scale training of multi‑modal deep learning methods, both in terms of dataset sizes and model complexity, context length extension, and efficient attention, distributed computing, etc. Application of machine learning and computer vision to embodied applications. The pay range for this position at commencement of employment is expected to be between $45 and $60/hour for California‑based roles. Base pay offered will depend on multiple individualized factors, including, but not limited to, a candidate's experience, skills, job‑related knowledge, and market location. TRI offers a generous benefits package including medical, dental, and vision insurance, and paid time off benefits (including holiday pay and sick time). Additional details regarding these benefit plans will be provided if an employee receives an offer of employment. Please reference the Candidate Privacy Notice to inform you of the categories of personal information that we collect from individuals who inquire about and/or apply to work for Toyota Research Institute, Inc. or its subsidiaries, including Toyota A.I. Ventures GP, L.P., and the purposes for which we use such personal information. TRI is fueled by a diverse and inclusive community of people with unique backgrounds, education and life experiences. We are dedicated to fostering an innovative and collaborative environment by living the values that are an essential part of our culture. We believe diversity makes us stronger and are proud to provide Equal Employment Opportunity for all, without regard to an applicant’s race, color, creed, gender, gender identity or expression, sexual orientation, national origin, age, physical or mental disability, medical condition, religion, marital status, genetic information, veteran status, or any other status protected under federal, state or local laws. It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability. Pursuant to the San Francisco Fair Chance Ordinance, we will consider qualified applicants with arrest and conviction records for employment. #J-18808-Ljbffr

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the AI Resident - Learning From Videos (LFV) in Los Altos, CA vacancy
  •  ...AI Residency At Toyota Research Institute At Toyota Research Institute (TRI), we're on a mission to improve the quality of...  ...robotics, driving, and material sciences. The Team The Learning From Videos (LFV) team in the Robotics division focuses on the development... 
    Resident
    Video
    Shift work

    Toyota Research Institute

    Los Altos, CA
    3 days ago
  •  ...targeting physical applications. Handle the ingestion of large-scale datasets for training, including streaming, online, and continual learning. Contribute innovative solutions at the intersection of machine learning, computer vision, and robotics to improve real-world task... 
    Resident
    Video

    Jobtailor

    Los Altos, CA
    4 days ago
  • $180k - $258.75k

     ...transformative shift in mobility, we’ve built a world-class team advancing the state of the art in AI, robotics, driving, and material sciences.The Learning From Videos (LFV) team develops world foundation models that leverage large-scale multi-modal data (RGB, depth,... 
    Video
    Full time
    Local area
    Shift work

    Toyota Research Institute

    Los Altos, CA
    8 hours ago
  • Do you want to be part of the AI revolution? Do you want to think out of the box, thriving...  ...are looking for a world-class Machine Learning System Architect (HW) to join our SoC team...  ..., and power optimizationsFamiliarity with video, DSP, Ethernet, and PCIeMS or PhD in Electrical... 
    Video
    Work at office

    Baidu

    Sunnyvale, CA
    8 hours ago
  • $175k - $275k

     ...About Abaka AI   Abaka AI is built on one mission: to be the world’s most trusted...  ...catalog of off-the-shelf datasets (image, video, multimodal, reasoning, 3D, and beyond) as...  ...Role   We’re hiring our first Machine Learning Engineer in the United States, a foundational... 
    Video
    Full time
    Immediate start
    Flexible hours

    Abaka Ai

    Palo Alto, CA
    1 day ago
  • $45 - $60 per hour

    Toyota Research Institute is seeking an AI Resident to join its Learning From Videos team in Los Altos, California. This year-long position is geared towards early-career researchers excited about tackling ambitious problems in embodied AI. Responsibilities include developing... 
    Resident
    Video
    Hourly pay

    Toyota Research Institute

    Los Altos, CA
    5 days ago
  •  ...Creatify is building the world’s first end-to-end AI advertising agent—a platform that automates the entire video ad lifecycle, from scripting and avatar-led...  ...AI. About this role We’re hiring a Machine Learning Engineer to design and scale advanced models and... 
    Video
    Full time

    Creatify Lab

    Mountain View, CA
    1 day ago
  • $229k - $343k

     ...to express themselves, live in the moment, learn about the world, and have fun together.The...  ...ML Platform team builds cutting-edge AI technologies that power creative, scalable...  ...Snapchatters worldwide. From multimodal LLMs and video generation to real-time AR, human... 
    Video
    Full time
    Live in
    Work at office
    Local area
    Worldwide

    Snap

    Palo Alto, CA
    4 hours ago
  •  ...research, nurture the next generation of AI builders, and drive transformative contributions...  ...for high-performance computing in deep learning, driving impactful discoveries that...  ...environments.   ~ Experience with large-scale video or multimodal data pipelines.   ~... 
    Video
    Visa sponsorship

    Institute of Foundation Models

    Sunnyvale, CA
    more than 2 months ago
  •  ..., fearless and high-impact talent who see AI as a teammate – leveraging it to move faster...  ...debate, experimentation and continuous learning, and we seek out people with different...  ...includes the following steps: 1. Video Introduction: Submit a brief video introducing... 
    Video
    Full time
    Flexible hours

    Gen Digital Inc.

    Mountain View, CA
    2 days ago
  • Matroid is seeking a world-class Deep Learning Software Engineer to advance our computer vision and DL platform at our Palo Alto office...  ...teams to push state-of-the-art models, including multimodal LLMs, video transformers, and edge devices, while maintaining high-quality... 
    Video
    Work at office

    Matroid

    Palo Alto, CA
    4 days ago
  • # Machine Learning Engineer, Multimodal AIGoogle DeepMind## Job Description### About Google DeepMindGoogle...  ...and technological challenges through AI.### Responsibilities- Build advanced...  ...of understanding vision, audio, text, and video.- Train, fine-tune, and optimize state-of-... 
    Video
    Flexible hours

    AI Breaking Wire

    Mountain View, CA
    4 days ago
  • $140k - $230k

     ..., a test course for mobility; and Cloud & AI, the digital infrastructure powering our collaborative...  ...tackles the core challenges of machine learning for 3D perception, sensor fusion, and...  ...Hands-on experience with world models, video prediction, or latent dynamics models for... 
    Video
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    Woven by Toyota

    Palo Alto, CA
    2 days ago
  • $122k - $179k

     ...CoreWeave is The Essential Cloud for AI™. Built for pioneers by...  ...(Nasdaq: CRWV) in March 2025. Learn more at What You'll Do CoreWeave...  ...in enterprise access control, video surveillance systems, and the...  ..., (ii) U.S. lawful permanent resident (green card holder), (iii)... 
    Resident
    Video
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    8 days ago
  •  ...PyTorch infrastructure for the multimodal video, training, and RL pipelines that power frontier...  ...systems / ML infra. About Orbifold AI Orbifold AI is building the foundational...  ...evaluation, model training, reinforcement learning, and the multimodal data systems that fuel... 
    Video
    Full time

    Orbifold AI

    Palo Alto, CA
    13 days ago
  • Direct message the job poster from GMI Cloud Focusing on LLM, AI Infrastructure, and AIGC. Opportunities across Silicon...  ...research institutions, and large enterprises worldwide. Machine Learning Engineer - Video Generation We are seeking a highly skilled Machine Learning... 
    Video
    Full time
    Worldwide

    GMI Cloud

    Mountain View, CA
    2 days ago
  •  ...Description: We are looking for a Machine Learning Engineer to join our core research and...  ...hand motion from egocentric (first-person) video. Human demonstration data is the fuel for...  ..., or demonstrated impact in applied ML or AI systems. What We Offer ~ Competitive... 
    Video
    Full time

    Maxinsights Corporation

    Santa Clara, CA
    1 day ago
  • $150.4k - $277.6k

    Applied Machine Learning Research Engineer - Multimodal for Human Understanding Sunnyvale, California...  ..., United States Machine Learning and AI We’re starting to see the incredible...  ...Research Engineer to join our team in the Video Computer Vision group and help us push the... 
    Video
    Worldwide
    Relocation

    Apple Inc.

    Sunnyvale, CA
    5 days ago
  • $232k

    Uber AI Solutions (UAIS) is a startup inside Uber, building the data and evaluation infrastructure...  ...as good as the data and feedback they learn from, and that is the work we do. We are...  ...every modality, from text, images, and video to audio and physical AI and AVs, and... 
    Video
    Full time
    Work experience placement
    Work at office
    Remote work

    Jobzhr

    Sunnyvale, CA
    4 days ago
  • $281k - $356k

     ...current solutions with future innovations. You'll build active learning and ML-aided labeling workflows to tackle rare, "longtail" issues...  ...Experience developing and productizing large-scale vision, video, or multi-modal foundation models. Familiarity with end-to-end... 
    Video
    Full time
    Temporary work
    Remote work

    Waymo

    Mountain View, CA
    1 day ago
  • $184k - $287.5k

    We are seeking a Senior Machine Learning Engineer to join our end‑to‑end autonomous driving team...  ...tapping into the unlimited potential of AI to define the next era of computing. An...  ...maintaining high‑quality multimodal datasets (e.g., video, sensor, language/action traces) tailored... 
    Video
    Full time

    Nvidia

    Santa Clara, CA
    8 hours ago
  • $111.07k - $166.4k

     ...Across enterprise, cloud and AI, and carrier architectures, our...  ...Marvell is a place to thrive, learn, and lead. Your Team, Your ImpactThe...  ...driven by AI, cloud services, video streaming megatrends requires...  ...S. citizens, lawful permanent residents, or protected individuals as... 
    Resident
    Video
    Permanent employment
    Full time
    Internship
    Work from home

    Marvell

    Santa Clara, CA
    3 days ago
  •  ...Description: We are seeking a highly motivated Machine Learning Engineer to join our core research and development team, focused on video understanding and segmentation. In this role,...  ...downstream training pipelines for embodied AI and robot learning. Evaluate and benchmark... 
    Video
    Full time

    Maxinsights Corporation

    Santa Clara, CA
    1 day ago
  • $108.22k - $162.1k

     ...Across enterprise, cloud and AI, and carrier architectures, our...  ...Marvell is a place to thrive, learn, and lead. Your Team, Your ImpactThe...  ...megatrends of cloud services, video streaming, 5G wireless and AI/...  ...S. citizens, lawful permanent residents, or protected individuals as... 
    Resident
    Video
    Permanent employment
    Full time
    Internship
    Work from home

    Marvell

    Santa Clara, CA
    2 days ago
  • $95k - $155k

     ...the team. The Instructional Designer creates high-impact, AI-optimized, modular learning experiences that help LinkedIn customers adopt new behaviors...  ...solutions. Script, record, and edit engaging training videos to support product adoption and customer enablement.  Partner... 
    Video
    For contractors
    Work experience placement
    Work at office
    Flexible hours

    Linkedin

    Sunnyvale, CA
    2 days ago
  • $165k - $206.5k

     ...business workflows with enterprise AI. We help companies thrive in...  ...evolve these tactics as you learn what resonates. What matters...  ...technical blog posts, webinars, demo videos, case studies). Work with...  .... If you are a California-resident, please read our California Applicant... 
    Resident
    Video
    Live in
    Work at office
    Shift work
    3 days per week

    Box

    Redwood City, CA
    2 days ago
  • At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own...  ...or ML Systems Engineer to make our robot- learning pipeline reliable, reproducible, and...  ...robot post-training, reinforcement learning, video models, multimodal models, or learned control... 
    Video

    Rhoda AI

    Mountain View, CA
    4 days ago
  • At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own...  ...knowledge to adapt our web-pretrained video model to real robot tasks. Post-training at...  ...policy performance beyond what imitation learning alone achieves — reward design, online data... 
    Video
    Shift work

    Rhoda AI

    Mountain View, CA
    22 hours ago
  • $174.72k - $295.68k

     ...forefront of innovation, integrating advanced AI and autonomous driving technologies into...  ...through cutting-edge R&D in AI, machine learning, and smart connectivity. We are seeking Machine...  ...diffusion and flow-matching models, video tokenizers, and transformer-based multimodal... 
    Video
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  •  ...officeChance to be perm?: YesPerformance Expectations: Technical Hiring Criteria (Must Haves) • Top 3 Required skills: Machine Learning, AI,Python• Years of experience in each of the must-have skills: 7+ Years • Any Certifications required: NoAny additional information... 
    Hourly pay
    Permanent employment
    Work at office
    Remote work
    3 days per week

    Spruce Infotech

    Sunnyvale, CA
    8 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Resident - Learning From Videos (LFV). Be the first to apply!