AI Resident - Learning From Videos (LFV)
Toyota Research Institute
AI Residency At Toyota Research Institute
At Toyota Research Institute (TRI), we're on a mission to improve the quality of human life. We're developing new tools and capabilities to amplify the human experience. To lead this transformative shift in mobility, we've built a world-class team advancing the state of the art in AI, robotics, driving, and material sciences.
The Team
The Learning From Videos (LFV) team in the Robotics division focuses on the development of foundation models capable of leveraging large-scale multi-modal (RGB, depth, flow, semantics, bounding boxes, tactile, audio, etc) data from multiple domains (driving, robotics, indoors, outdoors, etc) to improve downstream task performance.
Our approach emphasizes training scalability: by learning from multiple modalities, models can develop useful data-driven priors about 3D geometry, physics, and dynamics for world understanding.
Our research interests include, but are not limited to:
- Video Generation
- World Models
- 4D Reconstruction
- Multi-Modal Models
- Multi-View Geometry
- Data Augmentation
- Video-Language-Action Models
We focus primarily on embodied applications and aim to tackle some of the hardest scientific challenges in spatio-temporal reasoning, enabling autonomous agents to operate in real-world, unstructured environments.
The AI Resident
This year-long AI Residency is a research-focused position designed for early-career researchers and engineers who are excited to work on ambitious problems in embodied AI. The resident will be deeply integrated into the LFV team, contributing to both ongoing and new research efforts in areas including:
- 4D World Models
- Physical and Embodied Intelligence
- Multi-Modal Learning
As an AI Resident, you will collaborate closely with researchers and engineers at TRI on high-risk, pushing forward our understanding of spatio-temporal reasoning and zero-shot generalization. This is a research-focused position, targeting the development of methods and techniques that can solve real-world problems.
We welcome you to join a positive, friendly, and enthusiastic team of researchers, where you will contribute to helping people gain and maintain independence, access, and mobility. We work closely with other Toyota affiliates, and actively collaborate towards research publications and the productization of our developed technologies.
Responsibilities
- Develop, integrate, and deploy algorithms for Multi-Modal and 4D reasoning targeting physical applications.
- Handle the ingestion of large-scale datasets for training, including streaming, online, and continual learning.
- Contribute innovative solutions at the intersection of machine learning, computer vision, and robotics to improve real-world task performance.
- Work closely with robotics and machine learning researchers and engineers to understand theoretical and practical needs.
- Follow best practices producing maintainable code, both for internal use as well as for open-sourcing to the scientific community.
- Contribute to research publications and technical reports.
Qualifications
- Bachelor's or Master's degree in Computer Science, Electrical Engineering, Robotics, or a related technical field.
- Exceptional candidates with equivalent research experience (e.g., strong publication record, open-source contributions, or industry research experience) are encouraged to apply.
- Strong background in computer vision and its applications to robotics and embodied systems.
- Demonstrated research experience through publications, technical projects, or open-source contributions.
- Strong communication skills and a collaborative mindset, with the ability to learn quickly and contribute to team research efforts.
- Passionate about assisting and amplifying older adults and those in need through dexterous manipulation, human-robot collaboration, and physical assistance innovation.
Bonus Qualifications
- Spatio-temporal (4D) computer vision, including multi-view geometry, 3D/4D reconstruction, video generation, self-supervised learning, occlusion reasoning, etc.
- Large-scale training of multi-modal deep learning methods, both in terms of dataset sizes and model complexity, context length extension, and efficient attention, distributed computing, etc.
- Application of machine learning and computer vision to embodied applications.
$45 - $60 per hour
...class team advancing the state of the art in AI, robotics, driving, and material sciences. The Team The Learning From Videos (LFV) team in the Robotics division focuses on... ...Intelligence Multi‑Modal Learning The AI Resident This year‑long AI Residency is a research-focused...ResidentVideoHourly payLocal areaShift work$180k - $258.75k
...transformative shift in mobility, we’ve built a world-class team advancing the state of the art in AI, robotics, driving, and material sciences. The Learning From Videos (LFV) team develops world foundation models that leverage large-scale multi-modal data (RGB, depth,...VideoLocal areaShift work$150k - $180k
...Overview Job Title: Machine Learning Infrastructure Engineer – Computer Vision & AI Employer: The Mice Groups, Inc. Location Redwood City, CA (... ...skills Nice to Have Experience working in real-time video processing environments Familiarity with cloud-...VideoFull timeContract work- ...business workflows with enterprise AI. We help companies thrive in... ...evolve these tactics as you learn what resonates. What matters... ...technical blog posts, webinars, demo videos, case studies). Work with... .... If you are a California-resident, please read our California Applicant...ResidentVideoLive inWork at officeShift work3 days per week
- ...research, nurture the next generation of AI builders, and drive transformative contributions... ...for high-performance computing in deep learning, driving impactful discoveries that... ...environments. ~ Experience with large-scale video or multimodal data pipelines. ~...VideoVisa sponsorship
$204k - $259k
...Get AI-powered advice on this job and more exclusive features. Waymo is an autonomous driving technology company with the mission... ...data. In This Role, You Will Design, train, and deploy machine learning models to automate the creation of Waymo’s HD maps, unlocking scale...Full timeRemote work$152k - $241.5k
...inspirational demos, and sharing that code with the AI community. Our team is deeply committed to... ...aspire to this calling, we would love to learn more about you! What you'll be doing:... ...(e.g., demos, blogs, presentations, videos, etc.) and engage with developers,...VideoRemote work$34 per hour
...'ll be working directly with cutting-edge AI systems - evaluating outputs, identifying... ...), all new employees must complete a live video verification with their selected IDs and provide... ...solutions for NLP-enabled machine learning by blending technology and human intelligence...VideoFull timeContract workRemote workVisa sponsorship$26 - $28 per hour
...Labeling Analysts, supporting speech and voice AI systems. This is a high-impact... ...), all new employees must complete a live video verification with their selected IDs and provide... ...solutions for NLP-enabled machine learning by blending technology and human intelligence...VideoFull timeWork experience placementRemote workVisa sponsorship- ...knowledge is required to use cutting-edge AI techniques to solve real-world problems. At... ...and compliance. We're looking for a Deep Learning Field Engineer to operate at the forefront... ...execution systems, safety alert systems, and video management systems. Perform quantitative...VideoWork experience placementWork at officeFlexible hours
$117.2k - $313.7k
...About the Role Salesforce AI Research is seeking outstanding AI Research Scientists / Research... ...– LLM‑powered agents, reinforcement learning, autonomous workflows Multimodal & Computer Vision – Vision‑language models, video understanding, visual grounding for GUI agents...VideoFull time$117.2k - $313.7k
...Salesforce AI Research is looking for outstanding AI Research Scientists and Research... ...Reasoning: LLM‑powered agents, reinforcement learning (RL), reasoning and planning, autonomous... ...Computer Vision: Vision‑language models (VLM), video understanding, visual grounding for agents...Video- ...Research Engineer – ML Track Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and... ...Engineer – ML track, you'll build and optimise the large-scale learning systems that power our open-weight models. Working hand-in-hand...Relocation package
$188.5k - $282.7k
...About the Team & Role: We're building SAGE , Rubrik's Semantic AI Governance Engine, which is the first system designed to monitor... ...Education: A Bachelor's degree (or higher) in Computer Science, Machine Learning, Computer Engineering, Statistics, or a closely related...Permanent employmentLocal area$110k - $160k
...Abaka AI is built on one mission: to be the world’s most trusted data partner for AI companies... ...catalog of off-the-shelf datasets (image, video, multimodal, reasoning, 3D, and beyond) as... ..., or dataset preparation for machine learning. Experience working with international...VideoFlexible hours- ...Senior Machine Learning Engineer, Recommendation & AI Applications Mountain View, California, United States About NewsBreak Founded in 2015, NewsBreak is the Content Intelligence platform shaping the future content economy. With over 40 million monthly active users, our...Full timeLocal areaWork from home
- ...Data Infrastructure organization builds AI-native analytics platforms and first-party... ...of a little chaos, and we’re constantly learning. Our team cares deeply about how we build... ...or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under...ResidentPermanent employmentTemporary workCasual workWork at officeFlexible hours
$22.3 - $25.73 per hour
...Glassdoor's 100 Best Companies to work for. Learn from the best and accelerate your career... .... To learn more, watch the following videos from our employees: What is a... ...How do Vi Servers interact with our senior residents? Qualifications Qualified applicants...ResidentVideo- ...The Mission: As a Senior Machine Learning Engineer, you will be responsible for building machine... ...that deliver the power of Generative AI to our customers. You will work closely with... ...content types including text, image, audio and video enabling us to generate brand-affinitized,...VideoLocal area
- ...Development and Implementation: Develop and implement scalable machine learning models focusing on advanced ranking and recommendation systems,... ...projects. A fervent interest in exploring and applying AI and ML technologies. Strive to solve sophisticated engineering problems...Work at office
- ...Senior Machine Learning Engineer, Agentic Application Full-time Employee Type: Regular Region... ...Who we are Moveworks is the Agentic AI Assistant platform that empowers the entire... ...confirm the distance between your primary residence and the closest ServiceNow office using a...Full timeWork at officeRemote workFlexible hours
- ...The Opportunity We are seeking a Senior Software Engineer to join our Machine Learning team. The ideal candidate will have a strong background in machine learning, NLP, and generative AI technologies, with substantial experience in developing and deploying ML-driven products...Flexible hours3 days per week
$175k - $230k
...far more constrained. We’re changing that. Atoms builds Physical AI— real-world robots for the industries that move civilization forward... ...in a lab. We deploy them into real environments, operate them, learn from them, and improve them until they work at scale. We are...Full timeTemporary workWork at officeFlexible hours- ...team: Mind Robotics is building robots that learn from real-world experience — and the... ...into the high-quality data that grounds our AI models in physical reality. Both flow through... ...Hands‑on experience with sensor data (video, depth, IMU, force/torque) and the infrastructure...VideoImmediate startShift work
$120k - $250k
Get AI-powered advice on this job and more exclusive features. This range is provided... ...experience — talk with your recruiter to learn more. Base pay range $120,000.00/yr - $25... ...-scale multimodal datasets (images, text, video, structured data). Implement and refine MoE...VideoFull timeInternshipFlexible hours$169.8k - $233.5k
...Sr Ai Engineer Uniphore is one of the largest B2B AI-native companies—decades-proven... ...better than anyone how to capture voice, video and text and how to analyze all types of data... ...and end users by utilizing machine learning (ML) and Generative AI techniques. The...Video- ...Empowered with innovative tools, continuous learning and a global community of diverse talent... ...positively impacts the team in which one resides. Manages small teams and/or work... ...advanced OneStream capabilities; Sensible AI, CPM Express etc. You have experience...ResidentWork experience placementLive inWork at officeLocal area
$143.1k - $264.2k
...extraordinary team of Computer Graphics, Computer Vision and Deep Learning researchers and engineers to discover and build solutions to... ...applications impacting millions of users. Description Apple's Video Computer Vision (VCV) Face and Body technologies team is...VideoFull timeRelocation- ...infrastructure layer for the next generation of AI. We believe open‑source models are the... ...content including blog posts, tutorials, videos, sample applications, demos, and... ...expected. We stay persistent, adapt quickly, and learn as we go. We take setbacks seriously, but...Video
- ...government and security. We are looking for a world-class Deep Learning Software Engineer who is excited to operate at the forefront of... ...are excited to adapt the latest multimodal LLMs, or implement a video transformer model from scratch, or get realtime segmentation...VideoWork experience placementWork at officeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Resident - Learning From Videos (LFV). Be the first to apply!

