AI Resident - Learning From Videos (LFV)
$45 - $60 per hourTri
At Toyota Research Institute (TRI), we’re on a mission to improve the quality of human life. We’re developing new tools and capabilities to amplify the human experience. To lead this transformative shift in mobility, we’ve built a world‑class team advancing the state of the art in AI, robotics, driving, and material sciences. The Team The Learning From Videos (LFV) team in the Robotics division focuses on the development of foundation models capable of leveraging large-scale multi-modal (RGB, depth, flow, semantics, bounding boxes, tactile, audio, etc) data from multiple domains (driving, robotics, indoors, outdoors, etc) to improve downstream task performance. Our approach emphasizes training scalability: by learning from multiple modalities, models can develop useful data-driven priors about 3D geometry, physics, and dynamics for world understanding. Research Interests Video Generation World Models 4D Reconstruction Multi-Modal Models Multi-View Geometry Data Augmentation Video‑Language‑Action Models Areas of Focus 4D World Models Physical and Embodied Intelligence Multi‑Modal Learning The AI Resident This year‑long AI Residency is a research-focused position designed for early‑career researchers and engineers who are excited to work on ambitious problems in embodied AI. The resident will be deeply integrated into the LFV team, contributing to both ongoing and new research efforts in areas including: Responsibilities Develop, integrate, and deploy algorithms for Multi‑Modal and 4D reasoning targeting physical applications. Handle the ingestion of large-scale datasets for training, including streaming, online, and continual learning. Contribute innovative solutions at the intersection of machine learning, computer vision, and robotics to improve real‑world task performance. Work closely with robotics and machine learning researchers and engineers to understand theoretical and practical needs. Follow best practices producing maintainable code, both for internal use as well as for open‑sourcing to the scientific community. Contribute to research publications and technical reports. Qualifications Bachelor's or Master’s degree in Computer Science, Electrical Engineering, Robotics, or a related technical field. Exceptional candidates with equivalent research experience (e.g., strong publication record, open‑source contributions, or industry research experience) are encouraged to apply. Strong background in computer vision and its applications to robotics and embodied systems. Demonstrated research experience through publications, technical projects, or open‑source contributions. Strong communication skills and a collaborative mindset, with the ability to learn quickly and contribute to team research efforts. Passionate about assisting and amplifying older adults and those in need through dexterous manipulation, human‑robot collaboration, and physical assistance innovation. Spatio‑temporal (4D) computer vision, including multi‑view geometry, 3D/4D reconstruction, video generation, self‑supervised learning, occlusion reasoning, etc. Large‑scale training of multi‑modal deep learning methods, both in terms of dataset sizes and model complexity, context length extension, and efficient attention, distributed computing, etc. Application of machine learning and computer vision to embodied applications. Benefits and Compensation The pay range for this position at commencement of employment is expected to be between $45 and $60 per hour for California-based roles. Base pay offered will depend on multiple individualized factors, including, but not limited to, a candidate's experience, skills, job‑related knowledge, and market location. TRI offers a generous benefits package including medical, dental, and vision insurance, and paid time off benefits (including holiday pay and sick time). Additional details regarding these benefit plans will be provided if an employee receives an offer of employment. Please reference the Candidate Privacy Notice to inform you of the categories of personal information that we collect from individuals who inquire about and/or apply to work for Toyota Research Institute, Inc. or its subsidiaries, including Toyota A.I. Ventures GP, L.P., and the purposes for which we use such personal information. TRI is fueled by a diverse and inclusive community of people with unique backgrounds, education and life experiences. We are dedicated to fostering an innovative and collaborative environment by living the values that are an essential part of our culture. We believe diversity makes us stronger and are proud to provide Equal Employment Opportunity for all, without regard to an applicant’s race, color, creed, gender, gender identity or expression, sexual orientation, national origin, age, physical or mental disability, medical condition, religion, marital status, genetic information, veteran status, or any other status protected under federal, state or local laws. It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability. Pursuant to the San Francisco Fair Chance Ordinance, we will consider qualified applicants with arrest and conviction records for employment. #J-18808-Ljbffr Tri
- ...AI Residency At Toyota Research Institute At Toyota Research Institute (TRI), we're on a mission to improve the quality of... ...robotics, driving, and material sciences. The Team The Learning From Videos (LFV) team in the Robotics division focuses on the development...ResidentVideoShift work
$175k - $275k
...About Abaka AI Abaka AI is built on one mission: to be the world’s most trusted... ...catalog of off-the-shelf datasets (image, video, multimodal, reasoning, 3D, and beyond) as... ...Role We’re hiring our first Machine Learning Engineer in the United States, a foundational...VideoFull timeImmediate startFlexible hours- ...Creatify is building the world’s first end-to-end AI advertising agent—a platform that automates the entire video ad lifecycle, from scripting and avatar-led... ...AI. About this role We’re hiring a Machine Learning Engineer to design and scale advanced models and...VideoFull time
$180k - $258.75k
...transformative shift in mobility, we’ve built a world-class team advancing the state of the art in AI, robotics, driving, and material sciences. The Learning From Videos (LFV) team develops world foundation models that leverage large-scale multi-modal data (RGB, depth,...VideoLocal areaShift work$281k - $356k
...current solutions with future innovations. You'll build active learning and ML-aided labeling workflows to tackle rare, "longtail" issues... ...Experience developing and productizing large-scale vision, video, or multi-modal foundation models. Familiarity with end-to-end...VideoFull timeTemporary workRemote work$150k - $180k
...Overview Job Title: Machine Learning Infrastructure Engineer – Computer Vision & AI Employer: The Mice Groups, Inc. Location Redwood City, CA (... ...skills Nice to Have Experience working in real-time video processing environments Familiarity with cloud-...VideoFull timeContract work$165k - $206.5k
...business workflows with enterprise AI. We help companies thrive in... ...evolve these tactics as you learn what resonates. What matters... ...technical blog posts, webinars, demo videos, case studies). Work with... .... If you are a California‑resident, please read our California Applicant...ResidentVideoLive inWork at officeShift work3 days per week- ...research, nurture the next generation of AI builders, and drive transformative contributions... ...for high-performance computing in deep learning, driving impactful discoveries that... ...environments. ~ Experience with large-scale video or multimodal data pipelines. ~...VideoVisa sponsorship
$150k
...mandate is to advance research, nurture the next generation of AI builders, and drive transformative contributions to a knowledge-... ...establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation...Full timeWorldwideVisa sponsorship$150k
...mandate is to advance research, nurture the next generation of AI builders, and drive transformative contributions to a knowledge-... ...establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation...Full timeWork experience placementVisa sponsorship- ...About Mistral At Mistral AI, we believe in the power of AI to simplify tasks, save time, and enhance learning and creativity. Our technology is designed to integrate seamlessly into daily working life. We democratize AI through high-performance, optimized,...Full timeWork at officeVisa sponsorship
$300k
...mandate is to advance research, nurture the next generation of AI builders, and drive transformative contributions to a knowledge-... ...establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation...Full timeFlexible hours$204k - $259k
...Get AI-powered advice on this job and more exclusive features. Waymo is an autonomous driving technology company with the mission... ...data. In This Role, You Will Design, train, and deploy machine learning models to automate the creation of Waymo’s HD maps, unlocking scale...Full timeRemote work$152k - $241.5k
...inspirational demos, and sharing that code with the AI community. Our team is deeply committed to... ...aspire to this calling, we would love to learn more about you! What you'll be doing:... ...(e.g., demos, blogs, presentations, videos, etc.) and engage with developers,...VideoRemote work- ...long-term user value. You’ll work at the intersection of machine learning, product, and systems engineering: experimenting with modern... ..., backed by Andreessen Horowitz SR04 and top angels across the video AI space. We’re still early, and this is your chance to shape both...VideoFull time
$268.6k - $395k
...a Principal Engineer, you will lead the technical direction for AI-first experiences, including ranking and relevance systems that... ...using state-of-the-art techniques such as sequence modeling, deep learning, and large language models (LLMs). Your work will span query...Hourly payWork at officeLocal areaRemote workFlexible hours$26 - $28 per hour
...Labeling Analysts, supporting speech and voice AI systems. This is a high-impact... ...), all new employees must complete a live video verification with their selected IDs and provide... ...solutions for NLP-enabled machine learning by blending technology and human intelligence...VideoFull timeWork experience placementRemote workVisa sponsorship- ...Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable... ...aim high and stay humble in our pursuit of excellence, constantly learning and evolving as we pave the way for a smarter, safer future....Full timeWork at officeWork from home
$195k - $230k
...highly personalized local news and information powered by advanced AI, recommendation systems, and adtech. Recognized by Fast... ...visit About the Role We are looking for a Senior Machine Learning Engineer to help evolve our large-scale recommendation systems...Full timeLocal areaWork from home$137.1k - $201.6k
...their subscription. We are forming a new team that will leverage AI and advanced ML to power decision making in real-time – from... ...beyond. About the Role We’re looking for a Staff Machine Learning Engineer to drive the design and development of large-scale ML/...Hourly payWork at officeLocal areaRemote workFlexible hours- ...knowledge is required to use cutting-edge AI techniques to solve real-world problems. At... ...and compliance. We're looking for a Deep Learning Field Engineer to operate at the forefront... ...execution systems, safety alert systems, and video management systems. Perform quantitative...VideoWork experience placementWork at officeFlexible hours
$34 per hour
...'ll be working directly with cutting-edge AI systems evaluating outputs, identifying gaps... ...), all new employees must complete a live video verification with their selected IDs and... ...transformation solutions for NLP-enabled machine learning by blending technology and human...VideoFull timeContract workRemote workVisa sponsorship$117.2k - $313.7k
...Salesforce AI Research is looking for outstanding AI Research Scientists and Research... ...Reasoning: LLM‑powered agents, reinforcement learning (RL), reasoning and planning, autonomous... ...Computer Vision: Vision‑language models (VLM), video understanding, visual grounding for agents...Video$117.2k - $313.7k
...About the Role Salesforce AI Research is seeking outstanding AI Research Scientists / Research... ...– LLM‑powered agents, reinforcement learning, autonomous workflows Multimodal & Computer Vision – Vision‑language models, video understanding, visual grounding for GUI agents...VideoFull time- ...At Coram AI, we’re reimagining video security for the modern world. Our cloud-native platform uses computer vision and AI to help businesses stay... ...problems at the intersection of user experience, machine learning and infrastructure. It also means committing to excellence...VideoFull timeRemote work
- ...Research Engineer – ML Track Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and... ...Engineer – ML track, you'll build and optimise the large-scale learning systems that power our open-weight models. Working hand-in-hand...Relocation package
- ...AI Research Intern We're looking for an AI Research Intern to join our AI team and... ...Development Research and develop deep learning models in one or more of the following areas... ...priorities: Computer vision (e.g. video enhancement, super-resolution, restoration...VideoInternshipLocal areaRemote workWorldwideFlexible hours3 days per week
$110k - $160k
...Abaka AI is built on one mission: to be the world’s most trusted data partner for AI companies... ...catalog of off-the-shelf datasets (image, video, multimodal, reasoning, 3D, and beyond) as... ..., or dataset preparation for machine learning. Experience working with international...VideoFlexible hours$188.5k - $282.7k
...About the Team & Role: We're building SAGE , Rubrik's Semantic AI Governance Engine, which is the first system designed to monitor... ...Education: A Bachelor's degree (or higher) in Computer Science, Machine Learning, Computer Engineering, Statistics, or a closely related...Permanent employmentLocal area$120k - $250k
...Get AI-powered advice on this job and more exclusive features. This range is provided... ...experience — talk with your recruiter to learn more. Base pay range $120,000.00/yr - $250... ...-scale multimodal datasets (images, text, video, structured data). Implement and refine MoE...VideoFull timeInternshipFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Resident - Learning From Videos (LFV). Be the first to apply!


