Research Member of Technical Staff- Video Generation Modeling
Rhoda AI
At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots. Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design. We've raised over $450M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality. We're looking for Research Scientists and Research Engineers to push the frontier of large-scale pre-training for our video action model. Our approach formulates robot control as video prediction — we pre-train causal video generation models on web-scale video data, then adapt them to predict robot actions from real-world demonstrations. You'll work on the core architectures, training objectives, and scaling strategies that determine how well our models learn from internet-scale video. We hire across levels — from senior to staff — and welcome both research-track and engineering-track candidates. What You'll Do Design and train large-scale causal video generation models on web-scale video data Develop and validate training objectives, model architectures, and data mixtures for video prediction at scale Research scaling laws and data efficiency for web-scale video pretraining Investigate what properties of web video transfer most effectively to robotic control and action prediction Build systematic evaluations to measure video generation quality, long-horizon prediction fidelity, and downstream robot task performance Run rigorous ablations and benchmarking to understand what drives model quality at scale Collaborate closely with data & evaluation, post-training, and training systems teams to translate research ideas into working systems Publish and present work at top-tier ML and robotics venues (especially valued for RS track) What We're Looking For Strong background in large-scale generative modeling — either video generation (autoregressive video models, diffusion transformers, causal video architectures) or language model pretraining (LLMs, autoregressive transformers at scale) Hands-on experience training large generative models from scratch at scale Deep understanding of autoregressive modeling, causal architectures, and scaling behavior Fluency with modern ML frameworks (PyTorch required; JAX a plus) Ability to design experiments, interpret results, and iterate quickly Strong research taste: ability to identify high-leverage questions and cut through noise Comfort operating in a fast-moving, ambiguous startup environment Staff-level candidates are expected to define technical direction and drive research strategy independently; senior/MTS candidates execute complex projects with strong fundamentals and growing scope Nice to Have (But Not Required) PhD in ML, CS, Robotics, or a related field — or equivalent research/industry experience Strong publication record at NeurIPS, ICML, ICLR, CVPR, CoRL, etc. (especially valued for RS track) Prior work specifically on video generation models (autoregressive video, diffusion transformers, world models, or causal video architectures) Experience with large-scale autoregressive language model pretraining and scaling Familiarity with web-scale video datasets and video data curation pipelines Prior work connecting video generation to control, action prediction, or robotic learning Familiarity with distributed training and multi-node infrastructure Why This Role Work on a fundamentally different approach to robot learning — web-scale video pretraining rather than robot-data-only VLA models Your models give our robots the ability to understand and predict the visual world from internet-scale supervision Direct collaboration with data, post-training, and deployment teams with no silos High ownership and fast iteration in a small, elite team #J-18808-Ljbffr Rhoda AI
- ...AI, we’re building the next generation of generalist intelligent robots... ...-of-the-art foundation world models that control our robots. Our... ...possible by our cutting edge research and end-to-end system design.... ...Experience with efficient video or multimodal model architectures...Video
- Member of Technical Staff Physical AI (Robotics / World Models) Palo Alto, CA About Orbifold AI Orbifold AI... ...and world model research teams on the field's hardest... ...how multimodal data (video, image, audio, sensor,... ...multimodal reasoning, video generation models, etc.) and RL...VideoShift work
- Member of Technical Staff — Diffusion Model About the Role RadixArk is seeking a Member of Technical... ...advance the frontier of generative modeling. You will work... ...-based models for image, video, and multimodal... ...This role combines deep research thinking with strong engineering...VideoFlexible hours
- ...able to concisely and accurately share knowledge with their teammates. Description Image and video generation/editing. Candidate Criteria People worked on image or video models before. Also games, world models, simulation is good for this category. xAI is an equal...Video
$180k
...About the Role As a multimodal engineer on the Imagine Model Team, you will develop cutting‑edge AI experiences beyond... ...strong focus on enabling high‑fidelity understanding and generation across image and video modalities, while also incorporating audio where it enhances...VideoTemporary work- ...we’re building the next generation of generalist... ...the-art foundation world models that control our robots... ...possible by our cutting edge research and end-to-end system design... .... We're looking for a Staff / Principal ML Training... ...training (vision, video, proprioception, actions...Video
- ...AI, we’re building the next generation of generalist intelligent robots... ...-of-the-art foundation world models that control our robots. Our... ...possible by our cutting edge research and end-to-end system design.... ...pipelines (LLMs, VLMs, or video models) Understanding of parallelism...Video
- ...pipelines for web-scale video pretraining data:... ...causal video model capabilities: prediction... ...task performance Research and implement data... ...improve video generation quality and... ..., and actionable Staff-level candidates are... ...expected to define technical direction and drive...Video
- ...building the next generation of generalist intelligent... ...foundation world models that control our... ...our cutting edge research and end-to-end... ...— from senior to staff. What You'll Do Architect... ...billions of video clips with strong... ...expected to define technical direction and own...VideoImmediate start
- ...building the next generation of generalist intelligent... ...foundation world models that control our... ...our cutting edge research and end-to-end... ...our web-pretrained video model to real robot... ...— from senior to staff. What You'll Do... ...expected to define technical direction and drive...VideoShift work
$202.35k - $303.05k
...In this role, you will be a key member of the behavior team focused on... ...driving. You will be focusing on researching and developing state of the art generative models, with an emphasis on diffusion models... .... However, experiences in video generation, text-to-image generation...Video- Rhoda AI is seeking Research Scientists and Engineers in Palo Alto to develop large-scale causal video generation models for robotic control. The role involves designing training objectives and researching scaling laws for video pretraining. Candidates should have a strong...Video
- ...general-purpose world models—a new form of audio-... ...will power the next generation of gaming, education... ...for a deeply technical and creative researcher who thrives on invention... ...pioneer interactive video models capable of real... .... Who you are A staff-level or senior researcher...Video
- ...Odyssey is an AI lab pioneering general world models: causal, multimodal systems that learn to... ...’ve now brought together a world‑class research team from DeepMind, Tesla, Waymo, Meta,... ...to language models (DeepMind Gemini), video models (DeepMind Veo), world models (Wayve...VideoRemote workFlexible hours
$180k
...aims to provide an omni model that can understand the... ...through text, image, video and audio. To accomplish... ...through synthetic data generation. Optimize mid‑training... ...interview”) during which a member of our team will ask... ...which consists of four technical interviews: Coding...VideoTemporary workRelocation$180k
Member of Technical Staff - Imagine Safety ABOUT xAI xAI’s mission is to... ...ensure Grok’s multimodal generation capabilities (images, video, audio, and beyond) are... ...in collaboration with researchers, product, and policy... ...between user interactions, model outputs, and training...VideoTemporary workWorldwide$180k
...cutting-edge AI to enable seamless generation, processing, and delivery of images, video, audio, and beyond. Your work... ...millions, turning advanced multimodal models into production-grade features.... ...with frontend engineers, AI researchers, and product teams to deliver...VideoTemporary workWorldwide$180k
Member of Technical Staff - Multimodal Understanding About xAI xAI’s mission... ...understanding and generation across modalities—image, video, audio, and text—spanning... ...reasoning, world modeling, tool use, agentic behaviors... ...art performance. Build research tooling, user‑friendly...VideoTemporary work$148.5k - $223.9k
Senior Member of Technical Staff - AI ResearchSkip to main content#Senior Member of Technical Staff - AI Research page is loaded## Senior Member of Technical... ...code (Human or AI-generated) for correctness, quality... ...implementing and debugging model training, evaluation, and...Work at office- ...stack and using it to transform the video product experience - video playback... .... What You'll Do Build the next generation of large-scale video services Contribute... ...phone interview, during which a member of our team will ask technical questions. If you clear the phone...VideoTemporary workH1bRelocationWork visa
- ...Poke.com is seeking a Member of Technical Staff to join a fast-growing, early-stage team and help build... ...top engineers from quant finance and research labs to scale impact quickly. What you... ...Strava, and Ray-Ban. Optimize high-cost models from providers like Anthropic and...Monday to Friday
$148.5k - $223.9k
...standardizing AI systems. Critically evaluate code (human or AI-generated) for correctness, quality, security, and performance. Strong... ...learning frameworks and strong ML fundamentals; experience debugging model training, evaluation, and inference pipelines. Infrastructure...$324k - $396k
Pantera Capital is seeking a Member of Technical Staff in Palo Alto, CA. This role involves introducing innovative techniques in AI, stabilizing large language model training, and working on hands-on technical problems. Candidates should have relevant experience and a...Remote job- About the Role As a Member of Technical Staff [Research] at NeoCognition , you’ll be part of the core team advancing the frontier of LLM agents — systems... ...execute experiments, benchmark performance, and analyze model behaviors to identify failure modes and opportunities....
$150k
...their teammates. ABOUT THE ROLE: You will join the Grok Voice Model team to help build the world’s best voice AI. We deliver smooth... ...including collection of diverse real‑world audio, synthetic data generation, and automated annotation workflows to enable high‑quality...Temporary work$200k - $300k
...Department AI Perplexity is seeking top-tier AI Research Scientists and Engineers to advance our... ...and agent experiences through our Sonar models, Deep Research Agent, Comet Agent, and... ...Research Team (Horizontal) Focus on generating and improving base models that power all...Full time$160.36k - $240.54k
...flexible, partner-led business model, Nuro is working toward a... ...Develop and scale state-of-the-art generative models—especially diffusion... ..., in industry, or both. Research experiences in generative models... ...model optimization, video generation, text-to-image generation...VideoImmediate startFlexible hours$165.76k - $207.2k
...for a data-driven robotics researcher with a strong background... ...experienced existing faculty member looking to take a year-... ...using new capabilities in generative AI (e.g., recent results in... ...training large-scale foundation models (VLMs, text-to-video models, etc) Sufficient...Video- Job Description Prime Video offers customers a vast... ...thousands of devices. Prime members can customize their... ...apply state‑of‑the‑art generative AI, graphics, audio, and computer vision research to video‑centric digital... ...machine learning models for business applications...Video
$151.8k - $218.21k
...capabilities development, as well as research impact via open-source... ...large scale foundation models (VLMs, text-to-video models, etc). The ideal... ...we can create the next generation of AI-powered capable robots... ...and performance. Be a key member of the team and play a critical...Video
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Member of Technical Staff- Video Generation Modeling. Be the first to apply!
- research associate scientist Mountain View, CA
- research assistant linguistics Mountain View, CA
- art history research assistant Mountain View, CA
- research associate microbiology Mountain View, CA
- equity research associate Mountain View, CA
- social science research assistant Mountain View, CA
- computer science research assistant Mountain View, CA
- biochemistry research assistant Mountain View, CA
- research associate Mountain View, CA
- neuroscience research assistant Mountain View, CA

