Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Member of Technical Staff- Post-training & Robot Learning

Rhoda AI

At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots. Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design. We've raised over $400M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality. We're looking for Research Scientists and Research Engineers with deep robotics or autonomous systems domain knowledge to adapt our web-pretrained video model to real robot tasks. Post-training at Rhoda means taking a causal video generation model pretrained on internet-scale data and fine-tuning it on robot-collected demonstrations to produce reliable, generalizable behavior — with as little task-specific data as possible. We hire across levels — from senior to staff. What You'll Do Design and implement RL training pipelines to improve robot policy performance beyond what imitation learning alone achieves — reward design, online data collection, and policy optimization Develop and apply RL algorithms (PPO, GRPO, or similar) adapted to the video prediction setting, including reward modeling and feedback collection strategies for physical task performance Design and implement broader post-training pipelines: supervised fine-tuning, preference optimization, and behavioral alignment on robot-collected demonstration data Work on the inverse dynamics model that translates video predictions into executable robot actions Build evaluation frameworks for post-trained policies: task success, generalization to novel objects and environments, and failure mode analysis on real hardware Research methods to efficiently adapt models to new tasks with minimal demonstration data, including in-context generalization and few-shot adaptation Identify failure modes and systematic weaknesses in deployed robot policies and drive targeted improvements Iterate quickly between simulation and real robot evaluation to close the feedback loop Collaborate with the pre-training team to surface what capabilities are missing from the base model and need to be addressed upstream What We're Looking For Hands-on experience with robot systems, robotic policy learning, or autonomous systems in an industry or research setting (robotics, self-driving, or similar physical AI domains) Strong understanding of robot policy learning: imitation learning, behavior cloning, and how RL builds on top of it Practical familiarity with real robot hardware, deployment constraints, and sensor modalities (vision, proprioception) Solid ML skills with hands-on PyTorch experience Ability to diagnose policy failures, reason about distribution shift, and iterate effectively on data and training strategies Comfort with ambiguity and fast-changing research priorities Staff-level candidates are expected to define technical direction and drive research strategy independently; senior candidates execute complex projects with strong fundamentals and growing scope Nice to Have (But Not Required) Hands-on experience with reinforcement learning — reward design, policy optimization, and online RL training loops — applied to real or near-real environments (robotics, games, simulated physics, or similar); this is a significant plus Prior industry experience in robotics, autonomous driving, or physical AI (e.g., manipulation, mobile robotics, self-driving stacks) Experience with teleoperation systems or robot demonstration collection at scale Familiarity with robot middleware (ROS/ROS2) and real-time control systems Experience with simulation environments for robotics (MuJoCo, Isaac Sim, Genesis) Understanding of video generation models and how they connect to action prediction PhD in Robotics, ML, or a related field Publication record at ICRA, CoRL, RSS, NeurIPS, or related venues Why This Role Your work is what makes our robots actually perform tasks reliably in the real world — the direct connection between pre-trained capability and deployed behavior Work at a rare intersection: state-of-the-art video generation models applied to real robot hardware, not simulation Fast feedback loop between model changes and real robot performance High ownership on a small team where robotics domain expertise is core to the mission #J-18808-Ljbffr Rhoda AI

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Research Member of Technical Staff- Post-training & Robot Learning in Mountain View, CA vacancy
  •  ...generalist intelligent robots. We own the full...  ...our cutting edge research and end-to-end...  ...looking for a senior or staff-level Research...  ...to make our robot- learning pipeline reliable,...  ...generation, model training, inference, and...  ...and compilation, post-training, checkpoint... 
    Training

    Rhoda AI

    Mountain View, CA
    1 day ago
  • Member of Technical Staff Physical AI (Robotics / World Models) Palo Alto, CA About Orbifold AI...  ...robotics and world model research teams on the field's...  ..., evaluation, and model training. We design evaluation harnesses...  ...or reinforcement learning run. Each cycle compounds... 
    Training
    Shift work

    Bonfirevc

    Palo Alto, CA
    4 days ago
  •  ...Born out of Stanford Research, our team blends AI...  ...ll Do As a Founding Member of the Technical Staff at Architect, you'll...  ...at the forefront of training AI models for chip design...  ...the Reinforcement Learning environments and...  ...scaling, and improving post‑training techniques... 
    Training

    Kindredventures

    Palo Alto, CA
    4 days ago
  • About the Role As a Member of Technical Staff [Research] at NeoCognition , you’ll be part of the core team...  ...initiatives in the areas of LLM reasoning, post-training, and agentic system design....  ...Required Strong background in machine learning , natural language processing , or... 
    Training

    NeoCognition Inc.

    Palo Alto, CA
    2 days ago
  •  ...causal, multimodal systems that learn to predict and interact with...  ...promises to revolutionize robotics, science, healthcare, education...  ...together a world‑class research team from DeepMind, Tesla, Waymo...  ...simulators, feature extractors, and training grounds for real robot... 
    Training
    Remote work
    Flexible hours

    Odyssey

    Palo Alto, CA
    1 day ago
  •  ...multimodal systems that learn to predict and...  ...promises to revolutionize robotics, science, healthcare,...  ...together a world-class research team from DeepMind, Tesla...  ...for We hire deeply technical staff working in world models...  ...from scratch, scaling training runs, or working on inference... 
    Training

    Doist

    Palo Alto, CA
    2 days ago
  • Member of Technical Staff, Lead Researcher San Francisco, CA; Sunnyvale, CA About...  ...budgets for training and inference, sized...  ...model pre-training and post-training, RL training...  ...RL, continual learning, harness-based improvements...  ...interfaces Robotics and embodied AI for... 
    Training
    Local area

    DoorDash USA

    Sunnyvale, CA
    1 day ago
  •  ...multimodal systems that learn to predict and...  ...to revolutionize robotics, science,...  ...together a world-class research team from DeepMind...  ...intersection of deep technical research and scientific...  ...objectives, and training frameworks that...  .... Who you are A staff‑level or senior ML... 
    Training

    Odyssey

    Palo Alto, CA
    5 days ago
  • Member of Technical Staff, Research Evaluations Sanas is pioneering the future of human communication. Founded by a team of Stanford researchers and...  ...clear, actionable reporting Partner directly with model training and research teams to embed evaluation into the... 
    Training

    Sanas

    Palo Alto, CA
    2 days ago
  • $200k - $420k

     ...we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructure, next-generation UIs, and frontier deep learning research. Who we are We are scientists, engineers, and builders from the industry's top tech companies... 
    Training
    Local area
    Visa sponsorship
    Relocation package

    River AI Inc.

    Palo Alto, CA
    2 days ago
  • About Architect Architect is an AI research and product lab for chip design. We...  ...and implementing the Reinforcement Learning experiments (GRPO/PPO/DPO), training data mixes and reward signal explorations. Contribute to research on post‑training techniques, running... 
    Training
    Internship

    Architect Labs

    Palo Alto, CA
    4 days ago
  • Overview Perplexity is seeking top-tier AI Research Scientists and Engineers to advance our...  ...on foundational model capabilities, post-training techniques, building RL infra and...  ...the latest supervised and reinforcement learning techniques (SFT/DPO/GRPO) Leverage our... 
    Training

    Perplexity AI

    Palo Alto, CA
    3 days ago
  •  ...generation of generalist intelligent robots. We own the full robotics stack from...  ...cases, made possible by our cutting edge research and end-to-end system design. We've...  ...a reality. We're looking for a Staff / Principal ML Training Systems Engineer to own training systems... 
    Training

    Rhoda AI

    Mountain View, CA
    2 days ago
  •  ...rollout fidelity, and downstream robot task performance Research and implement data...  ...Collaborate closely with pre-training and post-training teams to ensure data...  ..., and actionable Staff-level candidates are expected to define technical direction and drive research... 
    Training

    Rhoda AI

    Mountain View, CA
    1 day ago
  •  ...next generation of generalist intelligent robots. We own the full robotics stack from...  ...cases, made possible by our cutting edge research and end-to-end system design. We've raised...  ...Engineer to build and maintain the training platform that powers our model development... 
    Training

    Rhoda AI

    Mountain View, CA
    2 days ago
  •  ...generalist intelligent robots. We own the full robotics...  ...by our cutting edge research and end‑to‑end system design...  ...traceability across training runs Build internal...  ...design to production Staff-level candidates are expected to define technical direction and own architectural... 
    Training
    Immediate start

    Rhoda AI

    Mountain View, CA
    1 day ago
  • Member of Technical Staff — Diffusion Model About the Role RadixArk is seeking a...  ...scalability. This role combines deep research thinking with strong...  ...novel algorithms to training and deploying models at scale...  ...Deep understanding of deep learning fundamentals and optimization... 
    Training
    Flexible hours

    RadixArk

    Palo Alto, CA
    2 days ago
  • About the Role As a Member of Technical Staff [Platform] at NeoCognition , you’ll design and...  ...that power everything we do — from research experiments to product...  ...to have: Experience with machine learning infrastructure , training pipelines, or model evaluation tooling... 
    Training

    NeoCognition

    Palo Alto, CA
    2 days ago
  •  ...causal, multimodal systems that learn to predict and interact with...  ...promises to revolutionize robotics, science, healthcare, education...  ...together a world-class research team from DeepMind, Tesla, Waymo...  ...features on modern GPUs to increase training and inference efficiency.... 
    Training

    Odyssey

    Palo Alto, CA
    2 days ago
  •  ...for We’re looking for a deeply technical and creative researcher who thrives on invention: someone...  ...do. Explore new architectures, learning objectives, and training paradigms that move beyond today...  ...generative models. Who you are A staff-level or senior researcher with... 
    Training

    Odyssey

    Santa Clara, CA
    2 days ago
  • Member of Technical Staff, LLM Post-Training, Applied Sanas is pioneering the future of human communication. Founded by a team of Stanford researchers and entrepreneurs with deep industry experience, Sanas...  ...building and deploying machine learning‑based services in a... 
    Training

    Sanas

    Palo Alto, CA
    2 days ago
  • $180k - $250k

    Member of Technical Staff -- TPU Systems (JAX / XLA / PALLAS) About the Role RadixArk...  ...-performance inference and training systems using JAX, XLA, and...  ..., or strong ability to learn quickly Responsibilities...  ...powers leading AI companies and research labs. Join us in building... 
    Training
    Full time
    Flexible hours

    RadixArk

    Palo Alto, CA
    2 days ago
  •  ...deliver industry‑leading training and inference speeds;...  ...We are seeking a Sr. Member of Technical Staff to design and develop...  ...execution, and post‑processing for real‑time...  ...cutting‑edge AI research. Work on one of the fastest...  ...through continuous learning, growth and support of... 
    Training

    Cerebras Systems, Inc.

    Sunnyvale, CA
    2 days ago
  •  ...The Role RadixArk is seeking a Member of Technical Staff, Developer Technology (DevTech) to make LLM inference and training dramatically faster, cheaper,...  ...leading AI companies and research labs, and Miles is our reinforcement-learning post-training framework for large-... 
    Training
    Flexible hours

    RadixArk

    Palo Alto, CA
    2 days ago
  • $300k - $350k

    Member of Technical Staff Level 1 - Engineering Bellevue | Hybrid NTT DATA AIVista, Inc., a wholly owned...  ..., Artificial Intelligence, Machine Learning, or a related field, or equivalent...  ...job‑related factors such as education, training, work experience, business needs, and... 
    Training
    Work experience placement
    Local area
    Flexible hours

    NTT DATA AIVista

    Palo Alto, CA
    1 day ago
  •  ...Machine Learning & AI for Mathematical Discovery and Reasoning...  ...successors to PatternBoost): set research agendas, design and run large...  ...junior researchers by providing technical guidance, rigorous code...  ...Deep expertise in large-scale training, reinforcement learning, program... 
    Training

    Axiom Math

    Palo Alto, CA
    5 days ago
  •  ...intelligence with a focus on physical self-replication. You will own the robot-learning research agenda, set technical direction, and shape hiring as the program grows. You will develop representations, training methods, and evaluations for coordinated physical behavior across... 
    Training

    GRAM

    Palo Alto, CA
    2 days ago
  •  ...causal, multimodal systems that learn to predict and interact with...  ...technology promises to revolutionize robotics, science, healthcare,...  ...brought together a world-class research team from DeepMind, Tesla, Waymo...  ...building. You will want to train them from scratch, at the scale... 
    Training

    Odyssey

    Palo Alto, CA
    3 days ago
  • Member of Technical Staff, ML Inference Engineering Sanas is pioneering the future...  ...by a team of Stanford researchers and entrepreneurs with deep...  ...across multi-node training and inference Analyze and...  ...Experience managing machine learning workloads on Kubernetes clusters... 
    Training

    Sanas

    Palo Alto, CA
    2 days ago
  • $180k

     ...infrastructure team is looking for an engineer to help develop our RL training framework.  RESPONSIBILITIES: Design and implement the...  ...training infrastructure Strong knowledge of reinforcement learning techniques Experience with RL numerics COMPENSATION AND... 
    Training
    Temporary work

    SpaceXAI

    Palo Alto, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Member of Technical Staff- Post-training & Robot Learning. Be the first to apply!