Reinforcement Learning Researcher (Humanoid)
However, REK Inc
About the job Reinforcement Learning Researcher (Humanoid) Location: San Francisco, CA (On-site at REK HQ) Reports to: CTO Company: REK Inc. About REK REK is pioneering a new global sport: Robot Kombat. Its VR-controlled humanoid robot fighting. We merge robotics, gaming, and live entertainment into the most advanced real-time human-machine competition on Earth. Our debut event, REK0 , was the most attended robot fight in Temple Nightclubs history. As we scale globally, were building a world-class robotics and AI team to push the limits of humanoid motion, balance, and control through reinforcement learning and sim-to-real transfer. The Role Were seeking a Humanoid Robot Reinforcement Learning Researcher to develop and train advanced control policies for our REK humanoid robots. Youll work closely with REKs robotics, simulation, and teleoperation teams to design reinforcement learning pipelines that improve movement quality, stability, responsiveness, and adaptability all optimized for real-time fighting performance. This role sits at the intersection of research and applied engineering: youll design algorithms, implement simulation environments, train agents, and test them on full-scale humanoid robots. Responsibilities RL Algorithm Development Design and implement state-of-the-art reinforcement learning algorithms for humanoid locomotion, balance, and reactive movement. Simulation Environments Build and customize physics-based simulation environments (Isaac Gym, MuJoCo, PyBullet, etc.) for efficient training and domain randomization. Sim-to-Real Transfer Develop robust transfer strategies that ensure trained policies perform reliably on physical humanoid robots. Policy Evaluation Define performance metrics, run experiments, and benchmark results for both simulated and real-world tests. Integration & Collaboration Work closely with teleoperation and control system teams to blend RL policies with operator input in hybrid control architectures. Research & Publication Stay current with cutting-edge humanoid and robotics RL research and contribute to internal whitepapers or external publications as appropriate. Data & Infrastructure Maintain scalable training pipelines and data logging systems using GPU clusters or cloud resources. Required Qualifications 2+ years of hands-on experience in reinforcement learning for humanoid robots at a university lab, research institute, or corporation. Deep understanding of deep RL algorithms (PPO, SAC, TD3, DDPG, etc.). Experience using Isaac Gym, MuJoCo, PyBullet, or Gazebo for simulation training. Strong software engineering skills in Python, PyTorch, and C++ . Understanding of robot kinematics, dynamics, and control systems. Ability to run large-scale experiments efficiently and interpret quantitative results. Strong communication skills and comfort working in an experimental, cross-disciplinary team. Preferred / Bonus Qualifications Experience with Unitree G1 EDU Background in teleoperation, imitation learning, or motion retargeting. Familiarity with low-latency communication protocols and embedded robot control. Understanding of VR systems, motion capture, or human-in-the-loop RL. Mandarin proficiency a major plus for collaboration with Chinese robotics partners and manufacturers. Compensation Competitive salary based on experience. Equity participation in a rapidly growing robotics entertainment startup. Full benefits and access to REKs humanoid robotics lab and training infrastructure. Why REK At REK, your research will come to life on stage powering real humanoid robots in front of live audiences. You wont just train agents in simulation youll see your policies executed in full-scale, physical combat, shaping the future of sport, robotics, and human-machine interaction. #J-18808-Ljbffr However, REK Inc
- An innovative robotics entertainment company in San Francisco is seeking a Reinforcement Learning Researcher to develop advanced control policies for humanoid robots. This on-site role involves creating state-of-the-art algorithms, running simulations, and ensuring policies...Suggested
$15k
...real-world experience applying ML to financial markets.Voleon Securities is looking to add an experienced and creative reinforcement learning researcher to our growing ML research group. We welcome researchers with both strong theoretical foundations and experience...SuggestedLocal areaImmediate startRelocationWork visa$192.6k - $344.85k
...Scientist & ManagerPost-Training · Alignment · Reinforcement LearningAutodesk AI Lab: London · San... ...-capable systems is still an open research problem.Autodesk touches more of the... ...preference data, we can ground reinforcement learning in the laws of physics and the...SuggestedFull timeFor contractorsRemote work$192.6k - $344.85k
## AI Research Manager/Scientist, Reinforcement LearningApplylocations: San Francisco, CA, USA: AMER - United States - Massachusetts - Boston - Drydock... ...OverviewAs an **AI Scientist Manager Reinforcement Learning** at Autodesk Research, you will be doing fundamental...SuggestedRemote work$250k - $350k
...architectures, running experiments, and turning research insights into products that ship, we'd... ...from exploring new architectures and learning methods to optimizing latency and... ...other post-training techniques Perform reinforcement learning research to improve model...SuggestedFull timeWork at office- Anthropic is seeking an education researcher for its Education Labs to study how people learn with AI and how to measure AI fluency. You will design instruments, build tools, run studies, and translate findings into changes across product experiments and action research...
- Embodied AI Research Scientist — Humanoid Robotics San Francisco Bay Area, CA · Full-time · On-site About... ..., multimodal AI, and large-scale learning systems — building the end-to-end... ...of: VLMs, VLA Models, World Models, Reinforcement Learning, Embodied AI, Robot Learning...Full time
$216.3k - $280.8k
...Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance... ...models, agentic AI, multimodal learning, reasoning systems, scalable training algorithms... ...research in advanced domains such as reinforcement learning, mechanistic interpretability,...Full timeTemporary workLocal areaFlexible hours$150k - $250k
...manufacturing, consumer goods, and global social organizations.We research and deploy technologies that power AI-native operations —... ...to enable autonomous evolution. This work bridges reinforcement learning, interpretability, and meta-optimization to pioneer continuously...Work at office3 days per week$204.61k - $306.91k
...at Hinge Health, you'll own the machine learning that decides what message each member... ...Computer Science, Statistics, Operations Research, Machine Learning, or a related... ...at consumer scaleContextual bandits or reinforcement learning operated in productionMulti-objective...Local areaImmediate start$150k
...continuously fuzz-testing them. We are looking for Research Engineers to help develop our reliability... ...! Some familiarity with ideas from active learning, weak supervision, synthetic data, functional verification, reinforcement learning, reward modeling, automated...Visa sponsorship$147.6k - $274k
...drug discovery and development. Roche’s Research and Early Development organisations at... ...Sciences Center of Excellence as a Machine Learning Scientist / Senior Machine Learning... ...with GNNs, sequence/language models and reinforcement learning.You are fluent in Python and...Full timeLocal areaWorldwideRelocation package- ...THE COMPANY We're building autonomous research agents for recursive self-improvement (... ...that propose, run, and analyze machine learning experiments). We're a small team based... ...supervised fine-tuning, preference data, reinforcement learning from human and AI feedback, reward...
$7.84k
...deliver novel zeroth-order mathematical optimization and machine learning algorithms, models, and software to accelerate scientific... ...mathematics, computer/computational science, data science, operations research, or a related field is required. Fundamental knowledge of...Full timeContract work- The Biological Computing Co. is seeking a Senior AI Researcher to lead research on video generation models and neurally optimized AI infrastructure. You will design scalable models, iterate from theory to implementation, and drive platform-level experiments in collaboration...
- ...then use that data to advance our own ML research, while also collaborating with leading... ..., feature extraction, representation learning, model training, evaluation, inference,... ...literature, and mechanistic evidence. Reinforcement learning for basic biology, chemistry,...
- ...real, paying customers almost immediately. No speculative research track here. If you want your work to hit production within... ...Experience fine‑tuning or training language or speech models with reinforcement learning Published research or open-source work in speech or...Permanent employmentFull timeImmediate start
- ...training, alignment evaluations, monitoring, and frontier‑risk research. We care most about monitorability where the stakes are... ...pre‑training, synthetic data, mid‑training, post‑training, reinforcement learning, and other interventions improve or degrade monitorability....Work at officeRelocation package3 days per week
- Patronus AI, Inc. in San Francisco seeks an Applied Researcher to drive foundational research in training and evaluating agentic... ...systems. You will solve critical research questions using reinforcement learning and simulations, working on real-world applications alongside...
$350k
Thinking Machines in San Francisco is looking for post-training researchers. This role involves developing and tuning post-training processes, debugging training configurations, and publishing impactful research. Candidates should have proficiency in Python and a relevant...Visa sponsorship- ...Office Employment Type Full time Department Engineering AI Researcher Job Description At AGI Inc., we’re not just redefining AI‑human... ...Novel Research: Lead research initiatives in areas such as reinforcement learning, multi‑agent systems, and natural language understanding to...Full timeWork at office
- ...Cartesia Our mission is to architect AI that learns from and interacts with the world like... ...This team is where customer needs meet research, and covers the full spectrum of... ...quality, experiment with finetuning and reinforcement learning approaches to refine model behavior...Work at officeVisa sponsorshipFlexible hours
- ML Researcher (World Models) We are seeking Researchers with expertize in World Models to join an an NVIDIA Inception & SPC backed lab building SOTA AI voice products. The team were founded by Berkeley alumni, who spent time honing their expertise in big tech, before...
- ...biology, computational neuroscience, AI research and software engineering to develop new... ...generation models that enable robots to learn, plan, and act through imagined futures.... ...autoregressive video, or sequence models Model-based reinforcement learning or planning System...
$218.4k - $273k
...Agent Capabilities & Environments (ACE) team, part of Scale’s Research organization, brings together customer-facing Researchers and Applied... ...cloud technology stack (eg. AWS or GCP) and developing machine learning models in a cloud environment.Our research interviews are...Full time- ...business in history. We work directly with frontier AI lab researchers to create evaluations, publish benchmarks, and push the... ...professional agent development. You'll also partner closely with our Reinforcement Learning Environments team, blending taxonomy research with...Full timeWork at officeRemote workFlexible hours
$142.7k - $270.95k
...what is possible in design!What You'll DoBuild innovative machine learning models that drive Agentic and other Generative AI scenarios for... ..., and production deployment.Collaborate closely with Adobe Research, engineering, and product teams to bring AI-powered features to...Full timeTemporary workLocal areaWorldwide- ...biology, computational neuroscience, AI research and software engineering to develop new... ...generation models that enable robots to learn, plan, and act through imagined futures.... ...autoregressive video, or sequence models Model‑based reinforcement learning or planning System...
- Goaly is seeking an AI Researcher to lead research on agentic AI, focusing on training specialized models and building orchestration... ...a Ph.D. or Master's in relevant fields, deep knowledge of reinforcement learning and LLMs, and are experienced with frameworks like PyTorch...
$116.2k - $146.7k
Associate Professional Researcher - Andino LabThe Andino Laboratory at UCSF is seeking an Associate Professional Researcher to join the lab... ...communication. • Demonstrated willingness and ability to learn and implement new methods and technical skills in response to evolving...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Reinforcement Learning Researcher (Humanoid). Be the first to apply!
- senior researcher San Francisco, CA
- machine learning researcher San Francisco, CA
- researcher San Francisco, CA
- senior design researcher San Francisco, CA
- design researcher San Francisco, CA
- qualitative researcher San Francisco, CA
- data collection researcher San Francisco, CA
- product researcher San Francisco, CA
- survey researcher San Francisco, CA
- legal researcher San Francisco, CA


