Research Scientist, Reinforcement Learning
RecruitSeq
Member of Technical Staff, Reinforcement Learning Research Our client is a well-funded, early-stage AI lab building a real-time, multimodal AI system designed to interact with the natural timing, emotion, and responsiveness of a human. The team is lean, highly credentialed, and moving fast on problems that are genuinely unsolved. About the Role Own RL and post-training for large-scale multimodal models at a frontier lab, from building the stack at 0→1 to scaling it to production. This is broader than a traditional RL algorithms role - you'll develop post-training methods and build the infrastructure needed to run them. The work spans rollout generation, reward modeling, policy optimization, evaluation, data feedback loops, serving, observability, and distributed execution. Responsibilities Build the RL/post-training stack from scratch: rollout generation, policy optimization, reward and reference model serving, data feedback loops, evaluation, checkpointing, and observability. Develop and scale post-training methods including PPO, GRPO, DPO, rejection sampling, RLHF/RLAIF, online RL, and model-based data improvement. Design systems abstractions connecting research ideas to production-scale RL runs: trainers, rollout workers, reward models, evaluators, data queues, experience buffers, and checkpoint promotion. Build evaluation and feedback loops for multimodal behavior: turn-taking, timing, emotional response, audiovisual coherence, instruction following, and real-time interaction quality. Optimize the end-to-end post-training loop across rollout throughput, serving latency, GPU utilization, policy update efficiency, and research iteration speed. Evolve the platform as algorithms, model architectures, reward definitions, data sources, and evaluation methods change. Qualifications PhD in ML, RL, CS, or a related field - completed or in final stretch Deep knowledge of RL and post-training methods: policy optimization, reward modeling, preference optimization, rejection sampling, KL control, and data feedback loops Ability to reason about training dynamics: reward hacking, unstable rewards, distribution shift, stale policies, mode collapse, over-optimization, noisy preferences, and evaluation mismatch Hands-on exposure to RL/post-training pipelines through research, internships, or open-source work; familiarity with frameworks such as verl, OpenRLHF, ms-swift, or equivalent, and rollout serving systems such as vLLM or SGLang Strong software engineering fundamentals with the appetite to build real systems, not just prototypes Must be willing to work on-site 5 days/week in Seattle, WA (relocation support available) Preferred Skills First-author publications at tier-1 ML journal/conference Internship experience at a top AI lab doing RL or post-training work Prior 0→1 experience building post-training systems, RL pipelines, agent training infrastructure, or evaluation platforms Hands-on experience with multimodal post-training for audio, video, or language models - particularly long-context or real-time interactive systems Experience with adjacent areas: distributed pretraining, inference serving, data infrastructure, simulation, or human/AI feedback collection Substantial open-source contributions in RL, post-training, alignment, or ML systems #J-18808-Ljbffr
$173k
Senior Machine Learning Scientist The Senior Machine Learning Scientist is responsible for building... ...projects: problem framing, ideation, research, prototyping, deployment, and post‑... ...and prototype advanced ML techniques (reinforcement learning, sequence modeling, transformers...SuggestedFlexible hours- ...deployment, constantly evolving to power the next generation AI use cases. Role Summary: We are looking for a talented Machine Learning Research Scientist to play a key role in designing our core AI inference platform. You will be actively conducting high impact research on...SuggestedWork at officeFlexible hours3 days per week
$150k - $173k
...will join a small but world‑class Applied Research and AI team and work on genuinely hard,... ...fluctuations, permitting hold‑ups — learning from historical project data to improve... ...patents, or equivalent deployed systems). Reinforcement Learning Depth: Hands‑on experience...SuggestedBi-weekly payShift work$167.03k - $260.57k
...are a creative, hands‑on AI researcher with a strong foundation in... ...will collaborate with human scientists, not replace them. Your Next... ...research scientists and machine learning engineers at Ai2, you will... ...or more of the following: reinforcement learning, experimental methodology...SuggestedContract workTemporary workWork at officeFlexible hoursWeekend work$176k - $255k
...and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise... ...in Computer Science, Machine Learning, AI, or a related field. Deep understanding of deep learning, reinforcement learning, and large-scale model fine...SuggestedFull timeShift work$290.25k
...Impact We are seeking highly skilled and innovative Machine Learning Scientists to join our AI team, focusing on AI applications (LLM and Computer... ...) in Cloud, Devices and Robotics. As a key member of our research and development efforts, you will play a crucial role in...Remote work$224k
...powered by data and machine learning provides secure, differentiated... ...Principal Machine Learning Scientist to join our growing Travel... ...role, you will: Drive the research, design, and deployment of... ...learning for recommendations, reinforcement learning, and causal...$173k
...is part of a focused machine learning science team within... .... Raise the bar: Mentor scientists through code reviews and design... ...Master’s or PhD in Operations Research, Applied Mathematics, Statistics... ...with deep learning, reinforcement learning, or multi-armed bandits...Worldwide$177.69k - $341.73k
Senior Machine Learning Engineer, Risk Data Mining - USDS The USDS-Platform and Community Integrity (PaCI) team is missioned to: Protect... ...such as deep neural nets, transfer/multi-task learning, reinforcement learning, time series or graph unsupervised learning. Ability...Temporary work$175k - $257.5k
...outcomes come from unconstrained thinking. Position overview: As a Research Scientist you will leverage your deep understanding of scientific methodologies, e.g. operations research (OR), Machine Learning (ML), software engineering, and problem solving to innovate...$190k - $320k
...come alive. About The Role New and novel approaches are needed to realize all of our product goals. As a Machine Learning Scientist at Sesame, you are a research-oriented person with experience in NLP, Speech, and/or Computer Vision with a focus on deep learning. You are...Full timeContract workFlexible hours$129.4k - $242.2k
...organization, and we know the next big idea could be yours! Adobe Research is looking for research scientists working on generative AI models for image and video... ...any other applicable characteristics protected by law. Learn more. Adobe aims to make Adobe.com accessible to any...Temporary workWorldwide$192k - $304.75k
...NVIDIA is seeking an Applied Deep Learning Research Scientist to join our Efficiency team. This role involves researching low-bit number representations and optimizing neural network performance. The ideal candidate will hold a PhD in a related field with over five years...$184.7k - $200.2k
...Research Scientist at Meta – Bellevue, WA Meta Platforms, Inc. (Meta), formerly known as Facebook Inc., builds technologies that help people... ..., and development of new core infrastructure. Use machine learning, statistics, or other data techniques to build algorithms. Suggest...Hourly payInternshipLive in$164.96k - $215.3k
...Develop and apply state of the art machine learning methods to large, multi source datasets... ...innovations. Mentor and coach junior scientists, including through lunch‑and‑learn sessions... ...three years of experience. Conducting research in machine learning, natural language processing...Work at office$125.5k - $169.8k
...Research Scientist, Selling Partner Experience Job ID: 10484849 | Amazon.com Services LLC We’re looking for a Research Scientist to join a... ...programming language such as Python, Java, C++ Knowledge of machine learning processing: computer vision, NLU, NLP or operations research...Remote workFlexible hoursShift workDay shift$167.1k - $226.1k
...capabilities. We are looking for a Senior Applied Scientist with deep expertise in quantum error... ...the FTQC Science Lead to translate research direction into implementable solutions:... ...matching, paid time off, and parental leave. Learn more about our benefits at USA, WA,...Flexible hours- ...A leading research organization in Seattle is seeking an AI Research Scientist II to develop and deploy large-scale machine learning models for the analysis of complex biological data. The ideal candidate will have a PhD in a relevant field and experience in AI/ML. This...Relocation package
- ...should have a PhD or equivalent experience with a strong background in training audio generation models and deep learning. You will work alongside a top-tier research team to create lifelike avatars that understand and respond to human nuances, helping bridge the emotional...
- ...Employer: AMAZON DEVELOPMENT CENTER U.S., INC. Offered Position: Research Scientist III Job Location: Seattle, Washington Job Number: AMZ1006159... ...Responsibilities Develop and apply state of the art machine learning methods to large, multi source datasets to build and...Work at office
$159.75k - $255.6k
...Your Impact We are seeking a skilled and innovative Senior AI Research Scientist to join a new team focusing on agentic video and multimodal... ...lives. You will advance the state-of-the-art in machine learning and multimodal technology and apply your research findings to...Work experience placementWork at officeRemote work- ...forward-thinking Senior Applied Scientist to join their Seattle-based... ...AI, agentic AI, and deep learning-based solutions. You’ll partner... ...who enjoys balancing research, experimentation, and hands‑... ...agentic AI, deep learning, reinforcement learning, and/or graph neural...Full timeWork at officeFlexible hours
- ...We are seeking an experienced Applied Scientist to drive research and development in Large Language Models (LLMs) and machine learning systems for business applications. The... ...particularly in generative AI, Transformers, reinforcement learning, and state-of-the-art NLP/...
$192.2k - $260k
...highly skilled and experienced Applied Scientist to support adoption and enable... ...customization, including supervised fine-tuning, reinforcement learning, and knowledge distillation across... ...Master's degree and 6+ years of applied research experience. Experience programming in...Flexible hours$171.6k - $222.2k
...Contribute to the design and development of GenAI, deep learning, multi-objective optimization and/or reinforcement learning empowered solutions to transform ad... .... Collaborate cross-functionally with other scientists, engineers, and product managers to bring scalable...Seasonal workLocal areaFlexible hours$167.1k - $226.1k
...for you. Watch this video to learn more about our organization,... ...for a Senior Applied Scientist to join us! OSS designs and... ...understanding about machine learning/reinforcement learning/GenAI models,... ...chain optimization, operations research, or vendor management systems...Flexible hours$142.7k - $270.95k
...Description Adobe Firefly’s Applied Science & Machine Learning (ASML) group is looking for research scientists and engineers focused on post‑training, alignment... ...interested in candidates with expertise in reinforcement learning from human feedback (RLHF), direct preference...Worldwide$200k - $250k
...this challenge. Your role as Applied Scientist is to build next-generation AI agents... ...approaches fall short. Your focus: Applied research from research to data to model fine-... ...supervised fine-tuning and reinforcement learning techniques. Collaborating with engineering...Remote workFlexible hours$142.8k - $193.2k
...talent pool. We are looking for a highly talented applied scientist to be part of our science team in Bellevue, WA. In this role... ...business decisions and set new policies. - Lead exploration of reinforcement learning techniques to augment optimization algorithms. A day in the...Full timeTemporary workSeasonal workFlexible hours$136k - $184k
...customer behavior through machine learning, artificial intelligence,... ...customers? Join our team of Scientists and Engineers developing... ...step optimization leading to reinforcement learning of the customer journey... ...and value of scientific research and have developed a strong...Temporary workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist, Reinforcement Learning. Be the first to apply!
- drug safety scientist Seattle, WA
- image scientist Seattle, WA
- molecular biology scientist Seattle, WA
- safety scientist Seattle, WA
- validation scientist Seattle, WA
- senior analytical scientist Seattle, WA
- support scientist Seattle, WA
- water quality scientist Seattle, WA
- scientist 1 Seattle, WA
- senior research scientist - machine learning Seattle, WA

