Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Scientist, Reinforcement Learning

RecruitSeq

Member of Technical Staff, Reinforcement Learning Research Our client is a well-funded, early-stage AI lab building a real-time, multimodal AI system designed to interact with the natural timing, emotion, and responsiveness of a human. The team is lean, highly credentialed, and moving fast on problems that are genuinely unsolved. About the Role Own RL and post-training for large-scale multimodal models at a frontier lab, from building the stack at 0→1 to scaling it to production. This is broader than a traditional RL algorithms role - you'll develop post-training methods and build the infrastructure needed to run them. The work spans rollout generation, reward modeling, policy optimization, evaluation, data feedback loops, serving, observability, and distributed execution. Responsibilities Build the RL/post-training stack from scratch: rollout generation, policy optimization, reward and reference model serving, data feedback loops, evaluation, checkpointing, and observability. Develop and scale post-training methods including PPO, GRPO, DPO, rejection sampling, RLHF/RLAIF, online RL, and model-based data improvement. Design systems abstractions connecting research ideas to production-scale RL runs: trainers, rollout workers, reward models, evaluators, data queues, experience buffers, and checkpoint promotion. Build evaluation and feedback loops for multimodal behavior: turn-taking, timing, emotional response, audiovisual coherence, instruction following, and real-time interaction quality. Optimize the end-to-end post-training loop across rollout throughput, serving latency, GPU utilization, policy update efficiency, and research iteration speed. Evolve the platform as algorithms, model architectures, reward definitions, data sources, and evaluation methods change. Qualifications PhD in ML, RL, CS, or a related field - completed or in final stretch Deep knowledge of RL and post-training methods: policy optimization, reward modeling, preference optimization, rejection sampling, KL control, and data feedback loops Ability to reason about training dynamics: reward hacking, unstable rewards, distribution shift, stale policies, mode collapse, over-optimization, noisy preferences, and evaluation mismatch Hands-on exposure to RL/post-training pipelines through research, internships, or open-source work; familiarity with frameworks such as verl, OpenRLHF, ms-swift, or equivalent, and rollout serving systems such as vLLM or SGLang Strong software engineering fundamentals with the appetite to build real systems, not just prototypes Must be willing to work on-site 5 days/week in Seattle, WA (relocation support available) Preferred Skills First-author publications at tier-1 ML journal/conference Internship experience at a top AI lab doing RL or post-training work Prior 0→1 experience building post-training systems, RL pipelines, agent training infrastructure, or evaluation platforms Hands-on experience with multimodal post-training for audio, video, or language models - particularly long-context or real-time interactive systems Experience with adjacent areas: distributed pretraining, inference serving, data infrastructure, simulation, or human/AI feedback collection Substantial open-source contributions in RL, post-training, alignment, or ML systems #J-18808-Ljbffr

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Research Scientist, Reinforcement Learning in Seattle, WA vacancy
  • $173k

    Senior Machine Learning Scientist The Senior Machine Learning Scientist is responsible for building...  ...projects: problem framing, ideation, research, prototyping, deployment, and post‑...  ...and prototype advanced ML techniques (reinforcement learning, sequence modeling, transformers... 
    Suggested
    Flexible hours

    NLP PEOPLE

    Seattle, WA
    3 days ago
  •  ...deployment, constantly evolving to power the next generation AI use cases. Role Summary: We are looking for a talented Machine Learning Research Scientist to play a key role in designing our core AI inference platform. You will be actively conducting high impact research on... 
    Suggested
    Work at office
    Flexible hours
    3 days per week

    ElastixAI INC.

    Seattle, WA
    5 days ago
  • $150k - $173k

     ...will join a small but world‑class Applied Research and AI team and work on genuinely hard,...  ...fluctuations, permitting hold‑ups — learning from historical project data to improve...  ...patents, or equivalent deployed systems). Reinforcement Learning Depth: Hands‑on experience... 
    Suggested
    Bi-weekly pay
    Shift work

    The Nuclear Company

    Seattle, WA
    3 days ago
  • $167.03k - $260.57k

     ...are a creative, hands‑on AI researcher with a strong foundation in...  ...will collaborate with human scientists, not replace them. Your Next...  ...research scientists and machine learning engineers at Ai2, you will...  ...or more of the following: reinforcement learning, experimental methodology... 
    Suggested
    Contract work
    Temporary work
    Work at office
    Flexible hours
    Weekend work

    Allen Institute for Artificial Intelligence

    Seattle, WA
    3 days ago
  • $176k - $255k

     ...and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise...  ...in Computer Science, Machine Learning, AI, or a related field. Deep understanding of deep learning, reinforcement learning, and large-scale model fine... 
    Suggested
    Full time
    Shift work

    Scale AI

    Seattle, WA
    more than 2 months ago
  • $290.25k

     ...Impact We are seeking highly skilled and innovative Machine Learning Scientists to join our AI team, focusing on AI applications (LLM and Computer...  ...) in Cloud, Devices and Robotics. As a key member of our research and development efforts, you will play a crucial role in... 
    Remote work

    National Society for Black Engineers

    Seattle, WA
    5 days ago
  • $224k

     ...powered by data and machine learning provides secure, differentiated...  ...Principal Machine Learning Scientist to join our growing Travel...  ...role, you will:  Drive the research, design, and deployment of...  ...learning for recommendations, reinforcement learning, and causal... 

    Expedia Group

    Seattle, WA
    1 day ago
  • $173k

     ...is part of a focused machine learning science team within...  .... Raise the bar: Mentor scientists through code reviews and design...  ...Master’s or PhD in Operations Research, Applied Mathematics, Statistics...  ...with deep learning, reinforcement learning, or multi-armed bandits... 
    Worldwide

    Expedia Group

    Seattle, WA
    1 day ago
  • $177.69k - $341.73k

    Senior Machine Learning Engineer, Risk Data Mining - USDS The USDS-Platform and Community Integrity (PaCI) team is missioned to: Protect...  ...such as deep neural nets, transfer/multi-task learning, reinforcement learning, time series or graph unsupervised learning. Ability... 
    Temporary work

    TikTok

    Seattle, WA
    4 days ago
  • $175k - $257.5k

     ...outcomes come from unconstrained thinking. Position overview: As a Research Scientist you will leverage your deep understanding of scientific methodologies, e.g. operations research (OR), Machine Learning (ML), software engineering, and problem solving to innovate... 

    Keystone Strategy

    Seattle, WA
    3 days ago
  • $190k - $320k

     ...come alive. About The Role New and novel approaches are needed to realize all of our product goals. As a Machine Learning Scientist at Sesame, you are a research-oriented person with experience in NLP, Speech, and/or Computer Vision with a focus on deep learning. You are... 
    Full time
    Contract work
    Flexible hours

    SESAME

    Bellevue, WA
    2 days ago
  • $129.4k - $242.2k

     ...organization, and we know the next big idea could be yours! Adobe Research is looking for research scientists working on generative AI models for image and video...  ...any other applicable characteristics protected by law. Learn more. Adobe aims to make Adobe.com accessible to any... 
    Temporary work
    Worldwide

    Adobe

    Seattle, WA
    2 days ago
  • $192k - $304.75k

     ...NVIDIA is seeking an Applied Deep Learning Research Scientist to join our Efficiency team. This role involves researching low-bit number representations and optimizing neural network performance. The ideal candidate will hold a PhD in a related field with over five years... 

    NVIDIA

    Seattle, WA
    3 days ago
  • $184.7k - $200.2k

     ...Research Scientist at Meta – Bellevue, WA Meta Platforms, Inc. (Meta), formerly known as Facebook Inc., builds technologies that help people...  ..., and development of new core infrastructure. Use machine learning, statistics, or other data techniques to build algorithms. Suggest... 
    Hourly pay
    Internship
    Live in

    Victrays

    Bellevue, WA
    5 days ago
  • $164.96k - $215.3k

     ...Develop and apply state of the art machine learning methods to large, multi source datasets...  ...innovations. Mentor and coach junior scientists, including through lunch‑and‑learn sessions...  ...three years of experience. Conducting research in machine learning, natural language processing... 
    Work at office

    Amazon

    Seattle, WA
    3 days ago
  • $125.5k - $169.8k

     ...Research Scientist, Selling Partner Experience Job ID: 10484849 | Amazon.com Services LLC We’re looking for a Research Scientist to join a...  ...programming language such as Python, Java, C++ Knowledge of machine learning processing: computer vision, NLU, NLP or operations research... 
    Remote work
    Flexible hours
    Shift work
    Day shift

    Amazon

    Seattle, WA
    3 days ago
  • $167.1k - $226.1k

     ...capabilities. We are looking for a Senior Applied Scientist with deep expertise in quantum error...  ...the FTQC Science Lead to translate research direction into implementable solutions:...  ...matching, paid time off, and parental leave. Learn more about our benefits at USA, WA,... 
    Flexible hours

    Socket

    Seattle, WA
    3 days ago
  •  ...A leading research organization in Seattle is seeking an AI Research Scientist II to develop and deploy large-scale machine learning models for the analysis of complex biological data. The ideal candidate will have a PhD in a relevant field and experience in AI/ML. This... 
    Relocation package

    Allen Institute

    Seattle, WA
    3 days ago
  •  ...should have a PhD or equivalent experience with a strong background in training audio generation models and deep learning. You will work alongside a top-tier research team to create lifelike avatars that understand and respond to human nuances, helping bridge the emotional... 

    Nuance Labs, Inc.

    Seattle, WA
    3 days ago
  •  ...Employer: AMAZON DEVELOPMENT CENTER U.S., INC. Offered Position: Research Scientist III Job Location: Seattle, Washington Job Number: AMZ1006159...  ...Responsibilities Develop and apply state of the art machine learning methods to large, multi source datasets to build and... 
    Work at office

    Amazon Science

    Seattle, WA
    2 days ago
  • $159.75k - $255.6k

     ...Your Impact We are seeking a skilled and innovative Senior AI Research Scientist to join a new team focusing on agentic video and multimodal...  ...lives. You will advance the state-of-the-art in machine learning and multimodal technology and apply your research findings to... 
    Work experience placement
    Work at office
    Remote work

    University of Georgia- FACS

    Seattle, WA
    3 days ago
  •  ...forward-thinking Senior Applied Scientist to join their Seattle-based...  ...AI, agentic AI, and deep learning-based solutions. You’ll partner...  ...who enjoys balancing research, experimentation, and hands‑...  ...agentic AI, deep learning, reinforcement learning, and/or graph neural... 
    Full time
    Work at office
    Flexible hours

    Designworks Talent LLC

    Seattle, WA
    5 days ago
  •  ...We are seeking an experienced Applied Scientist to drive research and development in Large Language Models (LLMs) and machine learning systems for business applications. The...  ...particularly in generative AI, Transformers, reinforcement learning, and state-of-the-art NLP/... 

    Veriipro

    Seattle, WA
    3 days ago
  • $192.2k - $260k

     ...highly skilled and experienced Applied Scientist to support adoption and enable...  ...customization, including supervised fine-tuning, reinforcement learning, and knowledge distillation across...  ...Master's degree and 6+ years of applied research experience. Experience programming in... 
    Flexible hours

    Amazon

    Bellevue, WA
    3 days ago
  • $171.6k - $222.2k

     ...Contribute to the design and development of GenAI, deep learning, multi-objective optimization and/or reinforcement learning empowered solutions to transform ad...  .... Collaborate cross-functionally with other scientists, engineers, and product managers to bring scalable... 
    Seasonal work
    Local area
    Flexible hours

    Amazon Science

    Seattle, WA
    3 days ago
  • $167.1k - $226.1k

     ...for you. Watch this video to learn more about our organization,...  ...for a Senior Applied Scientist to join us! OSS designs and...  ...understanding about machine learning/reinforcement learning/GenAI models,...  ...chain optimization, operations research, or vendor management systems... 
    Flexible hours

    Amazon Science

    Bellevue, WA
    3 days ago
  • $142.7k - $270.95k

     ...Description Adobe Firefly’s Applied Science & Machine Learning (ASML) group is looking for research scientists and engineers focused on post‑training, alignment...  ...interested in candidates with expertise in reinforcement learning from human feedback (RLHF), direct preference... 
    Worldwide

    Dormont Manufacturing Company

    Seattle, WA
    1 day ago
  • $200k - $250k

     ...this challenge. Your role as Applied Scientist is to build next-generation AI agents...  ...approaches fall short. Your focus: Applied research from research to data to model fine-...  ...supervised fine-tuning and reinforcement learning techniques. Collaborating with engineering... 
    Remote work
    Flexible hours

    techire ai

    Seattle, WA
    1 day ago
  • $142.8k - $193.2k

     ...talent pool. We are looking for a highly talented applied scientist to be part of our science team in Bellevue, WA. In this role...  ...business decisions and set new policies. - Lead exploration of reinforcement learning techniques to augment optimization algorithms. A day in the... 
    Full time
    Temporary work
    Seasonal work
    Flexible hours

    Amazon

    Bellevue, WA
    3 days ago
  • $136k - $184k

     ...customer behavior through machine learning, artificial intelligence,...  ...customers? Join our team of Scientists and Engineers developing...  ...step optimization leading to reinforcement learning of the customer journey...  ...and value of scientific research and have developed a strong... 
    Temporary work
    Flexible hours

    Amazon

    Seattle, WA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Scientist, Reinforcement Learning. Be the first to apply!