Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Research Scientist - Reinforcement Learning

$200k - $250k
Full-time

Centific

Role Description

What You'll Do:

  • Design simulation environments and digital twins for enterprise workflows
  • Post-train LLM agents using RLHF, DPO, GRPO, PPO, and emerging methods
  • Build pipelines that convert human-labeled traces and verifiable signals into training data
  • Architect multi-turn, tool-using agents with closed learning loops
  • Design reward functions and verifiers that resist reward hacking and reflect real task outcomes
  • Set the technical bar across the team — architecture, code review, engineering standards
  • Mentor researchers and engineers; drive technical direction through influence
  • Translate research into production; contribute to publications

Qualifications

  • 7+ years in ML/AI research or engineering; 3+ years at senior/staff level
  • MS or PhD in Computer Science, Machine Learning, or related field (or equivalent)
  • 5+ years hands-on RL — environment design, reward engineering, policy optimization — with at least one production deployment
  • 3+ years fine-tuning LLMs with hands-on RL post-training (RLHF, DPO, GRPO, PPO)
  • Expert-level implementation of RLHF pipelines, reward modeling (Bradley-Terry), DPO, and KTO
  • Working knowledge of modern post-training and rollout-serving libraries (TRL, veRL, OpenRLHF, SkyRL)
  • Experience building LLM-based agents: tool use, multi-turn reasoning, trajectory evaluation
  • Strong Python and software engineering skills — comfortable building production pipelines, not just notebooks
  • Deep expertise in MDPs, policy gradient methods (PPO, SAC), and temporal difference learning
  • Hands-on experience with Gymnasium-based environments and reward engineering (sparse vs. dense)

Preferred Qualifications

  • Publications at NeurIPS, ICML, ICLR, ACL, COLM, or similar venues
  • Open-source contributions to post-training or agent frameworks (TRL, veRL, OpenRLHF, SkyRL)
  • Experience with Offline RL (CQL, IQL), Model-based RL / World Models, or Hierarchical RL
  • Background in synthetic data generation, simulation, or world models
  • Domain experience in healthcare, finance, logistics, or compliance
  • Distributed training on GPU clusters

Benefits

  • Lead the frontier. Shape a new discipline at the intersection of post-training, simulation, and enterprise AI.
  • Ship your science. See your research power real systems across healthcare, finance, and safety-critical operations.
  • Collaborate with leaders. Work alongside NVIDIA, Microsoft, and the global AI community.
  • Build what matters. Create governed, compliant AI systems enterprises can actually trust.
  • Salary: $200k-$250k

How to Apply

Send your CV, a description of a technically complex system you personally built or led, and (if applicable) your publication list or open-source contributions to:

View email address on us.fitly.work

Subject: Senior Staff Research Scientist – RL

Company Description

Centific is a frontier AI data foundry that curates diverse, high-quality data, using our purpose-built technology platforms to empower the Magnificent Seven and our enterprise clients with safe, scalable AI deployment. Our team includes more than 150 PhDs and data scientists, along with more than 4,000 AI practitioners and engineers. We harness the power of an integrated solution ecosystem—comprising industry-leading partnerships and 1.8 million vertical domain experts in more than 230 markets—to create contextual, multilingual, pre-trained datasets; fine-tuned, industry-specific LLMs; and RAG pipelines supported by vector databases. Our zero-distance innovation™ solutions for GenAI can reduce GenAI costs by up to 80% and bring solutions to market 50% faster.

Our mission is to bridge the gap between AI creators and industry leaders by bringing best practices in GenAI to unicorn innovators and enterprise customers. We aim to help these organizations unlock significant business value by deploying GenAI at scale, helping to ensure they stay at the forefront of technological advancement and maintain a competitive edge in their respective markets.

Vacancy posted more than 2 months ago
Similar jobs that could be interesting for youBased on the Staff Research Scientist - Reinforcement Learning in Remote vacancy
  • $94.49k - $147.4k

     ...Leadership Computing Facility (ALCF) is seeking a Staff Scientist in Post-Training and Reinforcement Learning for AI for Science to help advance the next generation...  ...workflows.The successful candidate will conduct research on methods that improve the usefulness,... 
    Suggested
    Full time
    For contractors
    Remote work

    Argonne National Laboratory

    Lemont, IL
    2 days ago
  • $110k - $220k

     ...Science team is looking for a Staff Data Scientist to work alongside our team...  ...Team: The Applied AI team researches and develops advanced...  ...data scientists, machine learning engineers, operations research...  ...ML, deep learning, reinforcement learning), causal inference... 
    Suggested
    Full time
    Temporary work
    Part time
    Remote work

    Walmart

    Sunnyvale, CA
    13 hours ago
  • $225k - $250k

    As a Staff Machine Learning Research Scientist you will work on some of the deepest problems in machine learning. You will publish papers, attend conferences, and file patents. But you will also ship code right into the products used every day. We call it applied research... 
    Suggested
    Contract work
    Remote work

    Primer AI

    San Francisco, CA
    13 hours ago
  • $154.31k - $192.89k

    As a Staff ML Research Scientist on theMachine Learning Safety R&D Team, you will join a small cross-functional group developing machine learning models...  ...architectures such as Transformers, VLMs/VLAs, and Deep Reinforcement Learning.Knowledge of state of the art work in... 
    Suggested
    Work at office
    Visa sponsorship

    Boston Dynamics

    Waltham, MA
    2 days ago
  • $251k - $310k

     ...Waymo AI Foundations Team Research Scientist Waymo is an autonomous driving technology company...  ...team is to develop machine learning solutions addressing open problems in...  ...we are currently focusing on include reinforcement learning, learning from demonstration,... 
    Suggested
    Full time
    Temporary work
    Remote work

    Latent Logic

    New York, NY
    3 days ago
  • $251k - $310k

     ...team is to develop machine learning solutions addressing open problems...  ...collaborations with other research teams in Alphabet. AI...  ...currently focusing on include reinforcement learning, learning from demonstration...  ...to a Principal Research Scientist . You Will Use Multimodal... 
    Full time
    Temporary work
    Remote work

    Waymo

    New York, NY
    4 days ago
  • $231.5k - $405.1k

     ...DescriptionAbout the team Our Core AI Research team develops novel...  .... About the role As a Staff Research Scientist, you will independently...  ...major workstream in agent learning and recursive self-improvement...  ...learning, deep learning, reinforcement learning, and... 
    Work at office
    Immediate start
    Remote work
    Flexible hours
    Shift work

    ServiceNow

    Santa Clara, CA
    1 day ago
  • $135k - $170k

     ...Opportunities with Mitsubishi Electric Research Laboratories A great place to...  ...is seeking a Robotics Research Scientist to conduct research in learning, perception, motion planning, manipulation...  ...in imitation learning, reinforcement learning Vision-Language-Action (... 
    Local area
    Home office

    Mitsubishi Electric Research Laboratories

    Cambridge, MA
    13 hours ago
  •  ...Who You Are Research Scientists at Phaidra lead our efforts in developing novel algorithmic architecture towards the end goal of bringing...  ...from a variety of disciplines including model-based reinforcement learning , planning and optimal control, deep learning, world... 
    Full time

    Phaidra

    Remote
    19 days ago
  • $151k - $297k

     ...It is backed by a strong team of AI researchers from Stanford, MIT, Berkeley, Princeton...  ....Position OverviewWe are seeking a Staff Research Scientist to join our team and contribute to the...  ...problems at the intersection of machine learning research and practical deployment of... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Palo Alto, CA
    2 days ago
  • $218.8k - $335.3k

     ...mobility and shape the future of autonomous transportation? As a Staff Research Scientist specializing in Vision-Language Models (VLMs), Vision-...  ....Influence technical roadmaps and shape strategic machine learning priorities that align with safety requirements, core product... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    3 days ago
  • $203.5k - $299.3k

     ...causal question.About the RoleWe are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash...  ...operate across functions with ML engineers, economists, data scientists, product managers, and business leaders.CompensationThe... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    Seattle, WA
    13 hours ago
  • $135k - $170k

     ...Opportunities with Mitsubishi Electric Research Laboratories A great place to work....  ...posted here as they become available. Staff - Research Scientist - Multiphysical Systems Mitsubishi...  ...estimation, and applications of machine learning. Successful candidates will be... 
    Local area
    Home office

    Mitsubishi Electric Research Laboratories

    Cambridge, MA
    4 days ago
  • $200k - $270k

     ...industry trailblazers redefining media buying with our Deep Learning Advertising Platform. Since 2015, we have harnessed the...  ...industry. Now, we’re growing! The Role We are looking for a Staff Research Scientist to join Cognitiv’s New Initiatives team— the group responsible... 
    Work at office
    Remote work
    Work from home

    Cognitiv Corp

    Bellevue, WA
    13 hours ago
  • $151k - $297k

     ...It is backed by a strong team of AI researchers from Stanford, MIT, Berkeley, Princeton...  .... Position Overview We are seeking a Staff Research Scientist to join our team and contribute to...  ...problems at the intersection of machine learning research and practical deployment of... 
    Work at office
    Local area
    Remote work
    Worldwide

    MongoDB

    Palo Alto, CA
    2 days ago
  •  ...Introduction to the role: The PFAS team is looking for a Staff Materials Research Scientist to serve as the scientific bridge between our...  ...with internal dataset, computational chemistry, machine-learning, and generative-modeling teams to keep property targets,... 
    Full time
    Flexible hours

    SandboxAQ

    Remote
    a month ago
  • $200k - $270k

     ...industry trailblazers redefining media buying with our Deep Learning Advertising Platform. Since 2015, we have harnessed the...  ...industry. Now, we're growing! The Role We are seeking a Staff Research Scientist who can drive innovation through deep technical expertise... 
    Work at office
    Remote work
    Work from home

    Cognitiv

    San Mateo, CA
    5 days ago
  • $230k - $400k

     ...a quickly growing group of committed researchers, engineers, policy experts, and business...  ...including: Complex multimodal reinforcement learning environments. High-performance RPC...  ...hybrid policy: Currently, we expect all staff to be in one of our offices at least 2... 
    Work experience placement
    Work at office
    Home office
    Visa sponsorship
    Relocation package
    Flexible hours

    TalentPros.AI

    San Francisco, CA
    16 days ago
  • Founding Research Scientist, Robot Learning The Mission GRAM is a self replication company creating populations of insectoids for the physical economy...  ...a consequential technical direction in robot learning, reinforcement learning, imitation learning, or embodied foundation... 
    Immediate start
    Relocation package

    GRAM

    Palo Alto, CA
    3 days ago
  •  ...sacrificing craft or quality. We’re looking for a Senior Staff Machine Learning Scientist to help us solve challenging problems to address emerging...  ...Learning Scientist, you’ll:  Lead and drive ambitious research initiatives that advance the state of the art in computer... 
    Ongoing contract
    Permanent employment
    Full time
    Temporary work
    Fixed term contract
    Remote work
    Flexible hours

    Webflow

    Remote
    10 days ago
  •  ...company in the world. The Data Science team is seeking a Staff Machine Learning Scientist to lead and advance the development of personalization products...  ...Science, Machine Learning, Engineering, Operations Research, Statistics, or a related quantitative field, OR Master's... 
    Full time
    Temporary work
    Flexible hours

    Bertelsmann-Jobs

    Remote
    more than 2 months ago
  • $180k - $220k

     ...world. The Data Science team is seeking an experienced Machine Learning Scientist to drive business-critical forecasting products. We have a...  ...effectively. This role may be filled at the Senior or Staff level depending on experience and interview performance. Location... 
    Full time
    Remote work

    Bertelsmann-Jobs

    Remote
    a month ago
  • $131.5k - $219.1k

     ...looking for a Senior Applied Machine Learning Scientist to lead the development of advanced AI...  ...You'll Support the MissionLead applied research and development for AI/ML systems across...  ....Experience with deep learning, reinforcement learning, multi-armed bandits, causal... 
    Full time
    Part time

    Washington Post

    Washington DC
    4 days ago
  •  ...distributed systems at large scale and low latencies. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems and Auction...  ...About the role In this role you will work on applying SOTA research and conduct your own research to develop novel... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    Austin, TX
    2 days ago
  • $147.8k - $274.4k

     ...discovery and development. Roche’s Research and Early Development organisations...  ...Artificial Intelligence (AI) to assist our scientists in both pRED and gRED to deliver...  ...modeling, representation learning, LLMs, and/or reinforcement learning, with direct applications... 
    Full time
    Local area
    Worldwide
    Relocation package

    Genentech

    South San Francisco, CA
    3 days ago
  • $159.8k - $244.3k

    Job DescriptionWe are seeking a Senior Machine Learning and Artificial Intelligence Scientist to lead the development and production deployment of advanced...  ..., optimization, causal inference, simulation, reinforcement learning, recommender systems, NLP, computer vision,... 
    Full time
    Local area
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Austin, TX
    4 days ago
  • $159.75k - $255.6k

     ...at a company where you matter.Your ImpactWe are seeking a skilled and innovative Senior Research Scientist to join one of our AI teams focusing on computer vision and machine learning (CVML). As a research scientist at Axon you will play a crucial role in developing AI... 
    Work experience placement
    Work at office
    Remote work

    Axon

    Seattle, WA
    4 days ago
  •  ...Research Scientist ISEE is seeking full-time Research Scientists to join our team. The ideal candidate has several years of research/work...  ...in the areas of: Robotics, Artificial Intelligence, Machine Learning, Perception, Modeling, Simulation, Applied Mathematics,... 
    Full time
    Work experience placement
    Remote work

    ISEE

    United States
    4 days ago
  •  ...GroupSolver Market Research Tech Startup GroupSolver is a market research tech startup based in San Diego with offices in Utah and...  ...Basic Qualifications ~ MS/PhD in Computer Science, Machine Learning, Statistics, Applied Mathematics, or a related field. ~2+ years... 
    Remote work

    GROUPSOLVER INC

    United States
    4 days ago
  •  ...an innovative and hands-on Staff AI Scientist to join the Intuit AI team....  ...AI scientists and machine learning engineers and build models...  ...Learning, Bayesian Learning, Reinforcement Learning, or Deep Learning)...  ...DS solutions. Proactively researches, explores, and enables new... 
    Worldwide

    Intuit

    Atlanta, GA
    13 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Research Scientist - Reinforcement Learning. Be the first to apply!