Staff Research Scientist - Reinforcement Learning
$200k - $250kCentific
Role Description
What You'll Do:
- Design simulation environments and digital twins for enterprise workflows
- Post-train LLM agents using RLHF, DPO, GRPO, PPO, and emerging methods
- Build pipelines that convert human-labeled traces and verifiable signals into training data
- Architect multi-turn, tool-using agents with closed learning loops
- Design reward functions and verifiers that resist reward hacking and reflect real task outcomes
- Set the technical bar across the team — architecture, code review, engineering standards
- Mentor researchers and engineers; drive technical direction through influence
- Translate research into production; contribute to publications
Qualifications
- 7+ years in ML/AI research or engineering; 3+ years at senior/staff level
- MS or PhD in Computer Science, Machine Learning, or related field (or equivalent)
- 5+ years hands-on RL — environment design, reward engineering, policy optimization — with at least one production deployment
- 3+ years fine-tuning LLMs with hands-on RL post-training (RLHF, DPO, GRPO, PPO)
- Expert-level implementation of RLHF pipelines, reward modeling (Bradley-Terry), DPO, and KTO
- Working knowledge of modern post-training and rollout-serving libraries (TRL, veRL, OpenRLHF, SkyRL)
- Experience building LLM-based agents: tool use, multi-turn reasoning, trajectory evaluation
- Strong Python and software engineering skills — comfortable building production pipelines, not just notebooks
- Deep expertise in MDPs, policy gradient methods (PPO, SAC), and temporal difference learning
- Hands-on experience with Gymnasium-based environments and reward engineering (sparse vs. dense)
Preferred Qualifications
- Publications at NeurIPS, ICML, ICLR, ACL, COLM, or similar venues
- Open-source contributions to post-training or agent frameworks (TRL, veRL, OpenRLHF, SkyRL)
- Experience with Offline RL (CQL, IQL), Model-based RL / World Models, or Hierarchical RL
- Background in synthetic data generation, simulation, or world models
- Domain experience in healthcare, finance, logistics, or compliance
- Distributed training on GPU clusters
Benefits
- Lead the frontier. Shape a new discipline at the intersection of post-training, simulation, and enterprise AI.
- Ship your science. See your research power real systems across healthcare, finance, and safety-critical operations.
- Collaborate with leaders. Work alongside NVIDIA, Microsoft, and the global AI community.
- Build what matters. Create governed, compliant AI systems enterprises can actually trust.
- Salary: $200k-$250k
How to Apply
Send your CV, a description of a technically complex system you personally built or led, and (if applicable) your publication list or open-source contributions to:
View email address on us.fitly.work
Subject: Senior Staff Research Scientist – RL
Company Description
Centific is a frontier AI data foundry that curates diverse, high-quality data, using our purpose-built technology platforms to empower the Magnificent Seven and our enterprise clients with safe, scalable AI deployment. Our team includes more than 150 PhDs and data scientists, along with more than 4,000 AI practitioners and engineers. We harness the power of an integrated solution ecosystem—comprising industry-leading partnerships and 1.8 million vertical domain experts in more than 230 markets—to create contextual, multilingual, pre-trained datasets; fine-tuned, industry-specific LLMs; and RAG pipelines supported by vector databases. Our zero-distance innovation™ solutions for GenAI can reduce GenAI costs by up to 80% and bring solutions to market 50% faster.
Our mission is to bridge the gap between AI creators and industry leaders by bringing best practices in GenAI to unicorn innovators and enterprise customers. We aim to help these organizations unlock significant business value by deploying GenAI at scale, helping to ensure they stay at the forefront of technological advancement and maintain a competitive edge in their respective markets.
$94.49k - $147.4k
...Leadership Computing Facility (ALCF) is seeking a Staff Scientist in Post-Training and Reinforcement Learning for AI for Science to help advance the next generation... ...workflows.The successful candidate will conduct research on methods that improve the usefulness,...SuggestedFull timeFor contractorsRemote work$110k - $220k
...Science team is looking for a Staff Data Scientist to work alongside our team... ...Team: The Applied AI team researches and develops advanced... ...data scientists, machine learning engineers, operations research... ...ML, deep learning, reinforcement learning), causal inference...SuggestedFull timeTemporary workPart timeRemote work$225k - $250k
As a Staff Machine Learning Research Scientist you will work on some of the deepest problems in machine learning. You will publish papers, attend conferences, and file patents. But you will also ship code right into the products used every day. We call it applied research...SuggestedContract workRemote work$154.31k - $192.89k
As a Staff ML Research Scientist on theMachine Learning Safety R&D Team, you will join a small cross-functional group developing machine learning models... ...architectures such as Transformers, VLMs/VLAs, and Deep Reinforcement Learning.Knowledge of state of the art work in...SuggestedWork at officeVisa sponsorship$251k - $310k
...Waymo AI Foundations Team Research Scientist Waymo is an autonomous driving technology company... ...team is to develop machine learning solutions addressing open problems in... ...we are currently focusing on include reinforcement learning, learning from demonstration,...SuggestedFull timeTemporary workRemote work$251k - $310k
...team is to develop machine learning solutions addressing open problems... ...collaborations with other research teams in Alphabet. AI... ...currently focusing on include reinforcement learning, learning from demonstration... ...to a Principal Research Scientist . You Will Use Multimodal...Full timeTemporary workRemote work$231.5k - $405.1k
...DescriptionAbout the team Our Core AI Research team develops novel... .... About the role As a Staff Research Scientist, you will independently... ...major workstream in agent learning and recursive self-improvement... ...learning, deep learning, reinforcement learning, and...Work at officeImmediate startRemote workFlexible hoursShift work$135k - $170k
...Opportunities with Mitsubishi Electric Research Laboratories A great place to... ...is seeking a Robotics Research Scientist to conduct research in learning, perception, motion planning, manipulation... ...in imitation learning, reinforcement learning Vision-Language-Action (...Local areaHome office- ...Who You Are Research Scientists at Phaidra lead our efforts in developing novel algorithmic architecture towards the end goal of bringing... ...from a variety of disciplines including model-based reinforcement learning , planning and optimal control, deep learning, world...Full time
$151k - $297k
...It is backed by a strong team of AI researchers from Stanford, MIT, Berkeley, Princeton... ....Position OverviewWe are seeking a Staff Research Scientist to join our team and contribute to the... ...problems at the intersection of machine learning research and practical deployment of...Work at officeLocal areaRemote workWorldwideFlexible hours$218.8k - $335.3k
...mobility and shape the future of autonomous transportation? As a Staff Research Scientist specializing in Vision-Language Models (VLMs), Vision-... ....Influence technical roadmaps and shape strategic machine learning priorities that align with safety requirements, core product...Full timeLocal areaRemote workWork from homeRelocation packageFlexible hours$203.5k - $299.3k
...causal question.About the RoleWe are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash... ...operate across functions with ML engineers, economists, data scientists, product managers, and business leaders.CompensationThe...Hourly payWork at officeLocal areaRemote workFlexible hours$135k - $170k
...Opportunities with Mitsubishi Electric Research Laboratories A great place to work.... ...posted here as they become available. Staff - Research Scientist - Multiphysical Systems Mitsubishi... ...estimation, and applications of machine learning. Successful candidates will be...Local areaHome office$200k - $270k
...industry trailblazers redefining media buying with our Deep Learning Advertising Platform. Since 2015, we have harnessed the... ...industry. Now, we’re growing! The Role We are looking for a Staff Research Scientist to join Cognitiv’s New Initiatives team— the group responsible...Work at officeRemote workWork from home$151k - $297k
...It is backed by a strong team of AI researchers from Stanford, MIT, Berkeley, Princeton... .... Position Overview We are seeking a Staff Research Scientist to join our team and contribute to... ...problems at the intersection of machine learning research and practical deployment of...Work at officeLocal areaRemote workWorldwide- ...Introduction to the role: The PFAS team is looking for a Staff Materials Research Scientist to serve as the scientific bridge between our... ...with internal dataset, computational chemistry, machine-learning, and generative-modeling teams to keep property targets,...Full timeFlexible hours
$200k - $270k
...industry trailblazers redefining media buying with our Deep Learning Advertising Platform. Since 2015, we have harnessed the... ...industry. Now, we're growing! The Role We are seeking a Staff Research Scientist who can drive innovation through deep technical expertise...Work at officeRemote workWork from home$230k - $400k
...a quickly growing group of committed researchers, engineers, policy experts, and business... ...including: Complex multimodal reinforcement learning environments. High-performance RPC... ...hybrid policy: Currently, we expect all staff to be in one of our offices at least 2...Work experience placementWork at officeHome officeVisa sponsorshipRelocation packageFlexible hours- Founding Research Scientist, Robot Learning The Mission GRAM is a self replication company creating populations of insectoids for the physical economy... ...a consequential technical direction in robot learning, reinforcement learning, imitation learning, or embodied foundation...Immediate startRelocation package
- ...sacrificing craft or quality. We’re looking for a Senior Staff Machine Learning Scientist to help us solve challenging problems to address emerging... ...Learning Scientist, you’ll: Lead and drive ambitious research initiatives that advance the state of the art in computer...Ongoing contractPermanent employmentFull timeTemporary workFixed term contractRemote workFlexible hours
- ...company in the world. The Data Science team is seeking a Staff Machine Learning Scientist to lead and advance the development of personalization products... ...Science, Machine Learning, Engineering, Operations Research, Statistics, or a related quantitative field, OR Master's...Full timeTemporary workFlexible hours
$180k - $220k
...world. The Data Science team is seeking an experienced Machine Learning Scientist to drive business-critical forecasting products. We have a... ...effectively. This role may be filled at the Senior or Staff level depending on experience and interview performance. Location...Full timeRemote work$131.5k - $219.1k
...looking for a Senior Applied Machine Learning Scientist to lead the development of advanced AI... ...You'll Support the MissionLead applied research and development for AI/ML systems across... ....Experience with deep learning, reinforcement learning, multi-armed bandits, causal...Full timePart time- ...distributed systems at large scale and low latencies. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems and Auction... ...About the role In this role you will work on applying SOTA research and conduct your own research to develop novel...Work at officeLocal areaRemote workMonday to ThursdayFlexible hours
$147.8k - $274.4k
...discovery and development. Roche’s Research and Early Development organisations... ...Artificial Intelligence (AI) to assist our scientists in both pRED and gRED to deliver... ...modeling, representation learning, LLMs, and/or reinforcement learning, with direct applications...Full timeLocal areaWorldwideRelocation package$159.8k - $244.3k
Job DescriptionWe are seeking a Senior Machine Learning and Artificial Intelligence Scientist to lead the development and production deployment of advanced... ..., optimization, causal inference, simulation, reinforcement learning, recommender systems, NLP, computer vision,...Full timeLocal areaWork from homeRelocation packageFlexible hours$159.75k - $255.6k
...at a company where you matter.Your ImpactWe are seeking a skilled and innovative Senior Research Scientist to join one of our AI teams focusing on computer vision and machine learning (CVML). As a research scientist at Axon you will play a crucial role in developing AI...Work experience placementWork at officeRemote work- ...Research Scientist ISEE is seeking full-time Research Scientists to join our team. The ideal candidate has several years of research/work... ...in the areas of: Robotics, Artificial Intelligence, Machine Learning, Perception, Modeling, Simulation, Applied Mathematics,...Full timeWork experience placementRemote work
- ...GroupSolver Market Research Tech Startup GroupSolver is a market research tech startup based in San Diego with offices in Utah and... ...Basic Qualifications ~ MS/PhD in Computer Science, Machine Learning, Statistics, Applied Mathematics, or a related field. ~2+ years...Remote work
- ...an innovative and hands-on Staff AI Scientist to join the Intuit AI team.... ...AI scientists and machine learning engineers and build models... ...Learning, Bayesian Learning, Reinforcement Learning, or Deep Learning)... ...DS solutions. Proactively researches, explores, and enables new...Worldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Research Scientist - Reinforcement Learning. Be the first to apply!



