Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Scientist - RL Training

$200k - $350k

Snorkel AI

About SnorkelAt Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data.We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler!ABOUT THE ROLE We're looking for a Research Scientist to work on reinforcement learning for training and aligning large language models. This is a foundational research role focused on one of the most consequential open data problems in AI: how to generate the data, reward signals, and training procedures that steer LLM behavior in reliable and generalizable directions — and a core capability that directly differentiates Snorkel's data-as-a-service offering. You'll work closely with Snorkel's research, engineering, and delivery teams to advance our RL data capabilities — translating research ideas into the preference datasets, reward models, and RL-ready corpora we produce for frontier AI labs, and contributing to a research agenda that is central to Snorkel's long-term differentiation as a provider of bespoke training data. MAIN RESPONSIBILITIES Research and implement reinforcement learning techniques — including GRPO, RLHF, RLAIF, DPO, and reward modeling — and translate them into data products (preference datasets, reward signals, verifiable rewards) that customers can use to train and fine-tune large language models. Design and build data pipelines that generate high-quality training signal for RL workflows, including AI-assisted data annotation and curation data pipelines to improve model generalization to unseen benchmarks . Prototype and iterate on end-to-end RL training recipes that inform what data Snorkel ships as part of its data-as-a-service deliveries. Work closely with research scientists, ML engineers, and delivery teams to translate RL research into customer-ready data products.Stay current with the latest developments in large-scale muli-node LLM training, alignment research, and scalable RL methods (on complex environments such as Terminal-Bench), bringing relevant advances into Snorkel's data-as-a-service approach.Contribute to Snorkel's research publications and internal knowledge base in RL and model training.PREFERRED QUALIFICATIONS Deep expertise in reinforcement learning from human or AI feedback, reward modeling and credit attribution ideally with a clear perspective on what data makes these techniques work. Experience training or fine-tuning 30B+ large language models at scale, including familiarity with distributed training infrastructure. Strong proficiency in Python and ML frameworks, especially PyTorch and HuggingFace and hands-on experience with RL frameworks such as Verl and SkyRL. Solid software engineering fundamentals — you can build research prototypes that others can run, extend, and integrate into data production workflows. Familiarity with ML infrastructure and cloud platforms and tools (AWS, GCP, Kubernetes, Slurm, etc.); experience with large-scale RL training pipelines a strong plus. Comfort operating in a high-iteration environment with open-ended research questions and shifting, customer-driven technical constraints. Ph.D. in machine learning, reinforcement learning, or a related field strongly preferred; exceptional industry experience considered. Actual compensation will be determined based on factors including skills, qualifications, experience, and geographic location.Salary range(s) for this role$200,000—$350,000 USDBe Your Best at SnorkelJoining Snorkel AI means becoming part of a company that has market proven solutions, robust funding, and is scaling rapidly—offering a unique combination of stability and the excitement of high growth. As a member of our team, you’ll have meaningful opportunities to shape priorities and initiatives, influence key strategic decisions, and directly impact our ongoing success. Whether you’re looking to deepen your technical expertise, explore leadership opportunities, or learn new skills across multiple functions, you’re fully supported in building your career in an environment designed for growth, learning, and shared success.Snorkel AI is proud to be an Equal Employment Opportunity employer and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. Snorkel AI embraces diversity and provides equal employment opportunities to all employees and applicants for employment. Snorkel AI prohibits discrimination and harassment of any type on the basis of race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local law. All employment is decided on the basis of qualifications, performance, merit, and business need.We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Research Scientist - RL Training in Redwood City, CA vacancy
  • $204k - $259k

     ...foster collaborations with other research teams in Alphabet. AI...  ...you will report to a Principal Scientist. You will: Participate in Waymo...  ...’s Foundation World Model post-training and evaluation Research and develop cutting edge RL and Distillation techniques for... 
    Training
    Full time
    Remote work

    Waymo

    Mountain View, CA
    13 hours ago
  •  ...Develop autonomous research systems that ideate, experiment, and publish scientific papers. You'll be working on the first and best...  ...development team to build and deploy production‑ready research. RL post‑train and fine‑tune reasoning models to automate components of the... 
    Training
    Full time
    Flexible hours

    Autoscience Institute

    Menlo Park, CA
    2 days ago
  •  ...working across three frontier research institutes with leading...  ...biological foundation models: Trained on some of the largest GPU clusters...  ...discovery, including RL, reward modeling, reasoning,...  ...systems in collaboration with scientists, grounded in real biological... 
    Training

    Biohub

    Redwood City, CA
    4 days ago
  •  ...Google DeepMind, we’re a team of scientists, engineers, machine learning...  ...models. The role of the Research Scientist / Research Engineer...  ...Search.Key responsibilities:Post-training / instruction tuning state of...  ...at NeurIPS, ICLR, ICML, RL/DL, EMNLP, AAAI, UAIExperience... 
    Training

    DeepMind

    Mountain View, CA
    5 days ago
  • $180k - $300k

     ...are what they eat. But a large portion of training compute is wasted training on data that...  ...For more details, check out our recent research on synthetic data scaling (BeyondWeb) and...  ...the RoleWe're looking for a Research Scientist to investigate how intervening on training... 
    Training
    Work at office
    Work from home
    Relocation package

    DatologyAI

    Redwood City, CA
    1 day ago
  • $112.88k - $149.57k

     ...International is seeking multiple Applied and Computational Scientists to conduct research and development in the areas of mathematics of generative...  ...such as Python, C/C++, and libraries associated with training of AI models.Excellent written, oral, and interpersonal communication... 
    Training
    Permanent employment

    SRI International

    Menlo Park, CA
    1 day ago
  • $180k - $300k

     ...are what they eat. But a large portion of training compute is wasted training on data that...  ...For more details, check out our recent research on synthetic data scaling (BeyondWeb) and...  ...the RoleWe’re looking for a Research Scientist to lead work on post-training data curation... 
    Training
    Work at office
    Work from home
    Relocation package

    DatologyAI

    Redwood City, CA
    6 days ago
  • $117.2k - $313.7k

     ...Salesforce.The ExperienceSalesforce AI Research is looking for outstanding AI Research Scientists and Research Engineers. Our team...  ...agents, reinforcement learning (RL), reasoning and planning,...  ...processing.Core Modeling and Post-Training: Machine learning methodology, pre... 
    Training
    Full time

    Salesforce

    Palo Alto, CA
    1 day ago
  • $193.93k - $291.15k

     ...controllable agents to enable effective closed-loop training in simulation.If you are passionate about...  ...new problems, leading impactful research, and seeing your work deployed onto real...  ...training via Reinforcement Learning (RL).Mitigate accumulated uncertainties across... 
    Training
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    6 days ago
  •  ...message the job poster from GreyOrange Overview Title: Operations Research Scientist Location: Hybrid, Redwood City, CA Range: 200-220k About...  ..., recall, transfer, leaves of absence, compensation, and training. Seniority level Mid-Senior level Employment type Full-time... 
    Training
    Full time
    Local area
    Worldwide

    GreyOrange

    Redwood City, CA
    2 days ago
  • $115k - $130k

     ...The Research Scientist provides broad and in-depth scientific support of Translational Research, Clinical, and Medical Affairs activities....  ...under Research Grant Agreements. Key Requirements Education and Training: Advanced degree (MS or higher) in biological sciences or... 
    Training
    Full time
    Temporary work
    Local area
    Flexible hours

    Socket

    Redwood City, CA
    13 hours ago
  • $150k - $200k

     ...language models. Responsibilities As a Research Scientist/ ML Engineer, you will play a crucial role...  ...learning systems, such as distributed training of large models and/or ML performance...  ...state-of-the-art research on Alignment and RL. We help companies build AI systems... 
    Training
    Full time

    Collinear AI

    Mountain View, CA
    3 days ago
  •  ...enjoy multi-year runway. About The Role We’re looking for an AI Research Scientist to advance the methodological frontier of AI in healthcare....  ...strategy. Your work may include novel architectures, new training or evaluation techniques, long-horizon research bets, peer-reviewed... 
    Training
    Temporary work
    Work at office
    Monday to Friday
    Monday to Thursday
    Flexible hours

    Sprinter Health

    Menlo Park, CA
    3 days ago
  •  ...more than 150 PhDs and data scientists, along with more than 4,000 AI...  ...contextual, multilingual, pre-trained datasets; fine-tuned,...  ...engineering standards Mentor researchers and engineers; drive technical...  ...equivalent) 5+ years hands-on RL — environment design, reward... 
    Training

    Centific

    Palo Alto, CA
    1 day ago
  •  ...interpretation of results. Propose and provide input for the design of subsequent experiments.• May act as test method owner capable of training new hires• May act as study lead for test protocols• Write lab procedures, reports, and/or SOPs.• Work according to appropriate... 
    Training
    Immediate start

    Artech

    San Carlos, CA
    1 day ago
  • $174k - $253k

     ...content.Design and implement advanced pre-training, supervised fine tuning (SFT), and...  ...experience.2 years of experience leading a research agenda.Preferred qualifications:2 years...  ...influencing other researchers.As a Research Scientist, you will be responsible for improving... 
    Training

    Google

    Mountain View, CA
    1 day ago
  •  ...safe, autonomous, clinical conversations with patients. We have trained our own LLMs as part of our Polaris constellation, resulting...  ...and a team of physicians, hospital leaders, AI pioneers, and researchers from institutions like El Camino Health, Johns Hopkins, Washington... 
    Training
    Work at office

    Hippocratic AI

    Palo Alto, CA
    1 day ago
  • $185k - $400k

     ...generation and intelligent agentic platforms. We are looking for a staff or lead-level Research Engineer, Data to architect and scale data engineering systems supporting model training for our advanced multimodal foundation models. This pivotal role will strengthen our... 
    Training
    Remote work

    Pika

    Palo Alto, CA
    3 days ago
  • $190k - $238k

     ...signaling pathway.The Opportunity:As a Principal Scientist in the Cancer Pharmacology team within the Translational Research group, in the Biomedical Discovery Research...  ...experience, market dynamics, and relevant education or training.Please note that base pay salary range is one... 
    Training
    Full time
    Local area

    Revolution Medicines

    Redwood City, CA
    13 hours ago
  • $183.83k - $275.98k

     ...Behavior team, leveraging the cutting edge of machine learning research to solve challenging real-world robotics problems. This role is...  ...data and evaluation teams to build effective and efficient data, training, evaluation, and validation pipelines.Develop practical,... 
    Training
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    1 day ago
  • $147k - $211k

    SnapshotWe are seeking strong Research Scientists with expertise in AI research and experience in interdisciplinary sociotechnical modeling...  ...with modern prompting strategiesExperience finetuning and post-training LLMs using RLExperience with developing agentic AI solutions... 
    Training
    Full time

    DeepMind

    Mountain View, CA
    1 day ago
  • $165k - $238k

     ...that have the aspiration and riskiness of research with the speed and ambition of a startup...  ...: About the role:As a Senior Research Scientist you will be a key architect of the core...  ...to optimize system performanceExperience training, fine-tuning, or distilling models for specialized... 
    Training
    Full time
    Work at office
    3 days per week

    X Company

    Mountain View, CA
    1 day ago
  • $112k - $168k

    Scientist II, Field Applications Bioinformatics SupportPacBio (NASDAQ: PACB) is a premier...  ...solutions to help scientists and clinical researchers resolve genetically complex problems....  ...this role, they will be responsible to train customers for PacBio data analysis, support... 
    Training
    Full time
    Work at office
    Home office

    Pacific Biosciences

    Menlo Park, CA
    6 days ago
  • $132k - $166k

     ...pay salary is determined by multiple factors, including job-related skills, experience, market dynamics, and relevant education or training.Please note that base pay salary range is one part of the overall total rewards program at RevMed, which includes competitive cash... 
    Training
    Full time
    Local area
    Shift work

    Revolution Medicines

    Redwood City, CA
    6 days ago
  • $132k - $166k

     ...Opportunity:Revolution Medicines is seeking Scientist II to join the Drug Product team within...  ...supervision to conduct appropriate research and troubleshoot technical challenges, escalating...  ...dynamics, and relevant education or training.Please note that base pay salary range... 
    Training
    Full time
    Local area

    Revolution Medicines

    Redwood City, CA
    6 days ago
  • $170k - $212k

     ...Substance function, the position will be responsible for process research and development and scale-up of manufacturing in support of...  ...skills, experience, market dynamics, and relevant education or training.Please note that base pay salary range is one part of the overall... 
    Training
    Full time
    Work experience placement
    Local area

    Revolution Medicines

    Redwood City, CA
    3 days ago
  • $147k - $210k

    Author research papers to share and generate impact of research results across the team and...  ...specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy...  ..., experience, and relevant education or training. US: $147000 - $210000 (USD) + 15% bonus... 
    Training

    Google

    Mountain View, CA
    13 hours ago
  • $120k - $150k

     ...in the RAS signaling pathway.The Opportunity:The Senior Safety Scientist within Global Patient Safety Science is an individual...  ...skills, experience, market dynamics, and relevant education or training.Please note that base pay salary range is one part of the overall... 
    Training
    Full time
    Local area

    Revolution Medicines

    Redwood City, CA
    13 hours ago
  • $207k - $301k

     ...personalization, and ML efficiency. Drive new research ideas from conception, experimentation,...  ...of research experience (e.g., model training/deployment, efficiency optimization,...  ...emphasize specific types of work. As a Research Scientist, you'll setup large-scale tests and... 
    Training
    Shift work

    Google

    Mountain View, CA
    3 days ago
  • $251k - $295k

     ...We are seeking a Senior Machine Learning Scientist to help accelerate drug discovery...  ...datasets into actionable insights that guide research decisions.The Senior Machine Learning Scientist...  ...dynamics, and relevant education or training.Please note that base pay salary range... 
    Training
    Full time
    Local area

    Revolution Medicines

    Redwood City, CA
    13 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Scientist - RL Training. Be the first to apply!