Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

RL Research Scientist: Training & Aligning LLMs with Data

Remote Jobs

Snorkel AI in Redwood City is seeking a Research Scientist to advance reinforcement learning methods for training and aligning large language models. You will build data products, reward models, and end-to-end RL pipelines in collaboration with research, engineering, and delivery teams. The role focuses on developing data-as-a-service capabilities, translating research into practical datasets, signals, and corpora for frontier AI labs, while contributing to Snorkel's long-term differentiation. #J-18808-Ljbffr Remote Jobs

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the RL Research Scientist: Training & Aligning LLMs with Data in Redwood City, CA vacancy
  •  ...San Francisco seeks a Research Scientist to work on reinforcement learning for training and aligning large language models....  ...foundational role focuses on data, reward signals, and...  ...You will collaborate with Snorkel’s research,...  ...teams to advance RL data capabilities, translating... 
    Data
    Training

    Snorkel AI

    Redwood City, CA
    2 days ago
  •  ...California is seeking a Research Scientist to advance reinforcement learning for training and aligning large language models. You...  ...research ideas into practical data products and collaborate with research, engineering,...  ...teams to deliver RL‑ready corpora for client... 
    Data
    Training

    Snorkel AI, Inc.

    Redwood City, CA
    6 days ago
  • $200k - $350k

     ...AI doesn’t start with the model, it starts with the data.We’re on a mission...  ...Snorkel started as a research project in the...  ...organizations to empower scientists, engineers,...  ...reinforcement learning for training and aligning large language...  ...to advance our RL data capabilities... 
    Data
    Training
    Local area
    Shift work

    Snorkel AI

    Redwood City, CA
    3 days ago
  • $147k - $211k

    In accordance with Washington state law...  ...Experience conducting research and development,...  ...-tuning and post-training LLMs using Reinforcement Learning (RL) or other alignment methods....  ...information from different data types (e.g.,...  ...work. As a Research Scientist, you'll setup... 
    Data
    Training
    Temporary work

    Google DeepMind

    Mountain View, CA
    3 days ago
  • $147k - $211k

    SnapshotWe are seeking strong Research Scientists with expertise in AI research...  ...strong awareness of the AI alignment and safety landscape, and a...  ...strategiesExperience finetuning and post-training LLMs using RLExperience with...  ...information from different data types (e.g., vision, audio,... 
    Data
    Training
    Full time

    DeepMind

    Mountain View, CA
    2 days ago
  • $204k - $259k

     ...technology company with the mission to...  ...with other research teams in Alphabet...  ...to a Principal Scientist. You will:...  ...World Model post-training and evaluation...  ...develop cutting edge RL and...  ...manner such as Data parallel, FSDP...  ...policy learning or alignment with human preferences... 
    Data
    Training
    Full time
    Remote work

    Waymo

    Mountain View, CA
    6 hours ago
  •  ...lab capabilities with the ultimate...  ...three frontier research institutes with...  ...foundation models: Trained on some of the largest...  ...frontier LLMs with biological...  ...weights. Scientific data at unprecedented...  ...discovery, including RL, reward modeling...  ...with scientists, grounded in real... 
    Data
    Training

    Biohub

    Redwood City, CA
    5 days ago
  •  ...we’re a team of scientists, engineers,...  ...and collaborate with others on critical...  ...role of the Research Scientist / Research...  ...and develop data and algorithmic...  ...responsibilities:Post-training / instruction...  ...of the art LLMs, focusing on...  ...NeurIPS, ICLR, ICML, RL/DL, EMNLP, AAAI... 
    Data
    Training

    DeepMind

    Mountain View, CA
    1 day ago
  •  ...those locations, with priority determined...  ...What you’ll do As a Research Scientist at Simular, you...  ...interaction, and alignment (e.g. reward modeling...  ...end-to-end: from data collection and benchmarking...  ..., to model training and evaluation....  ...training, or fine-tuning LLMs/VLMs... 
    Data
    Training

    Simular

    Palo Alto, CA
    2 days ago
  • $138.23k - $207.58k

     ...Machine Learning Research Scientist: Generative Models...  ...combines AD hardware with our generalized AI...  ...improve through data, the Nuro Driver is...  .... This includes alignment with reward/cost models...  ...with closed-loop RL, etc. Research...  ...Scientist Training Program San Jose... 
    Data
    Training
    Full time
    Work at office

    Nuro

    Mountain View, CA
    4 hours ago
  • $156.56k - $180.04k

     ...Research Scientist - Interpretability (1 Year Fixed Term...  ...key resource for aligning artificial intelligence models with human-like neural...  ...dimensional neural data ~ Collaborate with...  ...experience in training, fine-tuning, and...  ...and quantization of LLMs or multimodal models... 
    Data
    Training
    Full time
    Fixed term contract
    Flexible hours

    Stanford University School of Medicine

    Palo Alto, CA
    4 hours ago
  • $209.5k - $283.5k

     ...communities we serve. With approximately 100...  ...hiring a Staff AI Research Scientist to join the Intuit...  ...and model training (pre and post) , and...  ...of domain experts, data scientists, and machine...  ...that enable LLMs to reason over structured...  ...goals that are in alignment with Intuit... 
    Data
    Training
    Work experience placement
    Worldwide

    Intuit

    Mountain View, CA
    2 days ago
  • $211.2k - $257.5k

     ...biomedical imaging with molecular and language data. By...  ...accessible to our global research community. Responsibilities...  ..., code, and train large-scale...  ..., Multimodal LLMs) that can embed...  ...with experimental scientists and software...  ...Multimodal Fusion (e.g., aligning image embeddings... 
    Data
    Training
    Local area

    Dormont Manufacturing Co

    Redwood City, CA
    5 days ago
  • $141k - $244k

    Research Scientist - Gemini Personal Intelligence Mountain...  ...research behind Gemini, with a mission to make AI...  ...Language Models (LLMs) to build the brain...  ...with users' personal data to solve real-world...  ...Driving research on post‑training techniques (e.g., RL, SFT, and preference... 
    Data
    Training
    Full time
    Relocation
    Flexible hours
    Shift work

    DeepMind

    Mountain View, CA
    4 days ago
  •  ...own the metrics that define success. It fits a researcher with deep generative-modeling or model-based-RL expertise who has trained models to the limits of a multi-node cluster....  ...Run scaling studies that show where compute, data, and architecture pay off. Publish at the... 
    Data
    Training

    Luma

    Redwood City, CA
    6 days ago
  • $128.25k - $266.88k

     ...consumer inbox with hundreds of millions...  ...messages and data at the petabyte...  ...team of Researchers, Engineers, Product...  ...in designing, training, and evaluating...  ...level research scientists and machine learning...  ...performance, alignment, and governance...  ...models (LLMs) for search, retrieval... 
    Data
    Training
    Work at office
    Flexible hours

    Yahoo Holdings Inc.

    Mountain View, CA
    3 days ago
  • $150k - $200k

     ...experience — talk with your recruiter to...  ...evaluation benchmarks and alignment techniques for...  ...As a Research Scientist/ ML Engineer, you...  ...reliability, and data curation. Your technical...  ...such as distributed training of large models and...  ...on Alignment and RL. We help companies... 
    Data
    Training
    Full time

    Collinear AI

    Mountain View, CA
    4 hours ago
  • $174k - $252k

    Scope and drive research efforts to improve complex frontier...  ....Curate and generate data to evaluate and...  ...development.Experience with LLM training and generative models....  ...of work. As a Research Scientist, you'll setup large-...  ...foundational models. Modern LLMs are required to... 
    Data
    Training

    Google

    Mountain View, CA
    2 days ago
  •  ...autonomous, clinical conversations with patients. We have trained our own LLMs as part of our Polaris...  ...hospital leaders, AI pioneers, and researchers from institutions like El Camino Health...  ...DoDesign, Develop, Evaluate and update data-driven models for Speech First applications... 
    Data
    Training
    Work at office

    Hippocratic AI

    Palo Alto, CA
    2 days ago
  • $129.5k - $166k

    Research Specialist / Scientist, Metabolic Phenotyping (In Vitro & In Vivo...  ...that make the team's data rigorous and reproducible...  ...in collaboration with the team, including animal...  ...discovery team to align in vitro mechanism-of...  ..., compensation and training. Thank you for your... 
    Data
    Training
    Contract work
    Local area

    Altos Labs

    Redwood City, CA
    6 days ago
  • $180k - $300k

     .... But a large portion of training compute is wasted training on data that are already learned,...  ...and allow smaller models with fewer than half the parameters...  ..., check out our recent research on synthetic data scaling...  ...looking for a Research Scientist to lead work on post-... 
    Data
    Training
    Work at office
    Work from home
    Relocation package

    DatologyAI

    Redwood City, CA
    2 days ago
  • $185k - $400k

     ...looking for a staff or lead-level Research Engineer, Data to architect and scale data...  ...systems supporting model training for our advanced multimodal...  ..., and video datasetsPartner with research and engineering...  ...engineering and ML data curation for LLMs, VLMs, or other large-scale... 
    Data
    Training
    Remote work

    Pika

    Palo Alto, CA
    4 days ago
  • $180k - $300k

     .... But a large portion of training compute is wasted training on data that are already learned,...  ...and allow smaller models with fewer than half the parameters...  ..., check out our recent research on synthetic data scaling...  ...looking for a Research Scientist to investigate how... 
    Data
    Training
    Work at office
    Work from home
    Relocation package

    DatologyAI

    Redwood City, CA
    2 days ago
  • $165k - $238k

     ...moonshot factory with a mission of inventing...  ...and riskiness of research with the speed and...  ...a Senior Research Scientist you will be a key...  ...of unstructured data.We look for engineers...  ...systems that use LLMs as building blocks...  ...training, fine-tuning, or distilling... 
    Data
    Training
    Full time
    Work at office
    3 days per week

    X Company

    Mountain View, CA
    2 days ago
  • ## Senior Staff Research Scientist, Agentic AI & RLApplylocations...  ...is a frontier AI data foundry that...  ...clients with safe, scalable AI...  ...multilingual, pre-trained datasets; fine-tuned...  ...industry-specific LLMs; and RAG pipelines...  ...’s Agentic AI and RL platform — designing... 
    Data
    Training

    Centific Global Solutions, Inc.

    Palo Alto, CA
    4 days ago
  • $174k - $252k

     ...generation capabilities with a focus on...  ...cross-functional research and engineering teams...  ...Language Models (LLMs).Experience coding...  ...working with audio data, speech processing...  ...work. As a Research Scientist, you'll setup...  ...relevant education or training. US: $174000 - $25... 
    Data
    Training

    Google

    Mountain View, CA
    4 days ago
  • $174k - $252k

    Conduct AI research and development in the field...  ...by interacting with potential collaborators...  ...and distributed training on accelerators....  ...models (e.g. LLMs, VLMs, time-series...  ...multimodal real-world data such as wearable sensor...  .... As a Research Scientist, you'll setup... 
    Data
    Training

    Google

    Mountain View, CA
    1 day ago
  • $207k - $301k

    Work with experts on LLMs, multi-modal, recommendation systems,...  ...efficiency. Drive new research ideas from conception...  ...work by defining the data structure, framework,...  ...(e.g., model training/deployment, efficiency...  ...work. As a Research Scientist, you'll setup large-scale... 
    Data
    Training
    Shift work

    Google

    Mountain View, CA
    4 days ago
  • $132k - $166k

     ...therapies for patients with RAS-addicted...  ...Medicines is seeking Scientist II to join the...  ...analyze and interpret data and present...  ...conduct appropriate research and troubleshoot technical...  ...teams to support aligned development...  ...relevant education or training.Please note that... 
    Data
    Training
    Full time
    Local area

    Revolution Medicines

    Redwood City, CA
    2 days ago
  • $126k - $248k

     ...efficient unstructured data search and...  ...strong team of AI researchers from Stanford, MIT...  ...-edge research on training embedding models....  ...embedding models with MongoDB's data platform...  ...a Senior Research Scientist to join our team and...  ..., from frontier LLMs to embedding models... 
    Data
    Training
    Local area
    Worldwide
    Flexible hours

    MongoDB

    Palo Alto, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to RL Research Scientist: Training & Aligning LLMs with Data. Be the first to apply!