RL Research Scientist: Training & Aligning LLMs with Data
Remote Jobs
Snorkel AI in Redwood City is seeking a Research Scientist to advance reinforcement learning methods for training and aligning large language models. You will build data products, reward models, and end-to-end RL pipelines in collaboration with research, engineering, and delivery teams. The role focuses on developing data-as-a-service capabilities, translating research into practical datasets, signals, and corpora for frontier AI labs, while contributing to Snorkel's long-term differentiation. #J-18808-Ljbffr Remote Jobs
- ...San Francisco seeks a Research Scientist to work on reinforcement learning for training and aligning large language models.... ...foundational role focuses on data, reward signals, and... ...You will collaborate with Snorkel’s research,... ...teams to advance RL data capabilities, translating...DataTraining
- ...California is seeking a Research Scientist to advance reinforcement learning for training and aligning large language models. You... ...research ideas into practical data products and collaborate with research, engineering,... ...teams to deliver RL‑ready corpora for client...DataTraining
$200k - $350k
...AI doesn’t start with the model, it starts with the data.We’re on a mission... ...Snorkel started as a research project in the... ...organizations to empower scientists, engineers,... ...reinforcement learning for training and aligning large language... ...to advance our RL data capabilities...DataTrainingLocal areaShift work$147k - $211k
In accordance with Washington state law... ...Experience conducting research and development,... ...-tuning and post-training LLMs using Reinforcement Learning (RL) or other alignment methods.... ...information from different data types (e.g.,... ...work. As a Research Scientist, you'll setup...DataTrainingTemporary work$147k - $211k
SnapshotWe are seeking strong Research Scientists with expertise in AI research... ...strong awareness of the AI alignment and safety landscape, and a... ...strategiesExperience finetuning and post-training LLMs using RLExperience with... ...information from different data types (e.g., vision, audio,...DataTrainingFull time$204k - $259k
...technology company with the mission to... ...with other research teams in Alphabet... ...to a Principal Scientist. You will:... ...World Model post-training and evaluation... ...develop cutting edge RL and... ...manner such as Data parallel, FSDP... ...policy learning or alignment with human preferences...DataTrainingFull timeRemote work- ...lab capabilities with the ultimate... ...three frontier research institutes with... ...foundation models: Trained on some of the largest... ...frontier LLMs with biological... ...weights. Scientific data at unprecedented... ...discovery, including RL, reward modeling... ...with scientists, grounded in real...DataTraining
- ...we’re a team of scientists, engineers,... ...and collaborate with others on critical... ...role of the Research Scientist / Research... ...and develop data and algorithmic... ...responsibilities:Post-training / instruction... ...of the art LLMs, focusing on... ...NeurIPS, ICLR, ICML, RL/DL, EMNLP, AAAI...DataTraining
- ...those locations, with priority determined... ...What you’ll do As a Research Scientist at Simular, you... ...interaction, and alignment (e.g. reward modeling... ...end-to-end: from data collection and benchmarking... ..., to model training and evaluation.... ...training, or fine-tuning LLMs/VLMs...DataTraining
$138.23k - $207.58k
...Machine Learning Research Scientist: Generative Models... ...combines AD hardware with our generalized AI... ...improve through data, the Nuro Driver is... .... This includes alignment with reward/cost models... ...with closed-loop RL, etc. Research... ...Scientist Training Program San Jose...DataTrainingFull timeWork at office$156.56k - $180.04k
...Research Scientist - Interpretability (1 Year Fixed Term... ...key resource for aligning artificial intelligence models with human-like neural... ...dimensional neural data ~ Collaborate with... ...experience in training, fine-tuning, and... ...and quantization of LLMs or multimodal models...DataTrainingFull timeFixed term contractFlexible hours$209.5k - $283.5k
...communities we serve. With approximately 100... ...hiring a Staff AI Research Scientist to join the Intuit... ...and model training (pre and post) , and... ...of domain experts, data scientists, and machine... ...that enable LLMs to reason over structured... ...goals that are in alignment with Intuit...DataTrainingWork experience placementWorldwide$211.2k - $257.5k
...biomedical imaging with molecular and language data. By... ...accessible to our global research community. Responsibilities... ..., code, and train large-scale... ..., Multimodal LLMs) that can embed... ...with experimental scientists and software... ...Multimodal Fusion (e.g., aligning image embeddings...DataTrainingLocal area$141k - $244k
Research Scientist - Gemini Personal Intelligence Mountain... ...research behind Gemini, with a mission to make AI... ...Language Models (LLMs) to build the brain... ...with users' personal data to solve real-world... ...Driving research on post‑training techniques (e.g., RL, SFT, and preference...DataTrainingFull timeRelocationFlexible hoursShift work- ...own the metrics that define success. It fits a researcher with deep generative-modeling or model-based-RL expertise who has trained models to the limits of a multi-node cluster.... ...Run scaling studies that show where compute, data, and architecture pay off. Publish at the...DataTraining
$128.25k - $266.88k
...consumer inbox with hundreds of millions... ...messages and data at the petabyte... ...team of Researchers, Engineers, Product... ...in designing, training, and evaluating... ...level research scientists and machine learning... ...performance, alignment, and governance... ...models (LLMs) for search, retrieval...DataTrainingWork at officeFlexible hours$150k - $200k
...experience — talk with your recruiter to... ...evaluation benchmarks and alignment techniques for... ...As a Research Scientist/ ML Engineer, you... ...reliability, and data curation. Your technical... ...such as distributed training of large models and... ...on Alignment and RL. We help companies...DataTrainingFull time$174k - $252k
Scope and drive research efforts to improve complex frontier... ....Curate and generate data to evaluate and... ...development.Experience with LLM training and generative models.... ...of work. As a Research Scientist, you'll setup large-... ...foundational models. Modern LLMs are required to...DataTraining- ...autonomous, clinical conversations with patients. We have trained our own LLMs as part of our Polaris... ...hospital leaders, AI pioneers, and researchers from institutions like El Camino Health... ...DoDesign, Develop, Evaluate and update data-driven models for Speech First applications...DataTrainingWork at office
$129.5k - $166k
Research Specialist / Scientist, Metabolic Phenotyping (In Vitro & In Vivo... ...that make the team's data rigorous and reproducible... ...in collaboration with the team, including animal... ...discovery team to align in vitro mechanism-of... ..., compensation and training. Thank you for your...DataTrainingContract workLocal area$180k - $300k
.... But a large portion of training compute is wasted training on data that are already learned,... ...and allow smaller models with fewer than half the parameters... ..., check out our recent research on synthetic data scaling... ...looking for a Research Scientist to lead work on post-...DataTrainingWork at officeWork from homeRelocation package$185k - $400k
...looking for a staff or lead-level Research Engineer, Data to architect and scale data... ...systems supporting model training for our advanced multimodal... ..., and video datasetsPartner with research and engineering... ...engineering and ML data curation for LLMs, VLMs, or other large-scale...DataTrainingRemote work$180k - $300k
.... But a large portion of training compute is wasted training on data that are already learned,... ...and allow smaller models with fewer than half the parameters... ..., check out our recent research on synthetic data scaling... ...looking for a Research Scientist to investigate how...DataTrainingWork at officeWork from homeRelocation package$165k - $238k
...moonshot factory with a mission of inventing... ...and riskiness of research with the speed and... ...a Senior Research Scientist you will be a key... ...of unstructured data.We look for engineers... ...systems that use LLMs as building blocks... ...training, fine-tuning, or distilling...DataTrainingFull timeWork at office3 days per week- ## Senior Staff Research Scientist, Agentic AI & RLApplylocations... ...is a frontier AI data foundry that... ...clients with safe, scalable AI... ...multilingual, pre-trained datasets; fine-tuned... ...industry-specific LLMs; and RAG pipelines... ...’s Agentic AI and RL platform — designing...DataTraining
$174k - $252k
...generation capabilities with a focus on... ...cross-functional research and engineering teams... ...Language Models (LLMs).Experience coding... ...working with audio data, speech processing... ...work. As a Research Scientist, you'll setup... ...relevant education or training. US: $174000 - $25...DataTraining$174k - $252k
Conduct AI research and development in the field... ...by interacting with potential collaborators... ...and distributed training on accelerators.... ...models (e.g. LLMs, VLMs, time-series... ...multimodal real-world data such as wearable sensor... .... As a Research Scientist, you'll setup...DataTraining$207k - $301k
Work with experts on LLMs, multi-modal, recommendation systems,... ...efficiency. Drive new research ideas from conception... ...work by defining the data structure, framework,... ...(e.g., model training/deployment, efficiency... ...work. As a Research Scientist, you'll setup large-scale...DataTrainingShift work$132k - $166k
...therapies for patients with RAS-addicted... ...Medicines is seeking Scientist II to join the... ...analyze and interpret data and present... ...conduct appropriate research and troubleshoot technical... ...teams to support aligned development... ...relevant education or training.Please note that...DataTrainingFull timeLocal area$126k - $248k
...efficient unstructured data search and... ...strong team of AI researchers from Stanford, MIT... ...-edge research on training embedding models.... ...embedding models with MongoDB's data platform... ...a Senior Research Scientist to join our team and... ..., from frontier LLMs to embedding models...DataTrainingLocal areaWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to RL Research Scientist: Training & Aligning LLMs with Data. Be the first to apply!
- scientist ii Redwood City, CA
- machine learning scientist Redwood City, CA
- scientist Redwood City, CA
- quality control scientist Redwood City, CA
- qc scientist Redwood City, CA
- regulatory scientist Redwood City, CA
- research scientist - biology Redwood City, CA
- applied scientist Redwood City, CA
- operations research scientist Redwood City, CA
- image scientist Redwood City, CA


