RL Research Scientist - Data for LLM Alignment
Snorkel AI, Inc.
Snorkel AI, Inc. in California is seeking a Research Scientist to advance reinforcement learning for training and aligning large language models. You will translate research ideas into practical data products and collaborate with research, engineering, and delivery teams to deliver RL‑ready corpora for client use. The role emphasizes building data pipelines, experimenting with RL recipes, and contributing to publications, with a strong emphasis on data-driven approaches and scalable, multi‑node #J-18808-Ljbffr Snorkel AI, Inc.
- ...San Francisco seeks a Research Scientist to work on reinforcement... ...learning for training and aligning large language models.... ...role focuses on data, reward signals, and training... ...procedures that steer LLM behavior in reliable directions... ...teams to advance RL data capabilities,...Data
$200k - $350k
...it starts with the data.We’re on a mission... ...Snorkel started as a research project in the... ...organizations to empower scientists, engineers,... ...learning for training and aligning large language... ...procedures that steer LLM behavior in reliable... ...teams to advance our RL data capabilities —...DataLocal areaShift work$204k - $259k
...collaborations with other research teams in Alphabet. AI... ...report to a Principal Scientist. You will: Participate... ...develop cutting edge RL and Distillation... ...manner such as Data parallel, FSDP and other... ...on-policy learning or alignment with human preferences...DataFull timeTemporary workRemote work$147k - $211k
...Research Scientist, Multimodal Alignment, Safety, and Fairness Kirkland, Washington, US; Mountain View, California... ...and post-training LLMs using RL Experience with developing agentic... ...integrating information from different data types (e.g., vision, audio, text)....DataFull time$147k - $211k
...systems. Experience conducting research and development, including... ...Reinforcement Learning (RL) or other alignment methods. Experience with developing... ...information from different data types (e.g., vision, audio,... ...of work. As a Research Scientist, you'll setup large-scale...DataTemporary work$214k - $375k
...Research Scientist, AI New York, NY (Hybrid) Biohub is the first large... ...model weights Scientific data at unprecedented scale: AI systems... ...drive biological discovery: RL, reward modeling, reasoning,... ...you if your experience aligns with the skills we seek for future...DataWorldwideRelocation package$117.2k - $313.7k
...Salesforce AI Research is looking for outstanding... ...AI Research Scientists and Research Engineers... ...Large language model (LLM)-powered agents,... ...reinforcement learning (RL), reasoning and... ...including techniques for alignment, safety, and... ...use your personal data and your rights, including...Data- ...DeepMind, we’re a team of scientists, engineers, machine... ...for a versatile Research Scientist at ease... ...Exploring data, reasoning and algorithmic... ...experience Significant LLM post-training... ...Safety, Fairness and Alignment Track record of publications... ..., ICLR, ICML, RL/DL, EMNLP, AAAI,...Data
$200k - $350k
...the model, it starts with the data.We’re on a mission to help enterprises... ...15, when Snorkel started as a research project in the Stanford AI Lab... ...organizations to empower scientists, engineers, financial experts,... ...utilizing techniques such as LLM as a JudgePrototype and build...DataLocal area- ...Research Scientist As a Research Scientist at Simular, you will: Shape... ...human-agent interaction, and alignment (e.g. reward modeling,... ...experiments end-to-end: from data collection and benchmarking,... ...Reinforcement learning and/or LLM-based agents Computer vision...Data
- Simular is seeking a Research Scientist to push the boundaries of AI research across planning, reinforcement learning, and multimodal reasoning. You will drive end-to-end experiments, from data collection to model evaluation, and collaborate with engineers to bring research...Data
- ...focused large language model (LLM) for the healthcare industry. Our team comprised of ex-researchers from Microsoft, Meta, Nvidia,... ...foundation model training and alignment to create AI-powered conversational... ...Develop, Evaluate and update data-driven models for Speech First...DataWork at office
$156.56k - $180.04k
Research Scientist - Interpretability (1 Year Fixed Term) Join to apply for the... ...as a key resource for aligning artificial intelligence models... ...analysis of high-dimensional neural data Collaborate with... ...Scientist / Research Engineer, LLM Evaluation Hayward, CA $85,00...DataFull timeFixed term contractFlexible hours$143.63k - $299.38k
...billions of messages and manage data on a petabyte scale. Using... ...You are a seasoned Applied ML Researcher who thrives at the intersection... ...don't just follow the latest LLM trends; you understand the mechanics... ..., preference optimization, or alignment techniques. Experience with...DataWork at officeFlexible hours$128.25k - $266.88k
...billions of messages and data at the petabyte scale.... ...this amazing team of Researchers, Engineers, Product... ...evaluation. Incorporate LLM-driven synthesis into... ...mid-level research scientists and machine learning engineers... ...for performance, alignment, and governance....DataWork at officeFlexible hours$209.5k - $283.5k
...OverviewIntuit is hiring a Staff AI Research Scientist to join the Intuit Foresight... ...reinforcement learning, and LLM based reasoning for real-... ...team of domain experts, data scientists, and machine learning... ...ambitious research goals that are in alignment with Intuit products, while...DataWork experience placementWorldwide$150k - $200k
...novel evaluation benchmarks and alignment techniques for enterprise-... ...models. Responsibilities As a Research Scientist/ ML Engineer, you will play a... ...for safety, reliability, and data curation. Your technical... ...art research on Alignment and RL. We help companies build AI systems...DataFull time- ...and own the metrics that define success. It fits a researcher with deep generative-modeling or model-based-RL expertise who has trained models to the limits of a... ...training. Run scaling studies that show where compute, data, and architecture pay off. Publish at the frontier...Data
- ABOUT THE ROLE This is a research-driven, high-impact role for ML researchers who want to push the boundaries of real-time AI. As a Founding... ...Background - You've worked on advanced ML problems (like LLM pre-training and post‑training, transcription model training, TTS...
- ...About The Role We’re looking for an AI Research Scientist to advance the methodological frontier... ...will develop and own a research agenda aligned with Sprinter’s company strategy. Your... ...including work that may involve IRB review, data governance, external validation, or...DataTemporary workWork at officeMonday to FridayMonday to ThursdayFlexible hours
$195.2k - $262.2k
...developers and enterprises from data and model training through to production... ...role Nebius Token Factory needs scientists who can turn frontier inference bottlenecks into research problems, publish credible work,... ...research programs in efficient LLM and VLM inference with...DataTemporary workImmediate startRemote work$174k - $253k
...detecting factual inaccuracies in all LLM generated content.Design and... ...algorithmic interventions and data curation strategies to improve... ...years of experience leading a research agenda.Preferred... ...other researchers.As a Research Scientist, you will be responsible for improving...Data- ...’s first healthcare‑only, safety‑focused LLM — a breakthrough platform designed to transform... ..., hospital leaders, AI pioneers, and researchers from institutions like El Camino Health,... ...DoDesign, Develop, Evaluate and update data-driven models for Speech First applications...DataWork at office
$180k - $300k
...of training compute is wasted training on data that are already learned, irrelevant, or... .... For more details, check out our recent research on synthetic data scaling (BeyondWeb) and... ...About the RoleWe're looking for a Research Scientist to investigate how intervening on...DataWork at officeWork from homeRelocation package$129.5k - $166k
Research Specialist / Scientist, Metabolic Phenotyping (In Vitro & In Vivo) Our Mission Our mission is to... ...notebook records that make the team's data rigorous and reproducible. Key Responsibilities... ...across the discovery team to align in vitro mechanism-of-action findings...DataContract workLocal area$116.8k - $140k
...SUMMARY: Natera is seeking an innovative Scientist, R&D to join Natera’s Oncology Product... ...high complexity experimentsPerform basic data analysis (e.g. Excel, JMP). Coordinate the... ..., outcomes, and next steps to ensure alignment and transparencyAuthor and review SOPs,...DataImmediate startWorldwideFlexible hours$180k - $300k
...of training compute is wasted training on data that are already learned, irrelevant, or... .... For more details, check out our recent research on synthetic data scaling (BeyondWeb) and... ...About the RoleWe’re looking for a Research Scientist to lead work on post-training data...DataWork at officeWork from homeRelocation package$108k - $150k
RESEARCH ASSOCIATE / ASSOCIENT SCIENTIST, NONCLINICAL DEVELOPMENT Key Responsibilities Coordinate and execute in... ...planning, timeline development, and data‑driven decision‑making to support program... ..., and ensure study end points align with program objectives. Perform necropsies...DataFull timeContract workTemporary workFlexible hours3 days per week$165k - $238k
...the aspiration and riskiness of research with the speed and ambition of... ...the role:As a Senior Research Scientist you will be a key architect of... ...large corpora of unstructured data.We look for engineers who... ...LLMs as building blocks, such as LLM-as-a-judge evaluation pipelines...DataFull timeWork at office3 days per week- ...the right time.Translational Scientist, Applied Machine Learning and... ...advanced Large Language Model (LLM) orchestration and computational... ...vast multimodal oncology data, helping to scale scientific discovery... ...and collaboration: Work with Research, Engineering & Data Science...DataFull timeRemote workShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to RL Research Scientist - Data for LLM Alignment. Be the first to apply!
- scientist ii Redwood City, CA
- scientist 1 Redwood City, CA
- image scientist Redwood City, CA
- qc scientist Redwood City, CA
- research scientist Redwood City, CA
- analytical scientist Redwood City, CA
- research scientist - biology Redwood City, CA
- genomics scientist Redwood City, CA
- quality control scientist Redwood City, CA
- applied scientist Redwood City, CA

