Research Scientist - Post Training
$250kProduct Pulse
About Us
We build training data and evaluation infrastructure that frontier AI labs use to improve their models. We partner with the world's leading labs to design high-signal datasets and run rigorous evaluations that go beyond static benchmarks. We're a small, early team (post–Series A) where individual contributors have direct impact on how the next generation of models learns and improves.
The Role
We're building out our post-training research team and hiring 2–3 Research Scientists to work together on this mission. Your job is to prove that our data works. You'll design and run training experiments that isolate the impact of our datasets on model behavior, including SFT and RL-based post-training, to measure how different data sources shift capability, generalization, and alignment. Working closely with partner labs, you'll turn our datasets into clear, defensible evidence: this data this improvement under these conditions. It's experimental, high- leverage work at the edge of model development.
What You'll Do
- Run controlled SFT and RL experiments to measure the impact of our datasets on model performance.
- Quantify lift across capabilities — reasoning, tool use, long-horizon tasks, and domain-specific workflows. Share findings directly with partner labs to deepen relationships and drive sales.
- Collaborate with internal SPLs to iterate on data quality based on your results.
- Work closely with the other Research Scientists on this team to build shared experimental infrastructure and benchmarks.
What We're Looking For
- Strong familiarity with LLM training and evaluation methodologies (SFT, RL post-training).
- Genuine obsession with how data structure, selection, and quality drive model behavior.
- Ability to design lightweight experiments, move fast, and extract actionable insights from messy results.
- Comfort working across domains — you'll touch finance, software engineering, policy, and more.
- A bias toward building over theorizing.
Must-Have Requirements
- Strong familiarity with LLM training and evaluation methodologies, including SFT and RL post-training.
- Genuine obsession with how data structure, selection, and quality drive model behavior.
- Ability to design lightweight experiments, move fast, and extract actionable insights from messy results. Comfort working across domains — finance, software engineering, policy, and more.
- Undergrad or master's research background; pre-PhD candidates preferred.
Nice-to-Have Requirements
- Prior work or internship at an RL environment company, AI safety org, or benchmarking org (METR, Artificial Analysis, or equivalent).
- Experience running controlled training experiments end-to-end.
- Published research on model evaluation, post-training, or data curation.
- Strong SWE chops alongside research instincts. Compensation
Compensation
$250K–$450K total compensation + equity
Requirements
- Run controlled SFT and RL experiments to measure dataset impact on model performance
- Quantify lift across capabilities including reasoning, tool use, long-horizon tasks, and domain- specific workflows
- Communicate findings with partner labs to drive sales
- Work with internal SPLs to iterate on data quality based on experimental results
- Strong familiarity with LLM training and evaluation methodologies
- Design lightweight experiments and extract actionable insights from messy results
- Work across multiple domains including finance, software engineering, and policy
- We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards... ...overview:We are seeking an exceptional Research Scientist to join our team, focusing on alignment and post-training techniques for large-scale video generation...TrainingRelocation
$216k - $270k
Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner... .... The range displayed on each job posting reflects the minimum and maximum target... ...performance, and relevant education or training. Scale employees in eligible roles are...TrainingFull time$120.7k - $238.6k
The OpportunityAdobe Research is looking for research scientists in Generative AI to join a world-class research... ...Experience on large-scale generative model training· Experience of working with large-... ...in Colorado (as listed on the job posting), the application window will...TrainingFull timeTemporary workLocal areaWorldwide- ...Researcher Position As one of our researchers - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep intuition on modern models and training. Like to argue how search, recommendations, and transformer models can...Training
- ...workersResponsibilitiesWe are looking for an exceptional AI Research Scientist to join our growing team. In this role, you will be responsible... ...work; hands‑on experience with large‑scale model training, transformer architectures, reinforcement‑learning techniques...TrainingRemote workFlexible hours
$160k - $220k
...enjoy multi-year runway.About the RoleWe’re looking for an AI Research Scientist to advance the methodological frontier of AI in healthcare.... ...strategy. Your work may include novel architectures, new training or evaluation techniques, long-horizon research bets, peer-reviewed...TrainingTemporary workWork at officeMonday to FridayMonday to Thursday- ...Research Scientist Engineering · Full-time · San Francisco; New York Our mission is to automate coding. The first step in our journey... ...Research Scientist Cursor is building the future of coding. We train frontier coding agents and scale RL on real user data to make...TrainingFull time
- ...Chai Discovery Chai is a research lab working on AI to unlock biology... ...We are hiring research scientists with outlier insight who can drive... ...the work we pursue. From pre-training large diffusion models and architecture design, to post-training and inference time scaling...Training
$150k - $250k
...About the Company Our client builds the training data and evaluation infrastructure that frontier... ...quant firms, big tech, and leading AI research labs. Founded 2025 · 11–50 people ·... ...datasets on model behavior (SFT and RL post-training), and turn the results into defensible...TrainingFull timeShift work$400k
...prototype and early commercial traction across several high-profile industry verticals. The role As a Senior Research Scientist, your focus is post-training - curating data, fine-tuning pre-trained speech models, and building the evaluation infrastructure that...TrainingRelocation packageShift work- ...Research Scientist ThirdLayer is solving one of the hardest problems in deploying agents: AI models are generic, but... ...everything we build. The Role We're building the training loop that makes agents specific: post-training models on the workflows, context, and...Training
- ...Overview Join our core R&D team building end-to-end automated research systems . Zochi publishes the first fully-AI generated... ...at the forefront of long-horizon agentic capabilities, post-training for open-ended goals, and environment development. Publish...Training
$234.3k - $349k
...future of work with AI. About the roleAI research at WRITER isn't just about publishing... ...deployments in the world. As an AI research scientist, you'll be at the center of that work.... ...possible. The work you do here — on post-training, planning, multi-step reasoning, and...TrainingFull timeWork at officeLocal area$225k - $300k
...Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access... ...on: Verifiable reinforcement learning at scale Mid-training and post-training of foundation models Novel objectives derived...TrainingWork at officeImmediate start- ...a world-class team of engineers, designers, marketers, sellers, researchers, and operational experts to achieve our mission. Job: As one of our Researchers - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep...TrainingWork at officeVisa sponsorshipFlexible hours
- ...sessions, it retains scraps at best. We train models to study your world and anticipate... ...You will join a small, focused team of researchers and engineers working at the frontier of learning and memory. As a Research Scientist, you'll design experiments, develop new recipes...TrainingWork at office
- ...Researcher Position at Hedra Hedra is building a world-class Physical AI research team to push the boundaries of action-conditioned... ...modeling for embodied systems Design novel architectures, training objectives, and evaluation frameworks for VLMs, VLAs, and world...TrainingWork at office
- ...Machine Learning Scientist Sesame believes in a future where computers are lifelike -... ...Learning Scientist at Sesame, you are a research-oriented person with experience in NLP,... ...architectures, data curation, model evaluation, training & inference infrastructure, research,...TrainingFull timeContract workFlexible hours
$200k - $335k
...The role As a research scientist, you will design, implement, and optimize the large-scale training infrastructure that powers our frontier reinforcement learning stack... ...Partner with researchers to bring frontier post-training capabilities into production deployments...TrainingFull timeWork at officeVisa sponsorshipRelocation package- ...AfterQuery AfterQuery is an applied research lab curating data solutions for foundation... ...data works . You will design and run training experiments that isolate the impact of our... ...behavior. This includes SFT and RL-based post-training, where you'll measure how different...TrainingShift work
$250k - $325k
...backgrounds in technology — from AI research to systems engineering to product... ...looking for a talented Research Scientist with a strong background in... ...plus : ~ Large-scale model training ~ Data curation for pretraining or post-training ~ Tokenizers and VAEs...TrainingFull time$125k - $225k
...FutureSearch is looking for exceptional Research Scientists to evaluate and improve state-of-the-art forecasting and agentic LLM web research... .... We work with frontier labs on research, evaluation, and training. You are a talented researcher with background in math, physics...TrainingRemote workFlexible hours- ...Senior Research Scientist Sciforium is an AI infrastructure company developing next-generation multimodal AI models and a proprietary,... ...generative media, model architecture, optimization, and scalable training systems. You will work hands-on with modern ML frameworks,...TrainingFlexible hours
- ...Research Scientist / Machine Learning Scientist Location: SF Bay Area/Hybrid / Remote Type: Full-Time About the Role: The Client... ...home here. We're looking for: Hands-on experience training large-scale models, including reward models, preference models...TrainingFull timeRemote work
- ...5, when Snorkel started as a research project in the Stanford AI Lab... ...organizations to empower scientists, engineers, financial experts... ...datasets that drive frontier model training and evaluation based on... ...externally through publications, blog posts, conference talks, and...TrainingFull timeLocal area
- ...physical goods as easily as they post online, and we're building... ...is Varun Jampani, a leading researcher who co-authored Dreambooth and... ..., was formerly a Principal Scientist at Amazon. Together, we're pioneering... ..., and author proprietary training paradigms when existing open-...Training
$300k
...applying to this role, you will be considered for Research Scientist positions at Stellon Labs pushing the frontier of efficient... ...breakthroughs in quantization, efficient training and inference, distillation, and post-training, as well as foundational research on training...TrainingFull time$285k - $380k
...is a quickly growing group of committed researchers, engineers, policy experts, and... ...We are seeking a Recruiting Research Scientist to join our People Data Solutions team.... ...an equivalent combination of education, training, and/or experience Required field of...TrainingFull timeWork at officeVisa sponsorshipFlexible hours$204k - $259k
...Research Scientist, RL for Autonomous Planning & World Modeling Waymo is an autonomous driving technology company with the mission to... ...will: Participate in Waymo's Foundation World Model post-training and evaluation Research and develop cutting edge RL and...TrainingFull timeTemporary workRemote work$300k - $320k
...Research Scientist Anthropic's mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial... ...improve model capabilities on scientific tasks through post-training, evaluation design, and RL environment development. As a...TrainingWork at officeVisa sponsorshipFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist - Post Training. Be the first to apply!
- scientist ii San Francisco, CA
- scientist 1 San Francisco, CA
- image scientist San Francisco, CA
- downstream processing scientist San Francisco, CA
- entry level research scientist San Francisco, CA
- qc scientist San Francisco, CA
- research scientist San Francisco, CA
- analytical scientist San Francisco, CA
- research scientist - biology San Francisco, CA
- genomics scientist San Francisco, CA


