Research Scientist - Post Training
$250kProduct Pulse
About Us
We build training data and evaluation infrastructure that frontier AI labs use to improve their models. We partner with the world's leading labs to design high-signal datasets and run rigorous evaluations that go beyond static benchmarks. We're a small, early team (post–Series A) where individual contributors have direct impact on how the next generation of models learns and improves.
The Role
We're building out our post-training research team and hiring 2–3 Research Scientists to work together on this mission. Your job is to prove that our data works. You'll design and run training experiments that isolate the impact of our datasets on model behavior, including SFT and RL-based post-training, to measure how different data sources shift capability, generalization, and alignment. Working closely with partner labs, you'll turn our datasets into clear, defensible evidence: this data this improvement under these conditions. It's experimental, high- leverage work at the edge of model development.
What You'll Do
- Run controlled SFT and RL experiments to measure the impact of our datasets on model performance.
- Quantify lift across capabilities — reasoning, tool use, long-horizon tasks, and domain-specific workflows. Share findings directly with partner labs to deepen relationships and drive sales.
- Collaborate with internal SPLs to iterate on data quality based on your results.
- Work closely with the other Research Scientists on this team to build shared experimental infrastructure and benchmarks.
What We're Looking For
- Strong familiarity with LLM training and evaluation methodologies (SFT, RL post-training).
- Genuine obsession with how data structure, selection, and quality drive model behavior.
- Ability to design lightweight experiments, move fast, and extract actionable insights from messy results.
- Comfort working across domains — you'll touch finance, software engineering, policy, and more.
- A bias toward building over theorizing.
Must-Have Requirements
- Strong familiarity with LLM training and evaluation methodologies, including SFT and RL post-training.
- Genuine obsession with how data structure, selection, and quality drive model behavior.
- Ability to design lightweight experiments, move fast, and extract actionable insights from messy results. Comfort working across domains — finance, software engineering, policy, and more.
- Undergrad or master's research background; pre-PhD candidates preferred.
Nice-to-Have Requirements
- Prior work or internship at an RL environment company, AI safety org, or benchmarking org (METR, Artificial Analysis, or equivalent).
- Experience running controlled training experiments end-to-end.
- Published research on model evaluation, post-training, or data curation.
- Strong SWE chops alongside research instincts. Compensation
Compensation
$250K–$450K total compensation + equity
Requirements
- Run controlled SFT and RL experiments to measure dataset impact on model performance
- Quantify lift across capabilities including reasoning, tool use, long-horizon tasks, and domain- specific workflows
- Communicate findings with partner labs to drive sales
- Work with internal SPLs to iterate on data quality based on experimental results
- Strong familiarity with LLM training and evaluation methodologies
- Design lightweight experiments and extract actionable insights from messy results
- Work across multiple domains including finance, software engineering, and policy
- ...time; have a track record of exceptional research or engineering achievement; move... ...continual skills learning, and small custom post-trained models (SFT and RLVR) using proprietary... ...’re seeking an exceptional AI Research Scientist to join our small team of elite...TrainingFull timeRelocation package
- ...THE ROLE As a Research Scientist - Post Training, you may work on projects that require strong execution, communication, analytical judgment, and the ability to move quickly in ambiguous environments. WHAT YOU MAY WORK ON Own projects, workflows, research, execution,...Training
$125k - $225k
...FutureSearch is looking for exceptional Research Scientists to evaluate and improve state-of-the-art forecasting and agentic LLM web research... .... We work with frontier labs on research, evaluation, and training. You are a talented researcher with background in math, physics...TrainingRemote workFlexible hours- ...upside. Make high-conviction bets - Try and fail. But succeed an unfair amount. Job: Our first dedicated research hire - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep intuition on modern models and training. Like...Training
$250k
...About Us We build training data and evaluation infrastructure that frontier... ...benchmarks. We're a small, early team (post‑Series A) where individual... ...re building out our post‑training research team and hiring 2–3 Research Scientists to work together on this mission....TrainingInternshipShift work- ...Scale Labs in San Francisco seeks a Research Scientist focused on Safety Post-Training to advance post-training methods and interpretability for frontier AI systems. You will design pipelines, evaluate safety properties, and help translate findings into practical guidelines...Training
$300k
...applying to this role, you will be considered for Research Scientist positions at Stellon Labs pushing the frontier of efficient... ...breakthroughs in quantization, efficient training and inference, distillation, and post-training, as well as foundational research on training...TrainingFull time$225k - $300k
...Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access, and those... ...work on: Verifiable reinforcement learning at scale Mid-training and post-training of foundation models Novel objectives derived...TrainingWork at officeImmediate start- ...About the Role You will build the base intelligence layer for robotics. We train large‑scale robot foundation models from massive multimodal datasets spanning video, proprioception, action traces, language, and more. You will design and run the core large‑scale training...Training
$250k
...About AfterQuery AfterQuery builds the training data and evaluation infrastructure that frontier... ...benchmarks. We are a small, early team (post Series A) where individual contributors... ...and measured. Working directly with research teams at top AI labs, you’ll experiment with...Training$285k - $380k
...is a quickly growing group of committed researchers, engineers, policy experts, and... ...Role We are seeking a People Research Scientist to join our People Data Solutions team.... ...an equivalent combination of education, training, and/or experience. Required Field of Study...TrainingVisa sponsorship$170k - $220k
...Introduction The Center for AI Safety is a research and field-building nonprofit located in... ...and technical research. As a research scientist, you will pursue a variety of research projects... ...). Have experience launching and training distributed ML jobs. Communicate clearly...TrainingWork at office- ...and externally. Collaborate with Product & Engineering to ship research‑backed features. Contribute to open‑source repos and author technical... ...experimental work; hands‑on experience with large‑scale model training, transformer architectures, reinforcement‑learning techniques,...Training
- ...realize all of our product goals. As a Machine Learning Scientist at Sesame, you are a research-oriented person with experience in NLP, Speech, and/or... ...model architectures, data curation, model evaluation, training & inference infrastructure, research, and experimentation...TrainingFull timeContract workFlexible hours
$150k - $250k
...grasping and more dexterous behaviors in unstructured environments Research and implement state-of-the-art robot learning policies,... ...production robot fleets Optimize robot policies for distributed training at scale and real-time edge deployment Ship production quality,...Training- ...mission to make robots commonplace. The team is looking for a Research Scientist to help architect and deploy the machine learning models... ...vision‑language‑action (VLA) models. What You’ll Be Doing: Training, deploying, and maintaining manipulation models on physical...Training
- ...professional programmers, using a combination of inventive research, design, and engineering. Our organization is very... ...debate, crazy ideas, and shipping code. Research Scientist Cursor is building the future of coding. We train frontier coding agents and scale RL on real user...Training
$120k - $250k
...grasping and more dexterous behaviors in unstructured environments Research and implement state-of-the-art robot learning policies,... ...production robot fleets Optimize robot policies for distributed training at scale and real-time edge deployment Ship production quality,...Training$250k - $400k
...opportunity to work on genuinely novel AI for Science research, combining frontier reasoning, post-training and reinforcement learning with a proprietary... ...Engineer, Applied Researcher or experienced Research Scientist. What matters most is hands‑on ownership of reasoning...Training$204k - $259k
...initiate and foster collaborations with other research teams in Alphabet. AI Foundations areas... ...hybrid role, you will report to a Principal Scientist. Responsibilities Participate in Waymo’s Foundation World Model post‑training and evaluation Research and develop cutting...TrainingTemporary workRemote work- ...Traverse is a research data lab building reinforcement learning environments... ...else has figured out how to train models on. We work directly... ...About the Role As a Research Scientist, you will design and build RL... ...integrate environments into their post-training pipelines Your...Training
- ...funding from major venture capital firms. About this role As a Research Scientist you will join a small, focused team of researchers and... ...prefix tuning, state‑space methods). Investigate synthetic training data generalization and develop self‑study pipelines that allow...TrainingWork at office
- ...About the role We’re looking for a top‑tier Research Scientist to join our tech team. Your core responsibilities will be to: Lead Research for... ...‑art retrieval and ranking systems for AI agents Prototype, train, and evaluate new models for factual search and multimodal understanding...TrainingRemote workFlexible hours
- ...founding team brought together leading researchers in this space and top silicon valley operators... ...more. About the role As an AI Research Scientist, you will conduct groundbreaking... ...and deep-learning frameworks Experience training and evaluating large models on protein,...Training
$300k
...Research Scientist — Frontier World Models & RL A stealth, exceptionally well-backed applied AI lab... ...with real users and a data flywheel to train against the moment beta launches. The open... ...in an autoregressive setting. RL / Post-Training Advance RL post-training for coding...TrainingRemote workVisa sponsorshipRelocation package- ...Velvet is a data research company building the datasets that power the next generation of multimodal AI.... ...more human by producing high-quality audiovisual training data for frontier labs. We’re hiring a Research Scientist to develop and fine‑tune models for video and audio...TrainingImmediate startShift work
- ...consistency, and genuine care. The Role Research at Aristotle works on the open problems in... ...that can be evaluated reliably at scale. Train models where prompting falls short, for student... ...experience with some of: LLM evaluation, post‑training/fine‑tuning, agent systems,...TrainingFull timeContract workSummer workFlexible hours
$176k - $304k
...; San Francisco, CA USA As a Machine Learning Research Scientist I/II in LLM Inference you will lead research on how we train and serve large language models for scientific... ...What You’ll Be Building Develop and optimize LLM post-training strategies including SFT, RLHF, and...Training- ...diverse architectures onto our platform, and we aim to both get more out of what they've already trained and shape how the next generation of models is designed. Department: Research Location: San Francisco What You'll Do Lead a research agenda on real‑time interactive...TrainingVisa sponsorshipRelocation package
- ...systems . This role is for an experienced scientist who thrives both in innovating... ...reasoning, and deep content extraction. Research, evaluate, and integrate the latest vision... ...knowledge of quantization/LoRA/efficient training. Proficiency with deep learning frameworks...Training
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist - Post Training. Be the first to apply!
- drug safety scientist San Francisco, CA
- image scientist San Francisco, CA
- molecular biology scientist San Francisco, CA
- safety scientist San Francisco, CA
- validation scientist San Francisco, CA
- senior analytical scientist San Francisco, CA
- support scientist San Francisco, CA
- water quality scientist San Francisco, CA
- scientist 1 San Francisco, CA
- scientist assay development San Francisco, CA

