Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Scientist - Post Training

$250k
Full-time

Product Pulse

About Us

We build training data and evaluation infrastructure that frontier AI labs use to improve their models. We partner with the world's leading labs to design high-signal datasets and run rigorous evaluations that go beyond static benchmarks. We're a small, early team (post–Series A) where individual contributors have direct impact on how the next generation of models learns and improves.

The Role

We're building out our post-training research team and hiring 2–3 Research Scientists to work together on this mission. Your job is to prove that our data works. You'll design and run training experiments that isolate the impact of our datasets on model behavior, including SFT and RL-based post-training, to measure how different data sources shift capability, generalization, and alignment. Working closely with partner labs, you'll turn our datasets into clear, defensible evidence: this data this improvement under these conditions. It's experimental, high- leverage work at the edge of model development.

What You'll Do

  1. Run controlled SFT and RL experiments to measure the impact of our datasets on model performance.
  2. Quantify lift across capabilities — reasoning, tool use, long-horizon tasks, and domain-specific workflows. Share findings directly with partner labs to deepen relationships and drive sales.
  3. Collaborate with internal SPLs to iterate on data quality based on your results.
  4. Work closely with the other Research Scientists on this team to build shared experimental infrastructure and benchmarks.

What We're Looking For

  1. Strong familiarity with LLM training and evaluation methodologies (SFT, RL post-training).
  2. Genuine obsession with how data structure, selection, and quality drive model behavior.
  3. Ability to design lightweight experiments, move fast, and extract actionable insights from messy results. 
  4. Comfort working across domains — you'll touch finance, software engineering, policy, and more.
  5. A bias toward building over theorizing.

Must-Have Requirements

  1. Strong familiarity with LLM training and evaluation methodologies, including SFT and RL post-training.
  2. Genuine obsession with how data structure, selection, and quality drive model behavior.
  3. Ability to design lightweight experiments, move fast, and extract actionable insights from messy results. Comfort working across domains — finance, software engineering, policy, and more.
  4. Undergrad or master's research background; pre-PhD candidates preferred.

Nice-to-Have Requirements

  1. Prior work or internship at an RL environment company, AI safety org, or benchmarking org (METR, Artificial Analysis, or equivalent).
  2. Experience running controlled training experiments end-to-end.
  3. Published research on model evaluation, post-training, or data curation.
  4. Strong SWE chops alongside research instincts. Compensation

Compensation

$250K–$450K total compensation + equity 

Requirements

  1. Run controlled SFT and RL experiments to measure dataset impact on model performance
  2. Quantify lift across capabilities including reasoning, tool use, long-horizon tasks, and domain- specific workflows
  3. Communicate findings with partner labs to drive sales
  4. Work with internal SPLs to iterate on data quality based on experimental results
  5. Strong familiarity with LLM training and evaluation methodologies
  6. Design lightweight experiments and extract actionable insights from messy results 
  7. Work across multiple domains including finance, software engineering, and policy

Vacancy posted 9 days ago
Similar jobs that could be interesting for youBased on the Research Scientist - Post Training in San Francisco, CA vacancy
  •  ...time; have a track record of exceptional research or engineering achievement; move...  ...continual skills learning, and small custom post-trained models (SFT and RLVR) using proprietary...  ...’re seeking an exceptional AI Research Scientist to join our small team of elite... 
    Training
    Full time
    Relocation package

    P-1 AI

    San Francisco, CA
    5 days ago
  •  ...THE ROLE As a Research Scientist - Post Training, you may work on projects that require strong execution, communication, analytical judgment, and the ability to move quickly in ambiguous environments. WHAT YOU MAY WORK ON Own projects, workflows, research, execution,... 
    Training

    Career-Launch.net

    San Francisco, CA
    5 days ago
  • $125k - $225k

     ...FutureSearch is looking for exceptional Research Scientists to evaluate and improve state-of-the-art forecasting and agentic LLM web research...  .... We work with frontier labs on research, evaluation, and training. You are a talented researcher with background in math, physics... 
    Training
    Remote work
    Flexible hours

    Future Research Corp

    San Francisco, CA
    15 hours ago
  •  ...upside. Make high-conviction bets - Try and fail. But succeed an unfair amount. Job: Our first dedicated research hire - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep intuition on modern models and training. Like... 
    Training

    Parallel Web Systems

    San Francisco, CA
    2 days ago
  • $250k

     ...About Us We build training data and evaluation infrastructure that frontier...  ...benchmarks. We're a small, early team (post‑Series A) where individual...  ...re building out our post‑training research team and hiring 2–3 Research Scientists to work together on this mission.... 
    Training
    Internship
    Shift work

    Ersilia

    San Francisco, CA
    4 days ago
  •  ...Scale Labs in San Francisco seeks a Research Scientist focused on Safety Post-Training to advance post-training methods and interpretability for frontier AI systems. You will design pipelines, evaluate safety properties, and help translate findings into practical guidelines... 
    Training

    Scale AI

    San Francisco, CA
    15 hours ago
  • $300k

     ...applying to this role, you will be considered for Research Scientist positions at Stellon Labs pushing the frontier of efficient...  ...breakthroughs in quantization, efficient training and inference, distillation, and post-training, as well as foundational research on training... 
    Training
    Full time

    Stellon Labs Inc

    San Francisco, CA
    5 days ago
  • $225k - $300k

     ...Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access, and those...  ...work on: Verifiable reinforcement learning at scale Mid-training and post-training of foundation models Novel objectives derived... 
    Training
    Work at office
    Immediate start

    Latent

    San Francisco, CA
    2 days ago
  •  ...About the Role You will build the base intelligence layer for robotics. We train large‑scale robot foundation models from massive multimodal datasets spanning video, proprioception, action traces, language, and more. You will design and run the core large‑scale training... 
    Training

    Generalist

    San Francisco, CA
    3 days ago
  • $250k

     ...About AfterQuery AfterQuery builds the training data and evaluation infrastructure that frontier...  ...benchmarks. We are a small, early team (post Series A) where individual contributors...  ...and measured. Working directly with research teams at top AI labs, you’ll experiment with... 
    Training

    AfterQuery

    San Francisco, CA
    2 days ago
  • $285k - $380k

     ...is a quickly growing group of committed researchers, engineers, policy experts, and...  ...Role We are seeking a People Research Scientist to join our People Data Solutions team....  ...an equivalent combination of education, training, and/or experience. Required Field of Study... 
    Training
    Visa sponsorship

    Anthropic

    San Francisco, CA
    4 days ago
  • $170k - $220k

     ...Introduction The Center for AI Safety is a research and field-building nonprofit located in...  ...and technical research. As a research scientist, you will pursue a variety of research projects...  ...). Have experience launching and training distributed ML jobs. Communicate clearly... 
    Training
    Work at office

    Center for AI Safety

    San Francisco, CA
    3 days ago
  •  ...and externally. Collaborate with Product & Engineering to ship research‑backed features. Contribute to open‑source repos and author technical...  ...experimental work; hands‑on experience with large‑scale model training, transformer architectures, reinforcement‑learning techniques,... 
    Training

    Workato

    San Francisco, CA
    15 hours ago
  •  ...realize all of our product goals. As a Machine Learning Scientist at Sesame, you are a research-oriented person with experience in NLP, Speech, and/or...  ...model architectures, data curation, model evaluation, training & inference infrastructure, research, and experimentation... 
    Training
    Full time
    Contract work
    Flexible hours

    SESAME

    San Francisco, CA
    3 days ago
  • $150k - $250k

     ...grasping and more dexterous behaviors in unstructured environments Research and implement state-of-the-art robot learning policies,...  ...production robot fleets Optimize robot policies for distributed training at scale and real-time edge deployment Ship production quality,... 
    Training

    Deft AI, Inc.

    San Francisco, CA
    2 days ago
  •  ...mission to make robots commonplace. The team is looking for a Research Scientist to help architect and deploy the machine learning models...  ...vision‑language‑action (VLA) models. What You’ll Be Doing: Training, deploying, and maintaining manipulation models on physical... 
    Training

    Cubiq Recruitment

    San Francisco, CA
    5 days ago
  •  ...professional programmers, using a combination of inventive research, design, and engineering. Our organization is very...  ...debate, crazy ideas, and shipping code. Research Scientist Cursor is building the future of coding. We train frontier coding agents and scale RL on real user... 
    Training

    Cursor

    San Francisco, CA
    5 days ago
  • $120k - $250k

     ...grasping and more dexterous behaviors in unstructured environments Research and implement state-of-the-art robot learning policies,...  ...production robot fleets Optimize robot policies for distributed training at scale and real-time edge deployment Ship production quality,... 
    Training

    Rainfall Ventures

    San Francisco, CA
    2 days ago
  • $250k - $400k

     ...opportunity to work on genuinely novel AI for Science research, combining frontier reasoning, post-training and reinforcement learning with a proprietary...  ...Engineer, Applied Researcher or experienced Research Scientist. What matters most is hands‑on ownership of reasoning... 
    Training

    techire ai

    San Francisco, CA
    5 days ago
  • $204k - $259k

     ...initiate and foster collaborations with other research teams in Alphabet. AI Foundations areas...  ...hybrid role, you will report to a Principal Scientist. Responsibilities Participate in Waymo’s Foundation World Model post‑training and evaluation Research and develop cutting... 
    Training
    Temporary work
    Remote work

    Waymo

    San Francisco, CA
    5 days ago
  •  ...Traverse is a research data lab building reinforcement learning environments...  ...else has figured out how to train models on. We work directly...  ...About the Role As a Research Scientist, you will design and build RL...  ...integrate environments into their post-training pipelines Your... 
    Training

    Traverse PC Inc

    San Francisco, CA
    1 day ago
  •  ...funding from major venture capital firms. About this role As a Research Scientist you will join a small, focused team of researchers and...  ...prefix tuning, state‑space methods). Investigate synthetic training data generalization and develop self‑study pipelines that allow... 
    Training
    Work at office

    Doist

    San Francisco, CA
    2 days ago
  •  ...About the role We’re looking for a top‑tier Research Scientist to join our tech team. Your core responsibilities will be to: Lead Research for...  ...‑art retrieval and ranking systems for AI agents Prototype, train, and evaluate new models for factual search and multimodal understanding... 
    Training
    Remote work
    Flexible hours

    Linkup Inc

    San Francisco, CA
    1 day ago
  •  ...founding team brought together leading researchers in this space and top silicon valley operators...  ...more. About the role As an AI Research Scientist, you will conduct groundbreaking...  ...and deep-learning frameworks Experience training and evaluating large models on protein,... 
    Training

    Chaidiscovery

    San Francisco, CA
    3 days ago
  • $300k

     ...Research Scientist — Frontier World Models & RL A stealth, exceptionally well-backed applied AI lab...  ...with real users and a data flywheel to train against the moment beta launches. The open...  ...in an autoregressive setting. RL / Post-Training Advance RL post-training for coding... 
    Training
    Remote work
    Visa sponsorship
    Relocation package

    Harnham

    San Francisco, CA
    4 days ago
  •  ...Velvet is a data research company building the datasets that power the next generation of multimodal AI....  ...more human by producing high-quality audiovisual training data for frontier labs. We’re hiring a Research Scientist to develop and fine‑tune models for video and audio... 
    Training
    Immediate start
    Shift work

    Velvet

    San Francisco, CA
    5 days ago
  •  ...consistency, and genuine care. The Role Research at Aristotle works on the open problems in...  ...that can be evaluated reliably at scale. Train models where prompting falls short, for student...  ...experience with some of: LLM evaluation, post‑training/fine‑tuning, agent systems,... 
    Training
    Full time
    Contract work
    Summer work
    Flexible hours

    Heyaristotle

    San Francisco, CA
    1 day ago
  • $176k - $304k

     ...; San Francisco, CA USA As a Machine Learning Research Scientist I/II in LLM Inference you will lead research on how we train and serve large language models for scientific...  ...What You’ll Be Building Develop and optimize LLM post-training strategies including SFT, RLHF, and... 
    Training

    Lila Sciences

    San Francisco, CA
    3 days ago
  •  ...diverse architectures onto our platform, and we aim to both get more out of what they've already trained and shape how the next generation of models is designed. Department: Research Location: San Francisco What You'll Do Lead a research agenda on real‑time interactive... 
    Training
    Visa sponsorship
    Relocation package

    Reactor.am

    San Francisco, CA
    15 hours ago
  •  ...systems . This role is for an experienced scientist who thrives both in innovating...  ...reasoning, and deep content extraction. Research, evaluate, and integrate the latest vision...  ...knowledge of quantization/LoRA/efficient training. Proficiency with deep learning frameworks... 
    Training

    Tensorlake Inc.

    San Francisco, CA
    15 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Scientist - Post Training. Be the first to apply!