Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Scientist - Post Training

$250k
Full-time

Product Pulse

About Us

We build training data and evaluation infrastructure that frontier AI labs use to improve their models. We partner with the world's leading labs to design high-signal datasets and run rigorous evaluations that go beyond static benchmarks. We're a small, early team (post–Series A) where individual contributors have direct impact on how the next generation of models learns and improves.

The Role

We're building out our post-training research team and hiring 2–3 Research Scientists to work together on this mission. Your job is to prove that our data works. You'll design and run training experiments that isolate the impact of our datasets on model behavior, including SFT and RL-based post-training, to measure how different data sources shift capability, generalization, and alignment. Working closely with partner labs, you'll turn our datasets into clear, defensible evidence: this data this improvement under these conditions. It's experimental, high- leverage work at the edge of model development.

What You'll Do

  1. Run controlled SFT and RL experiments to measure the impact of our datasets on model performance.
  2. Quantify lift across capabilities — reasoning, tool use, long-horizon tasks, and domain-specific workflows. Share findings directly with partner labs to deepen relationships and drive sales.
  3. Collaborate with internal SPLs to iterate on data quality based on your results.
  4. Work closely with the other Research Scientists on this team to build shared experimental infrastructure and benchmarks.

What We're Looking For

  1. Strong familiarity with LLM training and evaluation methodologies (SFT, RL post-training).
  2. Genuine obsession with how data structure, selection, and quality drive model behavior.
  3. Ability to design lightweight experiments, move fast, and extract actionable insights from messy results. 
  4. Comfort working across domains — you'll touch finance, software engineering, policy, and more.
  5. A bias toward building over theorizing.

Must-Have Requirements

  1. Strong familiarity with LLM training and evaluation methodologies, including SFT and RL post-training.
  2. Genuine obsession with how data structure, selection, and quality drive model behavior.
  3. Ability to design lightweight experiments, move fast, and extract actionable insights from messy results. Comfort working across domains — finance, software engineering, policy, and more.
  4. Undergrad or master's research background; pre-PhD candidates preferred.

Nice-to-Have Requirements

  1. Prior work or internship at an RL environment company, AI safety org, or benchmarking org (METR, Artificial Analysis, or equivalent).
  2. Experience running controlled training experiments end-to-end.
  3. Published research on model evaluation, post-training, or data curation.
  4. Strong SWE chops alongside research instincts. Compensation

Compensation

$250K–$450K total compensation + equity 

Requirements

  1. Run controlled SFT and RL experiments to measure dataset impact on model performance
  2. Quantify lift across capabilities including reasoning, tool use, long-horizon tasks, and domain- specific workflows
  3. Communicate findings with partner labs to drive sales
  4. Work with internal SPLs to iterate on data quality based on experimental results
  5. Strong familiarity with LLM training and evaluation methodologies
  6. Design lightweight experiments and extract actionable insights from messy results 
  7. Work across multiple domains including finance, software engineering, and policy

Vacancy posted 12 days ago
Similar jobs that could be interesting for youBased on the Research Scientist - Post Training in San Francisco, CA vacancy
  • We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards...  ...overview:We are seeking an exceptional Research Scientist to join our team, focusing on alignment and post-training techniques for large-scale video generation... 
    Training
    Relocation

    Genmo

    San Francisco, CA
    3 days ago
  • $216k - $270k

    Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner...  .... The range displayed on each job posting reflects the minimum and maximum target...  ...performance, and relevant education or training. Scale employees in eligible roles are... 
    Training
    Full time

    Scale AI

    San Francisco, CA
    4 days ago
  • $120.7k - $238.6k

    The OpportunityAdobe Research is looking for research scientists in Generative AI to join a world-class research...  ...Experience on large-scale generative model training· Experience of working with large-...  ...in Colorado (as listed on the job posting), the application window will... 
    Training
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    4 days ago
  •  ...Researcher Position As one of our researchers - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep intuition on modern models and training. Like to argue how search, recommendations, and transformer models can... 
    Training

    Parallel Web Systems

    San Francisco, CA
    2 days ago
  •  ...workersResponsibilitiesWe are looking for an exceptional AI Research Scientist to join our growing team. In this role, you will be responsible...  ...work; hands‑on experience with large‑scale model training, transformer architectures, reinforcement‑learning techniques... 
    Training
    Remote work
    Flexible hours

    Workato

    San Francisco, CA
    4 days ago
  • $160k - $220k

     ...enjoy multi-year runway.About the RoleWe’re looking for an AI Research Scientist to advance the methodological frontier of AI in healthcare....  ...strategy. Your work may include novel architectures, new training or evaluation techniques, long-horizon research bets, peer-reviewed... 
    Training
    Temporary work
    Work at office
    Monday to Friday
    Monday to Thursday

    Sprinter Health

    San Francisco, CA
    16 hours ago
  •  ...Research Scientist Engineering · Full-time · San Francisco; New York Our mission is to automate coding. The first step in our journey...  ...Research Scientist Cursor is building the future of coding. We train frontier coding agents and scale RL on real user data to make... 
    Training
    Full time

    Anysphere

    San Francisco, CA
    2 days ago
  •  ...Chai Discovery Chai is a research lab working on AI to unlock biology...  ...We are hiring research scientists with outlier insight who can drive...  ...the work we pursue. From pre-training large diffusion models and architecture design, to post-training and inference time scaling... 
    Training

    Chai Discovery

    San Francisco, CA
    2 days ago
  • $150k - $250k

     ...About the Company Our client builds the training data and evaluation infrastructure that frontier...  ...quant firms, big tech, and leading AI research labs. Founded 2025 · 11–50 people ·...  ...datasets on model behavior (SFT and RL post-training), and turn the results into defensible... 
    Training
    Full time
    Shift work

    David Joseph & Company

    San Francisco, CA
    23 days ago
  • $400k

     ...prototype and early commercial traction across several high-profile industry verticals. The role As a Senior Research Scientist, your focus is post-training - curating data, fine-tuning pre-trained speech models, and building the evaluation infrastructure that... 
    Training
    Relocation package
    Shift work

    techire ai

    San Francisco, CA
    1 day ago
  •  ...Research Scientist ThirdLayer is solving one of the hardest problems in deploying agents: AI models are generic, but...  ...everything we build. The Role We're building the training loop that makes agents specific: post-training models on the workflows, context, and... 
    Training

    Thirdlayer (yc W25)

    San Francisco, CA
    1 day ago
  •  ...Overview Join our core R&D team building end-to-end automated research systems . Zochi publishes the first fully-AI generated...  ...at the forefront of long-horizon agentic capabilities, post-training for open-ended goals, and environment development. Publish... 
    Training

    Intology

    San Francisco, CA
    2 days ago
  • $234.3k - $349k

     ...future of work with AI. About the roleAI research at WRITER isn't just about publishing...  ...deployments in the world. As an AI research scientist, you'll be at the center of that work....  ...possible. The work you do here — on post-training, planning, multi-step reasoning, and... 
    Training
    Full time
    Work at office
    Local area

    Writer

    San Francisco, CA
    16 hours ago
  • $225k - $300k

     ...Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access...  ...on: Verifiable reinforcement learning at scale Mid-training and post-training of foundation models Novel objectives derived... 
    Training
    Work at office
    Immediate start

    Latent

    San Francisco, CA
    4 days ago
  •  ...a world-class team of engineers, designers, marketers, sellers, researchers, and operational experts to achieve our mission. Job: As one of our Researchers - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep... 
    Training
    Work at office
    Visa sponsorship
    Flexible hours

    Parallel Web Systems Inc

    San Francisco, CA
    5 days ago
  •  ...sessions, it retains scraps at best. We train models to study your world and anticipate...  ...You will join a small, focused team of researchers and engineers working at the frontier of learning and memory. As a Research Scientist, you'll design experiments, develop new recipes... 
    Training
    Work at office

    Engram

    San Francisco, CA
    3 days ago
  •  ...Researcher Position at Hedra Hedra is building a world-class Physical AI research team to push the boundaries of action-conditioned...  ...modeling for embodied systems Design novel architectures, training objectives, and evaluation frameworks for VLMs, VLAs, and world... 
    Training
    Work at office

    HEDRA INC

    San Francisco, CA
    2 days ago
  •  ...Machine Learning Scientist Sesame believes in a future where computers are lifelike -...  ...Learning Scientist at Sesame, you are a research-oriented person with experience in NLP,...  ...architectures, data curation, model evaluation, training & inference infrastructure, research,... 
    Training
    Full time
    Contract work
    Flexible hours

    SESAME

    San Francisco, CA
    3 days ago
  • $200k - $335k

     ...The role As a research scientist, you will design, implement, and optimize the large-scale training infrastructure that powers our frontier reinforcement learning stack...  ...Partner with researchers to bring frontier post-training capabilities into production deployments... 
    Training
    Full time
    Work at office
    Visa sponsorship
    Relocation package

    Applied Compute

    San Francisco, CA
    1 day ago
  •  ...AfterQuery AfterQuery is an applied research lab curating data solutions for foundation...  ...data works . You will design and run training experiments that isolate the impact of our...  ...behavior. This includes SFT and RL-based post-training, where you'll measure how different... 
    Training
    Shift work

    AfterQuery

    San Francisco, CA
    4 days ago
  • $250k - $325k

     ...backgrounds in technology — from AI research to systems engineering to product...  ...looking for a talented Research Scientist with a strong background in...  ...plus : ~ Large-scale model training ~ Data curation for pretraining or post-training ~ Tokenizers and VAEs... 
    Training
    Full time

    World Labs

    San Francisco, CA
    1 day ago
  • $125k - $225k

     ...FutureSearch is looking for exceptional Research Scientists to evaluate and improve state-of-the-art forecasting and agentic LLM web research...  .... We work with frontier labs on research, evaluation, and training. You are a talented researcher with background in math, physics... 
    Training
    Remote work
    Flexible hours

    Future Research Corp

    San Francisco, CA
    3 days ago
  •  ...Senior Research Scientist Sciforium is an AI infrastructure company developing next-generation multimodal AI models and a proprietary,...  ...generative media, model architecture, optimization, and scalable training systems. You will work hands-on with modern ML frameworks,... 
    Training
    Flexible hours

    Sciforium

    San Francisco, CA
    1 day ago
  •  ...Research Scientist / Machine Learning Scientist Location: SF Bay Area/Hybrid / Remote Type: Full-Time About the Role: The Client...  ...home here. We're looking for: Hands-on experience training large-scale models, including reward models, preference models... 
    Training
    Full time
    Remote work

    Lead Allies Inc.

    San Francisco, CA
    5 days ago
  •  ...5, when Snorkel started as a research project in the Stanford AI Lab...  ...organizations to empower scientists, engineers, financial experts...  ...datasets that drive frontier model training and evaluation based on...  ...externally through publications, blog posts, conference talks, and... 
    Training
    Full time
    Local area

    Snorkel AI

    San Francisco, CA
    2 days ago
  •  ...physical goods as easily as they post online, and we're building...  ...is Varun Jampani, a leading researcher who co-authored Dreambooth and...  ..., was formerly a Principal Scientist at Amazon. Together, we're pioneering...  ..., and author proprietary training paradigms when existing open-... 
    Training

    Arcade Studio Inc

    San Francisco, CA
    2 days ago
  • $300k

     ...applying to this role, you will be considered for Research Scientist positions at Stellon Labs pushing the frontier of efficient...  ...breakthroughs in quantization, efficient training and inference, distillation, and post-training, as well as foundational research on training... 
    Training
    Full time

    Stellon Labs Inc

    San Francisco, CA
    2 days ago
  • $285k - $380k

     ...is a quickly growing group of committed researchers, engineers, policy experts, and...  ...We are seeking a Recruiting Research Scientist to join our People Data Solutions team....  ...an equivalent combination of education, training, and/or experience Required field of... 
    Training
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    14 days ago
  • $204k - $259k

     ...Research Scientist, RL for Autonomous Planning & World Modeling Waymo is an autonomous driving technology company with the mission to...  ...will: Participate in Waymo's Foundation World Model post-training and evaluation Research and develop cutting edge RL and... 
    Training
    Full time
    Temporary work
    Remote work

    Waymo

    San Francisco, CA
    3 days ago
  • $300k - $320k

     ...Research Scientist Anthropic's mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial...  ...improve model capabilities on scientific tasks through post-training, evaluation design, and RL environment development. As a... 
    Training
    Work at office
    Visa sponsorship
    Flexible hours

    Colorwave Inc

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Scientist - Post Training. Be the first to apply!