Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Scientist, Post-Training

$150k - $250k

David Joseph & Company

Job Description

Job Description

San Francisco, CA · On-site · Full-time Compensation: $150,000–$250,000 base + profit sharing (total cash ~$250K–$450K) + equity

About the Company

Our client builds the training data and evaluation infrastructure that frontier AI labs use to improve their models, partnering with leading labs to design high-signal datasets and run rigorous evaluations that go beyond static benchmarks. It's a small, early team where individual contributors have direct impact on how the next generation of models learns and improves. The founding team comes from top quant firms, big tech, and leading AI research labs.

Founded 2025 · 11–50 people · Industry: AI / ML

The Role

Prove that the client's data works — design and run training experiments that isolate the impact of its datasets on model behavior (SFT and RL post-training), and turn the results into defensible evidence for partner labs.

Tech stack: LLM post-training (SFT, RL).

What you'll be doing

  • Run controlled SFT and RL experiments to measure dataset impact on model performance
  • Quantify lift across capabilities including reasoning, tool use, long-horizon tasks, and domain-specific workflows
  • Communicate findings with partner labs to drive sales
  • Work with internal delivery leads to iterate on data quality based on experimental results
Requirements
  • Strong familiarity with LLM training and evaluation methodologies
  • Able to design lightweight experiments and extract actionable insights from messy results
  • Comfortable working across multiple domains including finance, software engineering, and policy
Nice to Haves
  • Has run controlled post-training experiments end to end and can point to a specific data intervention that shifted model behavior measurably
  • Comfortable reading messy experimental results without needing clean data to find signal
  • Strong quantitative instincts paired with SWE ability — can actually ship the experiment, not just design it
  • Has worked adjacent to or inside frontier labs or eval orgs, with a real sense of what "high-signal data" means
Why Join
  • Frontier-lab leverage: your work directly shapes the datasets leading AI labs use to train next-gen models
  • Strong backing and pedigree: a founding team from top quant firms, big tech, and leading AI research labs
  • High cash comp: base plus profit sharing pushes total cash to ~$250K–$450K, with equity on top
  • Build, don't theorize: experimental, high-leverage IC work at the edge of model development
Details
  • Location — San Francisco, CA
  • Work policy — On-site
  • Compensation — $150,000–$250,000 base + profit sharing (total cash ~$250K–$450K) + equity
  • Visa sponsorship — Not available
  • Employment type — Full-time

Vacancy posted 23 days ago
Similar jobs that could be interesting for youBased on the Research Scientist, Post-Training in San Francisco, CA vacancy
  • We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards...  ...overview:We are seeking an exceptional Research Scientist to join our team, focusing on alignment and post-training techniques for large-scale video generation... 
    Training
    Relocation

    Genmo

    San Francisco, CA
    3 days ago
  • $216k - $270k

    Scale Labs, Research Scientist — Safety Post TrainingAs the leading data and evaluation partner for frontier AI companies, Scale plays an integral...  ...this vision.As a Research Scientist working on Safety Post-Training you will develop and apply post-training methods and... 
    Training
    Full time

    Scale AI

    San Francisco, CA
    4 days ago
  • $120.7k - $238.6k

    The OpportunityAdobe Research is looking for research scientists in Generative AI to join a world-class research...  ...Experience on large-scale generative model training· Experience of working with large-...  ...in Colorado (as listed on the job posting), the application window will... 
    Training
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    4 days ago
  •  ...Researcher Position As one of our researchers - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep intuition on modern models and training. Like to argue how search, recommendations, and transformer models can... 
    Training

    Parallel Web Systems

    San Francisco, CA
    2 days ago
  •  ...workersResponsibilitiesWe are looking for an exceptional AI Research Scientist to join our growing team. In this role, you will be responsible...  ...work; hands‑on experience with large‑scale model training, transformer architectures, reinforcement‑learning techniques... 
    Training
    Remote work
    Flexible hours

    Workato

    San Francisco, CA
    4 days ago
  • $160k - $220k

     ...enjoy multi-year runway.About the RoleWe’re looking for an AI Research Scientist to advance the methodological frontier of AI in healthcare....  ...strategy. Your work may include novel architectures, new training or evaluation techniques, long-horizon research bets, peer-reviewed... 
    Training
    Temporary work
    Work at office
    Monday to Friday
    Monday to Thursday

    Sprinter Health

    San Francisco, CA
    15 hours ago
  •  ...Research Scientist Engineering · Full-time · San Francisco; New York Our mission is to automate coding. The first step in our journey...  ...Research Scientist Cursor is building the future of coding. We train frontier coding agents and scale RL on real user data to make... 
    Training
    Full time

    Anysphere

    San Francisco, CA
    2 days ago
  •  ...Chai Discovery Chai is a research lab working on AI to unlock biology...  ...We are hiring research scientists with outlier insight who can drive...  ...the work we pursue. From pre-training large diffusion models and architecture design, to post-training and inference time scaling... 
    Training

    Chai Discovery

    San Francisco, CA
    2 days ago
  •  ...Research Scientist ThirdLayer is solving one of the hardest problems in deploying agents: AI models are generic, but...  ...everything we build. The Role We're building the training loop that makes agents specific: post-training models on the workflows, context, and... 
    Training

    Thirdlayer (yc W25)

    San Francisco, CA
    1 day ago
  • $400k

     ...prototype and early commercial traction across several high-profile industry verticals. The role As a Senior Research Scientist, your focus is post-training - curating data, fine-tuning pre-trained speech models, and building the evaluation infrastructure that... 
    Training
    Relocation package
    Shift work

    techire ai

    San Francisco, CA
    1 day ago
  •  ...Overview Join our core R&D team building end-to-end automated research systems . Zochi publishes the first fully-AI generated...  ...at the forefront of long-horizon agentic capabilities, post-training for open-ended goals, and environment development. Publish... 
    Training

    Intology

    San Francisco, CA
    2 days ago
  • $234.3k - $349k

     ...future of work with AI. About the roleAI research at WRITER isn't just about publishing...  ...deployments in the world. As an AI research scientist, you'll be at the center of that work....  ...possible. The work you do here — on post-training, planning, multi-step reasoning, and... 
    Training
    Full time
    Work at office
    Local area

    Writer

    San Francisco, CA
    15 hours ago
  • $250k

     ...About Us We build training data and evaluation infrastructure that frontier...  .... We're a small, early team (post–Series A) where individual contributors...  ...re building out our post-training research team and hiring 2–3 Research Scientists to work together on this mission.... 
    Training
    Full time
    Internship
    Shift work

    Product Pulse

    San Francisco, CA
    12 days ago
  • $225k - $300k

     ...Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access...  ...on: Verifiable reinforcement learning at scale Mid-training and post-training of foundation models Novel objectives derived... 
    Training
    Work at office
    Immediate start

    Latent

    San Francisco, CA
    4 days ago
  •  ...a world-class team of engineers, designers, marketers, sellers, researchers, and operational experts to achieve our mission. Job: As one of our Researchers - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep... 
    Training
    Work at office
    Visa sponsorship
    Flexible hours

    Parallel Web Systems Inc

    San Francisco, CA
    5 days ago
  •  ...sessions, it retains scraps at best. We train models to study your world and anticipate...  ...You will join a small, focused team of researchers and engineers working at the frontier of learning and memory. As a Research Scientist, you'll design experiments, develop new recipes... 
    Training
    Work at office

    Engram

    San Francisco, CA
    3 days ago
  • $200k - $335k

     ...The role As a research scientist, you will design, implement, and optimize the large-scale training infrastructure that powers our frontier reinforcement learning stack...  ...Partner with researchers to bring frontier post-training capabilities into production deployments... 
    Training
    Full time
    Work at office
    Visa sponsorship
    Relocation package

    Applied Compute

    San Francisco, CA
    1 day ago
  •  ...Researcher Position at Hedra Hedra is building a world-class Physical AI research team to push the boundaries of action-conditioned...  ...modeling for embodied systems Design novel architectures, training objectives, and evaluation frameworks for VLMs, VLAs, and world... 
    Training
    Work at office

    HEDRA INC

    San Francisco, CA
    2 days ago
  •  ...Machine Learning Scientist Sesame believes in a future where computers are lifelike -...  ...Learning Scientist at Sesame, you are a research-oriented person with experience in NLP,...  ...architectures, data curation, model evaluation, training & inference infrastructure, research,... 
    Training
    Full time
    Contract work
    Flexible hours

    SESAME

    San Francisco, CA
    3 days ago
  • $125k - $225k

     ...FutureSearch is looking for exceptional Research Scientists to evaluate and improve state-of-the-art forecasting and agentic LLM web research...  .... We work with frontier labs on research, evaluation, and training. You are a talented researcher with background in math, physics... 
    Training
    Remote work
    Flexible hours

    Future Research Corp

    San Francisco, CA
    3 days ago
  •  ...AfterQuery AfterQuery is an applied research lab curating data solutions for foundation...  ...data works . You will design and run training experiments that isolate the impact of our...  ...behavior. This includes SFT and RL-based post-training, where you'll measure how different... 
    Training
    Shift work

    AfterQuery

    San Francisco, CA
    4 days ago
  • $250k - $325k

     ...backgrounds in technology — from AI research to systems engineering to product...  ...looking for a talented Research Scientist with a strong background in...  ...plus : ~ Large-scale model training ~ Data curation for pretraining or post-training ~ Tokenizers and VAEs... 
    Training
    Full time

    World Labs

    San Francisco, CA
    1 day ago
  •  ...physical goods as easily as they post online, and we're building...  ...is Varun Jampani, a leading researcher who co-authored Dreambooth and...  ..., was formerly a Principal Scientist at Amazon. Together, we're pioneering...  ..., and author proprietary training paradigms when existing open-... 
    Training

    Arcade Studio Inc

    San Francisco, CA
    2 days ago
  •  ...Senior Research Scientist Sciforium is an AI infrastructure company developing next-generation multimodal AI models and a proprietary,...  ...generative media, model architecture, optimization, and scalable training systems. You will work hands-on with modern ML frameworks,... 
    Training
    Flexible hours

    Sciforium

    San Francisco, CA
    1 day ago
  •  ...Research Scientist / Machine Learning Scientist Location: SF Bay Area/Hybrid / Remote Type: Full-Time About the Role: The Client...  ...home here. We're looking for: Hands-on experience training large-scale models, including reward models, preference models... 
    Training
    Full time
    Remote work

    Lead Allies Inc.

    San Francisco, CA
    5 days ago
  •  ...5, when Snorkel started as a research project in the Stanford AI Lab...  ...organizations to empower scientists, engineers, financial experts...  ...datasets that drive frontier model training and evaluation based on...  ...externally through publications, blog posts, conference talks, and... 
    Training
    Full time
    Local area

    Snorkel AI

    San Francisco, CA
    2 days ago
  • $300k

     ...applying to this role, you will be considered for Research Scientist positions at Stellon Labs pushing the frontier of efficient...  ...breakthroughs in quantization, efficient training and inference, distillation, and post-training, as well as foundational research on training... 
    Training
    Full time

    Stellon Labs Inc

    San Francisco, CA
    2 days ago
  • $285k - $380k

     ...is a quickly growing group of committed researchers, engineers, policy experts, and...  ...We are seeking a Recruiting Research Scientist to join our People Data Solutions team....  ...an equivalent combination of education, training, and/or experience Required field of... 
    Training
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    14 days ago
  • $204k - $259k

     ...initiate and foster collaborations with other research teams in Alphabet. AI Foundations areas...  ...role, you will report to a Principal Scientist. You will: Participate in Waymo's Foundation World Model post-training and evaluation Research and develop cutting... 
    Training
    Full time
    Temporary work
    Remote work

    Latent Logic

    San Francisco, CA
    1 day ago
  • $300k - $320k

     ...is a quickly growing group of committed researchers, engineers, policy experts, and...  ...We're seeking an exceptional Research Scientist to join our Life Sciences team at Anthropic...  ...capabilities on scientific tasks through post-training, evaluation design, and RL environment... 
    Training
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Scientist, Post-Training. Be the first to apply!