Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Post-Training AI Research Scientist: RLHF for Finance

Two Sigma

Two Sigma is advancing post-training with RLHF, DPO, and reward modeling to align LLMs with complex, multi-step financial workflows. You will define the research agenda, build scalable infrastructure, and guide evaluation frameworks for transforming quant research into production-grade AI capabilities. The role combines training, fine-tuning, context management, and model evaluation, shaping both the post-training capability and the broader research direction of the team. #J-18808-Ljbffr Two Sigma

Vacancy posted more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Post-Training AI Research Scientist: RLHF for Finance. Be the first to apply!