Post-Training AI Research Scientist: RLHF for Finance
Two Sigma
Two Sigma is advancing post-training with RLHF, DPO, and reward modeling to align LLMs with complex, multi-step financial workflows. You will define the research agenda, build scalable infrastructure, and guide evaluation frameworks for transforming quant research into production-grade AI capabilities. The role combines training, fine-tuning, context management, and model evaluation, shaping both the post-training capability and the broader research direction of the team. #J-18808-Ljbffr Two Sigma
Vacancy posted more than 2 months ago
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Post-Training AI Research Scientist: RLHF for Finance. Be the first to apply!
Related searches
- ai data scientist New York, NY
- ai scientist New York, NY
- scientist ii New York, NY
- machine learning scientist New York, NY
- scientist New York, NY
- quality control scientist New York, NY
- qc scientist New York, NY
- regulatory scientist New York, NY
- research scientist - biology New York, NY
- applied scientist New York, NY
