Research Scientist (post-training)
Genmo
We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.Role overview:We are seeking an exceptional Research Scientist to join our team, focusing on alignment and post-training techniques for large-scale video generation models. In this role, you will be at the forefront of ensuring our diffusion-based video models reliably produce high-quality, physically accurate and safe outputs that match human preferences and values.Key responsibilities:Lead research initiatives in alignment and post-training methods for video generation models, focusing on improved quality, reliability, and adherence to human intentDesign and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF) pipelines for video generation modelsDevelop robust evaluation frameworks to measure model alignment, safety, and output qualityCreate and optimize data collection pipelines for human feedback and preferencesDesign and conduct experiments to validate alignment techniques and their scaling propertiesCollaborate with cross-functional teams to integrate alignment improvements into our production pipelineStay at the cutting edge of the field by regularly reviewing academic literature in both generative AI and alignmentMentor junior researchers and foster a culture of responsible AI developmentWork closely with product teams to ensure alignment methods enhance rather than inhibit model capabilitiesQualifications:Ph.D. in Computer Science, Artificial Intelligence, Machine Learning, or a closely related fieldMust have:Strong publication record in top-tier conferences (e.g., NeurIPS, ICML, ICLR) with a focus on reinforcement learning, alignment, or generative modelsExtensive experience implementing and optimizing large-scale training pipelines using PyTorchDeep understanding of reinforcement learning techniques, particularly RLHFExperience with distributed training systems and large-scale experimentsProven track record in designing and implementing robust evaluation frameworksExcellent communication skills with the ability to explain complex technical concepts to diverse audiencesStrong software engineering skills and experience with complex shared codebasesIdeal candidate will have:Experience with diffusion models or other generative architecturesBackground in fine-tuning large language models or generative modelsExperience working with human feedback data collection and annotation pipelinesStrong aesthetic sense and understanding of video quality assessmentFamiliarity with alignment techniques such as constitutional AI or debateTrack record of successful collaboration with product teamsExperience with perceptual quality metrics and human evaluation designContributions to open-source projects in AI alignment or generative AIAdditional InformationThe role is based in the Bay Area (San Francisco). Candidates are expected to be located near the Bay Area or open to relocation.Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.LocationSan Francisco HQEmployment TypeFull timeDepartmentResearch
$180.6k - $225.75k
...'s leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. This role is on the evaluation pod within...TrainingFull time$120.7k - $238.6k
The OpportunityAdobe Research is looking for research scientists in Generative AI to join a world-class research... ...Experience on large-scale generative model training· Experience of working with large-... ...in Colorado (as listed on the job posting), the application window will...TrainingFull timeTemporary workLocal areaWorldwide$216k - $270k
Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner... .... The range displayed on each job posting reflects the minimum and maximum target... ...performance, and relevant education or training. Scale employees in eligible roles are...TrainingFull time$180.6k - $225.75k
...’s leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling). This role will focus on optimizing data curation and...TrainingFull time- ...Scale Labs in San Francisco seeks a Research Scientist focused on Safety Post-Training to advance post-training methods and interpretability for frontier AI systems. You will design pipelines, evaluate safety properties, and help translate findings into practical guidelines...Training
$90k - $120k
...lung, brain, muscle, heart, tumor, and other relevant organs. Support the team during surgeries and special in vivo procedures, with training provided as needed. Perform EEG surgeries in mice and support EEG studies. Who You Are You have experience working with laboratory...TrainingH1bVisa sponsorshipNight shiftDay shift- ...'re partnering with a well-funded, early-stage AI lab building research systems aimed at one of the hardest open problems in the field:... ...researchers and engineers. What you'll be working on: Build and train reinforcement learning agents capable of solving complex...Training
- ...Baseten seeks a senior ML researcher to pursue open problems at the intersection of post-training methodology and scalable inference. You will translate findings into Baseten's production platform, collaborating with research engineering to ship AI products for customers...Training
- ...professional programmers, using a combination of inventive research, design, and engineering. Our organization is very... ...debate, crazy ideas, and shipping code. Research Scientist Cursor is building the future of coding. We train frontier coding agents and scale RL on real user...Training
- ...and externally. Collaborate with Product & Engineering to ship research‑backed features. Contribute to open‑source repos and author technical... ...experimental work; hands‑on experience with large‑scale model training, transformer architectures, reinforcement‑learning techniques,...Training
$250k - $400k
...opportunity to work on genuinely novel AI for Science research, combining frontier reasoning, post-training and reinforcement learning with a proprietary... ...Engineer, Applied Researcher or experienced Research Scientist. What matters most is hands-on ownership of reasoning...Training- ...Traverse is a research data lab building reinforcement learning environments... ...else has figured out how to train models on. We work directly... ...About the Role As a Research Scientist, you will design and build RL... ...integrate environments into their post-training pipelines Your...Training
$147k - $210k
Drive projects by defining key research questions.Design, implement, and evaluate experiments... ...:3 years of experience with training, evaluating, and interpreting large language... ...simulation.We are looking for a Research Scientist to develop cutting-edge social simulation...Training- ...workersResponsibilitiesWe are looking for an exceptional AI Research Scientist to join our growing team. In this role, you will be responsible... ...work; hands‑on experience with large‑scale model training, transformer architectures, reinforcement‑learning techniques...TrainingRemote workFlexible hours
$180k - $250k
...massively accelerates certain kinds of probabilistic inference. Our ML team works on the science of training models in the thermodynamic paradigm, and we are looking for senior research and engineering talent to derive probabilistic ML theory, empirically demonstrate its...Training- ...consistency, and genuine care. The Role Research at Aristotle works on the open problems in... ...that can be evaluated reliably at scale. Train models where prompting falls short, for student... ...experience with some of: LLM evaluation, post‑training/fine‑tuning, agent systems,...TrainingFull timeContract workSummer workFlexible hours
- ...funding from major venture capital firms. About this role As a Research Scientist you will join a small, focused team of researchers and... ...prefix tuning, state‑space methods). Investigate synthetic training data generalization and develop self‑study pipelines that allow...TrainingWork at office
- ...Velvet is a data research company building the datasets that power the next generation of multimodal AI.... ...more human by producing high-quality audiovisual training data for frontier labs. We’re hiring a Research Scientist to develop and fine‑tune models for video and audio...TrainingImmediate startShift work
- ...forming a world-class team of engineers, designers, marketers, sellers, researchers, and operational experts to achieve our mission. Job: As one of our Researchers - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep...TrainingWork at officeVisa sponsorshipFlexible hours
- ...foundation models, by providing high-quality, fully licensed training datasets. As the AI data landscape undergoes a major shift,... ...larger models, but from better data. We are looking for an AI Research Scientist who is obsessed with the Data half of Algorithms. Your role...TrainingShift work
$300k
...Research Scientist — Frontier World Models & RL A stealth, exceptionally well-backed applied AI lab... ...with real users and a data flywheel to train against the moment beta launches. The open... ...in an autoregressive setting. RL / Post-Training Advance RL post-training for coding...TrainingRemote workVisa sponsorshipRelocation package- ...founding team brought together leading researchers in this space and top silicon valley operators... ...more. About the role As an AI Research Scientist, you will conduct groundbreaking... ...and deep-learning frameworks Experience training and evaluating large models on protein,...Training
- ...Mercor is seeking computational scientists specializing in atomistic and surface modeling to support a frontier AI research lab building models for materials science and the physical... ...energetics — to build high-quality training and evaluation data. Review and evaluate AI...TrainingRemote work
$200k - $325k
...Overview Hedra is building a world‑class Physical AI research team to push the boundaries of action‑conditioned world models and generative... ...modeling for embodied systems Design novel architectures, training objectives, and evaluation frameworks for VLMs, VLAs, and world...TrainingWork at office$250k
...About AfterQuery AfterQuery builds the training data and evaluation infrastructure that frontier... ...benchmarks. We are a small, early team (post Series A) where individual contributors... ...and measured. Working directly with research teams at top AI labs, you’ll experiment with...Training- ...Research Scientist, Real-Time Interactivity / Inference San Francisco · Research · Full Time Real-time interactivity can come from inference... ..., and we aim to both get more out of what they've already trained and shape how the next generation of models is designed. We'...TrainingFull timeVisa sponsorshipRelocation package
$160k - $220k
...enjoy multi-year runway.About the RoleWe’re looking for an AI Research Scientist to advance the methodological frontier of AI in healthcare.... ...strategy. Your work may include novel architectures, new training or evaluation techniques, long-horizon research bets, peer-reviewed...TrainingTemporary workWork at officeMonday to FridayMonday to Thursday- ...systems . This role is for an experienced scientist who thrives both in innovating... ...reasoning, and deep content extraction. Research, evaluate, and integrate the latest vision... ...knowledge of quantization/LoRA/efficient training. Proficiency with deep learning frameworks...Training
$129.4k - $242.2k
...next big idea could be yours! Position Overview Adobe Research is looking for research scientists working on generative AI models for image and video synthesis... ...editing Experience with large‑scale generative model training Preferred Experience Working with large‑scale datasets...Training- ...realize all of our product goals. As a Machine Learning Scientist at Sesame, you are a research-oriented person with experience in NLP, Speech, and/or... ...model architectures, data curation, model evaluation, training & inference infrastructure, research, and experimentation...TrainingFull timeContract workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist (post-training). Be the first to apply!
- molecular biology scientist San Francisco, CA
- water quality scientist San Francisco, CA
- machine learning scientist San Francisco, CA
- scientist antibody discovery San Francisco, CA
- image scientist San Francisco, CA
- machine learning research scientist San Francisco, CA
- hplc scientist San Francisco, CA
- materials scientist San Francisco, CA
- research associate scientist San Francisco, CA
- downstream processing scientist San Francisco, CA

