Research Scientist (post-training)
Genmo
We are Genmo, a research lab developing the world’s most sophisticated video world models to understand, simulate, and interact with the physical world. Our mission is to unlock the right brain of AGI. Join us in advancing physical intelligence and enabling robots to learn and act in a changing world.Role overview:We are seeking an exceptional Research Scientist to join our team, focusing on alignment and post-training techniques for large-scale video generation models. In this role, you will be at the forefront of ensuring our diffusion-based video models reliably produce high-quality, physically accurate and safe outputs that match human preferences and values.Key responsibilities:Lead research initiatives in alignment and post-training methods for video generation models, focusing on improved quality, reliability, and adherence to human intentDesign and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF) pipelines for video generation modelsDevelop robust evaluation frameworks to measure model alignment, safety, and output qualityCreate and optimize data collection pipelines for human feedback and preferencesDesign and conduct experiments to validate alignment techniques and their scaling propertiesCollaborate with cross-functional teams to integrate alignment improvements into our production pipelineStay at the cutting edge of the field by regularly reviewing academic literature in both generative AI and alignmentMentor junior researchers and foster a culture of responsible AI developmentWork closely with product teams to ensure alignment methods enhance rather than inhibit model capabilitiesQualifications:Ph.D. in Computer Science, Artificial Intelligence, Machine Learning, or a closely related fieldMust have:Strong publication record in top-tier conferences (e.g., NeurIPS, ICML, ICLR) with a focus on reinforcement learning, alignment, or generative modelsExtensive experience implementing and optimizing large-scale training pipelines using PyTorchDeep understanding of reinforcement learning techniques, particularly RLHFExperience with distributed training systems and large-scale experimentsProven track record in designing and implementing robust evaluation frameworksExcellent communication skills with the ability to explain complex technical concepts to diverse audiencesStrong software engineering skills and experience with complex shared codebasesIdeal candidate will have:Experience with diffusion models or other generative architecturesBackground in fine-tuning large language models or generative modelsExperience working with human feedback data collection and annotation pipelinesStrong aesthetic sense and understanding of video quality assessmentFamiliarity with alignment techniques such as constitutional AI or debateTrack record of successful collaboration with product teamsExperience with perceptual quality metrics and human evaluation designContributions to open-source projects in AI alignment or generative AIAdditional InformationThe role is based in the Bay Area (San Francisco). Candidates are expected to be located near the Bay Area or open to relocation.Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.LocationSan Francisco HQEmployment TypeFull timeDepartmentResearch
$225k - $300k
...Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access... ...on: Verifiable reinforcement learning at scale Mid-training and post-training of foundation models Novel objectives derived...TrainingFull timeWork at officeImmediate start$216k - $270k
Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner... .... The range displayed on each job posting reflects the minimum and maximum target... ...performance, and relevant education or training. Scale employees in eligible roles are...TrainingFull time$165.6k - $207k
...'s leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. This role is on the evaluation pod within...TrainingFull time$120.7k - $238.6k
The OpportunityAdobe Research is looking for research scientists in Generative AI to join a world-class research... ...Experience on large-scale generative model training· Experience of working with large-... ...in Colorado (as listed on the job posting), the application window will...TrainingFull timeTemporary workLocal areaWorldwide$117.2k - $313.7k
....The ExperienceSalesforce AI Research is a global leader in Enterprise... ...in multimodal AI; trained state-of-the-art large language... ...for entrepreneurial Research Scientists who want to build, ship, and... ...processing.Core Modeling and Post-Training: Pre-training and post...TrainingFull timeWorldwide- ...a world-class team of engineers, designers, marketers, sellers, researchers, and operational experts to achieve our mission. Job: As one of our Researchers - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep...TrainingWork at officeVisa sponsorshipFlexible hours
- ...Researcher Position at Hedra Hedra is building a world-class Physical AI research team to push the boundaries of action-conditioned... ...modeling for embodied systems Design novel architectures, training objectives, and evaluation frameworks for VLMs, VLAs, and world...TrainingWork at office
$200k - $335k
...The Role As a research scientist, you will design, implement, and optimize the large-scale training infrastructure that powers our frontier reinforcement learning stack.... ...Partner with researchers to bring frontier post-training capabilities into production deployments...TrainingFull timeWork at officeVisa sponsorshipRelocation package- ...AI David AI is the first audio data research company. We bring an R&D approach to data... ...and new use cases emerge, high-quality training data is the bottleneck. This is where David... .... About This Role As a Research Scientist at David AI you'll build cutting-edge speech...TrainingWork at office
- ...Senior Research Scientist Sciforium is an AI infrastructure company developing next-generation multimodal AI models and a proprietary,... ...generative media, model architecture, optimization, and scalable training systems. You will work hands-on with modern ML frameworks,...TrainingFlexible hours
$142.7k - $270.95k
...Opportunity The Speech AI Lab at Adobe Research is seeking a Senior Audio Research Scientist to join our speech generative AI... ...state-of-the-art AI models and training procedures, including foundation... ...Colorado (as listed on the job posting), the application window will...TrainingTemporary workLocal areaWorldwide- ...idler is a frontier data research lab. We build the evals and environments... ...labs use to measure and train their models. After raising... ...About the role As a Research Scientist at idler, you'll own measuring... ...labs - to design novel post-training recipes and data quality...TrainingWork at officeRelocation package
$155k - $269k
...realistic, scalable, controllable, and efficient simulation. As a Research Scientist in World Models, you will develop algorithms and... ...rich generative priors for downstream planning, testing, and training. You will... Conduct fundamental and applied research...TrainingFull timeWork at officeWork from homeFlexible hours$147k - $210k
Drive projects by defining key research questions.Design, implement, and evaluate experiments... ...:3 years of experience with training, evaluating, and interpreting large language... ...simulation.We are looking for a Research Scientist to develop cutting-edge social simulation...Training$204k - $259k
...initiate and foster collaborations with other research teams in Alphabet. AI Foundations areas... ...role, you will report to a Principal Scientist. Responsibilities Participate in Waymo’s Foundation World Model post‑training and evaluation Research and develop cutting...TrainingTemporary workRemote work$160k - $220k
...enjoy multi-year runway.About the RoleWe’re looking for an AI Research Scientist to advance the methodological frontier of AI in healthcare.... ...strategy. Your work may include novel architectures, new training or evaluation techniques, long-horizon research bets, peer-reviewed...TrainingTemporary workWork at officeMonday to FridayMonday to Thursday$234.3k - $349k
...future of work with AI. About the roleAI research at WRITER isn't just about publishing... ...deployments in the world. As an AI research scientist, you'll be at the center of that work.... ...possible. The work you do here — on post-training, planning, multi-step reasoning, and...TrainingFull timeWork at officeLocal area$184k - $274.6k
Please note this posting is to advertise potential job opportunities. This exact role may... ...AI, we are leading frontier AI research across Cisco. Our mission is to advance... ...We explore new model architectures, pre-training and post-training methods, inference optimization...TrainingFull timeTemporary workLocal areaFlexible hours$54 - $60 per hour
...problems lie in enterprise domains, behind closed doors. Our research team's goal is to push the frontier of "domain adaptation" -... ...together multiple methods to create new recipes for efficient post-training. Evaluation of LLMs and AI systems. Your qualifications...TrainingHourly payInternshipWorldwide$131.2k - $204.1k
Women’s Reproductive Health Research Scholar PositionPHYSICIAN SCIENTIST, Women’s Reproductive Health Research Scholar... ...year of postgraduate residency training in obstetrics-gynecologyHave... ...or PD/PI on other such awardsThe posted UC salary scales set the minimum pay...Training- ...Job Description Job Description Research Scientist / Research Engineer — San Francisco I’m working with an early-stage AI startup... ...researchers/engineers with depth in areas like: • LLM post-training • Reinforcement learning • Agents + reasoning • Long...Training
$290.4k - $363k
...operate at the intersection of cutting-edge research, large-scale engineering, and real-... ...advances in large-language models, post-training, evaluation, and agentic/RL environments... ..., measured, and deployed.As a Research Scientist Manager, you will lead a world-class team...TrainingFull time$127k - $333.7k
...Medicine faculty to lead clinical outcomes research in any subspecialty of medicine. The... ...Board Certified in their subspecialty. Training and experience in clinical research is required... ...qualifications upon submission. The posted UC salary scales set the minimum pay...Training- ...Neuroscience (formerly Forest Neurotech) Arbor is a nonprofit research organization working on one of the hardest open problems in... ...have: A strong quantitative background: a PhD, or comparable training in an equally demanding environment such as industry Real depth...Training
$230k - $400k
...as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together... ...team works across many parts of a large stack that includes training, inference, system design and data collection. Some of our core...TrainingWork experience placementWork at officeHome officeVisa sponsorshipRelocation packageFlexible hours$114.2k - $306.6k
...are not duplicating efforts.*Salesforce Research advances state-of-the-art AI techniques... ...is looking for outstanding AI Research Scientists / Research Engineers.**Our team... ...Modeling:** Machine learning methodology, pre-training/post-training (RLHF, DPO), and model distillation...TrainingFull time- ...preservation protocol Ensure accurate sample processing and on-time data reporting Qualifications You’ve completed graduate training or have significant industry experience in a STEM related field. You have 4+ years experience and an excellent record of...TrainingFlexible hours
- Research Scientist Engineering · Full-time · San Francisco; New York Our mission is to automate coding. The first step in our journey is to build... ...Scientist Cursor is building the future of coding. We train frontier coding agents and scale RL on real user data to make...TrainingFull time
$400k
...prototype and early commercial traction across several high-profile industry verticals. The role As a Senior Research Scientist, your focus is post-training - curating data, fine-tuning pre-trained speech models, and building the evaluation infrastructure that validates...TrainingRelocation packageShift work- ...the right moment. We are looking for a researcher who can turn that ambition into something... ...evaluation and harness design , not model training. You will: Define evaluation dimensions... ...can later support reward design and post-training This is not a role for running...Training
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist (post-training). Be the first to apply!
- scientist ii San Francisco, CA
- regulatory scientist San Francisco, CA
- principal scientist San Francisco, CA
- graduate scientist San Francisco, CA
- manufacturing scientist San Francisco, CA
- analytical scientist San Francisco, CA
- senior research scientist San Francisco, CA
- genomics scientist San Francisco, CA
- support scientist San Francisco, CA
- application scientist San Francisco, CA




