Lead RL Research Scientist: LLM Post-Training
Advanced Micro Devices
Advanced Micro Devices is seeking a Lead AI Research Scientist specializing in reinforcement learning to innovate in post-training for large generative models and related tasks. The role involves researching and developing RL methods, designing training recipes, and collaborating on scaling training efforts. The ideal candidate holds a Ph.D. in a relevant field and possesses a strong publication record in reinforcement learning. Benefits include a wide range of offerings as detailed on our benefits page. #J-18808-Ljbffr Advanced Micro Devices
- ...we shape the future of AI and beyond. The Role Lead AI Research Scientist, Reinforcement Learning (LLM) and Post‑Training. You specialize in reinforcement learning to advance... ...decision making). You will invent and analyze RL algorithms—policy optimization, preference‑based...Training
- AMD in Santa Clara seeks a Lead AI Research Scientist specialized in reinforcement learning. You will advance post-training algorithms and work on engineering tasks with cutting-edge technology. This role requires a PhD in relevant fields and a strong publication record...Training
- ...ideas and collective ingenuity. The Role Lead AI Research Scientist, Recursive Self Improvement, AI Safety... ..., or toolchains improve their own training signals while staying under explicit governance... ...closed‑loop training. Partner with RL scientists on intersections between...TrainingShift work
$184k - $287.5k
...desirable employers. We lead the way in High-... ...Senior Deep Learning Scientists to advance our efforts... ...come join our Nemotron LLM team. For more... ...fundamental and applied research to develop, train, fine-tune, and deploy... ...understanding.Advance post-training and alignment...TrainingFull timeWork experience placement- AMD is looking for a Lead AI Research Scientist specializing in Recursive Self Improvement and AI Safety in Santa Clara, California. This role involves researching self-improving training loops and partnering with reinforcement learning scientists to enhance AMD's AI projects...Training
$192.2k - $260k
...an elite team of world-class scientists and engineers to pioneer the... ...tools. Join the Amazon Kiro LLM-Training team and help create groundbreaking... ..., where cutting-edge research meets real-world application:... ...of reinforcement learning and post-training methodologies for large...TrainingWork at officeLocal areaWorldwideFlexible hours$204k - $259k
...foster collaborations with other research teams in Alphabet. AI... ...you will report to a Principal Scientist. You will: Participate... ...Waymo's Foundation World Model post-training and evaluation Research and develop cutting edge RL and Distillation techniques for...TrainingFull timeTemporary workRemote work$192k - $304.75k
...We are looking for a research scientist / engineer who is passionate... ...our next-generation post-training pipelines. You will... ...research for agentic RL 2) Data and training Infrastructure... ...vLLM, SGLang or TRT-LLM. Ways to stand out... ...learning for leading foundation models. Experience...Training- NeoCognition Inc. is seeking a Member of Technical Staff for research on LLM agents in Palo Alto, California. You will lead research projects and collaborate with engineers to create impactful AI systems. Essential qualifications include a solid foundation in machine learning...
- ...Solutions, Inc. is seeking a Senior Staff Research Scientist in Agentic AI & Reinforcement Learning to lead cutting-edge AI initiatives. This role... ...design responsibilities in building governed RL environments and LLM post-training pipelines. The ideal candidate will excel...Training
- ...creative, skilled, and motivated research scientists to join our founding team in advancing... ...new algorithms and methods for training AI models for enhancing the robot... ...across multiple disciplines (Robotics, RL/IL, control, perception, LLM, VLM, etc.). Work with large-scale...TrainingFull time
- A leading technology firm in California is seeking a passionate Research Scientist to advance next-generation AI hardware platforms. The role... ..., benchmarking innovative LLM architectures, and collaborating... ...in generative AI and LLM training. Competitive salary range of...Training
$224k - $356.5k
..., partner-facing Agentic AI Lead to drive the co-design of Generative... ...Product, Engineering, and Research, and your read on real... ...Nemotron, NIM, Dynamo, TensorRT-LLM) and ecosystem tools (vLLM,... .... Fine-tuning (PEFT, SFT), post-training and RL from verifiable rewards,...TrainingFull time- ...Models We are a dedicated research lab for building,... ...cutting‑edge foundation model training, alongside world‑class researchers, data scientists, and engineers, tackling... ...or related areas: LLM training/fine‑tuning, evaluations... ...of literature on RL, LLM reasoning, and tool...TrainingVisa sponsorship
$192k - $304.75k
...for an Applied Deep Learning Research Scientist, Efficiency!Join our ADLR -... ...optimize neural networks for training and deployment. Topics... ...architectures, optimizers and LLM training.Experience with modern... ...until February 8, 2026.This posting is for an existing vacancy....TrainingFull time- ..., we’re a team of scientists, engineers, machine... .... The role of the Research Scientist /... ...responsibilities:Post-training / instruction tuning... ...experience.Significant LLM post-training experienceIn... ..., ICLR, ICML, RL/DL, EMNLP, AAAI,... ...collaborating or leading an applied...Training
- ...ROLE: We are looking for an Applied Research Scientist experienced with training large language models, large... ...this role, you will explore novel LLM/LMM and image/video generation architectures... ...AI Policy” is available here.This posting is for an existing vacancy.Training
$192k - $304.75k
...re now looking for a Senior Research Scientist, Multi-Modal Language Models... ...modelsDeveloping recipes for training models that mix multiple... ...crowd:Specific multi-modal LLM research experienceExperience... ...until February 8, 2026.This posting is for an existing vacancy....TrainingFull time$262.5k - $299.6k
...Capital One has been leading the industry in using... ...touches every aspect of the research life cycle, from... ...functional team of data scientists, software engineers, machine... ..., from design through training, evaluation,... ...Engineering or related fields LLM PhD focus on NLP or...TrainingFull timePart timeLocal areaFlexible hours$137.7k - $275.4k
...AI Incubation team seeks a Research Scientist to advance AI-driven innovation... ....About the TeamZoom is the leading platform for video... ...operating at the frontier of LLM posttraining, agentic AI, and... ...multimodal architectures and post-training strategies that enable robust...TrainingFull timeWork at officeRemote workWorldwide- ...Clara, CA, headquarters 3 days per week.The role: Sr. Staff, ML Researcher - LLM Algorithmic OptimizationWhat You Will Do:d-Matrix is seeking... ...individual’s knowledge, skills, experience, education, and training. We also offer incentive opportunities that reward employees...Training3 days per week
$262.5k - $299.6k
Applied Researcher II (AI Foundations, LLM Core and Agentic AI) Overview:... ...Capital One has been leading the industry in... ...functional team of data scientists, software engineers... ...design through training, evaluation, validation... ...the time of this posting. Salaries for part-...TrainingFull timePart timeLocal areaFlexible hours- ...THE ROLE:We are hiring a AI Research Scientist - Infrastructure Engineer, Reinforcement... ...policy and value training, rollout generation, logging... ...large GPU fleets. You make RL scientists productive by improving... ...of RL training infra, LLM post-training pipelines, or large...Training
- ...ROLE:We are hiring Forward Deployed AI Research Scientist to help bring advanced AI capabilities... ..., human-in-the-loop review, and LLM-as-judge methods where appropriate.AI... ...methods, LLMs, agents, tool use, retrieval, post-training, evaluation, or applied ML systems.Strong...Training
$150k
...Research Scientist Specializing in Natural Language Processing (NLP) We are... ...-edge foundation model training, alongside world-class researchers... ...Key Responsibilities Lead the research of technology for... ...efficiency of Large Language Model (LLM) while performing target...TrainingVisa sponsorship$150k
...We are a dedicated research lab for building, understanding... ...-edge foundation model training, alongside world-class researchers, data scientists, and engineers,... ...large language model (LLM) development, your role... ...Key Responsibilities Lead research and implementation...TrainingVisa sponsorship$117.2k - $313.7k
...AI Research Scientist And Research Engineer Salesforce is the... ...career at the company leading workforce... ...Large language model (LLM)-powered agents, reinforcement learning (RL), reasoning and planning... ...Core Modeling and Post-Training: Machine learning methodology...Training$230k - $380k
...Wayve Labs Research Scientist Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software... ..., transformers, MoE, large-scale training) ~ Generative world modeling (e... ...learning (e.g., offline RL, RLHF, reward modeling) ~ Spatial...TrainingFull timeWork at officeVisa sponsorshipRelocation packageFlexible hours- ...Research Scientist Applied Intuition, Inc. is powering the future of physical... ...We have a group composed of leading experts from top... ...on reinforcement learning (RL) related topics including large-scale self-play RL, VLA post-training, large-scale closed-loop RL...TrainingFor contractorsFor subcontractorCasual workWork at officeImmediate startRemote workDay shift
$164k - $313.3k
...Senior Applied Scientist - Brand Intelligence... ...synthetic audiences, LLM-powered simulated... ...a fast-moving research literature into shipping... .../GRPO) and modern post-training tradeoffs.... ...with RLHF, RLAIF, or RL-based state alignment... ...Adobe's industry-leading offerings including...TrainingTemporary workLocal areaWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lead RL Research Scientist: LLM Post-Training. Be the first to apply!
- scientist 1 Santa Clara, CA
- image scientist Santa Clara, CA
- qc scientist Santa Clara, CA
- research scientist Santa Clara, CA
- analytical scientist Santa Clara, CA
- research scientist - biology Santa Clara, CA
- quality control scientist Santa Clara, CA
- applied scientist Santa Clara, CA
- manufacturing scientist Santa Clara, CA
- machine learning research scientist Santa Clara, CA

