RL Research Scientist: Post-Training for LLMs & Code Models
Advanced Micro Devices
Advanced Micro Devices (AMD) is hiring an AI Research Scientist focused on reinforcement learning for post-training of large language and code models. The role involves designing RL methods, reward models, and training recipes, with emphasis on scalable experiments and safety. You will collaborate with infra and product teams to land practical methods that improve task success while maintaining stability, with a strong emphasis on publication quality. #J-18808-Ljbffr Advanced Micro Devices
$204k - $259k
...collaborations with other research teams in Alphabet. AI... ..., generative modeling, Bayesian inference, hierarchical... ...report to a Principal Scientist. You will:... ...Foundation World Model post-training and evaluation Research... ...develop cutting edge RL and Distillation techniques...TrainingFull timeTemporary workRemote work- AMD in Santa Clara seeks a Lead AI Research Scientist specialized in reinforcement learning. You will advance post-training algorithms and work on engineering tasks with cutting-edge technology. This role requires a PhD in relevant fields and a strong publication record...Training
$192k - $304.75k
...now looking for a Senior Research Scientist focused on Multimodal Foundation Models and Robotics! NVIDIA is... ...Develop large-scale AI training and inference methods... ...the following topics: LLMs; Large vision-language... ...January 13, 2026.This posting is for an existing vacancy...TrainingFull time$192k - $304.75k
...now looking for a Senior Research Scientist, Multi-Modal Language Models!NVIDIA is seeking a Senior... ...of users of multi-modal LLMs.What you'll be doing:Driving... ...recipes for training models that mix multiple... ...until February 8, 2026.This posting is for an existing vacancy...TrainingFull time$228.7k - $309.4k
...Principal Applied Scientist who will... ...Amazon Kiro LLM-Training team and help create... ...most efficient coding agents for Kiro... ...in agentic RL space:- Push the... ...reinforcement learning and post-training... ...large language models specialized in... ...transition research breakthroughs into...TrainingLocal areaWorldwideFlexible hours- ...CoreLLM team, focusing on the development of large language models. The role requires AI expertise to confront real-world... ...Applicants should have a solid background in Python, experience in training and fine-tuning LLMs, and a strong publication record. Competitive compensation...TrainingFlexible hours
- NVIDIA is seeking a Senior Research Scientist to advance multi-modal language models and push open-source efforts. You will... ...capabilities, design multi-modal training recipes, and collaborate with researchers... ...state-of-the-art in multi-modal LLMs. Equity and strong benefits...Training
- Institute of Foundation Models in Sunnyvale, CA, invites researchers and engineers to tackle training advanced agentic language models using reasoning and tool use. You'll develop datasets, evaluations, RL infrastructure, and scalable systems across research‑engineering...TrainingVisa sponsorship
- About the Institute of Foundation Models We are a dedicated research lab for building, understanding, using... ...data teams to tackle key challenges in training the world model on large‑scale... ...infrastructure for training multimodal LLMs and video diffusion models at massive...TrainingVisa sponsorship
$184k - $287.5k
Senior Applied Research Scientist, Multimodal Foundation Models - Healthcare (Finance) NVIDIA is... ...model architectures and training strategies for disease progression... .... Experience using AI coding assistants and agentic... ...August 17, 2026. This posting is for an existing...Training- ...We are hiring a AI Research Scientist, Reinforcement Learning (LLM) and Post-Training , specializing in... ...large generative models applied to demanding... ...-adjacent tasks (code, optimization, tool... ...and analyze RL algorithms—policy... ...for post-training LLMs and code models on...Training
$184k - $253k
...pretrain, fine‑tune, and align LLMs and generative models tailored for scientific... ...and workflows. Innovate post‑training methods, alignment, and... ...design. Collaborate with scientists, engineers, and cross‑functional... ..., and publish original research in top venues. Mentor...Training- The Role We are looking for a Research Scientist to join the Multi-Embodiment... ...MEGA is building foundation models for general-purpose robots... ...building scalable data and training pipelines, designing evaluations... ..., ICML or ICLR. Strong coding skills and hands-on experience...TrainingFull timeWork from home
$192k - $304.75k
...now looking for a Senior Research Scientist for Human‑AI Perception &... ...or C++.Experience with AI model training and evaluation frameworks,... ...packages, or technical blog posts with code).Experience with multi-node... ...models (such as recent LLMs, VLMs).A track record of applying...TrainingFull time$192k - $304.75k
We are seeking an outstanding Research Scientist or Research Engineer with a passion for... ...generation and its application to training autonomous driving models of tomorrow. As part of NVIDIA's... ...at least until August 4, 2026.This posting is for an existing vacancy. NVIDIA...TrainingFull time$174k - $252k
...long-term applied research agenda to... ...limitations of LLMs in long-horizon... ...RLGF, offline RL).One or more scientific... ...:2 years of coding experience.1... ...generative code models or agentic systems... ...building, training, and fine-tuning... ...As a Research Scientist, you'll setup large...Training- ...Content Representation Models team creates a single,... ...are looking for a Research Scientist specializing in embeddings... ...IDs, continuous pre-training, novel representation... ...methods for improving how LLMs and foundation models... ...military service.Job Posting Date:07-27-2026Job Requisition...TrainingHourly payFull timeImmediate startFlexible hours
$184.7k - $324.8k
Research Scientist / Engineer, Foundation Model Evaluation Cupertino, California, United States Software... ...works across the full training lifecycle: including pre‑... ...reasoning, knowledge, code, and agentic workflows.... ...accepts applications to this posting on an ongoing basis. #J-...TrainingRelocation$185k - $400k
...agentic platforms. We are seeking accomplished Research Scientists in Foundation Models with expertise in pre-training and mid-training large-scale multimodal foundation... .../mid-training of multimodal foundation models (LLMs, VLMs, Audio LMs, or similar), ideally at the staff...TrainingRemote work$117.2k - $313.7k
...Salesforce AI Research is looking for... ...outstanding AI Research Scientists and Research... ...develops novel models, and bridges... ...learning (RL), reasoning and... ...research agents, and coding agents.... ...Core Modeling and Post-Training: Machine learning... ...core focus areas: LLMs, Agentic...Training$165k - $185k
Company DescriptionThe Bosch Research and Technology Center North America with offices in... ...in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable... ...experience on foundation models, including training, fine-tuning, and promptingIn-depth...TrainingWork experience placementWorldwide- ...Institute of Foundation Models We are a dedicated research lab for building... ...model training, alongside world... ...researchers, data scientists, and engineers,... ...methodologies for LLMs and multimodal... ...synthesis pipelines for code, mathematics,... ...LLM evaluation, post‑training data,...TrainingWorldwideVisa sponsorship
- ...The Content Representation Models team develops unified foundation... ...the Role We are seeking a Research Scientist specializing in embeddings... ...including methods for improving how LLMs and foundation models... ...retrieval. Experience with LLM pre‑training, fine‑tuning, or...TrainingFull timeFlexible hours
$200k - $300k
...process, data, and code that no longer... ...the reason the research is interesting.... ...for a Research Scientist to set and... ...twelve reveals the model of the world... ...knowing up front: we post-train open-weight... ...or contested; RL formulations for... ...and RL for LLMs, agent memory and...TrainingFull time- Tencent’s Technology Engineering Group (TEG) seeks a research-focused engineer to advance large-scale video world models, including data set design, model pre-training, SFT, RL, and downstream applications. You will analyze R&D challenges, optimize training and inference...Training
- ...the Institute of Foundation Models We are a dedicated research lab for building, understanding... ...‑edge foundation model training, alongside world‑class researchers, data scientists, and engineers, tackling the... ...recipes for pre‑training and post‑training, and evaluation benchmarks...Training
$195.2k - $262.2k
...enterprises from data and model training through to... ...Token Factory needs scientists who can turn... ...bottlenecks into research problems, publish... ...experiments, write strong code, collaborate with... ...reports, blog posts, and open-source artifacts... ...understanding of LLMs, VLMs, transformer...TrainingTemporary workImmediate startRemote work$192k - $304.75k
...Applied Deep Learning Research Scientist, Efficiency!Join our ADLR... ...Nemotron series of models to make our state-of-the... ...neural networks for training and deployment. Topics... ...in Python, and solid coding/software-engineering practicesA... ...February 8, 2026.This posting is for an existing...TrainingFull time$174k - $252k
Scope and drive research efforts to... ...Experience in core ML model development.... ...Experience with LLM training and generative... ...:2 years of coding experience. 1... ...As a Research Scientist, you'll setup large... ...frontier of post-training for large... ...models. Modern LLMs are required to...Training$192k - $304.75k
...are now seeking a Sr Research Scientist for Autonomous Vehicles... ...can we: Enable novel training paradigms that leverage... ...including agent behavior models, end-to-end AV... ...work and open-source code as well as provide real... ...January 13, 2026.This posting is for an existing vacancy...TrainingFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to RL Research Scientist: Post-Training for LLMs & Code Models. Be the first to apply!
- machine learning scientist Santa Clara, CA
- scientist Santa Clara, CA
- quality control scientist Santa Clara, CA
- qc scientist Santa Clara, CA
- regulatory scientist Santa Clara, CA
- research scientist - biology Santa Clara, CA
- applied scientist Santa Clara, CA
- operations research scientist Santa Clara, CA
- image scientist Santa Clara, CA
- support scientist Santa Clara, CA

