Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

RL Research Scientist: Post-Training for LLMs & Code Models

Advanced Micro Devices

Advanced Micro Devices (AMD) is hiring an AI Research Scientist focused on reinforcement learning for post-training of large language and code models. The role involves designing RL methods, reward models, and training recipes, with emphasis on scalable experiments and safety. You will collaborate with infra and product teams to land practical methods that improve task success while maintaining stability, with a strong emphasis on publication quality. #J-18808-Ljbffr Advanced Micro Devices

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the RL Research Scientist: Post-Training for LLMs & Code Models in Santa Clara, CA vacancy
  • $204k - $259k

     ...collaborations with other research teams in Alphabet. AI...  ..., generative modeling, Bayesian inference, hierarchical...  ...report to a Principal Scientist. You will:...  ...Foundation World Model post-training and evaluation Research...  ...develop cutting edge RL and Distillation techniques... 
    Training
    Full time
    Temporary work
    Remote work

    Latent Logic

    Mountain View, CA
    1 day ago
  • AMD in Santa Clara seeks a Lead AI Research Scientist specialized in reinforcement learning. You will advance post-training algorithms and work on engineering tasks with cutting-edge technology. This role requires a PhD in relevant fields and a strong publication record... 
    Training

    AMD

    Santa Clara, CA
    3 days ago
  • $192k - $304.75k

     ...now looking for a Senior Research Scientist focused on Multimodal Foundation Models and Robotics! NVIDIA is...  ...Develop large-scale AI training and inference methods...  ...the following topics: LLMs; Large vision-language...  ...January 13, 2026.This posting is for an existing vacancy... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $192k - $304.75k

     ...now looking for a Senior Research Scientist, Multi-Modal Language Models!NVIDIA is seeking a Senior...  ...of users of multi-modal LLMs.What you'll be doing:Driving...  ...recipes for training models that mix multiple...  ...until February 8, 2026.This posting is for an existing vacancy... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $228.7k - $309.4k

     ...Principal Applied Scientist who will...  ...Amazon Kiro LLM-Training team and help create...  ...most efficient coding agents for Kiro...  ...in agentic RL space:- Push the...  ...reinforcement learning and post-training...  ...large language models specialized in...  ...transition research breakthroughs into... 
    Training
    Local area
    Worldwide
    Flexible hours

    Amazon

    Santa Clara, CA
    2 days ago
  •  ...CoreLLM team, focusing on the development of large language models. The role requires AI expertise to confront real-world...  ...Applicants should have a solid background in Python, experience in training and fine-tuning LLMs, and a strong publication record. Competitive compensation... 
    Training
    Flexible hours

    Victrays

    Santa Clara, CA
    20 hours ago
  • NVIDIA is seeking a Senior Research Scientist to advance multi-modal language models and push open-source efforts. You will...  ...capabilities, design multi-modal training recipes, and collaborate with researchers...  ...state-of-the-art in multi-modal LLMs. Equity and strong benefits... 
    Training

    NVIDIA

    Santa Clara, CA
    4 days ago
  • Institute of Foundation Models in Sunnyvale, CA, invites researchers and engineers to tackle training advanced agentic language models using reasoning and tool use. You'll develop datasets, evaluations, RL infrastructure, and scalable systems across research‑engineering... 
    Training
    Visa sponsorship

    Ifm Us

    Sunnyvale, CA
    20 hours ago
  • About the Institute of Foundation Models We are a dedicated research lab for building, understanding, using...  ...data teams to tackle key challenges in training the world model on large‑scale...  ...infrastructure for training multimodal LLMs and video diffusion models at massive... 
    Training
    Visa sponsorship

    Ifm Us

    Sunnyvale, CA
    20 hours ago
  • $184k - $287.5k

    Senior Applied Research Scientist, Multimodal Foundation Models - Healthcare (Finance) NVIDIA is...  ...model architectures and training strategies for disease progression...  .... Experience using AI coding assistants and agentic...  ...August 17, 2026. This posting is for an existing... 
    Training

    Nvidia Corporation in

    Santa Clara, CA
    4 days ago
  •  ...We are hiring a AI Research Scientist, Reinforcement Learning (LLM) and Post-Training , specializing in...  ...large generative models applied to demanding...  ...-adjacent tasks (code, optimization, tool...  ...and analyze RL algorithms—policy...  ...for post-training LLMs and code models on... 
    Training

    Advanced Micro Devices

    Santa Clara, CA
    1 day ago
  • $184k - $253k

     ...pretrain, fine‑tune, and align LLMs and generative models tailored for scientific...  ...and workflows. Innovate post‑training methods, alignment, and...  ...design. Collaborate with scientists, engineers, and cross‑functional...  ..., and publish original research in top venues. Mentor... 
    Training

    Applied Materials

    Santa Clara, CA
    2 days ago
  • The Role We are looking for a Research Scientist to join the Multi-Embodiment...  ...MEGA is building foundation models for general-purpose robots...  ...building scalable data and training pipelines, designing evaluations...  ..., ICML or ICLR. Strong coding skills and hands-on experience... 
    Training
    Full time
    Work from home

    Wayve

    Sunnyvale, CA
    4 days ago
  • $192k - $304.75k

     ...now looking for a Senior Research Scientist for Human‑AI Perception &...  ...or C++.Experience with AI model training and evaluation frameworks,...  ...packages, or technical blog posts with code).Experience with multi-node...  ...models (such as recent LLMs, VLMs).A track record of applying... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $192k - $304.75k

    We are seeking an outstanding Research Scientist or Research Engineer with a passion for...  ...generation and its application to training autonomous driving models of tomorrow. As part of NVIDIA's...  ...at least until August 4, 2026.This posting is for an existing vacancy. NVIDIA... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $174k - $252k

     ...long-term applied research agenda to...  ...limitations of LLMs in long-horizon...  ...RLGF, offline RL).One or more scientific...  ...:2 years of coding experience.1...  ...generative code models or agentic systems...  ...building, training, and fine-tuning...  ...As a Research Scientist, you'll setup large... 
    Training

    Google

    Mountain View, CA
    16 hours ago
  •  ...Content Representation Models team creates a single,...  ...are looking for a Research Scientist specializing in embeddings...  ...IDs, continuous pre-training, novel representation...  ...methods for improving how LLMs and foundation models...  ...military service.Job Posting Date:07-27-2026Job Requisition... 
    Training
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    1 day ago
  • $184.7k - $324.8k

    Research Scientist / Engineer, Foundation Model Evaluation Cupertino, California, United States Software...  ...works across the full training lifecycle: including pre‑...  ...reasoning, knowledge, code, and agentic workflows....  ...accepts applications to this posting on an ongoing basis. #J-... 
    Training
    Relocation

    Apple Inc.

    Cupertino, CA
    20 hours ago
  • $185k - $400k

     ...agentic platforms. We are seeking accomplished Research Scientists in Foundation Models with expertise in pre-training and mid-training large-scale multimodal foundation...  .../mid-training of multimodal foundation models (LLMs, VLMs, Audio LMs, or similar), ideally at the staff... 
    Training
    Remote work

    Pika

    Palo Alto, CA
    2 days ago
  • $117.2k - $313.7k

     ...Salesforce AI Research is looking for...  ...outstanding AI Research Scientists and Research...  ...develops novel models, and bridges...  ...learning (RL), reasoning and...  ...research agents, and coding agents....  ...Core Modeling and Post-Training: Machine learning...  ...core focus areas: LLMs, Agentic... 
    Training

    Salesforce.Com Inc

    Palo Alto, CA
    3 days ago
  • $165k - $185k

    Company DescriptionThe Bosch Research and Technology Center North America with offices in...  ...in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable...  ...experience on foundation models, including training, fine-tuning, and promptingIn-depth... 
    Training
    Work experience placement
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    16 hours ago
  •  ...Institute of Foundation Models We are a dedicated research lab for building...  ...model training, alongside world...  ...researchers, data scientists, and engineers,...  ...methodologies for LLMs and multimodal...  ...synthesis pipelines for code, mathematics,...  ...LLM evaluation, post‑training data,... 
    Training
    Worldwide
    Visa sponsorship

    Ifm Us

    Sunnyvale, CA
    20 hours ago
  •  ...The Content Representation Models team develops unified foundation...  ...the Role We are seeking a Research Scientist specializing in embeddings...  ...including methods for improving how LLMs and foundation models...  ...retrieval. Experience with LLM pre‑training, fine‑tuning, or... 
    Training
    Full time
    Flexible hours

    Netflix, Inc.

    Los Gatos, CA
    2 days ago
  • $200k - $300k

     ...process, data, and code that no longer...  ...the reason the research is interesting....  ...for a Research Scientist to set and...  ...twelve reveals the model of the world...  ...knowing up front: we post-train open-weight...  ...or contested; RL formulations for...  ...and RL for LLMs, agent memory and... 
    Training
    Full time

    Tessera Labs

    San Jose, CA
    1 day ago
  • Tencent’s Technology Engineering Group (TEG) seeks a research-focused engineer to advance large-scale video world models, including data set design, model pre-training, SFT, RL, and downstream applications. You will analyze R&D challenges, optimize training and inference... 
    Training

    Tencent

    Palo Alto, CA
    2 days ago
  •  ...the Institute of Foundation Models We are a dedicated research lab for building, understanding...  ...‑edge foundation model training, alongside world‑class researchers, data scientists, and engineers, tackling the...  ...recipes for pre‑training and post‑training, and evaluation benchmarks... 
    Training

    Ifm Us

    Sunnyvale, CA
    20 hours ago
  • $195.2k - $262.2k

     ...enterprises from data and model training through to...  ...Token Factory needs scientists who can turn...  ...bottlenecks into research problems, publish...  ...experiments, write strong code, collaborate with...  ...reports, blog posts, and open-source artifacts...  ...understanding of LLMs, VLMs, transformer... 
    Training
    Temporary work
    Immediate start
    Remote work

    Nebius

    Palo Alto, CA
    2 days ago
  • $192k - $304.75k

     ...Applied Deep Learning Research Scientist, Efficiency!Join our ADLR...  ...Nemotron series of models to make our state-of-the...  ...neural networks for training and deployment. Topics...  ...in Python, and solid coding/software-engineering practicesA...  ...February 8, 2026.This posting is for an existing... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $174k - $252k

    Scope and drive research efforts to...  ...Experience in core ML model development....  ...Experience with LLM training and generative...  ...:2 years of coding experience. 1...  ...As a Research Scientist, you'll setup large...  ...frontier of post-training for large...  ...models. Modern LLMs are required to... 
    Training

    Google

    Mountain View, CA
    3 days ago
  • $192k - $304.75k

     ...are now seeking a Sr Research Scientist for Autonomous Vehicles...  ...can we: Enable novel training paradigms that leverage...  ...including agent behavior models, end-to-end AV...  ...work and open-source code as well as provide real...  ...January 13, 2026.This posting is for an existing vacancy... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to RL Research Scientist: Post-Training for LLMs & Code Models. Be the first to apply!