Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Lead RL Research Scientist: LLM Post-Training

Advanced Micro Devices

Advanced Micro Devices is seeking a Lead AI Research Scientist specializing in reinforcement learning to innovate in post-training for large generative models and related tasks. The role involves researching and developing RL methods, designing training recipes, and collaborating on scaling training efforts. The ideal candidate holds a Ph.D. in a relevant field and possesses a strong publication record in reinforcement learning. Benefits include a wide range of offerings as detailed on our benefits page. #J-18808-Ljbffr Advanced Micro Devices

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Lead RL Research Scientist: LLM Post-Training in Santa Clara, CA vacancy
  •  ...we shape the future of AI and beyond. The Role Lead AI Research Scientist, Reinforcement Learning (LLM) and Post‑Training. You specialize in reinforcement learning to advance...  ...decision making). You will invent and analyze RL algorithms—policy optimization, preference‑based... 
    Training

    AMD

    Santa Clara, CA
    3 days ago
  • AMD in Santa Clara seeks a Lead AI Research Scientist specialized in reinforcement learning. You will advance post-training algorithms and work on engineering tasks with cutting-edge technology. This role requires a PhD in relevant fields and a strong publication record... 
    Training

    AMD

    Santa Clara, CA
    3 days ago
  •  ...ideas and collective ingenuity. The Role Lead AI Research Scientist, Recursive Self Improvement, AI Safety...  ..., or toolchains improve their own training signals while staying under explicit governance...  ...closed‑loop training. Partner with RL scientists on intersections between... 
    Training
    Shift work

    AMD

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...desirable employers. We lead the way in High-...  ...Senior Deep Learning Scientists to advance our efforts...  ...come join our Nemotron LLM team. For more...  ...fundamental and applied research to develop, train, fine-tune, and deploy...  ...understanding.Advance post-training and alignment... 
    Training
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    3 days ago
  • AMD is looking for a Lead AI Research Scientist specializing in Recursive Self Improvement and AI Safety in Santa Clara, California. This role involves researching self-improving training loops and partnering with reinforcement learning scientists to enhance AMD's AI projects... 
    Training

    AMD

    Santa Clara, CA
    3 days ago
  • $192.2k - $260k

     ...an elite team of world-class scientists and engineers to pioneer the...  ...tools. Join the Amazon Kiro LLM-Training team and help create groundbreaking...  ..., where cutting-edge research meets real-world application:...  ...of reinforcement learning and post-training methodologies for large... 
    Training
    Work at office
    Local area
    Worldwide
    Flexible hours

    Amazon

    Santa Clara, CA
    3 days ago
  • $204k - $259k

     ...foster collaborations with other research teams in Alphabet. AI...  ...you will report to a Principal Scientist. You will: Participate...  ...Waymo's Foundation World Model post-training and evaluation Research and develop cutting edge RL and Distillation techniques for... 
    Training
    Full time
    Temporary work
    Remote work

    Latent Logic

    Mountain View, CA
    1 day ago
  • $192k - $304.75k

     ...We are looking for a research scientist / engineer who is passionate...  ...our next-generation post-training pipelines. You will...  ...research for agentic RL 2) Data and training Infrastructure...  ...vLLM, SGLang or TRT-LLM. Ways to stand out...  ...learning for leading foundation models. Experience... 
    Training

    NVIDIA

    Santa Clara, CA
    3 days ago
  • NeoCognition Inc. is seeking a Member of Technical Staff for research on LLM agents in Palo Alto, California. You will lead research projects and collaborate with engineers to create impactful AI systems. Essential qualifications include a solid foundation in machine learning... 

    NeoCognition Inc.

    Palo Alto, CA
    1 day ago
  •  ...Solutions, Inc. is seeking a Senior Staff Research Scientist in Agentic AI & Reinforcement Learning to lead cutting-edge AI initiatives. This role...  ...design responsibilities in building governed RL environments and LLM post-training pipelines. The ideal candidate will excel... 
    Training

    Centific Global Solutions, Inc.

    Palo Alto, CA
    5 days ago
  •  ...creative, skilled, and motivated research scientists to join our founding team in advancing...  ...new algorithms and methods for training AI models for enhancing the robot...  ...across multiple disciplines (Robotics, RL/IL, control, perception, LLM, VLM, etc.). Work with large-scale... 
    Training
    Full time

    Dexmate

    Santa Clara, CA
    2 days ago
  • A leading technology firm in California is seeking a passionate Research Scientist to advance next-generation AI hardware platforms. The role...  ..., benchmarking innovative LLM architectures, and collaborating...  ...in generative AI and LLM training. Competitive salary range of... 
    Training

    Jobleads-US

    Palo Alto, CA
    4 days ago
  • $224k - $356.5k

     ..., partner-facing Agentic AI Lead to drive the co-design of Generative...  ...Product, Engineering, and Research, and your read on real...  ...Nemotron, NIM, Dynamo, TensorRT-LLM) and ecosystem tools (vLLM,...  .... Fine-tuning (PEFT, SFT), post-training and RL from verifiable rewards,... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...Models We are a dedicated research lab for building,...  ...cutting‑edge foundation model training, alongside world‑class researchers, data scientists, and engineers, tackling...  ...or related areas: LLM training/fine‑tuning, evaluations...  ...of literature on RL, LLM reasoning, and tool... 
    Training
    Visa sponsorship

    Ifm Us

    Sunnyvale, CA
    5 days ago
  • $192k - $304.75k

     ...for an Applied Deep Learning Research Scientist, Efficiency!Join our ADLR -...  ...optimize neural networks for training and deployment. Topics...  ...architectures, optimizers and LLM training.Experience with modern...  ...until February 8, 2026.This posting is for an existing vacancy.... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ..., we’re a team of scientists, engineers, machine...  .... The role of the Research Scientist /...  ...responsibilities:Post-training / instruction tuning...  ...experience.Significant LLM post-training experienceIn...  ..., ICLR, ICML, RL/DL, EMNLP, AAAI,...  ...collaborating or leading an applied... 
    Training

    DeepMind

    Mountain View, CA
    2 days ago
  •  ...ROLE: We are looking for an Applied Research Scientist experienced with training large language models, large...  ...this role, you will explore novel LLM/LMM and image/video generation architectures...  ...AI Policy” is available here.This posting is for an existing vacancy.
    Training

    AMD

    San Jose, CA
    3 days ago
  • $192k - $304.75k

     ...re now looking for a Senior Research Scientist, Multi-Modal Language Models...  ...modelsDeveloping recipes for training models that mix multiple...  ...crowd:Specific multi-modal LLM research experienceExperience...  ...until February 8, 2026.This posting is for an existing vacancy.... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $262.5k - $299.6k

     ...Capital One has been leading the industry in using...  ...touches every aspect of the research life cycle, from...  ...functional team of data scientists, software engineers, machine...  ..., from design through training, evaluation,...  ...Engineering or related fields LLM PhD focus on NLP or... 
    Training
    Full time
    Part time
    Local area
    Flexible hours

    Capital One

    San Jose, CA
    5 days ago
  • $137.7k - $275.4k

     ...AI Incubation team seeks a Research Scientist to advance AI-driven innovation...  ....About the TeamZoom is the leading platform for video...  ...operating at the frontier of LLM posttraining, agentic AI, and...  ...multimodal architectures and post-training strategies that enable robust... 
    Training
    Full time
    Work at office
    Remote work
    Worldwide

    Zoom

    San Jose, CA
    1 day ago
  •  ...Clara, CA, headquarters 3 days per week.The role: Sr. Staff, ML Researcher - LLM Algorithmic OptimizationWhat You Will Do:d-Matrix is seeking...  ...individual’s knowledge, skills, experience, education, and training. We also offer incentive opportunities that reward employees... 
    Training
    3 days per week

    d-Matrix

    Santa Clara, CA
    4 days ago
  • $262.5k - $299.6k

    Applied Researcher II (AI Foundations, LLM Core and Agentic AI) Overview:...  ...Capital One has been leading the industry in...  ...functional team of data scientists, software engineers...  ...design through training, evaluation, validation...  ...the time of this posting. Salaries for part-... 
    Training
    Full time
    Part time
    Local area
    Flexible hours

    Capital One

    San Jose, CA
    4 days ago
  •  ...THE ROLE:We are hiring a AI Research Scientist - Infrastructure Engineer, Reinforcement...  ...policy and value training, rollout generation, logging...  ...large GPU fleets. You make RL scientists productive by improving...  ...of RL training infra, LLM post-training pipelines, or large... 
    Training

    AMD

    Santa Clara, CA
    1 day ago
  •  ...ROLE:We are hiring Forward Deployed AI Research Scientist to help bring advanced AI capabilities...  ..., human-in-the-loop review, and LLM-as-judge methods where appropriate.AI...  ...methods, LLMs, agents, tool use, retrieval, post-training, evaluation, or applied ML systems.Strong... 
    Training

    AMD

    Santa Clara, CA
    5 days ago
  • $150k

     ...Research Scientist Specializing in Natural Language Processing (NLP) We are...  ...-edge foundation model training, alongside world-class researchers...  ...Key Responsibilities Lead the research of technology for...  ...efficiency of Large Language Model (LLM) while performing target... 
    Training
    Visa sponsorship

    Institute of Foundation Models

    Sunnyvale, CA
    1 day ago
  • $150k

     ...We are a dedicated research lab for building, understanding...  ...-edge foundation model training, alongside world-class researchers, data scientists, and engineers,...  ...large language model (LLM) development, your role...  ...Key Responsibilities Lead research and implementation... 
    Training
    Visa sponsorship

    Institute of Foundation Models

    Sunnyvale, CA
    1 day ago
  • $117.2k - $313.7k

     ...AI Research Scientist And Research Engineer Salesforce is the...  ...career at the company leading workforce...  ...Large language model (LLM)-powered agents, reinforcement learning (RL), reasoning and planning...  ...Core Modeling and Post-Training: Machine learning methodology... 
    Training

    Salesforce

    Palo Alto, CA
    1 day ago
  • $230k - $380k

     ...Wayve Labs Research Scientist Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software...  ..., transformers, MoE, large-scale training) ~ Generative world modeling (e...  ...learning (e.g., offline RL, RLHF, reward modeling) ~ Spatial... 
    Training
    Full time
    Work at office
    Visa sponsorship
    Relocation package
    Flexible hours

    Wayve

    Sunnyvale, CA
    1 day ago
  •  ...Research Scientist Applied Intuition, Inc. is powering the future of physical...  ...We have a group composed of leading experts from top...  ...on reinforcement learning (RL) related topics including large-scale self-play RL, VLA post-training, large-scale closed-loop RL... 
    Training
    For contractors
    For subcontractor
    Casual work
    Work at office
    Immediate start
    Remote work
    Day shift

    Applied Intuition

    Sunnyvale, CA
    5 days ago
  • $164k - $313.3k

     ...Senior Applied Scientist - Brand Intelligence...  ...synthetic audiences, LLM-powered simulated...  ...a fast-moving research literature into shipping...  .../GRPO) and modern post-training tradeoffs....  ...with RLHF, RLAIF, or RL-based state alignment...  ...Adobe's industry-leading offerings including... 
    Training
    Temporary work
    Local area
    Worldwide

    Adobe

    San Jose, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Lead RL Research Scientist: LLM Post-Training. Be the first to apply!