Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Model Training Engineer: RLHF & LLM Fine-Tuning

$175k - $350k

Inflection AI

A pioneering AI company is seeking a Model Training Engineer to design and scale post-training pipelines for large language models. The ideal candidate will have hands-on experience with training and fine-tuning large transformer models. Responsibilities include end-to-end workflow contributions, alignment techniques prototyping, and automating training at scale. Competitive salary range is $175,000 – $350,000 per year based on experience and location, along with equity and benefits including unlimited paid time off. #J-18808-Ljbffr Inflection AI

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Model Training Engineer: RLHF & LLM Fine-Tuning in Palo Alto, CA vacancy
  • $238k - $302k

     ...from a diverse set of sensors, enabling engineers like you to (1) develop methods for...  ...real-world data, to (2) develop models and model training at scale, to (3) analyze real-world behavior...  .... You will: Lead the multi-task fine-tuning of multi-billion parameter models, and... 
    Training
    Full time

    Neura Market

    Mountain View, CA
    2 days ago
  • $175k - $350k

     ...with human-centered AI models that unite emotional intelligence...  ...the Role As a Model Training engineer, you will design, build,...  ...that turn a general LLM into a brand-fluent, production...  .... Your innovations in fine-tuning and preference optimization (RLHF, DPO, GRPO, RLAIF) will... 
    Training

    Inflection AI

    Palo Alto, CA
    3 days ago
  • $195.2k - $262.2k

     ...and enterprises from data and model training through to production...  ...ML infrastructure. Built by engineers, for engineers. From large-scale...  ...research programs in efficient LLM and VLM inference with measurable...  ...post-training, SFT, DPO, RLHF, RLAIF, preference optimization... 
    Training
    Temporary work
    Immediate start
    Remote work

    Nebius

    Palo Alto, CA
    17 hours ago
  • Waymo, a leading autonomous driving technology company, seeks a senior ML/CV leader to guide multi-billion parameter model fine-tuning and production release. You will architect distillation of expert models, manage deployments, and build high-performing teams while shaping... 
    Suggested

    Neura Market

    Mountain View, CA
    2 days ago
  • $272k - $431.25k

     ...is seeking a Principal Engineer to drive the performance of large-scale AI training and post-training workloads...  ...frontier-scale LLM workloads running on thousands...  ...software, tools, models, benchmarks, and analysis...  ...reinforcement learning, fine-tuning, or other post-training... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $150k - $230k

     ...hands-on Machine Learning Engineer to drive the post-training of our large language models, with a strong emphasis...  ...(CPT), supervised fine-tuning (SFT), and RL — along with...  ...primary focus (e.g., RLHF, PPO, GRPO, DPO, and related...  ....RequirementsHands-on LLM post-training... 
    Training
    Full time
    Local area
    Work from home

    News Break

    Mountain View, CA
    2 days ago
  • Dormont Manufacturing Co is seeking experts in pre-training and fine-tuning large language models. The role includes modifying models for human interaction...  ...a strong understanding of generative AI, technical engineering skills for production software, and collaborative... 
    Training

    Dormont Manufacturing Company

    Palo Alto, CA
    4 days ago
  • $174.72k - $295.68k

     ...time Machine Learning Engineer / Research Scientist to drive the modeling and algorithmic...  ...experts to design, train, and deploy large-scale...  ...pretraining and fine-tuning strategies leveraging...  ...driving models, or LLM/VLM architectures (...  ...).Familiarity with RLHF/DPO/GRPO,... 
    Training
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

     ...now looking for a Senior High-Performance LLM Training Engineer!NVIDIA is seeking experienced engineers...  ...architecture.Proven experience analyzing and tuning application performance & processor and system-level performance modelling.Programming skills in C++, Python, and... 
    Training
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    4 days ago
  • $185k - $400k

     ...Research Scientists in Foundation Models with expertise in pre-training and mid-training large-scale multimodal...  ...You will collaborate closely with engineering and product teams, shaping the...  ...for large-scale pre-training and fine-tuning.Identify, create, and leverage large... 
    Training
    Remote work

    Pika

    Palo Alto, CA
    1 day ago
  •  ...autonomous, clinical conversations with patients. We have trained our own LLMs as part of our Polaris constellation,...  .... About the Role We're seeking an experienced LLM Inference Engineer to optimize our large language model (LLM) serving infrastructure. The ideal candidate... 
    Training

    Hippocratic AI

    Palo Alto, CA
    4 days ago
  •  ...practical deployments. You will stabilize large language model training, enable pipeline parallelism, and fine-tune models using truthful data. You will design and...  ...techniques, collaborate with AI researchers and engineers, and contribute to projects spanning machine... 
    Training
    Remote work

    Neura Market

    Palo Alto, CA
    2 days ago
  • $224k - $356.5k

     ...searching for a senior or principal engineer who specializes in building...  ...for large-scale foundation model training in the Generalist Embodied...  ...model training and fine-tuning on massive datasets.Implement...  ...experience at building large-scale LLM and multimodal LLM training... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $165.2k - $223.6k

     ...unparalleled ML inference and training performance.The...  ...a wide range of models and supporting novel architecture...  ...boundary, our engineers build systematic infrastructure...  ...every compute unit is fine tuned for optimal...  ...of a wide variety of LLM model families, including... 
    Training
    Work experience placement
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $170k - $260k

     ...for the Deep Learning Engineer role at GenBio AI Base...  ...Large Biological Models (LBM), we are pioneering...  ...biomedicine, with our LBM training leading to ground‑...  ...team and leadership in LLM and generative AI position...  ..., pre‑training, fine‑tuning, serving) Build and maintain... 
    Training
    Full time
    Work at office

    GenBio AI

    Palo Alto, CA
    2 days ago
  • Job You will own the training pipeline behind the models that power both Parallel’s search stack and Parallel’s agents. On the search side, that means...  ...from real product usage to high‑quality training data, fine‑tune and evaluate these models rigorously, and ship them... 
    Training
    Work at office
    Visa sponsorship

    Parallel Web Systems

    Palo Alto, CA
    5 days ago
  • $150k

     ...highly motivated, and focused on engineering excellence. This...  ...You will join the Grok Voice Model team to help build the world...  ...scenarios. We own the full training pipeline: massive data curation...  ...enhancements through supervised fine-tuning, reinforcement learning, and... 
    Training
    Temporary work

    SpaceXAI

    Palo Alto, CA
    9 days ago
  • $182.5k - $343.2k

     ...: Research and development of large‑scale video world models, including design and construction of training datasets, foundational model algorithm design, optimization related to pre‑training, supervised fine‑tuning, and reinforcement learning, model capability evaluation... 
    Training
    Relocation package

    Lightspeed Studios

    Palo Alto, CA
    3 days ago
  • $224k - $356.5k

     ...performance computing. As a Senior / Principal Deep Learning Engineer — Model Evaluation & AI Systems, you will play a meaningful role in crafting...  ...the roadmap, and sharing best practices.Work alongside model training, inference, and product divisions to provide trusted... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $207k - $301k

     ...journey.Design, develop and tune LLM-based content...  ...Formats. Guide junior engineers and provide technical leadership...  ...infrastructure (e.g., model deployment, model...  ...processing, debugging, fine tuning).3 years of experience...  ...relevant education or training. US: $207000 - $301000... 
    Training

    Google

    Mountain View, CA
    5 days ago
  •  ...Systems in Palo Alto is seeking a professional to own the training pipeline behind models that power both their search stack and agents. The role...  ...connections from product usage to training data, fine-tuning models, and ensuring safe deployment. Ideal candidates have... 
    Training

    Parallel Web Systems

    Palo Alto, CA
    5 days ago
  • $184k - $287.5k

     ...platform upon which every new AI-powered application is built. We are seeking a senior vision language model engineer to design and build agentic data and training workflows for Autonomous Vehicles, Robotics, and Medical applications. The right person for this role brings... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

     ...NVIDIA, we are seeking exceptional engineers to join our autonomous driving team...  ...together!What You’ll Be Doing:Design and train innovative large-scale models—including generative, imitation,...  ...systems.Build, pre-train, and fine-tune LLM/VLM/VLA systems for deployment in real... 
    Training
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...experienced software engineer with machine learning...  ...and Graph of Thoughts), fine tunings of LLMs for tool use and...  ...approaches such as RLHF, RLAIF, and DPO), agent...  ...infrastructure to ensure model performance across...  ...meticulous standards for training and evaluation... 
    Training
    Work at office
    Immediate start
    Remote work
    Flexible hours

    ServiceNow

    Mountain View, CA
    4 days ago
  • $165k - $185k

     ...Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable...  ...Cloud Robotics, Data Science, AI System Engineering, Time-series Analysis. We develop...  ...experience on foundation models, including training, fine-tuning, and promptingIn-depth experiences in... 
    Training
    Work experience placement
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    1 day ago
  • $192k - $278k

    Lead model releases for Search, evaluating DeepMind release applicants...  ...with Generative AI or LLM model releases. Preferred qualifications...  ..., and key enablers of engineering velocity.In this role, you...  ...experience, and relevant education or training. US: $192000 - $278000 (USD)... 
    Training
    Shift work

    Google

    Mountain View, CA
    1 day ago
  • $188.5k - $282.7k

     ...Semantic AI Governance Engine, which is the...  ...custom small language models act as judges on...  ...core, SAGE is "LLM-as-judge" applied...  ...lifecycle: curating data, training small models,...  ...DutiesTraining, Fine-Tuning, and Distilling Production...  ...(DPO, RLAIF, or RLHF).Production... 
    Training
    Permanent employment
    Local area

    Rubrik

    Palo Alto, CA
    1 day ago
  • $400k

     ...Principal Research Engineer, Model Training & Post-TrainingPalo Alto, California, United StatesAbout...  ...strategy, including supervised fine-tuning, RLHF, DPO, GRPO, RLAIF, reward modeling, preference...  ...contributor to, large-scale LLM, multimodal, or foundation-model training... 
    Training
    Work at office

    Inflection AI

    Palo Alto, CA
    3 days ago
  •  ...analyze, profile, and optimize AI training workloads on innovative...  ...Computer Science, Electrical Engineering or Computer Engineering and 5...  ...Proven experience analyzing and tuning application performance &...  ...and system‑level performance modelling. Programming skills in C++, Python... 
    Training
    Work experience placement

    NVIDIA Gruppe

    Santa Clara, CA
    1 day ago
  • $159k - $277k

     ...experts in energy, AI, software, engineering, and product to build tools...  ...Tapestry’s end-to-end grid model strategy. You will lead efforts...  ...we can individually.Always fine-tune: We stay curious, seek feedback...  ..., and relevant education or training. Your recruiter can share... 
    Training
    Flexible hours

    X Company

    Mountain View, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Model Training Engineer: RLHF & LLM Fine-Tuning. Be the first to apply!