Research Engineer - Post-Training [Remote]
Pluralis Research
- Remote job
Pluralis Research works on Protocol Learning: training and serving large models in a fully decentralized way on small consumer-grade devices connected via the internet. Despite being dismissed as infeasible, we have made significant advances on this problem, most recently Agora, a permissionless run that pretrained an 8B model from scratch on consumer GPUs spread over the internet, with no single participant ever holding the full weights ([tech report]( While many of the core research problems have been solved, Protocol Learning unlocks a series of new challenges. For the mission in full, read [A Third Path: Protocol Learning](
Agora gave us a pretrained 8B model. Post-training is how we make it useful for agentic use-cases. But every post-training stack you've seen assumes a datacenter — synchronous rollouts, fast interconnects, trusted workers. Ours gets none of that. It has to run on consumer GPUs, and Macs spread across the public internet, training a model whose weights no single participant ever holds, with rollouts arriving from a geo-distributed inference pipeline at high latencies. Your primary role is to make RL post-training work here anyway — the algorithms and the system, end-to-end.
Key Responsibilities
- Build the post-training stack : You build the RL training loop end-to-end: rollout ingestion from the geo-distributed inference pipeline, reward computation, policy updates, and getting updated weights back out to the network. You set the direction, and you make things happen.
- Invent the algorithms : Standard RL recipes assume on-policy rollouts from fast, trusted hardware. You adapt them to asynchronous, high-latency, partially trusted generation: staleness tolerance, off-policy corrections, and communication-efficient policy updates.
- Ship first post-trained models : You build the evals that show the models are improving, and you take the first decentralized post-trained release from run to public artifact.
What We're Looking For
- Hands-on RL post-training : You've run RL post-training on large language models — RLHF, RLVR, or reasoning-focused RL — and touched the systems layer yourself: rollout generation, async training loops, weight synchronization. Not just launched jobs on someone else's stack.
- Strong engineering : Production-quality Python and PyTorch: concurrency, failure handling, profiling before optimizing.
- Research ability : Publications in RL post-training, asynchronous or distributed RL, or nearby fields are a strong signal. So is unpublished work you can defend in detail.
- Mission alignment : You believe Protocol Learning is the viable third path for collective, trustless, and sovereign AI.
Nice to Have
- Experience training over slow networks, or with decentralized or federated setups.
- Familiarity with serving-engine internals such as vLLM or SGLang — our rollout pipeline is a serving system.
- Experience with reward modeling or building verifiable-reward datasets.
- Experience with P2P networking and NAT traversal.
- Experience at proprietary, open-weight and open-source AI labs
Compensation & Benefits
- Equity-Heavy Package : We offer significant ownership for key technical contributors in addition to a high base salary.
- Remote-First Culture : Flexible work environment with team members distributed globally.
- Visa Sponsorship : Optional full visa sponsorship and relocation support to either Australia or the US.
- Open Problems : Training and serving frontier models on hardware you don't control, over networks you don't own, mostly has no published answers yet. You'll write some of the first ones.
FYI's
- We work remotely across the world, with the main teams in Australia and North America. You'll need to be comfortable working across timezones.
- Applicants must have professional-level English proficiency (written and spoken).
- Recruiters: we aren't looking for agency support at this time. We'll reach out if we need help.
- ...Job Description Job Description **Senior Research Engineer, Post-Training** **About CHAI AI** CHAI is a social AI platform where millions of people create, share, and talk to AI characters. As of Sep 2026, it has 10 million active users and $100 million ARR....TrainingFull timeWork at office
- ...real-world environments. Job Overview We are seeking a Research Engineer, Robotics & Embodied AI to bridge cutting-edge AI research... ...research from leading robotics and AI conferences. Build training and evaluation pipelines for robot learning models....TrainingFull time
$140k
...driven applications.Build data pipelines and training frameworks to support experimentation... ...cross-functional teams to integrate research outcomes into production systems.Analyze... ...protected] with the position title “Research Engineer” in the email subject.CompensationBase...TrainingWork at office$147k - $210k
...and energy systems modeling and optimizationAt Google, research-focused Software Engineers are embedded throughout the company, allowing them to setup... ...-related skills, experience, and relevant education or training. US: $147000 - $210000 (USD) + 15% bonus target + equity...Training$150k - $250k
...goods, and global social organizations.We research and deploy technologies that power AI-... ...We Are Looking ForAt Distyl, Research Engineers build the bridge between frontier AI... ...production.Key ResponsibilitiesDesign and run post-training workflows that improve the behavior,...TrainingWork at office3 days per week$147k - $210k
...ability to perform long-horizon software engineering tasks, and launch entirely new projects... ...within DeepMind aimed at improving Gemini’s training data and evaluations (evals).Minimum... ...tools, and technical stacks.At Google, research-focused Software Engineers are embedded...Training$207k - $300k
Design, prototype, scale engineering solutions, and run experiments to address emerging research priorities within the Responsibility portfolio.Work with Research Scientists... ...skills, experience, and relevant education or training. US: $207000 - $300000 (USD) + 20% bonus target...Training$192k - $304.75k
...technology—and outstanding people! We are seeking a world-class engineer to drive applied research at the intersection of AI and ASIC design. Large... ....Hands-on experience with LLMs, RL, RLHF/RLAIF, post-training, evaluation, graders, synthetic data, model training, coding...TrainingFull time$224k - $356.5k
...searching for a senior or principal engineer who specializes in physics... ...Generalist Embodied Agent Research (GEAR) group. Our team is... ...performance for large-scale training workloads.Import, configure,... ...until January 13, 2026.This posting is for an existing vacancy. NVIDIA...TrainingFull time$50 per hour
...and outside of work. This is a place for engineers, scientists, and problem-solvers who are... ...is seeking a full-time Plasma Physics Research Engineer. In this role, you will develop... ...candidate's work experience, education/ training, key skills as well as market(work location...TrainingFull timeTemporary workWork experience placementCasual workFlexible hours$164.6k - $313.3k
...looking for a driven Data/ML engineer to push the boundaries of audio... ...collaborative and efficient research team looking for highly... ...hiring a data lead to own the training data behind our generative audio... ...Colorado (as listed on the job posting), the application window will...TrainingFull timeTemporary workLocal areaWorldwide$224k - $356.5k
...Tools organization is seeking a Senior Research Engineer to join our Research team, where we build... ..., curate, and validate synthetic training and evaluation data for CUDA programmingDeliver... ..., influential papers, or blog posts with demonstrated real-world impactExperience...TrainingFull timeShift work$100.32k - $144.21k
Valencia, CA (Hybrid) Senior Engineer, Research Advanced Bionics is seeking a Senior Engineer, Research to support the design and development... ...mentor to lower-level engineers by providing guidance, training, and knowledge transferMore about you: • Bachelor’s degree in...TrainingTemporary workWork experience placementRemote workFlexible hours$174k - $252k
Scope and drive research efforts to improve complex frontier Gemini capabilities, such... ...practical experience.5 years of software engineering experience using Python.1 year of experience working on large language model post-training (e.g., SFT, RLHF, DPO, PPO), model...Training$174k - $255k
...a hands-on Machine Learning engineer, contributing to all aspects... ...world products, and is not a research-based role.Our team is small... ...experience, and relevant education or training. Your recruiter can share... ...details listed in US role postings reflect the base salary only,...TrainingFull time$197.3k - $313.7k
...talented software and platform engineers to embed in our AI team to... ...directly enable world-class research and products used by millions... ...AI tools and opt out options.Posting StatementSalesforce is an equal... ...compensation, promotion, benefits, training, assessment of job...TrainingFull time$193.3k - $261.5k
...key member of this team, you will lead research and development efforts in generative AI... ...innovation, and deep collaboration with engineering teams to bring research into production.... ...fundamentals, including transformer architecture, training/inference lifecycles, and optimization...TrainingInternshipLocal areaFlexible hours$184k - $287.5k
We are recruiting top research engineers in the Autonomous Vehicles Research team at NVIDIA with... ...programming skills, a solid track record of training deep learning models at scale, and a... ...at least until August 9, 2026.This posting is for an existing vacancy. NVIDIA uses...TrainingFull time$150k - $200k
Santa Clara, CASoftware Engineering - Controls /Full-time /HybridPlusAI is a Physical AI company... ...to join its fast-growing teams.As a Research Engineer, you will deliver mission-critical... ...infrastructure for dataset generation, training, and evaluation to drive advancements in...TrainingFull time- ...faster deliveries.About the role:We're currently looking for research engineers with specialized skills in LiDAR, camera, and radar... ...initiatives to collect, augment, and utilize large-scale datasets for training and validating perception models under various driving...Training
- ...drive this mission, we are looking for a Research Engineer to tackle the industry's toughest... ...The salary range displayed on this job posting reflects a minimum and maximum target.... ...experience, and relevant education or training. Please note that this range reflects the...TrainingContract workWork from homeHome officeFlexible hours
$141.4k - $204.4k
...rewarding.We are looking for a Senior Research Engineer to join our Worlds research pillar and... ...engineering work around integration, scaling, training, development, and support.Build... ...in these locations at the time of this posting. If you reside in a different location,...TrainingFull timeWork at officeLocal area3 days per week$120k - $243k
HPE Labs - Research EngineerThis role has been designed as ‘’Onsite’ with an expectation... ...Computer Science or Electrical & Computer Engineering. Excellent networked systems building... ...geographic location, work experience, education/training, and/or skill level. - United States of...TrainingFull timeWork experience placementWork at officeLocal areaImmediate start$180k - $258.75k
...Development /Full-time /HybridAt Toyota Research Institute (TRI), we’re on a mission to improve... ....Our team is looking for a Research Engineer to help develop and deploy our world... ...and optimizing large-scale distributed training of diffusion and transformer models; maintaining...TrainingFull timeLocal areaShift work- ...of previous Stanford professors, SAIL researchers, Olympiad medalists (IPhO, IOI, etc.),... ...Your work will enable large‑scale model training, inference, and reinforcement learning... .... Working closely with researchers and engineers, you’ll help make Voltai the world’s leading...Training
- ...Requirements Strong general software engineering skills Thorough knowledge of the deep learning literature Experience with pre- and post-training of LLMs Ability to come up with and evaluate research ideas Experience working with large distributed systems...Training
- ...Pantograph is training general models that start by watching internet-scale video and end up on robots. We think the path to capable... ...fleet of affordable, durable robots. We're looking for a research engineer to help us train increasingly capable models across enormous...Training
- ...The Role As a Research Engineer, you'll build the systems that let us post-train models continuously: the training pipelines, environment infrastructure, and evaluation harnesses that turn research into something that runs every day. You'll sit between research and...Training
- ...of care. About the Role As a Research Engineer, you’ll be responsible for building industry... ...iterative milestones and roadmaps Train, fine-tune, validate, and further... ...engineering or research ~ Prior experience post-training and deploying LLMs in...TrainingWork at officeFlexible hours
- ...Proximal is building the research systems needed to identify what models can’t yet do, build... .... About the role As a research engineer, you will work on open-ended research problems... ...correctness Modify our RL training stack to support end-to-end training with...Training
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Engineer - Post-Training [Remote]. Be the first to apply!
- deep learning research engineer California
- research engineer California
- research programmer California
- remote social worker California
- remote nurse practitioner California
- remote coding manager California
- customer service rep remote California
- part-time virtual/remote assistant California
- senior network engineer remote California
- frontend internship remote California



