Research Engineer, Post-Training Data
NxT Level
Job Description
Job Description
Research Engineer, Post-Training Data
Location: San Fran
Employment Type: Full-time
Focus: Post-Training, Reinforcement Learning, Data Generation, Research Environments, Frontier AI
Profile: Researcher-engineer with strong ML and software engineering depth
About Our Client
Our client is building the next generation of post-training data infrastructure for frontier AI labs.
The company's core belief is that research and data production are inseparable. The best post-training data is not created through brute-force labeling or headcount alone — it requires domain expertise, research judgment, ML fluency, and the ability to build systems that scale data generation superlinearly.
Inbound demand from frontier AI customers is growing faster than the current team can support, and our client is hiring early technical talent to help meet that demand.
This is an opportunity to join a team building high-quality post-training data, reinforcement learning environments, long-horizon tasks, and domain-specific evaluation workflows for some of the most advanced AI systems in the world.
About the Role
Our client is hiring a Research Engineer, Post-Training Data to own the full lifecycle of post-training model and data work.
This role blends AI research, ML engineering, software engineering, and data production into one function. The ideal candidate is not just a researcher and not just an engineer — they are someone who can understand a domain deeply, identify what makes a task realistic and economically valuable, build the environment or data-generation system, and improve the model through high-quality post-training data.
The company is indexing heavily on research and ML horsepower, strong software ability, high slope, and data taste.
What You'll Do
Own post-training workflows end to end, from infrastructure provisioning through data generation and curation
Build and curate high-quality post-training data from model and environment generation processes
Produce reinforcement learning environments and long-horizon tasks across complex domains
Work on research-loop, science, chip, physics, and other technically deep task environments
Build systems that scale data generation superlinearly through self-improving developer processes
Improve data generation through better tools, workflows, automation, evaluation, and model feedback loops
Exercise and develop strong data taste around what problems matter in a given domain
Identify realistic, economically valuable tasks and environments that frontier AI labs will actually want
Blend research and data production into a single technical function
Debug ML pipelines, understand unfamiliar code quickly, and solve open-ended technical problems
Help define the standards for what high-quality post-training data should look like across domains
What We're Looking For
Genuine spike in ML, AI research, or a technical research domain
Experience or strong interest in post-training, reinforcement learning, evaluations, interpretability, or related areas
Strong software engineering ability and comfort building production-quality tools or research systems
Ability to post-train models and build or curate high-quality post-training data
Strong judgment around data quality, task design, and what makes a problem valuable to AI labs
Fast problem-solving ability and strong code comprehension
Ability to work across unfamiliar domains and ramp quickly
High-slope learning profile with strong technical curiosity
Comfort operating in ambiguous, research-heavy environments
Evidence of deep commitment, strong output, and public artifacts or meaningful technical work
Ideal Background
Our client is especially interested in candidates who combine:
Domain or research expertise in a field such as neuroscience, physics, chemistry, systems, chip design, science, or another technical discipline
Strong ML or software engineering core
Experience with PyTorch, ML pipelines, model debugging, or research tooling
Experience building environments, evaluations, or data-generation systems
Exposure to post-training, RL, long-horizon tasks, or frontier model workflows
Pedigree can be a helpful signal, but it is not the primary filter. The team cares more about slope, research horsepower, problem-solving speed, and the ability to build.
Bonus Experience
Experience at frontier AI labs, AI infrastructure companies, or high-talent technical teams
Experience with post-training data, RL environments, evals, model behavior, or interpretability
Experience building data products, research tooling, or developer workflows for AI teams
Experience in domains like chip design, physics, chemistry, biology, systems, or scientific computing
Public artifacts, research projects, open-source work, technical writing, or demos that show exceptional ability
Experience creating tasks or environments that require long-horizon reasoning
Experience scaling data generation through automation rather than manual labor
Who Will Thrive Here
A researcher-engineer who treats data as a research problem
A high-slope polymath with a real technical spike
Someone who can figure out unfamiliar domains quickly
A builder who understands that the best data comes from strong research judgment
Someone obsessed with creating realistic, economically valuable AI training and evaluation environments
A deeply committed, results-driven technical operator
Someone excited by frontier lab customers and the thesis that better systems can scale data generation superlinearly
Why This Opportunity
Join an early team serving fast-growing demand from frontier AI customers
Work on the core bottleneck behind better post-training outcomes: high-quality data
Build RL environments and long-horizon task systems across technically deep domains
Shape how research and data production come together as one function
Work on systems designed to scale data generation through process, tooling, and model feedback — not simply more people
Build at the frontier of post-training, evaluations, RL environments, and data quality
Step into a role where research taste, engineering speed, and domain curiosity all matter
Ideal Candidate Profile
The ideal candidate is a research-minded engineer with a real spike in ML, AI research, or a technical domain.
They can read and understand code quickly, debug ML pipelines, build tools, reason about task quality, and identify what makes a data environment valuable to a frontier AI lab. They are not looking for a narrow research role or a pure software role — they want to build the systems and data that make models better.
This person is high-slope, obsessive, technically broad, and excited to help define what great post-training data looks like.
$150k - $250k
...global social organizations.We research and deploy technologies that... ...ForAt Distyl, Research Engineers build the bridge between frontier... ...ResponsibilitiesDesign and run post-training workflows that improve the... ..., reward modeling, synthetic data, evals, or related post-training...DataTrainingWork at office3 days per week- ...Workspace.Sierra’s Applied Research team advances the quality, speed... ...of tasks. Our work spans post-training open-weight models for specialized... ...working deeply with data, experimentation, evaluation... ...TypeFull timeLocation TypeOn-siteDepartmentEngineeringPlatform EngineeringDataTrainingFull timeFlexible hours
$197.3k - $313.7k
...software and platform engineers to embed in our AI team... ...enable world-class research and products used by millions... ...we use your personal data and your rights,... ...tools and opt out options.Posting StatementSalesforce is... ...promotion, benefits, training, assessment of job...DataTrainingFull time$193.3k - $261.5k
...structured and unstructured data across enterprise systems. As... ...of this team, you will lead research and development efforts in generative... ...and deep collaboration with engineering teams to bring research into... ...transformer architecture, training/inference lifecycles, and...DataTrainingInternshipLocal areaFlexible hours- ...Requirements Strong general software engineering skills Thorough knowledge... ...Experience with pre- and post-training of LLMs Ability to come up with and evaluate research ideas Experience working... ...capabilities Build out internet-scale data pipelines and crawlers...DataTraining
- ...megaprojects: power plants, factories, data centers, and the physical... .... What you'll do The research that matters most to us comes... ...end to end, building agents, training models, and designing the benchmarks... ...about current PhDs and post-docs in continual learning, test...DataTraining
- ...The Role As a Research Engineer, you'll build the systems that let us post-train models continuously: the training pipelines, environment infrastructure, and evaluation... ...building the machinery for trace collection, data curation, environment generation, and the RL...DataTraining
- ...Pantograph is training general models that start by watching internet-scale video and... ...durable robots. We're looking for a research engineer to help us train increasingly capable models... ...learning, reinforcement learning, data processing, evaluation, and the infrastructure...DataTraining
- ...team includes ex-founders and engineers who have built and scaled... ...Overview We are hiring a Founding Research Engineer to design, prototype... .... You will turn messy health data into accurate, cited, and... ...engineering: dataset curation, model training and evaluation, retrieval and...DataTraining
$175k - $250k
...Research Engineer About Scorecard We’re a small, nimble team backed by top-tier investors... ...everything you build; your simusers run inside training environments where their failures... ...evaluations with baselines, held-out data, and calibrated judges, and you notice...DataTrainingWork at office- ...real-world industry and economy use cases. As a Research Engineer on our Physical AI team, you will lead pre-training and post-training on action-conditioned world models,... ..., FSDP, and DeepSpeed Work with multimodal data pipelines involving video, sensory inputs, and...DataTrainingWork at office
- ...critical software vulnerabilities. We are training and scaling security AI agents to... ...founding team includes deep expertise in data, infrastructure, security and LLMs with... ...role We’re seeking an experienced Research Engineer to join our effort in building and training...DataTrainingFull timeWork at office
- ...Industry: AI infrastructure / Reinforcement Learning (RL) training data & evaluations Compensation: Competitive (range not provided... .... The Opportunity Our partner is hiring a Research Engineer to help scale the quality assurance (QA) systems behind training...DataTrainingRemote work
- ...of care. About the Role As a Research Engineer, you’ll be responsible for building... ...milestones and roadmaps Train, fine-tune, validate, and... ...or research Prior experience post-training and deploying LLMs in... ...training, evaluation, inference, data pipelines) Track record of...DataTrainingWork at officeFlexible hours
- ...AfterQuery is an applied research lab curating data solutions for foundation model... ...opportunity to shape the engineering organization and lead major... ...Engineer, you will design and run training experiments that isolate... ...SFT and RL - base post-training experiments, you will...DataTrainingLocal area
$180k - $250k
...Open role Research Engineer San Francisco (On-site), Full-time About Us Constellation... ...a Research Engineer to sit between our data and our models and make the whole loop... ...faster. You will orchestrate and optimize training runs on long-horizon multimodal...DataTrainingFull timeWork at officeRelocation package- ...Agentic AI that empowers software engineers by automating production... ...workflows end‑to‑end, balancing research and engineering to create... ...AI models Build and optimize data pipelines to process high‑volume... ...and unstructured data for training and evaluation Design and execute...DataTrainingFull timeWork at officeVisa sponsorshipFlexible hours
- ...Mecka AI Mecka AI is building the data infrastructure layer for robotics and embodied... ...AI labs and robotics companies to train and validate humanoid and embodied AI systems... ...deployed hardware. The Role The Research Engineer, RL Env role at Mecka AI involves...DataTraining
$140k - $250k
...Work Model: Remote Industry: AI training data infrastructure Compensation: $140K-... ...top hiring priority and a genuinely hard research problem. Because data flows through a decentralized... ...bottleneck to growth. As a Research Engineer, you will build the automated systems...DataTrainingFull timeWork at officeRemote work$250k - $300k
# Research EngineerArtificial intelligence company* LocationSan Francisco... ...models after initial training, from setting up compute to preparing... ...research with hands-on engineering and calls for sound judgment... ...Background at a leading AI lab or a data-focused technology company.##...DataTraining- ...About Mecka AI Mecka AI is building the data infrastructure layer for robotics and... ...AI labs and robotics companies to train and validate humanoid and embodied AI systems... ...Mecka Labs. This is a hands‑on research engineering role for someone exceptional in simulation...DataTraining
$200k - $250k
...the team knows whether each training run or prompt tweak helped, and... ...and tools that speed up research iterations and make leadership... ...really care about), product engineers (tracking how real users behave... ...Evals built in a real setting: a data or eval company, or a...DataTrainingFull timeRelocation package- ...Francisco, California. The Role: As a Research Engineer - Agency and Reasoning , you will be... ...research in reinforcement learning, post-training, and human preference learning, and... ...Interest in grappling in detail with data and spending significant time involved...DataTrainingWork at officeRelocation package
$130k - $225k
...anonymization systems that help make sensitive real-world data safe and useful for AI training. You will develop end-to-end methods to protect... ...information in unexpected fields. Partner with engineering, research, operations, and customers to turn privacy requirements...DataTrainingVisa sponsorship$231k - $340k
...getting started. Role Overview Post-training is how Harvey turns expert feedback and... ...better at legal work. We are looking for a research engineer who can help scale that loop: defining... ...research partners to build better data, environments, graders, and training...DataTraining$211k - $251k
...Experience working deeply with data, experimentation, evaluation,... ...recommendation, or advertising systems; training and deploying language, speech... ...or contributing to applied ML research or benchmarks.... ...machine learning challenges in post-training, voice AI, and long-horizon...DataTrainingFull timeFlexible hours- ...based in San Francisco, California. The Role: As a Research Engineer - Audio & Speech Models , you will be a core contributor... ...models. You will be deeply involved in the entire model training process, from data gathering and processing to designing novel...DataTrainingWork at officeRelocation package
$200k - $350k
...building at the intersection of research, product, and creativity .... ...Machine Learning Research Engineer , you’ll own end-to-end research... ..., running, and analyzing post‑training experiments that help models... ..., iterative experiments from data to evaluation Curiosity for...DataTraining- ...specialized models, proprietary training signals, and evaluation... ...against enterprise data at scale, every day. We see exactly where research meets production and... ...elite and competitive engineering minds. Translate... ...reliability, observability); post-training and model...DataTrainingTemporary workWork at officeRelocation
$350k
...is a quickly growing group of committed researchers, engineers, policy experts, and business leaders... ...redesign how Claude interacts with external data sources. Many of the paradigms for how... ...for how information is organized, and train language models to optimally use those...DataTrainingRemote jobWork at officeVisa sponsorshipFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Engineer, Post-Training Data. Be the first to apply!
- deep learning research engineer San Francisco, CA
- junior machine learning research engineer San Francisco, CA
- research software engineer San Francisco, CA
- research engineer San Francisco, CA
- research programmer San Francisco, CA
- senior research engineer San Francisco, CA
- ai research engineer San Francisco, CA
- data engineer machine learning San Francisco, CA
- aws data engineer San Francisco, CA
- data systems engineer San Francisco, CA



