Remote Senior SWE: AI Training Data & RL Environments
YO AI Labs
- Remote job
YO AI Labs is seeking experienced Senior Software Engineers to support an AI training project by creating reinforcement learning environments that evaluate AI models on complex software tasks using MCP tools. You will design reproducible environments, deterministic verification, and reference solutions for tasks such as bug fixing, feature implementation, codebase refactoring, and performance optimization. No prior AI experience is required. Contractor role (~15 hours/week) with remote work. #J-18808-Ljbffr YO AI Labs
- YO AI Labs is seeking experienced Senior Software Engineers to support an AI training project by creating reinforcement learning environments that evaluate AI models on complex software tasks using MCP tools. You will design reproducible environments, deterministic verification...Remote jobDataSeniorTraining
- YO AI Labs is seeking experienced Senior Software Engineers to support an AI training project by creating reinforcement learning environments that evaluate AI models on complex software engineering tasks using Model Context Protocol (MCP) tools. You will design reproducible...Remote jobDataSeniorTraining
- YO AI Labs is seeking a Senior Software Engineer for a remote contractor role (~15 hours/week). You will build RL environments to evaluate AI models on software tasks using MCP tools. The work focuses on reproducible environments, deterministic verification, and reference...Remote jobSeniorTrainingFor contractors
$245k - $300k
...virtuous cycle: human insight improves AI, and better AI expands what people can... ...that turns that judgment into the data, evals, and RL environments frontier models learn from. We work with... ...own the RL environments frontier labs train on, end to end. Scope the problem with...DataSeniorTraining$184k - $287.5k
Reinforcement learning post-training is driving some of the... ...capability gains in AI today. It is the... ...challenges in the field. RL requires inference, rollout... ...model must interact with environments, tools, and other... ...design (service boundaries, data flows, consistency...DataSeniorTrainingFull time- YO AI Labs is seeking a Senior Software Engineer to support an AI training project by creating reinforcement learning environments that evaluate AI models on complex software tasks using MCP tools. You will design reproducible environments, deterministic verification, and...Remote jobSeniorTraining
- YO AI Labs is seeking a Senior Software Engineer to support an AI training project by creating reinforcement learning environments that evaluate AI models on complex software engineering tasks using Model Context Protocol (MCP) tools. You will design reproducible environments...Remote jobSeniorTrainingContract work
- ...United States is seeking a Machine Learning Engineer / Data Scientist to join our team, working on agent harness research... ...and implements algorithms for agent harness and post-training pipelines, develops RL environments and reward models, and conducts training runs to...DataSeniorTraining
- YO AI Labs in the United States is seeking a Senior Software Engineer to support an AI training project by creating reinforcement learning environments that evaluate AI models on complex software tasks using... ...performance optimization. Remote work available. #J-18808-Ljbffr...Remote jobSeniorTraining
- YO AI Labs seeks a Senior Software Engineer (Contractor, ~15 hours/week) to build reinforcement learning environments for software tasks using MCP tools. You will design reproducible scenarios... .... Responsibilities include creating RL environments, building verification systems...Remote jobSeniorFor contractors
- YO AI Labs is seeking a Senior Software Engineer on a contracting basis (≈15 hours/week) to support an AI training project by building reinforcement learning environments for software engineering tasks using MCP tools. You will design reproducible environments, deterministic...Remote jobSeniorTrainingContract work
- YO AI Labs is seeking a Senior Software Engineer on a remote, contractor basis (~15 hours/week) to support an AI training project. You will create reinforcement learning environments that evaluate AI models on complex software engineering tasks using MCP tools. Responsibilities...Remote jobSeniorTrainingFor contractors
- YO AI Labs is seeking a Senior Software Engineer contractor to join an AI training project. You will create reinforcement learning environments that evaluate AI models on software engineering tasks using... ...across large codebases. Remote work with ~15 hours per week and...Remote jobSeniorTrainingFor contractors
- Astera seeks a Senior/Staff AI Research Scientist for Polytope Bio residency... ..., experimental design, and training strategy at a high-leverage stage... ...with substantial compute and data resources. You will publish... ...founder roles in the future. Remote work is possible for the...Remote jobDataSeniorTraining
- YO AI Labs is seeking experienced Senior Software Engineers to support an AI training project by creating reinforcement learning environments using MCP tools. You will design reproducible environments,... ...experience required. The position is remote as a contractor (~15 hours/week...Remote jobSeniorTrainingFor contractors
- YO AI Labs is seeking a Senior Software Engineer to support an AI training project by building reinforcement learning environments that evaluate models on complex software tasks using MCP tools.... ...refactoring, and performance optimization. Remote work from the United States is...Remote jobSeniorTrainingFor contractors
$272k - $431.25k
...Reinforcement learning post-training is where modern AI systems learn to... ...in AI: a single RL run ties together inference... ...We are looking for a Senior Software Engineering... ...an inclusive work environment and proud to be an equal... ...Santa Clara; US, DC, Remote; US, NY, Remote; US,...Remote workSeniorTrainingFull time- ...Job Description Location: Remote (United States) Work Model... ...Remote Industry: Applied AI / AI research data Compensation: $180K-$220... ...reinforcement-learning environments and agents sold to the world... .... The Opportunity As an RL Environment Software Engineer...Remote workData
- YO AI Labs is seeking a Senior Software Engineer for a remote, contractor role (~15 hours per week) to support an AI training project. You will build reinforcement learning environments that evaluate AI models on complex software engineering tasks using MCP tools. You...Remote jobSeniorTrainingFor contractors
- Traverse is a research data lab building reinforcement learning environments for frontier AI labs. As a Research Scientist, you will design and build RL environments that teach models to do work... ...domain knowledge, turning that into training signals that actually work. We...DataTraining
$155k - $180k
...company delivering cloud, AI, data, and enterprise solutions... ...Machine Learning Engineer – RL Location: 100% Remote (U.S.) Position Type:... ...Engineer - RL to design, train, and deploy RL-based systems... ...learning algorithms, simulation environments, reward modeling, and the...Remote workDataTrainingFull timeH1bLocal areaImmediate startVisa sponsorship- ...States | Posted on 06/15/2026 AI Talent Now, LLC is a... ...infrastructure that powers frontier data creation for agentic and hard... ...and quant firms. We build the training data and evaluation... ...models learn and improve. As an RL Environment Engineer you will design datasets...DataTrainingH1bWork at officeVisa sponsorshipFlexible hoursNight shift
$281k - $356k
...machine learning and data systems,... ...models to deliver training and evaluation... ...our embodied AI applications... ...iteration of novel RL algorithms, reward... ...simulation environments ~ Deep understanding... ..., influencing senior stakeholders,... ...be performed remote, the specific...Remote workDataSeniorTrainingFull time- Pareto in San Francisco builds production RL environments and data platforms for frontier labs. You will own environments end to end, from build... ...shipping fast, owning the release path, and turning training signals into product improvements. Equity is part of the package...DataSeniorTraining
$213k - $263k
...foundation for training and validating the... ...and generative AI to automatically... ..., and power the data engine that... ...will report to a Senior Staff Technical... ...Reinforcement Learning (RL) techniques to... ...fast-paced R&D environment. We... ...can be performed remote, the specific salary...Remote workDataSeniorTrainingFull time$184k - $287.5k
...highly motivated Senior DevTech Compute Engineer... ...Compression and Data Processing! Would... ...the middle of the RL training run defines the... ...what can be done in AI.#LI-HybridYour... ...an inclusive work environment and proud to be an... ...Santa Clara; US, NY, Remote; US, CA, Remote; US...Remote workDataSeniorTrainingFull time- About Bespoke Labs Bespoke Labs is an applied AI research lab pioneering data and RL environment curation for training and evaluating agents. Recently, we curated Open... ...keep making the next unit cheaper to produce. SWE and RL Environment skills Strong software engineering...DataTrainingFlexible hours
- ...Digital Space LLC is seeking an AI Product Manager to own the... ...vertical within our Agents Data & Reinforcement Learning Environments team. You will lead the development of RL environments and the data-as... ...workflows into high-quality training data and simulations for AI...DataTraining
- ...design, build and operate data centerswe are enabling... ...of humanity.The Role: Senior Regional Environmental... ...conduct environmental training and technical guidance... ...priorities in a fast-paced environment.Proficiency with... ...development.Flexibility & Remote Opportunities Whether in...Remote workDataSeniorTrainingFor contractorsFor subcontractorWork at officeLocal area
- ...Senior Data Scientist Actifai is seeking a Senior... ...for a hybrid or remote position. Actifai... ...by Foundry.ai, a technology fund... ...in a fast-moving environment. Key Responsibilities... ...Create and maintain training datasets from... ...contextual bandits, RL, or policy optimization...Remote workDataSeniorTrainingWork at officeImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Remote Senior SWE: AI Training Data & RL Environments. Be the first to apply!
- software engineering manager remote Atlanta, GA
- system engineer remote Atlanta, GA
- remote technical project manager Atlanta, GA
- remote recruitment consultant Atlanta, GA
- remote property manager Atlanta, GA
- remote reviewer Atlanta, GA
- remote insurance agent Atlanta, GA
- remote no experience Atlanta, GA
- remote life underwriter Atlanta, GA
- remote medical coding supervisor Atlanta, GA


