Reinforcement Learning Engineer
$100k - $150kBright Vision Technologies
Reinforcement Learning Engineer - Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. Job Title: Reinforcement Learning Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000-$150,000 Annually
Experience Required: 6+ years Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. Key Responsibilities
Required Qualifications
Preferred Qualifications
How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at View phone number on click.appcast.io. Learn more about Bright Vision Technologies at Bright Vision Technologies is an Equal Opportunity Employer.
Equal Employment Opportunity (EEO) Statement Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall. BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. Job Title: Reinforcement Learning Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000-$150,000 Annually
Experience Required: 6+ years Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. Key Responsibilities
- Design and implement reinforcement learning solutions for sequential decision-making problems in real and simulated environments.
- Develop, calibrate, and maintain simulation environments suitable for large-scale agent training.
- Implement and evaluate modern RL algorithms including policy gradient, actor-critic, off-policy, and offline RL methods.
- Engineer reward functions and shaping strategies that align agent behavior with desired outcomes and safety constraints.
- Apply offline RL and imitation learning techniques where exploration is costly or unsafe.
- Use RLHF, DPO, and related techniques for fine-tuning large language models when relevant.
- Build scalable training infrastructure for distributed RL, including efficient experience collection and replay systems.
- Optimize training stability and sample efficiency through algorithmic and engineering improvements.
- Design rigorous evaluation protocols, including out-of-distribution and adversarial test cases.
- Implement safety mechanisms such as constraint enforcement, conservative policies, and human-in-the-loop oversight.
- Collaborate with applied scientists and product teams to identify high-value RL use cases.
- Monitor deployed policies and models in production for drift, regression, and unintended behaviors, building the alerting and dashboards that surface issues before they meaningfully affect users.
- Document methodology, design decisions, and operational characteristics for internal stakeholders.
- Stay current with RL research and translate promising techniques into production-ready solutions.
Required Qualifications
- Master's or PhD in Computer Science, Machine Learning, or a related field; or equivalent applied experience.
- Six or more years of combined RL research and engineering experience.
- Strong proficiency in Python and modern deep learning frameworks.
- Hands-on experience with at least one major RL library or in-house RL stack.
- Solid understanding of probability, optimization, and the theoretical foundations of RL.
- Experience designing and tuning reward functions in non-trivial environments.
- Familiarity with simulation environments and large-scale experience collection.
- Experience training neural network policies on GPU clusters.
- Strong written and verbal communication skills.
- Track record of shipping or publishing impactful RL work.
Preferred Qualifications
- Experience with RLHF for large language models.
- Familiarity with multi-agent RL or hierarchical RL.
- Exposure to robotics, control systems, or autonomous driving.
- Publications in RL or related research venues.
- Open-source contributions to RL libraries or environments.
How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at View phone number on click.appcast.io. Learn more about Bright Vision Technologies at Bright Vision Technologies is an Equal Opportunity Employer.
Equal Employment Opportunity (EEO) Statement Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall. BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Reinforcement Learning Engineer in United States vacancy
$170k - $200k
...Job Description Job Description Senior Reinforcement Learning & Autonomous Decision Systems Engineer Huntsville, AL Who We Are Aurex is a mission-focused aerospace and defense company building the next frontier of deterrence. From hypersonics and missile defense...SuggestedContract workWork experience placementWork at office$96k - $120k
...Reinforcement Learning Engineer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and...SuggestedFull timeH1bRemote workVisa sponsorship- ...Reinforcement Learning Engineer As a Reinforcement Learning Engineer, you will be the architect of the core intelligence for Hammerhead's ORCA platform. Reporting to the Head of AI / Reinforcement Learning Engineering, you will design, train, and deploy the Orchestrated...SuggestedShift work
- ...Reinforcement Learning Engineer Job Title Reinforcement Learning Engineer Job Summary We are seeking a talented Reinforcement Learning (RL) Engineer to design, develop, and optimize intelligent agents capable of learning through interaction...SuggestedFlexible hours
- ...Reinforcement Learning Expert Dexmate is building the foundation for physical AI — a unified platform that combines high-quality robotic... ...simulation and real-world environments Collaborate with robotics engineers to integrate RL models into production systems Conduct...Suggested
- ...To support advanced reinforcement learning research, the part-time contract Reinforcement Learning Engineer will develop and test tool-based environments for knowledge work applications, focusing on creating robust simulation frameworks for model interactions with complex...Contract workPart timeRemote work
$155.28k - $200k
Do you want to build the scalable reinforcement learning framework that powers the next generation of humanoid and quadruped robots? As a Staff RL Research Engineer, you'll own the RL stack, including massively parallel simulation, domain randomization, policy optimization...Hourly payFull time$298k - $368k
...This role is at the intersection of robotics and machine learning, driving the next generation of operational efficiency for Waymo... ...advanced robotics. Key work involves leveraging foundation models, reinforcement learning, simulation, and integrating ML models in production...Full timeRemote work- ...challenges like safety, commercialization, and mass production to change the world for the better. JOB SUMMARY The Senior Reinforcement Learning Engineer is a key, hands-on role focused on achieving state-of-the-art performance on our humanoid robots. This engineer will...Full timeLocal area
- ...Helix team is responsible for developing the core AI systems that power humanoid autonomy. We are looking for a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience....Full timeWork at office
- ...startups. We’ve raised $16M from top VCs and were YC W25. About the role We’re looking for a Full-Stack Software Engineer, Reinforcement Learning to build the product surfaces, backend systems, and internal tools that power HUD’s RL data engine. You’ll own...Full timeWork at officeRemote workRelocationVisa sponsorship
- ...of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital... ...research —including deep learning, generative AI, and reinforcement learning techniques— with large-scale engineering to bridge experimentation and production; you'll...Full time
- ...Earth. We are seeking a talented Senior Autonomy Machine Learning Engineer . As a Senior Autonomy Machine Learning Engineer,... ...foundation models ~ Robot learning, imitation learning, or reinforcement learning ~ Natural-language planning, tool use, or...Full time
- ...benefits in a high-growth, mission-driven environment. Learn from an experienced team that has built and sold startups... ...Economics of AI News & Blogs Role Description As a Reinforcement Learning Engineer, you will be the architect of the core intelligence for Hammerhead...Shift work
$224k - $356.5k
...driving vehicle technology by bringing to bear the power of Deep Reinforcement Learning (RL). As a world leader in AI and high-performance... ...world production. We are looking for a Reinforcement Learning Engineer to join our mission in building intelligent, safe, and efficient...Full time$100k - $150k
...AI Learning Systems Engineer - Remote Bright Vision Technologies is a technology consulting and software development company delivering... ...insufficient. The role requires deep familiarity with modern reinforcement learning algorithms, simulation environments, reward...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...Responsibilities Design, develop, and deploy learned behavior planning models using approaches such as reinforcement learning, behavior cloning, and imitation... ...Required skills MS or PhD in Computer/Software Engineering, Robotics, Machine Learning or related field....
$174.72k - $295.68k
...through cutting-edge R&D in AI, machine learning, and smart connectivity.Job... ...-to-end robot motion controllers using reinforcement learning, imitation learning, or other... ...the latest advancements in academic and engineering research for humanoid robotics.Minimum...Full time$112.7k - $169.1k
...Unity's Vector AI team builds the machine learning systems that decide which ads reach... ...monthly users on the world's leading game engine. Recommendation and ranking systems are... ...frontier has shifted — large language models, reinforcement learning from human feedback, and...InternshipWork at officeWorldwideRelocation packageShift work- ...develop, and implement end-to-end machine learning pipelines, from data ingestion and... ...Collaborate with the general software engineering team to integrate ML models into existing... ...including supervised, unsupervised, and reinforcement learning. Experience with various machine...
$230k - $345k
...Most Innovative Companies, and Forbes World’s Best Banks. Visit our Institutional Page We're looking for a Staff Machine Learning Engineer to help lead the technical direction of our recommendation systems. This is a hands-on senior individual contributor role for...Full timeWork at officeWork from homeRelocation packageFlexible hours- ...take initiative, streamline complex workflows, and continuously learn and adapt. Moveworks is trusted by over 5.5 million employees... ...ServiceNow’s leading workflow automation with Moveworks’ Reasoning Engine and natural language capabilities, we deliver the AI platform...Full timeWork at officeRemote workFlexible hours
- ...take initiative, streamline complex workflows, and continuously learn and adapt. Moveworks is trusted by over 5.5 million employees... ...ServiceNow’s leading workflow automation with Moveworks’ Reasoning Engine and natural language capabilities, we deliver the AI platform...Full timeWork at officeImmediate startRemote workFlexible hours
$187k - $292k
...differentiated by its expertise in imagining, engineering, and delivering robots with advanced... ...The AI Controls team builds high-rate learned controllers that let Digit move... ...Controls Engineer, you'll develop and deploy reinforcement learning policies across humanoid...Full timeTemporary workRelocation packageFlexible hours$193.93k - $291.15k
...the Role Our robotics team is growing and we are looking for a Software Engineer to join our Sensor Data and Calibration team. We are searching for an engineer with robotics and machine learning expertise to develop synthetic sensor simulation models and algorithms....Full timeImmediate startFlexible hours- ...Generation is building closed-loop decision systems that use machine learning to operate consumer businesses more intelligently. We are... ...marketplace, credit, or other decision systems; bandits, reinforcement learning, optimization, or active learning; uncertainty...Full time
- ...Together, we’re helping build a financial system that is open to everyone. Join us. The Role As a Staff Applied Machine Learning Engineer focused on Intelligent Data, Signals & Systems, you will build production ML systems that transform customer behavior, product...Full timeTemporary workLocal areaFlexible hours
- ...take initiative, streamline complex workflows, and continuously learn and adapt. Moveworks is trusted by over 5.5 million employees... ...ServiceNow’s leading workflow automation with Moveworks’ Reasoning Engine and natural language capabilities, we deliver the AI platform...Full timeWork at officeImmediate startRemote workFlexible hours
$130k - $250k
...team of high-performers. We convert audience attention into action through data, machine learning, and continuous optimization. We’re hiring a Machine Learning Engineer (Recommendation Systems) to build the personalization engine behind our portfolio of brands....Full timeRemote work$251k - $310k
...billions in simulation across 15+ U.S. states. As a Staff Software Engineer (L6) on the Labeling Team, you will lead the technical strategy... ...in the field of software engineering and applied machine learning ~ Experience programming in C++ or Python ~ Experience...Full timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Reinforcement Learning Engineer. Be the first to apply!



