Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Reinforcement Learning Engineer

$100k - $150k
Full-time

Bright Vision Technologies

Reinforcement Learning Engineer - Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. Job Title: Reinforcement Learning Engineer Location: 100% Remote (U.S.) Position Type: Full-time, Direct W2 Salary Range: $100,000–$150,000 Annually Experience Required: 6+ years Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. Job Summary: We are looking for a Reinforcement Learning Engineer to design, train, and deploy RL-based systems for high-impact decision-making problems where supervised learning alone is insufficient. The role requires deep familiarity with modern reinforcement learning algorithms, simulation environments, reward modeling, and the engineering complexity of training and evaluating policies at scale. The ideal candidate has both research depth and engineering pragmatism, with experience taking RL solutions out of the lab and into production where stability, safety, and ongoing improvement are critical. Key Responsibilities * Design and implement reinforcement learning solutions for sequential decision-making problems in real and simulated environments. * Develop, calibrate, and maintain simulation environments suitable for large-scale agent training. * Implement and evaluate modern RL algorithms including policy gradient, actor-critic, off-policy, and offline RL methods. * Engineer reward functions and shaping strategies that align agent behavior with desired outcomes and safety constraints. * Apply offline RL and imitation learning techniques where exploration is costly or unsafe. * Use RLHF, DPO, and related techniques for fine-tuning large language models when relevant. * Build scalable training infrastructure for distributed RL, including efficient experience collection and replay systems. * Optimize training stability and sample efficiency through algorithmic and engineering improvements. * Design rigorous evaluation protocols, including out-of-distribution and adversarial test cases. * Implement safety mechanisms such as constraint enforcement, conservative policies, and human-in-the-loop oversight. * Collaborate with applied scientists and product teams to identify high-value RL use cases. * Monitor deployed policies and models in production for drift, regression, and unintended behaviors, building the alerting and dashboards that surface issues before they meaningfully affect users. * Document methodology, design decisions, and operational characteristics for internal stakeholders. * Stay current with RL research and translate promising techniques into production-ready solutions. Required Qualifications * Master’s or PhD in Computer Science, Machine Learning, or a related field; or equivalent applied experience.

  • Six or more years of combined RL research and engineering experience.
  • Strong proficiency in Python and modern deep learning frameworks.
  • Hands-on experience with at least one major RL library or in-house RL stack.
  • Solid understanding of probability, optimization, and the theoretical
foundations of RL.
  • Experience designing and tuning reward functions in non-trivial environments.
  • Familiarity with simulation environments and large-scale experience
collection.
  • Experience training neural network policies on GPU clusters.
  • Strong written and verbal communication skills.
  • Track record of shipping or publishing impactful RL work.
Preferred Qualifications
  • Experience with RLHF for large language models.
  • Familiarity with multi-agent RL or hierarchical RL.
  • Exposure to robotics, control systems, or autonomous driving.
  • Publications in RL or related research venues.
  • Open-source contributions to RL libraries or environments.
How to Apply Would you like to know more about this opportunity? For immediate consideration, please send your resume to View email address on click.appcast.io or contact us at View phone number on click.appcast.io. Learn more about Bright Vision Technologies at [ Bright Vision Technologies is an Equal Opportunity Employer. Equal Employment Opportunity (EEO) Statement Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall. BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Reinforcement Learning Engineer in United States vacancy
  • $100k - $150k

     ...Reinforcement Learning Engineer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and... 
    Suggested
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    United States
    2 days ago
  •  ...Job Description Job Description Reinforcement Learning (RL) Engineer Location: New York (Office) On-site | Full-time Compensation: Competitive Our client is an elite development firm and a high-growth software company responsible for building the infrastructure... 
    Suggested
    Full time
    Work at office
    Immediate start

    MLabs

    New York, NY
    8 days ago
  • $220k

     ...As a Research Engineer, you will operate at the forefront of Reinforcement Learning (RL), developing innovative environments, training pipelines, and evaluation systems that enhance the capabilities of modern AI models. This role uniquely combines research and production... 
    Suggested
    Full time
    Local area
    Remote work

    SaidGig

    United States
    2 days ago
  •  ...in a high-growth, mission-driven environment. &##128640;Learn from an experienced team that has built and sold startups...  ...Economics of AI News & Blogs Role Description As a Reinforcement Learning Engineer, you will be the architect of the core intelligence for... 
    Suggested
    Shift work

    Hammerhead AI

    Redwood City, CA
    13 days ago
  • $298k - $368k

     ...This role is at the intersection of robotics and machine learning, driving the next generation of operational efficiency for Waymo...  ...advanced robotics. Key work involves leveraging foundation models, reinforcement learning, simulation, and integrating ML models in production... 
    Suggested
    Full time
    Remote work

    Waymo

    Remote
    19 hours ago
  • $100k - $150k

     ...AI Learning Systems Engineer – Remote Bright Vision Technologies is a technology consulting and software development company delivering...  ...insufficient. The role requires deep familiarity with modern reinforcement learning algorithms, simulation environments, reward... 
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    United States
    19 hours ago
  •  ...challenges like safety, commercialization, and mass production to change the world for the better. JOB SUMMARY The Senior Reinforcement Learning Engineer is a key, hands-on role focused on achieving state-of-the-art performance on our humanoid robots. This engineer will... 
    Full time
    Local area

    Apptronik

    Texas
    15 days ago
  • $244.14k - $413.16k

     ...transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.We are looking for exceptional Research Engineers / Scientists to design learning systems...  ....This role sits at the intersection of reinforcement learning, large language models, and real-... 
    Full time
    Part time

    XPENG Motors

    Santa Clara, CA
    4 hours ago
  • $200k - $300k

     ...humanoid robots with human level intelligence. Its robots are engineered to perform a variety of tasks in the home and commercial...  ...collaboration. It’s time to build. We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate advanced reinforcement... 
    Full time
    Work at office

    Figure

    California
    26 days ago
  •  ...Helix team is responsible for developing the core AI systems that power humanoid autonomy. We are looking for a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience.... 
    Full time
    Work at office

    Figure

    Remote
    19 hours ago
  •  ...of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital...  ...research —including deep learning, generative AI, and reinforcement learning techniques— with large-scale engineering to bridge experimentation and production; you'll... 
    Full time

    Roblox

    Remote
    19 hours ago
  • $225k - $280k

     ...quality, and trust. The Role We’re looking for a Machine Learning Engineer to design and build feedback driven learning systems that...  ...into reliable learning signals Design lightweight reinforcement learning / bandit-style approaches where appropriate Partner... 
    Remote job
    Full time
    Flexible hours

    Wizard

    Remote
    19 hours ago
  • $215.28k - $364.32k

     ...Staff Robotics Engineer / Tech Lead – Whole-Body Control & Robot Learning Santa Clara, CA XPENG is a leading smart technology company at the forefront...  ...related field. ~5+ years of relevant experience in reinforcement learning, robotics, robot learning, control... 
    Full time

    XPENG

    Santa Clara, CA
    4 days ago
  • $65k - $82k

     ...management platform in Charlotte, NC is seeking an Associate QA Engineer to join their Quality Engineering team. This entry-level role...  ...on building foundations in testing and automation, where you'll learn from experienced professionals. Key responsibilities include executing... 

    AssetMark

    Charlotte, NC
    4 days ago
  • $240k

     ...defense layer for the AI age and are looking for an exceptional ML engineer to stabilize the system that turns raw signal into decisions —...  ...government officials, and executive protection units. What we learn there becomes the foundation for a civilization-defining... 
    Full time
    Flexible hours

    Sweep360

    New York, NY
    19 hours ago
  •  ...take initiative, streamline complex workflows, and continuously learn and adapt. Moveworks is trusted by over 5.5 million employees...  ...ServiceNow’s leading workflow automation with Moveworks’ Reasoning Engine and natural language capabilities, we deliver the AI platform... 
    Full time
    Work at office
    Immediate start
    Remote work
    Flexible hours

    Servicenow

    Remote
    19 hours ago
  •  ...take initiative, streamline complex workflows, and continuously learn and adapt. Moveworks is trusted by over 5.5 million employees...  ...ServiceNow’s leading workflow automation with Moveworks’ Reasoning Engine and natural language capabilities, we deliver the AI platform... 
    Full time
    Work at office
    Remote work
    Flexible hours

    Servicenow

    Remote
    19 hours ago
  • $281.2k - $401.71k

     ...systems. As part of this shift, we’re creating a shared Agent Engine that powers agent-based experiences across Spotify. You’ll join...  ...partners working at the intersection of distributed systems, machine learning, and user experience to shape how millions of listeners... 
    Remote job
    Full time
    Flexible hours
    Shift work

    Spotify

    New York, NY
    19 hours ago
  • $193.93k - $291.15k

     ...the Role Our robotics team is growing and we are looking for a Software Engineer to join our Sensor Data and Calibration team. We are searching for an engineer with robotics and machine learning expertise to develop synthetic sensor simulation models and algorithms.... 
    Full time
    Immediate start
    Flexible hours

    Nuro

    Remote
    19 hours ago
  •  ...customers and partners in manufacturing, aerospace, transportation, security, safety, and compliance. We're looking for a Deep Learning Field Engineer to operate at the forefront of CV deployment in industry - building best-in-class CV systems that leverage deep learning... 
    Work experience placement
    Work at office
    Flexible hours

    Matroid

    Palo Alto, CA
    19 hours ago
  •  ...Together, we’re helping build a financial system that is open to everyone. Join us. The Role As a Staff Applied Machine Learning Engineer focused on Intelligent Data, Signals & Systems, you will build production ML systems that transform customer behavior, product... 
    Full time
    Temporary work
    Local area
    Flexible hours

    Block

    United States
    19 hours ago
  • $146.16k - $219.24k

     ...audiences and our employees - and aim to leave a positive mark on culture. Overview We're hiring a Senior Machine Learning Operations Engineer to own the operational layer around our personalization and recommendation Machine Learning (ML) systems. Our models... 
    Full time
    Shift work

    Paramount

    New York, NY
    12 hours ago
  • $170k - $216k

     ...5+ U.S. states. The Perception team builds the system which learns the spatial-temporal representation and their semantic meanings...  ...miles of driving data from a diverse set of sensors, enabling engineers like you to (1) develop methods for efficiently and continuously... 
    Full time
    Remote work

    Waymo

    Remote
    19 hours ago
  • $204k - $259k

     ...5+ U.S. states. The Perception team builds the system which learns the spatial-temporal representation and their semantic meanings...  ...miles of driving data from a diverse set of sensors, enabling engineers like you to (1) develop methods for efficiently and continuously... 
    Full time
    Remote work

    Waymo

    Remote
    19 hours ago
  • $130k - $250k

     ...team of high-performers. We convert audience attention into action through data, machine learning, and continuous optimization. We’re hiring a Machine Learning Engineer (Recommendation Systems) to build the personalization engine behind our portfolio of brands.... 
    Full time
    Remote work

    Launch Potato

    Miami, FL
    19 hours ago
  • $117.5k - $234.5k

     ...motivated principal investigators, systems engineers, and researchers who will be a part of...  ...translating ambiguity into structured learning, reducing technical risk, and ensuring...  ...such as model predictive control, reinforcement learning, etc. Pay Range The annual salary... 
    Full time
    Temporary work
    Local area

    Carrier

    East Syracuse, NY
    19 hours ago
  • $126k - $423k

     ...Stockholm; Bangalore; Seoul; and Tokyo. Learn more at applied.co. We are an in-office...  ...for multiple passionate Research Engineers to join the Research Group at Applied Intuition...  ..., You Will Conduct research on reinforcement learning (RL) and its training infrastructure... 
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Immediate start
    Remote work
    Day shift

    Applied Intuition

    California
    25 days ago
  • $251k - $310k

     ...billions in simulation across 15+ U.S. states. As a Staff Software Engineer (L6) on the Labeling Team, you will lead the technical strategy...  ...in the field of software engineering and applied machine learning ~ Experience programming in C++ or Python ~ Experience... 
    Full time
    Remote work

    Waymo

    Remote
    19 hours ago
  • $184k - $287.5k

     ...Deep Learning And Computer Vision Engineer For Autonomous Vehicles Team We are looking for a deep learning and computer vision engineer for our autonomous vehicles team. The role involves applying state-of-the-art techniques to build ground truth for autonomous vehicles... 
    Night shift

    NVIDIA

    Redmond, WA
    2 days ago
  • $200k - $237k

     ...this means leveraging cloud delivery, modern tech stacks, machine learning, and hand-crafted native app experiences on all of our...  ...and multi-user playback experiences. Senior Machine Learning Engineer (Recommendation Systems) Philo’s recommendation system improves... 
    Full time
    For contractors
    Work at office
    Home office
    Flexible hours
    3 days per week

    Philo

    San Francisco, CA
    19 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Reinforcement Learning Engineer. Be the first to apply!