Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Engineer for Reinforcement Learning

$220k

SaidGig

This role focuses on advancing the capabilities of modern AI models through the development of novel Reinforcement Learning (RL) environments, training pipelines, and evaluation systems. As a key member of the Research Engineering Core team, you will operate at the intersection of research and production, translating experimental ideas into scalable, high-performance systems. Key Responsibilities

  • Architect self-contained RL environments that capture complex, real-world tasks, including reward functions, verifiers, and evaluation logic.
  • Design and scale episode pipelines and multi-component training processes (MCPs) to support reproducible experimentation.
  • Build automated data generation systems, leveraging synthetic data to accelerate training cycles without compromising quality.
  • Develop and integrate AI-driven evaluation and quality assurance systems for automated grading, validation, and feedback loops.
  • Fine-tune and optimize open-source RL models using internally generated datasets and custom training strategies.
  • Establish benchmarking frameworks to measure model capability, robustness, and data quality across tasks.
  • Contribute to the release and analysis of evaluations on internal and external benchmark platforms.
Qualifications
  • Deep experience in Reinforcement Learning, including environment design and training dynamics.
  • Strong track record of building and scaling RL systems, pipelines, or experimentation frameworks.
  • Proficient in automation and data generation, including synthetic data pipelines.
  • Familiar with automated evaluation systems, model validation, and quality assurance workflows.
  • Experienced in fine-tuning and evaluating open-source ML models.
  • Clear, concise communicator with strong technical writing skills.
  • Comfortable operating in fast-paced, research-driven, and highly collaborative environments.
Work Terms

This is a full-time remote position.

Compensation

The national pay range for this position is a base salary of $220, 000, $500, 000 USD per year. All employees are eligible for equity compensation, and performance-based bonuses may also be available, depending on the role and company policies.

Eligibility

micro1 is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, sexual orientation, or gender identity), national origin, age, disability, genetic information, veteran status, or any other characteristic protected by applicable local laws and regulations.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Research Engineer for Reinforcement Learning in United States vacancy
  • $126k - $423k

     ...Stockholm; Bangalore; Seoul; and Tokyo. Learn more at applied.co. We are an in-...  ...We are looking for multiple passionate Research Engineers to join the Research Group at Applied...  ...Intuition, You Will Conduct research on reinforcement learning (RL) and its training... 
    Suggested
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Immediate start
    Remote work
    Day shift

    Applied Intuition

    California
    22 days ago
  •  ...Platform (shared infra & clean code) and Embedded (inside research squads). Engineers can move along the researchproduction spectrum as needs or...  ...- ML track, you'll build and optimise the large-scale learning systems that power our open-weight models. Working hand-in... 
    Suggested
    Relocation package

    Mistral AI

    Palo Alto, CA
    23 hours ago
  • $170k - $216k

     ...team is to develop machine learning solutions addressing open problems...  ...collaborations with other research teams in Alphabet. AI...  ...currently focusing on include reinforcement learning, learning from demonstration...  ...in collaboration with engineering teams across Waymo... 
    Suggested
    Full time
    Remote work

    Waymo

    Remote
    19 hours ago
  •  ...giving countless hours back to our customers so they can spend more time on the things they value most. As a Machine Learning Research Engineer, you will work on the software and algorithms that enable our robots to complete dexterous manipulation tasks in home environments... 
    Suggested

    Sunday

    Redwood City, CA
    2 days ago
  • $200k - $350k

     ...stage AI company that’s redefining how models learn to understand subjective quality , from...  ...deeply curious—building at the intersection of research, product, and creativity . The Role As a Machine Learning Research Engineer , you’ll own end-to-end research cycles—designing... 
    Suggested

    Coders Connect

    San Francisco, CA
    4 days ago
  • $176k - $255k

     ...accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training...  ...in Computer Science, Machine Learning, AI, or a related field....  ...of deep learning, reinforcement learning, and large-scale model... 
    Full time
    Shift work

    Scale AI

    New York, NY
    more than 2 months ago
  • $208k - $260k

     ...This position will be a key contributor in conducting applied research in Robotics and developing ML pipelines for training and...  ...robotics, computer vision, embodied AI, sim-to-real, imitation learning, reinforcement learning, and vision language actions models  ~ PhD or... 
    Full time
    Shift work

    Scale AI

    San Francisco, CA
    more than 2 months ago
  •  ...Research Engineer We believe that software is the foundation of modern civilization - yet vulnerabilities threaten its integrity...  ...intuition, experience in model evaluation, and benchmarks. Reinforcement Learning experience is a plus. Your work will play a crucial role... 
    Full time
    Work at office

    DepthFirst

    San Francisco, CA
    3 days ago
  •  ...Voltai AI + Semiconductor Research Position Voltai is developing...  ...world models, and agents to learn, evaluate, plan, experiment,...  ...training, inference, and reinforcement learning systems that reason...  ...closely with researchers and engineers, you'll help make Voltai the... 

    Voltai

    Palo Alto, CA
    3 days ago
  •  ...About Mbodi Mbodi makes industrial robots learn like humans. Our AI software plugs into...  ...In This Role, You Will Conduct applied research at the intersection of generative AI and...  ...vision language models, imitation learning, reinforcement learning, or robotic foundation models.... 

    Mbodi AI

    New York, NY
    3 days ago
  •  ...Research Engineer Bespoke Labs is an applied AI research lab pioneering data and RL environment curation for training and evaluating...  ..., and taught agents to do multi-turn tool-calling with reinforcement learning. Bespoke is uniquely positioned to capture a large market... 
    Remote work

    Bespoke Labs

    United States
    3 days ago
  •  ...Archive Human Archive is a research lab backed by Y Combinator...  ...embodied intelligence as a learned model. To achieve this, we...  ...Opportunity As a Research Engineer, you'll work on multimodal...  ...systems Experience with reinforcement learning, real-world robot deployments... 
    Shift work

    Human Archive

    San Francisco, CA
    4 days ago
  • $180k - $340k

     ...Research Engineer You'll own the quality of AI across everything Gamma creates. As our Research Engineer, you'll design evaluation...  ...with post-training techniques for LLMs including reinforcement learning and supervised fine-tuning ~ Exceptional attention to detail... 
    Full time
    Work at office
    Work from home

    Gamma

    San Francisco, CA
    2 days ago
  •  ...The Role: We are looking for Research Engineers to build AI systems that use agent interaction...  ...at scale, and improve them through learning and feedback. Your research will...  ...settings You have a strong background in reinforcement learning, agents, or machine learning... 
    Immediate start

    Judgment Labs

    San Francisco, CA
    1 day ago
  • $300k - $350k

     ...is committed to world‑class research. We empower exceptional talents...  ...quantitative researchers, engineers, and ML experts leading...  ...experience with modern deep learning frameworks (PyTorch and/or JAX...  ...acceleration (FPGA/ASIC). Knowledge of reinforcement learning, post‑training, or... 

    Jump Trading

    New York, NY
    3 days ago
  • $120k - $250k

     ...talk with your recruiter to learn more. Base pay range $120,00...  ...high-performance, AI-ready data engines. Your work here will directly...  .... About the Role As a Research Engineer at Orbifold AI, you...  ...boundaries of training, RAG, and reinforcement learning for enterprise AI applications... 
    Full time
    Internship
    Flexible hours

    Orbifold AI

    Los Altos, CA
    4 days ago
  •  ...building Agentic AI that empowers software engineers by automating production engineering...  ...workflows end‑to‑end, balancing research and engineering to create production‑ready...  ...with novel techniques, including reinforcement learning, retrieval‑augmented generation, and autonomous... 
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Resolve AI

    San Francisco, CA
    23 hours ago
  • $100k - $300k

     ...massive scale through data-driven machine learning is the key to unlocking these...  ...projects. Position Overview We are hiring Research Engineers to develop scalable robotic systems aimed...  ...research experience in deep learning, reinforcement and/or imitation learning, robotics,... 
    Full time

    Skild AI

    San Francisco, CA
    23 hours ago
  • $350k

     ...Research Engineer On Alignment Science Anthropic's mission is to create reliable, interpretable...  ...and run elegant and thorough machine learning experiments to help us understand and...  ...our interventions. Run multi-agent reinforcement learning experiments to test out... 
    Work at office
    Visa sponsorship
    Flexible hours

    Colorwave Inc

    San Francisco, CA
    2 days ago
  •  ...About the Role We are looking for Research Scientists and Research Engineers to help us build the data engine for...  ...environments where agents learn to navigate ambiguity and maintain...  ...Expertise: Prior experience with Reinforcement Learning (RLHF/RLAIF), simulation... 

    Gravity Engineering Services Pvt Ltd.

    Sunnyvale, CA
    3 days ago
  •  ...’ll work at the intersection of machine learning infrastructure, applied AI, and distributed...  ...Plaid. As a Staff Machine Learning Engineer, you will lead the technical strategy...  ...scalable, repeatable pipelines that translate research into production impact. You will also... 
    Full time
    Work experience placement
    Local area
    Immediate start

    Plaid Inc.

    San Francisco, CA
    19 hours ago
  •  ...ll share more once we meet. About the Role As an ML Research Engineer at Maple, you'll be a part of our core product team...  ...Fine-tune LLMs with retrieval-augmented generation (RAG), reinforcement learning (RL), and prompt engineering for dynamic, context-aware conversations... 
    Work at office
    Local area

    Maple AI, Inc

    New York, NY
    23 hours ago
  • $159k - $296k

     ...impact the world in a positive way. To learn more visit: The Motion Planning...  ...trajectories for our self-driving trucks. As a research engineer for Learnable Planner you will support...  .../decision making (e.g., imitation and reinforcement learning, optimization-based... 
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    Pittsburgh, PA
    9 days ago
  • $96k - $214.5k

     ...and systematically explore their role in health and disease. The Role We are seeking a highly motivated Senior Machine Learning Engineer / Data Scientist to join our AI/ML team. This individual will play a central role in designing and implementing advanced AI/ML... 
    Full time

    Flagship Pioneering, Inc.

    Boston, MA
    19 hours ago
  • $250k - $300k

     ...looking for a Principal Software Engineer to architect the systems...  ...at the intersection of deep learning, computational geometry, and...  ...to apply cutting-edge AI research to one of humanity's most fundamental...  ...optimization. Create reinforcement learning systems for multi-... 
    Full time
    Work at office
    Home office

    Zero Rfi

    San Francisco, CA
    19 hours ago
  •  ...mission to democratize AI. Leveraging proprietary AI Studio and AI Engines, the company helps drive the clients’ AI Enterprise...  ...Remote Role Overview We’re hiring a mid-to-senior Machine Learning Engineer / Data Scientist to build and deploy machine learning... 
    Remote job
    Full time
    H1b
    Local area

    Fusemachines

    Remote
    19 hours ago
  •  ...NeurIPS, ICML, and ACL - and see that research deployed to a user base of over 80...  ...As an AI Agents Applied Research - Engineering Executive Director in our The Digital...  ...latency, memory, and cost. Apply reinforcement learning and preference optimization to improve... 

    J.P. Morgan

    Palo Alto, CA
    20 days ago
  • $170k - $240k

     .... The Role: As an Applied Machine Learning Engineer, you will serve as a vital bridge between cutting-edge AI research and practical, real-world applications. Your...  ...supervised fine-tuning (SFT) and reinforcement learning from human feedback (RLHF or RFT)... 
    Full time

    Fireworks Ai

    Remote
    19 hours ago
  •  ...the digital infrastructure that modern life runs on. They’ve learned—correctly—that those attacks rarely produce consequences....  ...Twenty is seeking an exceptionally skilled Offensive Cyber Research Engineer for an in-office position in its Arlington, VA office to lead... 
    Full time
    Work at office
    Worldwide
    Flexible hours

    Twenty

    Washington DC
    23 days ago
  • $220k - $292k

     ...providing needed context for our users. AI Engineers on Anduril’s Frontier AI team build edge-compatible...  ...Anduril to help discover and scope new research problems REQUIRED QUALIFICATIONS BS in Computer Science, Machine Learning, Electrical Engineering, or related field 5+... 
    Full time
    Work experience placement

    Anduril

    Costa Mesa, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Engineer for Reinforcement Learning. Be the first to apply!