Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Engineer - Scalable QC for RL Training Data

HUD

HUD is seeking Research Engineers to automate QC for training data generated through our ML infrastructure, building systems that scale quality for RL training. You’ll implement robust data validation pipelines, sampling strategies, and feedback loops while collaborating with data vendors. Ideal candidates have strong Python, Docker, and Linux skills, startup experience, and a knack for designing metrics and experiments to assess data quality and task realism. #J-18808-Ljbffr HUD

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Research Engineer - Scalable QC for RL Training Data in San Francisco, CA vacancy
  • $150k - $180k

     ...evaluation, synthetic data, and...  ...measure and improve training data quality...  ...shaping internal research culture around...  ...of tasks across RL environments, synthetic...  ...by building QC systems,...  ...with research engineers, domain experts...  ...convert it into scalable review or generation... 
    Data
    Training
    Relocation
    Visa sponsorship

    Clera

    San Francisco, CA
    5 days ago
  •  ...building infrastructure to create RL training data and evals for frontier AI agents, as...  ...About the role We're looking for Research Engineers to automate QC for training data created by companies...  ...means and how to measure it Built scalable data validation pipelines and... 
    Data
    Training
    Full time
    Work at office
    Remote work
    Relocation
    Visa sponsorship

    HUD

    San Francisco, CA
    4 days ago
  •  ...Description We are looking for a hybrid Systems Engineer and AI Researcher to lead the development of our agent evaluation framework and post-training data pipelines. You will design sandboxed...  ...safely. Scale Post-Training & RL Pipelines: Implement high-throughput... 
    Data
    Training
    Work at office

    Hyphen Connect

    San Francisco, CA
    17 days ago
  • $100k - $200k

     ...Research Engineer, QC Automation Location: San Francisco Bay Area, CA — On-site Employment...  ...infrastructure for companies creating training data for AI agents. As demand grows, we need...  ...requirements into measurable, scalable systems. What You’ll Do Build... 
    Data
    Training
    Full time
    Visa sponsorship
    Relocation package

    Dreams 2 Reality Recruitment

    San Francisco, CA
    10 days ago
  • $100k - $200k

     ...infrastructure to create reinforcement-learning training data and evaluations for frontier AI...  ...: Our client is seeking a Research Engineer, QC Automation to automate quality control...  ...understand them Experience building scalable data-validation pipelines or... 
    Data
    Training
    Full time
    Visa sponsorship
    Relocation package

    Invictus Direct

    San Francisco, CA
    12 days ago
  •  ...strong infrastructure engineer to build the systems layer for RL at scale. You will enable...  ...experiments, tackling training orchestration, data pipelines, and...  ...while aligning with ML researchers to translate complex experiments...  ...into durable, scalable platforms within our San... 
    Data
    Training
    Work at office

    Vmax AI Corp

    San Francisco, CA
    2 days ago
  •  ...strengthening frontier models for cyberdefense tasks. You will contribute to post-training models for adversarial capabilities using reinforcement learning in a distributed setting. You’ll handle data generation, environment design, and evaluations across the training... 
    Data
    Training

    General Analysis

    San Francisco, CA
    5 days ago
  • $250k - $350k

     ...environments, and frontier research benchmarks that...  ...in software engineering, enterprise knowledge...  ...and longest-running data provider in the category...  ...learning (RL) environments that power post-training for the world’s leading...  ...promising ideas into scalable applications.... 
    Data
    Training
    Full time
    Work at office

    Turing

    San Francisco, CA
    4 days ago
  • $100k - $300k

     ...believe massive scale through data-driven machine learning is...  ...Overview We are hiring Research Engineers to develop scalable robotic systems aimed at...  ...implement new algorithms for training and optimizing general-...  ...disciplines (Perception, Robotics, RL/IL, Machine Learning, etc.... 
    Data
    Training
    Full time

    Skild AI

    San Francisco, CA
    1 day ago
  •  ...RSI) team works across research, engineering, product, and...  ...designing evaluations, and training models to develop missing...  ...and model failures into data and evaluation...  ...harnesses, synthetic data, RL environments, and model...  ...into rigorous, reliable, scalable results. You might... 
    Data
    Training

    OpenAI

    San Francisco, CA
    2 days ago
  • $275k

     ...Research Engineer Magic's mission is to build safe AGI that accelerates humanity's progress...  ...combines frontier-scale pre-training, domain-specific RL, ultra-long context, and inference-time...  ...capabilities Build out internet-scale data pipelines and crawlers Design,... 
    Data
    Training
    Relocation
    Visa sponsorship

    Magic Inc

    San Francisco, CA
    4 days ago
  • $175k - $250k

     ...Research Engineer About Scorecard We’re a small, nimble team backed by...  ...build; your simusers run inside training environments where their...  ...evaluations with baselines, held-out data, and calibrated judges, and...  ...annotation pipelines. RL experience: building RL environments... 
    Data
    Training
    Work at office

    Kindredventures

    San Francisco, CA
    1 day ago
  •  ...Research Engineer We believe that software is the foundation of modern...  ...software vulnerabilities. We are training and scaling security AI...  ...includes deep expertise in data, infrastructure, security and...  ...solutions. Building reliable and scalable products, making right trade... 
    Data
    Training
    Full time
    Work at office

    DepthFirst

    San Francisco, CA
    4 days ago
  •  ...Research Engineer New York - Hybrid; San Francisco Bay Area - Hybrid...  ...AI platform structures messy data, automates digital workflows,...  ...methodology, benchmarks, and RL environments for frontier labs...  ...systems, RL environments, or training and inference infrastructure... 
    Data
    Training
    Full time
    Work at office
    Local area
    Remote work

    Invisible Technologies Inc. Defunct

    San Francisco, CA
    5 days ago
  •  ...AfterQuery is an applied research lab curating data solutions for foundation model...  ...opportunity to shape the engineering organization and lead major...  ...You will design and run training experiments that isolate the...  ...Through controlled SFT and RL - base post-training experiments... 
    Data
    Training
    Local area

    AfterQuery

    San Francisco, CA
    3 days ago
  •  ...become a massive tax of engineering velocity. Resolve AI...  ...end-to-end, balancing research and engineering to create...  ...Build and optimize data pipelines to process high...  ...unstructured data for training and evaluation...  ...translating research into scalable AI-powered solutions... 
    Data
    Training
    Work at office
    Visa sponsorship
    Flexible hours

    Resolve AI

    San Francisco, CA
    1 day ago
  • $200k - $350k

     ...spanning Ex-Moonshot (post-training & agents, diffusion...  ...time technical founders, engineers that made 100+ games...  ...generation pipelines, data-efficient training methods...  .... Our current research spans: Distributed...  ...vision, world models, or RL systems. Strong Candidates... 
    Data
    Training
    Visa sponsorship
    Relocation package

    ROAM

    San Francisco, CA
    5 days ago
  • $175k - $250k

     ...push the boundaries of AI research and development. Their mission...  .... The Role: As a Research Engineer in Pre-Training, you'll develop and...  ...strategies, and large-scale data pipelines to drive AI innovation...  ...-world applications. Build scalable pre-training pipelines for... 
    Data
    Training
    Full time
    Relocation package

    HartleyCo

    San Francisco, CA
    1 day ago
  •  ...California. The Role: As a Research Engineer - Agency and Reasoning , you...  ...reinforcement learning, post-training, and human preference...  ...reasoning or more classical RL tasks Experience with language...  ...in grappling in detail with data and spending significant time... 
    Data
    Training
    Work at office
    Relocation package

    Zyphra

    San Francisco, CA
    a month ago
  • $197.3k - $313.7k

    ## Applied Research EngineerApplyremote type: Office...  ...software and platform engineers to embed in our AI team...  ...building substantive, scalable infrastructure is what...  ...we use your personal data and your rights, including...  ...promotion, benefits, training, assessment of job... 
    Data
    Training
    Work at office

    Salesforce, Inc.

    San Francisco, CA
    4 days ago
  • $155k - $269k

     ...World, which delivers realistic, scalable, controllable, and efficient simulation. As a Research Engineer in the World Models team, you...  ...planning, testing, and training.   You will… - Design, implement...  ...datasets. - Build large scale data pipelines to build high... 
    Data
    Training
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    a month ago
  • $200k - $330k

     ...experienced Machine Learning Engineer to build and improve...  .... Develop highly scalable ETL pipelines to...  ...petabyte-scale protein data for model pretraining Optimize model training and inference code to maximize...  ...to prototype research ideas and bring them into... 
    Data
    Training

    Profluent

    Emeryville, CA
    1 day ago
  •  ...-of-the-art perception and foundation-model driven robotics. As a Member of Technical Staff, Research, you will advance perception models, data pipelines, and scalable training workflows for real-world robots deployed in manufacturing. You will collaborate across hardware... 
    Data
    Training

    Industrial Next (YC W22)

    San Francisco, CA
    4 days ago
  •  ...encouraged to apply. Mission Design, train, ship, iterate on, and innovate on...  ...The Path’s AI Therapist. Combine research, data science, and engineering to create models, orchestration, and...  ..., training data curation, building RL environments, new model... 
    Data
    Training

    Monograph

    San Francisco, CA
    1 day ago
  •  ...America’s surging demand for housing, data centers, and manufacturing...  ...from data collection and model training through edge deployment on Jetson AGX Orin. Every research project will have a deployment milestone...  ...two of: imitation learning, RL, vision-language models, robot... 
    Data
    Training

    Origin

    San Francisco, CA
    a month ago
  •  ...AI Research Engineer (Robot Learning) San Francisco AI & Software In...  ...robotic hands and cutting edge AI trained on human observations, bring...  ...AI model development and data flywheel project management for...  ...Experience with RL fine-tuning of generative models... 
    Data
    Training
    Full time
    Work at office
    Immediate start

    Software Engineering, Data Science

    San Francisco, CA
    more than 2 months ago
  •  ...Francisco, California | Primarily On-site We are seeking an Research Engineer – RL Infrastructure & Agent Environments to build the...  ..., evaluation systems, and supporting infrastructure used to train and assess long-horizon enterprise AI agents. The Opportunity... 
    Training

    MaxIT Consulting - Max Corporate Group

    San Francisco, CA
    11 days ago
  • $200k - $350k

     ..., a new primitive for training efficient, large-scale...  ...innovation and systems engineering paired with a design-...  ...About the Role Data is the lifeblood of our...  ...we are looking for a Research Engineer, Data Infrastructure...  ...write performant, scalable infrastructure to acquire... 
    Data
    Training
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Cartesia

    San Francisco, CA
    4 days ago
  •  ...Taste Labs is building the data and infrastructure layer for...  ...two sides: building the post-training data and RL environments that teach taste...  ...problems you’d work on Research different grading methods and...  ...and context layers Work on scalable RL infra Work with... 
    Data
    Training
    Full time

    Taste Labs

    San Francisco, CA
    17 days ago
  • $180k - $240k

     ...is hiring a talented ML/AI Research Engineer to join their team in San Francisco...  ...will lead the design, training, evaluation and optimization...  ...on top of their enterprise data fabric. Responsibilities ~...  ...and decision-making. ~Create scalable evaluation frameworks,... 
    Data
    Training
    Full time

    Alldus International Consulting Ltd

    San Francisco, CA
    more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Engineer - Scalable QC for RL Training Data. Be the first to apply!