Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer - RL Environments

$200k
Full-time

AfterQuery

About AfterQuery

AfterQuery builds the training data and evaluation infrastructure that frontier AI labs use to make their models better. We work with the world's leading labs to design high signal datasets and run rigorous evaluations that go beyond static benchmarks. We are a small, early team (post Series A) where individual contributors have a direct impact on how the next generation of models learn and improve.

The Role

As a SWE (Environments), you will design the datasets and evaluation rubrics that directly influence how frontier models learn. You'll work hands-on with research teams at top AI labs, experimenting with data collection strategies, diagnosing model failure modes, and developing the metrics that determine whether a model is actually improving. You'll go from hypothesis to live experiment quickly, and your output will feed directly into model training runs at scale.

Day to day, you will design data slices that expose meaningful failure modes across domains like finance, code, and enterprise workflows. You will build and refine reward signals for RLHF and RLVR pipelines. You will develop quantitative frameworks for measuring dataset quality, diversity, and downstream impact on alignment and capability. You will partner with lab research teams to translate their training objectives into concrete data and evaluation specifications.

What You'll Do

  • Design data slides and explore data shapes that expose meaningful model failure modes across domains like finance, code, and enterprise workflows

  • Build and refine evaluation rubrics and reward signals for RLHF and RLVR training pipelines

  • Model annotator behavior and run experiments to improve different model capabilities

  • Develop quantitative frameworks for measuring dataset quality, diversity, and downstream impact on model alignment and capability

  • Create and manage both real world & synthetic data pipelines

  • Partner with lab research teams to translate their training objectives into concrete data and evaluation specifications

What We're Looking For

  • 1-4 YOE

  • Major plus if they've worked for/interned for any RL environment companies in the past or any AI safety or benchmarking orgs like METR, Artificial Analysis, etc..

  • Genuine obsession with how data structure, selection, and quality drive model behavior

  • Ability to design lightweight experiments, move fast, and extract actionable insights from messy results

  • Former founders and early engineers at early stage startups are a plus. We don't filter on pedigree. We want people who can demonstrate they work hard, learn fast, and care deeply about getting the details right.

Compensation Structure:

$200k base + profit share (around 150% of base) + competitive equity

Vacancy posted 22 hours ago
Similar jobs that could be interesting for youBased on the Software Engineer - RL Environments in San Francisco, CA vacancy
  • $245k - $300k

     ...platform that turns that judgment into the data , evals , and RL environments frontier models learn from. We work with leading frontier...  ...keep it healthy once it's running. You sit between Pareto's engineering team and the researchers at the labs we work with, close... 
    Suggested
    Full time

    Pareto B.v.

    San Francisco, CA
    22 hours ago
  •  ...research lab that builds high-quality reinforcement-learning environments and agents sold to the world's leading AI labs. In under...  ...expanding into new domains. The Opportunity As an RL Environment Software Engineer, you will sit at the intersection of research... 
    Suggested
    Remote work

    talentpluto

    San Francisco, CA
    7 days ago
  • RippleMatch Inc. is seeking an innovative and motivated individual to design and refine reinforcement learning tasks in San Francisco. This role requires a strong command of Python and the ability to work independently with coding agents. Responsibilities include the full...
    Suggested

    RippleMatch

    San Francisco, CA
    2 days ago
  • $252k - $315k

     ...agentic tool use, and domain expertise. Reinforcement learning environments are now the center of gravity for that work: the...  ...signals it was trained against. Responsibilities As a Staff Software Engineer, RL Environments, you'll own the technical foundation for how... 
    Suggested
    Full time
    Work experience placement
    Remote work

    Scale AI

    San Francisco, CA
    2 days ago
  •  ...integrations that should go into the final RL run and deciding what can make it in, (2...  ..., and unblocked. You will work across engineering and infrastructure problems as they...  ...system. - Are energized by fast-moving environments where reliability, speed, and judgment matter... 
    Suggested
    Full time

    OpenAI

    San Francisco, CA
    22 hours ago
  • Pareto Inc. is hiring for a role that owns RL environments end-to-end, from scoping with researchers to shipping and maintaining production. You will work between Pareto's engineering team and labs, ensuring specs are buildable and delivered quickly. The candidate should... 
    Relocation

    Pareto Inc.

    San Francisco, CA
    1 day ago
  • $200k - $275k

    Halluminate is seeking a Platform Engineer to own our platform engineering, develop frontier RL environments, and help scale our engineering team from the ground up. The role is based in San Francisco with 5 days in the office, offering a base salary of $200-275k plus... 
    Work at office
    Relocation package

    Halluminate

    San Francisco, CA
    2 days ago
  • David Joseph & Company is seeking a founding Member of Technical Staff for platform engineering in San Francisco. You will own the RL environment infrastructure, build scalable inference pipelines, and help stand up the engineering org and its culture from scratch. We expect... 

    David Joseph & Company

    San Francisco, CA
    12 hours ago
  • Jack & Jill in San Francisco builds RL environments and robust backend services to stress-test frontier AI models. You will collaborate with...  ...Researchers to translate designs into production-grade software and evaluation systems. The role demands hands-on work with React... 

    Jack & Jill

    San Francisco, CA
    4 days ago
  • HeyMilo AI is hiring a Research Engineer to join our Applied AI team in San Francisco. You’ll design and build reinforcement learning environments with verifiable rewards for real-world use cases, creating simulators, reward functions, and evaluation harnesses to measure... 

    HeyMilo AI

    San Francisco, CA
    1 day ago
  •  ...research organization in San Francisco is seeking a Research Engineer focused on pushing the boundaries of frontier models and AGI/ASI...  ...mindset, and the ability to thrive in a fast-paced research environment. This role offers the opportunity to shape innovative AI measurements... 

    OpenAI

    San Francisco, CA
    1 day ago
  • $144k - $198k

     ...join the adventure? Loft’s Onboard Software group provides the on-orbit software that...  ...orbit. The Deployment & Runtime Environments team is where software meets the satellite...  ...for someone scrappy — the kind of engineer who digs into an unfamiliar system, maps... 
    Full time
    Temporary work
    Work at office
    Relocation package
    Flexible hours

    Loft Orbital Solutions

    San Francisco, CA
    13 hours ago
  • YO AI Labs is seeking a Senior Software Engineer to support an AI training project by creating reinforcement learning environments using Model Context Protocol (MCP) tools. You will design reproducible environments, deterministic verification, and reference solutions for... 
    Remote job

    YO AI Labs

    San Francisco, CA
    12 hours ago
  • Simplify in San Francisco is hiring a junior software engineer to design and build reinforcement-learning environments for software engineering tasks. You will own the full lifecycle from concept to evaluation against frontier models, using Python and Docker/Linux tools... 

    Simplify

    San Francisco, CA
    1 day ago
  • Patronus AI, Inc. seeks a Member of Technical Staff - Product Engineer in San Francisco to build high‑quality simulations used to train and evaluate AI agents. You will own complex agent environments end‑to‑end, collaborate with researchers, and improve tooling to support... 

    Patronus AI, Inc.

    San Francisco, CA
    2 days ago
  • YO AI Labs is seeking a Senior Software Engineer to support an AI training project by creating reinforcement learning environments that evaluate AI models on complex software engineering tasks using MCP tools. You will design reproducible environments, deterministic verification... 
    Remote job
    For contractors

    YO AI Labs

    San Francisco, CA
    12 hours ago
  • Fundamental AI Research Institute in San Francisco is seeking researchers experienced in LLM post-training, agentic infrastructure, RL environments, and evaluation harnesses to accelerate flagship AI research. Responsibilities include building and scaling post-training... 

    Storm3

    San Francisco, CA
    3 days ago
  •  ...is building infrastructure to create RL training data and evals for frontier AI...  ...role We’re looking for a Full-Stack Software Engineer, Reinforcement Learning to build the product...  ...product-facing tools for browsing environments, inspecting trajectories, reviewing task... 
    Full time
    Work at office
    Remote work
    Relocation
    Visa sponsorship

    Hud

    San Francisco, CA
    22 hours ago
  • A leading technology company in San Francisco is seeking a Senior Software Engineer to develop core infrastructure for AI systems. The successful candidate will have a strong background in full-stack or infrastructure systems and a degree from a top-tier university. This... 

    Amigos

    San Francisco, CA
    1 day ago
  •  ...What we do Idler builds reinforcement learning environments that teach AI models to code like 0.01% engineers. Our training environments are based on real-world coding...  ...Design and build scaleable systems that generate RL environments Create automated QA systems to... 
    Full time
    Contract work
    Relocation package

    Idler

    San Francisco, CA
    22 hours ago
  •  ...Senior Full Stack Software Engineer Mariana Minerals is looking for an experienced Senior Full...  ...and implement production-ready ML, RL, and LLM-powered features, ensuring robust...  ...record of working in ambiguous, fast-paced environments and bringing structure to complex... 

    Mariana Minerals

    San Francisco, CA
    1 day ago
  •  ...the time to join!). What you'll do (Software Engineer): Build pipelines that score, filter...  ...with founders in a fast-iteration environment, independently owning systems end-to-end...  ...to train in-house reward models used in RL pipelines, for both enterprise clients... 
    Full time
    Temporary work
    Part time
    For contractors
    Work at office
    Weekday work

    Andreessen Horowitz

    San Francisco, CA
    1 day ago
  • Together AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures, and scheduling for...  ...and memory to improve latency and cost. You will also co-design RL and post-training pipelines, drive performance improvements, and... 

    Together

    San Francisco, CA
    2 days ago
  • Preference Model is seeking Research Engineers or Research Scientists to advance self-directed learning in AI. The role involves training and evaluating models within proprietary RL environments and optimizing ML infrastructure. Candidates will benefit from competitive... 

    Preference Model

    San Francisco, CA
    12 hours ago
  • Frontier Arc in San Francisco seeks a Research Engineer to push the core RL research agenda. You’ll own the end-to-end loop—from designing policies...  ...to real hardware. Expect on-site work, a fast-paced lab environment, and the opportunity to tackle difficult #J-18808-Ljbffr... 

    Frontier Arc

    San Francisco, CA
    1 day ago
  •  ...Founding Software Engineer As a Founding Software Engineer, you'll have end-to-end ownership over projects pushing the frontier of AI...  ...This isn't a narrow role. One week you might prototype a new RL environment from a research paper, the next you'll deploy distributed... 
    Work at office
    Visa sponsorship

    RainesDev

    San Francisco, CA
    1 day ago
  • $300k

     ...Mechanize builds reinforcement learning environments that frontier AI labs use to train and...  ...the complex, judgment-heavy parts of software engineering. We build the environments that expose...  .... You'll design, build, and refine RL tasks. Each task is a self-contained... 

    Mechanize

    San Francisco, CA
    22 hours ago
  • $400k

     ...Mechanize builds reinforcement learning environments that frontier AI labs use to train and...  ...the complex, judgment-heavy parts of software engineering. We build the environments that expose...  ...You'll design, build, and refine RL tasks, owning the full lifecycle from... 

    Mechanize, Inc

    San Francisco, CA
    3 days ago
  • $350k

     ...Mechanize builds reinforcement learning environments that frontier AI labs use to train and...  ...the complex, judgment-heavy parts of software engineering. We build the environments that expose...  ...'ll design, build, and quality-assure RL tasks. Each task is a self-contained... 

    Mechanize, Inc

    San Francisco, CA
    3 days ago
  • CoffeeSpace is recruiting a Platform Engineer in San Francisco to build the infrastructure and tooling for frontier RL environments in financial services. You will own significant parts of the ML infrastructure, collaborate with ML engineers and product teams, and help... 
    Work at office
    Visa sponsorship

    CoffeeSpace

    San Francisco, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer - RL Environments. Be the first to apply!