Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Engineer - Environments, Data and Post-Training

$250k
Full-time

Mercor

About Mercor

Mercor's mission is to organize human intelligence to power the AI economy. We're a leading AI data company, building the layer between human expertise and frontier models. Millions of domain experts on the platform are paid over $4 million per day to train frontier AI models. Mercor's APEX benchmark family measures AI's real-world impact on professional work. Mercor Enterprise brings this same infrastructure to Fortune 500 companies: helping companies capture how their best people actually work, translating that expertise directly back into agents.

Mercor is creating a new category of work where expertise powers AI advancement. Achieving this requires an ambitious, fast-paced and deeply committed team. You’ll work alongside researchers, operators, and AI companies at the forefront of shaping the systems that are redefining society. Mercor is a profitable Series C company valued at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices.

About the Role

As a Research Scientist at Mercor, you will work at the intersection of research and engineering on frontier post-training. You will develop new training and evaluation methods, test them through rigorous experiments, and implement successful approaches at scale.

Working with researchers, engineers, domain experts, and customers, you will investigate how data, rewards, environments, and optimization methods shape model behavior. Your work will influence frontier models, Mercor’s products, and the broader research community through releasing blogposts, technical reports, and papers.

What You’ll Do

  • Implement novel post-training methods that improve model reasoning, tool use, and agentic behavior.

  • Develop new training recipes for frontier open models.

  • Design and run experiments across datasets, reward functions, environments, and optimization strategies, including methods such as GRPO and DAPO.

  • Build reinforcement learning with verifiable rewards (RLVR) and other post-training pipelines at scale.

  • Investigate model capabilities and failure modes, then develop targeted training interventions.

  • Create methods for measuring data quality, usability, and causal impact on model performance.

  • Build scalable pipelines for data generation, filtering, augmentation, and selection.

  • Develop rubrics, evaluators, benchmarks, and scoring systems that inform training decisions.

  • Translate open-ended research questions into rigorous experiments and production systems.

  • Collaborate with researchers, applied AI teams, engineers, and domain experts producing training data.

  • Contribute to open-source post-training tools and research.

What We’re Looking For

  • Demonstrated experience training and evaluating machine learning models.

  • A strong research record in post-training, reinforcement learning, language-model evaluation, data-centric ML, or a closely related field.

  • Ability to reason rigorously about model behavior, experimental results, and data quality.

  • Strong programming skills and experience implementing machine learning systems.

  • Knowledge of the current AI research landscape and important open problems.

  • Excitement to work in person in San Francisco, five days a week (with optional remote Saturdays), and thrive in a high-intensity, high-ownership environment.

Nice To Have

  • Experience on an industry post-training or frontier-model team.

  • Main authorship of publications at top-tier conferences (NeurIPS, ICML, ACL).

  • Experience with synthetic-data generation

  • Experience building large-scale evaluation or data-generation infrastructure.

  • Solid foundations in distributed or backend systems, and experimental design.

  • Familiarity with APIs, databases, and cloud infrastructure.

Benefits

  • Bi-annual performance bonus structure

  • Generous equity grant vested over 4 years

  • Up to $15k Relocation bonus

  • $10K housing bonus (if you live within 0.5 miles of our office)

  • $1.5K monthly stipend for meals

  • Free Equinox membership

  • $200 monthly laundry reimbursement

  • $200 monthly personal wellness reimbursement

  • Health, Dental, Vision insurance

Vacancy posted 6 days ago
Similar jobs that could be interesting for youBased on the Research Engineer - Environments, Data and Post-Training in San Francisco, CA vacancy
  • $150k - $180k

     ...evaluation, synthetic data, and reinforcement...  ...and improve training data quality for frontier...  ...shaping internal research culture around...  ...of tasks across RL environments, synthetic data...  ...Partner with research engineers, domain experts,...  ...understanding of AI evals and post-training, going... 
    Training
    Full time
    Relocation
    Visa sponsorship

    Clera

    San Francisco, CA
    2 days ago
  •  ...Job Description We are looking for a hybrid Systems Engineer and AI Researcher to lead the development of our agent evaluation framework and post-training data pipelines. You will design sandboxed execution environments, high-throughput reinforcement learning feedback... 
    Training
    Work at office

    Hyphen Connect

    San Francisco, CA
    13 days ago
  •  ...Primarily On-site We are seeking an Research Engineer – RL Infrastructure & Agent Environments to build the environments,...  ...infrastructure used to train and assess long-horizon enterprise...  ...behind realistic agent environments, post-training systems, and reliable evaluation... 
    Training

    MaxIT Consulting - Max Corporate Group

    San Francisco, CA
    7 days ago
  • $197.3k - $313.7k

     ...foundation that makes agentic engineering on Salesforce Database...  ...evaluation with data curation.Your work...  ...Java in a Unix/Linux environment; solid grasp of distributed...  ...and opt out options.Posting StatementSalesforce is...  ...promotion, benefits, training, assessment of job performance... 
    Training
    Full time
    Local area

    Salesforce

    San Francisco, CA
    1 day ago
  • $150k - $250k

     ...global social organizations.We research and deploy technologies that...  ...ForAt Distyl, Research Engineers build the bridge between frontier...  ...that work inside real customer environments.Research Engineers operate at...  ...and build data systems that power reliable AI... 
    Suggested
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    1 day ago
  •  ...knowledge, and skills and can provide valuable training and experience that translates directly...  ...to a major, as listed. Environmental Engineer: You must have one of the following:...  ...accommodation is any change in the work environment or in the way things are customarily done... 
    Training
    Full time
    Part time
    Work experience placement
    Work at office
    Local area
    Relocation

    Environmental Protection Agency

    San Francisco, CA
    2 days ago
  • $150k - $250k

     ...social organizations.We research and deploy...  ...ForAt Distyl, Research Engineers build the bridge between...  ...inside real customer environments.Research Engineers operate...  ...and run post-training workflows that improve...  ...modeling, synthetic data, evals, or related post... 
    Training
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    1 day ago
  • $164.6k - $313.3k

     ...SODA) is looking for a driven Data/ML engineer to push the boundaries of...  ...collaborative and efficient research team looking for highly motivated...  ...a data lead to own the training data behind our generative audio...  ...(as listed on the job posting), the application window will... 
    Training
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    9 days ago
  • $197.3k - $313.7k

     ...software and platform engineers to embed in our AI...  ...you thrive in an environment where your...  ...enable world-class research and products used...  ...use your personal data and your rights, including...  ...opt out options.Posting...  ...promotion, benefits, training, assessment of job... 
    Training
    Full time

    Salesforce

    San Francisco, CA
    5 days ago
  •  ...OpenAI is seeking an exceptional researcher to build frontier health capabilities and translate them into measurable impact at scale...  ...hands‑on builder, and possess a track record in advancing pretraining, RL/post‑training, or health‑focused AI problems. #J-18808-Ljbffr... 
    Training

    OpenAI

    San Francisco, CA
    2 days ago
  •  ...Mistral is seeking a Research Engineer – ML track to build and optimise large-scale learning systems powering open-weight models. You will...  ...hands-on with PyTorch, JAX or TensorFlow, plus distributed training tools. You thrive with collaboration and rigorous software practices... 
    Training

    Mistral

    San Francisco, CA
    18 hours ago
  •  ...hands. We build high-fidelity research infrastructure that evaluates...  ...About the role As a Research Engineer, you will be responsible for post-training models for adversarial...  ...entire training pipeline from data generation and environment design to evaluations. You will... 
    Training

    General Analysis

    San Francisco, CA
    1 day ago
  •  ...Research Engineer Lotus Health is a groundbreaking primary care app that integrates your medical...  ...Lotus. You will turn messy health data into accurate, cited, and actionable guidance...  ...engineering: dataset curation, model training and evaluation, retrieval and tool use,... 
    Training

    Lotus Health

    San Francisco, CA
    5 days ago
  •  ...Archive Human Archive is a research lab backed by Y Combinator...  ...publish research. Today, our data is used for robotics and world...  ...As a Research Engineer, you'll work on multimodal sensing...  ...and used for downstream VLA training. You'll research emerging... 
    Training
    Shift work

    Human Archive

    San Francisco, CA
    1 day ago
  •  ...Phonic is a product and research lab focused on...  ...Role As a Research Engineer at Phonic, you'll sit...  ...human. You'll design, train, and iterate on models...  ...high trust, in-person environment in our SF office to push...  ...experiments end-to-end: data processing, training,... 
    Training
    Work at office

    Phonic

    San Francisco, CA
    4 days ago
  • $175k - $250k

     ...Research Engineer About Scorecard We’re a small, nimble team backed...  ...simusers to take action in the environment. Convert transcripts into...  ...; your simusers run inside training environments where their...  ...evaluations with baselines, held-out data, and calibrated judges, and... 
    Training
    Work at office

    Kindredventures

    San Francisco, CA
    2 days ago
  •  ...Pantograph is training general models that start by watching internet-scale video and...  ...durable robots. We're looking for a research engineer to help us train increasingly capable models...  ...learning, reinforcement learning, data processing, evaluation, and the infrastructure... 
    Training

    Pantograph

    San Francisco, CA
    2 days ago
  • $275k

     ...Research Engineer Magic's mission is to build safe AGI that accelerates...  ...combines frontier-scale pre-training, domain-specific RL, ultra-long...  ...large GPU clusters Curate post-training datasets to improve...  ...Build out internet-scale data pipelines and crawlers Design... 
    Training
    Relocation
    Visa sponsorship

    Magic Inc

    San Francisco, CA
    5 days ago
  •  ...Research Engineer We believe that software is the foundation of modern...  ...software vulnerabilities. We are training and scaling security AI...  ...includes deep expertise in data, infrastructure, security and...  ...and reinforcement learning environments to train security coding agents... 
    Training
    Full time
    Work at office

    DepthFirst

    San Francisco, CA
    5 days ago
  •  ...Research Engineer Sesame believes in a future where computers are lifelike - with the ability...  ...models honest in production. Harness the data — create tooling for safe, versioned,...  ...partner with research and infra to prototype, train, and deploy state-of-the-art voice... 
    Training
    Full time
    Contract work
    Flexible hours
    Shift work

    SESAME

    San Francisco, CA
    4 days ago
  • $100k - $300k

     ...failing. We believe massive scale through data-driven machine learning is the key to...  .... Position Overview We are hiring Research Engineers to develop scalable robotic systems aimed...  ...Develop and implement new algorithms for training and optimizing general-purpose robot foundation... 
    Training
    Full time

    Skild AI

    San Francisco, CA
    2 days ago
  • $180k - $340k

     ...Research Engineer As a Research Engineer at Gamma, you'll build models for visual communication...  ...the opportunity to build evals and training data for a field where there isn't much of...  ...systematically evaluating them ~ Experience with post-training techniques including... 
    Training
    Full time
    Work at office
    Work from home

    Gamma

    San Francisco, CA
    4 days ago
  • $180k - $250k

     ...Open role Research Engineer San Francisco (On-site), Full-time About...  ...Engineer to sit between our data and our models and make the...  ...orchestrate and optimize training runs on long-horizon multimodal...  ...config, dataset version and environment recoverable from any result.... 
    Training
    Full time
    Work at office
    Relocation package

    Breakout Ventures

    San Francisco, CA
    1 day ago
  • $200k - $400k

     ...Research Engineer Decagon is the leading conversational AI...  ...frontier approaches for training, evaluation, and...  ...~ Prior experience post-training and deploying LLMs in production environments. ~ Fluency in Python...  ...training, evaluation, data pipelines) ~ Track record... 
    Training
    Full time
    Work at office
    Local area

    Decagon

    San Francisco, CA
    3 days ago
  •  ...become a massive tax of engineering velocity. Resolve AI...  ...end-to-end, balancing research and engineering to create...  ...Build and optimize data pipelines to process high...  ...unstructured data for training and evaluation...  ...collaborative, high-trust environment. ~ Accelerate Your Career... 
    Training
    Work at office
    Visa sponsorship
    Flexible hours

    Resolve AI

    San Francisco, CA
    2 days ago
  • $162k - $243.1k

     ...naturally committed to the environment.Help Shape the Future of Coastal...  ...planning, coastal engineering, and ecosystem restoration,...  ...: YesSchedule: Full timeJob Posting: 15/07/2026 06:07:00Req ID:...  ...compensation, fringe benefits, job training, terminations or any other condition... 
    Training
    Full time
    Temporary work
    Part time
    Casual work
    Local area
    Worldwide
    Flexible hours

    Stantec

    San Francisco, CA
    3 days ago
  • $200k - $350k

     ...at the intersection of research, product, and...  ...Machine Learning Research Engineer , you’ll own end-to-end...  ...running, and analyzing post‑training experiments that help...  ...iterative experiments from data to evaluation...  ...in a small, in‑person environment Bonus Experience working... 
    Training

    Coders Connect

    San Francisco, CA
    2 days ago
  •  ...London, and Amsterdam.Making data-driven decisions is key to Plaid...  ...and tooling to teams across engineering, product, and business and...  ...pay range shown on each job posting is the minimum and maximum target...  ...like skills, experience, and relevant education or training.
    Training
    Full time
    Work experience placement
    Work at office
    Local area
    Flexible hours

    Plaid Financial

    San Francisco, CA
    2 days ago
  • $105.4k - $207.8k

     ...practice. I&DT brings an engineering- and innovation-led...  ...technology-enabled solutions. Data Studio is Converge for...  ...and delivery teams post-sale to prove and...  ...capabilities in each client's environment, powering provider and...  .... Support client training, onboarding, and... 
    Training
    Local area

    Deloitte

    San Francisco, CA
    5 days ago
  • $88.97k - $287.91k

     ...is hiring Forward Deployed Engineers focused on Data 360, Salesforce’s real-...  ...inside enterprise customer environments. You will turn approved solution...  ...tools and opt out options.Posting StatementSalesforce is an...  ..., promotion, benefits, training, assessment of job... 
    Training
    Full time

    Salesforce

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Engineer - Environments, Data and Post-Training. Be the first to apply!