Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

STEM Explanation AI Evaluator [Remote]

AuraOne Human Data

Remote
  • Remote job

STEM Explanation AI Evaluator is a remote review track for evaluating AI outputs across stem research reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method so the modeling team can train on it.

Why this role matters

STEM research models live or die on whether their derivations actually hold up under scrutiny. AuraOne uses scientific specialists to grade outputs the way a peer reviewer would — checking assumptions, reproducing key steps, and capturing the right method alongside the wrong one.

Responsibilities

  • Review AI outputs against current stem research methods, conventions, and prior work for STEM Explanation AI Evaluator assignments.
  • Reproduce or sanity-check key derivations, calculations, or experimental claims.
  • Flag dimensional, methodological, and citation errors with structured severity tags.
  • Capture the corrected reasoning or worked example so the modeling team can train on it.
  • Adjudicate disputed answers against textbooks, papers, or community standards.
  • Maintain reviewer-quality scores in inter-rater calibration cycles.

Qualifications

  • Graduate-level training or equivalent applied experience in stem research or a closely related field for STEM Explanation AI Evaluator work.
  • Hands-on experience publishing, teaching, or advising on the topic at a professional level.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that cites methods, papers, or worked examples.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Reproduce a stem research derivation from a model output and flag any algebraic or dimensional errors.
  • Grade a model's literature summary against the cited papers and rate the citation quality.
  • Adjudicate a disputed answer between two reviewers using textbook methods.
  • Audit a 25-row batch for rubric consistency and report drift to the program lead.

Nice to have

  • PhD, postdoc, or industry research experience in the topic area.
  • Prior work reviewing AI-assisted research tooling and its failure modes.
  • Multilingual fluency for non-English papers and corpora.

Skills

  • Scientific reasoning
  • Method validation
  • Citation review
  • Quantitative analysis
  • STEM research
  • Learning design
  • Assessment review
  • Pedagogy
  • STEM
  • Explanation

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the STEM Explanation AI Evaluator [Remote] in Remote vacancy
  • A leading AI research accelerator is hiring a position focused on contributing to projects that evaluate and enhance AI systems. You will design community service scenarios, write structured explanations, and evaluate AI accuracy. The ideal candidate will have 4+ years... 
    Suggested
    Remote job
    Full time
    For contractors

    Turing

    New York, NY
    2 days ago
  • $40 - $100 per hour

    About OpenTrain OpenTrain AI is the hiring and contracting organization for this opportunity...  ...work About AI Training and Scientific Evaluation AI training is the human side of...  ...when models handle calculations, technical explanations, uncertainty, and limits of inference.... 
    Suggested
    Hourly pay
    Contract work
    Part time
    For contractors
    Remote work

    OpenTrain AI

    Brooklyn, NY
    21 hours ago
  •  ...remote, hourly contractor role supporting AI data and language projects on a project-...  ...to support AI training datasets. LLM evaluation: reviewing AI-generated responses for...  ...gaps, methodological errors, and unclear explanations even when language is fluent.... 
    Suggested
    Hourly pay
    For contractors
    Remote work
    Flexible hours

    CNTXT AI

    Brooklyn, NY
    10 days ago
  •  ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional... 
    Suggested
    Work at office
    Remote work

    Obsidian

    New York, NY
    3 days ago
  • $20 per hour

    A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and... 
    Suggested
    Remote work
    Flexible hours

    DataAnnotation

    United States
    1 day ago
  • $20 - $30 per hour

    A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong... 
    Remote job
    Hourly pay

    Crossing Hurdles

    New York, NY
    2 days ago
  • YO AI Labs is seeking an Investment & Finance Expert to work remotely as a contractor. The role focuses on evaluating AI outputs related to valuation, financial modeling, markets, and investments, and crafting expert prompts plus reference answers to improve model reasoning... 
    Remote job
    For contractors

    YO AI Labs

    Austin, TX
    1 day ago
  • MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers... 
    Remote job
    Contract work
    Temporary work
    Immediate start

    MERIT Beauty

    New York, NY
    21 hours ago
  • Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates will leverage their subject matter expertise to assess various documents, spreadsheets, and slide decks.... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    4 days ago
  • $50 per hour

    Prolific seeks Product Designers and UX Specialists to help train AI models using your expertise. You'll evaluate AI-generated designs and ensure usability and accessibility while working from home. Ideal candidates hold a BS, MS, or PhD in a relevant field and have at... 
    Remote job
    Work from home
    Flexible hours

    Prolific

    Sacramento, CA
    3 days ago
  • $24 per hour

    Prolific is seeking an AI Trainer with advanced Tamil fluency to evaluate AI models' understanding of the Tamil language's emotional and cultural nuances. Responsibilities include assessing audio clips, reviewing tones for cultural relevancy, and ensuring quality control... 
    Remote job
    Flexible hours

    Prolific

    New York, NY
    4 days ago
  • A talent marketplace is seeking soccer experts to evaluate live soccer games. The role involves assessing AI-generated commentary, scoring performance, and providing feedback. Qualified candidates will have deep expertise in soccer, strong analytical and communication... 
    Contract work

    Mercor

    New York, NY
    2 days ago
  • $14.5 per hour

     ...ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years...  ...performance and provide insights on relevance and quality. Evaluate and rate the effectiveness of search engine results to ensure... 
    Hourly pay
    Part time
    Immediate start
    Remote work
    Work from home
    10 hours per week
    Flexible hours

    Kanz.us

    Dallas, TX
    21 hours ago
  • Dorado is seeking expert Evaluators in program management / implementation planning to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote,... 
    Remote job
    Hourly pay
    Work at office

    Dorado

    New York, NY
    2 days ago
  • $14.5 per hour

    A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical... 
    Remote job
    Hourly pay
    Part time
    Flexible hours

    Welo Data

    New York, NY
    4 days ago
  • Feedinkoo is seeking a Graphic Designer to help train and improve AI models focused on UI/UX design, visuals, and user experiences. You will critique AI outputs, provide feedback, and guide model improvements to better reflect aesthetics and usability for designers. The... 
    Remote work

    Feedinkoo

    New York, NY
    2 days ago
  • Obsidian is seeking expert Evaluators in Humanities / arts / culture to review AI-generated work products for accuracy, rigor, and domain quality. You will apply subject-matter expertise to grade outputs and provide structured feedback. This is a remote, hourly engagement... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    3 days ago
  • Mercor is seeking experts in Spreadsheet QA and workbook maintenance to review AI-generated documents, spreadsheets, and slide decks for accuracy and quality. This is a remote, hourly engagement. Ideal candidates have 5+ years in Spreadsheet QA, fluent English, and strong... 
    Remote job
    Hourly pay
    Work at office

    Mercor

    Miami, FL
    1 day ago
  • Obsidian is hiring expert Evaluators in Healthcare operations for a remote, hourly engagement. You will review AI-generated work products for accuracy, rigor, and domain quality, leveraging your extensive subject-matter expertise in the field. The ideal candidate has over... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    San Francisco, CA
    2 days ago
  • Mercor is seeking expert Evaluators in Compliance / regulatory response with financial-services AI to review AI-generated outputs (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. This remote, hourly engagement leverages your subject-matter... 
    Remote job
    Hourly pay

    Mercor

    New York, NY
    10 hours ago
  •  ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise... 
    Hourly pay
    Work at office
    Remote work

    Mercor Inc

    New York, NY
    2 days ago
  • $30 per hour

     ...Location: Remote Commitment: 10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous agent systems...  ...stakeholders. Requirements Strong experience in LLM evaluation, AI output analysis, QA/testing, UX research, or similar analytical... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    2 days ago
  • $50 - $60 per hour

     ...OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI is recruiting...  ...intelligence. Contributors review examples, test model behavior, evaluate responses, and identify errors so AI systems become more accurate... 
    Hourly pay
    Part time
    For contractors
    Remote work
    Worldwide
    Flexible hours

    OpenTrain AI

    Brooklyn, NY
    1 day ago
  • Mercor is seeking expert Evaluators in Clinical / biomedical / pharma to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. This is a remote, hourly engagement. You will apply deep subject-matter... 
    Remote job
    Hourly pay

    Mercor

    New York, NY
    2 days ago
  • BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance... 
    Remote job

    BAM Ventures

    New York, NY
    4 days ago
  • $140 per hour

     ...matter expertise to improve how next-generation AI systems learn, reason, and perform. As an AI Domain Expert you will evaluate AI outputs, create challenging prompts,...  ...professional field, such as finance, healthcare, STEM engineering, software, law, or similar. Strong... 
    Hourly pay
    Part time
    For contractors
    Remote work
    Visa sponsorship
    Work visa
    Free visa

    SaidGig

    United States
    8 days ago
  • YO AI Labs is seeking Turkish bilingual experts for a contract-based remote role focusing on language and AI training. You will evaluate Turkish audio samples, assess quality, and provide precise feedback to help improve AI systems. No prior AI experience is required, but... 
    Remote job
    Contract work

    YO AI Labs

    Seattle, WA
    1 day ago
  • Mercor is hiring experienced musicians to evaluate generative music AI models, focusing on Urdu and English lyrics. You will assess AI-generated music across genres, rating quality, creativity, and originality against detailed standards. Start date is immediate, duration... 
    Immediate start
    Remote work
    Flexible hours

    Mercor

    New York, NY
    2 days ago
  •  ...Prolific is seeking Product Designers and UX Specialists to join our Expert Network, contributing to the training and evaluation of cutting-edge AI models. This role requires expertise in usability, design systems, and user research, providing essential feedback on design... 
    Remote work
    Work from home
    Flexible hours

    Prolific

    Tucson, AZ
    3 days ago
  • YO AI Labs is seeking a PhD and academic expert to support AI research projects remotely. You will apply subject-matter expertise to evaluate and improve AI model responses across technical and humanities disciplines. You will design expert prompts, reference answers,... 
    Remote job

    YO AI Labs

    Miami, FL
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to STEM Explanation AI Evaluator [Remote]. Be the first to apply!