Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Misconception Detection AI Evaluator [Remote]

AuraOne Human Data

Remote
  • Remote job

Misconception Detection AI Evaluator is a remote evaluation track for reviewing misconception detection ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.

Why this role matters

AI data reviewers help turn misconception detection ai evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.

Responsibilities

  • Evaluate misconception detection ai evaluation model outputs against a versioned rubric and assign severity tags for Misconception Detection AI Evaluator assignments.
  • Compare paired responses and pick the stronger answer with a written rationale.
  • Label hallucinations, instruction-following failures, and unsafe content with structured tags.
  • Capture ambiguous prompts and route them back to the program team for rubric updates.
  • Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
  • Document recurring failure modes so the modeling team can target them in the next training run.

Qualifications

  • Prior evaluation, annotation, or human-rater experience on misconception detection ai evaluation or adjacent content for Misconception Detection AI Evaluator work.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that names the issue and the rubric clause being applied.
  • Strong attention to detail and the ability to flag when a prompt itself is the problem.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Compare two misconception detection ai evaluation model responses to the same prompt and pick the stronger one with rationale.
  • Tag an unsafe response with the correct policy category and severity.
  • Audit a 50-row batch for rubric consistency and report drift to the program lead.
  • Propose a rubric clarification after spotting a recurring failure mode.

Nice to have

  • Background in linguistics, content moderation, or trust & safety review.
  • Experience with inter-rater agreement metrics and calibration cycles.
  • Domain expertise that lets you spot subject-matter errors automated checks miss.

Skills

  • Model output evaluation
  • Rubric-based annotation
  • Severity tagging
  • Inter-rater calibration
  • Misconception Detection AI evaluation
  • Learning design
  • Assessment review
  • Pedagogy
  • Misconception
  • Detection

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Misconception Detection AI Evaluator [Remote] in Remote vacancy
  • $20 - $160 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Annotator — Health AI Conversation Quality Evaluation Type: Contract Compensation: $...  ...understandable language for non-experts. Detect sycophancy in AI interactions. Identify... 
    Suggested
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    11 days ago
  • $14.5 per hour

    A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical... 
    Suggested
    Hourly pay
    Part time
    Remote work
    Flexible hours

    Welo Data

    United States
    2 days ago
  •  ...Weekday 1 is seeking expert Evaluators to review AI-generated real estate, hospitality and events outputs for accuracy, rigor and domain quality. You will apply deep expertise to grade documents, spreadsheets and slide decks. Requirements include 5+ years in Real estate... 
    Suggested
    Hourly pay
    Weekly pay
    Contract work
    Work at office
    Remote work
    Weekday work

    Weekday 1

    United States
    4 days ago
  • $11.5 per hour

     ...position as an Online Task Contributor. In this role, you will evaluate and provide feedback on content to enhance search engine results...  ...11.50 hourly, based on task completion, with a supportive community of contributors involved in AI advancements. #J-18808-Ljbffr... 
    Suggested
    Hourly pay
    Part time
    Remote work

    University of Delaware

    United States
    2 days ago
  •  ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers... 
    Suggested
    Contract work
    Temporary work
    Immediate start
    Remote work

    MERIT Beauty

    United States
    2 days ago
  •  ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional... 
    Work at office
    Remote work

    Obsidian

    New York, NY
    4 days ago
  •  ...About the role Swahili Evaluation AI Evaluator is a remote evaluation track for reviewing swahili generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback... 
    Hourly pay
    For contractors
    Remote work
    10 hours per week

    AI Trainer Jobs

    New York, NY
    2 days ago
  •  ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and... 
    Full time
    Contract work
    Remote work
    Flexible hours

    Virtual Vocations Inc

    United States
    1 day ago
  • $20 per hour

    A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and... 
    Remote work
    Flexible hours

    DataAnnotation

    United States
    2 days ago
  •  ...Obsidian is hiring expert Evaluators in Real estate, hospitality, and events to review AI-generated work for accuracy, rigor, and domain quality. This remote position requires deep expertise and involves grading outputs like documents and presentations. Applicants must... 
    Work at office
    Remote work

    Obsidian

    San Francisco, CA
    3 days ago
  • $14.5 per hour

     ...AI Web Search Evaluator Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years of experience in... 
    Bi-weekly pay
    Hourly pay
    Part time
    Immediate start
    Remote work
    Work from home
    Flexible hours

    Welo Data

    United States
    5 days ago
  •  ...Alignerr is seeking Creative Writing Evaluators to assess AI-generated stories, essays, poetry, and other writings to shape how AI learns compelling prose. This fully remote, flexible contract roles welcomes avid readers and writers with no publishing credits required... 
    Contract work
    Remote work
    Flexible hours

    Alignerr Corp.

    United States
    4 days ago
  • $14.5 per hour

     ...AI Web Search Evaluator As a Web Search Evaluator, you will play a key role in improving the quality of search engine results, ensuring users find the most relevant and useful information. Your work will directly impact the development of AI algorithms, making search... 
    Hourly pay
    Part time
    Currently hiring
    Immediate start
    Remote work
    Work from home
    10 hours per week
    Flexible hours

    Welo Data

    United States
    5 days ago
  •  ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise... 
    Hourly pay
    Work at office
    Remote work

    Mercor Inc

    New York, NY
    3 days ago
  • Obsidian is seeking expert Evaluators in Humanities / arts / culture to review AI-generated work products for accuracy, rigor, and domain quality. You will apply subject-matter expertise to grade outputs and provide structured feedback. This is a remote, hourly engagement... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    4 days ago
  • AI Trainer Jobs is seeking licensed clinicians to remotely evaluate AI outputs in internal medicine clinical reviews. You will assess differential diagnoses, dosing logic, and guideline adherence, and document corrected reasoning for model training. Ideal candidates hold... 
    Remote job
    Hourly pay
    10 hours per week

    AI Trainer Jobs

    New York, NY
    5 days ago
  • $18 per hour

    We are looking for AI Linguistic Evaluators to assess how effectively an AI application performs in Malayalam for a short collaboration. Job Type: Freelance Location: Remote from the United States Work Schedule: Flexible Start Date: Immediately Duration: Approximately... 
    Freelance
    Immediate start
    Remote work
    Flexible hours

    Lever, Inc.

    Dallas, TX
    2 days ago
  • $20 - $30 per hour

    A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong... 
    Remote job
    Hourly pay

    Crossing Hurdles

    New York, NY
    3 days ago
  • Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates will leverage their subject matter expertise to assess various documents, spreadsheets, and slide decks.... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    5 days ago
  • $150 per hour

    The Work You will help improve AI systems by creating expert-level quantitative finance training data and reviewing AI-generated answers. Your feedback will focus on whether responses are accurate, relevant, and consistent with accepted quantitative finance methods. The... 
    Hourly pay
    Ongoing contract
    Contract work
    Part time
    Remote work
    Flexible hours

    OpenTrain AI, Inc.

    Eastern, KY
    2 days ago
  • $20 per hour

    A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI...  ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with... 
    Hourly pay
    Remote work
    Flexible hours

    SupportFinity

    United States
    2 days ago
  • Prolific is seeking fluent Norwegian speakers to act as evaluators for AI training, performing side-by-side assessments of text and voice snippets to judge naturalness and authenticity. You will listen to audio clips and rate how naturally the AI speaks, providing detailed... 
    Remote job
    Work from home
    Flexible hours

    Prolific Academic Ltd

    New York, NY
    3 days ago
  • AI Trainer Jobs seeks a remote independent contractor to review AI outputs for fintech operations evaluation. You will assess workflow adherence, tone, and escalation logic, assigning severity tags and documenting next steps for model training. The role requires experience... 
    Remote job
    Hourly pay
    For contractors
    Flexible hours

    AI Trainer Jobs

    New York, NY
    5 days ago
  • Archangel Health AI is seeking Clinical AI Evaluators to review AI-generated clinical outputs, benchmark diagnostic reasoning, and refine responses to real-world medical queries. You will perform clinical accuracy auditing, RLHF ranking, error and harm identification,... 
    Remote work
    Flexible hours
    Shift work

    Modern MedEd LLC

    New York, NY
    4 days ago
  • $18 per hour

    rws is seeking AI Linguistic Evaluators to assess Marathi performance of an AI application for a short collaboration. This freelance role offers flexible hours with a total of about 10-15 hours (roughly 5 hours daily) and a rate of 18 USD per hour. Candidates should be... 
    Remote job
    Hourly pay
    Daily paid
    Freelance
    Immediate start
    Flexible hours

    RWS

    New York, NY
    2 days ago
  • Rex.zone is seeking a Senior AI data annotator to perform data labeling and evaluation for NLP tasks, RLHF assessments, and prompt QA to improve training data quality and model performance. This is a US-based remote, full-time role aligned with Miami talent demand. You... 
    Remote job
    Full time

    Rex.zone

    Miami, FL
    3 days ago
  • Mercor is seeking expert Evaluators in Clinical / biomedical / pharma to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. This is a remote, hourly engagement. You will apply deep subject-matter... 
    Remote job
    Hourly pay

    Mercor

    New York, NY
    3 days ago
  • Alignerr is seeking a Population Health Informaticist for an AI training project. You will evaluate AI outputs on health data, assess population metrics and disparities, and provide structured feedback to improve AI reasoning in health contexts. The role is a fully remote... 
    Remote job
    Hourly pay
    Contract work
    10 hours per week
    Flexible hours

    Alignerr Corp.

    Dallas, TX
    1 day ago
  • $30 per hour

     ...Location: Remote Commitment: 10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous agent systems...  ...stakeholders. Requirements Strong experience in LLM evaluation, AI output analysis, QA/testing, UX research, or similar analytical... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    3 days ago
  • OpenTrain AI, Inc. seeks a senior dermatology reviewer to provide expert clinical judgment on complex dermatology cases and evaluate longitudinal data for AI systems. You will document commentary according to guidelines and participate in consensus processes, with occasional... 
    Remote job
    Part time
    For contractors

    OpenTrain AI, Inc.

    Brooklyn, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Misconception Detection AI Evaluator [Remote]. Be the first to apply!