Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Socratic Dialogue AI Evaluator [Remote]

AuraOne Human Data

Remote
  • Remote job

Socratic Dialogue AI Evaluator is a remote evaluation track for reviewing socratic dialogue ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.

Why this role matters

AI data reviewers help turn socratic dialogue ai evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.

Responsibilities

  • Evaluate socratic dialogue ai evaluation model outputs against a versioned rubric and assign severity tags for Socratic Dialogue AI Evaluator assignments.
  • Compare paired responses and pick the stronger answer with a written rationale.
  • Label hallucinations, instruction-following failures, and unsafe content with structured tags.
  • Capture ambiguous prompts and route them back to the program team for rubric updates.
  • Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
  • Document recurring failure modes so the modeling team can target them in the next training run.

Qualifications

  • Prior evaluation, annotation, or human-rater experience on socratic dialogue ai evaluation or adjacent content for Socratic Dialogue AI Evaluator work.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that names the issue and the rubric clause being applied.
  • Strong attention to detail and the ability to flag when a prompt itself is the problem.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Compare two socratic dialogue ai evaluation model responses to the same prompt and pick the stronger one with rationale.
  • Tag an unsafe response with the correct policy category and severity.
  • Audit a 50-row batch for rubric consistency and report drift to the program lead.
  • Propose a rubric clarification after spotting a recurring failure mode.

Nice to have

  • Background in linguistics, content moderation, or trust & safety review.
  • Experience with inter-rater agreement metrics and calibration cycles.
  • Domain expertise that lets you spot subject-matter errors automated checks miss.

Skills

  • Model output evaluation
  • Rubric-based annotation
  • Severity tagging
  • Inter-rater calibration
  • Socratic Dialogue AI evaluation
  • Learning design
  • Assessment review
  • Pedagogy
  • Socratic
  • Dialogue

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 8 hours ago
Similar jobs that could be interesting for youBased on the Socratic Dialogue AI Evaluator [Remote] in Remote vacancy
  • $20 per hour

    A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and... 
    Suggested
    Remote work
    Flexible hours

    DataAnnotation

    United States
    5 days ago
  •  ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and... 
    Suggested
    Full time
    Contract work
    Remote work
    Flexible hours

    Virtual Vocations Inc

    United States
    4 days ago
  •  ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise... 
    Suggested
    Hourly pay
    Work at office
    Remote work

    Mercor Inc

    New York, NY
    1 day ago
  • Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates will leverage their subject matter expertise to assess various documents, spreadsheets, and slide decks.... 
    Suggested
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    3 days ago
  • YO AI Labs is seeking an Investment & Finance Expert to work remotely as a contractor. The role focuses on evaluating AI outputs related to valuation, financial modeling, markets, and investments, and crafting expert prompts plus reference answers to improve model reasoning... 
    Suggested
    Remote job
    For contractors

    YO AI Labs

    Austin, TX
    5 days ago
  • $14.5 per hour

     ...ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years...  ...performance and provide insights on relevance and quality.Evaluate and rate the effectiveness of search engine results to ensure... 
    Part time
    Remote work
    Work from home
    10 hours per week
    Flexible hours

    Kanz

    Dallas, TX
    4 days ago
  • MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers... 
    Remote job
    Contract work
    Temporary work
    Immediate start

    MERIT Beauty

    New York, NY
    4 days ago
  • $20 - $30 per hour

    A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong... 
    Remote job
    Hourly pay

    Crossing Hurdles

    New York, NY
    6 days ago
  • $30 per hour

     ...Location: Remote Commitment: 10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous agent systems...  ...stakeholders. Requirements Strong experience in LLM evaluation, AI output analysis, QA/testing, UX research, or similar analytical... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    6 days ago
  • Mercor is seeking expert Evaluators in Clinical / biomedical / pharma to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. This is a remote, hourly engagement. You will apply deep subject-matter... 
    Remote job
    Hourly pay

    Mercor

    New York, NY
    6 days ago
  • BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance... 
    Remote job

    BAM Ventures

    New York, NY
    3 days ago
  • Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional... 
    Remote job
    Work at office

    Obsidian

    New York, NY
    3 days ago
  • A leading AI research accelerator is hiring a position focused on contributing to projects that evaluate and enhance AI systems. You will design community service scenarios, write structured explanations, and evaluate AI accuracy. The ideal candidate will have 4+ years... 
    Remote job
    Full time
    For contractors

    Turing

    New York, NY
    6 days ago
  • $14.5 per hour

    Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages... 
    Hourly pay
    Part time
    Immediate start
    Remote work
    Work from home
    10 hours per week
    Flexible hours

    Welo Data

    New York, NY
    3 days ago
  • $14.5 per hour

    A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical... 
    Remote job
    Hourly pay
    Part time
    Flexible hours

    Welo Data

    New York, NY
    3 days ago
  • Obsidian is hiring expert Evaluators in Healthcare operations for a remote, hourly engagement. You will review AI-generated work products for accuracy, rigor, and domain quality, leveraging your extensive subject-matter expertise in the field. The ideal candidate has over... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    San Francisco, CA
    6 days ago
  • $24 per hour

    Prolific is seeking an AI Trainer with advanced Tamil fluency to evaluate AI models' understanding of the Tamil language's emotional and cultural nuances. Responsibilities include assessing audio clips, reviewing tones for cultural relevancy, and ensuring quality control... 
    Remote job
    Flexible hours

    Prolific

    New York, NY
    3 days ago
  • $50 per hour

    Prolific seeks Product Designers and UX Specialists to help train AI models using your expertise. You'll evaluate AI-generated designs and ensure usability and accessibility while working from home. Ideal candidates hold a BS, MS, or PhD in a relevant field and have at... 
    Remote job
    Work from home
    Flexible hours

    Prolific

    Sacramento, CA
    2 days ago
  • YO AI Labs is seeking a Healthcare Expert for a remote contract to support AI training and evaluation projects. You will apply clinical knowledge to assess AI outputs, workflows, and real-world use of healthcare systems. Key tasks include evaluating EHR applications, reviewing... 
    Remote job
    Contract work

    YO AI Labs

    San Jose, CA
    6 days ago
  • A talent marketplace is seeking soccer experts to evaluate live soccer games. The role involves assessing AI-generated commentary, scoring performance, and providing feedback. Qualified candidates will have deep expertise in soccer, strong analytical and communication... 
    Contract work

    Mercor

    New York, NY
    6 days ago
  • Dorado is seeking expert Evaluators in program management / implementation planning to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote,... 
    Remote job
    Hourly pay
    Work at office

    Dorado

    New York, NY
    6 days ago
  • Obsidian is seeking expert Evaluators in Humanities / arts / culture to review AI-generated work products for accuracy, rigor, and domain quality. You will apply subject-matter expertise to grade outputs and provide structured feedback. This is a remote, hourly engagement... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    2 days ago
  • Feedinkoo is seeking a Graphic Designer to help train and improve AI models focused on UI/UX design, visuals, and user experiences. You will critique AI outputs, provide feedback, and guide model improvements to better reflect aesthetics and usability for designers. The... 
    Remote work

    Feedinkoo

    New York, NY
    6 days ago
  •  ...Supporting diverse AI data and language projects, the hourly contractor AI Trainer and Evaluator will work remotely to generate content, annotate data, and evaluate AI responses for accuracy and cultural relevance. Key responsibilities Generate high-quality prompts and... 
    Hourly pay
    For contractors
    Remote work

    Virtual Vocations Inc

    United States
    1 day ago
  • Mercor is hiring experienced musicians to evaluate generative music AI models, focusing on Urdu and English lyrics. You will assess AI-generated music across genres, rating quality, creativity, and originality against detailed standards. Start date is immediate, duration... 
    Immediate start
    Remote work
    Flexible hours

    Mercor

    New York, NY
    6 days ago
  • YO AI Labs seeks experienced Specialized Industries Experts to evaluate AI-powered workflows across Government Contracting, Grants, Nonprofit operations, and EdTech. You will test AI agents in real-world scenarios and provide practical feedback to improve next-generation... 
    Remote job

    YO AI Labs

    Los Angeles, CA
    6 days ago
  • Mercor is hiring expert Evaluators in Healthcare operations to review AI-generated outputs (documents, spreadsheets, and slide decks) for accuracy and domain quality. This is a remote, hourly engagement. You will apply deep subject-matter expertise to grade outputs, identify... 
    Remote job
    Hourly pay

    Mercor

    New York, NY
    2 days ago
  • Mercor is seeking expert Evaluators in People ops / recruiting to review AI-generated work products for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote, hourly engagement. You will evaluate artifacts... 
    Remote job
    Hourly pay

    Mercor

    New York, NY
    6 days ago
  • $20 - $80 per hour

     ...Role Overview Help improve next-generation AI systems by supplying precise, real-world evaluation, annotation, and feedback. This remote contractor role focuses on how AI models learn, reason, and perform across diverse subject areas. Key Responsibilities Evaluate... 
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    United States
    a month ago
  • $30 per hour

    About Prolific Prolific is not just another player in the AI space – we are building the biggest pool of quality human data in the...  ...Visual Designers to act as Domain Experts for a high-level AI evaluation project. AI models are evolving beyond simple image generation and... 
    Remote work
    Work from home
    Flexible hours

    Prolific

    New York, NY
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Socratic Dialogue AI Evaluator [Remote]. Be the first to apply!