Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Trainer & Evaluator [Remote]

AuraOne Human Data

Remote
  • Remote job

AI Trainer & Evaluator is a remote evaluation track for reviewing ai trainer evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.

Why this role matters

AI data reviewers help turn ai trainer evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.

Responsibilities

  • Evaluate ai trainer evaluation model outputs against a versioned rubric and assign severity tags for AI Trainer & Evaluator assignments.
  • Compare paired responses and pick the stronger answer with a written rationale.
  • Label hallucinations, instruction-following failures, and unsafe content with structured tags.
  • Capture ambiguous prompts and route them back to the program team for rubric updates.
  • Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
  • Document recurring failure modes so the modeling team can target them in the next training run.

Qualifications

  • Prior evaluation, annotation, or human-rater experience on ai trainer evaluation or adjacent content for AI Trainer & Evaluator work.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that names the issue and the rubric clause being applied.
  • Strong attention to detail and the ability to flag when a prompt itself is the problem.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Compare two ai trainer evaluation model responses to the same prompt and pick the stronger one with rationale.
  • Tag an unsafe response with the correct policy category and severity.
  • Audit a 50-row batch for rubric consistency and report drift to the program lead.
  • Propose a rubric clarification after spotting a recurring failure mode.

Nice to have

  • Background in linguistics, content moderation, or trust & safety review.
  • Experience with inter-rater agreement metrics and calibration cycles.
  • Domain expertise that lets you spot subject-matter errors automated checks miss.

Skills

  • Model output evaluation
  • Rubric-based annotation
  • Severity tagging
  • Inter-rater calibration
  • AI Trainer evaluation
  • AI model evaluation

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 14 days ago
Similar jobs that could be interesting for youBased on the AI Trainer & Evaluator [Remote] in Remote vacancy
  • $50 - $60 per hour

     ...DataAnnotation is seeking a Clinical Specialist to assist in training AI models, focusing on evaluating healthcare-related problems. The role is suitable for healthcare professionals, including physicians and advanced practice clinicians, and allows for a flexible schedule... 
    Suggested
    Hourly pay
    For contractors
    Remote work
    Work from home
    Flexible hours

    DataAnnotation

    Salt Lake City, UT
    4 days ago
  •  ...Remote | Work from Home Employment Type: Project-based | Contract  We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists with strong proficiency in English. In this role, you will support AI/ML... 
    Suggested
    Contract work
    Remote work
    Work from home
    Monday to Friday
    Day shift

    iMerit Technologies

    San Jose, CA
    a month ago
  • $50 - $60 per hour

     ...DataAnnotation is seeking a Clinical Specialist to aid in training AI models by providing diverse healthcare problems and evaluating outputs. This independent contractor position allows you to work remotely, on your schedule, and offers hourly rates starting at $50-$60... 
    Suggested
    Hourly pay
    For contractors
    Remote work

    DataAnnotation

    New York, NY
    4 days ago
  • $50 - $190 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Remote Commitment: 20+ hours/week Role Responsibilities Evaluate AI systems on complex personal workflows, including personal... 
    Suggested
    Hourly pay
    Contract work
    For contractors
    Summer work
    Remote work
    Trial period

    Mercor

    New York, NY
    5 days ago
  • $50 - $60 per hour

     ...DataAnnotation is seeking a Clinical Specialist in Tennessee to help train AI models by providing diverse healthcare-related problems for AI chatbots. This role allows you to work from home on your own schedule. The ideal candidate should have a current or in-progress... 
    Suggested
    Hourly pay
    Remote work
    Work from home

    DataAnnotation

    Nashville, TN
    5 days ago
  •  ...Summary This is a fully remote, hourly contractor role supporting AI data and language projects on a project-based, flexible hour...  ..., and other content to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy, reasoning quality... 
    Hourly pay
    For contractors
    Remote work
    Flexible hours

    CNTXT AI

    Brooklyn, NY
    8 days ago
  • $50 - $60 per hour

     ...DataAnnotation is looking for a Clinical Specialist to assist in training AI models. This independent contractor role allows you to work from home with flexible hours, evaluating diverse healthcare-related problems for AI chatbots and ensuring medical accuracy. The ideal... 
    Hourly pay
    For contractors
    Remote work
    Work from home
    Flexible hours

    DataAnnotation

    Jackson, MS
    5 days ago
  •  ...DataAnnotation is seeking a Clinical Specialist to train AI models by providing complex healthcare-related problems for evaluation. This remote independent contractor position allows for flexible scheduling and projects paid hourly, starting at $50–$60, with bonuses on... 
    Hourly pay
    For contractors
    Remote work
    Flexible hours

    DataAnnotation

    Hartford, CT
    5 days ago
  • $70 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...: $70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses to ensure accuracy and depth in reasoning... 
    Contract work
    Summer work
    Remote work

    Mercor

    Boston, MA
    more than 2 months ago
  • $50 - $60 per hour

     ...DataAnnotation is seeking a Clinical Specialist in Wisconsin to train AI models focused on healthcare. You’ll provide complex medical problems to AI chatbots and evaluate their performance for accuracy and quality. The ideal candidate will have a medical degree or be... 
    Hourly pay
    For contractors
    Remote work
    Flexible hours

    DataAnnotation

    Madison, WI
    5 days ago
  •  ...DataAnnotation is looking for a Clinical Specialist to help train AI models in the United States. The role involves evaluating AI outputs related to healthcare and ensuring their accuracy. Candidates must be fluent in English and hold a current or in-progress medical... 
    Hourly pay
    For contractors
    Remote work
    Flexible hours

    DataAnnotation

    United States
    4 days ago
  •  ...Supporting AI data and language projects, the hourly contractor AI Trainer and Evaluator will work remotely on a flexible basis, focusing on content generation, data annotation, and evaluating AI-generated responses for accuracy and cultural appropriateness. Key responsibilities... 
    Hourly pay
    For contractors
    Remote work
    Flexible hours

    Virtual Vocations Inc

    United States
    2 days ago
  • $20 - $80 per hour

     ...Role Overview Train and evaluate next-generation AI systems by scoring model outputs, annotating real-world content, and delivering clear, actionable feedback that improves model accuracy and reasoning across diverse domains. About the company micro1 is an AI data... 
    Hourly pay
    For contractors
    Remote work

    SaidGig

    United States
    15 days ago
  • $80 - $120 per hour

    Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks. ~Provide clear, structured written feedback to improve AI outputs. ~Apply... 
    Part time
    Work at office
    Remote work

    Mercor

    Remote
    4 days ago
  • $80 - $120 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...and Jack Dorsey . Position: Procurement / vendor management Evaluator Type: Contract Compensation: $80–$120/hour Location... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    Dallas, TX
    5 days ago
  • $150 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...consistent. Preferred ~ Prior experience with AI training , evaluation, or human-data projects. Application Process (Takes 20–30... 
    Contract work
    Summer work
    Remote work

    Mercor

    Dallas, TX
    5 days ago
  • $80 - $150 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...quickly on work. Work independently and asynchronously to evaluate and improve AI model performance . Qualifications Must-Have... 
    Contract work
    Summer work
    Remote work

    Mercor

    Pennsauken, NJ
    5 days ago
  • $80 - $120 per hour

    Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks. ~Provide clear, structured written feedback to improve AI model outputs.... 
    Part time
    Work at office
    Remote work

    Mercor

    Remote
    a month ago
  • $20 per hour

    A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI...  ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with... 
    Remote job
    Hourly pay
    Flexible hours

    SupportFinity™

    Raleigh, NC
    5 days ago
  • YO IT Consulting is seeking an AI Trainer & Evaluator for a remote contract role. You will train next-generation AI systems by providing high-quality real-world input and precise scoring against rubrics. This position emphasizes data annotation, rubric refinement, and... 
    Remote job
    Contract work

    YO IT Consulting

    Chicago, IL
    5 days ago
  • CNTXT AI is seeking a fully remote, hourly contractor to support AI data and language projects on a flexible, project-based schedule. The role involves content generation, data annotation, LLM evaluation, and localization QA across diverse topics. Ideal candidates are native... 
    Remote job
    Hourly pay
    For contractors
    Flexible hours

    CNTXT AI

    New York, NY
    5 days ago
  • $30 per hour

    Prolific is seeking Advanced Dutch Speakers in Chicago, IL to train AI models. You will complete AI tasks and assess AI performance in...  ...competitive rates and direct payment through PayPal. Pass the evaluation, and you can start within 15 minutes. #J-18808-Ljbffr Prolific
    Remote job
    Work from home

    Prolific

    Chicago, IL
    1 day ago
  • YO IT Consulting is seeking an AI Trainer & Evaluator for a remote contract role. You will apply your domain expertise to train next-generation AI systems, shaping how models learn, reason, and perform through high-quality real-world input. Responsibilities include scoring... 
    Remote job
    Contract work

    YO IT Consulting

    New York, NY
    5 days ago
  • Alignerr is seeking a Population Health Informaticist to help shape AI systems that interpret population-level health data. You will evaluate health analytics, identify discrepancies, and provide expert feedback on health informatics concepts. Remote, hourly contract work... 
    Remote job
    Hourly pay
    Contract work
    Flexible hours

    Alignerr

    Seattle, WA
    5 days ago
  • Feitong Buke is hiring a Lead AI Trainer to oversee and enhance the quality of AI model dialogues with users. The role involves reviewing datasets for accuracy, providing feedback to annotators, and validating AI model outputs to ensure high production quality. Candidates... 
    Remote job
    Full time

    Feitong Buke

    New York, NY
    3 days ago
  • $80 - $120 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...and Jack Dorsey . Position: Public health communications Evaluator Type: Contract Compensation: $80–$120/hour Location... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    New York, NY
    5 days ago
  • $80 - $120 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Jack Dorsey . Position: Personal finance / consumer planning Evaluator Type: Contract Compensation: $80–$120/hour Location... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    Charlotte, NC
    5 days ago
  • $80 - $120 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Jack Dorsey . Position: Finance operations / audit support Evaluator Type: Contract Compensation: $80–$120/hour Location... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    Charlotte, NC
    5 days ago
  • $85 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    a month ago
  • $65 per hour

    Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience... 
    Hourly pay
    Self employment
    Work from home
    Flexible hours

    Prolific

    Las Vegas, NV
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Trainer & Evaluator [Remote]. Be the first to apply!