AI Trainer & Evaluator [Remote]
AuraOne Human Data
- Remote job
AI Trainer & Evaluator is a remote evaluation track for reviewing ai trainer evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn ai trainer evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate ai trainer evaluation model outputs against a versioned rubric and assign severity tags for AI Trainer & Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on ai trainer evaluation or adjacent content for AI Trainer & Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two ai trainer evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- AI Trainer evaluation
- AI model evaluation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$20 per hour
A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI... ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with...SuggestedHourly payRemote workFlexible hours$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Summers , and Jack Dorsey . Position: Cybersecurity / IT GRC Evaluator Type: Contract Compensation: $80–$120/hour...SuggestedContract workSummer workWork at officeRemote work$30 per hour
Prolific is seeking Advanced Dutch Speakers in Chicago, IL to train AI models. You will complete AI tasks and assess AI performance in... ...competitive rates and direct payment through PayPal. Pass the evaluation, and you can start within 15 minutes. #J-18808-Ljbffr ProlificSuggestedRemote jobWork from home$60 per hour
Prolific is looking for Biology Experts and Life Science Professionals in Las Vegas, NV to evaluate AI-generated science. Successful candidates will work flexibly from home, earning up to $60 per hour for paid tasks. Responsibilities include reviewing scientific data accuracy...SuggestedRemote jobHourly payWork from homeFlexible hours$11 - $30.65 per hour
Meridial is seeking contractors to evaluate advanced agentic audio models by simulating realistic customer service interactions across multiple domains. You will contribute to developing diverse datasets and assess model performance using various metrics. The role requires...SuggestedRemote jobHourly payFor contractors$90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...genuinely correct replications from those that merely look correct. Evaluate responsive behavior and semantic quality, ensuring proper use of...Contract workSummer workLocal areaRemote work$100 - $150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...00–$150/hour Location: Remote Role Responsibilities Evaluate AI-generated artifacts for usability in a consulting context....Contract workSummer workWork at officeRemote work$80 - $150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Summers , and Jack Dorsey . Position: MS Excel / Google Sheets Evaluator Type: Contract Compensation: $80–$150/hour...Contract workSummer workRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ..., and Jack Dorsey . Position: Process improvement / SOPs Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$100 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... .... Position: Clinical Mental Health Expert — AI Conversation Evaluation Type: Contract Compensation: $100/hour Location:...Contract workSummer workRemote work- Prolific Academic Ltd is seeking Licensed Pharmacists to assist in AI model training and evaluation from a home office. Successful candidates join as Domain Experts and will be paid to train and evaluate models, with tests to assess suitability. Experts review AI-driven...Remote workHome officeFlexible hours
$65 per hour
Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience...Hourly paySelf employmentWork from homeFlexible hours$80 - $150 per hour
Prolific Academic Ltd is recruiting Medical Doctors to help train and evaluate AI models. You’ll complete a quick test to assess suitability and, if successful, join as a Domain Expert, paid to work on AI tasks. Researchers pay $80-$150 per hour per completed task, with...Hourly paySelf employmentWork from home- OpenTrain AI, Inc. is seeking a remote, part-time contractor to develop and evaluate psychiatric materials used to train AI systems. This role blends clinical expertise, Malay cultural knowledge, bilingual writing, and careful review of sensitive mental health content....Remote jobPart timeFor contractors
- ...Summary This is a fully remote, hourly contractor role supporting AI data and language projects on a project-based, flexible hour... ..., and other content to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy, reasoning quality...Hourly payTemporary workFor contractorsRemote workFlexible hours
$70 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...: $70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses to strengthen reasoning and rigor in model...Contract workSummer workRemote work$20 - $160 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Position: Generalist Annotator — Health AI Conversation Quality Evaluation Type: Contract Compensation: $20–$160/hour Location...Contract workSummer workRemote work$20 - $80 per hour
...role, you''ll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and... ...through high-quality, real-world input. Key Responsibilities: Evaluate and score AI-generated responses using well-defined rubrics and...Hourly payContract workRemote work$150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...consistent. Preferred ~ Prior experience with AI training , evaluation, or human-data projects. Application Process (Takes 20–30...Contract workSummer workRemote work$120 - $175 per hour
...Job Description Job Description About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers...Contract workSummer workRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... .... Position: User/customer research and feedback synthesis Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$80 - $150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Summers , and Jack Dorsey . Position: MS Excel / Google Sheets Evaluator Type: Contract Compensation: $80–$150/hour...Contract workSummer workRemote work$50 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...: $50/hour Location: Remote Role Responsibilities Evaluate the quality and accuracy of LLM-generated English text across...Contract workSummer workRemote work$85 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-...Contract workSummer workRemote work$22 - $27 per hour
...Job Description Job Description About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers...Contract workSummer workRemote work$60 - $70 per hour
...growing community of professionals advancing the next wave of AI. As an AI Trainer, you’ll play a hands-on role by analyzing and providing... ...experts only Engagement: Part-time, project-based expert evaluation work Work Type: Remote Project Summary...Hourly payFull timeContract workPart timeRemote work$30 per hour
...AI Trainer - Advanced Japanese Fluency About Prolific Prolific is not just another player in the AI space – we are building the... ...We’re looking for Advanced Japanese Speakers to help train and evaluate cutting-edge AI models. If you have the necessary experience,...Self employmentRemote workWork from homeFlexible hours$100 - $150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...00–$150/hour Location: Remote Role Responsibilities Evaluate AI-generated artifacts for usability in a software engineering...Contract workSummer workWork at officeRemote work- ...Meridial is looking for a Hungarian Trilingual Language Specialist – AI Trainer to work remotely. The role involves reviewing AI-generated text for accuracy, idiomatic expressions, and cultural appropriateness. Candidates should have fluency in Hungarian and English,...Hourly payContract workRemote work
$23 per hour
...A leading AI project company is seeking an annotator for a remote, freelance role focused on helping train and improve artificial intelligence. Candidates must be proficient in Portuguese and advanced in English. Responsibilities include reviewing and labeling data for...Part timeFreelanceRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Trainer & Evaluator [Remote]. Be the first to apply!





