AI Trainer & Evaluator [Remote]
AuraOne Human Data
- Remote job
AI Trainer & Evaluator is a remote evaluation track for reviewing ai trainer evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn ai trainer evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate ai trainer evaluation model outputs against a versioned rubric and assign severity tags for AI Trainer & Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on ai trainer evaluation or adjacent content for AI Trainer & Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two ai trainer evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- AI Trainer evaluation
- AI model evaluation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$50 - $60 per hour
...DataAnnotation is seeking a Clinical Specialist to assist in training AI models, focusing on evaluating healthcare-related problems. The role is suitable for healthcare professionals, including physicians and advanced practice clinicians, and allows for a flexible schedule...SuggestedHourly payFor contractorsRemote workWork from homeFlexible hours- ...Remote | Work from Home Employment Type: Project-based | Contract We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists with strong proficiency in English. In this role, you will support AI/ML...SuggestedContract workRemote workWork from homeMonday to FridayDay shift
$50 - $60 per hour
...DataAnnotation is seeking a Clinical Specialist to aid in training AI models by providing diverse healthcare problems and evaluating outputs. This independent contractor position allows you to work remotely, on your schedule, and offers hourly rates starting at $50-$60...SuggestedHourly payFor contractorsRemote work$50 - $190 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Remote Commitment: 20+ hours/week Role Responsibilities Evaluate AI systems on complex personal workflows, including personal...SuggestedHourly payContract workFor contractorsSummer workRemote workTrial period$50 - $60 per hour
...DataAnnotation is seeking a Clinical Specialist in Tennessee to help train AI models by providing diverse healthcare-related problems for AI chatbots. This role allows you to work from home on your own schedule. The ideal candidate should have a current or in-progress...SuggestedHourly payRemote workWork from home- ...Summary This is a fully remote, hourly contractor role supporting AI data and language projects on a project-based, flexible hour... ..., and other content to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy, reasoning quality...Hourly payFor contractorsRemote workFlexible hours
$50 - $60 per hour
...DataAnnotation is looking for a Clinical Specialist to assist in training AI models. This independent contractor role allows you to work from home with flexible hours, evaluating diverse healthcare-related problems for AI chatbots and ensuring medical accuracy. The ideal...Hourly payFor contractorsRemote workWork from homeFlexible hours- ...DataAnnotation is seeking a Clinical Specialist to train AI models by providing complex healthcare-related problems for evaluation. This remote independent contractor position allows for flexible scheduling and projects paid hourly, starting at $50–$60, with bonuses on...Hourly payFor contractorsRemote workFlexible hours
$70 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...: $70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses to ensure accuracy and depth in reasoning...Contract workSummer workRemote work$50 - $60 per hour
...DataAnnotation is seeking a Clinical Specialist in Wisconsin to train AI models focused on healthcare. You’ll provide complex medical problems to AI chatbots and evaluate their performance for accuracy and quality. The ideal candidate will have a medical degree or be...Hourly payFor contractorsRemote workFlexible hours- ...DataAnnotation is looking for a Clinical Specialist to help train AI models in the United States. The role involves evaluating AI outputs related to healthcare and ensuring their accuracy. Candidates must be fluent in English and hold a current or in-progress medical...Hourly payFor contractorsRemote workFlexible hours
- ...Supporting AI data and language projects, the hourly contractor AI Trainer and Evaluator will work remotely on a flexible basis, focusing on content generation, data annotation, and evaluating AI-generated responses for accuracy and cultural appropriateness. Key responsibilities...Hourly payFor contractorsRemote workFlexible hours
$20 - $80 per hour
...Role Overview Train and evaluate next-generation AI systems by scoring model outputs, annotating real-world content, and delivering clear, actionable feedback that improves model accuracy and reasoning across diverse domains. About the company micro1 is an AI data...Hourly payFor contractorsRemote work$80 - $120 per hour
Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks. ~Provide clear, structured written feedback to improve AI outputs. ~Apply...Part timeWork at officeRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...and Jack Dorsey . Position: Procurement / vendor management Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...consistent. Preferred ~ Prior experience with AI training , evaluation, or human-data projects. Application Process (Takes 20–30...Contract workSummer workRemote work$80 - $150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...quickly on work. Work independently and asynchronously to evaluate and improve AI model performance . Qualifications Must-Have...Contract workSummer workRemote work$80 - $120 per hour
Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks. ~Provide clear, structured written feedback to improve AI model outputs....Part timeWork at officeRemote work$20 per hour
A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI... ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with...Remote jobHourly payFlexible hours- YO IT Consulting is seeking an AI Trainer & Evaluator for a remote contract role. You will train next-generation AI systems by providing high-quality real-world input and precise scoring against rubrics. This position emphasizes data annotation, rubric refinement, and...Remote jobContract work
- CNTXT AI is seeking a fully remote, hourly contractor to support AI data and language projects on a flexible, project-based schedule. The role involves content generation, data annotation, LLM evaluation, and localization QA across diverse topics. Ideal candidates are native...Remote jobHourly payFor contractorsFlexible hours
$30 per hour
Prolific is seeking Advanced Dutch Speakers in Chicago, IL to train AI models. You will complete AI tasks and assess AI performance in... ...competitive rates and direct payment through PayPal. Pass the evaluation, and you can start within 15 minutes. #J-18808-Ljbffr ProlificRemote jobWork from home- YO IT Consulting is seeking an AI Trainer & Evaluator for a remote contract role. You will apply your domain expertise to train next-generation AI systems, shaping how models learn, reason, and perform through high-quality real-world input. Responsibilities include scoring...Remote jobContract work
- Alignerr is seeking a Population Health Informaticist to help shape AI systems that interpret population-level health data. You will evaluate health analytics, identify discrepancies, and provide expert feedback on health informatics concepts. Remote, hourly contract work...Remote jobHourly payContract workFlexible hours
- Feitong Buke is hiring a Lead AI Trainer to oversee and enhance the quality of AI model dialogues with users. The role involves reviewing datasets for accuracy, providing feedback to annotators, and validating AI model outputs to ensure high production quality. Candidates...Remote jobFull time
$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...and Jack Dorsey . Position: Public health communications Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Personal finance / consumer planning Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Finance operations / audit support Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$85 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-...Contract workSummer workRemote work$65 per hour
Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience...Hourly paySelf employmentWork from homeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Trainer & Evaluator [Remote]. Be the first to apply!



