AI Trainer & Evaluator [Remote]
AuraOne Human Data
- Remote job
AI Trainer & Evaluator is a remote evaluation track for reviewing ai trainer evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn ai trainer evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate ai trainer evaluation model outputs against a versioned rubric and assign severity tags for AI Trainer & Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on ai trainer evaluation or adjacent content for AI Trainer & Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two ai trainer evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- AI Trainer evaluation
- AI model evaluation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$20 per hour
...A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training... ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with excellent...SuggestedHourly payRemote workFlexible hours- ...Supporting diverse AI data and language projects, the hourly contractor AI Trainer and Evaluator will work remotely to generate content, annotate data, and evaluate AI responses for accuracy and cultural relevance. Key responsibilities Generate high-quality prompts and...SuggestedHourly payFor contractorsRemote work
$30 per hour
...Prolific is seeking Advanced Dutch Speakers in Chicago, IL to train AI models. You will complete AI tasks and assess AI performance in... ...home with competitive rates and direct payment through PayPal. Pass the evaluation, and you can start within 15 minutes. #J-18808-Ljbffr...SuggestedRemote workWork from home- **Job Title: AI Trainer || Image Quality Evaluator || English** **Location**: Remote | Work from Home **Employment Type:** Project-based | Contract We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists...SuggestedContract workRemote workWork from homeMonday to FridayDay shift
$80 - $120 per hour
Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks. ~Provide clear, structured written feedback to improve AI outputs. ~Apply...SuggestedPart timeWork at officeRemote work$20 - $80 per hour
...Role Overview Help improve next-generation AI systems by supplying precise, real-world evaluation, annotation, and feedback. This remote contractor role focuses on how AI models learn, reason, and perform across diverse subject areas. Key Responsibilities Evaluate...Hourly payContract workFor contractorsRemote work- ...Summary This is a fully remote, hourly contractor role supporting AI data and language projects on a project-based, flexible hour... ..., and other content to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy, reasoning quality...Hourly payFor contractorsRemote workFlexible hours
- Mercor is seeking expert Evaluators in Finance operations / audit support to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. This is a remote, hourly engagement, requiring strong subject-matter expertise to grade outputs...Remote jobHourly pay
- Feitong Buke is hiring a Lead AI Trainer to oversee and enhance the quality of AI model dialogues with users. The role involves reviewing datasets for accuracy, providing feedback to annotators, and validating AI model outputs to ensure high production quality. Candidates...Remote jobFull time
$65 per hour
Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience...Hourly paySelf employmentWork from homeFlexible hours- ...contributors for flexible, project-based roles that require a strong background in finance. You will design complex tasks that challenge AI agents and enhance their learning. Ideal candidates have a Bachelor's or Master's degree, extensive experience in finance, and...Remote jobFlexible hours
- ...Key Resources is seeking experienced healthcare operations professionals for a fully remote 12-week consulting engagement focused on AI tools in healthcare. You will apply your domain expertise to improve AI-generated content and guide model behavior. Ideal candidates...Remote job
$20 - $80 per hour
...role, you''ll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and... ...through high-quality, real-world input. Key Responsibilities: Evaluate and score AI-generated responses using well-defined rubrics and...Hourly payContract workRemote work$135 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Measure accuracy against a held-out set and assess data quality. Evaluate the realistic performance ceiling of the computer vision system...Hourly payFull timeContract workFor contractorsSummer workRemote work$60 - $90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...expert-level prompts across specialized cybersecurity topics. Evaluate and annotate model responses for technical accuracy, helpfulness...Full timeContract workSummer workImmediate startRemote work$70 - $90 per hour
...About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position...Full timeContract workSummer workRemote work$100 - $150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...exfiltration , ransomware , worms , and exploits . Evaluate POC exploit development to determine boundaries between security...Hourly payWeekly payFull timeContract workFor contractorsSummer workRemote work$120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...hour Location: Remote Role Responsibilities Evaluate complex technical tasks using deep language expertise in Scala...Hourly payWeekly payFull timeContract workFor contractorsSummer workRemote work$20 per hour
...indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What...Hourly payPermanent employmentTemporary workPart timeFreelanceRemote work10 hours per week$23 per hour
...Annotators connects individuals with Generative AI projects from leading tech innovators.... ...such as rating AI-generated content, evaluating factual accuracy, or comparing responses... ...expertise required. On this project, AI trainers earn up to $23 per hour equivalent. Qualifications...Hourly payPart timeFreelanceRemote work$23 per hour
...Toloka is seeking an on-demand annotator to join our global freelance QA network. This part-time, remote role offers flexible hours on AI project tasks and compensation up to $23 per hour equivalent. You will review data types (text, images, or videos), label content per...Hourly payPart timeFreelanceRemote workFlexible hours$23 per hour
...and improve artificial intelligence. As an annotator, you may be invited to take part in online projects such as rating AI-generated content, evaluating factual accuracy, or comparing responses - when projects are available. This is a project-based, part-time, remote...Hourly payPart timeFreelanceRemote work- ...Toloka is seeking online annotators to join AI-related projects, rating content, evaluating factual accuracy, and comparing responses as projects become available. This role is remote, part-time, freelance, and compensation varies by project scope. Applicants should have...Part timeFreelanceRemote work
$104k - $208k
...Ireland, Australia, or New Zealand. Responsibilities: We ask you to design and solve a wide range of coding challenges used to train AI systems. We ask you to write clear, well-structured code examples and thorough explanations. We ask you to assess AI-generated...Hourly payFull timeRemote work- ...The Toloka Annotators connects individuals with Generative AI projects from leading tech innovators. Our mission is to unlock the... ...take part in online projects such as rating AI-generated content, evaluating factual accuracy, or comparing responses – when projects are available...FreelanceRemote work
$60 per hour
A data science company is seeking a Freelance Data Science Expert (Python & SQL) to create computational problems for Generative AI projects. The ideal candidate has a Master’s or PhD in a relevant field, at least 5 years of experience, and expertise in statistical analysis...FreelanceRemote workFlexible hours$75 - $100 per hour
...About the Company Ajaia is a rapidly growing AI consultancy and product studio dedicated to helping organizations move from experimentation to real results. We operate across four core pillars: AI Strategy and Advisory, AI Engineering and Automation, Workforce Training...For contractorsImmediate startRemote workShift work- ...Job : AI Prompting Trainer Location : This role will require frequent on site travel to San Diego, California, It will be remote besides the in person training sessions ( bi weekly/monthly) Someone who is relatively close (within a few hours) is preferred but...Part timeRemote work
- A leading AI project company is seeking annotators for freelance, part-time remote work. You will be responsible for reviewing and labeling data across various projects. Candidates should have at least a Bachelor's degree, be proficient in Portuguese, and have strong English...Part timeFreelanceRemote workFlexible hours
$23 per hour
A leading AI project company is seeking annotators for various Generative AI projects. Responsibilities include reviewing and labeling data, with a focus on accuracy and detail. Candidates should be proficient in Portuguese and have advanced English skills. This role offers...Part timeFreelanceRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Trainer & Evaluator [Remote]. Be the first to apply!




