Lesson Plan AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Lesson Plan AI Evaluator is a remote evaluation track for reviewing lesson plan ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn lesson plan ai evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate lesson plan ai evaluation model outputs against a versioned rubric and assign severity tags for Lesson Plan AI Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on lesson plan ai evaluation or adjacent content for Lesson Plan AI Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two lesson plan ai evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Lesson Plan AI evaluation
- Learning design
- Assessment review
- Pedagogy
- Lesson
- Plan
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- AI Trainer Jobs is seeking a Lesson Plan AI Evaluator for a remote, contractor role. You will review prompts and responses against AuraOne's quality rubric, label edge cases, and provide structured feedback to help retrain models. Responsibilities include evaluating model...SuggestedRemote jobFor contractors
$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Dorsey . Position: Operations / inventory / capacity planning Evaluator Type: Contract Compensation: $80–$120/hour Location...SuggestedContract workSummer workWork at officeRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Dorsey . Position: Operations / inventory / capacity planning Evaluator Type: Contract Compensation: $80–$120/hour Location...SuggestedContract workSummer workWork at officeRemote work$14.5 per hour
...AI Web Search Evaluator As a Web Search Evaluator, you will play a key role in improving the quality of search engine results, ensuring users... ...Illness, Hospital Indemnity Insurance ~401(k) Retirement Plan Federal Law Compliance: Verify identity and eligibility...SuggestedHourly payPart timeCurrently hiringImmediate startRemote workWork from home10 hours per weekFlexible hours$14.5 per hour
...diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data... ...provide insights on relevance and quality. Evaluate and rate the effectiveness of search engine... ...Indemnity Insurance 401(k) Retirement Plan Federal Law Compliance Verify identity and...SuggestedHourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours- Surgical Planning Safety Evaluator is a remote evaluation track for reviewing surgical planning safety evaluation prompts and responses against AuraOne... ...structured feedback the modeling team can use to retrain. AI data reviewers help turn surgical planning safety evaluation...Remote job
$37.5 per hour
...looking to create a golden data set for AI generated ad copy for various advertisers... ...In this role you will be responsible for evaluating advertising copy generated by a large language... ...may be subject to specific elections, plan, or program terms. If eligible, the...Contract workTemporary workRemote work$14.5 per hour
A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical...Hourly payPart timeRemote workFlexible hours- ...Weekday 1 is seeking expert Evaluators to review AI-generated real estate, hospitality and events outputs for accuracy, rigor and domain quality. You will apply deep expertise to grade documents, spreadsheets and slide decks. Requirements include 5+ years in Real estate...Hourly payWeekly payContract workWork at officeRemote workWeekday work
- ...About the role Swahili Evaluation AI Evaluator is a remote evaluation track for reviewing swahili generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback...Hourly payFor contractorsRemote work10 hours per week
$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...Remote workFlexible hours- ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and...Full timeContract workRemote workFlexible hours
- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Work at officeRemote work
$11.5 per hour
...position as an Online Task Contributor. In this role, you will evaluate and provide feedback on content to enhance search engine results... ...11.50 hourly, based on task completion, with a supportive community of contributors involved in AI advancements. #J-18808-Ljbffr...Hourly payPart timeRemote work$30 per hour
...Prolific is not just another player in the AI space - we are building the biggest pool... ...as Domain Experts for a high-level AI evaluation project. AI models are evolving beyond simple... ...for recruiting and global organisation planning. Prolific's Candidate Privacy Notice explains...Remote jobWork from homeFlexible hours- ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise...Hourly payWork at officeRemote work
$20 per hour
A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI... ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with...Hourly payRemote workFlexible hours$20 - $30 per hour
A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong...Remote jobHourly pay- YO AI Labs is seeking an AI Evaluation Specialist to remotely assess enterprise AI outputs against detailed rubrics, measuring accuracy, relevance and quality. You will provide concise, actionable feedback to drive improvements. Responsibilities include identifying reasoning...Remote jobFor contractors
- YO AI Labs is seeking an Investment & Finance Expert to work remotely as a contractor. The role focuses on evaluating AI outputs related to valuation, financial modeling, markets, and investments, and crafting expert prompts plus reference answers to improve model reasoning...Remote jobFor contractors
- YO AI Labs is seeking Content Writers & Editors to contribute to a customer project focused on text quality and AI evaluation. You will apply your writing and editing expertise to help train and evaluate next-generation AI systems through high-quality input. No prior AI...Remote job
- AI Trainer Jobs seeks a remote independent contractor to review AI outputs for fintech operations evaluation. You will assess workflow adherence, tone, and escalation logic, assigning severity tags and documenting next steps for model training. The role requires experience...Remote jobHourly payFor contractorsFlexible hours
- Archangel Health AI is seeking Clinical AI Evaluators to review AI-generated clinical outputs, benchmark diagnostic reasoning, and refine responses to real-world medical queries. You will perform clinical accuracy auditing, RLHF ranking, error and harm identification,...Remote workFlexible hoursShift work
- Brazilian Portuguese AI Evaluator is a remote evaluation track for reviewing portuguese generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling...Hourly payFor contractorsRemote work10 hours per week
- AuraOne is seeking a Children Safety Content AI Evaluator for a remote, contractor-based role. You will evaluate model outputs, compare responses, and label content with structured tags according to our quality rubric. Ideal candidates have prior evaluation/annotation...Remote jobFor contractors10 hours per week
$80 per hour
prolificacademicltd seeks senior AI/ML engineers to join an expert network contributing to large language model training and evaluation. Roles are task-based, with participants paid per study for applying deep technical judgment to model outputs and ML artifacts. Responsibilities...Remote jobHourly payFlexible hours$40 - $100 per hour
About OpenTrain OpenTrain AI is the hiring and contracting organization for this opportunity. OpenTrain is the #1 platform for finding... ...English-language work About AI Training and Scientific Evaluation AI training is the human side of building modern artificial intelligence...Hourly payContract workPart timeFor contractorsRemote work- Alignerr is seeking Wildlife and Habitat Conservation Scientists to evaluate AI-trained biodiversity protection content. This remote, hourly contract role offers flexible hours (10-40 per week) and a chance to shape how AI understands ecological challenges on a global scale...Remote jobHourly payContract workFlexible hours
- Unknown is seeking a Polish Bilingual Expert (contractor) to evaluate Polish audio content for AI training. The role emphasizes nativeness, fluency, and linguistic quality, with feedback delivered in English. Strong Polish language skills and clear written and verbal English...Remote jobFor contractors
- Prolific Academic Ltd seeks a remote Punjabi evaluator for refining AI language outputs. This role is based in the United States and involves reviewing Punjabi text and audio, flagging anything unnatural, and providing structured feedback to support AI training. You will...Remote jobWork from home
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lesson Plan AI Evaluator [Remote]. Be the first to apply!



