Tutor Conversation AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Tutor Conversation AI Evaluator is a remote review track for evaluating AI outputs across tutor conversation ai specialist operations workflows. Reviewers grade workflow correctness, policy adherence, and stakeholder fit; flag operational risk; and document the right next step so the modeling team can train on it.
Why this role matters
Tutor Conversation AI specialist operations AI has to fit into an actual day at work. AuraOne uses experienced operators to grade outputs the way a senior peer would — checking workflow, policy, and the unwritten rules that decide whether a task actually gets done.
Responsibilities
- Review AI outputs against current tutor conversation ai specialist operations workflows, playbooks, and firm policy for Tutor Conversation AI Evaluator assignments.
- Grade tone, escalation logic, and stakeholder fit on a structured rubric.
- Flag operational risk, missed escalations, and policy-adherence gaps with severity tags.
- Capture the right next step so the modeling team can train on it.
- Adjudicate disputed treatments against published playbooks or firm guidance.
- Maintain reviewer-quality scores in inter-rater calibration cycles.
Qualifications
- Direct working experience in tutor conversation ai specialist operations on real teams for Tutor Conversation AI Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the policy or workflow being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Grade a model's response to a real tutor conversation ai specialist operations ticket and rate workflow, tone, and escalation.
- Flag a missed escalation with the right severity tag and corrected next step.
- Adjudicate a disputed playbook call between two reviewers using firm guidance.
- Audit a 25-row batch for rubric consistency and report drift to the program lead.
Nice to have
- Prior experience training, calibrating, or QA-ing operations teams.
- Familiarity with AI-assisted workflow tooling and its failure modes.
- Bilingual experience for cross-region operations.
Skills
- Operational review
- Policy adherence
- Workflow judgment
- Stakeholder communication
- Tutor Conversation AI specialist operations
- Learning design
- Assessment review
- Pedagogy
- Tutor
- Conversation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$20 - $160 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Dorsey . Position: Generalist Annotator — Health AI Conversation Quality Evaluation Type: Contract Compensation: $20–$160/hour...SuggestedFull timeContract workSummer workRemote work$100 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Dorsey . Position: Clinical Mental Health Expert — AI Conversation Evaluation Type: Contract Compensation: $100/hour Location...SuggestedContract workSummer workRemote work- ...An innovative education provider seeks a freelance teacher to guide students in Conversation Techniques. In this flexible role, you will craft engaging course materials and support students through their learning journey. With the freedom to shape your curriculum and...SuggestedFreelanceRemote workFlexible hours
- About the role AI Safety and Red-Team Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers... ...Score model defenses across single-turn and multi-turn conversations. Triage emerging attack vectors and route them to the safety...SuggestedHourly payFor contractorsRemote work10 hours per week
$35 - $75 per hour
SpaceXAI’s mission is to create AI systems that can accurately... ...teammates. ABOUT THE ROLE: You will evaluate, refine, and create elite-... ...impacts (e.g., 30%+ conversion boosts, engagement/revenue growth... ...LOCATION & OTHER EXPECTATIONS: Tutor roles may be offered as full-time...SuggestedFull timePart timeFor contractorsRemote workWorldwide10 hours per week$50 - $100 per hour
...Apply your teaching and tutoring expertise to help improve next-generation AI systems through realistic, one-on-one educational conversations. This remote contract role focuses on providing clear, adaptive academic support that reflects real-world tutoring and teaching...Hourly payContract workFor contractorsRemote work$72.8k - $93.6k
...SpaceXAI’s mission is to create AI systems that can accurately... ...THE ROLE: As an AI Tutor specialized in multilingual audio... ...multilingual audio content, including evaluating speech accuracy, cultural... ...UI, product copy, or conversational AI text. Professional experience...Full timePart timeFor contractorsRemote workWorldwide10 hours per week- ...About the role Swahili Evaluation AI Evaluator is a remote evaluation track for reviewing swahili generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback...Hourly payFor contractorsRemote work10 hours per week
- ...Weekday 1 is seeking expert Evaluators to review AI-generated real estate, hospitality and events outputs for accuracy, rigor and domain quality. You will apply deep expertise to grade documents, spreadsheets and slide decks. Requirements include 5+ years in Real estate...Hourly payWeekly payContract workWork at officeRemote workWeekday work
- ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and...Full timeContract workRemote workFlexible hours
$14.5 per hour
A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical...Hourly payPart timeRemote workFlexible hours$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...Remote workFlexible hours- ...Alignerr is seeking Creative Writing Evaluators to assess AI-generated stories, essays, poetry, and other writings to shape how AI learns compelling prose. This fully remote, flexible contract roles welcomes avid readers and writers with no publishing credits required...Contract workRemote workFlexible hours
- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Work at officeRemote work
$14.5 per hour
...AI Web Search Evaluator Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years of experience in...Bi-weekly payHourly payPart timeImmediate startRemote workWork from homeFlexible hours$14.5 per hour
...Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data...Hourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours- A national online tutoring platform is seeking Instant Tutors for online Calculus sessions. Enjoy competitive pay, flexibility to work whenever you are available, and the ability to help students immediately with their learning needs. Ideal candidates will have strong...Immediate startRemote workFlexible hours
$35 - $45 per hour
...A forward-thinking tech company is seeking an AI Tutor specializing in multilingual audio capabilities. This role involves curating and annotating audio data to enhance AI interactions globally. Candidates must have native Norwegian proficiency, a strong command of English...Hourly payRemote workFlexible hours- ...A leading online tutoring platform is seeking an Instant College Application Essays Tutor to provide immediate support to students through their Live Learning Platform. Candidates should have strong communication skills and expertise in college essays. The role allows...Immediate startRemote workFlexible hours
- An innovative educational platform is seeking an Entry Level Instant Tutor to provide on-demand HTML tutoring. This role offers competitive pay, flexible hours, and the opportunity to make an impact by helping students when they need it most. The ideal candidate will possess...Remote workFlexible hours
$35 - $45 per hour
...A tech company is seeking an AI Tutor specialized in multilingual audio capabilities to enhance voice interactions for its AI system. The role involves curating and annotating audio data and requires native proficiency in Hebrew and strong English skills. Candidates will...Hourly payRemote workFlexible hours$35 - $45 per hour
...xAI is seeking an AI Tutor specialized in multilingual audio to enhance Grok's voice interactions. The role involves curating and annotating audio data to improve speech recognition globally. Candidates should have native proficiency in Italian and be proficient in English...Hourly payRemote workFlexible hours- ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise...Hourly payWork at officeRemote work
$37.44k
...We are looking for AI Linguistic Evaluators to assess how effectively an AI application performs in Bengali for a short collaboration. Job Type: Freelance Location: Remote from the United States Work Schedule: Flexible Start Date: Immediately...Temporary workFreelanceImmediate startRemote workFlexible hours- MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers...Remote jobContract workTemporary workImmediate start
$20 - $30 per hour
A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong...Remote jobHourly pay$30 per hour
...Location: Remote Commitment: 10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous agent systems... ...stakeholders. Requirements Strong experience in LLM evaluation, AI output analysis, QA/testing, UX research, or similar analytical...Remote jobHourly payContract work- BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance...Remote job
$20 per hour
A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI... ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with...Hourly payRemote workFlexible hours- Position Summary In this remote, hourly contractor role, you will evaluate AI-generated financial content and develop cases that test analytical reasoning accuracy. Your work directly improves how leading AI models handle financial information, making them more accurate...Remote jobHourly payFull timeFor contractors
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Tutor Conversation AI Evaluator [Remote]. Be the first to apply!



