Tutor Conversation AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Tutor Conversation AI Evaluator is a remote review track for evaluating AI outputs across tutor conversation ai specialist operations workflows. Reviewers grade workflow correctness, policy adherence, and stakeholder fit; flag operational risk; and document the right next step so the modeling team can train on it.
Why this role matters
Tutor Conversation AI specialist operations AI has to fit into an actual day at work. AuraOne uses experienced operators to grade outputs the way a senior peer would — checking workflow, policy, and the unwritten rules that decide whether a task actually gets done.
Responsibilities
- Review AI outputs against current tutor conversation ai specialist operations workflows, playbooks, and firm policy for Tutor Conversation AI Evaluator assignments.
- Grade tone, escalation logic, and stakeholder fit on a structured rubric.
- Flag operational risk, missed escalations, and policy-adherence gaps with severity tags.
- Capture the right next step so the modeling team can train on it.
- Adjudicate disputed treatments against published playbooks or firm guidance.
- Maintain reviewer-quality scores in inter-rater calibration cycles.
Qualifications
- Direct working experience in tutor conversation ai specialist operations on real teams for Tutor Conversation AI Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the policy or workflow being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Grade a model's response to a real tutor conversation ai specialist operations ticket and rate workflow, tone, and escalation.
- Flag a missed escalation with the right severity tag and corrected next step.
- Adjudicate a disputed playbook call between two reviewers using firm guidance.
- Audit a 25-row batch for rubric consistency and report drift to the program lead.
Nice to have
- Prior experience training, calibrating, or QA-ing operations teams.
- Familiarity with AI-assisted workflow tooling and its failure modes.
- Bilingual experience for cross-region operations.
Skills
- Operational review
- Policy adherence
- Workflow judgment
- Stakeholder communication
- Tutor Conversation AI specialist operations
- Learning design
- Assessment review
- Pedagogy
- Tutor
- Conversation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- ...counseling, communication, qualitative research, HCI, conflict resolution, or related advisory disciplines to support an AI conversation-evaluation project. The role involves reviewing conversations between people and AI systems, assessing whether the AI gathered...SuggestedContract work
- ...Conversational AI Evaluator is a remote evaluation track for reviewing conversational ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...An innovative education provider seeks a freelance teacher to guide students in Conversation Techniques. In this flexible role, you will craft engaging course materials and support students through their learning journey. With the freedom to shape your curriculum and...SuggestedFreelanceRemote workFlexible hours
- ...Hiring PhD-level quantitative finance experts, the full-time Remote AI Research Evaluator will assess AI-generated financial content, craft relevant questions, and evaluate model responses in a flexible contract role. Key responsibilities Assessing the factuality and...SuggestedFull timeContract workRemote workFlexible hours
$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...SuggestedRemote workFlexible hours- ...Alignerr is seeking a Search Quality Evaluator to assess search engine results, AI-generated answers, and content recommendations. This fully remote, flexible contract role values strong critical thinking and quality instincts, with no technical background required beyond...Contract workRemote workFlexible hours
- ...'s multilingual audio capabilities, the part-time or full-time AI Tutor - Tamil will curate and annotate audio data, ensuring high-quality... ...Demonstrated ability to handle multilingual audio content and evaluate speech accuracy Experience in providing high-quality voice...Full timePart timeRemote work
- An innovative educational platform is seeking an Entry Level Instant Tutor to provide on-demand HTML tutoring. This role offers competitive pay, flexible hours, and the opportunity to make an impact by helping students when they need it most. The ideal candidate will possess...Remote workFlexible hours
$100 per hour
...looking for a highly experienced software engineer (SR+) to help evaluate the quality of interactions with modern coding agents such as... ...whether the model thinks like a great engineer. You will assess how AI coding agents behave in real-world scenarios — focusing on:...Contract workImmediate start$35 - $45 per hour
...AI Tutor - Polish Remote SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity... ...ability to handle multilingual audio content, including evaluating speech accuracy, cultural vocal expressions, and contextual interpretation...Full timePart timeFor contractorsRemote workWorldwide10 hours per week$35 - $45 per hour
...A tech company is seeking an AI Tutor specialized in multilingual audio capabilities to enhance voice interactions for its AI system. The role involves curating and annotating audio data and requires native proficiency in Hebrew and strong English skills. Candidates will...Hourly payRemote workFlexible hours$35 - $45 per hour
...A forward-thinking tech company is seeking an AI Tutor specializing in multilingual audio capabilities. This role involves curating and annotating audio data to enhance AI interactions globally. Candidates must have native Norwegian proficiency, a strong command of English...Hourly payRemote workFlexible hours$35 - $45 per hour
xAI is seeking an AI Tutor specialized in multilingual audio to enhance Grok's voice interactions. The role involves curating and annotating audio data to improve speech recognition globally. Candidates should have native proficiency in Italian and be proficient in English...Hourly payRemote workFlexible hours$35 - $45 per hour
...A technology company is seeking an AI Tutor specializing in multilingual audio capabilities. This role involves training AI for voice interactions and improving speech recognition across various languages. The ideal candidate will have native proficiency in Korean and...Hourly payRemote workFlexible hours- A leading online tutoring platform is seeking an Instant College Application Essays Tutor to provide immediate support to students through their Live Learning Platform. Candidates should have strong communication skills and expertise in college essays. The role allows...Immediate startRemote workFlexible hours
$35 - $45 per hour
...A leading tech firm in the United States is seeking an AI Tutor specializing in multilingual audio capabilities to enhance AI interactions. Candidates should have native proficiency in Chinese and proficient English, along with strong auditory perception. The role involves...Hourly payRemote workFlexible hours- ...Prolific is seeking talented Product Designers and UX Specialists to join our Expert Network for training and evaluation of cutting-edge AI models. Candidates should have a relevant educational background and at least one year of experience in design-related fields. Responsibilities...Remote workWork from homeFlexible hours
- ...Prolific is seeking Product Designers and UX Specialists to join our Expert Network, contributing to the training and evaluation of cutting-edge AI models. This role requires expertise in usability, design systems, and user research, providing essential feedback on design...Remote workWork from homeFlexible hours
- ...AI Policy Compliance AI Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack... ...Score model defenses across single-turn and multi-turn conversations. Triage emerging attack vectors and route them to the...Remote jobHourly payFor contractors10 hours per week
$60 - $70 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...$60–$70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance...Contract workSummer workRemote work$80 - $120 per hour
Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks. ~Provide clear, structured written feedback to improve AI outputs. ~Apply...Part timeWork at officeRemote work- ...Remote | Work from Home Employment Type: Project-based | Contract We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists with strong proficiency in English. In this role, you will support AI/ML...Contract workRemote workWork from homeMonday to FridayDay shift
$50 - $190 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Remote Commitment: 20+ hours/week Role Responsibilities Evaluate AI systems on complex personal workflows, including personal...Hourly payContract workFor contractorsSummer workRemote workTrial period- ...Summary This is a fully remote, hourly contractor role supporting AI data and language projects on a project-based, flexible hour... ..., and other content to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy, reasoning quality...Hourly payFor contractorsRemote workFlexible hours
$90k - $200k
...About xAI xAI’s mission is to create AI systems that can accurately understand the... ...teammates. About the Role As an AI Tutor specialized in personality and behavior,... ...consistency, deliver sharp humor, excel in casual conversations across diverse contexts, and infuse...Remote jobHourly payPart timeCasual workWork at office$50 - $175 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Remote Commitment: 20–40 hours/week Role Responsibilities Evaluate JSON-formatted AI model outputs against defined grading...Contract workSummer workStart working todayRemote work$20 - $80 per hour
...Role Overview Train and evaluate next-generation AI systems by scoring model outputs, annotating real-world content, and delivering clear, actionable feedback that improves model accuracy and reasoning across diverse domains. About the company micro1 is an AI data...Hourly payFor contractorsRemote work- ...seeking detail-oriented human reviewers with a strong understanding of their local cultural context to support a range of AI training and evaluation projects . In this role, you will work across diverse task types, including evaluating prompts and AI-generated...Extra incomeFull timeFor contractorsLocal area10 hours per week
$70 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...: $70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses to strengthen reasoning and rigor in model...Contract workSummer workRemote work- ...We are seeking native German speakers to help evaluate and improve the next generation of AI-powered voice assistants in collaboration with a leading... ...required application form. Participate in a 25-minute conversational interview discussing your background, experience,...Contract workFor contractorsH1bImmediate startRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Tutor Conversation AI Evaluator [Remote]. Be the first to apply!







