Call Center QA Conversation Evaluator [Remote]
AuraOne Human Data
- Remote job
Call Center QA Conversation Evaluator is a remote evaluation track for reviewing call center qa conversation evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn call center qa conversation evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate call center qa conversation evaluation model outputs against a versioned rubric and assign severity tags for Call Center QA Conversation Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on call center qa conversation evaluation or adjacent content for Call Center QA Conversation Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two call center qa conversation evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Call Center QA Conversation evaluation
- Speech evaluation
- Voice QA
- Audio review
- Call
- Center
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- Rex.zone is seeking a Senior AI data annotator to perform data labeling and evaluation for NLP tasks, RLHF assessments, and prompt QA to improve training data quality and model performance. This is a US-based remote, full-time role aligned with Miami talent demand. You...SuggestedRemote jobFull time
- YO AI Labs is seeking an experienced Adobe Marketing Technology Expert to support an AI training and evaluation project focused on enterprise marketing operations. The contractor will test marketing workflows using Adobe Workfront, AEM, CJA, Analytics, and Experience Cloud...SuggestedRemote jobFor contractors
- Dorado is seeking an AI Language Quality Evaluator fluent in Greek and English for an ongoing, task-based project. This remote freelance role... ...translations. You will work with a global team on localization QA tasks, with flexible hours and a 1099 contractor setup through...SuggestedRemote jobFor contractorsFreelanceFlexible hours
- YO AI Labs is seeking Polish bilingual experts for a remote, contract-based role focused on evaluating Polish audio content and providing detailed feedback to improve AI language models. The ideal candidate has native Polish proficiency, strong English communication skills...SuggestedRemote jobContract work
- YO AI Labs is seeking a remote Adobe Marketing Technology Expert to test and evaluate marketing workflows within enterprise environments. You will review workflows across Adobe Workfront, AEM, CJA, Analytics, and Experience Cloud, identifying issues and providing actionable...SuggestedRemote job
- YO AI Labs is seeking Tamil bilingual experts to contribute to language and AI training projects focused on evaluating Tamil audio quality and linguistic authenticity. You will assess nativeness, fluency, pronunciation, and overall accuracy, providing clear feedback in...Remote job
- ...Tamil bilingual experts to contribute to an AI training project focused on Tamil language understanding and generation. You will evaluate Tamil audio content, assess linguistic quality, and provide detailed feedback based on established guidelines. No prior AI experience...Remote job
- Unknown is seeking a Polish Bilingual Expert (contractor) to evaluate Polish audio content for AI training. The role emphasizes nativeness, fluency, and linguistic quality, with feedback delivered in English. Strong Polish language skills and clear written and verbal English...Remote jobFor contractors
- Adobe Marketing Technology Expert is sought to support an AI training and evaluation project focused on enterprise marketing operations. The role involves testing marketing technology workflows and evaluating AI-generated outcomes based on real-world experience. Responsibilities...Remote job
- YO AI Labs is seeking Polish bilingual experts to contribute to an AI training project. You will evaluate Polish audio content, assess nativeness, fluency, pronunciation, and linguistic authenticity, and provide clear feedback in English. No prior AI experience required...Remote job
$25 - $35 per hour
...Expert (Real Estate and Leasing) to work part-time and remotely. Responsibilities include writing prompts for real estate scenarios, evaluating AI outputs, and providing evidence-based evaluations. The ideal candidate has expertise in real estate or leasing, attention to...Remote jobHourly payPart time- Mercor is seeking expert Evaluators in Legal/compliance to review AI-generated documents, spreadsheets, and slide decks for accuracy and quality. This is a remote, hourly engagement requiring deep subject-matter expertise and careful judgment. You will assess artifacts...Remote jobHourly payWork at office
- YO AI Labs is seeking a Thai Bilingual Expert to evaluate Thai audio content for quality and authenticity on a remote basis. You will assess nativeness, fluency, pronunciation, and linguistic accuracy, while documenting feedback in English. No AI experience is required...Remote jobFor contractorsFlexible hours
- YO AI Labs seeks Hebrew bilingual experts to contribute to an AI training project by evaluating Hebrew audio content and assessing linguistic quality. You will provide detailed feedback on nativeness, fluency, pronunciation, and authenticity, with clear explanations in...Remote jobFor contractorsFlexible hours
- YO AI Labs is seeking Polish Bilingual Experts to contribute to an AI training project by evaluating Polish audio content and providing feedback on nativeness, fluency, and linguistic quality. You will document assessments in English, follow guidelines, and collaborate...Remote job
- YO AI Labs is seeking a Hebrew Bilingual Expert to contribute to a language and AI training project focused on Hebrew audio evaluation. You will assess nativeness, fluency, and linguistic quality, and provide feedback to help improve model understanding and generation....Remote jobContract work
- YO AI Labs is seeking a Dutch Bilingual Expert for a remote, contractor role. You will evaluate AI-generated Dutch speech, assess nativeness and accuracy, and provide detailed feedback to support language model improvements. No prior AI experience is required; strong Dutch...Remote jobFor contractorsFlexible hours
- ...to contribute to a language and AI training project focused on improving Italian-language understanding and generation. You will evaluate Italian audio content, assess nativeness and fluency, and provide detailed feedback based on established guidelines to help refine...Remote job
- ...counseling, communication, qualitative research, HCI, conflict resolution, or related advisory disciplines to support an AI conversation-evaluation project. The role involves reviewing conversations between people and AI systems, assessing whether the AI gathered enough...Contract work
- YO AI Labs is seeking Dutch bilingual experts to contribute to a global language and AI training project. You will evaluate AI-generated Dutch speech, assess nativeness and quality, and provide detailed feedback to help improve language models. This remote contractor role...Remote jobFor contractors
- Obsidian is seeking expert Evaluators in Biology/environmental science to review and assess AI-generated work products for accuracy and quality. In this remote, hourly role, you will leverage your expertise to provide feedback on documents and presentations, ensuring they...Remote jobHourly pay
- ...the Admissions Manager and admission staff, as well as nursing and other internal and external staff to facilitate the referral conversion. Qualifications Education and Training: Degree in a health-related field from an accredited college or university preferred...Full timeLocal area
- ...Tutor Conversation AI Evaluator is a remote review track for evaluating AI outputs across tutor conversation... .... Adjudicate a disputed playbook call between two reviewers using firm... ...Prior experience training, calibrating, or QA-ing operations teams. Familiarity with...Remote jobHourly payFor contractorsWork experience placement10 hours per week
- ...Tamil Bilingual Expert to contribute to a language and AI training project focused on Tamil understanding and generation. You will evaluate Tamil audio content, assess linguistic quality, and provide detailed feedback based on established guidelines. No prior AI...Remote jobFor contractors
$80 per hour
A technology company is seeking a part-time QA contributor for an AI project to validate autonomous agents. Candidates must have strong analytical thinking, attention to detail, and experience with logical problem-solving. The role includes reviewing tasks, identifying...Remote jobPart timeFlexible hours- YO AI Labs is seeking Thai bilingual experts to evaluate Thai audio content and provide detailed feedback to improve AI language models. You will assess nativeness, fluency, and linguistic authenticity while documenting evaluations in English. Strong Thai proficiency and...Remote jobFor contractors
- A virtual AI evaluation firm is seeking individuals to review and evaluate AI-generated responses in therapeutic conversations. The ideal candidate will possess strong written communication and analytical skills, as well as a keen attention to detail for assessing tone...Remote jobImmediate start
- ...Systems is seeking highly motivated and detail-oriented Clinical Evaluators to join our team on a part-time contract basis. You will play... ...by evaluating the quality and accuracy of healthcare conversations conducted by our AI agents. This hourly contract role offers...Remote jobHourly payContract workPart timeFlexible hours
$8 per hour
Dorado is seeking a Speech AI Evaluation Specialist to help improve AI-generated content in Thai or Chinese Simplified. This freelance... ...and a rate of 8 USD/hour. You will engage in short voice conversations with AI models, follow prompts, and rate AI responses for quality...Remote jobPart timeFreelanceImmediate start- Dorado is seeking Speech AI Evaluation Specialists based in Thailand to help improve Vietnamese-language AI content. This is a remote... ...flexible, part-time schedule. You will participate in short voice conversations with AI models, follow given prompts, and evaluate responses...Remote jobPart timeFreelanceImmediate start10 hours per weekFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Call Center QA Conversation Evaluator [Remote]. Be the first to apply!

