Pronunciation Conversation Evaluator [Remote]
AuraOne Human Data
- Remote job
Pronunciation Conversation Evaluator is a remote evaluation track for reviewing pronunciation conversation evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn pronunciation conversation evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate pronunciation conversation evaluation model outputs against a versioned rubric and assign severity tags for Pronunciation Conversation Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on pronunciation conversation evaluation or adjacent content for Pronunciation Conversation Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two pronunciation conversation evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Pronunciation Conversation evaluation
- Speech evaluation
- Voice QA
- Audio review
- Pronunciation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$15 - $20 per hour
...identifies strengths and specific areas for improvement. Your evaluations will be used to help create the "perfect AI-generated... ...responses. Verify that model responses align with expected conversational behavior and applicable system guidelines. Write concise,...SuggestedHourly payContract workFor contractorsRemote work$40 per hour
...AI models. You will measure the progress of these AI chatbots, evaluate their logic, and solve problems to improve the quality of each... ...ask for any money from you. PayPal will handle any currency conversions from USD. Only applicants in the United States will be considered...SuggestedHourly payFull timeContract workPart timeRemote work- ...Tutor Conversation AI Evaluator is a remote review track for evaluating AI outputs across tutor conversation ai specialist operations workflows. Reviewers grade workflow correctness, policy adherence, and stakeholder fit; flag operational risk; and document the right next...SuggestedRemote jobHourly payFor contractorsWork experience placement10 hours per week
$15 - $20 per hour
...sources and external tools. ~Generate high-quality human evaluation data by identifying response strengths, areas for improvement... ...of responses. ~Ensure model responses align with expected conversational behavior and system guidelines. ~Work independently and asynchronously...SuggestedPart timeSummer work$25 per hour
...Language Specialist to work on training AI models in Washington, D.C. You will be responsible for evaluating the performance of AI chatbots through engaging in conversations and providing feedback. Candidates should be fluent in both English and Korean, possess a strong...SuggestedRemote jobHourly payContract workFlexible hours$25 per hour
A leading tech company is seeking a Language Specialist to improve AI chatbots by evaluating their progress and teaching them conversational skills in English and Japanese. This role allows flexibility with remote work and the choice of projects. Candidates should have...Hourly payRemote work$40 per hour
...AI models. You will measure the progress of these AI chatbots, evaluate their logic, and solve problems to improve the quality of each... ...ask for any money from you PayPal will handle any currency conversions from USD Only applicants in the United States will be...Hourly payFull timeContract workPart timeRemote work$20 per hour
...public sources and external tools. Generate high-quality human evaluation data by identifying response strengths, areas for improvement... ...of responses. Ensure model responses align with expected conversational behavior and system guidelines. Work independently and asynchronously...Remote jobContract workPart timeSummer work$25 per hour
...proficient in both Korean and English. This position involves training AI chatbots by measuring their progress and writing new conversations. Ideal candidates are detail-oriented with excellent grammar skills. This independent contract role offers flexibility with project...Remote jobHourly payContract work$40 per hour
...AI models. You will measure the progress of these AI chatbots, evaluate their logic, and solve problems to improve the quality of each... ...ask for any money from you. PayPal will handle any currency conversions from USD. Only applicants in the United States will be considered...Hourly payFull timeContract workPart timeRemote work$50 - $60 per hour
...Responsibilities Give AI chatbots diverse and complex problems and evaluate their outputs Evaluate the quality produced by AI models... ...ask for any money from you. PayPal will handle any currency conversions from USD. Only applicants in United States will be considered...Hourly payFull timeContract workPart timeWork experience placementRemote workFlexible hours$19 per hour
...is building realistic, high-fidelity simulated environments to evaluate and train AI models on real-world procurement workflows for a... ...19.00 hourly Temporary Seasonal (with potential for permanent conversion).Monday – Friday, 7:30 AM – 4:30 PM.Occasional Saturday work may...Hourly payPermanent employmentFull timeTemporary workCasual workSeasonal workWork at officeRemote workWork from homeMonday to Friday$20 per hour
...write high-quality responses to demonstrate excellence, and evaluate different model outputs based on accuracy and style guidelines... ...high-quality work Responsibilities Come up with diverse conversations over a range of topics Write high-quality answers when...Hourly payFull timeContract workPart timeFor contractorsSelf employmentFreelanceRemote work$9 - $30 per hour
ALTA Language Services, Inc. is looking for a remote Testing Evaluator in Atlanta, GA. This part-time position involves assisting with language proficiency examinations and evaluating tests according to set criteria. Candidates should have strong communication skills, professionalism...Part timeRemote work$40 per hour
A leading data annotation service is seeking a Biology Expert to train AI models. The role involves measuring AI progress, evaluating chatbot outputs, and ensuring quality improvement. Ideal candidates will have a strong background in biology or biochemistry. This independent...Hourly payContract workRemote workFlexible hours$28.26 - $42.4 per hour
Role Description The Evaluator applies expertise in the relevant content area to review students’ assessments and determine mastery of learning objectives and course competencies. The Evaluator provides accurate and helpful feedback on assessments to help students develop...Part timeWork experience placementFlexible hoursAfternoon shift- ...Blueprint Technologies, LLC. is seeking a detail-oriented Labeler / Annotator to evaluate AI responses in French. This remote role focuses on side-by-side evaluation across real-world scenarios, not translation, requiring strong judgment and attention to detail. You...Remote work
$80 - $120 per hour
Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks. ~Provide clear, structured written feedback to improve AI-generated work...Part timeWork at officeRemote work$16 - $20 per hour
...The research team at Penn State Ross and Carol Nese College of Nursing is hiring part-time research evaluators for projects focused on dementia care in assisted living settings. The research evaluator will assist with in-person recruitment, consent, data collection, and...Hourly payPart timeSummer work- ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers...Contract workTemporary workImmediate startRemote work
$40 per hour
A technology company specializing in AI is seeking a Physics Graduate Assistant to enhance AI models. You will evaluate AI chatbot outputs and provide complex physics problems. Ideal candidates will have a strong understanding of classical mechanics, E&M, thermodynamics...Hourly payFull timePart timeRemote work$40 per hour
...company is seeking a Physics Graduate Assistant to join their team in a remote position. The role involves training AI models by evaluating their outputs and improving their quality through complex physics problems. Candidates should have a strong grasp of classical mechanics...Hourly payContract workRemote work$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...Remote workFlexible hours$55 per hour
...Charter Trauma Clinic, a division of The Key Program, Inc., is seeking to hire a part-time (15-20 hrs. per week) fee-for-service evaluator. Founded in 1985, the Forensic Services Team has a long-standing reputation as the leading provider of Child Trauma and Parental Capacity...Hourly payPart timeRemote workFlexible hours- ...alongside iconic names like Louis Vuitton, Dolce & Gabbana, Bentley, Prada, Versace, and more. About the Role As a luxury brand evaluator, you will step into the world of luxury to discreetly assess customer experiences, providing critical feedback that helps brands...Contract workWorldwideFlexible hours
$20 per hour
A tech company specializing in AI is looking for a Digital Web Designer to evaluate AI-generated designs and help train models for better aesthetic understanding. This is a remote position open solely to applicants in the United States. The role involves assessing UI/UX...Remote work$20 per hour
A tech company specializing in AI is seeking an Experience Designer to evaluate and enhance AI-generated UI/UX designs. This remote position allows for flexible scheduling and offers competitive pay rates starting at $20/hour for general projects and $40/hour for design...Remote workFlexible hours$45 - $85 per hour
...Ixolabs is seeking an Animation Quality Evaluator to refine AI-generated animations. The ideal candidate will evaluate motion fluidity and naturalness, ensuring animations convey believable narratives. The role is part-time (15-25 hours/week) and offers competitive compensation...Part timeRemote work- ...Seeking a PhD-level expert in Quantitative Finance, the full-time Remote AI Research Evaluator will assess AI-generated financial content, craft relevant questions, and evaluate responses, contributing to the training of advanced AI models in a flexible contract role....Full timeContract workRemote workFlexible hours
$40 per hour
...We are looking for experienced cybersecurity professionals to join our team to help train AI models. In this role, you will evaluate AI-generated security content, solve technical cybersecurity problems, and provide feedback to improve how AI systems reason about real...Hourly payFull timePart timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Pronunciation Conversation Evaluator [Remote]. Be the first to apply!




