Hebrew Music & Lyrics Expert - AI Output Evaluation [Remote]
$17 - $42 per hourAuraOne Human Data
- Remote job
Hebrew Music & Lyrics Expert - AI Output Evaluation is a remote evaluation track for reviewing hebrew generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn hebrew generalist evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate hebrew generalist evaluation model outputs against a versioned rubric and assign severity tags for Hebrew Music & Lyrics Expert - AI Output Evaluation assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on hebrew generalist evaluation or adjacent content for Hebrew Music & Lyrics Expert - AI Output Evaluation work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two hebrew generalist evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Hebrew generalist evaluation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
$17–$42 / hr
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$17 - $42 per hour
...Role Overview You will evaluate AI-generated music and lyrics across a wide range of genres, rating outputs against detailed quality standards... ...and requires assessments in Hebrew and English. Key... ...Commitment: flexible, most experts work around 20 hours per week...Hebrew language skillsHourly payImmediate startRemote workFlexible hours$17 - $45 per hour
...Turkish Music & Lyrics Expert - AI Audio Evaluation is a remote review track for evaluating AI outputs across music lyrics turkish specialist operations workflows. Reviewers grade workflow correctness, policy adherence, and stakeholder fit; flag operational risk; and...SuggestedRemote jobFor contractorsWork experience placement10 hours per week- Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback. Candidates should have at least 5 years of relevant experience...SuggestedRemote jobHourly payWork at office
$13 - $17 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Summers , and Jack Dorsey . Position: Music & Lyrics Expert - English (India) Type: Contract... ...hours/week Role Responsibilities Evaluate AI-generated music across various...SuggestedContract workSummer workImmediate startRemote work$17 - $42 per hour
...Role Overview Evaluate AI-generated music across many genres for a leading AI research partner... ...and label musical attributes and lyrical accuracy in Hebrew and English to help improve generative... ...Commitment: Flexible, most experts work around 20 hours per week, there...Hebrew language skillsHourly payImmediate startRemote workFlexible hours$39 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ..., and Jack Dorsey . Position: Music Audio Expert - Swedish Type: Contract... ...hours/week Role Responsibilities Evaluate AI model output lyrics, voice generation, and other standards...Contract workSummer workImmediate startRemote work$60 - $75 per hour
Role Description Join an advanced AI research initiative focused on improving how next... ...design high-quality benchmark tasks that evaluate AI performance across software... ..., and generate accurate, well-structured outputs. This is a fully remote, independent contractor...Weekly payContract workPart timeFor contractorsRemote workFlexible hours- A leading research services provider is seeking a Chemistry Expert (PhD) to evaluate complex chemistry problems and review AI-generated outputs for accuracy. This remote, hourly contract role requires deep subject-matter expertise and excellent communication skills. Candidates...Remote jobHourly payContract workFlexible hours
- ...experienced finance, audit, insurance, compliance, and legal professionals to join our on-call AI Evaluation Specialists network. This remote, project-based role involves evaluating AI outputs, applying critical thinking, and providing high-quality feedback across multiple...Remote jobFlexible hours
- AuraOne is seeking a Rubrics Expert to remotely review AI outputs and evaluate them against established rubrics, policies, and workflows. You will identify gaps, rate tone and escalation logic, and document next steps to guide model training. As a remote contractor eligible...Remote jobFor contractors10 hours per week
- A leading AI Data Services company is seeking a bilingual content evaluator to review AI-generated responses and create training content. The ideal candidate will have native or near-native Hebrew proficiency, minimum C1 English proficiency, and a bachelor's degree in...Hebrew language skillsRemote jobFlexible hours
- A growing AI Data Services company is seeking a contractor for a fully remote role focusing on reviewing... ...quality bilingual training content. You will create and evaluate AI responses, ensuring accuracy and clarity in Hebrew and English. The ideal candidate should possess a...Hebrew language skillsRemote jobFor contractorsFlexible hours
$14 per hour
...and technical talent with leading AI research labs. Headquartered in San... ...and Jack Dorsey . Position: Music Audio Expert - Telugu Type: Contract Compensation... .../week Role Responsibilities Evaluate AI model output lyrics , voice generation, and other...Contract workSummer workImmediate startRemote work$80 - $120 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Position: Biology / environmental science Evaluator Type: Contract Compensation: $... ...structured written feedback to improve AI model outputs . Review and assess documents,...Contract workSummer workWork at officeRemote work$80 - $120 per hour
Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation... ...clear, structured written feedback to improve AI model outputs. ~Apply deep subject-matter expertise in Clinical / biomedical...Part timeWork at officeRemote work$60 - $70 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Location: Remote Role Responsibilities Evaluate AI-generated responses for safety,... ...safety benchmarking . Identify unsafe outputs, hallucinations, reasoning failures, and...Contract workSummer workRemote work$30 per hour
...Prolific is not just another player in the AI space - we are building the biggest... ...Visual Designers to act as Domain Experts for a high-level AI evaluation project. AI models are evolving... ...and structural integrity of design outputs. This is a flexible, ad hoc project...Remote jobWork from homeFlexible hours- ...Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant...Remote workFlexible hours
- ...A leading AI research firm is seeking Expert Prompt Curators to design challenging prompts for evaluating advanced AI models. The role requires advanced knowledge in diverse fields and offers flexible hours, remote work, and a competitive hourly wage. Ideal candidates...Hourly payTemporary workRemote workFlexible hours
- Prolific is seeking Mental Health Professionals in Indianapolis, Indiana, to assist in training and evaluating AI models. Ideal candidates will have a verified status as a Mental Health Professional, an understanding of psychological theory, and the ability to focus on...Remote jobHourly payFlexible hours
- A leading AI research accelerator is seeking an Associate to leverage expertise in internal or emergency medicine. This role involves designing and evaluating clinical scenarios to enhance AI diagnostics. Candidates should have an MD and active practice within the last...Remote job
- Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. This role involves designing statistics problems, evaluating...Permanent employmentTemporary workPart timeRemote work
$80 - $150 per hour
...leading behavioral health consulting firm is seeking Senior Behavioral Health Experts to work part-time and remotely on frontier AI research projects. You will be responsible for designing evaluations and testing AI systems in critical mental health contexts. The ideal...Hourly payPart timeRemote work$30 per hour
Prolific is seeking fluent Hindi speakers to join their Expert Network, helping to train and evaluate AI models with real legal expertise. Responsibilities include analyzing and writing tasks in Hindi, judging AI’s performance, and aiding in the improvement of AI models...Remote jobHourly payFlexible hours$8 - $65 per unit
Prolific is seeking Mental Health Professionals to train and evaluate AI models. In this role, you will review AI responses, analyze psychology-related content, and provide feedback to improve models. Applicants must have verified status as a Mental Health Professional...Remote jobFlexible hours$150 per hour
...Senior Design Expert — Paid AI Design Evaluation Study is a remote review track for evaluating AI outputs across senior design paid ai design research study creative review workflows. Reviewers grade craft, constraints, and tooling adherence; flag production-readiness...Remote jobFor contractors10 hours per week- Mercor, in partnership with Crossing Hurdles, is looking for a Multimedia Expert to work remotely on project-based tasks. The role involves reviewing and editing multimedia outputs, creating assets for AI model training, and providing feedback to ensure quality standards....Remote jobPart time10 hours per week
$8 - $65 per hour
Prolific is hiring Mental Health Professionals in New York to train and evaluate AI models. As a Domain Expert, you will be responsible for reviewing AI-generated responses, completing tasks related to psychology, and improving AI models based on your expertise. Pay rates...Remote jobHourly payWork from homeFlexible hours$60 per hour
Prolific in Seattle, WA is searching for Chemistry Experts and Chemical Engineers to join our Expert Network to train AI models with your expertise. The role involves evaluating AI-generated chemistry, fact-checking chemical reactions, and auditing technical documentation...Remote jobHourly payWork from homeFlexible hours$8 - $65 per hour
Prolific is seeking Mental Health Professionals to train and evaluate AI models from home. Responsibilities include reviewing AI responses and enhancing model performance using psychological expertise. Ideal candidates will have verified professional status, a solid understanding...Remote jobFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Hebrew Music & Lyrics Expert - AI Output Evaluation [Remote]. Be the first to apply!






