Indonesian Bilingual Expert — AI Evaluation & Annotation (Remote)
AuraOne Human Data
- Remote job
Indonesian Bilingual Expert — AI Evaluation & Annotation (Remote) is a remote evaluation track for reviewing indonesian generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn indonesian generalist evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate indonesian generalist evaluation model outputs against a versioned rubric and assign severity tags for Indonesian Bilingual Expert — AI Evaluation & Annotation (Remote) assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on indonesian generalist evaluation or adjacent content for Indonesian Bilingual Expert — AI Evaluation & Annotation (Remote) work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two indonesian generalist evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Indonesian generalist evaluation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$60 - $65 per hour
...Role Overview Help improve next-generation AI language systems by evaluating AI-generated Indonesian speech and providing linguistically informed feedback. This remote contractor role focuses on the quality, authenticity, and cultural appropriateness of Indonesian audio...Remote workBilingualIndonesian language skillsHourly payFor contractors$11 - $30 per hour
...Evaluate AI-generated music and lyrics in Indonesian and English, helping assess outputs across a broad range of genres... ...subgenres, and artists. Work Terms Remote, hourly engagement with an... ...the application, including a bilingual Indonesian competency interview....Remote workBilingualIndonesian language skillsHourly payImmediate startFlexible hours$11 - $30 per hour
...Role Overview Evaluate generative music AI outputs across diverse genres... ...quality in Indonesian and English. Work directly... ...Terms Location: Remote. Employment type:... ...Commitment: flexible. Most experts work around 20 hours... ...steps, including a Bilingual Competency Interview...Remote workBilingualIndonesian language skillsHourly payContract workImmediate startFlexible hours- ...train next-generation AI systems. Your work will... ..., axes, legends, and annotations in all outputs. # Create... .... # Participate in remote collaboration,... ...rigorous training and evaluation of advanced AI models.... ...and verbal English. # Bilingual proficiency in English...Remote workBilingualTemporary work
$60 - $100 per hour
...Legal Domain Expert — AI Training & Evaluation is a remote review track for evaluating AI outputs in legal review workflows. Reviewers grade citation accuracy... ...AI-assisted legal-research or compliance tooling. Bilingual experience for cross-jurisdiction matters. Skills...Remote jobBilingualFor contractors10 hours per week$70 - $100 per hour
...Bioinformatics & Single-Cell Genomics Expert — AI Evaluation is a remote review track for evaluating AI outputs across bioinformatics computational... ...AI-assisted workflow tooling and its failure modes. Bilingual experience for cross-region operations. Skills Operational...Remote jobBilingualFor contractorsWork experience placement10 hours per week$30 - $65 per hour
...accurate, high-quality data that supports AI and machine learning projects. In this remote contract role, you will help refine language resources, evaluate linguistic quality, and contribute... ...nuance. Key Responsibilities Annotate, label, and review Portuguese-...Remote workHourly payContract work$17 - $25 per hour
...technical talent with leading AI research labs.... .... Position: AI Safety Experts — English & Indonesian Type: Contract Compensation... ...7–$25/hour Location: Remote Role Responsibilities... ...Generate high-quality human data. Annotate failures, classify...Remote workIndonesian language skillsContract workSummer work- ...solving to improve and evaluate large language models.... ...Accelerate frontier AI research by contributing... ...training pipelines and expert researchers who specialize... ...and steps. Review and annotate model generated... ...work independently in a remote setting. Technical requirements...Remote workContract workFor contractorsFreelance
- Mercor is seeking a remote Senior Red Team AI Specialist to test conversational agents and AI models against adversarial inputs. You will annotate vulnerabilities, surface systemic risks, and produce reproducible attack cases to help customers strengthen their AI systems...Remote workBilingual
$30 - $65 per hour
...Role Overview Help improve next-generation AI systems by applying your Vietnamese language expertise to evaluate real-world audio and provide precise linguistic feedback. This remote contractor role focuses on how naturally and accurately Vietnamese is spoken and understood...Remote workBilingualHourly payFor contractors$30 - $65 per hour
...Role Overview Help improve next-generation AI systems by applying your Tamil language expertise to real-world audio evaluations and written feedback. This remote contractor role focuses on assessing how naturally and accurately Tamil is spoken and generated. No previous...Remote workBilingualHourly payFor contractors- A leading AI Data Services company is seeking a bilingual content evaluator to review AI-generated responses and create training content. The ideal candidate will have... ..., and optimizing AI performance. This fully remote role offers flexible hours and requires at least...Remote jobBilingualFlexible hours
$30 - $65 per hour
...Help improve next-generation AI systems by applying your Turkish language expertise to real-world audio evaluations and written feedback. This remote contractor role focuses on assessing how naturally and accurately Turkish is spoken and generated. Key Responsibilities...Remote workBilingualHourly payFor contractors$35 - $42 per hour
Location: USA (Remote) Compensation: $35.00-$42.00 USD... ...Shape the Future of AI in Finance, Audit & Risk... ...growing network ofon-call AI Evaluation Specialistssupporting... ...of subject matter experts. As client projects become... ...learning, or data annotation projects Compensation...Remote workHourly payExtra incomePermanent employmentContract workTemporary workFlexible hours- A growing AI Data Services company is seeking a contractor for a fully remote role focusing on reviewing and generating high-quality bilingual training content. You will create and evaluate AI responses, ensuring accuracy and clarity in Hebrew and English. The ideal candidate...Remote jobBilingualFor contractorsFlexible hours
$35 - $62 per hour
...Role Overview Evaluate and rate AI-generated music across a wide... ...will compare songs, annotate musical attributes,... ...Work Terms Location: Remote. Employment type:... ...: flexible. Most experts work around 20 hours... ...steps, which include a bilingual competency interview...Remote workBilingualHourly payImmediate startFlexible hours$14 - $42 per hour
...Role Overview Evaluate AI-generated music across diverse... ...or mix quality. Annotate and label tracks for genre... ...Terms Location: Remote. Employment type: hourly... ...: flexible, most experts work around 20 hours per... ...steps, including a Bilingual Competency Interview conducted...Remote workBilingualHourly payImmediate startFlexible hours$15 - $19 per hour
...Role Overview Evaluate AI-generated music in Tamil and... ...compare model outputs, annotate musical characteristics... .... Work Terms Remote, hourly engagement. Start... ...commitment is flexible; most experts work around 20 hours... ...tasks, including a Bilingual Competency Interview...Remote workBilingualHourly payImmediate startFlexible hours- Investment Banking AI Subject-Matter Expert is a remote review track for evaluating AI outputs across banking workflows. Reviewers grade calculations, narrative... ..., audit, or risk tooling and its failure modes. Bilingual experience for cross-jurisdiction reviews. Skills...Remote workBilingualHourly payFor contractorsWork experience placement10 hours per week
$35 - $62 per hour
...Role Overview Evaluate AI-generated music across many... ...production and mix quality. Annotate tracks with detailed... ...Terms Location: Remote. Start date:... ...Commitment: flexible, most experts work around 20 hours per... ..., which include a bilingual competency interview conducted...Remote workBilingualHourly payImmediate startFlexible hours$15 per hour
...technical talent with leading AI research labs.... ...Position: Music & Lyrics Expert - Malayalam Type:... ...5/hour Location: Remote Duration: Up to 6... ...Role Responsibilities Evaluate AI-generated music across... .... Complete the Bilingual Competency Interview in...Remote workBilingualContract workSummer workImmediate startFlexible hours$25 - $30 per hour
...Bilingual Traditional Chinese AI Evaluation Specialist is a remote Chinese specialist track for evaluating chinese evaluation outputs against native-speaker standards... ...10 hours per week. Prior model-evaluation, annotation, or human-rater experience is a plus. Example...Remote workBilingualFor contractors10 hours per week$42 - $78 per hour
...Role Overview Evaluate and rate AI-generated music to help improve... ...when rating and annotating audio Qualifications... ...Terms Location, remote Employment type, hourly... ..., flexible; most experts work around 20 hours... ...application steps, including a Bilingual Competency Interview...Remote workBilingualHourly payImmediate startFlexible hours$150 - $180 per hour
Prolific in Sacramento is seeking Mental Health Professionals to help train and evaluate advanced AI models. You will review AI-generated responses and engage in various training tasks, earning competitive pay rates up to $150-180/hr. Ideal candidates must hold a verified...Remote jobWork from home$150 - $180 per hour
...Virginia Beach is seeking Mental Health Professionals to train and evaluate AI models. The role involves reviewing AI responses to... ...attention to detail, and a reliable internet connection. Join our Expert Network to influence future AI innovations and work flexibly from...Remote jobWork from home- ...Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant...Remote workFlexible hours
$8 - $65 per hour
Prolific is hiring Mental Health Professionals in New York to train and evaluate AI models. As a Domain Expert, you will be responsible for reviewing AI-generated responses, completing tasks related to psychology, and improving AI models based on your expertise. Pay rates...Remote jobHourly payWork from homeFlexible hours- ...Prolific is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities...Remote workHourly payFlexible hours
- Prolific in New York, NY, is seeking Chemistry Experts and Chemical Engineers to join its Expert Network. In this role, you will help train and evaluate AI models using your chemical expertise. Duties include evaluating AI-generated responses for accuracy and validating...Remote jobWork from homeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Indonesian Bilingual Expert — AI Evaluation & Annotation (Remote). Be the first to apply!
- technology expert Remote
- subject matter expert Remote
- fulfillment expert Remote
- guest service support expert Remote
- indonesian Remote
- bilingual spanish receptionist Remote
- bilingual spanish virtual assistant Remote
- customer service bilingual Remote
- bilingual nurse practitioner Remote
- bilingual work from home Remote



