Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Arabic Localization AI Evaluator [Remote]

AuraOne Human Data

Remote
  • Remote job

Arabic Localization AI Evaluator is a remote evaluation track for reviewing arabic generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.

Why this role matters

AI data reviewers help turn arabic generalist evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.

Responsibilities

  • Evaluate arabic generalist evaluation model outputs against a versioned rubric and assign severity tags for Arabic Localization AI Evaluator assignments.
  • Compare paired responses and pick the stronger answer with a written rationale.
  • Label hallucinations, instruction-following failures, and unsafe content with structured tags.
  • Capture ambiguous prompts and route them back to the program team for rubric updates.
  • Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
  • Document recurring failure modes so the modeling team can target them in the next training run.

Qualifications

  • Prior evaluation, annotation, or human-rater experience on arabic generalist evaluation or adjacent content for Arabic Localization AI Evaluator work.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that names the issue and the rubric clause being applied.
  • Strong attention to detail and the ability to flag when a prompt itself is the problem.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Compare two arabic generalist evaluation model responses to the same prompt and pick the stronger one with rationale.
  • Tag an unsafe response with the correct policy category and severity.
  • Audit a 50-row batch for rubric consistency and report drift to the program lead.
  • Propose a rubric clarification after spotting a recurring failure mode.

Nice to have

  • Background in linguistics, content moderation, or trust & safety review.
  • Experience with inter-rater agreement metrics and calibration cycles.
  • Domain expertise that lets you spot subject-matter errors automated checks miss.

Skills

  • Model output evaluation
  • Rubric-based annotation
  • Severity tagging
  • Inter-rater calibration
  • Arabic generalist evaluation
  • Localization review
  • Cultural context
  • Language evaluation
  • Arabic
  • Localization

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 8 days ago
Similar jobs that could be interesting for youBased on the Arabic Localization AI Evaluator [Remote] in Remote vacancy
  •  ...hourly contractor role supporting AI data and language projects on...  ...AI training datasets. LLM evaluation: reviewing AI-generated...  ...linguistic appropriateness. Localization QA: ensuring terminology,...  ...products, with deep expertise in Arabic-native and secure, sovereign... 
    Arabic language skills
    Hourly pay
    For contractors
    Remote work
    Flexible hours

    CNTXT AI

    Brooklyn, NY
    13 days ago
  •  ...Bring characters and stories to life in Arabic across a range of domains, using vocal performance...  ...right emotion, personality, and tone for AI-focused audio projects. Key...  ...Qualifications Experience with dubbing, ADR, or localization. A demo reel demonstrating a range of... 
    Arabic language skills
    Full time
    Contract work
    Part time
    Remote work
    Flexible hours

    SaidGig

    United States
    more than 2 months ago
  •  ...Legal Translation AI Evaluator is a remote review track for evaluating AI outputs in legal review workflows. Reviewers grade citation...  ...Statutory analysis Policy reasoning Legal review Localization review Cultural context Language evaluation Legal... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    8 days ago
  • $60 - $65 per hour

     ...Role Overview Help improve next generation AI language systems by evaluating AI generated Indonesian speech. You will use your native level Indonesian...  ...requirements on time. Experience in translation, localization, linguistics, content moderation, or language quality... 
    Suggested
    Hourly pay
    For contractors
    Remote work

    SaidGig

    United States
    24 days ago
  •  ...Medical Translation AI Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers grade...  ...Structured clinical writing Clinical review Localization review Cultural context Language evaluation Medical... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  • $14.5 per hour

    A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical... 
    Hourly pay
    Part time
    Remote work
    Flexible hours

    Welo Data

    United States
    3 days ago
  • $20 per hour

    A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and... 
    Remote work
    Flexible hours

    DataAnnotation

    United States
    3 days ago
  •  ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and... 
    Full time
    Contract work
    Remote work
    Flexible hours

    Virtual Vocations Inc

    United States
    2 days ago
  •  ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional... 
    Work at office
    Remote work

    Obsidian

    New York, NY
    5 days ago
  • $14.5 per hour

     ...ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years...  ...we'd like you to join us! Job Description As a Web Search Evaluator, you will play a key role in improving the quality of search... 
    Part time
    Currently hiring
    Immediate start
    Remote work
    Work from home
    10 hours per week
    Flexible hours

    Jobs for Humanity

    United States
    4 days ago
  •  ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers... 
    Contract work
    Temporary work
    Immediate start
    Remote work

    MERIT Beauty

    New York, NY
    2 days ago
  • $14.5 per hour

     ...AI Web Search Evaluator Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years of experience in... 
    Bi-weekly pay
    Hourly pay
    Part time
    Immediate start
    Remote work
    Work from home
    Flexible hours

    Welo Data

    United States
    13 hours ago
  • $8 - $65 per hour

     ...Meridial is seeking an Arabic (Gulf) Language Expert for a freelance AI trainer project. This remote position allows you to leverage your expertise to influence the next generation of AI. You will verify linguistic accuracy and explore various topics in Gulf Arabic, documenting... 
    Arabic language skills
    Hourly pay
    Freelance
    Remote work

    Meridial

    United States
    3 days ago
  •  ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise... 
    Hourly pay
    Work at office
    Remote work

    Mercor Inc

    New York, NY
    3 days ago
  • $20 per hour

     ...Role Overview: Bring Arabic-language characters and scripted content...  ...with dubbing, ADR, or localization work. Demo reel showcasing...  ...Opportunity to contribute to advanced AI projects with leading AI...  ...and performance align. Evaluation Process: Applications are... 
    Arabic language skills
    Full time
    Contract work
    Part time
    For contractors
    Freelance
    Remote work
    Flexible hours

    SaidGig

    United States
    2 days ago
  •  ...Invisible Technologies Inc. is seeking an experienced Arabic voice actor to power AI training data for education, entertainment, and accessibility...  ...delivery. You will record scripted material, evaluate AI outputs for naturalness, and provide feedback to improve... 
    Arabic language skills
    Hourly pay
    Contract work
    Remote work

    Invisible Technologies Inc. Defunct

    United States
    14 hours ago
  • $20 per hour

    A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI...  ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with... 
    Hourly pay
    Remote work
    Flexible hours

    SupportFinity

    United States
    3 days ago
  •  ...Prolific is seeking talented Product Designers and UX Specialists to join our Expert Network for training and evaluation of cutting-edge AI models. Candidates should have a relevant educational background and at least one year of experience in design-related fields. Responsibilities... 
    Remote work
    Work from home
    Flexible hours

    Prolific

    Arizona City, AZ
    1 day ago
  • $20 - $80 per hour

     ...Role Overview Help improve next-generation AI systems by supplying precise, real-world evaluation, annotation, and feedback. This remote contractor role focuses on how AI models learn, reason, and perform across diverse subject areas. Key Responsibilities Evaluate... 
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    United States
    more than 2 months ago
  • $20 - $160 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Position: Generalist Annotator — Health AI Conversation Quality Evaluation Type: Contract Compensation: $20–$160/hour Location... 
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    2 days ago
  •  ...Prolific is seeking Product Designers and UX Specialists to join our Expert Network, contributing to the training and evaluation of cutting-edge AI models. This role requires expertise in usability, design systems, and user research, providing essential feedback on design... 
    Remote work
    Work from home
    Flexible hours

    Prolific

    Tucson, AZ
    5 days ago
  • $15 per hour

     ...We are looking for AI Business Conversation Analysts to support the development and improvement of AI models in Saudi Arabic (Najran ). Job Type: Freelance Location: Work from home...  ..., business reporting, research/policy evaluation, and data-driven decision support. Conduct... 
    Arabic language skills
    Full time
    Part time
    Freelance
    Immediate start
    Work from home
    10 hours per week
    Flexible hours

    Rws

    Remote
    10 days ago
  •  ...underpinning the emerging trillion-dollar Voice AI economy, providing real-time APIs for...  ...THE OPPORTUNITY Deepgram is investing in Arabic-language capabilities, and the customers...  ...discovery, proof-of-concepts, and evaluations that earn the technical win for regional... 
    Arabic language skills
    Full time

    Deepgram

    Remote
    22 days ago
  • $70 - $110 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...weeks Commitment: 20+ hours/week Role Responsibilities Evaluate AI-generated financial plans , budgets, and forecasts for quality... 
    Hourly pay
    Contract work
    Summer work
    Work at office
    Immediate start
    Remote work

    Mercor

    New York, NY
    2 days ago
  • $20 per hour

    A leading AI technology firm is seeking a Graphic Design Expert to evaluate graphic design elements for AI model training. This part-time contract role offers remote work flexibility and requires expertise in design principles. Ideal candidates should have formal training... 
    Contract work
    Part time
    Remote work

    HumanSignal

    Austin, TX
    1 day ago
  • $50 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Remote Commitment: 20–40 hours/week Role Responsibilities Evaluate the effectiveness of AI systems in handling personalized, real... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    2 days ago
  •  ...Executive ElevenLabs is an AI research and product company...  ...stakeholders on compliance, localization, and responsible AI deployment...  ...at all levels: from technical evaluators to ministerial leadership,...  ...Professional proficiency in Arabic We are an equal opportunity... 
    Arabic language skills
    Remote work

    Eleven Labs

    United States
    5 days ago
  • $30 - $50 per hour

     ...A leading AI training company is seeking a Remote Annotator to support human-in-the-loop AI training workflows for large language...  ...models. This role involves reviewing labeled datasets, performing evaluations for helpfulness and safety, and ensuring quality in training... 
    Hourly pay
    Remote work

    Rex USA

    United States
    4 days ago
  • $26 - $28 per hour

     ...supporting speech and voice AI systems. This is a high-impact...  ...more execution-focused than evaluation-heavy roles, it still...  ...~ Native-level fluency in Arabic  ~ Strong written communication...  ...transformation services in translation, localization, and adaptation for over 250... 
    Arabic language skills
    Full time
    Work experience placement
    Remote work
    Visa sponsorship

    Welo Data

    Boston, MA
    11 days ago
  • $90 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...replications from those that merely look correct. Evaluate responsive behavior and semantic quality,...  .... ~ Comfort in a terminal for local server setup. ~ Strong written English and... 
    Contract work
    Summer work
    Local area
    Remote work

    Mercor

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Arabic Localization AI Evaluator [Remote]. Be the first to apply!