Cross-Cultural Pragmatics AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Cross-Cultural Pragmatics AI Evaluator is a remote evaluation track for reviewing cross cultural pragmatics ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn cross cultural pragmatics ai evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate cross cultural pragmatics ai evaluation model outputs against a versioned rubric and assign severity tags for Cross-Cultural Pragmatics AI Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on cross cultural pragmatics ai evaluation or adjacent content for Cross-Cultural Pragmatics AI Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two cross cultural pragmatics ai evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Cross Cultural Pragmatics AI evaluation
- Localization review
- Cultural context
- Language evaluation
- Cross
- Cultural
- Pragmatics
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- ...Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence,... ...candidates are native or fluent Turkish speakers with strong cultural insight, attention to detail, and the ability to give...SuggestedContract workTemporary workImmediate startRemote work
$14.5 per hour
...Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are... ...English (written and spoken) Strong understanding of pop culture in the US Reliable computer system and internet connection...SuggestedHourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours$20 - $30 per hour
...firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing... ...commitment between 10 to 40 hours each week, compensated hourly at $20-$30. #J-18808-Ljbffr Crossing HurdlesSuggestedRemote jobHourly pay$30 per hour
...: 10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous... ...Strong experience in LLM evaluation, AI output analysis, QA/testing, UX research,... ...structured evaluation workflows and evolving guidelines. #J-18808-Ljbffr Crossing HurdlesSuggestedRemote jobHourly payContract work- BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with... ...should be native Swedish speakers with a keen attention to detail and a strong grasp of cultural nuances. #J-18808-Ljbffr BAM VenturesSuggestedRemote job
- ...Legal Translation AI Evaluator is a remote review track for evaluating AI outputs in legal review... ...tooling. Bilingual experience for cross-jurisdiction matters. Skills Legal... ...Legal review Localization review Cultural context Language evaluation Legal...Remote jobHourly payFor contractors10 hours per week
$70 - $110 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...0+ hours/week Role Responsibilities Evaluate AI-generated operational plans , staff schedules... ...soundness, financial rigor, and cross-department feasibility. Review operational...Hourly payContract workSummer workImmediate startRemote work- ...remote, hourly contractor role supporting AI data and language projects on a project-... ...to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy... ..., reasoning quality, coherence, and cultural/linguistic appropriateness. Localization...Hourly payFor contractorsRemote workFlexible hours
- YO AI Labs is seeking a Korean Language Expert contractor to contribute linguistic expertise... ...toward improving AI systems. You will evaluate, translate, and annotate Korean content... ...data, develop guidelines, and ensure culturally accurate communication. #J-18808-Ljbffr...Remote jobFor contractors
$50 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Location: Remote Role Responsibilities Evaluate the quality and accuracy of LLM-... ...to meet high linguistic, stylistic, and cultural standards. Annotate English text for grammatical...Contract workSummer workRemote work$14.5 per hour
A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical...Hourly payPart timeRemote workFlexible hours- ...Weekday 1 is seeking expert Evaluators to review AI-generated real estate, hospitality and events outputs for accuracy, rigor and domain quality. You will apply deep expertise to grade documents, spreadsheets and slide decks. Requirements include 5+ years in Real estate...Hourly payWeekly payContract workWork at officeRemote workWeekday work
$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...Remote workFlexible hours- ...Alignerr is seeking Creative Writing Evaluators to assess AI-generated stories, essays, poetry, and other writings to shape how AI learns compelling prose. This fully remote, flexible contract roles welcomes avid readers and writers with no publishing credits required...Contract workRemote workFlexible hours
- ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and...Full timeContract workRemote workFlexible hours
- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Work at officeRemote work
- A leading AI research accelerator is hiring a position focused on contributing to projects that evaluate and enhance AI systems. You will design community service scenarios, write structured explanations, and evaluate AI accuracy. The ideal candidate will have 4+ years...Remote jobFull timeFor contractors
- Alignerr is seeking Wildlife and Habitat Conservation Scientists to evaluate AI-trained biodiversity protection content. This remote, hourly contract role offers flexible hours (10-40 per week) and a chance to shape how AI understands ecological challenges on a global scale...Remote jobHourly payContract workFlexible hours
- YO AI Labs is seeking a Developer & Infrastructure Expert for a remote contractor role focused on evaluating AI-powered workflows across software development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands and configurations...Remote jobFor contractors
- ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise...Hourly payWork at officeRemote work
$50 per hour
Prolific seeks Product Designers and UX Specialists to help train AI models using your expertise. You'll evaluate AI-generated designs and ensure usability and accessibility while working from home. Ideal candidates hold a BS, MS, or PhD in a relevant field and have at...Remote jobWork from homeFlexible hours$14.5 per hour
...ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years... ...performance and provide insights on relevance and quality. Evaluate and rate the effectiveness of search engine results to ensure...Hourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours- YO AI Labs is seeking a Healthcare Expert for a remote contract to support AI training and evaluation projects. You will apply clinical knowledge to assess AI outputs, workflows, and real-world use of healthcare systems. Key tasks include evaluating EHR applications, reviewing...Remote jobContract work
- Mercor is hiring expert Evaluators in General Sales / GTM to review AI-generated outputs for accuracy and quality. This remote hourly role requires applying deep domain knowledge to assess documents, spreadsheets, and slide decks. Ideal candidates have 5+ years in General...Remote jobHourly payWork at office
- Obsidian is hiring expert Evaluators in Healthcare operations for a remote, hourly engagement. You will review AI-generated work products for accuracy, rigor, and domain quality, leveraging your extensive subject-matter expertise in the field. The ideal candidate has over...Remote jobHourly payWork at office
$20 per hour
A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI... ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with...Hourly payRemote workFlexible hours- ...Insurance Policy AI Evaluator is a remote review track for evaluating AI outputs across insurance workflows. Reviewers grade calculations... ...risk tooling and its failure modes. Bilingual experience for cross-jurisdiction reviews. Skills Financial analysis Audit...Remote jobHourly payFor contractorsWork experience placement10 hours per week
$60 - $65 per hour
...Role Overview Help improve next generation AI language systems by evaluating AI generated Indonesian speech. You will use your native level Indonesian expertise to assess fluency, authenticity, and cultural fit, then provide clear feedback that supports model training...Hourly payFor contractorsRemote work$50 - $75 per hour
...Overview Apply your consumer and lifestyle expertise to help evaluate and improve advanced AI systems. Your judgment will help establish the standards... ...assess AI responses in areas where practical knowledge, cultural context, and real-world constraints matter. Key...Hourly payFor contractorsLocal areaRemote work- ...Patent Claim AI Evaluator is a remote review track for evaluating AI outputs in patent law workflows. Reviewers grade citation accuracy,... ...legal-research or compliance tooling. Bilingual experience for cross-jurisdiction matters. Skills Legal research Citation...Remote jobHourly payFor contractors10 hours per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Cross-Cultural Pragmatics AI Evaluator [Remote]. Be the first to apply!



