Cross-Cultural Pragmatics AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Cross-Cultural Pragmatics AI Evaluator is a remote evaluation track for reviewing cross cultural pragmatics ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn cross cultural pragmatics ai evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate cross cultural pragmatics ai evaluation model outputs against a versioned rubric and assign severity tags for Cross-Cultural Pragmatics AI Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on cross cultural pragmatics ai evaluation or adjacent content for Cross-Cultural Pragmatics AI Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two cross cultural pragmatics ai evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Cross Cultural Pragmatics AI evaluation
- Localization review
- Cultural context
- Language evaluation
- Cross
- Cultural
- Pragmatics
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$80 - $120 per hour
...Role Overview Review AI-generated documents, spreadsheets, and slide decks using your expertise in humanities, arts, and culture. You will assess each work product for accuracy, rigor... ...quality outputs. Key Responsibilities Evaluate AI-generated artifacts against domain-...SuggestedHourly payWork at officeRemote work- ...remote, hourly contractor role supporting AI data and language projects on a project-... ...to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy... ..., reasoning quality, coherence, and cultural/linguistic appropriateness. Localization...SuggestedHourly payFor contractorsRemote workFlexible hours
- ...Opportunity We are seeking detail-oriented human reviewers with a strong understanding of their local cultural context to support a range of AI training and evaluation projects . In this role, you will work across diverse task types, including evaluating prompts...SuggestedExtra incomeFull timeFor contractorsLocal area10 hours per week
- ...Hiring PhD-level quantitative finance experts, the full-time Remote AI Research Evaluator will assess AI-generated financial content, craft relevant questions, and evaluate model responses in a flexible contract role. Key responsibilities Assessing the factuality and...SuggestedFull timeContract workRemote workFlexible hours
- ...Legal Translation AI Evaluator is a remote review track for evaluating AI outputs in legal review... ...tooling. Bilingual experience for cross-jurisdiction matters. Skills Legal... ...Legal review Localization review Cultural context Language evaluation Legal...SuggestedRemote jobHourly payFor contractors10 hours per week
$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...Remote workFlexible hours- ...Alignerr is seeking a Search Quality Evaluator to assess search engine results, AI-generated answers, and content recommendations. This fully remote, flexible contract role values strong critical thinking and quality instincts, with no technical background required beyond...Contract workRemote workFlexible hours
$100 per hour
...looking for a highly experienced software engineer (SR+) to help evaluate the quality of interactions with modern coding agents such as... ...whether the model thinks like a great engineer. You will assess how AI coding agents behave in real-world scenarios — focusing on:...Contract workImmediate start$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ..., and Jack Dorsey . Position: Humanities / arts / culture Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ..., and Jack Dorsey . Position: Humanities / arts / culture Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$16.5 per hour
...Duration: TBC Rate: 16.5 USD per hour What You'll Do Media Evaluation: Watch and analyze short-form video clips to identify all... ...that the media is authentic to the French [FR] linguistic and cultural context). Data Labeling: Generate highly accurate ground...Hourly payPart timeFreelanceImmediate startWork from home10 hours per weekFlexible hours- ...Prolific is seeking Product Designers and UX Specialists to join our Expert Network, contributing to the training and evaluation of cutting-edge AI models. This role requires expertise in usability, design systems, and user research, providing essential feedback on design...Remote workWork from homeFlexible hours
- ...Prolific is seeking talented Product Designers and UX Specialists to join our Expert Network for training and evaluation of cutting-edge AI models. Candidates should have a relevant educational background and at least one year of experience in design-related fields. Responsibilities...Remote workWork from homeFlexible hours
$60 - $70 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...$60–$70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance...Contract workSummer workRemote work$80 - $120 per hour
Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks. ~Provide clear, structured written feedback to improve AI outputs. ~Apply...Part timeWork at officeRemote work$70 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...: $70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses to strengthen reasoning and rigor in model...Contract workSummer workRemote work- ...Remote | Work from Home Employment Type: Project-based | Contract We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists with strong proficiency in English. In this role, you will support AI/ML...Contract workRemote workWork from homeMonday to FridayDay shift
$50 - $190 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Remote Commitment: 20+ hours/week Role Responsibilities Evaluate AI systems on complex personal workflows, including personal...Hourly payContract workFor contractorsSummer workRemote workTrial period$50 - $175 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Remote Commitment: 20–40 hours/week Role Responsibilities Evaluate JSON-formatted AI model outputs against defined grading...Contract workSummer workStart working todayRemote work$20 - $80 per hour
...Role Overview Train and evaluate next-generation AI systems by scoring model outputs, annotating real-world content, and delivering clear, actionable feedback that improves model accuracy and reasoning across diverse domains. About the company micro1 is an AI data...Hourly payFor contractorsRemote work- ...position of Office Specialist VI: Transcript Evaluator. The Office Specialist VI: Transcript... ...was founded by the Congregation of Holy Cross, from which it acquired distinguishing... ...educational opportunities for students of varied cultural, religious, educational and economic...InternshipWork at officeImmediate startFlexible hours
- ...Poughkeepsie, NYAllied Health Prof/TechnicalFull TimeVariableVariedThe Mental Health Evaluator (MHE) is a critical role dedicated to delivering comprehensive, clinically appropriate, culturally competent, and trauma-informed patient behavioral health interventions in the...
$80 - $120 per hour
Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks. ~Provide clear, structured written feedback to improve AI model outputs....Part timeWork at officeRemote work$150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...consistent. Preferred ~ Prior experience with AI training , evaluation, or human-data projects. Application Process (Takes 20–30...Contract workSummer workRemote work$120 - $175 per hour
...Job Description Job Description About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers...Contract workSummer workRemote work- ...Trust and Safety Policy AI Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft... ...public adversarial-ML research. Multilingual fluency for cross-language attack testing. Skills Adversarial prompting...Remote jobHourly payFor contractors10 hours per week
- ...Privacy Law AI Evaluator is a remote review track for evaluating AI outputs in privacy law workflows. Reviewers grade citation accuracy,... ...legal-research or compliance tooling. Bilingual experience for cross-jurisdiction matters. Skills Legal research Citation...Remote jobHourly payFor contractors10 hours per week
- ...Audit and Controls AI Evaluator is a remote review track for evaluating AI outputs across audit workflows. Reviewers grade calculations,... ...risk tooling and its failure modes. Bilingual experience for cross-jurisdiction reviews. Skills Financial analysis Audit...Remote jobHourly payFor contractorsWork experience placement10 hours per week
- ...Position Summary In this remote, hourly contractor role, you will evaluate AI-generated financial content and develop cases that test... ...reasoning across multi-document financial scenarios, including cross-referencing data across balance sheets, reports, and operational...Hourly payFull timeTemporary workFor contractorsRemote work
$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...and Jack Dorsey . Position: Biology / environmental science Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Cross-Cultural Pragmatics AI Evaluator [Remote]. Be the first to apply!






