Swahili Evaluation AI Evaluator
AI Trainer Jobs
About the role Swahili Evaluation AI Evaluator is a remote evaluation track for reviewing swahili generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain. Category: Frontier Model Evaluation · Pay: Hourly rate confirmed after the interview process · Location: Remote — US-eligible · Contractor Swahili Evaluation AI Evaluator is a remote evaluation track for reviewing swahili generalist evaluation prompts and responses against AuraOne's quality rubric. About the role Swahili Evaluation AI Evaluator is a remote evaluation track for reviewing swahili generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain. AI data reviewers help turn swahili generalist evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data. Review frontier model outputs. Judge benchmark failures and calibrate other evaluators. Responsibilities Evaluate swahili generalist evaluation model outputs against a versioned rubric and assign severity tags for Swahili Evaluation AI Evaluator assignments. Compare paired responses and pick the stronger answer with a written rationale. Label hallucinations, instruction-following failures, and unsafe content with structured tags. Capture ambiguous prompts and route them back to the program team for rubric updates. Maintain reviewer-quality scores by calibrating against gold-standard examples each week. Role details Track Evaluation & annotation Work model Remote · Independent specialist contractor Compensation Hourly rate confirmed after the interview process. Eligible from US What you should bring Prior evaluation, annotation, or human-rater experience on swahili generalist evaluation or adjacent content for Swahili Evaluation AI Evaluator work. Comfort applying multi-page rubrics consistently across long batches. Clear written reasoning that names the issue and the rubric clause being applied. Strong attention to detail and the ability to flag when a prompt itself is the problem. Reliable async availability for at least 10 hours per week. Role signals Example tasks Compare two swahili generalist evaluation model responses to the same prompt and pick the stronger one with rationale. Tag an unsafe response with the correct policy category and severity. Audit a 50-row batch for rubric consistency and report drift to the program lead. Propose a rubric clarification after spotting a recurring failure mode. Useful experience Background in linguistics, content moderation, or trust & safety review. Experience with inter-rater agreement metrics and calibration cycles. Domain expertise that lets you spot subject-matter errors automated checks miss. Compensation and schedule Hourly rate confirmed after the interview process. Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation. Skills used in matching Model output evaluation Rubric-based annotation Severity tagging Inter-rater calibration Swahili generalist evaluation Localization review Cultural context Language evaluation Swahili Evaluation Application boundary Creating a specialist profile records your experience and preferences. Starting role intake is a separate action that attaches this role to your candidate record. Specialist intake The intake preserves your chosen role, the visible terms, and source attribution for reviewer context. 01 Confirm profile and eligibility details. 02 Attach this role deliberately. 03 Receive a human review decision or follow-up. Placement timing depends on program demand and reviewer confirmation. #J-18808-Ljbffr
- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...SuggestedWork at officeRemote work
- ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers...SuggestedContract workTemporary workImmediate startRemote work
$14.5 per hour
...Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data...SuggestedHourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours- ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise...SuggestedHourly payWork at officeRemote work
- Alignerr is seeking Creative Writing Evaluators to assess AI-generated stories, essays, poetry, and other writings to shape how AI learns compelling prose. This fully remote, flexible contract roles welcomes avid readers and writers with no publishing credits required—just...SuggestedRemote jobContract workFlexible hours
- BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance...Remote job
- Mercor is seeking experienced musicians to evaluate generative music AI models, working in Bengali and English. You will assess AI-generated lyrics across genres and rate them against detailed quality standards. Responsibilities include comparing lyrics to published songs...Part timeImmediate startFlexible hours
$14.5 per hour
A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical...Remote jobHourly payPart timeFlexible hours- A leading AI research accelerator is hiring a position focused on contributing to projects that evaluate and enhance AI systems. You will design community service scenarios, write structured explanations, and evaluate AI accuracy. The ideal candidate will have 4+ years...Remote jobFull timeFor contractors
$20 - $160 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Position: Generalist Annotator — Health AI Conversation Quality Evaluation Type: Contract Compensation: $20–$160/hour Location...Contract workSummer workRemote work$60 - $70 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...$60–$70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance...Contract workSummer workRemote work- About the Opportunity A leading AI research organization is seeking advanced LLM power users with strong experience using MCP and (... ...connectors for real-world personal life tasks. This project focuses on evaluating how well AI systems handle personalized, multi-step life tasks...Trial period
- Mercor is hiring expert Evaluators in Media, journalism, and communications to review AI-generated work products for accuracy, rigor, and domain quality. This is a remote, hourly engagement. Role requires 5+ years in media/journalism/communications, native/professional...Remote jobHourly payWork at office
$85 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-...Contract workSummer workRemote work$30 - $50 per hour
A leading AI training company is seeking a Remote Annotator to support human-in-the-loop AI training workflows for large language... ...models. This role involves reviewing labeled datasets, performing evaluations for helpfulness and safety, and ensuring quality in training...Remote jobHourly pay- Mercor is seeking experienced Adult Inpatient Nurses (RNs) to train and evaluate AI systems used in clinical and healthcare settings. This role leverages frontline nursing expertise to improve AI accuracy, safety, and reliability. You’ll work on projects requiring deep...
$70 - $110 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...weeks Commitment: 20+ hours/week Role Responsibilities Evaluate AI-generated financial plans , budgets, and forecasts for quality...Hourly payContract workSummer workWork at officeImmediate startRemote work$70 - $110 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...weeks Commitment: 20+ hours/week Role Responsibilities Evaluate AI-generated operational plans , staff schedules, and performance...Hourly payContract workSummer workImmediate startRemote work- ...required between testing sites in Bronx, Brooklyn, Queens, and Manhattan. Job Description: An employer is looking for Nurse Aide Evaluators. These individuals will be responsible for proctoring written exams and evaluating skills of students testing to become a CNA....Flexible hours
$160k - $210k
BVAL (Bloomberg's Evaluated Pricing Service) Evaluator - US Agency Structured Products Location New York Business Area Product Ref # 10053894 Description & Requirements Bloomberg’s Evaluated Pricing Service, BVAL, provides transparent and accurate...Price workTemporary workFor contractorsWork experience placement$80 - $120 per hour
...This role is for one of our clients Compensation: $80 - $120 per hour We are hiring expert Evaluators in Special education / IEP to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality....Hourly payContract workFor contractorsWork at officeRemote work- ...A leading research accelerator is seeking a contractor to evaluate North American teen humor. The role involves reviewing short-form content, rating based on cultural relevance, and explaining humor dynamics clearly. Ideal candidates are 18 to 19 years old, familiar with...Contract workFor contractorsFreelanceRemote workFlexible hours
$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... .... Position: User/customer research and feedback synthesis Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$80 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Responsibilities Use frontier AI coding agents to complete and evaluate complex data engineering tasks. Review model-generated...Contract workSummer workRemote work$75 per hour
...Commitment: 10–40 hours/week Role Responsibilities Develop and deliver sociology content through AI training initiatives for postsecondary education. Evaluate and review AI-generated coursework, assignments, and simulated academic papers. Create lectures,...Hourly payContract workRemote work- ...Health, part of CVS Health, is seeking a Part-Time Clinician (Nurse Practitioner or Physician Assistant) to provide in-home health evaluations. Visiting members in their own homes, you will review medical history, perform a physical exam and support gaps in care. Visits...Part time
- ...leading digital solutions provider is seeking a Personalized Ads Evaluator for a part-time position. This entry-level role involves... ...to 20 hours of remote work weekly and requires passing a basic qualification exam. #J-18808-Ljbffr TELUS Digital AI Data SolutionsRemote jobPart time
- ...Straive is a global leader in enterprise‑grade data analytics and AI solutions, committed to empowering businesses across various... ...fueled by our partnership with EQT. Website: Linkedin Job Title: Evaluator - Political Science Location: Remote (USA) Job Type: Contract...Contract workRemote workWorldwide
$150 per hour
JOB DESCRIPTION The Home and Community-Based Program evaluates preschoolers for special education for the Committee on Special Education (CPSE). Education Evaluations are one type of evaluation completed. Education Evaluations include assessment of children in five learning...Immediate startWork from home- About the Role We’re looking for contract, remote Turkish-speaking annotators to evaluate AI-generated content with a sharp eye for detail and cultural nuance. You’ll evaluate whether outputs are coherent, consistent, and aligned with real-world expectations in Turkish....Contract workTemporary workFreelanceImmediate startRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Swahili Evaluation AI Evaluator. Be the first to apply!




