Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Swahili Evaluation AI Evaluator

AI Trainer Jobs

About the role Swahili Evaluation AI Evaluator is a remote evaluation track for reviewing swahili generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain. Category: Frontier Model Evaluation · Pay: Hourly rate confirmed after the interview process · Location: Remote — US-eligible · Contractor Swahili Evaluation AI Evaluator is a remote evaluation track for reviewing swahili generalist evaluation prompts and responses against AuraOne's quality rubric. About the role Swahili Evaluation AI Evaluator is a remote evaluation track for reviewing swahili generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain. AI data reviewers help turn swahili generalist evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data. Review frontier model outputs. Judge benchmark failures and calibrate other evaluators. Responsibilities Evaluate swahili generalist evaluation model outputs against a versioned rubric and assign severity tags for Swahili Evaluation AI Evaluator assignments. Compare paired responses and pick the stronger answer with a written rationale. Label hallucinations, instruction-following failures, and unsafe content with structured tags. Capture ambiguous prompts and route them back to the program team for rubric updates. Maintain reviewer-quality scores by calibrating against gold-standard examples each week. Role details Track Evaluation & annotation Work model Remote · Independent specialist contractor Compensation Hourly rate confirmed after the interview process. Eligible from US What you should bring Prior evaluation, annotation, or human-rater experience on swahili generalist evaluation or adjacent content for Swahili Evaluation AI Evaluator work. Comfort applying multi-page rubrics consistently across long batches. Clear written reasoning that names the issue and the rubric clause being applied. Strong attention to detail and the ability to flag when a prompt itself is the problem. Reliable async availability for at least 10 hours per week. Role signals Example tasks Compare two swahili generalist evaluation model responses to the same prompt and pick the stronger one with rationale. Tag an unsafe response with the correct policy category and severity. Audit a 50-row batch for rubric consistency and report drift to the program lead. Propose a rubric clarification after spotting a recurring failure mode. Useful experience Background in linguistics, content moderation, or trust & safety review. Experience with inter-rater agreement metrics and calibration cycles. Domain expertise that lets you spot subject-matter errors automated checks miss. Compensation and schedule Hourly rate confirmed after the interview process. Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation. Skills used in matching Model output evaluation Rubric-based annotation Severity tagging Inter-rater calibration Swahili generalist evaluation Localization review Cultural context Language evaluation Swahili Evaluation Application boundary Creating a specialist profile records your experience and preferences. Starting role intake is a separate action that attaches this role to your candidate record. Specialist intake The intake preserves your chosen role, the visible terms, and source attribution for reviewer context. 01 Confirm profile and eligibility details. 02 Attach this role deliberately. 03 Receive a human review decision or follow-up. Placement timing depends on program demand and reviewer confirmation. #J-18808-Ljbffr

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Swahili Evaluation AI Evaluator in New York, NY vacancy
  •  ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional... 
    Suggested
    Work at office
    Remote work

    Obsidian

    New York, NY
    4 days ago
  •  ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers... 
    Suggested
    Contract work
    Temporary work
    Immediate start
    Remote work

    MERIT Beauty

    New York, NY
    2 days ago
  • $14.5 per hour

     ...Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data... 
    Suggested
    Hourly pay
    Part time
    Immediate start
    Remote work
    Work from home
    10 hours per week
    Flexible hours

    Welo Data

    New York, NY
    5 days ago
  •  ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise... 
    Suggested
    Hourly pay
    Work at office
    Remote work

    Mercor Inc

    New York, NY
    3 days ago
  • Alignerr is seeking Creative Writing Evaluators to assess AI-generated stories, essays, poetry, and other writings to shape how AI learns compelling prose. This fully remote, flexible contract roles welcomes avid readers and writers with no publishing credits required—just... 
    Suggested
    Remote job
    Contract work
    Flexible hours

    Alignerr Corp.

    New York, NY
    4 days ago
  • BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance... 
    Remote job

    BAM Ventures

    New York, NY
    5 days ago
  • Mercor is seeking experienced musicians to evaluate generative music AI models, working in Bengali and English. You will assess AI-generated lyrics across genres and rate them against detailed quality standards. Responsibilities include comparing lyrics to published songs... 
    Part time
    Immediate start
    Flexible hours

    Obsidian

    New York, NY
    5 days ago
  • $14.5 per hour

    A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical... 
    Remote job
    Hourly pay
    Part time
    Flexible hours

    Welo Data

    New York, NY
    5 days ago
  • A leading AI research accelerator is hiring a position focused on contributing to projects that evaluate and enhance AI systems. You will design community service scenarios, write structured explanations, and evaluate AI accuracy. The ideal candidate will have 4+ years... 
    Remote job
    Full time
    For contractors

    Turing

    New York, NY
    3 days ago
  • $20 - $160 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Position: Generalist Annotator — Health AI Conversation Quality Evaluation Type: Contract Compensation: $20–$160/hour Location... 
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    6 days ago
  • $60 - $70 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...$60–$70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance... 
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    6 days ago
  • About the Opportunity A leading AI research organization is seeking advanced LLM power users with strong experience using MCP and (...  ...connectors for real-world personal life tasks. This project focuses on evaluating how well AI systems handle personalized, multi-step life tasks... 
    Trial period

    Obsidian

    New York, NY
    4 days ago
  • Mercor is hiring expert Evaluators in Media, journalism, and communications to review AI-generated work products for accuracy, rigor, and domain quality. This is a remote, hourly engagement. Role requires 5+ years in media/journalism/communications, native/professional... 
    Remote job
    Hourly pay
    Work at office

    Mercor Inc

    New York, NY
    5 days ago
  • $85 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-... 
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    6 days ago
  • $30 - $50 per hour

    A leading AI training company is seeking a Remote Annotator to support human-in-the-loop AI training workflows for large language...  ...models. This role involves reviewing labeled datasets, performing evaluations for helpfulness and safety, and ensuring quality in training... 
    Remote job
    Hourly pay

    Rex.zone

    New York, NY
    3 days ago
  • Mercor is seeking experienced Adult Inpatient Nurses (RNs) to train and evaluate AI systems used in clinical and healthcare settings. This role leverages frontline nursing expertise to improve AI accuracy, safety, and reliability. You’ll work on projects requiring deep... 

    Mercor

    New York, NY
    6 days ago
  • $70 - $110 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...weeks Commitment: 20+ hours/week Role Responsibilities Evaluate AI-generated financial plans , budgets, and forecasts for quality... 
    Hourly pay
    Contract work
    Summer work
    Work at office
    Immediate start
    Remote work

    Mercor

    New York, NY
    14 days ago
  • $70 - $110 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...weeks Commitment: 20+ hours/week Role Responsibilities Evaluate AI-generated operational plans , staff schedules, and performance... 
    Hourly pay
    Contract work
    Summer work
    Immediate start
    Remote work

    Mercor

    New York, NY
    17 days ago
  •  ...required between testing sites in Bronx, Brooklyn, Queens, and Manhattan. Job Description: An employer is looking for Nurse Aide Evaluators. These individuals will be responsible for proctoring written exams and evaluating skills of students testing to become a CNA.... 
    Flexible hours

    Insight Global

    New York, NY
    15 days ago
  • $160k - $210k

    BVAL (Bloomberg's Evaluated Pricing Service) Evaluator - US Agency Structured Products Location New York Business Area Product Ref # 10053894 Description & Requirements Bloomberg’s Evaluated Pricing Service, BVAL, provides transparent and accurate... 
    Price work
    Temporary work
    For contractors
    Work experience placement

    Bloomberg

    New York, NY
    1 day ago
  • $80 - $120 per hour

     ...This role is for one of our clients Compensation: $80 - $120 per hour We are hiring expert Evaluators in Special education / IEP to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality.... 
    Hourly pay
    Contract work
    For contractors
    Work at office
    Remote work

    Weekday 1

    New York, NY
    2 days ago
  •  ...A leading research accelerator is seeking a contractor to evaluate North American teen humor. The role involves reviewing short-form content, rating based on cultural relevance, and explaining humor dynamics clearly. Ideal candidates are 18 to 19 years old, familiar with... 
    Contract work
    For contractors
    Freelance
    Remote work
    Flexible hours

    Turing Inc

    New York, NY
    4 days ago
  • $80 - $120 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  .... Position: User/customer research and feedback synthesis Evaluator Type: Contract Compensation: $80–$120/hour Location... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    New York, NY
    28 days ago
  • $80 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Responsibilities Use frontier AI coding agents to complete and evaluate complex data engineering tasks. Review model-generated... 
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    17 days ago
  • $75 per hour

     ...Commitment: 10–40 hours/week Role Responsibilities Develop and deliver sociology content through AI training initiatives for postsecondary education. Evaluate and review AI-generated coursework, assignments, and simulated academic papers. Create lectures,... 
    Hourly pay
    Contract work
    Remote work

    Crossing Hurdles

    New York, NY
    2 days ago
  •  ...Health, part of CVS Health, is seeking a Part-Time Clinician (Nurse Practitioner or Physician Assistant) to provide in-home health evaluations. Visiting members in their own homes, you will review medical history, perform a physical exam and support gaps in care. Visits... 
    Part time

    8205 Signify Health Medical Associates, PLLC

    New York, NY
    3 days ago
  •  ...leading digital solutions provider is seeking a Personalized Ads Evaluator for a part-time position. This entry-level role involves...  ...to 20 hours of remote work weekly and requires passing a basic qualification exam. #J-18808-Ljbffr TELUS Digital AI Data Solutions
    Remote job
    Part time

    TELUS Digital AI Data Solutions

    New York, NY
    2 days ago
  •  ...Straive is a global leader in enterprise‑grade data analytics and AI solutions, committed to empowering businesses across various...  ...fueled by our partnership with EQT. Website: Linkedin Job Title: Evaluator - Political Science Location: Remote (USA) Job Type: Contract... 
    Contract work
    Remote work
    Worldwide

    Straive

    New York, NY
    3 days ago
  • $150 per hour

    JOB DESCRIPTION The Home and Community-Based Program evaluates preschoolers for special education for the Committee on Special Education (CPSE). Education Evaluations are one type of evaluation completed. Education Evaluations include assessment of children in five learning... 
    Immediate start
    Work from home

    US Diversity Job Search

    New York, NY
    4 days ago
  • About the Role We’re looking for contract, remote Turkish-speaking annotators to evaluate AI-generated content with a sharp eye for detail and cultural nuance. You’ll evaluate whether outputs are coherent, consistent, and aligned with real-world expectations in Turkish.... 
    Contract work
    Temporary work
    Freelance
    Immediate start
    Remote work

    BAM Ventures

    New York, NY
    6 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Swahili Evaluation AI Evaluator. Be the first to apply!