Remote Strategy Evaluator for AI Output Quality
Mercor
- Remote job
Mercor is seeking a General business strategy / management Evaluator for a remote position. You will evaluate AI-generated documents and provide structured feedback to enhance quality. The ideal candidate has 5+ years experience in a relevant field and professional fluency in English. Required skills include proficiency in Microsoft Office and Google Workspace, particularly in Slides. An advanced degree is preferred. Join us to contribute your expertise and help improve AI outputs. #J-18808-Ljbffr Mercor
- Obsidian is looking for expert Evaluators in Finance operations/audit support to review AI-generated work products for accuracy and quality. This remote hourly position requires a minimum of 5... ...Your role will involve evaluating outputs and providing structured feedback....Remote jobQualityHourly payWork at office
- Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback... ...is essential. This is a remote, hourly engagement. #J-188...Remote jobQualityHourly payWork at office
- Obsidian is hiring expert Evaluators in Investment analysis / valuation / credit to review AI-generated work products for accuracy and quality. This remote, hourly position requires deep subject-matter expertise and professional fluency in English to provide structured...Remote workQualityHourly payWork at office
$80 - $120 per hour
...This role evaluates AI-generated documents, spreadsheets, and slide decks... ...accuracy, logical rigor, domain quality, and presentation polish. You... ...-tool skills to grade outputs and deliver clear, structured feedback. This is a remote, hourly engagement. Key Responsibilities...Remote workQualityHourly payWork at office$80 - $120 per hour
...Role Overview Review AI-generated documents, spreadsheets... ...accuracy, rigor, domain quality, and presentation, then... ...to support high-quality outputs. Key Responsibilities Evaluate AI-generated artifacts... ...preferred. Work Terms ~ Remote, hourly engagement....Remote workQualityHourly payWork at office$80 - $120 per hour
...Role Overview Assess AI-generated government and... ...rigor, and domain-specific quality. Use subject-matter expertise to score outputs and deliver clear,... ...spreadsheets, and slide decks. Evaluate outputs using domain-... .... Work Terms Remote engagement. Paid on an...Remote workQualityHourly payWork at office$80 - $120 per hour
...Role Overview Assess AI-generated data analysis deliverables... ...rigor, and domain quality. You will use deep... ...matter expertise to score outputs against rubrics and deliver... ...institution. Work Terms Remote, hourly engagement. You will evaluate AI-generated deliverables...Remote workQualityHourly payWork at office$80 - $120 per hour
...healthcare and clinical expertise to evaluate AI-generated documents, spreadsheets,... ...decks for accuracy, rigor, and domain quality. You will grade outputs against domain-specific quality... ...reputable institution. Work Terms ~ Remote, hourly engagement....Remote workQualityHourly payWork at office$80 - $120 per hour
...Role Overview Evaluate AI-generated operations work products, applying... ...accuracy, rigor, and domain quality. The role focuses on reviewing... ...capacity planning. Assess outputs for factual accuracy,... ...institution. Work Terms Remote engagement. Hourly contract...Remote workQualityHourly payContract workWork at office$80 - $120 per hour
...Role Overview Assess AI-generated procurement and... ..., rigor, and domain quality, applying deep subject-... ...matter expertise to grade outputs and provide actionable... ...Key Responsibilities Evaluate AI-generated artifacts... ...institution. Work Terms Remote engagement, paid on an...Remote workQualityHourly payContract workWork at office$80 - $120 per hour
...Role Overview Assess AI-generated people operations... ...rigor, and domain quality. Use your subject-matter expertise to grade outputs and deliver actionable,... ...Work Terms This is a remote, hourly engagement. Candidates... ...on an hourly basis for evaluation work. Compensation...Remote workQualityHourly payWork at office$80 - $120 per hour
...Role Overview Assess AI-generated documents, spreadsheets... ..., rigor, and domain quality within customer success... ...expertise to grade outputs and deliver concise,... ...written feedback. This is a remote, hourly engagement.... ...and overall quality. Evaluate outputs against domain-...Remote workQualityHourly payWork at office- Obsidian is seeking expert Evaluators in Privacy / regulatory compliance to review AI-generated work products for accuracy and quality. The role is remote and requires 5+ years of relevant experience... ...apply your expertise to grade outputs, identify errors, and provide...Remote jobQualityHourly payWork at office
- ...Social Science Research Assistant — AI Research Output Evaluator is a remote review track for evaluating AI outputs across social science research assistant... ..., papers, or community standards. Maintain reviewer-quality scores in inter-rater calibration cycles....Remote jobQualityHourly payFor contractors10 hours per week
$20 - $30 per hour
A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong...Remote jobQualityHourly pay$80 - $120 per hour
...contracts / diligence / redlines Evaluator to assess AI-generated artifacts. This role... ...ideal candidate will evaluate quality rubrics, provide structured... ...work independently to improve AI outputs. Compensation is $80-$120/hour with a fully remote work environment. #J-18808-...Remote jobQuality- AuraOne is seeking a Government Benefits AI Evaluator to remotely review prompts and model outputs against AuraOne's quality rubric. You will label edge cases and provide structured feedback to retrain the model. This contractor role emphasizes clear written reasoning,...Remote jobQualityFor contractors10 hours per week
- Turing is seeking detail-oriented AI Analysts based in the United States for a Google Wallet evaluation project. This role allows you to engage with advanced AI tools... ...AI. You will evaluate model responses, review output quality, and provide structured feedback. The position...Remote jobQualityFull timeContract work
- AuraOne is seeking a remote Classroom Scenario AI Evaluator to review prompts and responses against our quality rubric. You will compare paired outputs, label edge cases, and provide structured feedback to retrain the model. This contractor role requires prior evaluation...Remote jobQualityHourly payFor contractors10 hours per week
$24 per hour
Prolific is seeking an AI Trainer with advanced Tamil fluency to evaluate AI models' understanding of the Tamil... ...relevancy, and ensuring quality control in audio outputs. The role requires fluent Tamil... ...tasks, with flexible hours and remote work options. #J-18808-Ljbffr...Remote jobQualityFlexible hours$30 per hour
...Compensation: $20-$30/hour Location: Remote Commitment: 10-40 hours/week Role Responsibilities Evaluate outputs from large language models and... ...using defined rubrics and quality standards. Review multi-step... ...experience in LLM evaluation, AI output analysis, QA/testing,...Remote jobQualityHourly payContract work- ...role We are hiring expert Evaluators in Compliance / regulatory... ...response with financial-services AI to review and assess AI-... ..., rigor, and domain quality. You will apply deep subject... ...‑matter expertise to grade outputs. This is a remote, hourly engagement. Requirements...Remote workQualityHourly payWork at office
- Obsidian is hiring expert Evaluators in Real estate, hospitality, and events to review AI-generated work for accuracy, rigor, and domain quality. This remote position requires deep expertise and involves grading outputs like documents and presentations. Applicants must...Remote jobQualityWork at office
- ...Investment Banking SME to support the development, evaluation, and improvement of advanced AI models in finance. You will assess AI-generated... ...guiding model refinements and prompts for high-quality outputs. The role is remote, project-based, and demands strong domain...Remote jobQuality
- ...oriented contractors to develop prompts, test AI chatbots, and evaluate outputs. You will craft diverse conversations, write high-quality responses, and assess model performance... .... Projects offer 5-40 hours weekly with remote work and flexible scheduling. #J-18808-Ljbffr...Remote jobQualityFor contractorsSelf employmentFlexible hours
$80 - $120 per hour
Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and... ...written feedback to improve AI outputs. ~Apply deep subject-matter... ...Compensation: $80–$120/hour ~Location: Remote Application Process ~...Remote workQualityPart timeWork at office- ...project that defines what excellent AI-assisted marketing work looks like... ..., with strong writing, brand strategy, and familiarity with AI training or evaluation. This is a highly collaborative, fast-moving effort focused on quality and defensible scoring. #J-18808-...Quality
$60 - $70 per hour
...technical talent with leading AI research labs.... ...$70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses... ...compliance, and overall quality. Review content involving... ...benchmarking . Identify unsafe outputs, hallucinations,...Remote workQualityContract workSummer work$70 per hour
...technical talent with leading AI research labs.... ...$70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses... ...reasoning and rigor in model outputs . Provide structured... ...improve training data quality and downstream performance...Remote workQualityContract workSummer work$80 - $120 per hour
Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and... ...feedback to improve AI model outputs. ~Apply deep subject-matter... ...Compensation: $80–$120/hour ~Location: Remote Application Process ~...Remote workQualityPart timeWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Remote Strategy Evaluator for AI Output Quality. Be the first to apply!
- work from home social media evaluator New York, NY
- work from home web search evaluator New York, NY
- social media evaluator New York, NY
- evaluator New York, NY
- quality evaluator New York, NY
- education evaluator New York, NY
- clinical evaluator New York, NY
- ai evaluator New York, NY
- program evaluator New York, NY
- senior devops engineer remote New York, NY




