Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Assessment Rubric AI Evaluator [Remote]

AuraOne Human Data

Remote
  • Remote job

Assessment Rubric AI Evaluator is a remote review track for evaluating AI outputs across assessment rubric ai specialist operations workflows. Reviewers grade workflow correctness, policy adherence, and stakeholder fit; flag operational risk; and document the right next step so the modeling team can train on it.

Why this role matters

Assessment Rubric AI specialist operations AI has to fit into an actual day at work. AuraOne uses experienced operators to grade outputs the way a senior peer would — checking workflow, policy, and the unwritten rules that decide whether a task actually gets done.

Responsibilities

  • Review AI outputs against current assessment rubric ai specialist operations workflows, playbooks, and firm policy for Assessment Rubric AI Evaluator assignments.
  • Grade tone, escalation logic, and stakeholder fit on a structured rubric.
  • Flag operational risk, missed escalations, and policy-adherence gaps with severity tags.
  • Capture the right next step so the modeling team can train on it.
  • Adjudicate disputed treatments against published playbooks or firm guidance.
  • Maintain reviewer-quality scores in inter-rater calibration cycles.

Qualifications

  • Direct working experience in assessment rubric ai specialist operations on real teams for Assessment Rubric AI Evaluator work.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that names the policy or workflow being applied.
  • Strong attention to detail and the ability to flag when a prompt itself is the problem.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Grade a model's response to a real assessment rubric ai specialist operations ticket and rate workflow, tone, and escalation.
  • Flag a missed escalation with the right severity tag and corrected next step.
  • Adjudicate a disputed playbook call between two reviewers using firm guidance.
  • Audit a 25-row batch for rubric consistency and report drift to the program lead.

Nice to have

  • Prior experience training, calibrating, or QA-ing operations teams.
  • Familiarity with AI-assisted workflow tooling and its failure modes.
  • Bilingual experience for cross-region operations.

Skills

  • Operational review
  • Policy adherence
  • Workflow judgment
  • Stakeholder communication
  • Assessment Rubric AI specialist operations
  • Learning design
  • Assessment review
  • Pedagogy
  • Assessment
  • Rubric

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Assessment Rubric AI Evaluator [Remote] in Remote vacancy
  • $30 per hour

     ...hours/week Role Responsibilities Evaluate outputs from large language...  ...autonomous agent systems using defined rubrics and quality standards. Review...  ...and reasoning traces, to assess accuracy and completeness....  ...experience in LLM evaluation, AI output analysis, QA/testing, UX... 
    Suggested
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    3 days ago
  • $15 - $20 per hour

     ...Role Overview Assess Marathi AI-generated responses for factual accuracy, reasoning, clarity, tone, and completeness, and produce clear...  ...identifies strengths and specific areas for improvement. Your evaluations will be used to help create the "perfect AI-generated response... 
    Suggested
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    United States
    4 days ago
  • $15 - $20 per hour

     ...Role Overview Evaluate AI-generated responses in Bengali, identify factual errors and areas for improvement, and produce clear English...  ...strengths, areas for improvement, and factual inaccuracies. Assess reasoning quality, clarity, tone, and completeness of model responses... 
    Suggested
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    13 days ago
  • $20 per hour

    A tech company specializing in AI is looking for a Digital Web Designer to evaluate AI-generated designs and help train models for better aesthetic understanding...  ...applicants in the United States. The role involves assessing UI/UX designs and providing feedback to improve AI... 
    Suggested
    Remote work

    DataAnnotation

    Helena, MT
    1 day ago
  • $40 per hour

     ...cybersecurity professionals to join our team to help train AI models. In this role, you will evaluate AI-generated security content, solve technical...  ...outputs. You will work directly with advanced AI models to assess their accuracy, strengthen their reasoning, and contribute... 
    Suggested
    Hourly pay
    Full time
    Part time
    Remote work

    DataAnnotation

    Indiana, PA
    5 days ago
  •  ...Seeking a PhD-level expert in Quantitative Finance, the full-time Remote AI Research Evaluator will assess AI-generated financial content, craft relevant questions, and evaluate responses, contributing to the training of advanced AI models in a flexible contract role.... 
    Full time
    Contract work
    Remote work
    Flexible hours

    Virtual Vocations Inc

    United States
    1 day ago
  • $40 per hour

     ...cybersecurity training company is seeking experienced professionals to evaluate AI-generated security content and solve technical problems. This...  ...in the US, Canada, UK, Ireland, Australia, or New Zealand and must pass a short assessment to access paid work. #J-18808-Ljbffr
    Hourly pay
    Remote work

    DataAnnotation

    Lincoln, NE
    5 days ago
  • $40 per hour

    A cutting-edge AI design firm is seeking a Digital Designer to evaluate and enhance AI-generated designs. This role involves critiquing UI/UX visuals, assessing aesthetic quality, and contributing to various design-focused projects. Candidates should have a strong background... 
    Remote work

    DataAnnotation

    Providence, RI
    5 days ago
  • $20 per hour

    A cutting-edge AI company is seeking a Digital Designer to evaluate and enhance AI understanding of design principles. This position offers flexible remote work...  ...Designer, you'll review AI-generated designs and assess their quality, helping shape the future of AI tools... 
    Remote work
    Flexible hours

    DataAnnotation

    Vermont
    5 days ago
  • $20 per hour

    A leading technology company is seeking a Digital Designer to assess and improve AI-generated designs while working remotely. The role involves reviewing UI/UX outputs, providing critical feedback, and training AI models to enhance aesthetics and usability. Candidates are... 
    Hourly pay
    Remote work
    Flexible hours

    DataAnnotation

    Florida, NY
    5 days ago
  • $60 per hour

     ...and contribute to developing cutting-edge AI systems, while enjoying the flexibility...  ...-of-the-art AI models on tasks like evaluating AI-generated quantitative analysis, solving...  ...up for an account, you'll take a short assessment (this serves as our version of an interview... 
    Hourly pay
    Full time
    Remote work
    Flexible hours

    DataAnnotation

    Kansas City, MO
    5 days ago
  • $80 - $120 per hour

    Role Description ~Evaluate AI-generated artifacts against domain-specific quality rubrics. ~Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks. ~Provide clear, structured written feedback to improve AI model outputs.... 
    Summer work
    Work at office

    Mercor

    Remote
    18 days ago
  • $20 - $80 per hour

     ...apply your expertise to help train next-generation AI systems. Your work will shape how models learn,...  ...quality, real-world input. Key Responsibilities: Evaluate and score AI-generated responses using well-defined rubrics and metrics across diverse subject areas.... 
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    1 day ago
  • RWS - TrainAI is seeking a part-time Search Quality Rater to help assess how our clients’ search engines respond to everyday queries. Based in the United States, you will work remotely up to 29 hours per week, maintaining quality standards and meeting SLAs while using... 
    Remote job
    Part time
    Work from home
    Home office

    aitrainer

    New York, NY
    4 days ago
  • A cybersecurity solutions company is seeking experienced professionals to train AI models through evaluating and improving AI-generated security content. Responsibilities include assessing technical cybersecurity problems, providing feedback for AI systems, and contributing... 
    Remote job
    Flexible hours

    DataAnnotation

    Madison, WI
    1 day ago
  • $40 per hour

    A cybersecurity technology company in New York is seeking experienced professionals to evaluate AI-generated security content. The role involves assessing AI models, solving technical problems, and contributing to AI systems' improvement. Candidates should have 2+ years... 
    Remote job
    Hourly pay

    DataAnnotation

    Florida, NY
    1 day ago
  • Alignerr is seeking a Population Health Informaticist to collaborate with AI research teams and enhance health AI understanding of large-scale data. You will assess AI outputs in epidemiology, health disparities, and surveillance data, providing structured feedback to... 
    Remote job
    Hourly pay
    Flexible hours

    Alignerr

    Tampa, FL
    1 day ago
  • $20 - $30 per hour

    A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong... 
    Remote job
    Hourly pay

    Crossing Hurdles

    New York, NY
    3 days ago
  • $80 - $120 per hour

    Mercor is seeking a Remote Clinical Evaluator to evaluate AI-generated artifacts against domain-specific quality rubrics. The ideal candidate should have over 5 years of experience in the Clinical/biomedical/pharma field and proficiency in Microsoft Office and Google Workspace... 
    Remote job
    Contract work
    Work at office

    Mercor

    New Bremen, OH
    3 days ago
  • $60 per hour

     ...and contribute to developing cutting-edge AI systems, while enjoying the flexibility...  ...-of-the-art AI models on tasks like evaluating AI-generated security content, solving technical...  ...up for an account, you'll take a short assessment (this serves as our version of an... 
    Remote job
    Hourly pay
    Full time
    Flexible hours

    DataAnnotation

    Nevada, IA
    4 days ago
  •  ...Educational Safety AI Evaluator is a remote evaluation track for reviewing educational safety...  ...responses against AuraOne's quality rubric. Reviewers compare paired outputs, label...  ...Safety AI evaluation Learning design Assessment review Pedagogy Educational... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    2 days ago
  • $40 per hour

     ...is seeking experienced cybersecurity professionals for a remote role focused on evaluating AI-generated security content and solving technical problems. Responsibilities include assessing AI outputs for accuracy and providing necessary feedback to enhance AI systems. Candidates... 
    Remote job
    Hourly pay
    Flexible hours

    DataAnnotation

    Sioux Falls, SD
    1 day ago
  • $40 per hour

    A cybersecurity firm is seeking experienced professionals to evaluate AI-generated security content and solve technical cybersecurity problems. You'll train AI models by assessing their outputs and providing feedback to improve real-world threat handling. The ideal candidate... 
    Remote job
    Hourly pay
    Full time
    Part time

    DataAnnotation

    Saint Paul, MN
    3 days ago
  •  ...Audit and Controls AI Evaluator is a remote review track for evaluating AI outputs across...  ...experience. Comfort applying multi-page rubrics consistently across long batches....  ...review Regulatory standards Risk assessment Audit Quantitative finance and economics... 
    Remote job
    Hourly pay
    For contractors
    Work experience placement
    10 hours per week

    AuraOne Human Data

    Remote
    5 days ago
  • A leading AI evaluation firm is seeking medical experts to train AI models and evaluate their performance. This remote role allows you to...  ..., and strong attention to detail. Responsibilities include assessing AI chatbot outputs related to healthcare, ensuring medical accuracy... 
    Remote job
    Hourly pay

    DataAnnotation

    New York, NY
    5 days ago
  •  ...seeking a Generalist fluent in English and Assamese to work remotely. You will be responsible for fact-checking, generating evaluation data, and assessing model responses for clarity and completeness. The ideal candidate holds a Bachelor's degree, has significant experience... 
    Remote job

    Mercor

    San Francisco, CA
    5 days ago
  •  ...Healthcare Compliance AI Evaluator is a remote clinical-review track for evaluating AI outputs...  ...and red-flag handling on a structured rubric. Flag patient-safety issues with...  ...reading patient cases and writing structured assessments. Comfort applying multi-page rubrics... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    2 days ago
  • $50 per hour

     ...transcription, annotation, and evaluation of Hindi audio and video to...  ...contexts, and create detailed rubrics and grading guidelines in both...  ...prompts through language models, assess generated outputs against...  ...language variation. Interest in AI, language models, or applied... 
    Hourly pay
    Temporary work
    Remote work
    10 hours per week

    SaidGig

    United States
    6 days ago
  • $70 per hour

     ...Position: AI Model Assessment Specialist Type: Contract Compensation: $22 - $70/hour Location: Remote Commitment: 10-40 hrs/week Role Responsibilities Evaluate and critique the performance and accuracy of AI-generated content across various domains. Deliver detailed, constructive... 
    Contract work
    Remote work

    Crossing Hurdles

    New York, NY
    5 days ago
  • $14.5 per hour

     ...diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data...  ...? Join our team as a Web Search Evaluator and help shape the future of search engines...  ...Web Search Evaluator, you will: Review and assess internet search results, ensuring users receive... 
    Bi-weekly pay
    Hourly pay
    Part time
    Immediate start
    Remote work
    Work from home
    Flexible hours

    Jobs for Humanity

    Jacksonville, FL
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Assessment Rubric AI Evaluator [Remote]. Be the first to apply!