Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Scene Understanding Evaluator [Remote]

AuraOne Human Data

Remote
  • Remote job

Scene Understanding Evaluator is a remote evaluation track for reviewing scene understanding evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.

Why this role matters

AI data reviewers help turn scene understanding evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.

Responsibilities

  • Evaluate scene understanding evaluation model outputs against a versioned rubric and assign severity tags for Scene Understanding Evaluator assignments.
  • Compare paired responses and pick the stronger answer with a written rationale.
  • Label hallucinations, instruction-following failures, and unsafe content with structured tags.
  • Capture ambiguous prompts and route them back to the program team for rubric updates.
  • Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
  • Document recurring failure modes so the modeling team can target them in the next training run.

Qualifications

  • Prior evaluation, annotation, or human-rater experience on scene understanding evaluation or adjacent content for Scene Understanding Evaluator work.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that names the issue and the rubric clause being applied.
  • Strong attention to detail and the ability to flag when a prompt itself is the problem.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Compare two scene understanding evaluation model responses to the same prompt and pick the stronger one with rationale.
  • Tag an unsafe response with the correct policy category and severity.
  • Audit a 50-row batch for rubric consistency and report drift to the program lead.
  • Propose a rubric clarification after spotting a recurring failure mode.

Nice to have

  • Background in linguistics, content moderation, or trust & safety review.
  • Experience with inter-rater agreement metrics and calibration cycles.
  • Domain expertise that lets you spot subject-matter errors automated checks miss.

Skills

  • Model output evaluation
  • Rubric-based annotation
  • Severity tagging
  • Inter-rater calibration
  • Scene Understanding evaluation
  • Computer vision
  • Spatial reasoning
  • Video QA
  • Scene
  • Understanding

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Scene Understanding Evaluator [Remote] in Remote vacancy
  • $50 per hour

    AI Trainer Jobs is seeking reviewers for SWE Front-End — Visual Document Understanding in a remote contractor role. You will evaluate document annotation outputs, apply a structured rubric, and provide clear rationales to help retrain models. Candidates should have strong... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    Flexible hours

    AI Trainer Jobs

    New York, NY
    4 days ago
  • Receipt and Invoice Understanding Model Evaluator is a remote evaluation track for reviewing receipt and invoice understanding model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind... 
    Suggested
    Hourly pay
    For contractors
    Remote work
    10 hours per week

    AI Trainer Jobs

    New York, NY
    4 days ago
  •  ...Scientific Figure Understanding Model Evaluator is a remote review track for evaluating AI outputs across scientific figure understanding model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    14 days ago
  • $37.5 per hour

     ...within Amazon originals. In this role you will be responsible for evaluating advertising copy generated by a large language model (LLM) to...  ...evaluating advertising content and creative qualityAn understanding of:Brand voice and brand guidelinesMarketing messagingCreative... 
    Suggested
    Contract work
    Temporary work
    Remote work

    TEKsystems

    New York, NY
    4 hours ago
  • $20 per hour

     ...company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and excellent... 
    Suggested
    Remote work
    Flexible hours

    DataAnnotation

    United States
    4 days ago
  • $52k - $57k

     ...POSITION TITLE: Program Evaluator REPORTS TO: Director of Quality      BROAD FUNCTION:   Collects, analyzes, and reports on...  ...PROFICIENCY: Articulates and applies historical context of racism and understands the current reality of consumers and communities of color in... 
    Full time
    Summer work
    Local area

    Family Connections

    East Orange, NJ
    1 day ago
  •  ...Hebrew Bilingual Experts to contribute to an AI training project focused on improving Hebrew-language understanding and generation. In this contractor role, you will evaluate Hebrew audio content, assess nativeness and linguistic quality, and provide detailed feedback... 
    Remote job
    For contractors

    YO AI Labs

    Atlanta, GA
    3 hours ago
  • Freelance Luxury Brand Evaluator Automotive Project - Greater Los Angeles Are you a luxury automobile enthusiast who appreciates the...  ...giants. Requirements: Must be 18 years of age or older. Good understanding of the automobile industry. Passionate about automobiles and... 
    Freelance
    Worldwide
    Flexible hours

    CXG

    Los Angeles, CA
    1 day ago
  •  ...seeking a Thai Bilingual Expert to contribute to AI training by evaluating Thai audio content for nativeness and quality. This...  ...English, follow guidelines, and help improve models' Thai-language understanding. No prior AI experience is required, but attention to detail... 
    Remote job
    For contractors

    YO AI Labs

    Dallas, TX
    2 hours ago
  •  ...AI Labs is seeking a Hebrew bilingual expert to contribute to an AI training project focused on Hebrew language understanding and generation. You will evaluate Hebrew audio, judge nativeness and linguistic quality, and provide detailed feedback.No prior AI experience... 
    Remote job
    For contractors
    Flexible hours

    YO AI Labs

    Seattle, WA
    2 hours ago
  •  ...AI Labs invites a Tamil Bilingual Expert to contribute to a language and AI training project focused on Tamil understanding and generation. You will evaluate Tamil audio content, assess linguistic quality, provide feedback, and ensure alignment with guidelines. No prior... 
    Remote job
    For contractors
    Flexible hours

    YO AI Labs

    New York, NY
    2 hours ago
  •  ...experience assessment company seeks a Freelance Luxury Brand Evaluator in Newtown Square. This role requires evaluating high-end automotive...  ...assessments. Candidates should be 18+, possess a strong understanding of the automobile industry, and have keen observational... 
    Freelance
    Flexible hours

    CXG

    Newtown Square, PA
    1 day ago
  • $36 per hour

     ...:Up to 6 monthsCommitment:20+ hours/week Role Responsibilities Evaluate AI-generated music across various genres and rate against detailed...  ..., or music journalist. Knowledge of the Hungarian music scene. Headphones or studio monitors suitable for critical listening.... 
    Remote job
    Summer work
    Immediate start
    Flexible hours

    United States Digital Space LLC

    New York, NY
    2 days ago
  • AI Trainer Jobs is seeking a remote contractor to evaluate receipt and invoice understanding model outputs against AuraOne's rubric. You will review prompts, compare paired responses, label edge cases, and write structured feedback the modeling team can use to retrain.... 
    Remote job
    For contractors

    AI Trainer Jobs

    New York, NY
    4 days ago
  •  ...expert to support a language and AI training project. You will evaluate Turkish audio for nativeness, fluency, pronunciation, and...  ...and actionable linguistic insights, facilitating improvements in Turkish language understanding and generation.J-18808-Ljbffr YO AI Labs
    Remote job
    For contractors

    YO AI Labs

    Dallas, TX
    4 days ago
  • $14.5 per hour

    Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that...  ...Requirements Fluent in English (written and spoken) Strong understanding of pop culture in the US Reliable computer system and internet... 
    Hourly pay
    Part time
    Immediate start
    Remote work
    Work from home
    10 hours per week
    Flexible hours

    Welo Data

    New York, NY
    2 days ago
  • Alignerr is seeking Wildlife and Habitat Conservation Scientists to evaluate AI-trained biodiversity protection content. This remote,...  ...flexible hours (10-40 per week) and a chance to shape how AI understands ecological challenges on a global scale. Ideal candidates have... 
    Remote job
    Hourly pay
    Contract work
    Flexible hours

    Alignerr Corp.

    Miami, FL
    3 days ago
  • $20 per hour

    Prolific is seeking fluent Norwegian speakers to act as evaluators for AI training. You will perform side-by-side evaluations of text and...  ..., assessing naturalness and authenticity to help models understand Norwegian language nuances. Pay is up to $20/hr with many tasks... 
    Remote job
    Flexible hours

    Prolific

    New York, NY
    2 hours ago
  •  ...Labs is seeking Dutch Bilingual Experts to contribute to a global AI training project focused on Dutch language understanding and generation. You will evaluate AI-generated speech and provide detailed feedback to support language model improvement. No prior AI experience... 
    Remote job

    YO AI Labs

    Seattle, WA
    2 hours ago
  •  ...Experts to support a language and AI training project. You will evaluate Turkish audio and language content, assess linguistic quality...  ...accurate audio reviews on time, contributing to better Turkish language understanding and generation. #J-18808-Ljbffr YO AI Labs
    Remote job
    For contractors

    YO AI Labs

    Phoenix, AZ
    2 days ago
  • $35.81 - $42 per hour

     ...Responsibilities Acentra Health is looking for a Part-time PASRR Evaluator to join our growing team. Job Summary: Join Acentra Health as...  ..., emotional, behavioral, and physical functioning. Read, understand, and adhere to all corporate policies including policies related... 
    Contract work
    Part time
    Work at office
    Local area
    Remote work
    Work from home

    Acentra

    Brooklyn, NY
    2 hours ago
  • $24 per hour

    Prolific is seeking an AI Trainer with advanced Tamil fluency to evaluate AI models' understanding of the Tamil language's emotional and cultural nuances. Responsibilities include assessing audio clips, reviewing tones for cultural relevancy, and ensuring quality control... 
    Remote job
    Flexible hours

    Prolific

    New York, NY
    2 days ago
  •  ...cutting-edge technology and extensive data-supported analysis. This role is ideal for a Certified residential appraiser with an understanding of valuation techniques. The Appraiser has the ability to complete valuations for most types of properties and complexities. They... 
    Full time
    Remote work
    Long distance

    True Footage

    Flint, MI
    3 days ago
  •  ...changes, and cost-to-cure items. Conduct field inspections and evaluate property characteristics, improvements, access, utilities,...  ...eminent-domain law, and right-of-way acquisition practices. Understanding of partial acquisitions, larger-parcel analysis, easements, severance... 
    Immediate start
    Remote work
    Relocation

    CORRE Inc

    Madison, WI
    5 days ago
  • $60k - $100k

     ...appraisal firm, founded and led by experienced appraisers who understand the industry firsthand. As we expand into new markets and strengthen...  ...report compensation. Appraisers begin with a 90-day period to evaluate performance, reliability, and alignment with company standards... 
    For contractors
    Remote work

    Banks Appraisal Group

    Seattle, WA
    3 days ago
  • $20 - $160 per hour

     ...Position: Generalist Annotator — Health AI Conversation Quality Evaluation Type: Contract Compensation: $20–$160/hour...  ...clarity of AI responses . Ensure the AI answers the question in understandable language for non-experts. Detect sycophancy in AI... 
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    8 days ago
  • $60k - $100k

     ...appraisal firm, founded and led by experienced appraisers who understand the industry firsthand. As we expand into new markets and strengthen...  ...report compensation. Appraisers begin with a 90-day period to evaluate performance, reliability, and alignment with company standards... 
    For contractors
    Immediate start
    Remote work

    Banks Valuation

    United States
    5 days ago
  • $60k - $100k

     ...Albany Valdosta About Us Banks Valuation is a growing appraisal firm, founded and led by experienced appraisers who understand the industry firsthand. As we expand into new markets and strengthen our presence in existing ones, we're looking for appraisers... 
    For contractors
    Remote work

    Banks Valuation

    United States
    2 days ago
  • Overview At PAM Health, we care for chronically and critically ill patients who require extended hospital care. PAM Health has over 80 hospital locations and employs over 11,000 people across the country. Our teams work together to deliver the highest level of compassionate...
    Full time
    Local area

    PAM Health

    Humble, TX
    4 days ago
  •  ...Physician Services PC · Mental HealthValhalla, NYAllied Health Prof/TechnicalPer DiemAll ShiftsAs neededJob Summary: The Mental Health Evaluator (MHE) is a critical role dedicated to delivering comprehensive, clinically appropriate, culturally competent, and trauma-informed... 

    HealthAlliance of the Hudson Valley

    Valhalla, NY
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Scene Understanding Evaluator [Remote]. Be the first to apply!