Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Receipt and Invoice Understanding Model Evaluator [Remote]

AuraOne Human Data

Remote
  • Remote job

Receipt and Invoice Understanding Model Evaluator is a remote evaluation track for reviewing receipt and invoice understanding model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.

Why this role matters

AI data reviewers help turn receipt and invoice understanding model evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.

Responsibilities

  • Evaluate receipt and invoice understanding model evaluation model outputs against a versioned rubric and assign severity tags for Receipt and Invoice Understanding Model Evaluator assignments.
  • Compare paired responses and pick the stronger answer with a written rationale.
  • Label hallucinations, instruction-following failures, and unsafe content with structured tags.
  • Capture ambiguous prompts and route them back to the program team for rubric updates.
  • Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
  • Document recurring failure modes so the modeling team can target them in the next training run.

Qualifications

  • Prior evaluation, annotation, or human-rater experience on receipt and invoice understanding model evaluation or adjacent content for Receipt and Invoice Understanding Model Evaluator work.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that names the issue and the rubric clause being applied.
  • Strong attention to detail and the ability to flag when a prompt itself is the problem.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Compare two receipt and invoice understanding model evaluation model responses to the same prompt and pick the stronger one with rationale.
  • Tag an unsafe response with the correct policy category and severity.
  • Audit a 50-row batch for rubric consistency and report drift to the program lead.
  • Propose a rubric clarification after spotting a recurring failure mode.

Nice to have

  • Background in linguistics, content moderation, or trust & safety review.
  • Experience with inter-rater agreement metrics and calibration cycles.
  • Domain expertise that lets you spot subject-matter errors automated checks miss.

Skills

  • Model output evaluation
  • Rubric-based annotation
  • Severity tagging
  • Inter-rater calibration
  • Receipt and Invoice Understanding Model evaluation
  • Multimodal evaluation
  • Cross-modal reasoning
  • Grounding review
  • Receipt
  • Invoice
  • Understanding

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 9 days ago
Similar jobs that could be interesting for youBased on the Receipt and Invoice Understanding Model Evaluator [Remote] in Remote vacancy
  • $20 per hour

     ...public sources and external tools. Generate high-quality human evaluation data by identifying response strengths, areas for improvement,...  ...quality, clarity, tone, and completeness of responses. Ensure model responses align with expected conversational behavior and system... 
    Suggested
    Remote job
    Contract work
    Part time
    Summer work

    Mercor

    New York, NY
    5 days ago
  • $65 per hour

    Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience... 
    Suggested
    Hourly pay
    Self employment
    Work from home
    Flexible hours

    Prolific

    Las Vegas, NV
    4 days ago
  •  ...Medical Document OCR Model Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    22 days ago
  •  ...Formal Logic Model Evaluator is a remote review track for evaluating AI outputs across formal logic model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    17 days ago
  •  ...Refusal Preference Reward Model Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    9 days ago
  •  ...Policy Preference Reward Model Evaluator is a remote review track for evaluating AI outputs in policy review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    15 days ago
  • $52k - $57k

     ...Job Description Job Description POSITION TITLE: Program Evaluator REPORTS TO: Director of Quality      BROAD FUNCTION:  Collects...  ...: Articulates and applies historical context of racism and understands the current reality of consumers and communities of color in... 
    Full time
    Summer work
    Local area

    Family Connections, Inc.

    East Orange, NJ
    9 days ago
  • $79.4k - $119.1k

     ...development team environment as Project Lead on a mix of complex Full Model Change (FMC) and Minor Model Change (MMC) developments....  ...effectively with staff and management ~ Conceptual understanding of product development flow (Design Development Supply Chain and... 
    Full time
    Temporary work
    Work experience placement
    Work at office
    Remote work
    Relocation package

    Honda Dev. and Mfg. of Am.,LLC

    Raymond, OH
    23 days ago
  •  ...customers across the United States and globally at WOMEN'S FIT MODEL: This is a part-time opportunity to be a part of the design...  ...Prior fit modeling experience is a plus but not required. • Understanding of brand history and its aesthetic. • Ability to provide... 
    Part time
    Flexible hours

    The Normal Brand

    Saint Louis, MO
    5 days ago
  •  ...We are hiring for: Family Model Provider Type: Family Model Provider (TN) - Independent Contractor If you are a positive...  ...include, but are not limited to: Seeking to understand the individual in the context of their personal history, their... 
    Full time
    Contract work
    For contractors
    Live in
    Work from home

    RHA Health Services

    Kingsport, TN
    2 days ago
  •  ...processing a variety of documents, invoices, orders, etc.; processing a...  ...inquiries Receive and receipt monies Operate a...  ...assigned tasks Ability to understand and complete oral and...  ...within acceptable ranges. Evaluation The subsitute shall... 
    Permanent employment
    Work at office

    MARYSVILLE SCHOOL DISTRICT 25

    Marysville, WA
    3 days ago
  • $123.33k - $161.76k

     ...Forensic Evaluator Title: Forensic Evaluator State Role Title: Psych III/Psychology...  ...knowledge of psychopathology and its treatment, models of behavior change and management, the...  ...You will be provided a confirmation of receipt when your application and/or résumé is... 
    Work experience placement
    Work at office
    Local area
    Remote work

    Virginia Department of Human Resource Management

    Crewe, VA
    2 days ago
  • $14.5 per hour

     ...AI Web Search Evaluator As a Web Search Evaluator, you will play a key role in improving the quality of search engine results, ensuring...  ...Fluent in English (written and spoken) Strong understanding of pop culture in the US Reliable computer system and internet... 
    Hourly pay
    Part time
    Currently hiring
    Immediate start
    Remote work
    Work from home
    10 hours per week
    Flexible hours

    Welo Data

    United States
    4 days ago
  •  ...demanding AI workloads. We provide high-performance GPU compute and Model API services, enabling AI companies, research labs, and...  ...compute, AI/ML platforms, or a related technical domain. ~ Strong understanding of the GPU cloud market, AI/ML workflows, model training/... 
    Remote job
    Full time
    Flexible hours

    Yotta Labs

    United States
    10 hours ago
  • $60 - $90 per hour

     ...Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation...  ...in Python and Git . Preferred Basic understanding of reinforcement learning . Experience in AI training... 
    Full time
    Contract work
    Summer work
    Remote work

    Mercor

    Remote
    10 hours ago
  •  ...Product Manager to own AI inference and model serving for k0rdent AI, our control plane...  ...performance infrastructure products and understands how production systems behave under real...  ..., identifying performance bottlenecks, evaluating system design trade-offs, and translating... 
    Remote work

    Mirantis

    United States
    4 days ago
  • $298k - $368k

     ...continuously learning from large scale real-world data, to (2) develop models and model training at scale, to (3) analyze real-world...  ...in low-latency on-device inference techniques and a deep understanding of hardware acceleration. ~ Extensive experience with deep learning... 
    Full time
    Remote work

    Waymo

    San Francisco, CA
    10 hours ago
  • $204k - $259k

     ...S. states. The core challenge within Model Lifecycle is accelerating Waymo's ML development...  ...ready for efficient model training and evaluation. Develop infrastructure to produce...  ..., and core infrastructure teams to understand user needs and deliver impactful ML workflows... 
    Full time
    Temporary work
    Remote work

    Waymo

    Remote
    10 hours ago
  •  ...remote position, open to Missouri Residents only.  The Project Evaluator-GLS plays a key role in evaluating the Garrett Lee Smith (GLS)...  ...and a people-centered perspective. Value collaboration and understand the importance of listening to the experiences and... 
    Remote work

    Compass Health Network

    Clinton, MO
    5 days ago
  •  ...JOB DESCRIPTION Build Your Future at DTCC As a Model Risk Management Intern, you will gain exposure to DTCC’s enterprise-wide...  ...documentation, and governance activities while building a practical understanding of model risk frameworks and standards. At DTCC, interns... 
    Hourly pay
    Full time
    Summer work
    Internship
    Summer internship
    Work at office
    Remote work

    DTCC

    Jersey City, NJ
    10 hours ago
  •  ...Thai Bilingual Expert to contribute to AI training by evaluating Thai audio content for nativeness and quality. This contractor...  ...in English, follow guidelines, and help improve models' Thai-language understanding. No prior AI experience is required, but attention to detail... 
    Remote job
    For contractors

    YO AI Labs

    Dallas, TX
    2 days ago
  •  ...experience assessment company seeks a Freelance Luxury Brand Evaluator in Newtown Square. This role requires evaluating high-end automotive...  ...assessments. Candidates should be 18+, possess a strong understanding of the automobile industry, and have keen observational... 
    Freelance
    Flexible hours

    CXG

    Newtown Square, PA
    3 days ago
  • Freelance Luxury Brand Evaluator Automotive Project - Newtown Square Are you a luxury automobile enthusiast who appreciates the finer...  .... Requirements Must be 18 years of age or older. Good understanding of the automobile industry. Passionate about automobiles and... 
    Freelance
    Worldwide
    Flexible hours

    CXG

    Newtown Square, PA
    3 days ago
  •  ...by our partnership with EQT. Website: Linkedin Job Title: Evaluator - Political Science Location: Remote (USA) Job Type: Contract...  ...responses. Familiarity with AI and Technology: have a basic understanding of generative AI usage pertaining to history information processing... 
    Contract work
    Remote work
    Worldwide

    Straive

    New York, NY
    2 days ago
  •  ...System Security Officer (ISSO) / Control Evaluator to support a comprehensive enterprise...  ...cybersecurity professional with a deep understanding of NIST security frameworks, federal information...  ...within 10 business days of receipt, and escalate stakeholder unresponsiveness... 
    Work at office
    Remote work

    VIATEQ Corporation

    Washington DC
    18 days ago
  •  ...contract, remote French-speaking annotators to evaluate AI-generated content with a sharp eye...  ...structured feedback that improves model performance over time What We’re Looking...  ...Native or fluent French speaker Strong understanding of language and cultural context Ability... 
    Contract work
    Temporary work
    Freelance
    Immediate start
    Remote work

    BAM Ventures

    New York, NY
    5 days ago
  •  ...AI Labs is seeking a Polish Bilingual Expert to contribute to an AI training project focused on Polish language understanding and generation. You will evaluate Polish audio content, assess linguistic quality, and provide detailed feedback in English. No AI experience is... 
    Remote job

    YO AI Labs

    Phoenix, AZ
    2 days ago
  • $24 per hour

    Prolific is seeking an AI Trainer with advanced Tamil fluency to evaluate AI models' understanding of the Tamil language's emotional and cultural nuances. Responsibilities include assessing audio clips, reviewing tones for cultural relevancy, and ensuring quality control... 
    Remote job
    Flexible hours

    Prolific

    New York, NY
    4 days ago
  •  ...to contribute to a global AI training project focused on Dutch language understanding and generation. You will evaluate AI-generated speech and provide detailed feedback to support language model improvement. No prior AI experience is required. Strong Dutch language expertise... 
    Remote job

    YO AI Labs

    Seattle, WA
    2 days ago
  • $20 per hour

     ...company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and excellent... 
    Remote job
    Flexible hours

    DataAnnotation

    New York, NY
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Receipt and Invoice Understanding Model Evaluator [Remote]. Be the first to apply!