Receipt and Invoice Understanding Model Evaluator [Remote]
AuraOne Human Data
- Remote job
Receipt and Invoice Understanding Model Evaluator is a remote evaluation track for reviewing receipt and invoice understanding model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn receipt and invoice understanding model evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate receipt and invoice understanding model evaluation model outputs against a versioned rubric and assign severity tags for Receipt and Invoice Understanding Model Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on receipt and invoice understanding model evaluation or adjacent content for Receipt and Invoice Understanding Model Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two receipt and invoice understanding model evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Receipt and Invoice Understanding Model evaluation
- Multimodal evaluation
- Cross-modal reasoning
- Grounding review
- Receipt
- Invoice
- Understanding
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$20 per hour
...public sources and external tools. Generate high-quality human evaluation data by identifying response strengths, areas for improvement,... ...quality, clarity, tone, and completeness of responses. Ensure model responses align with expected conversational behavior and system...SuggestedRemote jobContract workPart timeSummer work$65 per hour
Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience...SuggestedHourly paySelf employmentWork from homeFlexible hours- ...Refusal Preference Reward Model Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Operations Research Model Evaluator is a remote review track for evaluating AI outputs across operations research model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Medical Document OCR Model Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Policy Preference Reward Model Evaluator is a remote review track for evaluating AI outputs in policy review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can...Remote jobHourly payFor contractors10 hours per week
$52k - $57k
...Job Description Job Description POSITION TITLE: Program Evaluator REPORTS TO: Director of Quality BROAD FUNCTION: Collects... ...: Articulates and applies historical context of racism and understands the current reality of consumers and communities of color in...Full timeSummer workLocal area$79.4k - $119.1k
...development team environment as Project Lead on a mix of complex Full Model Change (FMC) and Minor Model Change (MMC) developments.... ...effectively with staff and management ~ Conceptual understanding of product development flow (Design Development Supply Chain and...Full timeTemporary workWork experience placementWork at officeRemote workRelocation package$14.5 per hour
...AI Web Search Evaluator Welo Data works with technology companies to provide datasets... ..., and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages... ...English (written and spoken) Strong understanding of pop culture in the US Reliable...Hourly payPart timeCurrently hiringImmediate startRemote workWork from home10 hours per weekFlexible hours- ...customers across the United States and globally at WOMEN'S FIT MODEL: This is a part-time opportunity to be a part of the design... ...Prior fit modeling experience is a plus but not required. • Understanding of brand history and its aesthetic. • Ability to provide...Part timeFlexible hours
- ...Family Model Provider Type: Family Model Provider (TN) - Independent Contractor If you are a positive and personable individual... ...include, but are not limited to: Seeking to understand the individual in the context of their personal history, their...Full timeContract workFor contractorsLive inWork from home
- ...in career opportunities offered at MICA. General purpose: Models will pose for art students for the purpose of studying the human... ...sustained poses for the duration of the class as requested Understand all topics covered in the model training session Work within...
$60 - $90 per hour
...Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation... ...in Python and Git . Preferred Basic understanding of reinforcement learning . Experience in AI training...Full timeContract workSummer workRemote work- ...demanding AI workloads. We provide high-performance GPU compute and Model API services, enabling AI companies, research labs, and... ...compute, AI/ML platforms, or a related technical domain. ~ Strong understanding of the GPU cloud market, AI/ML workflows, model training/...Remote jobFull timeFlexible hours
$298k - $368k
...continuously learning from large scale real-world data, to (2) develop models and model training at scale, to (3) analyze real-world... ...in low-latency on-device inference techniques and a deep understanding of hardware acceleration. ~ Extensive experience with deep learning...Full timeRemote work$204k - $259k
...S. states. The core challenge within Model Lifecycle is accelerating Waymo's ML development... ...ready for efficient model training and evaluation. Develop infrastructure to produce... ..., and core infrastructure teams to understand user needs and deliver impactful ML workflows...Full timeTemporary workRemote work$220k - $320k
...trains and hosts specialized language models for companies that need frontier-quality... ...everything end-to-end: distillation, training, evaluation, and planet-scale hosting. We are a... ..., TensorRT-LLM, or similar) ~ Deep understanding of GPU architecture and experience...Full timeWork at office- ...JOB DESCRIPTION Build Your Future at DTCC As a Model Risk Management Intern, you will gain exposure to DTCC’s enterprise-wide... ...documentation, and governance activities while building a practical understanding of model risk frameworks and standards. At DTCC, interns...Hourly payFull timeSummer workInternshipSummer internshipWork at officeRemote work
- ...Thai Bilingual Expert to contribute to AI training by evaluating Thai audio content for nativeness and quality. This contractor... ...in English, follow guidelines, and help improve models' Thai-language understanding. No prior AI experience is required, but attention to detail...Remote jobFor contractors
- ...by our partnership with EQT. Website: Linkedin Job Title: Evaluator - Political Science Location: Remote (USA) Job Type: Contract... ...responses. Familiarity with AI and Technology: have a basic understanding of generative AI usage pertaining to history information processing...Contract workRemote workWorldwide
- OverviewFreelance Luxury Brand Evaluator Automotive Project - Pittsburgh - Join to apply for this role at CXG. Are you a luxury automobile... ...feedback.RequirementsMust be 18 years of age or older.Good understanding of the automobile industry.Passionate about automobiles and...FreelanceFlexible hours
- ...to contribute to a global AI training project focused on Dutch language understanding and generation. You will evaluate AI-generated speech and provide detailed feedback to support language model improvement. No prior AI experience is required. Strong Dutch language expertise...Remote job
$20 per hour
...company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and excellent...Remote jobFlexible hours- ...System Security Officer (ISSO) / Control Evaluator to support a comprehensive enterprise... ...cybersecurity professional with a deep understanding of NIST security frameworks, federal information... ...within 10 business days of receipt, and escalate stakeholder unresponsiveness...Work at officeRemote work
- ...experience assessment company seeks a Freelance Luxury Brand Evaluator in Newtown Square. This role requires evaluating high-end automotive... ...assessments. Candidates should be 18+, possess a strong understanding of the automobile industry, and have keen observational...FreelanceFlexible hours
$24 per hour
Prolific is seeking an AI Trainer with advanced Tamil fluency to evaluate AI models' understanding of the Tamil language's emotional and cultural nuances. Responsibilities include assessing audio clips, reviewing tones for cultural relevancy, and ensuring quality control...Remote jobFlexible hours- ...contract, remote French-speaking annotators to evaluate AI-generated content with a sharp eye... ...structured feedback that improves model performance over time What We’re Looking... ...Native or fluent French speaker Strong understanding of language and cultural context Ability...Contract workTemporary workFreelanceImmediate startRemote work
- ...AI Labs is seeking a Polish Bilingual Expert to contribute to an AI training project focused on Polish language understanding and generation. You will evaluate Polish audio content, assess linguistic quality, and provide detailed feedback in English. No AI experience is...Remote job
- ..., remote Turkish-speaking annotators to evaluate AI-generated content with a sharp eye for... ...structured feedback that improves model performance over time What We’re Looking... ...Native or fluent Turkish speaker Strong understanding of Turkish language and cultural context...Contract workTemporary workFreelanceImmediate startRemote work
$50 - $60 per hour
...artificial intelligence. Contributors review examples, test model behavior, evaluate responses, and identify errors so AI systems become more... ...and sound judgment in rubric-based evaluation and QA Understanding of KPI definitions and business or finance reporting practices...Hourly payPart timeFor contractorsRemote workWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Receipt and Invoice Understanding Model Evaluator [Remote]. Be the first to apply!





