Receipt and Invoice Understanding Model Evaluator [Remote]
AuraOne Human Data
- Remote job
Receipt and Invoice Understanding Model Evaluator is a remote evaluation track for reviewing receipt and invoice understanding model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn receipt and invoice understanding model evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate receipt and invoice understanding model evaluation model outputs against a versioned rubric and assign severity tags for Receipt and Invoice Understanding Model Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on receipt and invoice understanding model evaluation or adjacent content for Receipt and Invoice Understanding Model Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two receipt and invoice understanding model evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Receipt and Invoice Understanding Model evaluation
- Multimodal evaluation
- Cross-modal reasoning
- Grounding review
- Receipt
- Invoice
- Understanding
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$20 per hour
...public sources and external tools. Generate high-quality human evaluation data by identifying response strengths, areas for improvement,... ...quality, clarity, tone, and completeness of responses. Ensure model responses align with expected conversational behavior and system...SuggestedRemote jobContract workPart timeSummer work$65 per hour
Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience...SuggestedHourly paySelf employmentWork from homeFlexible hours- ...Medical Document OCR Model Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Formal Logic Model Evaluator is a remote review track for evaluating AI outputs across formal logic model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Refusal Preference Reward Model Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Policy Preference Reward Model Evaluator is a remote review track for evaluating AI outputs in policy review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can...Remote jobHourly payFor contractors10 hours per week
$52k - $57k
...Job Description Job Description POSITION TITLE: Program Evaluator REPORTS TO: Director of Quality BROAD FUNCTION: Collects... ...: Articulates and applies historical context of racism and understands the current reality of consumers and communities of color in...Full timeSummer workLocal area$79.4k - $119.1k
...development team environment as Project Lead on a mix of complex Full Model Change (FMC) and Minor Model Change (MMC) developments.... ...effectively with staff and management ~ Conceptual understanding of product development flow (Design Development Supply Chain and...Full timeTemporary workWork experience placementWork at officeRemote workRelocation package- ...customers across the United States and globally at WOMEN'S FIT MODEL: This is a part-time opportunity to be a part of the design... ...Prior fit modeling experience is a plus but not required. • Understanding of brand history and its aesthetic. • Ability to provide...Part timeFlexible hours
- ...We are hiring for: Family Model Provider Type: Family Model Provider (TN) - Independent Contractor If you are a positive... ...include, but are not limited to: Seeking to understand the individual in the context of their personal history, their...Full timeContract workFor contractorsLive inWork from home
- ...processing a variety of documents, invoices, orders, etc.; processing a... ...inquiries Receive and receipt monies Operate a... ...assigned tasks Ability to understand and complete oral and... ...within acceptable ranges. Evaluation The subsitute shall...Permanent employmentWork at office
$123.33k - $161.76k
...Forensic Evaluator Title: Forensic Evaluator State Role Title: Psych III/Psychology... ...knowledge of psychopathology and its treatment, models of behavior change and management, the... ...You will be provided a confirmation of receipt when your application and/or résumé is...Work experience placementWork at officeLocal areaRemote work$14.5 per hour
...AI Web Search Evaluator As a Web Search Evaluator, you will play a key role in improving the quality of search engine results, ensuring... ...Fluent in English (written and spoken) Strong understanding of pop culture in the US Reliable computer system and internet...Hourly payPart timeCurrently hiringImmediate startRemote workWork from home10 hours per weekFlexible hours- ...demanding AI workloads. We provide high-performance GPU compute and Model API services, enabling AI companies, research labs, and... ...compute, AI/ML platforms, or a related technical domain. ~ Strong understanding of the GPU cloud market, AI/ML workflows, model training/...Remote jobFull timeFlexible hours
$60 - $90 per hour
...Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation... ...in Python and Git . Preferred Basic understanding of reinforcement learning . Experience in AI training...Full timeContract workSummer workRemote work- ...Product Manager to own AI inference and model serving for k0rdent AI, our control plane... ...performance infrastructure products and understands how production systems behave under real... ..., identifying performance bottlenecks, evaluating system design trade-offs, and translating...Remote work
$298k - $368k
...continuously learning from large scale real-world data, to (2) develop models and model training at scale, to (3) analyze real-world... ...in low-latency on-device inference techniques and a deep understanding of hardware acceleration. ~ Extensive experience with deep learning...Full timeRemote work$204k - $259k
...S. states. The core challenge within Model Lifecycle is accelerating Waymo's ML development... ...ready for efficient model training and evaluation. Develop infrastructure to produce... ..., and core infrastructure teams to understand user needs and deliver impactful ML workflows...Full timeTemporary workRemote work- ...remote position, open to Missouri Residents only. The Project Evaluator-GLS plays a key role in evaluating the Garrett Lee Smith (GLS)... ...and a people-centered perspective. Value collaboration and understand the importance of listening to the experiences and...Remote work
- ...JOB DESCRIPTION Build Your Future at DTCC As a Model Risk Management Intern, you will gain exposure to DTCC’s enterprise-wide... ...documentation, and governance activities while building a practical understanding of model risk frameworks and standards. At DTCC, interns...Hourly payFull timeSummer workInternshipSummer internshipWork at officeRemote work
- ...Thai Bilingual Expert to contribute to AI training by evaluating Thai audio content for nativeness and quality. This contractor... ...in English, follow guidelines, and help improve models' Thai-language understanding. No prior AI experience is required, but attention to detail...Remote jobFor contractors
- ...experience assessment company seeks a Freelance Luxury Brand Evaluator in Newtown Square. This role requires evaluating high-end automotive... ...assessments. Candidates should be 18+, possess a strong understanding of the automobile industry, and have keen observational...FreelanceFlexible hours
- Freelance Luxury Brand Evaluator Automotive Project - Newtown Square Are you a luxury automobile enthusiast who appreciates the finer... .... Requirements Must be 18 years of age or older. Good understanding of the automobile industry. Passionate about automobiles and...FreelanceWorldwideFlexible hours
- ...by our partnership with EQT. Website: Linkedin Job Title: Evaluator - Political Science Location: Remote (USA) Job Type: Contract... ...responses. Familiarity with AI and Technology: have a basic understanding of generative AI usage pertaining to history information processing...Contract workRemote workWorldwide
- ...System Security Officer (ISSO) / Control Evaluator to support a comprehensive enterprise... ...cybersecurity professional with a deep understanding of NIST security frameworks, federal information... ...within 10 business days of receipt, and escalate stakeholder unresponsiveness...Work at officeRemote work
- ...contract, remote French-speaking annotators to evaluate AI-generated content with a sharp eye... ...structured feedback that improves model performance over time What We’re Looking... ...Native or fluent French speaker Strong understanding of language and cultural context Ability...Contract workTemporary workFreelanceImmediate startRemote work
- ...AI Labs is seeking a Polish Bilingual Expert to contribute to an AI training project focused on Polish language understanding and generation. You will evaluate Polish audio content, assess linguistic quality, and provide detailed feedback in English. No AI experience is...Remote job
$24 per hour
Prolific is seeking an AI Trainer with advanced Tamil fluency to evaluate AI models' understanding of the Tamil language's emotional and cultural nuances. Responsibilities include assessing audio clips, reviewing tones for cultural relevancy, and ensuring quality control...Remote jobFlexible hours- ...to contribute to a global AI training project focused on Dutch language understanding and generation. You will evaluate AI-generated speech and provide detailed feedback to support language model improvement. No prior AI experience is required. Strong Dutch language expertise...Remote job
$20 per hour
...company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and excellent...Remote jobFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Receipt and Invoice Understanding Model Evaluator [Remote]. Be the first to apply!





