Table Image Reasoning Model Evaluator [Remote]
AuraOne Human Data
- Remote job
Table Image Reasoning Model Evaluator is a remote evaluation track for reviewing table image reasoning model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn table image reasoning model evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate table image reasoning model evaluation model outputs against a versioned rubric and assign severity tags for Table Image Reasoning Model Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on table image reasoning model evaluation or adjacent content for Table Image Reasoning Model Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two table image reasoning model evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Table Image Reasoning Model evaluation
- Multimodal evaluation
- Cross-modal reasoning
- Grounding review
- Table
- Image
- Reasoning
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- ...Quant Finance Reasoning Model Evaluator is a remote review track for evaluating AI outputs across finance and risk workflows. Reviewers grade calculations, narrative reasoning, and policy adherence; flag compliance and reconciliation issues; and document the correct treatment...SuggestedRemote jobHourly payFor contractorsWork experience placement10 hours per week
- Handshake is seeking an AI Image Evaluator to assess prompts and generated images for quality, accuracy, and policy compliance... ...visual details, compare outputs, and explain your reasoning to help train evaluation models. Remote US role with flexible scheduling, part of...SuggestedRemote jobFlexible hours
$20 per hour
...external tools. Generate high-quality human evaluation data by identifying response strengths, areas... ...improvement, and factual inaccuracies. Assess reasoning quality, clarity, tone, and completeness of responses. Ensure model responses align with expected conversational...SuggestedRemote jobContract workPart timeSummer work$50 - $60 per hour
...remote team in the United States. This role involves training AI models, measuring their progress, and evaluating outputs to enhance their quality. Ideal candidates should possess expert-level financial reasoning and proficiency in financial analysis. This is a flexible...SuggestedRemote jobHourly payFlexible hours- micro1 is seeking an AI Image & Video Evaluation Specialist (remote, contractor) to generate and compare AI-generated visuals across platforms. You will assess realism, composition, lighting, color, anatomy, and text rendering, delivering structured analyses. No formal...SuggestedRemote jobFor contractors
- ...Medical Imaging Safety Evaluator is a remote clinical-review track for evaluating AI outputs that... ...review. Reviewers grade differential reasoning, dosing logic, and guideline adherence... ...corrected clinical reasoning so the modeling team can close the gap. Why this role...Remote jobHourly payFor contractors10 hours per week
- ...Job Description Job Description Breast Imager, Northern California, University Town Flexible Work Arrangement Hybrid Model, 1099 vs W2, Part Time or Full Time Join a physician-owned group committed to early cancer detection and precision diagnostics. This...Permanent employmentFull timeTemporary workPart timeWork at officeRemote workWork from homeRelocation packageFlexible hours2 days per week3 days per week
- ...Scientific Figure Understanding Model Evaluator is a remote review track for evaluating AI outputs across scientific figure understanding model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results...Remote jobHourly payFor contractors10 hours per week
- ...Policy Preference Reward Model Evaluator is a remote review track for evaluating AI outputs in policy review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can...Remote jobHourly payFor contractors10 hours per week
- **Job Title: AI Trainer || Image Quality Evaluator || English** **Location**: Remote | Work from Home **Employment Type:** Project-based | Contract We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists...Contract workRemote workWork from homeMonday to FridayDay shift
- Mercor is hiring Legal Experts to evaluate AI-generated responses for employment and labor law scenarios. This fully remote, hourly contract... ...week. You will assess accuracy, provide feedback to improve model behavior and participate in calibration sessions. Requirements...Remote jobHourly payContract workFlexible hours
$50 - $60 per hour
A technology company specializing in AI and finance is seeking a Wealth Advisor to help train AI models. In this independent contract role, you will measure the effectiveness of AI chatbots by solving complex financial problems. The ideal candidate should be fluent in...Remote jobHourly payContract work$65 per hour
Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience...Hourly paySelf employmentWork from homeFlexible hours- ...Job Description Job Description Breast Imager, Northern California, University Town Flexible Work Arrangement Hybrid Model, 1099 vs W2, Part Time or Full Time Join a physician-owned group committed to early cancer detection and precision diagnostics. This...Full timePart timeWork at officeRemote workWork from homeRelocation packageFlexible hours2 days per week3 days per week
$23.2 per hour
...help us improve Roblox systems. As a Human Evaluator you will have the opportunity to provide... ...the review and classification of text, image, video, scripts, and audio Track and... ...state or local laws. Roblox also provides reasonable accommodations to candidates with...Hourly payFull timeContract workWork experience placementWork at officeLocal areaRemote workMonday to Friday- ...Refusal Preference Reward Model Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety...Remote jobHourly payFor contractors10 hours per week
- ...writing high-quality prompts and model responses, or recording high-... ...classifying, and structuring documents, tables, and other content to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy, reasoning quality, coherence, and cultural...Hourly payFor contractorsRemote workFlexible hours
- ...skills• Thorough knowledge of psychopathology and its treatment, models of behavior change and management, the psychotherapeutic... ...to be reviewed within five (5) days of the date of this posting.Reasonable accommodations are available to persons with disabilities during...Work experience placementWork at officeLocal areaRemote work
- ...property characteristics, measuring buildings, and capturing property images. The appraiser confirms property locations using maps and... ...protected under local, state, or federal laws. If you require reasonable accommodation for any part of the application or hiring process...Work experience placementWork at officeLocal area
- ...reinforcement learning, and large language models. We offer generous relocation... ...Analyze large multi-domain datasets such as images, text and/or graph data, to identify statistically... ...move/lift items weighing up to 25 lbs. Reasonable accommodations may be made to enable...Full timeTemporary workWork at officeLocal areaRemote workVisa sponsorshipRelocation packageFlexible hours
- ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and...Full timeContract workRemote workFlexible hours
$52k - $57k
...POSITION TITLE: Program Evaluator REPORTS TO: Director of Quality BROAD FUNCTION: Collects, analyzes, and reports on data to evaluate... ...encounters while performing the essential functions of this job. Reasonable accommodations may be made to enable individuals with...Full timeSummer workLocal area$115.5k - $218.1k
In this position...New Model Introduction Supervisor, Vehicle and Powertrain LaunchThe New Model Program Supervisor is responsible for... ...protected veteran status. In the United States, if you need a reasonable accommodation for the online application process due to a...Immediate startFlexible hours$115k - $200k
...The AVP, Acquisition Fraud Strategy and Model Monitoring, is a multi-functional role within... ...responsibilities include supporting the evaluation of new fraud models, fraud and... ...have an award-winning culture for all. Reasonable Accommodation Notice: Federal law requires...Full timeWork experience placementWork from homeVisa sponsorshipWork visaMonday to Friday$185k - $400k
.... We are seeking accomplished Research Scientists in Foundation Models with expertise in pre-training and mid-training large-scale multimodal... ...for large-scale multimodal pre-training/mid-training (text, image, audio, and video), and drive innovative approaches for...Remote work$160k - $327k
...role advancements and constructive feedback through regular evaluations.• Work-Life Balance Support• Strong and inclusive organizational... ...reside in the USWhere required by law, NTT DATA provides a reasonable range of compensation for specific roles. The starting pay range...Temporary workRemote workFlexible hours$80 - $120 per hour
...Compensation: $80 - $120 per hour We are hiring expert Evaluators in Special education / IEP to review and assess AI-generated... ...without regard to legally protected characteristics and provide reasonable accommodations upon request. Contract and Payment Terms...Hourly payContract workFor contractorsWork at officeRemote work- ...We are hiring for: Family Model Provider Type: Family Model Provider (TN) - Independent Contractor If you are a positive... ...RHA is an equal opportunity employer. In addition, we provide reasonable accommodation to qualified employees who have protected disabilities...Full timeContract workFor contractorsLive inWork from home
$15 - $80 per hour
...anatomy expertise to annotate detailed hand and finger landmarks in image and video data, helping improve advanced AI systems through... ...anatomy expertise. Excellent attention to detail, visual reasoning, and image and video analysis skills. Ability to follow protocols...Hourly payContract workRemote work- ...A leading research accelerator is seeking a contractor to evaluate North American early to mid-teen humor. The role involves reviewing short-form humorous content and articulating the reasons behind its appeal using defined criteria. Ideal candidates should be within the...For contractorsFreelanceRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Table Image Reasoning Model Evaluator [Remote]. Be the first to apply!




