Table Image Reasoning Model Evaluator [Remote]
AuraOne Human Data
- Remote job
Table Image Reasoning Model Evaluator is a remote evaluation track for reviewing table image reasoning model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn table image reasoning model evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate table image reasoning model evaluation model outputs against a versioned rubric and assign severity tags for Table Image Reasoning Model Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on table image reasoning model evaluation or adjacent content for Table Image Reasoning Model Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two table image reasoning model evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Table Image Reasoning Model evaluation
- Multimodal evaluation
- Cross-modal reasoning
- Grounding review
- Table
- Image
- Reasoning
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- ...Quant Finance Reasoning Model Evaluator is a remote review track for evaluating AI outputs across finance and risk workflows. Reviewers grade calculations, narrative reasoning, and policy adherence; flag compliance and reconciliation issues; and document the correct treatment...SuggestedRemote jobHourly payFor contractorsWork experience placement10 hours per week
$40 per hour
...firm is looking for experienced quantitative professionals to evaluate AI-generated work and design problems for AI training. This fully... ...AI systems' development. Join us to shape the future of AI reasoning while working from anywhere in the US, Canada, UK, Ireland, Australia...SuggestedHourly payRemote workFlexible hours$60 per hour
...professionals to help advance AI development. AI models are increasingly capable of performing complex analytical and scientific reasoning — but these systems still need... ...state-of-the-art AI models on tasks like evaluating AI-generated quantitative analysis, solving...SuggestedHourly payFull timeRemote workFlexible hours$15 - $20 per hour
...external tools. ~Generate high-quality human evaluation data by identifying response strengths,... ...improvement, and factual inaccuracies. ~Assess reasoning quality, clarity, tone, and completeness of responses. ~Ensure model responses align with expected conversational...SuggestedPart timeSummer work- ...seeking an experienced Investment Banking SME to support the development, evaluation, and improvement of advanced AI models in finance. You will assess AI-generated analyses for accuracy, reasoned judgments, and data integrity, guiding model refinements and prompts for...SuggestedRemote job
$50 - $60 per hour
...team. The role involves training AI chatbots, providing complex problems for evaluation, and ensuring high-quality outputs from financial models. Ideal candidates have strong financial reasoning abilities and may hold advanced degrees. This contract position offers...Hourly payContract workTemporary workRemote workFlexible hours$40 per hour
A forward-thinking AI development firm seeks experienced quantitative professionals to evaluate AI-generated work, applying their skills in statistical analysis, predictive modeling, and technical writing. This fully remote opportunity offers a flexible schedule and projects...Hourly payRemote workFlexible hours$40 per hour
A leading AI development company is seeking experienced quantitative professionals to evaluate AI-generated analyses and design quantitative problems for AI training. This fully remote role offers flexibility in project selection and scheduling, with competitive pay starting...Hourly payRemote work$40 per hour
A leading AI development company is seeking experienced quantitative professionals to work remotely. In this role, you'll evaluate AI-generated quantitative work and solve technical problems while providing feedback to shape AI systems. Qualifications include 2+ years...Hourly payRemote workFlexible hours$40 per hour
...A leading AI development firm is seeking experienced quantitative professionals to join their team remotely. The role involves evaluating AI-generated quantitative work, providing insights, and shaping the future of AI systems. Candidates should have over two years of...Hourly payRemote work$40 per hour
...A leading AI development firm is seeking experienced quantitative professionals to evaluate AI-generated quantitative analysis and provide impactful feedback. This fully remote role allows for flexible scheduling and competitive pay starting at $40 per hour. Candidates...Hourly payRemote workFlexible hours$40 per hour
A leading AI development firm is seeking experienced quantitative professionals to evaluate AI-generated quantitative work and provide critical feedback. This role offers the flexibility of remote work, allowing you to set your own schedule while focusing on impactful...Hourly payRemote work$40 per hour
...A leading AI company in the United States is seeking experienced quantitative professionals to evaluate and validate AI-generated analytical work. This fully remote position allows you to set your own schedule, with competitive hourly pay starting at $40 USD. Responsibilities...Hourly payRemote work- micro1 is seeking an AI Image & Video Evaluation Specialist (remote, contractor) to generate and compare AI-generated visuals across platforms. You will assess realism, composition, lighting, color, anatomy, and text rendering, delivering structured analyses. No formal...Remote jobFor contractors
- AuraOne is seeking an Arabic Legal AI Evaluation Specialist — Multiple-Choice QA Review (Remote) to assess AI-generated legal outputs. You will review citations, statutory reasoning, and policy adherence, flag risk, and document corrected analyses for training data. Remote...Remote jobFor contractors10 hours per week
$20 - $30 per hour
...Role Overview Evaluate images to train next-generation AI systems by applying careful visual review and clear written reasoning. You will support a client image assessment project as a remote... ...image assessments that shape how models learn and perform. No prior AI experience...Remote jobHourly payFor contractors- ...Medical Imaging Safety Evaluator is a remote clinical-review track for evaluating AI outputs that... ...review. Reviewers grade differential reasoning, dosing logic, and guideline adherence... ...corrected clinical reasoning so the modeling team can close the gap. Why this role...Remote jobHourly payFor contractors10 hours per week
- ...Location: Remote | Work from Home Employment Type: Project-based | Contract We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists with strong proficiency in English. In this role, you will support...Contract workRemote workWork from homeMonday to FridayDay shift
- Productive Playhouse seeks AI Evaluators to support evaluating AI chatbots by interacting with models, assessing capabilities, safety, and usefulness. This is a project-based, task-based engagement with flexible hours and batch deliveries. Open to freelancers outside the...Remote jobFreelanceFlexible hours
- ...Scientific Figure Understanding Model Evaluator is a remote review track for evaluating AI outputs across scientific figure understanding model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results...Remote jobHourly payFor contractors10 hours per week
$80 - $120 per hour
...Role Overview Evaluate AI-generated finance and accounting deliverables for technical accuracy, analytical rigor, and presentation... ...against domain-specific quality rubrics for accuracy, completeness, reasoning, and presentation. Identify factual errors, calculation...Hourly payWork at officeRemote work$15 - $20 per hour
...external tools . Generate high-quality human evaluation data by identifying response strengths,... ..., and factual inaccuracies. Assess reasoning quality, clarity, tone, and completeness of responses. Ensure model responses align with expected conversational...Contract workSummer workRemote work- ...Policy Preference Reward Model Evaluator is a remote review track for evaluating AI outputs in policy review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can...Remote jobHourly payFor contractors10 hours per week
- AuraOne is seeking a remote Probability Model Evaluator to review prompts and model outputs against a quality rubric. You will compare paired... ...to help retrain the system. The role emphasizes clear reasoning, attention to detail, and the ability to calibrate against gold...Remote jobFor contractorsFlexible hours
$50 - $100 per hour
DataAnnotation is seeking experienced legal professionals to join as contract contributors to annotate and improve AI-generated legal research outputs. This fully remote, project-by-project role offers pay between $50 and $100+ per hour and is open to candidates in the...Hourly payContract workTemporary workRemote work- Alignerr is seeking mechanical engineering specialists to help train and improve cutting-edge AI models. Your deep technical knowledge will shape how AI reasons about fluid dynamics, thermodynamics, structural mechanics, and related fields as you collaborate with top AI...Remote jobHourly payContract work10 hours per week
$20 per hour
...contractors to join our team and teach AI chatbots. You will develop complex prompts to test AI models, write high-quality responses to demonstrate excellence, and evaluate different model outputs based on accuracy and style guidelines. This role is ideal for professionals...Hourly payFull timeContract workPart timeFor contractorsSelf employmentFreelanceRemote work$20 per hour
...A technology company specializing in AI is seeking a Copywriter to evaluate and improve AI model outputs. This role is remote, offering both full-time and part-time options with an hourly starting pay of $20+. The ideal candidate is detail-oriented, fluent in English,...Hourly payFull timeContract workPart timeRemote work$50 - $60 per hour
...A tech-focused company is seeking a Private Banker to train AI models and evaluate their performance. The ideal candidate has a strong background in finance, including skills in financial analysis and modeling. This is a remote position offering flexible hours with payment...Hourly payRemote workFlexible hours$85 per hour
...$85/hour Location: Remote Role Responsibilities Use frontier AI coding agents to complete and evaluate complex engineering tasks. Review model-generated mobile application code for correctness, quality, maintainability, and performance. Identify bugs...Contract workSummer workRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Table Image Reasoning Model Evaluator [Remote]. Be the first to apply!




