Preference Taxonomy Reward Model Evaluator [Remote]
AuraOne Human Data
- Remote job
Preference Taxonomy Reward Model Evaluator is a remote evaluation track for reviewing preference taxonomy reward model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn preference taxonomy reward model evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate preference taxonomy reward model evaluation model outputs against a versioned rubric and assign severity tags for Preference Taxonomy Reward Model Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on preference taxonomy reward model evaluation or adjacent content for Preference Taxonomy Reward Model Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two preference taxonomy reward model evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Preference Taxonomy Reward Model evaluation
- Preference ranking
- RLHF
- Rater calibration
- Preference
- Taxonomy
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- ...Refusal Preference Reward Model Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack... ...with prompt-injection, jailbreak, and policy-bypass taxonomies. Reliable async availability for at least 10 hours per...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Policy Preference Reward Model Evaluator is a remote review track for evaluating AI outputs in policy review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can...SuggestedRemote jobHourly payFor contractors10 hours per week
$20 per hour
...external tools. Generate high-quality human evaluation data by identifying response strengths,... ...and completeness of responses. Ensure model responses align with expected... ...requiring structured analytical thinking Preferred Experience with RLHF, model evaluation,...SuggestedRemote jobPart timeSummer work$182.4k - $273.6k
...scalable, production ready models and AI driven decision... ...for quality, evaluation rigor, and production... ...governance of metric taxonomies, thresholds, validation... ...degree. Master's or Ph.D. preferred in Machine Learning, Applied... ...for employees. Other rewards may include short-term...SuggestedTemporary workWork at officeRemote workShift work3 days per week- Dorado is seeking an experienced Investment Banking SME to support the development, evaluation, and improvement of advanced AI models in finance. You will assess AI-generated analyses for accuracy, reasoned judgments, and data integrity, guiding model refinements and prompts...SuggestedRemote job
- ...Medical Document OCR Model Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers... ...safety issues with severity tags that match clinical-incident taxonomies. Capture the corrected clinical reasoning so the modeling...Remote jobHourly payFor contractors10 hours per week
$178.5k - $257.83k
...Distinguished Scientist, Transgenic Model and TechnologyLocation:... ...efficacy studies• Identify, evaluate, and implement new technologies... ...rare diseases, neuroscience preferred)Leadership & Strategic... ...Enjoy a thoughtful, well-crafted rewards package that recognizes your...Full timeContract workWork at officeRemote workFlexible hours$19.25 per hour
...operate a company vehicle. Assist in Fruit Evaluation Management team duties Monitor crop... ...with a specific emphasis in agriculture preferably within a production/manufacturing setting... ...engaging employee community groups cash rewards for healthy habits and fitness reimbursements...Seasonal workWork at officeLocal areaWorldwideFlexible hoursAfternoon shift$60 - $90 per hour
...Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract... ...learning ideas, focusing on reward functions and training behavior. Evaluate... ...Proficiency in Python and Git . Preferred Basic understanding of...Full timeContract workSummer workRemote work- ...Audit function! It is our vision to be a preferred advisor to the business by building... ...more about you!Support Internal Audit’s evaluation of model and artificial intelligence (AI) risk... ...appreciation for our teams, who are rewarded with highly competitive pay and generous...InternshipMonday to Friday
$50 - $60 per hour
A technology company specializing in AI and finance is seeking a Wealth Advisor to help train AI models. In this independent contract role, you will measure the effectiveness of AI chatbots by solving complex financial problems. The ideal candidate should be fluent in...Remote jobHourly payContract work$65 per hour
Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience...Hourly paySelf employmentWork from homeFlexible hours- A luxury brand evaluation company is seeking a Luxury Brand Evaluator to assess customer experiences with premium brands. The role offers flexibility in choosing assignments and entails evaluating services in-store or online. Ideal candidates are detail-oriented, observant...Flexible hours
$36 - $40 per hour
...Acentra Health is looking for a Clinical Evaluator - PRN to join our growing team Job... ...commonly requiring long-term care placement. Preferred Qualifications 2 years of experience... ...Benefits Benefits are a key component of your rewards package. Our benefits are designed to...Hourly payContract workReliefWork at officeLocal areaRemote workWork from homeHome office$265k - $360k
...is seeking an experienced and motivated Model Sales and Strategy Lead to join our U.S.... ...Bachelor’s degree required; MBA and or CFA preferred.Knowledge of fixed income, multi-sector... ...a total compensation approach when rewarding employees which includes a base salary and...Full timeHome officeFlexible hours$110.6k - $178k
Principal Model Based Design Engineer - HybridThis is a hybrid position that requires onsite... ...Control Systems, or a related field (PhD preferred).10+ years of experience in control... ...that salary is only one component of total rewards at Stanley Black & Decker. The salary range...Full timeLocal area- ...vehicle design. As a Senior Virtual Hardware Model Engineer, you have the unique... ...related to software/hardware connectivity. Preferred experience in translating physical phenomena... ...your ambitions. Learn how GM supports a rewarding career that rewards you personally by...Full timeLocal areaWork from homeRelocation package
$75.33k - $125.5k
...Model Risk AnalystAt the Federal Home Loan Bank of Chicago, employees come first - that... ...or interest-rate model techniques preferred;Experience using machine learning techniques... ...preferredAt FHLBank Chicago, we believe in rewarding our high performing workforce. We offer...Work experience placementWork at officeRemote work$20 per hour
Feedinkoo is looking for a Web Developer/Designer to enhance AI models by evaluating design work, including interfaces and user experiences. This role involves reviewing AI‑generated visuals and providing feedback to improve users' experience with AI tools. Working remotely...Remote job- ...Formal Logic Model Evaluator is a remote review track for evaluating AI outputs across formal logic model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method...Remote jobHourly payFor contractors10 hours per week
- ...Quant Finance Reasoning Model Evaluator is a remote review track for evaluating AI outputs across finance and risk workflows. Reviewers grade calculations, narrative reasoning, and policy adherence; flag compliance and reconciliation issues; and document the correct treatment...Remote jobHourly payFor contractorsWork experience placement10 hours per week
$120k - $140k
SENIOR MANAGER OF SALES/NEW MODEL REQUIRED: Automotive experience... ...equipment. · Monitor and evaluate warranty concerns to ensure that... ...arise. HOW YOU WILL BE REWARDED: · Medical, Dental, Vision... ...or related Engineering field, preferred · 15 years of automotive...Work at officeRemote workVisa sponsorshipNight shift$60 - $90 per hour
...multi-step machine learning evaluation tasks for a leading generative... ...to determine where frontier models succeed or fail. Typical tasks... ...example modifying how an RL reward is computed and specifying success... ...and policy training is preferred. Prior experience in AI training...Hourly payFull timeFreelanceRemote work- ...This is an opportunity for experienced Model Based System Engineers who are highly motivated... ...Developing and driving MBSE model taxonomy and ontology consensus with stakeholders... ...based experience in a team environment Preferred: Experience with the practical...Full timeWork at officeRemote work
- ...The Mental Health Evaluator (MHE) is a critical role dedicated to delivering comprehensive, clinically appropriate, culturally competent... ...services setting as a member of a multi-disciplinary team, preferred. EDUCATION ~ Masters Degree in Social Work or Mental Health...
$135.6k - $237.4k
Role Description The Director, Model Engineering & Operations is... ...monitoring practices. ~Guide the evaluation and responsible adoption of... ..., Data Science, Engineering preferred. ~Equivalent years of... ...Substantial and comprehensive total rewards package. Working...Full timeWork experience placementWork at office$117.3k - $226.9k
...control, and range systems. As the CAMEO Model-Based Systems Engineer - Space Force... ...Advanced degree in a STEM field is strongly preferred Directed multi‑disciplinary teams... ...competitive compensation package where you’ll be rewarded based on your performance and recognized...Full timeWork at officeImmediate startRemote workRelocation packageFlexible hours$86.84k - $139.36k
...industry best practices. The Non-Model/End-User-Computing Tool (EUC)... ...lead, plan, implement, and evaluate program/project activities to... ...information with discretion Preferred Qualifications Non-Model/EUC... ...- and so will you. Our Total Rewards Package Our Total Rewards package...Local areaWork from homeFlexible hours- ...conversion. Qualifications Education and Training: Degree in a health-related field from an accredited college or university preferred. Current licensure in physical therapy or occupational therapy is required. Prior marketing and/or rehabilitation/LTACH experience...Full timeLocal area
- ...skills• Thorough knowledge of psychopathology and its treatment, models of behavior change and management, the psychotherapeutic... ...Values Veterans (V3) certified employer, that provides hiring preference to qualified veterans and service members. We highly encourage...Work experience placementWork at officeLocal areaRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Preference Taxonomy Reward Model Evaluator [Remote]. Be the first to apply!





