Product management / roadmap / PRD Evaluator
$80 - $120 per hourAI Trainer Jobs
Product management / roadmap / PRD Evaluator is a remote evaluation track for reviewing product management / roadmap / prd evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain. Category: Frontier Model Evaluation · Pay: $80–$120 / hr · Location: Remote — US-eligible · Contractor About the role Product management / roadmap / PRD Evaluator is a remote evaluation track for reviewing product management / roadmap / prd evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain. AI data reviewers help turn product management / roadmap / prd evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data. Review frontier model outputs. Judge benchmark failures and calibrate other evaluators. Responsibilities Evaluate product management / roadmap / prd evaluation model outputs against a versioned rubric and assign severity tags for Product management / roadmap / PRD Evaluator assignments. Compare paired responses and pick the stronger answer with a written rationale. Label hallucinations, instruction-following failures, and unsafe content with structured tags. Capture ambiguous prompts and route them back to the program team for rubric updates. Role details Track Evaluation & annotation Work model Remote · Independent specialist contractor Compensation $80–$120 / hr Eligible from US What you should bring Prior evaluation, annotation, or human-rater experience on product management / roadmap / prd evaluation or adjacent content for Product management / roadmap / PRD Evaluator work. Comfort applying multi-page rubrics consistently across long batches. Clear written reasoning that names the issue and the rubric clause being applied. Strong attention to detail and the ability to flag when a prompt itself is the problem. Reliable async availability for at least 10 hours per week. Example tasks Compare two product management / roadmap / prd evaluation model responses to the same prompt and pick the stronger one with rationale. Tag an unsafe response with the correct policy category and severity. Audit a 50-row batch for rubric consistency and report drift to the program lead. Propose a rubric clarification after spotting a recurring failure mode. Useful experience Background in linguistics, content moderation, or trust & safety review. Experience with inter-rater agreement metrics and calibration cycles. Domain expertise that lets you spot subject-matter errors automated checks miss. Compensation and schedule $80–$120 / hr Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation. Skills used in matching Model output evaluation Rubric-based annotation Severity tagging Inter-rater calibration Product management / roadmap / PRD evaluation Application boundary Creating a specialist profile records your experience and preferences. Starting role intake is a separate action that attaches this role to your candidate record. Specialist intake The intake preserves your chosen role, the visible terms, and source attribution for reviewer context. 01 Confirm profile and eligibility details. 02 Attach this role deliberately. 03 Receive a human review decision or follow-up. Placement timing depends on program demand and reviewer confirmation. #J-18808-Ljbffr AI Trainer Jobs
- AuraOne is seeking a Workflow Annotator—Product Management & Marketing to remotely review evaluation prompts and responses against the company's quality rubric. You will compare paired outputs, label edge cases, and provide structured feedback the modeling team can use...SuggestedRemote jobFor contractors
$160k - $210k
BVAL (Bloomberg's Evaluated Pricing Service) Evaluator - US Agency Structured Products Location New York Business Area Product Ref # 10053894... ...clients — including mutual funds, hedge funds, money managers, internal pricing teams, and auditors — rely on...SuggestedPrice workTemporary workFor contractorsWork experience placement$80 - $120 per hour
Incident management / reliability / SRE Evaluator is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit test the model...SuggestedFor contractorsRemote work10 hours per week- About the roleWe are hiring expert Evaluators in Document/deck production QA to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs...SuggestedRemote jobHourly payWork at office
- ...serve our clients globally by providing evaluated pricing and analytics on over 3 million... ...financial concepts with quantitative skills to manage large volumes of data to develop pricing... ...and processes. You will leverage your product and market knowledge to identify growth...SuggestedWorldwide
- ...methodologiesExhibit an ability to pay attention to detail in order to provide a higher-quality valuation product for the clientExercise the ability to effectively manage daily activities in order to efficiently establish a reasonable capacity for assignmentsEnsure high...Full timeRemote workLong distance
$75k - $80k
The Vocational Evaluation Supervisor manages and supervises the Vocational Evaluation team, oversees the Individualized Vocational Assessment Plans... .... Implements internal controls to confirm that work and production are consistent with regular policies, procedures, and...- ...our clients globally by providing them evaluated pricing on over two million fixed income... ...client interaction Continuously improve product and service quality through daily market... ...skills to interact with clients, portfolio managers, traders, research, and sales Strong...
- ...serve our clients globally by providing evaluated pricing and analytics on over 3 million... ...financial concepts with quantitative skills to manage large volumes of data to develop pricing... ...and processes. You will leverage your product and market knowledge to identify growth...Worldwide
$80 - $120 per hour
Product launch / experiment readiness Evaluator is a remote evaluation track for reviewing product launch / experiment readiness evaluation prompts and responses against AuraOne's quality rubric. Category: Frontier Model Evaluation · Pay: $80-$120 / hr · Location: Remote...For contractorsRemote work10 hours per week- Mercor is seeking expert Evaluators in product launch / experiment readiness to review AI-generated artifacts (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs.This is a remote...Remote jobHourly payWork at office
- Mercor is partnering with a leading AI research organization to engage experienced UI/UX and product designers for a project focused on evaluating how well AI systems perform real-world digital product design work. You will define what excellent work looks like: designing...
- Mercor is partnering with a leading AI research organization to engage experienced UI/UX and product designers for a project that evaluates how well AI systems perform real-world digital product design work. You will define what excellent work looks like by designing task...
$80 - $120 per hour
...Summers , and Jack Dorsey . Position: Incident management / reliability / SRE Evaluator Type: Contract Compensation: $80–$120/hour... ...structured written feedback to improve AI-generated work products. Apply deep subject-matter expertise in Incident management...Contract workSummer workWork at officeRemote work$80 - $120 per hour
...Summers , and Jack Dorsey . Position: Procurement / vendor management Evaluator Type: Contract Compensation: $80–$120/hour... ...clear, structured written feedback to improve AI-generated work products. Apply deep subject-matter expertise to grade outputs in...Contract workSummer workWork at officeRemote work$37.5 per hour
...within Amazon originals. In this role you will be responsible for evaluating advertising copy generated by a large language model (LLM) to... ...adherenceContextual relevance between the advertised brand/product and Prime Video contentCreativity and diversity of generated copyStructural...Contract workTemporary workRemote work$80 - $120 per hour
...role is for one of our clients Compensation: $80 - $120 per hour We are hiring expert Evaluators in Special education / IEP to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You...Hourly payContract workFor contractorsWork at officeRemote work- AI Trainer Jobs is seeking a remote Management Consultant — AI Workflow Evaluator to review prompts and model outputs against AuraOne's quality rubric. You will compare paired outputs, label edge cases, and write structured feedback the modeling team can use to retrain....Remote jobFor contractors
$20 - $30 per hour
...focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI... ...workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong analytical skills,...Remote jobHourly pay- Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates will leverage their subject matter expertise to assess various documents, spreadsheets, and slide decks....Remote jobHourly payWork at office
- Obsidian is seeking expert Evaluators in Humanities / arts / culture to review AI-generated work products for accuracy, rigor, and domain quality. You will apply subject-matter expertise to grade outputs and provide structured feedback. This is a remote, hourly engagement...Remote jobHourly payWork at office
- Obsidian is looking for expert Evaluators to review AI-generated work products in Public-sector procurement and RFI response. In this remote hourly role, you will assess the accuracy and quality of documents, spreadsheets, and slide decks, applying your deep subject-matter...Remote jobHourly payWork at office
- ...deployment of advanced AI systems. We are seeking a contractor to evaluate chatbot responses across real-world small business scenarios,... ..., data-driven insights, and transparent transcripts to inform product decisions. You will work remotely on a 16-week project, collaborating...Remote jobFor contractorsFreelance
- Mercor is seeking expert Evaluators in Clinical / biomedical / pharma to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. This is a remote, hourly engagement. You will apply deep subject-matter...Remote jobHourly pay
$30 per hour
...Location: Remote Commitment: 10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous agent systems... ..., actionable feedback to support model refinement and product improvements. Participate in calibration sessions to ensure consistent...Remote jobHourly payContract work- Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Remote jobWork at office
- AI Trainer Jobs is seeking an Equity Research AI Evaluator to remotely review AI outputs in finance and risk workflows. You will grade calculations, narrative reasoning, and policy adherence, flag issues, and document corrected workpapers for training. Prerequisites include...Remote jobFor contractors
- Mercor is seeking expert Evaluators in People ops / recruiting to review AI-generated work products for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote, hourly engagement. You will evaluate artifacts...Remote jobHourly pay
- ...and rubrics for AI training in journalism, editing, and video production. You will draft and source reference materials, then grade AI... ...editing, writing, or video production experience and comfort evaluating professional-grade work to benchmark AI performance. #J-18808...
- ...Trainer Jobs in the United States seeks individuals with strong baseball knowledge to evaluate AI assistants during MLB postseason games. You will ask questions live, compare two AI products, capture conversations, and rate responses for usefulness and accuracy. Applicants...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Product management / roadmap / PRD Evaluator. Be the first to apply!
- emergency disaster management New York, NY
- agency management specialist New York, NY
- provider data management New York, NY
- pain management New York, NY
- director talent management New York, NY
- identity & access management New York, NY
- wealth management relationship manager New York, NY
- director global product management New York, NY
- creed management New York, NY
- learning management system specialist New York, NY



