Product launch / experiment readiness Evaluator
$80 - $120 per hourAI Trainer Jobs
Product launch / experiment readiness Evaluator is a remote evaluation track for reviewing product launch / experiment readiness evaluation prompts and responses against AuraOne's quality rubric. Category: Frontier Model Evaluation · Pay: $80–$120 / hr · Location: Remote — US-eligible · Contractor Product launch / experiment readiness Evaluator is a remote evaluation track for reviewing product launch / experiment readiness evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain. About the role Product launch / experiment readiness Evaluator is a remote evaluation track for reviewing product launch / experiment readiness evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain. AI data reviewers help turn product launch / experiment readiness evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data. Review frontier model outputs. Judge benchmark failures and calibrate other evaluators. Responsibilities Evaluate product launch / experiment readiness evaluation model outputs against a versioned rubric and assign severity tags for Product launch / experiment readiness Evaluator assignments. Compare paired responses and pick the stronger answer with a written rationale. Label hallucinations, instruction-following failures, and unsafe content with structured tags. Capture ambiguous prompts and route them back to the program team for rubric updates. Role details Track Evaluation & annotation Work model Remote · Independent specialist contractor Compensation $80–$120 / hr Eligible from US What you should bring Prior evaluation, annotation, or human-rater experience on product launch / experiment readiness evaluation or adjacent content for Product launch / experiment readiness Evaluator work. Comfort applying multi-page rubrics consistently across long batches. Clear written reasoning that names the issue and the rubric clause being applied. Strong attention to detail and the ability to flag when a prompt itself is the problem. Reliable async availability for at least 10 hours per week. Example tasks Compare two product launch / experiment readiness evaluation model responses to the same prompt and pick the stronger one with rationale. Tag an unsafe response with the correct policy category and severity. Audit a 50-row batch for rubric consistency and report drift to the program lead. Propose a rubric clarification after spotting a recurring failure mode. Useful experience Background in linguistics, content moderation, or trust & safety review. Experience with inter-rater agreement metrics and calibration cycles. Domain expertise that lets you spot subject-matter errors automated checks miss. Compensation and schedule $80–$120 / hr Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation. Skills used in matching Model output evaluation Rubric-based annotation Severity tagging Inter-rater calibration Product launch / experiment readiness evaluation #J-18808-Ljbffr AI Trainer Jobs
- Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates... ...must have over 5 years of experience in product launch readiness and be fluent in English. Proficiency...SuggestedRemote jobHourly payWork at office
- Mercor is seeking expert Evaluators in product launch / experiment readiness to review AI-generated artifacts (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs.This is a remote...SuggestedRemote jobHourly payWork at office
$80 - $120 per hour
..., General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: Product launch / experiment readiness Evaluator Type: Contract Compensation: $80–$120/hour Location: Remote Role Responsibilities Evaluate...SuggestedContract workSummer workWork at officeRemote work$160k - $210k
BVAL (Bloomberg's Evaluated Pricing Service) Evaluator - US Agency Structured Products Location New York Business Area Product Ref # 10053894... ...ideal candidate should have 5+ years of market experience in the sectors below, with a background in trading...SuggestedPrice workTemporary workFor contractorsWork experience placement- About the roleWe are hiring expert Evaluators in Document/deck production QA to review and assess AI-generated work products (documents, spreadsheets... ...Requirements (must have)5+ years of relevant professional experience in Document/deck production QA.Native or professional...SuggestedRemote jobHourly payWork at office
- ...individuals with strong baseball knowledge to evaluate AI assistants during MLB postseason... ...will ask questions live, compare two AI products, capture conversations, and rate... ...based with at least 1 year of professional experience, superb English, and attention to detail...
$80 - $120 per hour
Product management / roadmap / PRD Evaluator is a remote evaluation track for reviewing product management / roadmap / prd evaluation prompts and responses... ...bring Prior evaluation, annotation, or human-rater experience on product management / roadmap / prd evaluation or...For contractorsRemote work10 hours per week- A global leader in customer experience is seeking a luxury brand evaluator. This role involves assessing customer experiences at luxury retailers and providing feedback to help improve service quality. You'll have the flexibility to engage with brands such as Louis Vuitton...Flexible hours
- AuraOne is seeking a Workflow Annotator—Product Management & Marketing to remotely review evaluation prompts and responses against the company's quality rubric. You will compare paired outputs, label edge cases, and provide structured feedback the modeling team can use...Remote jobFor contractors
- ...training in journalism, editing, and video production. You will draft and source reference... ...your rubric to ensure publication-ready quality. Preferred candidates bring... ..., writing, or video production experience and comfort evaluating professional-grade work to benchmark...
$37.5 per hour
...In this role you will be responsible for evaluating advertising copy generated by a large... ...relevance between the advertised brand/product and Prime Video contentCreativity and diversity... ..., media, marketing, or copywriting-Experience evaluating advertising content and creative...Contract workTemporary workRemote work$80 - $120 per hour
...0 - $120 per hour We are hiring expert Evaluators in Special education / IEP to review and assess AI-generated work products (documents, spreadsheets, and slide decks)... ...have) #5+ years of relevant professional experience in Special education / IEP. # Native or...Hourly payContract workFor contractorsWork at officeRemote work- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring... ...should have over 5 years of professional experience, fluency in English, and proficiency in...Work at officeRemote work
- Mercor is partnering with a leading AI research organization to engage experienced UI/UX and product designers for a project focused on evaluating how well AI systems perform real-world digital product design work. You will define what excellent work looks like: designing...
- Mercor is partnering with a leading AI research organization to engage experienced UI/UX and product designers for a project that evaluates how well AI systems perform real-world digital product design work. You will define what excellent work looks like by designing task...
- Obsidian is looking for expert Evaluators to review AI-generated work products in Public-sector procurement and RFI response. In this remote hourly role... ...The ideal candidate will have 5+ years of relevant experience, native or professional fluency in English, and high...Remote jobHourly payWork at office
$80 - $120 per hour
Incident management / reliability / SRE Evaluator is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs... ...you should bring Strong day-job engineering experience — you can read, run, and debug unfamiliar code...For contractorsRemote work10 hours per week- Mercor is seeking expert Evaluators in Clinical / biomedical / pharma to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy,... ...Ideal candidates have 5+ years of relevant experience, native or professional English fluency,...Remote jobHourly pay
$30 per hour
...10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous... ...to support model refinement and product improvements. Participate in calibration... ...relevant stakeholders. Requirements Strong experience in LLM evaluation, AI output analysis,...Remote jobHourly payContract work- ...serve our clients globally by providing evaluated pricing and analytics on over 3 million... ...models and processes. You will leverage your product and market knowledge to identify growth... ...fixed income, and data sciencePrevious experience working in fixed income, primarily...Worldwide
- ...PositionTrue Footage is seeking an Appraiser in New York, NY who is ready to redefine the profession through cutting-edge technology and... ...to detail in order to provide a higher-quality valuation product for the clientExercise the ability to effectively manage daily activities...Full timeRemote workLong distance
- ...our clients globally by providing them evaluated pricing on over two million fixed income... ...client interaction Continuously improve product and service quality through daily market... ...technology teams Market or quantitative experience in derivatives pricing with exposure to...
- ...DC What You’ll Do: Complete commercial valuation, alternative products and consulting assignments with precision and professionalism.... ...efficiently while adhering to deadlines. What You Bring: 5+ years of experience in commercial appraising, with expertise across multiple...Full timePart time
$75k - $80k
The Vocational Evaluation Supervisor manages and supervises the Vocational Evaluation team... ...the work skills, abilities, work readiness, and vocational potential through administration... ...controls to confirm that work and production are consistent with regular policies, procedures...- ...seeking advanced LLM power users with strong experience using MCP and (more importantly)... ...personal life tasks. This project focuses on evaluating how well AI systems handle personalized... ...a week Heavy personal usage of LLM products An active, rich LLM account with...Trial period
- ...sales engineers for a project focused on evaluating how well AI systems perform real-world... ...deliverables (technical discovery plans, tailored product demonstrations, proof-of-concept and... ...Qualifications 4+ years of relevant experience as a Sales Engineer, Solutions Engineer,...
$50 - $70 per hour
...Location: Remote Role Responsibilities Review and evaluate documents, slides, spreadsheets, and similar materials for quality... ...or formatting. Work independently across common productivity tools with flexible hours . Qualifications Must-Have...Contract workSummer workRemote workFlexible hours- ...serve our clients globally by providing evaluated pricing and analytics on over 3 million... ...models and processes. You will leverage your product and market knowledge to identify growth... ...income, and data science Previous experience working in fixed income, primarily Municipal...Worldwide
$75 per hour
...through AI training initiatives for postsecondary education. Evaluate and review AI-generated coursework, assignments, and simulated... ...degree in Sociology or a closely related field. Strong experience in postsecondary teaching with an academic or research background...Hourly payContract workRemote work$30 per hour
...study participants with a wide variety of experiences, knowledge, and skills. The role We... ...act as Domain Experts for a high-level AI evaluation project. AI models are evolving beyond... ...between "consumer-grade" and "enterprise-ready" content, ensuring outputs are professional...Remote jobWork from homeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Product launch / experiment readiness Evaluator. Be the first to apply!




