Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Product launch / experiment readiness Evaluator

$80 - $120 per hour

AI Trainer Jobs

Product launch / experiment readiness Evaluator is a remote evaluation track for reviewing product launch / experiment readiness evaluation prompts and responses against AuraOne's quality rubric. Category: Frontier Model Evaluation · Pay: $80–$120 / hr · Location: Remote — US-eligible · Contractor Product launch / experiment readiness Evaluator is a remote evaluation track for reviewing product launch / experiment readiness evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain. About the role Product launch / experiment readiness Evaluator is a remote evaluation track for reviewing product launch / experiment readiness evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain. AI data reviewers help turn product launch / experiment readiness evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data. Review frontier model outputs. Judge benchmark failures and calibrate other evaluators. Responsibilities Evaluate product launch / experiment readiness evaluation model outputs against a versioned rubric and assign severity tags for Product launch / experiment readiness Evaluator assignments. Compare paired responses and pick the stronger answer with a written rationale. Label hallucinations, instruction-following failures, and unsafe content with structured tags. Capture ambiguous prompts and route them back to the program team for rubric updates. Role details Track Evaluation & annotation Work model Remote · Independent specialist contractor Compensation $80–$120 / hr Eligible from US What you should bring Prior evaluation, annotation, or human-rater experience on product launch / experiment readiness evaluation or adjacent content for Product launch / experiment readiness Evaluator work. Comfort applying multi-page rubrics consistently across long batches. Clear written reasoning that names the issue and the rubric clause being applied. Strong attention to detail and the ability to flag when a prompt itself is the problem. Reliable async availability for at least 10 hours per week. Example tasks Compare two product launch / experiment readiness evaluation model responses to the same prompt and pick the stronger one with rationale. Tag an unsafe response with the correct policy category and severity. Audit a 50-row batch for rubric consistency and report drift to the program lead. Propose a rubric clarification after spotting a recurring failure mode. Useful experience Background in linguistics, content moderation, or trust & safety review. Experience with inter-rater agreement metrics and calibration cycles. Domain expertise that lets you spot subject-matter errors automated checks miss. Compensation and schedule $80–$120 / hr Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation. Skills used in matching Model output evaluation Rubric-based annotation Severity tagging Inter-rater calibration Product launch / experiment readiness evaluation #J-18808-Ljbffr AI Trainer Jobs

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Product launch / experiment readiness Evaluator in New York, NY vacancy
  • Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates...  ...must have over 5 years of experience in product launch readiness and be fluent in English. Proficiency... 
    Suggested
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    2 days ago
  • Mercor is seeking expert Evaluators in product launch / experiment readiness to review AI-generated artifacts (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs.This is a remote... 
    Suggested
    Remote job
    Hourly pay
    Work at office

    Mercor

    New York, NY
    2 days ago
  • $80 - $120 per hour

     ..., General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: Product launch / experiment readiness Evaluator Type: Contract Compensation: $80–$120/hour Location: Remote Role Responsibilities Evaluate... 
    Suggested
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    New York, NY
    18 days ago
  • $160k - $210k

    BVAL (Bloomberg's Evaluated Pricing Service) Evaluator - US Agency Structured Products Location New York Business Area Product Ref # 10053894...  ...ideal candidate should have 5+ years of market experience in the sectors below, with a background in trading... 
    Suggested
    Price work
    Temporary work
    For contractors
    Work experience placement

    Bloomberg

    New York, NY
    3 days ago
  • About the roleWe are hiring expert Evaluators in Document/deck production QA to review and assess AI-generated work products (documents, spreadsheets...  ...Requirements (must have)5+ years of relevant professional experience in Document/deck production QA.Native or professional... 
    Suggested
    Remote job
    Hourly pay
    Work at office

    Mercor

    New York, NY
    8 hours ago
  •  ...individuals with strong baseball knowledge to evaluate AI assistants during MLB postseason...  ...will ask questions live, compare two AI products, capture conversations, and rate...  ...based with at least 1 year of professional experience, superb English, and attention to detail... 

    AI Trainer Jobs

    New York, NY
    8 hours ago
  • $80 - $120 per hour

    Product management / roadmap / PRD Evaluator is a remote evaluation track for reviewing product management / roadmap / prd evaluation prompts and responses...  ...bring Prior evaluation, annotation, or human-rater experience on product management / roadmap / prd evaluation or... 
    For contractors
    Remote work
    10 hours per week

    AI Trainer Jobs

    New York, NY
    4 days ago
  • A global leader in customer experience is seeking a luxury brand evaluator. This role involves assessing customer experiences at luxury retailers and providing feedback to help improve service quality. You'll have the flexibility to engage with brands such as Louis Vuitton... 
    Flexible hours

    CXG

    New York, NY
    2 days ago
  • AuraOne is seeking a Workflow Annotator—Product Management & Marketing to remotely review evaluation prompts and responses against the company's quality rubric. You will compare paired outputs, label edge cases, and provide structured feedback the modeling team can use... 
    Remote job
    For contractors

    AI Trainer Jobs

    New York, NY
    4 days ago
  •  ...training in journalism, editing, and video production. You will draft and source reference...  ...your rubric to ensure publication-ready quality. Preferred candidates bring...  ..., writing, or video production experience and comfort evaluating professional-grade work to benchmark... 

    BAM Ventures

    New York, NY
    2 days ago
  • $37.5 per hour

     ...In this role you will be responsible for evaluating advertising copy generated by a large...  ...relevance between the advertised brand/product and Prime Video contentCreativity and diversity...  ..., media, marketing, or copywriting-Experience evaluating advertising content and creative... 
    Contract work
    Temporary work
    Remote work

    TEKsystems

    New York, NY
    10 hours ago
  • $80 - $120 per hour

     ...0 - $120 per hour We are hiring expert Evaluators in Special education / IEP to review and assess AI-generated work products (documents, spreadsheets, and slide decks)...  ...have) #5+ years of relevant professional experience in Special education / IEP. # Native or... 
    Hourly pay
    Contract work
    For contractors
    Work at office
    Remote work

    Weekday 1

    New York, NY
    18 hours ago
  •  ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring...  ...should have over 5 years of professional experience, fluency in English, and proficiency in... 
    Work at office
    Remote work

    Obsidian

    New York, NY
    1 day ago
  • Mercor is partnering with a leading AI research organization to engage experienced UI/UX and product designers for a project focused on evaluating how well AI systems perform real-world digital product design work. You will define what excellent work looks like: designing... 

    Obsidian

    New York, NY
    1 day ago
  • Mercor is partnering with a leading AI research organization to engage experienced UI/UX and product designers for a project that evaluates how well AI systems perform real-world digital product design work. You will define what excellent work looks like by designing task... 

    Mercor

    New York, NY
    2 days ago
  • Obsidian is looking for expert Evaluators to review AI-generated work products in Public-sector procurement and RFI response. In this remote hourly role...  ...The ideal candidate will have 5+ years of relevant experience, native or professional fluency in English, and high... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    2 days ago
  • $80 - $120 per hour

    Incident management / reliability / SRE Evaluator is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs...  ...you should bring Strong day-job engineering experience — you can read, run, and debug unfamiliar code... 
    For contractors
    Remote work
    10 hours per week

    AI Trainer Jobs

    New York, NY
    4 days ago
  • Mercor is seeking expert Evaluators in Clinical / biomedical / pharma to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy,...  ...Ideal candidates have 5+ years of relevant experience, native or professional English fluency,... 
    Remote job
    Hourly pay

    Mercor

    New York, NY
    8 hours ago
  • $30 per hour

     ...10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous...  ...to support model refinement and product improvements. Participate in calibration...  ...relevant stakeholders. Requirements Strong experience in LLM evaluation, AI output analysis,... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    18 hours ago
  •  ...serve our clients globally by providing evaluated pricing and analytics on over 3 million...  ...models and processes. You will leverage your product and market knowledge to identify growth...  ...fixed income, and data sciencePrevious experience working in fixed income, primarily... 
    Worldwide

    JP Morgan Chase

    New York, NY
    1 day ago
  •  ...PositionTrue Footage is seeking an Appraiser in New York, NY who is ready to redefine the profession through cutting-edge technology and...  ...to detail in order to provide a higher-quality valuation product for the clientExercise the ability to effectively manage daily activities... 
    Full time
    Remote work
    Long distance

    True Footage

    New York, NY
    4 days ago
  •  ...our clients globally by providing them evaluated pricing on over two million fixed income...  ...client interaction Continuously improve product and service quality through daily market...  ...technology teams Market or quantitative experience in derivatives pricing with exposure to... 

    J.P. Morgan

    New York, NY
    8 hours ago
  •  ...DC What You’ll Do: Complete commercial valuation, alternative products and consulting assignments with precision and professionalism....  ...efficiently while adhering to deadlines. What You Bring: 5+ years of experience in commercial appraising, with expertise across multiple... 
    Full time
    Part time

    Opteon USA

    New York, NY
    1 day ago
  • $75k - $80k

    The Vocational Evaluation Supervisor manages and supervises the Vocational Evaluation team...  ...the work skills, abilities, work readiness, and vocational potential through administration...  ...controls to confirm that work and production are consistent with regular policies, procedures... 

    The Fedcap Group

    New York, NY
    3 days ago
  •  ...seeking advanced LLM power users with strong experience using MCP and (more importantly)...  ...personal life tasks. This project focuses on evaluating how well AI systems handle personalized...  ...a week Heavy personal usage of LLM products An active, rich LLM account with... 
    Trial period

    Obsidian

    New York, NY
    1 day ago
  •  ...sales engineers for a project focused on evaluating how well AI systems perform real-world...  ...deliverables (technical discovery plans, tailored product demonstrations, proof-of-concept and...  ...Qualifications 4+ years of relevant experience as a Sales Engineer, Solutions Engineer,... 

    Obsidian

    New York, NY
    3 days ago
  • $50 - $70 per hour

     ...Location: Remote Role Responsibilities Review and evaluate documents, slides, spreadsheets, and similar materials for quality...  ...or formatting. Work independently across common productivity tools with flexible hours . Qualifications Must-Have... 
    Contract work
    Summer work
    Remote work
    Flexible hours

    Mercor

    New York, NY
    23 days ago
  •  ...serve our clients globally by providing evaluated pricing and analytics on over 3 million...  ...models and processes. You will leverage your product and market knowledge to identify growth...  ...income, and data science Previous experience working in fixed income, primarily Municipal... 
    Worldwide

    JPMorgan Chase & Co.

    New York, NY
    16 days ago
  • $75 per hour

     ...through AI training initiatives for postsecondary education. Evaluate and review AI-generated coursework, assignments, and simulated...  ...degree in Sociology or a closely related field. Strong experience in postsecondary teaching with an academic or research background... 
    Hourly pay
    Contract work
    Remote work

    Crossing Hurdles

    New York, NY
    4 days ago
  • $30 per hour

     ...study participants with a wide variety of experiences, knowledge, and skills. The role We...  ...act as Domain Experts for a high-level AI evaluation project. AI models are evolving beyond...  ...between "consumer-grade" and "enterprise-ready" content, ensuring outputs are professional... 
    Remote job
    Work from home
    Flexible hours

    Prolific

    New York, NY
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Product launch / experiment readiness Evaluator. Be the first to apply!