Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Retail Professional for AI Model Evaluation

$60 - $80 per hour

SaidGig

Apply deep retail expertise to help develop and evaluate advanced large language models. In this role, you will bring practical judgment in merchandising, category management, buying, planning, and retail operations to create rigorous training data and improve AI reasoning. Key Responsibilities

  • Advise research and engineering teams on retail merchandising, category management, and operations knowledge gaps.
  • Create challenging retail-focused tasks and accurate, well-reasoned solutions based on real-world retail practice.
  • Assess AI model responses using structured rubrics and provide written feedback on correctness, judgment, and reasoning quality.
  • Create and refine retail-specific evaluation guidelines and scoring rubrics.
  • Partner with other subject-matter experts to maintain consistency and accuracy across training data.
Qualifications
  • At least 8 years of dedicated professional retail experience in areas such as merchandising, category management, retail operations, buying, or planning.
  • Experience at a recognized top-tier organization, such as Amazon, Walmart, Target, Nike, Costco, Home Depot, or an equivalent company.
  • Hands-on experience evaluating LLM or AI model outputs against rubrics or structured scoring criteria is required. Describe this experience in your application.
  • Demonstrated career progression, such as advancing from Category Manager to Senior Manager to Director of Merchandising.
  • Strong verbal and written communication, problem-solving, and interpersonal skills.
Work Terms
  • United States-based, hourly W-2 employment.
  • Reliable weekday availability of at least 35 hours per week is required.
Compensation

$60 to $80 per hour.

Equal Employment Opportunity

Employment decisions are made without discrimination based on race, religion, color, national origin, sex, including pregnancy, childbirth, reproductive health decisions, related medical conditions, sexual orientation, gender identity, gender expression, age, protected veteran status, disability, genetic information, political views or activity, or any other legally protected characteristic.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Retail Professional for AI Model Evaluation in United States vacancy
  • $100 per hour

     ...finance expertise to improve AI-driven financial applications...  ...rigorous, real-world analysis, evaluation, and feedback. This remote,...  ...produce accurate, clear, and professionally relevant financial content. Prior...  ...financial data, reports, and model outputs using detailed... 
    Suggested
    Hourly pay
    Contract work
    Part time
    For contractors
    Remote work

    SaidGig

    United States
    more than 2 months ago
  • $60 - $80 per hour

     ...help develop advanced large language models by bringing real-world underwriting,...  ...claims, and risk-assessment judgment to AI training and evaluation work. Key Responsibilities...  ...Qualifications At least 8 years of dedicated professional insurance experience in underwriting,... 
    Suggested
    Hourly pay
    Weekday work

    SaidGig

    United States
    more than 2 months ago
  • $100 - $150 per hour

     ...Apply senior finance expertise to improve how frontier AI systems reason through real-world financial work. You...  ...research team to define high-quality finance tasks, evaluate model performance, and translate professional judgment into rigorous standards. Role Overview... 
    Suggested
    Hourly pay
    Full time
    Live in
    Relocation
    Relocation package

    SaidGig

    California
    a month ago
  • $45 - $70 per hour

     ...Role Overview Help shape how AI models respond to sensitive everyday conversations involving...  ...and difficult life decisions. You will evaluate model interactions and define standards...  ...field, or equivalent mental-health professional experience. At least 3 years of... 
    Suggested
    Hourly pay
    Part time
    Weekday work

    SaidGig

    United States
    4 days ago
  • $100 per hour

     ...domain expertise to improve AI-driven financial applications...  ...reviewing, annotating, and refining model outputs and prompts. You will...  .... Develop, refine, and evaluate prompts related to financial...  ...robustness, and adherence to professional financial standards.... 
    Suggested
    Hourly pay
    Part time
    For contractors
    Remote work

    SaidGig

    United States
    9 days ago
  • $50 per hour

     ...the performance of large language models on real-world finance tasks by evaluating model outputs, designing assessment...  ..., and working directly with AI researchers to shape training and...  ...Qualifications Minimum 2 years of professional experience in one or more of the following... 
    Hourly pay
    Contract work
    Remote work
    10 hours per week
    Flexible hours

    SaidGig

    United States
    4 days ago
  • $60 - $80 per hour

     ...underwriting and claims judgment, and evaluate large language model outputs against structured rubrics to...  ...and claims practice. Evaluate AI model outputs against structured rubrics...  ...Qualifications Minimum 8 years of dedicated professional experience in insurance, such as... 
    Hourly pay
    Weekday work

    SaidGig

    United States
    3 days ago
  • $100 - $150 per hour

     ...Overview Help shape how next-generation AI models perform real financial work by...  ...appear plausible but would not survive professional scrutiny. Instruction specs and golden...  ...: Design challenging finance tasks and evaluation sets, and collaborate with researchers... 
    Hourly pay
    Full time
    Live in
    Relocation
    Relocation package

    SaidGig

    Bay County, FL
    3 days ago
  •  ...We are seeking an experienced Engineer, AI – AI Evaluation & Model Risk Lead to lead how AI models are evaluated, cleared, monitored, and...  ...technical guidance. Required Qualifications ~2+ years of professional experience in relevant AI, machine learning, software... 
    Full time
    Work experience placement

    BrickRed Systems

    Washington DC
    4 days ago
  • $110 per hour

     ...physician talent network supporting AI labs and companies with medical expertise...  ...real-world clinical expertise to model development and evaluation. Key Responsibilities Train...  ...frontier AI research. Qualifications Professional experience in clinical diagnosis and... 
    Hourly pay
    Contract work
    Remote work

    Confidential

    Spring Valley, TX
    2 days ago
  • $110 - $150 per hour

     ...Role Overview Help advance frontier AI models by bringing professional finance judgment to the evaluation, design, and improvement of financial knowledge-work tasks. You will work closely with AI research and program teams to define what high-quality financial reasoning... 
    Hourly pay
    Full time
    Live in
    Relocation
    Relocation package

    SaidGig

    Bay County, FL
    a month ago
  • $60 - $80 per hour

     ...develop advanced large language models. In this role, you will bring...  ..., and campaign judgment to AI training data, partnering...  ...quality. Develop and improve evaluation guidelines and scoring rubrics...  ...least 8 years of dedicated professional marketing experience, such as... 
    Hourly pay
    Weekday work

    SaidGig

    United States
    more than 2 months ago
  • $45 - $70 per hour

     ...Recruiting is hiring a remote Behavioral Health Expert, AI Safety and Model Evaluation contractor (pay $45-$70/hr). Contribute to frontier AI research...  ...evaluation work. Ideal candidates: Behavioral health professionals (therapists, counselors, social workers, psychologists,... 
    Temporary work
    Part time
    For contractors
    Remote work

    Gridnaut Recruiting

    Remote
    5 days ago
  • Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant... 
    Remote job
    Flexible hours

    Prolific

    Charlotte, NC
    2 days ago
  •  ...Role Overview Use your investment and finance expertise to evaluate and improve AI model performance on financial reasoning, valuation, markets, and real-world investment scenarios. Key Responsibilities Assess AI model outputs on valuation, financial modeling, markets... 
    For contractors
    Remote work

    SaidGig

    United States
    3 days ago
  • Prolific is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities... 
    Remote job
    Hourly pay
    Flexible hours

    Prolific

    Jacksonville, FL
    2 days ago
  • $60 per hour

    Prolific, located in Arizona, is seeking Biology Experts and Life Science Professionals to join their Expert Network. This role involves evaluating and training AI models with real scientific expertise. Successful candidates will review AI-generated content for accuracy... 
    Hourly pay

    Prolific

    Arizona City, AZ
    1 day ago
  • $60 per hour

    Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    San Jose, CA
    3 days ago
  • $60 - $100 per hour

     ...Help advance frontier AI systems by bringing practicing insurance and actuarial judgment into the evaluation, training, and improvement of models used for real insurance work. You will work...  ...and responses that would not meet professional standards. Write detailed instruction... 
    Hourly pay
    Full time
    Live in
    Relocation
    Relocation package

    SaidGig

    California
    a month ago
  • $60 - $90 per hour

     ...engineering judgment to improve how advanced AI models reason through real-world engineering...  ...correct solutions, and create rigorous evaluations grounded in industry practice. Key...  ...or lead engineering responsibility. A Professional Engineer license is preferred but not... 
    Hourly pay
    Full time
    Remote work

    SaidGig

    United States
    21 days ago
  • A leading AI company is seeking a legal professional for a contractor role focused on evaluating AI model outputs in legal contexts. Candidates must hold a Juris Doctor (J.D.) and have more than 3 years of experience in law. The role involves reviewing complex legal hypotheticals... 
    For contractors
    10 hours per week

    Turing

    Los Angeles, CA
    5 days ago
  • $80 per hour

     ...and technical talent with leading AI research labs. Headquartered in...  ...coding agents to complete and evaluate complex data engineering tasks. Review model-generated implementations involving...  ...and weaknesses. Apply professional engineering judgment to realistic... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    10 days ago
  • $65 - $105 per hour

     ...Help advance frontier AI models by bringing rigorous life sciences research judgment into task design, evaluation, and model improvement. You will work closely with an AI lab...  ...Hands-on use of large language models in professional work and sound judgment in distinguishing... 
    Hourly pay
    Full time
    Freelance
    Live in
    Relocation
    Relocation package

    SaidGig

    California
    a month ago
  • $70 - $110 per hour

     ...Overview Shape how advanced AI systems reason about real clinical...  ...with an AI research team to evaluate medical knowledge tasks,...  ...benchmarks that measure meaningful model improvement. Key...  ...or Chief Medical Officer. Professional, hands-on experience using large... 
    Hourly pay
    Full time
    Live in
    Relocation
    Relocation package

    SaidGig

    California
    a month ago
  • $400 per month

     ...Mercor is partnering with a leading AI research lab to support a Frontier...  ...Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical...  ...their strengths and weaknesses. Apply professional engineering judgment to realistic... 

    Mercor Inc

    Doral, FL
    4 days ago
  • $85 per hour

     ...technical talent with leading AI research labs. Headquartered in...  ...agents to complete and evaluate complex infrastructure engineering tasks. Review model-generated implementations involving...  ...strengths and weaknesses. Apply professional engineering judgment to... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    10 days ago
  • Cincinnatus LLC is seeking behavioral health experts to help evaluate how AI models respond in everyday conversations about relationships, wellbeing and beliefs. This part-time role focuses on neutrality, safety and balanced guidance, with 20-40 hours per week and a W-... 
    Part time

    Annotation Academy

    New York, NY
    5 days ago
  • $60 - $100 per hour

     ...accounting and audit judgment to improve how advanced AI models handle real-world accounting and audit work. You will...  ...defining high-quality tasks, correct solutions, and evaluation standards grounded in professional practice. Key Responsibilities Review accounting... 
    Hourly pay
    Full time
    Live in
    Relocation
    Relocation package

    SaidGig

    California
    3 days ago
  • $60 - $80 per hour

     ...administration subject matter expertise to AI research and product work through short...  .... Key Responsibilities Train and evaluate AI models on HR and administration topics....  ...each engagement. Qualifications Professional experience in one or more of the following... 
    Hourly pay
    Contract work
    Temporary work
    For contractors
    Remote work

    SaidGig

    United States
    4 days ago
  • $70 - $100 per hour

     ...expertise to help improve how next-generation AI models complete spreadsheet-based work. This...  ...and is designed for versatile Excel professionals with experience across multiple...  ...training, human data annotation, or model evaluation is a strong plus. Demonstrable career... 
    Hourly pay
    Weekday work

    SaidGig

    United States
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Retail Professional for AI Model Evaluation. Be the first to apply!