Retail SME - AI Evaluation Expert
Obsidian
1. Overview Join a leading AI lab's cutting-edge GenAI team and help build foundational AI models from the ground up. We’re seeking talented Retail subject-matter experts (SMEs) with deep domain expertise and hands-on experience evaluating AI model outputs against rubrics to bring rigor and real-world retail judgment to our AI training data. This is a W-2 employment position with Cincinnatus LLC, with the opportunity to be placed at a leading AI Lab as part of their extended workforce. 2. Key Responsibilities Guide research and engineering teams to close knowledge gaps in retail merchandising, category management, and operations reasoning. Design challenging, domain-relevant retail tasks and write accurate, well-reasoned solutions grounded in real retail practice. Evaluate AI model outputs against structured rubrics and provide clear, written feedback on correctness, judgment, and reasoning quality. Develop and refine evaluation guidelines and scoring rubrics specific to retail tasks. Collaborate with other subject matter experts to ensure consistency and accuracy in training data. 3. Core Qualifications 8+ years of dedicated professional experience in retail (e.g., merchandising, category management, retail operations, buying/planning) at a recognized, top-tier organization (e.g., Amazon, Walmart, Target, Nike, Costco, Home Depot, or equivalent). Prior hands-on experience evaluating LLM/AI model outputs against rubrics or structured scoring criteria — mandatory; please describe this experience in your application. Demonstrable career progression (e.g., Category Manager - Senior Manager - Director of Merchandising). Ability to engage reliably for at least 35 hours/week during weekdays. Verbal and written communication skills, problem-solving skills, and interpersonal skills. About Cincinnatus LLC: Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives. Equal Employment Opportunity: Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic. #J-18808-Ljbffr Obsidian
$60 - $80 per hour
...technical talent with leading AI research labs. Headquartered in... ...Jack Dorsey . Position: Retail Specialist Type: Contract... ...in real retail practice. Evaluate AI model outputs against structured... ...with other subject matter experts to ensure consistency and...SuggestedContract workSummer workRemote workWeekday work- SME Careers is seeking a Tax & Audit Subject Matter Expert to review AI-generated tax and audit responses. You will create expert accounting content, ensuring accuracy... .... This remote position offers flexibility to evaluate AI outputs across time zones. #J-18808-Ljbffr...SuggestedRemote work
- ...A leading AI research firm is seeking Expert Prompt Curators to design challenging prompts for evaluating advanced AI models. The role requires advanced knowledge in diverse fields and offers flexible hours, remote work, and a competitive hourly wage. Ideal candidates...SuggestedHourly payTemporary workRemote workFlexible hours
- SME Careers is seeking a General Nurse Subject Matter Expert for a remote contractor role to review and create healthcare content... .... Your expertise will enhance AI models' accuracy in clinical... ...model integrity through rigorous evaluation of responses. #J-18808-Ljbffr SME...SuggestedRemote jobFor contractors
- A leading research services provider is seeking a Chemistry Expert (PhD) to evaluate complex chemistry problems and review AI-generated outputs for accuracy. This remote, hourly contract role requires deep subject-matter expertise and excellent communication skills. Candidates...SuggestedRemote jobHourly payContract workFlexible hours
- Prolific in New York, NY, is seeking Chemistry Experts and Chemical Engineers to join its Expert Network. In this role, you will help train and evaluate AI models using your chemical expertise. Duties include evaluating AI-generated responses for accuracy and validating...Remote jobWork from homeFlexible hours
$8 - $65 per hour
Prolific is hiring Mental Health Professionals in New York to train and evaluate AI models. As a Domain Expert, you will be responsible for reviewing AI-generated responses, completing tasks related to psychology, and improving AI models based on your expertise. Pay rates...Remote jobHourly payWork from homeFlexible hours$30 per hour
Prolific is seeking fluent Hindi speakers to join their Expert Network, helping to train and evaluate AI models with real legal expertise. Responsibilities include analyzing and writing tasks in Hindi, judging AI’s performance, and aiding in the improvement of AI models...Remote jobHourly payFlexible hours- A tech company focusing on AI research is looking for experienced Krita users for a flexible, project-based contract opportunity. This role allows you to earn while evaluating AI-generated content related to digital painting and concept art. Candidates should have at least...Remote jobContract workFlexible hours
$80 - $150 per hour
Prolific in New York, NY, is seeking Medical Doctors to join our Expert Network for training and evaluating AI models. As a Domain Expert participant, you will review and rate AI-generated clinical responses, providing crucial feedback for AI development. You must hold...Remote jobHourly payWork from homeFlexible hours$30 per hour
About Prolific Prolific is not just another player in the AI space - we are building the biggest pool of quality human... ...Graphic and Visual Designers to act as Domain Experts for a high-level AI evaluation project. AI models are evolving beyond simple image generation...Remote jobWork from homeFlexible hours- A leading AI Data Services company is seeking a bilingual content evaluator to review AI-generated responses and create training content. The ideal candidate will have... ...offers flexible hours and requires at least 3 years of relevant experience. #J-18808-Ljbffr SME CareersRemote jobFlexible hours
$30 per hour
Prolific is looking for Fluent Urdu Speakers to join our Expert Network to help train AI models using real legal expertise. The role involves completing tasks that require one hour of uninterrupted work, with pay rates of up to $30/hr. Candidates must possess advanced Urdu...Remote jobWork from homeFlexible hours- LILT (Production) is seeking a Subject Matter Expert to contribute to an AI benchmarking project in the healthcare sector. You will design realistic scenarios reflecting patient care and hospital operations, ensuring context and cultural appropriateness. Ideal candidates...For contractorsFlexible hours
- Mercor, in partnership with Crossing Hurdles, is looking for a Multimedia Expert to work remotely on project-based tasks. The role involves reviewing and editing multimedia outputs, creating assets for AI model training, and providing feedback to ensure quality standards....Remote jobPart time10 hours per week
- Handshake AI is seeking experienced CAD professionals with 2+ years of hands-on experience using SolidWorks and related CAD software, and professional proficiency in Mandarin Chinese, to support AI research through flexible, hourly contract work. This ongoing, project-...Remote jobHourly payContract workFlexible hours
$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ..., and Jack Dorsey . Position: Software / AI / IT / data Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$50 - $100 per hour
DataAnnotation is seeking an experienced Legal Expert to help train AI models. You will tackle diverse legal problems, measure chatbot reasoning, and improve model quality from home on a flexible schedule. This independent contractor role pays hourly from $50 to $100+ per...Remote jobHourly payFor contractorsFlexible hours$60 - $70 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...$60–$70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance...Contract workSummer workRemote work- ...Remote Mathematics Expert (AI/LLM) - 34877 Remote Mathematics Expert (AI/LLM) - 34877 1 week ago Be among the first 25 applicants Turing... ...Collaborate with LLM researchers to align problem types with evaluation goals, particularly in areas where models commonly struggle (e...Hourly payContract workPart timeFor contractorsFreelanceInternshipRemote workWork from homeWorldwideAfternoon shift
- ...Korean Language Expert (AI Training) About the Role We're looking for Korean language experts to help evaluate and improve AI systems built for Korean language learning and instruction. Your linguistic insight and teaching background will directly shape how AI...Hourly payOngoing contractContract workFreelanceRemote workWorldwideFlexible hours
$100 - $150 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... .... Position: Cybersecurity Labeling Expert Type: Contract Compensation... ...ransomware , worms , and exploits . Evaluate POC exploit development to determine...Hourly payWeekly payFull timeContract workFor contractorsSummer workRemote work$130 - $160 per hour
Mercor connects elite accounting talent with AI research labs. The Accounting Expert contract is remote, offering $130-$160/hour for candidates with deep US GAAP expertise and strong close-process skills. Responsibilities include designing grading criteria for accounting...Remote jobContract work- ...education and research sector is seeking a Social Sciences PhD Expert for a remote part-time role, requiring strong analytical... ...The position involves creating historically relevant prompts, evaluating AI outputs, and contributing to research initiatives. Ideal candidates...Part timeRemote workFlexible hours
- SME Careers is looking for a Contract Review Subject Matter Expert (SME) to review AI-generated contract analyses. You will create expert content and provide feedback on contract quality. This role requires 5+ years of professional experience in contract review across various...Remote jobContract work
- AuraOne is seeking a Criminology Expert to remotely review AI outputs in criminology research, reproducing key steps and documenting methods for training data. You will evaluate derivations, ensure consistency with literature, and flag errors with structured severity tags...Remote job
- Mindrift connects specialists with project-based AI opportunities for leading tech companies, focusing on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. Contributors may design original computational mathematics...Permanent employmentTemporary workPart time
- Investment Banking AI Subject-Matter Expert is a remote review track for evaluating AI outputs across banking workflows. Reviewers grade calculations, narrative reasoning, and policy adherence; flag compliance and reconciliation issues; and document the correct treatment...Hourly payFor contractorsWork experience placementRemote work10 hours per week
- ...hiring PhD‑level biologists to help make advanced AI models safer. You'll apply your scientific expertise to evaluate and strengthen how these models handle... ...train you on the workflow. Responsibilities Write expert‑level prompts across specialized life‑science topics...Part timeImmediate start
- Alignerr is seeking a Physics Subject Matter Expert (AI Training) to design and evaluate physics problems used to train advanced AI systems. This fully remote hourly contract offers flexible hours and the chance to influence how AI reasoning develops across mechanics, electromagnetism...Remote jobHourly payContract workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Retail SME - AI Evaluation Expert. Be the first to apply!
- subject matter expert New York, NY
- guest service support expert New York, NY
- technology expert New York, NY
- fulfillment expert New York, NY
- sql expert New York, NY
- cvs health retail New York, NY
- full time retail sales merchandiser New York, NY
- store manager retail New York, NY
- cookies retail New York, NY
- retail jobs New York, NY


