Retail SME - AI Evaluation Expert
Obsidian
1. Overview Join a leading AI lab's cutting-edge GenAI team and help build foundational AI models from the ground up. We’re seeking talented Retail subject-matter experts (SMEs) with deep domain expertise and hands-on experience evaluating AI model outputs against rubrics to bring rigor and real-world retail judgment to our AI training data. This is a W-2 employment position with Cincinnatus LLC, with the opportunity to be placed at a leading AI Lab as part of their extended workforce. 2. Key Responsibilities Guide research and engineering teams to close knowledge gaps in retail merchandising, category management, and operations reasoning. Design challenging, domain-relevant retail tasks and write accurate, well-reasoned solutions grounded in real retail practice. Evaluate AI model outputs against structured rubrics and provide clear, written feedback on correctness, judgment, and reasoning quality. Develop and refine evaluation guidelines and scoring rubrics specific to retail tasks. Collaborate with other subject matter experts to ensure consistency and accuracy in training data. 3. Core Qualifications 8+ years of dedicated professional experience in retail (e.g., merchandising, category management, retail operations, buying/planning) at a recognized, top-tier organization (e.g., Amazon, Walmart, Target, Nike, Costco, Home Depot, or equivalent). Prior hands-on experience evaluating LLM/AI model outputs against rubrics or structured scoring criteria — mandatory; please describe this experience in your application. Demonstrable career progression (e.g., Category Manager - Senior Manager - Director of Merchandising). Ability to engage reliably for at least 35 hours/week during weekdays. Verbal and written communication skills, problem-solving skills, and interpersonal skills. About Cincinnatus LLC: Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives. Equal Employment Opportunity: Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic. #J-18808-Ljbffr Obsidian
- Volga Partners is seeking experienced finance, audit, insurance, compliance, and legal professionals to join our on-call AI Evaluation Specialists network. This remote, project-based role involves evaluating AI outputs, applying critical thinking, and providing high-quality...SuggestedRemote jobFlexible hours
- SME Careers is seeking a General Nurse Subject Matter Expert for a remote contractor role to review and create healthcare content... .... Your expertise will enhance AI models' accuracy in clinical... ...model integrity through rigorous evaluation of responses. #J-18808-Ljbffr SME...SuggestedRemote jobFor contractors
$80 - $90 per hour
**Pay: $80–$90 per hour** We are seeking experienced **Biology Experts with a PhD** to contribute their subject-matter expertise to projects focused on improving and evaluating advanced AI systems. In this remote contractor role, you will apply your knowledge of biology...SuggestedHourly payFor contractorsRemote work$8 - $65 per hour
Prolific is hiring Mental Health Professionals in New York to train and evaluate AI models. As a Domain Expert, you will be responsible for reviewing AI-generated responses, completing tasks related to psychology, and improving AI models based on your expertise. Pay rates...SuggestedRemote jobHourly payWork from homeFlexible hours- Mercor is partnering with a leading AI lab to design and evaluate disability adjudication processes for frontier AI models. You will craft realistic disability scenarios, generate determinations and evidence summaries, and assess AI-generated responses against established...Suggested
- A leading research services provider is seeking a Chemistry Expert (PhD) to evaluate complex chemistry problems and review AI-generated outputs for accuracy. This remote, hourly contract role requires deep subject-matter expertise and excellent communication skills. Candidates...Remote jobHourly payContract workFlexible hours
- Prolific in New York, NY, is seeking Chemistry Experts and Chemical Engineers to join its Expert Network. In this role, you will help train and evaluate AI models using your chemical expertise. Duties include evaluating AI-generated responses for accuracy and validating...Remote jobWork from homeFlexible hours
$30 per hour
Prolific is seeking fluent Hindi speakers to join their Expert Network, helping to train and evaluate AI models with real legal expertise. Responsibilities include analyzing and writing tasks in Hindi, judging AI’s performance, and aiding in the improvement of AI models...Remote jobHourly payFlexible hours- ...A leading AI research firm is seeking Expert Prompt Curators to design challenging prompts for evaluating advanced AI models. The role requires advanced knowledge in diverse fields and offers flexible hours, remote work, and a competitive hourly wage. Ideal candidates...Hourly payTemporary workRemote workFlexible hours
- Mercor is seeking a Music Audio Expert - German for a remote, project-based engagement. You will evaluate AI output lyrics and voice generation, score training data quality, and write music in the domain language listed in the title. Applicants should have 3+ years as...Remote job10 hours per week
- ...A tech company focusing on AI research is looking for experienced Krita users for a flexible, project-based contract opportunity. This role allows you to earn while evaluating AI-generated content related to digital painting and concept art. Candidates should have at...Contract workRemote workFlexible hours
- Mercor is seeking experienced musicians to evaluate generative music AI models, collaborating with a leading AI lab. You will assess AI-generated music and lyrics in Telugu and English, applying detailed quality standards across genres. Responsibilities include comparing...Part timeImmediate start
- Mercor is seeking experienced Medical and Health Services Managers to evaluate and improve AI-generated healthcare operations content and workflows. You will leverage your expertise in directing clinical services, personnel, budgets, and compliance across healthcare facilities...
- Rise Data Labs is seeking advanced Mathematics and Statistics experts to support the training and evaluation of state-of-the-art AI systems. We need subject-matter experts who can apply deep quantitative knowledge to AI evaluation problems, assess AI-generated reasoning...Remote jobContract workImmediate startFlexible hours
- ...Physics Specialist to contribute deep scientific expertise to AI model evaluation. You will craft and assess challenging physics problems,... ...electromagnetism, use adversarial prompting to surface errors, and provide expert critique of AI responses while working with project #J-1880...Remote job
$80 - $150 per hour
Prolific in New York, NY, is seeking Medical Doctors to join our Expert Network for training and evaluating AI models. As a Domain Expert participant, you will review and rate AI-generated clinical responses, providing crucial feedback for AI development. You must hold...Remote jobHourly payWork from homeFlexible hours- Mercor is seeking experienced musicians to evaluate generative music AI models in collaboration with a leading AI lab. You will assess AI-generated music across genres and rate it against detailed quality standards, working in Punjabi and English. Requirements include...
$30 per hour
About Prolific Prolific is not just another player in the AI space - we are building the biggest pool of quality human... ...Graphic and Visual Designers to act as Domain Experts for a high-level AI evaluation project. AI models are evolving beyond simple image generation...Remote jobWork from homeFlexible hours- A leading AI Data Services company is seeking a bilingual content evaluator to review AI-generated responses and create training content. The ideal candidate will have... ...offers flexible hours and requires at least 3 years of relevant experience. #J-18808-Ljbffr SME CareersRemote jobFlexible hours
$30 per hour
Prolific is looking for Fluent Urdu Speakers to join our Expert Network to help train AI models using real legal expertise. The role involves completing tasks that require one hour of uninterrupted work, with pay rates of up to $30/hr. Candidates must possess advanced Urdu...Remote jobWork from homeFlexible hours- Mercor, in partnership with Crossing Hurdles, is looking for a Multimedia Expert to work remotely on project-based tasks. The role involves reviewing and editing multimedia outputs, creating assets for AI model training, and providing feedback to ensure quality standards....Remote jobPart time10 hours per week
- ...Exists At Mercor, we believe the safest AI is the one that’s already been attacked —... ...a red team for this project - human data experts who probe AI models with adversarial inputs... ...that strengthen customer AI systems Evaluation coverage expands: more scenarios tested,...Remote work
$100 - $150 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... .... Position: Cybersecurity Labeling Expert Type: Contract Compensation... ...ransomware , worms , and exploits . Evaluate POC exploit development to determine...Hourly payWeekly payFull timeContract workFor contractorsSummer workRemote work- YO AI Labs is seeking a remote-based Business Intelligence & Analytics Expert (contractor) to evaluate and improve AI-powered analytics workflows. You will review dashboards, KPIs, reports, and data visualizations to ensure accuracy and actionable insights. You will build...Remote jobFor contractors
- Visa Hunt seeks a Chemical Safety & Toxicology Expert (contractor, remote) to contribute to a client project enhancing chemical safety evaluation frameworks for AI training. You will apply domain knowledge in toxicology and regulated materials to shape model learning,...Remote jobFor contractors
- Alignerr is seeking biology experts to design and evaluate AI training content. You will craft challenging biology questions and provide detailed, logically structured solutions to help AI systems reason through advanced biology topics. This remote, flexible hourly contract...Remote jobHourly payContract workFlexible hours
- Mercor seeks experts in atomic layer deposition and thin‑film processes to support AI research for semiconductors and physical sciences. You will generate, structure, and evaluate scientific data to guide how models reason about deposition and processing. This is hands...Remote job
$137.5k - $186k
Overview The Expert Network's ability to operate efficiently and stay ahead of what's possible... ...infrastructure in place. As our AI & Technology Enablement lead, you'll own... ...improve COE and expert operations. Test and evaluate new capabilities — particularly AI tools...Seasonal work- We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You... ...will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area...
- SME Careers is looking for a Contract Review Subject Matter Expert (SME) to review AI-generated contract analyses. You will create expert content and provide feedback on contract quality. This role requires 5+ years of professional experience in contract review across various...Remote jobContract work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Retail SME - AI Evaluation Expert. Be the first to apply!
- subject matter expert New York, NY
- fulfillment expert New York, NY
- guest service support expert New York, NY
- sql expert New York, NY
- technology expert New York, NY
- retail sales full time New York, NY
- vice president of retail New York, NY
- retail planner New York, NY
- full time retail sales merchandiser New York, NY
- retail sales advisor New York, NY



