Behavioral Health Expert, AI Safety and Model Evaluation
$45 - $70 per hourGridnaut Recruiting
Gridnaut Recruiting is hiring a remote Behavioral Health Expert, AI Safety and Model Evaluation contractor (pay $45–$70/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Behavioral health professionals (therapists, counselors, social workers, psychologists, case managers, peer support specialists) helping evaluate how AI models handle sensitive everyday conversations.; Review conversations on relationships, family dynamics, emotional wellbeing, and personal beliefs; assess whether responses are balanced, neutral, appropriately bounded, and safe.; Write clear rationales, build rubrics and reference responses, and design test scenarios grounded in counseling and behavioral health practice.; Degree in psychology, counseling, social work, or related field (or equivalent experience) plus 3+ years supporting people in a mental health setting; licensure valued but not required.; Part-time W-2 role, 20+ hours/week (up to 40); remote, US-based..
- Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant...SuggestedRemote jobFlexible hours
- Prolific is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities...SuggestedRemote jobHourly payFlexible hours
$60 per hour
Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing...SuggestedRemote jobHourly payWork from homeFlexible hours$70 - $110 per hour
...Overview Shape how advanced AI systems reason about real clinical... ...with an AI research team to evaluate medical knowledge tasks,... ...benchmarks that measure meaningful model improvement. Key Responsibilities... ...and model outputs for missing behaviors, weak reasoning, unsafe...SuggestedHourly payFull timeLive inRelocationRelocation package$65 per hour
Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience...SuggestedHourly paySelf employmentWork from homeFlexible hours$110 per hour
...Apply to join a physician talent network supporting AI labs and companies with medical expertise. This is an open application... ...projects by bringing real-world clinical expertise to model development and evaluation. Key Responsibilities Train and evaluate AI models...Hourly payContract workRemote work- ...Overview Help assess frontier AI systems at the boundary between legitimate radiological safety work and potentially dangerous... ..., dual-use, or adversarial. Evaluate AI responses against a defined... ...writing, published research, or expert-witness experience is also valued...For contractorsRemote work
$60 - $70 per hour
...Role Overview Assess the safety, alignment, and overall quality of frontier AI model outputs on complex, policy sensitive, and ambiguous "grey area" topics. Work through structured evaluations to identify unsafe behavior, reasoning failures, hallucinations, and policy...Hourly payRemote work$80 - $150 per hour
Prolific Academic Ltd is recruiting Medical Doctors to help train and evaluate AI models. You’ll complete a quick test to assess suitability and, if successful, join as a Domain Expert, paid to work on AI tasks. Researchers pay $80-$150 per hour per completed task, with...Hourly paySelf employmentWork from home$18 - $50 per hour
...AI Safety & Policy Expert (AI Community) About the Role: At TELUS Digital... ...complex scenarios to ensure the model learns what is right, what... ...prevents harmful AI behaviors before they ever reach a user... ...Responsibilities): Safety Annotation: Evaluate AI-generated responses...Hourly payRemote workFlexible hours- ...clinical expertise to the development and evaluation of next-generation AI systems. This remote, contract... ...realistic clinical scenarios, develop expert reference responses, and evaluate AI... ...clinical accuracy, ethical practice, safety, and risk management. #J-18808-Ljbffr...Remote jobContract work
$80 - $150 per hour
Prolific in New York, NY, is seeking Medical Doctors to join our Expert Network for training and evaluating AI models. As a Domain Expert participant, you will review and rate AI-generated clinical responses, providing crucial feedback for AI development. You must hold...Remote jobHourly payWork from homeFlexible hours$75 - $110 per hour
AuraOne seeks Clinical Review contractors to evaluate AI outputs in Travel, Hospitality & Culinary Expert track. Work remote with US eligibility;... ...document corrected reasoning to support model improvement, with emphasis on safety and guideline adherence. Active clinical...Remote jobHourly payFor contractors10 hours per week$70 - $90 per hour
...Role Overview Help evaluate Neuron Kernel Interface development tasks that support the training and evaluation of advanced AI models. You will assess kernel quality, numerical correctness... ...accumulation order, rounding behavior, and mixed-precision semantics across...Hourly payRemote work$60 - $90 per hour
...technical talent with leading AI research labs. Headquartered... ...Machine Learning Engineer — Model Evaluation & Experimentation Type:... ...reward functions and training behavior. Evaluate models to identify... ...Collaborate with researchers and experts to maintain task consistency...Contract workSummer workRemote work$80 - $100 per hour
...expertise to improve next-generation AI systems through practical... ...documentation, and rigorous evaluation of AI-generated solutions. This... ...Responsibilities Provide expert analysis, feedback, and practical... ...data challenges for AI model development. Document advanced...Hourly payContract workRemote work$20 - $70 per hour
...expertise to help improve next-generation AI systems through accurate, real-world... ...role focuses on reviewing, evaluating, and refining CAD work, with no prior... ...varied tasks and workflows. Provide expert feedback on 2D and 3D models against project requirements and...Hourly payContract workFor contractorsRemote work$50 per hour
...Help evaluate and improve advanced AI language systems by applying expert English-language judgment to model outputs and training data. This fully remote role focuses on assessing how well large language models understand, generate, and use English across a range of real...Hourly payRemote workFlexible hours$40 - $90 per hour
...Apply clinical expertise to help train next-generation AI systems through rigorous, real-world medical evaluation work. You will create and assess challenging... ...Familiarity with AI or machine-learning systems and model evaluation is helpful but not required. Work Terms...Hourly payContract workRemote work$110 per hour
...Overview Provide clinical expertise to help train, evaluate, and shape medical AI systems by joining a Physician Expert Network. This is an open application to be... ...Key Responsibilities Train and evaluate AI models in medical and clinical contexts. Create tasks...Hourly payContract workRemote work- ...a Software Engineer L5/L6 — Model Evaluations & Data Curation (MEDC) based... ...infrastructure that accelerates how AI teams create, evaluate, and... ...decisions influence model behavior and performance. Your work... ...each year. ~ Comprehensive health insurance plans and mental health...Full timeRemote workFlexible hours
$17 - $25 per hour
...strengthen conversational AI by testing it from an... ...role, you will probe models and agents for weaknesses, create actionable safety data, and help identify... ...misinformation, and harmful behavior. Participation in... ...AI systems. Expand evaluation coverage by identifying...Hourly payRemote work- AI Trainer Jobs seeks a Research Physics Expert for a remote review track evaluating AI outputs across physics reasoning, calculations, and research workflows. Reviewers grade... ..., and document the correct method so the modeling team can train on it. Graduate-level physics...Remote jobFor contractors10 hours per week
$48 - $62 per hour
...Probe conversational AI systems for weaknesses before... ..., create high-quality safety data, and produce... ...misinformation, and harmful behaviors. Participation in... ...team conversational AI models and agents through jailbreaks... ...risks. Help expand evaluation coverage by testing...Hourly payRemote work$60 per hour
Prolific is seeking Chemistry Experts and Chemical Engineers to join its Expert Network to help train and evaluate AI models using real‑world chemical expertise. After a quick 10- to 15‑minute test, you may be invited to join Prolific as a participant, earning pay for...Remote jobHourly payFlexible hours$8 - $65 per hour
Prolific is looking for Mental Health Professionals to help train and evaluate AI models in Phoenix, Arizona. As a Domain Expert participant, you'll review AI-generated psychological... ...status and an understanding of human behavior. Join Prolific to contribute to ethical...Remote jobFlexible hours- ...we believe the safest AI is the one that’s... ...project - human data experts who probe AI models with adversarial inputs... ...misinformation, or harmful behaviors. All work is text-... ...AI systems - Evaluation coverage expands: more... ...customers trust the safety of their AI because you...Part timeRemote work
$60 - $70 per hour
...technical talent with leading AI research labs. Headquartered in... ...Dorsey . Position: AI Safety Practitioner Type: Contract... ...Role Responsibilities Evaluate AI-generated responses for safety... ...feedback to improve model alignment and safety performance...Contract workSummer workRemote work- ...Prolific is seeking Mental Health Professionals to train and evaluate AI models. You’ll start with a quick test to assess suitability and, if successful, join as a Domain Expert participant, paid to work on AI tasks from home. You'll review AI-generated psychology...Remote workWork from home
- ...Prolific is recruiting Mental Health Professionals to train and evaluate AI models as Domain Experts. You will be invited after a quick test and paid to complete AI tasks requiring focused hours. Expect earnings per task ranging with flexible hours and remote work....Remote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Behavioral Health Expert, AI Safety and Model Evaluation. Be the first to apply!




