Remote AI Image Policy Evaluator — Calibrate Models
Apply
- Remote job
Handshake is seeking an AI Image Evaluator to assess prompts and generated images for quality, accuracy, and policy compliance. You will weigh visual details, compare outputs, and explain your reasoning to help train evaluation models. Remote US role with flexible scheduling, part of Handshake AI's data-focused team. Prior experience with generative tools and content moderation is valuable, but a strong eye for detail can come from related fields. #J-18808-Ljbffr Apply
- Handshake is seeking an AI Image Evaluator to assess prompts and generated images for quality, accuracy, and policy compliance. You will weigh visual details, compare outputs... ...your reasoning to help train evaluation models. Remote US role with flexible scheduling, part...Remote jobPolicyFlexible hours
- Handshake is seeking an AI Image Evaluator to assess generated images against prompts, check... ...quality, and flag issues that violate policies. You will compare paired outputs, justify... ...choice with clear evidence, and help calibrate AI models. The role favors a trained eye from...Policy
- ...Table Image Reasoning Model Evaluator is a remote evaluation track for reviewing table image... ...Why this role matters AI data reviewers help turn... ...reviewer-quality scores by calibrating against gold-standard examples... ...with the correct policy category and severity....Remote jobPolicyHourly payFor contractors10 hours per week
- Handshake in Seattle, WA is seeking an AI Image Evaluator to help assess prompts and generated images for adherence, quality, and policy compliance. You will review outputs,... ...apply rubric-based judgments, and help calibrate models for safety and quality in Handshake AI...Policy
$45 - $55 per hour
...non-engineering content-policy evaluation role. Applicants must... ...we started Handshake AI and built the fastest-... ...labs currently improve model capabilities with... ..., precedent, and team calibration. What You Will Do Evaluate... ...Details Location: Remote, US Compensation: $45...Remote jobPolicyMonday to FridayShift work- Mercor is hiring Legal Experts to evaluate AI-generated responses for employment... ...and labor law scenarios. This fully remote, hourly contract offers flexible 6... ..., provide feedback to improve model behavior and participate in calibration sessions. Requirements include a J...Remote jobHourly payContract workFlexible hours
- Handshake is seeking an AI Policy Specialist on the Violence & Fiction team to evaluate user requests and model responses within a full conversation context, distinguishing violence... ...reasoning clearly to train models. This remote role is US-based, with 8:00 AM-5:00 PM PT...Remote jobPolicy
- ...Procurement Compliance AI Evaluator is a remote review track for evaluating AI... ..., statutory reasoning, and policy adherence; flag risk; and document... ...corrected analysis so the modeling team can train on it. Why... ...scores in inter-rater calibration cycles. Qualifications...Remote jobPolicyHourly payFor contractors10 hours per week
- ...Privacy Law AI Evaluator is a remote review track for evaluating AI outputs in... ..., statutory reasoning, and policy adherence; flag risk; and document... ...corrected analysis so the modeling team can train on it. Why... ...scores in inter-rater calibration cycles. Qualifications...Remote jobPolicyHourly payFor contractors10 hours per week
- ...AI Safety and Red-Team Evaluator is a remote red-team track for stress-testing AI systems against... ...how AuraOne hardens AI models before they ship to customers... ...classes (jailbreak, policy bypass, prompt injection)... ...new red-team rubrics. Calibrate against the broader red-team...Remote jobPolicyHourly payFor contractors10 hours per week
- ...Special Education AI Evaluator is a remote review track for evaluating AI outputs... ...workflow correctness, policy adherence, and stakeholder... ...the right next step so the modeling team can train on it. Why... ...quality scores in inter-rater calibration cycles. Qualifications...Remote jobPolicyHourly payFor contractorsWork experience placement10 hours per week
- ...Healthcare Compliance AI Evaluator is a remote clinical-review track for evaluating... ...clinical reasoning so the modeling team can close the gap.... ...scores in weekly inter-rater calibration cycles. Qualifications... ...review Legal reasoning Policy review Risk analysis...Remote jobPolicyHourly payFor contractors10 hours per week
- ...Insurance Policy AI Evaluator is a remote review track for evaluating AI outputs across insurance workflows... ...the correct treatment so the modeling team can train on it. Why this role... ...reviewer-quality scores in inter-rater calibration cycles. Qualifications Direct...Remote jobPolicyHourly payFor contractorsWork experience placement10 hours per week
- OpenTrain AI is seeking an insurance policy operations specialist to craft high-quality reasoning data for AI evaluation. You will design realistic workflows across... ...answers, and assess model outputs against... ...for improved AI behavior. Remote contractor position with...Remote workPolicyFor contractorsFlexible hours
$45 - $55 per hour
...partner is looking for an AI Safety Policy Evaluator, Violence & Threats based... ...helping improve how advanced AI models handle violent and... ...testing, policy refinement, calibration, and the identification of... ...required. Ability to work remotely in the United States on a...Remote jobPolicyHourly payFull timeMonday to Friday- micro1 is seeking an AI Image & Video Evaluation Specialist (remote, contractor) to generate and compare AI-generated visuals across platforms. You will assess realism, composition, lighting, color, anatomy, and text rendering, delivering structured analyses. No formal...Remote jobFor contractors
- **Job Title: AI Trainer || Image Quality Evaluator || English** **Location**: Remote | Work from Home **Employment Type:** Project-based | Contract We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists...Remote workContract workWork from homeMonday to FridayDay shift
$50 - $60 per hour
A technology company is seeking an Investment Partner to join their remote team in the United States. This role involves training AI models, measuring their progress, and evaluating outputs to enhance their quality. Ideal candidates should possess expert-level financial...Remote jobHourly payFlexible hours$50 - $60 per hour
A technology company specializing in AI and finance is seeking a Wealth Advisor to help train AI models. In this independent contract role, you will measure the effectiveness... .... This position offers flexibility to work remotely and competitive hourly pay starting at $50-$60...Remote jobHourly payContract work$43 - $47 per hour
...Overview Help improve how advanced AI models respond to sensitive topics in Czech. In this remote, hourly role, you will use... ...fluency and cultural judgment to evaluate model behavior and support... ...and safety, content moderation, policy evaluation, or adversarial testing...Remote jobPolicyHourly payImmediate start- Mercor is seeking experienced AI Safety Practitioners to assess the safety, quality, and alignment of frontier AI models across complex, policy-sensitive topics. You will evaluate AI-generated responses, apply safety policies, and provide structured feedback to improve...Policy
$18 - $22 per hour
...improve the safety of advanced AI models by applying Thai language fluency and cultural judgment to evaluate how models respond to sensitive... ...and safety, content moderation, policy evaluation, or adversarial testing. Work Terms Remote, with Southeast Asia preferred....Remote jobPolicyHourly payImmediate start$43 - $47 per hour
...Role Overview Help evaluate and strengthen how advanced AI systems respond to sensitive topics... ...analysis to support safer model behavior. Training is... ...safety, content moderation, policy evaluation, or adversarial... ...testing. Work Terms Remote, hourly engagement. Immediate...Remote jobPolicyHourly payImmediate start$48 - $52 per hour
...improve the safety of advanced AI systems by applying Chinese... ...fluency and cultural judgment to evaluate how models respond to sensitive topics.... ...and safety, content moderation, policy evaluation, or adversarial testing. Work Terms Remote, with East Asia preferred....Remote jobPolicyHourly payImmediate start- ...HeartFlow, Inc. is looking for an Imaging Analyst to create 3D models of coronary arteries from CT scans using proprietary software. This position is remote but requires residency in Austin, TX, with possible office visits. The role involves interpreting CT data, performing...Remote work
$30 per hour
...: $20-$30/hour Location: Remote Commitment: 10-40 hours/week... ...Role Responsibilities Evaluate outputs from large language models and autonomous agent systems... .... Participate in calibration sessions to ensure consistent... ...experience in LLM evaluation, AI output analysis, QA/...Remote jobHourly payContract work- Handshake is seeking an AI Policy Generalist in Seattle to translate complex customer... ...into consistent, well-reasoned evaluations of AI model behavior. You will read user requests... ...customer expectations, contribute to calibration discussions, and write precise rationales...Policy
- Evaluate generated images against prompts for adherence, composition... ...images under customer policies covering sexual... ...refine prompts to test model quality and safety boundaries... ...teams Participate in calibration discussions and... ...are not required Prior AI evaluation experience...PolicyMonday to Friday
$55 - $90 per hour
...is a non-engineering content-policy evaluation role. Applicants must... ...2025, we started Handshake AI and built the fastest-growing... ...Frontier AI labs currently improve model capabilities with various... ...context, precedent, and team calibration. What You Will Do Evaluate...PolicyMonday to FridayShift work- Join to apply for the Data Evaluator/Analyst role at MELE Associates... ...statistical and econometric models, and prepares preliminary findings... ..., statistics, public policy, data science, or a related field... ...Benefits MELE Offers Hybrid remote/office work environment. Employer...Remote workPolicyContract workFor contractorsWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Remote AI Image Policy Evaluator — Calibrate Models. Be the first to apply!
- education evaluator Brooklyn, NY
- program evaluator Brooklyn, NY
- social media evaluator Brooklyn, NY
- work from home web search evaluator Brooklyn, NY
- ai evaluator Brooklyn, NY
- clinical evaluator Brooklyn, NY
- evaluator Brooklyn, NY
- quality evaluator Brooklyn, NY
- junior remote developer Brooklyn, NY
- college summer internship remote Brooklyn, NY




