Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Remote AI Image Policy Evaluator — Calibrate Models

Apply

Brooklyn, NY
  • Remote job

Handshake is seeking an AI Image Evaluator to assess prompts and generated images for quality, accuracy, and policy compliance. You will weigh visual details, compare outputs, and explain your reasoning to help train evaluation models. Remote US role with flexible scheduling, part of Handshake AI's data-focused team. Prior experience with generative tools and content moderation is valuable, but a strong eye for detail can come from related fields. #J-18808-Ljbffr Apply

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Remote AI Image Policy Evaluator — Calibrate Models in Brooklyn, NY vacancy
  • Handshake is seeking an AI Image Evaluator to assess prompts and generated images for quality, accuracy, and policy compliance. You will weigh visual details, compare outputs...  ...your reasoning to help train evaluation models. Remote US role with flexible scheduling, part... 
    Remote job
    Policy
    Flexible hours

    Apply

    Seattle, WA
    2 days ago
  • Handshake is seeking an AI Image Evaluator to assess generated images against prompts, check...  ...quality, and flag issues that violate policies. You will compare paired outputs, justify...  ...choice with clear evidence, and help calibrate AI models. The role favors a trained eye from... 
    Policy

    handshake

    New York, NY
    2 days ago
  •  ...Table Image Reasoning Model Evaluator is a remote evaluation track for reviewing table image...  ...Why this role matters AI data reviewers help turn...  ...reviewer-quality scores by calibrating against gold-standard examples...  ...with the correct policy category and severity.... 
    Remote job
    Policy
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    7 days ago
  • Handshake in Seattle, WA is seeking an AI Image Evaluator to help assess prompts and generated images for adherence, quality, and policy compliance. You will review outputs,...  ...apply rubric-based judgments, and help calibrate models for safety and quality in Handshake AI... 
    Policy

    Handshake

    Seattle, WA
    4 days ago
  • $45 - $55 per hour

     ...non-engineering content-policy evaluation role. Applicants must...  ...we started Handshake AI and built the fastest-...  ...labs currently improve model capabilities with...  ..., precedent, and team calibration. What You Will Do Evaluate...  ...Details Location: Remote, US Compensation: $45... 
    Remote job
    Policy
    Monday to Friday
    Shift work

    Apply

    Brooklyn, NY
    4 days ago
  • Mercor is hiring Legal Experts to evaluate AI-generated responses for employment...  ...and labor law scenarios. This fully remote, hourly contract offers flexible 6...  ..., provide feedback to improve model behavior and participate in calibration sessions. Requirements include a J... 
    Remote job
    Hourly pay
    Contract work
    Flexible hours

    Intellerzone

    New York, NY
    3 days ago
  • Handshake is seeking an AI Policy Specialist on the Violence & Fiction team to evaluate user requests and model responses within a full conversation context, distinguishing violence...  ...reasoning clearly to train models. This remote role is US-based, with 8:00 AM-5:00 PM PT... 
    Remote job
    Policy

    Apply

    Brooklyn, NY
    4 days ago
  •  ...Procurement Compliance AI Evaluator is a remote review track for evaluating AI...  ..., statutory reasoning, and policy adherence; flag risk; and document...  ...corrected analysis so the modeling team can train on it. Why...  ...scores in inter-rater calibration cycles. Qualifications... 
    Remote job
    Policy
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    7 days ago
  •  ...Privacy Law AI Evaluator is a remote review track for evaluating AI outputs in...  ..., statutory reasoning, and policy adherence; flag risk; and document...  ...corrected analysis so the modeling team can train on it. Why...  ...scores in inter-rater calibration cycles. Qualifications... 
    Remote job
    Policy
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  •  ...AI Safety and Red-Team Evaluator is a remote red-team track for stress-testing AI systems against...  ...how AuraOne hardens AI models before they ship to customers...  ...classes (jailbreak, policy bypass, prompt injection)...  ...new red-team rubrics. Calibrate against the broader red-team... 
    Remote job
    Policy
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  •  ...Special Education AI Evaluator is a remote review track for evaluating AI outputs...  ...workflow correctness, policy adherence, and stakeholder...  ...the right next step so the modeling team can train on it. Why...  ...quality scores in inter-rater calibration cycles. Qualifications... 
    Remote job
    Policy
    Hourly pay
    For contractors
    Work experience placement
    10 hours per week

    AuraOne Human Data

    Remote
    7 days ago
  •  ...Healthcare Compliance AI Evaluator is a remote clinical-review track for evaluating...  ...clinical reasoning so the modeling team can close the gap....  ...scores in weekly inter-rater calibration cycles. Qualifications...  ...review Legal reasoning Policy review Risk analysis... 
    Remote job
    Policy
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  •  ...Insurance Policy AI Evaluator is a remote review track for evaluating AI outputs across insurance workflows...  ...the correct treatment so the modeling team can train on it. Why this role...  ...reviewer-quality scores in inter-rater calibration cycles. Qualifications Direct... 
    Remote job
    Policy
    Hourly pay
    For contractors
    Work experience placement
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  • OpenTrain AI is seeking an insurance policy operations specialist to craft high-quality reasoning data for AI evaluation. You will design realistic workflows across...  ...answers, and assess model outputs against...  ...for improved AI behavior. Remote contractor position with... 
    Remote work
    Policy
    For contractors
    Flexible hours

    OpenTrain AI

    Brooklyn, NY
    3 days ago
  • $45 - $55 per hour

     ...partner is looking for an AI Safety Policy Evaluator, Violence & Threats based...  ...helping improve how advanced AI models handle violent and...  ...testing, policy refinement, calibration, and the identification of...  ...required. Ability to work remotely in the United States on a... 
    Remote job
    Policy
    Hourly pay
    Full time
    Monday to Friday

    jobgether

    United States
    5 days ago
  • micro1 is seeking an AI Image & Video Evaluation Specialist (remote, contractor) to generate and compare AI-generated visuals across platforms. You will assess realism, composition, lighting, color, anatomy, and text rendering, delivering structured analyses. No formal... 
    Remote job
    For contractors

    Aiworkexpert

    Brooklyn, NY
    2 days ago
  • **Job Title: AI Trainer || Image Quality Evaluator || English** **Location**: Remote | Work from Home **Employment Type:** Project-based | Contract We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists... 
    Remote work
    Contract work
    Work from home
    Monday to Friday
    Day shift

    iMerit Technology

    Louisiana
    a month ago
  • $50 - $60 per hour

    A technology company is seeking an Investment Partner to join their remote team in the United States. This role involves training AI models, measuring their progress, and evaluating outputs to enhance their quality. Ideal candidates should possess expert-level financial... 
    Remote job
    Hourly pay
    Flexible hours

    DataAnnotation

    Raleigh, NC
    1 day ago
  • $50 - $60 per hour

    A technology company specializing in AI and finance is seeking a Wealth Advisor to help train AI models. In this independent contract role, you will measure the effectiveness...  .... This position offers flexibility to work remotely and competitive hourly pay starting at $50-$60... 
    Remote job
    Hourly pay
    Contract work

    DataAnnotation

    Brooklyn, NY
    4 days ago
  • $43 - $47 per hour

     ...Overview Help improve how advanced AI models respond to sensitive topics in Czech. In this remote, hourly role, you will use...  ...fluency and cultural judgment to evaluate model behavior and support...  ...and safety, content moderation, policy evaluation, or adversarial testing... 
    Remote job
    Policy
    Hourly pay
    Immediate start

    SaidGig

    Remote
    3 days ago
  • Mercor is seeking experienced AI Safety Practitioners to assess the safety, quality, and alignment of frontier AI models across complex, policy-sensitive topics. You will evaluate AI-generated responses, apply safety policies, and provide structured feedback to improve... 
    Policy

    Mercor

    San Francisco, CA
    5 days ago
  • $18 - $22 per hour

     ...improve the safety of advanced AI models by applying Thai language fluency and cultural judgment to evaluate how models respond to sensitive...  ...and safety, content moderation, policy evaluation, or adversarial testing. Work Terms Remote, with Southeast Asia preferred.... 
    Remote job
    Policy
    Hourly pay
    Immediate start

    SaidGig

    Remote
    2 days ago
  • $43 - $47 per hour

     ...Role Overview Help evaluate and strengthen how advanced AI systems respond to sensitive topics...  ...analysis to support safer model behavior. Training is...  ...safety, content moderation, policy evaluation, or adversarial...  ...testing. Work Terms Remote, hourly engagement. Immediate... 
    Remote job
    Policy
    Hourly pay
    Immediate start

    SaidGig

    Europe
    1 day ago
  • $48 - $52 per hour

     ...improve the safety of advanced AI systems by applying Chinese...  ...fluency and cultural judgment to evaluate how models respond to sensitive topics....  ...and safety, content moderation, policy evaluation, or adversarial testing. Work Terms Remote, with East Asia preferred.... 
    Remote job
    Policy
    Hourly pay
    Immediate start

    SaidGig

    Remote
    2 days ago
  •  ...HeartFlow, Inc. is looking for an Imaging Analyst to create 3D models of coronary arteries from CT scans using proprietary software. This position is remote but requires residency in Austin, TX, with possible office visits. The role involves interpreting CT data, performing... 
    Remote work

    HeartFlow

    Austin, TX
    5 days ago
  • $30 per hour

     ...: $20-$30/hour Location: Remote Commitment: 10-40 hours/week...  ...Role Responsibilities Evaluate outputs from large language models and autonomous agent systems...  .... Participate in calibration sessions to ensure consistent...  ...experience in LLM evaluation, AI output analysis, QA/... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    1 day ago
  • Handshake is seeking an AI Policy Generalist in Seattle to translate complex customer...  ...into consistent, well-reasoned evaluations of AI model behavior. You will read user requests...  ...customer expectations, contribute to calibration discussions, and write precise rationales... 
    Policy

    Apply

    Seattle, WA
    4 days ago
  • Evaluate generated images against prompts for adherence, composition...  ...images under customer policies covering sexual...  ...refine prompts to test model quality and safety boundaries...  ...teams Participate in calibration discussions and...  ...are not required Prior AI evaluation experience... 
    Policy
    Monday to Friday

    Jobtailor

    Seattle, WA
    4 days ago
  • $55 - $90 per hour

     ...is a non-engineering content-policy evaluation role. Applicants must...  ...2025, we started Handshake AI and built the fastest-growing...  ...Frontier AI labs currently improve model capabilities with various...  ...context, precedent, and team calibration. What You Will Do Evaluate... 
    Policy
    Monday to Friday
    Shift work

    Handshake

    Seattle, WA
    4 days ago
  • Join to apply for the Data Evaluator/Analyst role at MELE Associates...  ...statistical and econometric models, and prepares preliminary findings...  ..., statistics, public policy, data science, or a related field...  ...Benefits MELE Offers Hybrid remote/office work environment. Employer... 
    Remote work
    Policy
    Contract work
    For contractors
    Work at office

    MELE Associates

    Baltimore, MD
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Remote AI Image Policy Evaluator — Calibrate Models. Be the first to apply!