Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Model Policy Trainer, Image Evaluation - Seattle Onsite

$45 - $55 per hour

Handshake

Job Description

Job Description

About Handshake

Handshake was founded on a simple belief that everyone deserves a path to a great career, regardless of where they went to school or who they know. Today, we power 25 million job seekers, 1 million+ employers, and 1,600 educational institutions.

In 2025, we started Handshake AI and built the fastest-growing AI data business in history. We work directly with frontier AI lab researchers to create evaluations, publish benchmarks, and push the boundary of data. We've grown from $0 to ~$1B run rate and pay ~$60M to over 30K individuals every month.

Why join Handshake now:

  • Shape how every career evolves in the AI economy, at global scale, with impact your friends, family and peers can see and feel

  • Partner hand-in-hand with world-class AI labs, Fortune 500 partners and the world's top educational institutions

  • Work together with engineers, scientists, operators, and more from Palantir, Meta, Scale AI, and former YC founders

  • Build a massive, fast-growing business with billions in revenue

About Handshake AI

Human data is the core infrastructure to AI advancement. Frontier AI labs currently improve model capabilities with various data-intensive post-training techniques. We believe that data spend for AI training will increase by 3-5x in the next few years and continue for much longer as models take on new domains. Handshake AI supports all of the frontier AI labs, working on their most complex data at the largest scale.

About the Role

As an AI Image Evaluator, you will help image generation models learn two things at once: what a good image is, and what an acceptable image is.

You will look at prompts and the images a model produced from them, then answer questions like: Did the image do what the prompt asked? Is it well made, or does it have the hands, lighting, text, and anatomy problems that give generated images away? Which of two images is better, and why? Does the image violate the customer's content policy, and if so, which category and how severely? Does it depict a real person, a protected brand, or a minor in a way the policy does not allow?

The interesting cases are the close ones. Two images that look nearly identical until you notice one has a logo in the background. A stylized nude that is fine as figure study and not fine with one change of pose. A prompt that asked for "a realistic photo of a senator" and a model that complied. A beautiful image that ignored half the prompt, next to an ugly one that nailed it.

We are looking for people who already see images critically, whether that came from photography, illustration, design, years inside Midjourney and Stable Diffusion, or moderating visual content at scale. You do not need all of these. You need one deep, and the judgment to learn the rest.

This is not rote annotation. Rubrics cannot anticipate every image, and good evaluators do not apply them mechanically. You will balance the rubric's text and intent with customer expectations, precedent, and team calibration, and you will explain your reasoning clearly enough that it can train a model.

What You Will Do
  • Evaluate generated images against their prompts for adherence, composition, realism, style consistency, and technical defects

  • Compare images side by side and select the stronger one with a clear, evidence-based rationale

  • Classify images against customer content policies covering sexual content, violence, hate symbols, real-person likeness, intellectual property, and depictions of minors

  • Select the most defensible classification when an image is genuinely ambiguous, and write concise rationales that cite rubric language and specific visual details

  • Distinguish "I do not like this" from "this fails the prompt" from "this violates policy," and keep those judgments separate

  • Write and refine prompts that probe where a model's quality or safety behavior breaks down

  • Identify rubric gaps, contradictions, and emerging edge cases, and raise them with project leads and policy teams

  • Participate actively in calibration discussions; challenge interpretations respectfully and update your judgment when stronger reasoning emerges

  • Apply customer policy consistently without substituting personal taste or personal beliefs for the standard

  • Maintain accuracy and consistency across hundreds of visually similar evaluations

You May Be a Fit If
  • You have a trained eye from photography, illustration, concept art, art direction, retouching, photo editing, VFX, or visual design, and you can say precisely why one image is better than another

  • You use generative image tools heavily (Midjourney, Stable Diffusion, ComfyUI, Flux, DALL-E, Ideogram) and know their failure modes, their prompt quirks, and how their safety filters get bypassed

  • You have moderated or reviewed visual content at scale and have applied a policy taxonomy to borderline images under time pressure

  • You notice small details: an extra finger, a mismatched shadow, a brand mark, a face that is a little too familiar

  • You can hold a rubric steady across a long session of near-identical images

  • You can hold a strong opinion without becoming attached to being right

  • You explain judgment calls clearly enough that another person can audit your reasoning

  • You can separate your personal taste from the standard a customer has asked you to apply

  • You communicate clearly and precisely in writing

  • You treat sensitive imagery and difficult subject matter with maturity and sound judgment

Strong candidates may come from photography, illustration, graphic or UX design, art direction, photo editing, animation or VFX, game art, trust and safety, content moderation, brand or IP enforcement, ad review, or art education. We care more about how you see and how you reason than where you learned to do it. A degree and a technical background are not required.

Nice to Have
  • A public portfolio, publication credits, or a body of generative work (Civitai, Discord communities, LoRA or model training, published prompt work)

  • Experience judging images comparatively: portfolio review, photo competition judging, creative A/B testing, art school critique

  • Formal training in anatomy, color, lighting, or composition

  • Content moderation or trust and safety experience on an image-heavy platform

  • Working knowledge of copyright, trademark, and right-of-publicity basics

  • Prior work in AI evaluation, RLHF, image labeling, or data annotation

  • Familiarity with calibration sessions, inter-rater agreement, or adjudication workflows

Prior AI evaluation experience is helpful, but it is not required.

Sensitive-Content Notice

This role involves regular and deliberate engagement with sensitive imagery. Depending on the project, evaluations may include sexual content and nudity, graphic violence and gore, hate symbols, self-harm, and depictions of real people and of minors in contexts that must be assessed against policy. Some of this material is disturbing by design, because the purpose of the work is to teach models not to produce it.

The work is conducted within structured evaluation frameworks and professional guidelines, with exposure limits, content rotation, mandatory reporting protocols for illegal material, and access to mental health support. Candidates must be able to engage with this material carefully, responsibly, and sustainably while maintaining sound judgment and consistent work quality.

Role Details
  • Location: Seattle, WA

  • Compensation: $45-55/hr

  • Employment classification: W-2

  • Schedule: 8AM - 5PM PT

  • Weekly commitment: M-F

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the AI Model Policy Trainer, Image Evaluation - Seattle Onsite in Seattle, WA vacancy
  • $55 - $75 per hour

     ..., we started Handshake AI and built the fastest-growing...  ...researchers to create evaluations, publish benchmarks,...  ...labs currently improve model capabilities with...  ...About the Role As an AI Policy Generalist, you will turn...  ...Details Location: Seattle, WA Compensation: $... 
    Policy

    Handshake

    Seattle, WA
    1 day ago
  • $55 - $120 per hour

     ...a non-engineering content-policy evaluation role. Applicants must demonstrate...  ...2025, we started Handshake AI and built the fastest-...  ...AI labs currently improve model capabilities with various data...  ...Role Details Location: Seattle, WA, onsite Monday-Friday... 
    Policy
    Monday to Friday
    Shift work

    Handshake

    Seattle, WA
    3 days ago
  •  ...we started Handshake AI and built the fastest-...  ...researchers to create evaluations, publish benchmarks, and...  ...currently improve model capabilities with various...  ...Role As an AI Model Policy Trainer focused on mental health...  ...position is based in Seattle, Washington. Candidates... 
    Policy
    Immediate start
    Relocation

    Handshake

    Seattle, WA
    10 days ago
  • Handshake is seeking an AI Policy Generalist in Seattle to translate complex customer policies into consistent, well-reasoned evaluations of AI model behavior. You will read user requests, model outputs, and history to determine the correct policy category and cite evidence... 
    Policy

    Apply

    Seattle, WA
    14 hours ago
  • Handshake in Seattle, WA is seeking an AI Policy Generalist to turn complex customer policies into consistent evaluations of AI model behavior. You will read user requests, model responses, and conversation history, determine policy category, and craft concise, evidence... 
    Policy

    Handshake

    Seattle, WA
    4 days ago
  • $166k - $258k

     ...relationships with our customers.Nordstrom Technology is moving to an AI Native operating model, and we're rebuilding the processes, tools, and standards...  ...job postings open for at least one day after the posting date. 2026 Nordstrom, IncSummaryLocation: Seattle, WAType: Full time
    Full time

    Nordstrom

    Seattle, WA
    2 days ago
  • $200k - $300k

    Member of Technical Staff — Model Optimization and Inference (New Grad) Seattle, Washington About Nuance Labs Nuance Labs...  ...photorealistic, real-time AI avatars with emotional intelligence...  ...conversations, including eviction policies, compression, and memory-efficient... 
    Policy
    Internship
    H1b
    Work at office
    Visa sponsorship

    Nuance Labs

    Seattle, WA
    4 days ago
  •  ...Talent LLC seeks a Machine Learning Engineer - Generative Imaging to own the model layer for an AI-powered image production platform. You will focus on...  ...preserving production intelligence and workflows. You will evaluate architectures quickly and craft an abstraction layer to... 

    Fuel Talent LLC

    Seattle, WA
    1 day ago
  •  ...efficiency), and are aligned at the intersections of assets, processes, policies and people delivering value.ProSidian clients represent a broad...  ...52) for in-person meetings.Group Meeting Facilitator - HNRTC | Seattle, WA - GSSC Candidates shall work to support requirements for FY... 
    Policy
    For contractors
    Work at office
    Remote work

    Prosidian Consultng

    Seattle, WA
    2 days ago
  •  ...efficiency), and are aligned at the intersections of assets, processes, policies and people delivering value.ProSidian clients represent a broad...  ...9352) for in-person meetings.Group Meeting Facilitator - HAB | Seattle, WA - GSSC Candidates shall work to support requirements for FY... 
    Policy
    For contractors
    Work at office
    Local area
    Remote work

    Prosidian Consultng

    Seattle, WA
    2 days ago
  • $7,641 - $8,771 per month

     ...Urology job at University of Washington. Seattle, WA. The Department of Urology at the...  ...Radiation oncology Uropathology Urologic imaging techniques; Ability to perform complex...  ...urology training program) including referee evaluation form Society of Urologic Oncology... 
    Full time
    Traineeship
    H1b

    Shell Lubricants Hub Hamburg

    Seattle, WA
    1 day ago
  • $124k - $280k

     ...Strategy Consulting - Business Model Reinvention - Senior Manager...  ...set forth within the following policy: more about how we work: only...  ...closely with team members. We evaluate these factors thoughtfully to...  ...San Francisco; PA-Philadelphia; WA-Seattle; TX-HoustonType: Full time
    Policy
    Full time
    H1b

    PwC

    Seattle, WA
    5 days ago
  • $99k - $232k

     ...OpportunityAs a Strategy& - Business Model Reinvention - Manager, you...  ...forth within the following policy: more about how we work: only...  ...closely with team members. We evaluate these factors thoughtfully to...  ...San Francisco; PA-Philadelphia; WA-Seattle; TX-HoustonType: Full time
    Policy
    Full time
    H1b

    PwC

    Seattle, WA
    5 days ago
  • $77k - $202k

     ...Strategy Consulting Business Model Reinvention - Senior Associate...  ...set forth within the following policy: more about how we work: only...  ...closely with team members. We evaluate these factors thoughtfully to...  ...San Francisco; PA-Philadelphia; WA-Seattle; TX-HoustonType: Full time
    Policy
    Full time
    H1b

    PwC

    Seattle, WA
    5 days ago
  • $155k - $410k

     ...Strategy Consulting Business Model Reinvention - Director, you will...  ...forth within the following policy: more about how we work: only...  ...closely with team members. We evaluate these factors thoughtfully to...  ...San Francisco; PA-Philadelphia; WA-Seattle; TX-HoustonType: Full time
    Policy
    Full time
    H1b

    PwC

    Seattle, WA
    6 days ago
  • Welo Data is hiring a Data Labeling Associate in Washington with expertise in evaluating AI systems. This role focuses on providing structured feedback on model outputs and requires strong writing skills and attention to detail. As a full-time employee, you will engage... 
    Full time

    Welo Data

    Seattle, WA
    14 hours ago
  •  ...building machine learning models and systems to protect...  ...challenges: evolving policies, surging complexity in...  ...for multimodality (text/image/video/audio), Unified Understanding...  ...the frontier of AI today. This topic...  ...technologies.- Experience with evaluation of AI systems, LLM... 
    Policy
    Flexible hours
    Shift work

    TikTok

    Seattle, WA
    4 days ago
  • Welo Data is seeking a Data Labeling Associate in Washington State. This full-time role involves evaluating AI model outputs and improving data quality. The ideal candidate should have native-level Australian English proficiency, a bachelor’s degree, and strong analytical... 
    Full time

    Welo Data

    Seattle, WA
    14 hours ago
  • $300k - $320k

     ...role: We are seeking a Technical Program Manager to lead our AI model evaluation initiatives across multiple workstreams. This role will be...  ...with our Research, Trust & Safety, Frontier Redteaming, and Policy teams, you will drive high-priority evaluation projects to build... 
    Policy
    Work at office
    Home office
    Visa sponsorship
    Relocation package

    Anthropic

    Seattle, WA
    2 days ago
  • $100 per hour

    A leading technology firm is seeking finance experts to enhance AI models. Responsibilities include evaluating performance in capital markets and creating assessment rubrics. Candidates should have 2+ years in finance fields like investment banking and possess strong financial... 
    Remote job
    Hourly pay
    10 hours per week

    Turing

    Seattle, WA
    14 hours ago
  • Girls Soccer JV Assistant Coach, Bishop Blanchet HS Seattle JobID: 10903 Date Posted: 5/13/2026 Reports to Athletic Director Bishop...  ...Compliance with all Bishop Blanchet HS, Metro League, and WIAA policies, procedures, and coaches certification/education. Supervision... 
    Policy
    Monday to Friday

    Archdiocese of Seattle Catholic Schools

    Seattle, WA
    1 day ago
  • $305k

     ...seeking a Product Manager in Seattle to lead model launches for Claude Code and...  ...performance and design evaluations that meet real-world coding...  ...candidate has a background in AI, experience in product management...  ...$460,000, with a hybrid work policy. #J-18808-Ljbffr Anthropic
    Policy

    Anthropic

    Seattle, WA
    3 days ago
  • $182.75k - $330k

     ...We are the industry leader in image-guided therapy, helping to improve...  ...at least 3 days per week. Onsite roles require full-time presence...  ...commuting distance to the Seattle WA Area #LI-Field #LI-PH1...  ...#ImageGuidedTherapy It is the policy of Philips to provide equal employment... 
    Policy
    Full time
    Work at office
    Local area
    Work visa
    Relocation package
    3 days per week

    Philips

    Seattle, WA
    1 day ago
  •  ...Location: Washington, Seattle Workplace Flexibility:...  ...of Olympus’ digital and AI-powered endoscopy platform...  ...execution, including evaluation, product demonstrations...  ...centric mindsetOffers onsite, hybrid and field work...  ...potential, together.It is the policy of Olympus to extend... 
    Policy
    Local area
    Worldwide

    Olympus

    Seattle, WA
    6 days ago
  •  ...industry. We currently have a need for Onsite Contract Interpreters in the Seattle, WA Metro who have a sincere...  ...clients Follows all Propio L.S. policies and procedures related to information...  ...interpreting experience Propio’s evaluation process conforms to interpreting... 
    Policy
    Contract work
    Seasonal work
    Local area
    Remote work

    Propio Language Services

    Seattle, WA
    2 days ago
  • $180k

     ...SpaceXAI's mission is to create AI systems that can accurately...  ...multimodal engineer on the Imagine Model Team, you will develop cutting...  ...and generation across image and video modalities, while also...  ...visual and audio data.  Design evaluation frameworks, metrics,... 
    Temporary work

    SpaceXAI

    Seattle, WA
    2 days ago
  • $232.56k - $427.5k

    Research Scientist in AI Foundation Model Infrastructure - Seed - Graduates - 2027 Start (PhD) Location: Seattle Team: Technology Employment Type: Regular Job Code: A224405...  ...infrastructure for large-scale model training, evaluation, and inference. Optimize distributed... 
    Temporary work
    Local area

    ByteDance

    Seattle, WA
    2 days ago
  • $305k

     ...interpretable, and steerable AI systems. We want AI to...  ..., engineers, policy experts, and business leaders...  ...: San Francisco and Seattle only. About The Role As...  ...Manager on Claude Code's model performance team, you...  ...prompt engineering, and evaluation methodology. Are a systems... 
    Policy
    Visa sponsorship

    Anthropic

    Seattle, WA
    3 days ago
  • $24 per hour

     ...Trainer - Full Time We offer a full benefits package, PTO, weekly...  ...Primary Location: Downtown Seattle; occasional travel to SeaTac,...  ...classroom and hands-on training, evaluating performance, and fostering a...  ...and sick time or under a PTO policy, depending on local... 
    Policy
    Weekly pay
    Full time
    Traineeship
    Work at office
    Local area

    Securitas Security Services USA, Inc.

    Seattle, WA
    1 day ago
  • Job Description Automation Onsite Administrator - Seattle, WA The Automated Onsite...  ...effective presentations and evaluation of re-training needs for...  ...the pharmacy care delivery model, making patient care safer...  ...paramount. We adhere to strict policies to prevent discrimination... 
    Policy
    Worldwide
    Monday to Friday
    Shift work

    Omnicell, Inc.

    Seattle, WA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Model Policy Trainer, Image Evaluation - Seattle Onsite. Be the first to apply!