Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Model Policy Trainer, Image Evaluation - Seattle Onsite

$45 - $55 per hour

Handshake

Job Description

Job Description

About Handshake

Handshake was founded on a simple belief that everyone deserves a path to a great career, regardless of where they went to school or who they know. Today, we power 25 million job seekers, 1 million+ employers, and 1,600 educational institutions.

In 2025, we started Handshake AI and built the fastest-growing AI data business in history. We work directly with frontier AI lab researchers to create evaluations, publish benchmarks, and push the boundary of data. We've grown from $0 to ~$1B run rate and pay ~$60M to over 30K individuals every month.

Why join Handshake now:

  • Shape how every career evolves in the AI economy, at global scale, with impact your friends, family and peers can see and feel

  • Partner hand-in-hand with world-class AI labs, Fortune 500 partners and the world's top educational institutions

  • Work together with engineers, scientists, operators, and more from Palantir, Meta, Scale AI, and former YC founders

  • Build a massive, fast-growing business with billions in revenue

About Handshake AI

Human data is the core infrastructure to AI advancement. Frontier AI labs currently improve model capabilities with various data-intensive post-training techniques. We believe that data spend for AI training will increase by 3-5x in the next few years and continue for much longer as models take on new domains. Handshake AI supports all of the frontier AI labs, working on their most complex data at the largest scale.

About the Role

As an AI Image Evaluator, you will help image generation models learn two things at once: what a good image is, and what an acceptable image is.

You will look at prompts and the images a model produced from them, then answer questions like: Did the image do what the prompt asked? Is it well made, or does it have the hands, lighting, text, and anatomy problems that give generated images away? Which of two images is better, and why? Does the image violate the customer's content policy, and if so, which category and how severely? Does it depict a real person, a protected brand, or a minor in a way the policy does not allow?

The interesting cases are the close ones. Two images that look nearly identical until you notice one has a logo in the background. A stylized nude that is fine as figure study and not fine with one change of pose. A prompt that asked for "a realistic photo of a senator" and a model that complied. A beautiful image that ignored half the prompt, next to an ugly one that nailed it.

We are looking for people who already see images critically, whether that came from photography, illustration, design, years inside Midjourney and Stable Diffusion, or moderating visual content at scale. You do not need all of these. You need one deep, and the judgment to learn the rest.

This is not rote annotation. Rubrics cannot anticipate every image, and good evaluators do not apply them mechanically. You will balance the rubric's text and intent with customer expectations, precedent, and team calibration, and you will explain your reasoning clearly enough that it can train a model.

What You Will Do
  • Evaluate generated images against their prompts for adherence, composition, realism, style consistency, and technical defects

  • Compare images side by side and select the stronger one with a clear, evidence-based rationale

  • Classify images against customer content policies covering sexual content, violence, hate symbols, real-person likeness, intellectual property, and depictions of minors

  • Select the most defensible classification when an image is genuinely ambiguous, and write concise rationales that cite rubric language and specific visual details

  • Distinguish "I do not like this" from "this fails the prompt" from "this violates policy," and keep those judgments separate

  • Write and refine prompts that probe where a model's quality or safety behavior breaks down

  • Identify rubric gaps, contradictions, and emerging edge cases, and raise them with project leads and policy teams

  • Participate actively in calibration discussions; challenge interpretations respectfully and update your judgment when stronger reasoning emerges

  • Apply customer policy consistently without substituting personal taste or personal beliefs for the standard

  • Maintain accuracy and consistency across hundreds of visually similar evaluations

You May Be a Fit If
  • You have a trained eye from photography, illustration, concept art, art direction, retouching, photo editing, VFX, or visual design, and you can say precisely why one image is better than another

  • You use generative image tools heavily (Midjourney, Stable Diffusion, ComfyUI, Flux, DALL-E, Ideogram) and know their failure modes, their prompt quirks, and how their safety filters get bypassed

  • You have moderated or reviewed visual content at scale and have applied a policy taxonomy to borderline images under time pressure

  • You notice small details: an extra finger, a mismatched shadow, a brand mark, a face that is a little too familiar

  • You can hold a rubric steady across a long session of near-identical images

  • You can hold a strong opinion without becoming attached to being right

  • You explain judgment calls clearly enough that another person can audit your reasoning

  • You can separate your personal taste from the standard a customer has asked you to apply

  • You communicate clearly and precisely in writing

  • You treat sensitive imagery and difficult subject matter with maturity and sound judgment

Strong candidates may come from photography, illustration, graphic or UX design, art direction, photo editing, animation or VFX, game art, trust and safety, content moderation, brand or IP enforcement, ad review, or art education. We care more about how you see and how you reason than where you learned to do it. A degree and a technical background are not required.

Nice to Have
  • A public portfolio, publication credits, or a body of generative work (Civitai, Discord communities, LoRA or model training, published prompt work)

  • Experience judging images comparatively: portfolio review, photo competition judging, creative A/B testing, art school critique

  • Formal training in anatomy, color, lighting, or composition

  • Content moderation or trust and safety experience on an image-heavy platform

  • Working knowledge of copyright, trademark, and right-of-publicity basics

  • Prior work in AI evaluation, RLHF, image labeling, or data annotation

  • Familiarity with calibration sessions, inter-rater agreement, or adjudication workflows

Prior AI evaluation experience is helpful, but it is not required.

Sensitive-Content Notice

This role involves regular and deliberate engagement with sensitive imagery. Depending on the project, evaluations may include sexual content and nudity, graphic violence and gore, hate symbols, self-harm, and depictions of real people and of minors in contexts that must be assessed against policy. Some of this material is disturbing by design, because the purpose of the work is to teach models not to produce it.

The work is conducted within structured evaluation frameworks and professional guidelines, with exposure limits, content rotation, mandatory reporting protocols for illegal material, and access to mental health support. Candidates must be able to engage with this material carefully, responsibly, and sustainably while maintaining sound judgment and consistent work quality.

Role Details
  • Location: Seattle, WA

  • Compensation: $45-55/hr

  • Employment classification: W-2

  • Schedule: 8AM - 5PM PT

  • Weekly commitment: M-F

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the AI Model Policy Trainer, Image Evaluation - Seattle Onsite in Washington DC vacancy
  • $55 - $75 per hour

     ..., we started Handshake AI and built the fastest-growing...  ...researchers to create evaluations, publish benchmarks,...  ...labs currently improve model capabilities with...  ...About the Role As an AI Policy Generalist, you will turn...  ...Details Location: Seattle, WA Compensation: $... 
    Policy

    Handshake

    Washington DC
    1 day ago
  • $145k - $200k

     ...expertise in enabling ML models in production. We deploy AI models to run in variety of...  ...and the ability to quickly evaluate and integrate new models...  ...the posting is specified as Onsite, you are required to work...  ...will be processed by Palantir, please see our Privacy Policy.
    Policy
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Washington DC
    3 days ago
  •  ...we started Handshake AI and built the fastest-...  ...researchers to create evaluations, publish benchmarks, and...  ...currently improve model capabilities with various...  ...Role As an AI Model Policy Trainer focused on mental health...  ...position is based in Seattle, Washington. Candidates... 
    Policy
    Immediate start
    Relocation

    Handshake

    Washington DC
    7 days ago
  • $195k - $315k

     ...and oversight of our enterprise AI governance framework. The ideal...  ...to leading and enhancing GEICO’s model and AI governance program, including frameworks, policies and processes, to address emerging...  ...monitoring and testing to evaluate model performance in line with existing... 
    Policy
    Hourly pay
    Work experience placement
    Local area
    3 days per week

    GEICO

    Bethesda, MD
    20 hours ago
  • $75k - $95k

    Training Evaluation & Performance Specialist Arlington, VA Dynamis is seeking an analytical...  ...evaluation frameworks, including the Kirkpatrick Model or comparable evaluation methodologies,...  ..., Inc.’s Equal Employment Opportunity policy, we do not discriminate on the basis of... 
    Policy
    Contract work
    For contractors
    For subcontractor
    Work at office
    Remote work

    Dynamis, Inc.

    Arlington, VA
    4 days ago
  •  ...Cost Estimating, with additional benchmark capabilities in AI/ML & Decision Analytics, Modeling and Simulation, and Future Force Assessments. SPA has a...  ...,000.00/Yr.Job SummaryID: 2026-23100Category: Strategy, Policy, & Advisory ServicesSecurity Clearance Requirement: Top... 
    Policy
    Flexible hours

    MCR

    Alexandria, VA
    2 days ago
  •  ...the intersections of assets, processes, policies, and people delivering value. See Link To...  ...a Technical Training and Development Evaluation Specialist | Technical Training and Development...  ...evaluation frameworks (Kirkpatrick Model), conduct performance assessments, analyze... 
    Policy
    Full time
    Contract work
    Temporary work
    For contractors
    H1b
    Work at office
    Flexible hours

    Prosidian Consultng

    Arlington, VA
    2 days ago
  • Technical Training and Development Evaluation Specialist | Technical Training and Development...  ...the intersections of assets, processes, policies, and people delivering value. See Link...  ...Develop evaluation frameworks (Kirkpatrick Model), conduct performance assessments,... 
    Policy
    Full time
    Contract work
    Temporary work
    For contractors
    Work at office
    Flexible hours

    ProSidian Consulting, LLC

    Arlington, VA
    1 day ago
  • $180k

     ...SpaceXAI's mission is to create AI systems that can accurately...  ...multimodal engineer on the Imagine Model Team, you will develop cutting...  ...and generation across image and video modalities, while also...  ...visual and audio data.  Design evaluation frameworks, metrics,... 
    Temporary work

    SpaceXAI

    Washington DC
    2 days ago
  • $28 per hour

     ...Description Job Description Why work with YWCA Seattle King Snohomish? YWCA SKS is the region...  .... Clients have access to a range of onsite services, including case management, life...  ...reports. Understand and follow all policies in the EDNSC Staff Handbook as well as... 
    Policy
    Hourly pay
    Permanent employment
    Full time
    Work at office
    Immediate start
    Night shift
    Weekend work

    YWCA Seattle King Snohomish

    Washington DC
    1 day ago
  • $28 per hour

     ...Description Job Description Why work with YWCA Seattle King Snohomish?   YWCA SKS is the...  ...Demonstrated application of established policies, procedures, laws and regulations ~...  ...time, and Occasionally under 20% time #LI-onsite       YWCA encourages applicants... 
    Policy
    Hourly pay
    Permanent employment
    Full time
    Casual work
    Work at office
    Local area
    Immediate start
    Weekend work

    YWCA Seattle King Snohomish

    Washington DC
    8 days ago
  •  ...Corporate TrainerThe Corporate Trainer will develop and oversee training for franchisees...  ...brand ambassador in company policies, procedures, standards, and goals. Maintain...  ...opening, additional support or evaluation, follow up)Conduct onsite reviews and audits (Prime visits) of... 
    Policy

    Toastique

    Washington DC
    2 days ago
  • $115k - $135k

     ...research, development, test, and evaluation (RDT&E) support. As we...  ...accurate, and aligned with DHS policy.ResponsibilitiesAs a Training...  ...independentlyAbility to work onsite in Washington, DC with flexibility...  ...Together with Flexibility model allows you to work 60% in-... 
    Policy
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Flexible hours

    Battelle Memorial Institute

    Arlington, VA
    2 days ago
  •  ...instructional materials in response to evaluations from learners and organizational changes...  ...Leadership to maintain awareness of changes in policies, procedures, regulations, and...  ...to change. ** Some positions require onsite presence and/or structured hours/shifts.... 
    Policy
    Traineeship
    Shift work

    Mango Health

    Washington DC
    20 hours ago
  • $55.3k - $79k

     ...Summary As a DC CX Trainer at Gainwell, you will support...  ...support of business priorities. Evaluate training effectiveness...  ...environment with a combination of onsite and remote work; training delivery...  ...generous, flexible vacation policy, a  , and educational... 
    Policy
    Full time
    Remote work
    Flexible hours

    Gainwell Technologies LLC

    Washington DC
    more than 2 months ago
  • $63.31k - $85.66k

     ...into performance evaluations and development goals...  ...aligns with NIH policies, security requirements...  ..., and support models (e.g., NED, NIH network...  ...). HDI Trainer certification....  ...~ Bethesda, MD (onsite with potential hybrid...  ...ready capabilities in AI, cloud, cyber and... 
    Policy
    Contract work
    Temporary work
    Immediate start
    Work from home
    Worldwide
    Flexible hours

    General Dynamics Information Technology

    Bethesda, MD
    4 days ago
  •  ...Reviews, updates, and enhances formal Test & Evaluation (T&E) curriculum, including revising...  ...Acquisition Institute (HSAI), both onsite and in virtual environments. Collaborates...  ...Security Acquisition and Systems Engineering policy Security Clearance DHS Suitability... 
    Policy
    Full time
    Flexible hours

    Joint Research and Development , Inc.

    Washington DC
    4 days ago
  • $35 - $40.5 per hour

     ...full-time contract role and onsite, supporting multiple locations...  ...successfully complete the Train-the-Trainer program for Daily Operations...  ...effectiveness through evaluations, feedback, performance metrics...  ...EEO Statement It is the policy of Computer Aid, Inc.(CAI) not... 
    Policy
    Hourly pay
    Full time
    Contract work
    Apprenticeship
    Work at office
    Local area
    Worldwide

    CAI

    Washington DC
    2 days ago
  •  ...systems engineering (SE) and testing and evaluation (T&E) programs in support of the...  ...Excel, PowerPoint, Word). Ability to work onsite at SPA’s Headquarters in Alexandria, VA...  ...opportunity. It is, and will continue to be, the policy of the company to afford equal... 
    Policy
    Full time
    Contract work
    Work at office
    Local area

    Systems Planning & Analysis

    Alexandria, VA
    4 days ago
  • $79.7k - $113.5k

     ...research, development, test, and evaluation (RDT&E) support. As we...  ...accurate, and aligned with DHS policy. As a Training and Development...  ...independentlyAbility to work onsite in Washington, DC with flexibility...  ...Together with Flexibility model allows you to work 60% in-... 
    Policy
    Full time
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    ClearanceJobs

    Arlington, VA
    20 hours ago
  •  ...at the end of the class. Consult with onsite supervisors regarding available performance...  ...leadership for further action.  Evaluate operational deficiencies, recommend improvements...  ...Verification Program. Compensation Policy: Salary ranges displayed are typical for... 
    Policy
    Work at office
    Flexible hours

    Morgan Business Consulting, LLC

    Arlington, VA
    2 days ago
  • $152.66k - $261.71k

     ...days per week. Responsibilities The Model Risk Management Officer aids the Board of...  ...team to identify potential risks and evaluate the impact of stress events on the bank's...  ...documentation in compliance with internal policies and regulatory guidelines.Collaboration and... 
    Policy
    Full time
    Bank staff
    Remote work
    Flexible hours

    EagleBank

    Bethesda, MD
    6 days ago
  • $30 per hour

     ...Prolific is seeking Fluent Korean Speakers to join their Expert Network and help train AI models. This role requires advanced Korean skills and strong English capabilities. Participants will complete AI training tasks, influence AI development, and enjoy competitive pay... 
    Hourly pay
    Remote work
    Work from home

    Prolific

    Washington DC
    4 days ago
  •  ...SystemServe as an Observer, Controller, Trainer at exercises and training...  ...Homeland Security Exercise and Evaluation Program...  ...Biological Sciences, Engineering, Policy, and Operations. Additionally, CORTEK has provided onsite analytical support for the Department... 
    Policy
    Daily paid
    Full time
    Contract work
    For contractors
    Summer work
    Work at office
    Remote work

    Cortek Inc

    Washington DC
    10 days ago
  •  ...Serve as an Observer, Controller, Trainer at exercises and training...  ...Homeland Security Exercise and Evaluation Program certification· Experience...  ...Biological Sciences, Engineering, Policy, and Operations. Additionally, CORTEK has provided onsite analytical support for the... 
    Policy
    Daily paid
    Full time
    Contract work
    For contractors
    Work at office
    Remote work
    Relocation

    Cortek Inc

    Washington DC
    10 days ago
  •  ...swimmers' improvement. Mentorship: Evaluate swimmers' progress and provide individualized...  ...Employee Excellence: Adhere to club policies, procedures, and safety guidelines....  ...description, and actual requirements may vary. Onsite work only. No relocation offered for... 
    Policy
    Part time
    Relocation
    Flexible hours
    Shift work
    Weekend work
    Afternoon shift

    Gold's Gym Washington

    Washington DC
    10 days ago
  • $100k - $150k

     ...Large Language Model Specialist - Remote    Bright Vision Technologies...  ...company delivering cloud, AI, data, and enterprise...  ...dataset construction, rigorous evaluation methodology, and the engineering...  ...Implement safety, refusal, and policy evaluations to track model... 
    Policy
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Washington DC
    a month ago
  • Accenture Federal Services is seeking an AI Tool Support Engineer to assist with AI-powered coding assistants, model services, and IAM policies. You will support vendor platforms, provide T1/T2 helpdesk assistance, and deliver white-glove application services including... 
    Policy

    JobCubby

    Arlington, VA
    1 day ago
  •  ...community connections and employment leads throughout the work week  ~ Evaluate, monitor, and document program participants’ progress in accordance with funding source regulations and 1Axium’s policies and procedures  ~ Complete and submit daily and/or monthly... 
    Policy
    Self employment
    Local area
    Work from home

    1Axium LLC

    Beltsville, MD
    5 days ago
  •  ...will provide training to customers, requiring frequent travel and onsite work at the customer site. The qualified candidate will partner...  ...of compensation to certain job titles or levels, per internal policy or contractual designation. Additional compensation may be in the... 
    Policy
    Full time
    Temporary work
    Local area
    Relocation package

    KBR

    Lanham, MD
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Model Policy Trainer, Image Evaluation - Seattle Onsite. Be the first to apply!