AI Model Policy Trainer, Generalist (Seattle)
$35 - $100 per hourCacheflow
About Handshake Handshake was founded on a simple belief that everyone deserves a path to a great career, regardless of where they went to school or who they know. Today, we power 25 million job seekers, 1 million+ employers, and 1,600 educational institutions. In 2025, we started Handshake AI and built the fastest-growing AI data business in history. We work directly with frontier AI lab researchers to create evaluations, publish benchmarks, and push the boundary of data. We’ve grown from $0 to ~$1B run rate and pay ~$60M to over 30K individuals every month. Why join Handshake now: Shape how every career evolves in the AI economy, at global scale, with impact your friends, family and peers can see and feel Partner hand-in-hand with world-class AI labs, Fortune 500 partners and the world’s top educational institutions Work together with engineers, scientists, operators, and more from Palantir, Meta, Scale AI, and former YC founders Build a massive, fast-growing business with billions in revenue About Handshake AI Human data is the core infrastructure to AI advancement. Frontier AI labs currently improve model capabilities with various data-intensive post-training techniques. We believe that data spend for AI training will increase by 3-5x in the next few years and continue for much longer as models take on new domains. Handshake AI supports all of the frontier AI labs, working on their most complex data at the largest scale. About the Role As an AI Policy Generalist, you will turn complex customer policies into consistent, well-reasoned evaluations of AI model behavior. You will read user requests, model responses, and relevant conversation history, then determine which policy category best applies. The most interesting cases will not have obvious answers. Two examples may look almost identical until a single word, contextual detail, or difference in intent changes the correct classification. We are looking for people who enjoy splitting hairs in a healthy way. You form clear opinions, explain precisely why two cases should be treated differently, challenge interpretations respectfully, and change your mind when better evidence emerges. You understand that productive disagreement is not about winning an argument. It is how a team finds the most accurate and consistent interpretation. This is not rote annotation. Policies cannot anticipate every possible edge case, and good evaluators do not apply them mechanically. You will balance the policy’s text and intent with customer expectations, conversation context, precedent, and team calibration. The subject matter will vary. One project may involve distinguishing benign assistance from meaningful facilitation of harm. Another may require evaluating whether an interaction reflects ordinary emotional support or unhealthy reliance. A third may focus on nuanced boundaries within sexual-safety policy. Success requires learning each customer’s framework on its own terms rather than carrying assumptions from one domain into another. What You Will Do Learn new customer policies, definitions, taxonomies, and evaluation rubrics quickly Evaluate user requests and AI model responses within the full relevant conversation context Distinguish between closely related labels, severity levels, and policy boundaries Select the most defensible classification when a case is genuinely ambiguous Write concise, evidence-based rationales that cite relevant policy language and conversation details Identify policy gaps, contradictions, unclear definitions, and emerging edge cases Raise thoughtful questions when existing guidance does not resolve a case Participate actively in calibration discussions with evaluators, project leads, policy teams, and researchers Challenge interpretations respectfully and update your judgment when new guidance or stronger reasoning emerges Apply customer policy consistently without substituting personal beliefs for the policy standard Maintain accuracy and attention to detail across repeated evaluations Incorporate feedback quickly and apply clarified guidance to future work Help improve evaluation frameworks, examples, decision rules, and quality standards Move effectively between projects covering different policy domains and customer needs You May Be a Fit If You enjoy making precise distinctions between cases that other people might consider equivalent You notice when one word, contextual detail, or change in intent materially affects the answer You can hold a strong opinion without becoming attached to being right You explain judgment calls clearly enough that another person can audit your reasoning You ask productive questions when a policy is ambiguous instead of guessing or forcing certainty You can separate your personal views from the standard a customer has asked you to apply You are comfortable discussing disagreement directly, respectfully, and without making it personal You can follow the letter of a policy while also understanding its purpose and underlying logic You remain careful and consistent during repetitive, feedback-heavy work You learn unfamiliar subject matter quickly and know when additional context is needed You are intellectually curious, self-directed, and comfortable working in a fast-changing environment You communicate clearly and precisely in writing You treat sensitive information and difficult subject matter with maturity and sound judgment Strong candidates may come from quality assurance, research, editing, law, teaching, operations, trust and safety, content moderation, social science, policy, investigations, compliance, customer support, or other fields that require careful interpretation and defensible decision-making. We care more about how you reason than where you learned to reason. Nice to Have Experience evaluating or comparing outputs from ChatGPT, Claude, Gemini, or other language models in a professional capacity Prior work in AI evaluation, data annotation, RLHF, model quality, trust and safety, policy operations, or content moderation Experience applying detailed rubrics, taxonomies, regulatory language, editorial standards, or quality frameworks Familiarity with calibration sessions, inter-rater agreement, quality audits, or adjudication workflows Experience writing policy guidance, decision trees, evaluation examples, or structured rationales Comfort working with long conversations, incomplete context, and conflicting evidence Familiarity with AI safety, responsible AI, or the ways language models can assist, mislead, or cause harm Prior AI evaluation experience is helpful, but it is not required. Sensitive-Content Notice This role involves regular and deliberate engagement with sensitive material. Depending on the project, evaluations may include sexual content, emotional distress, self-harm, suicide, violence, weapons, abuse, exploitation, discrimination, and other potentially disturbing subjects. The work is conducted within structured evaluation frameworks and professional guidelines. Candidates must be able to engage with this material carefully, responsibly, and sustainably while maintaining sound judgment and consistent work quality. Role Details Location: Seattle, WA Compensation: $35-$100 Employment classification: W-2 Schedule: 8AM - 5PM PT Weekly commitment: M-F California eligibility: We are unable to hire candidates residing in California for this role. #J-18808-Ljbffr Cacheflow
$30 - $85 per hour
...institutions. Handshake AI works directly with frontier... ...benchmarks, and improve AI models through human expertise.... ...About the Role As an AI Model Policy Trainer, Generalist, you will turn complex customer... ...Details Location: Onsite in Seattle, WA, Monday–Friday....PolicyHourly payFull timeMonday to FridayFlexible hours$45 - $65 per hour
...educational institutions. Handshake AI works directly with frontier... ...benchmarks, and improve AI models through human expertise.... ...Location: Onsite in Seattle, WA, Monday-Friday. Compensation... ...About the Role As an AI Model Policy Trainer focused on mental health, you...PolicyHourly payFull timeImmediate startRelocationMonday to FridayFlexible hours$200k - $300k
Member of Technical Staff — Model Optimization and Inference (New Grad) Seattle, Washington About Nuance Labs Nuance Labs... ...photorealistic, real-time AI avatars with emotional intelligence... ...conversations, including eviction policies, compression, and memory-efficient...PolicyInternshipH1bWork at officeVisa sponsorship$300k - $320k
...a Technical Program Manager to lead our AI model evaluation initiatives across multiple workstreams... ...& Safety, Frontier Redteaming, and Policy teams, you will drive high-priority... ...who are comfortable acting as adaptable generalists who add value fast. We excel at maintaining...PolicyWork at officeHome officeVisa sponsorshipRelocation package- ...create reliable, interpretable, and steerable AI systems. We want AI to be safe and... ...group of committed researchers, engineers, policy experts, and business leaders working together... ...the role Anthropic's production models undergo sophisticated post-training processes...PolicyFull timeWork at officeVisa sponsorshipFlexible hours
- Handshake in Seattle, WA is seeking an AI Policy Generalist to turn complex customer policies into consistent evaluations of AI model behavior. You will read user requests, model responses, and conversation history, determine policy category, and craft concise, evidence...Policy
- SkyChefs in Seattle, WA, seeks an HR Generalist to support day-to-day HR operations across onboarding, employee relations, leaves, workers' compensation... ...a high-performing, inclusive work environment, ensuring policy and regulatory compliance while maintaining...Policy
- Neighborcare Health seeks a Senior HR Generalist to administer policies, advocate for employees, and align HR programs with business needs in Seattle. This role balances employee relations with compliance and supports leaders on employment law, pay equity, and talent development...Policy
- ByteDance seeks a Seed Infrastructure Intern in Seattle to contribute to large-scale AI infrastructure, including training platforms and inference systems. You will work with researchers and engineers to optimize system modules and tooling across distributed components....Internship
- Pangleglobal is seeking a Student Researcher in Seattle to conduct research on infrastructure for AI foundation models. This role requires pursuing a PhD in computer science and strong programming skills, focusing on efficiency and reliability in large-scale systems. Interns...Internship
- Distinguished, Data Scientist - Agentic AI Systems Engineering & Model Post-Training This role exists to build those systems end to end—and to improve... ..., resume, replay, or recover safely from failure. Build a policy-first agent runtime and control plane with deterministic...PolicyContract work
$57 per hour
Student Researcher (AI Foundation Model Infrastructure - Seed) - 2027 Start (PhD) Location: Seattle Team: Technology Employment Type: Intern Job Code: A258974A Responsibilities Conduct research on infrastructure and systems for large‑scale models. Explore methods...Hourly payInternshipLocal area- Nuance Labs in Seattle is seeking a Member of Technical Staff for RL Research, aimed at recent PhD graduates in AI or ML. In this impactful role, you will own the RL and post-training for large-scale omni models and contribute to developing advanced AI systems. You will...
$305k
...interpretable, and steerable AI systems. We want AI to be safe... ...committed researchers, engineers, policy experts, and business leaders... ...Manager on Claude Code's model performance team, you will drive... ...solving puzzles San Francisco and Seattle only The annual compensation...PolicyWork at officeVisa sponsorshipFlexible hours- TeachMe.To in Seattle is seeking skilled Soccer Instructors to join our growing peer-to-peer lessons platform. You can set your own schedule, choose your rates, and work with enthusiastic students eager to improve. Deliver personalized coaching, develop drills and plans...
$22 - $38.5 per hour
## Regional Maintenance Trainer HVAC/R IIIApplylocations: Seattle, WA: Los Angeles, CAtime type: Full timeposted on: Posted Todayjob requisition id: R... ...determinable under the terms and conditions of the applicable policies and plans. The amount and availability of any bonus,...PolicyHourly pay$29.68 - $32.91 per hour
Culinary Program Trainer Department: Programs Status: Non-Exempt, Full-Time Location: Seattle, WA. Pay Rate: $29.68-$32.91/hour ABOUT THE ORGANIZATION FareStart has... ...our communities, build business practices and policies, and engage in intentional partnerships and philanthropic...PolicyFull timeWork at officeLocal areaFlexible hoursNight shiftAfternoon shift$15 - $19.52 per hour
...caring, customer service-oriented installer/trainer with a passion for helping people with... ...areas are encouraged to apply: Seattle, WA Great Falls, MT Helena, MT... ...assigned Adhere to strict compliance policies set by the company This position has...PolicyHourly payPart timeFor contractorsWork at officeImmediate startFlexible hoursWeekend workAfternoon shift$177.69k - $416.1k
Machine Learning Engineer-Model Serving Infrastructure Machine Learning Engineer-Model Serving... ...AML team is to push the next-generation AI infrastructure and recommendation platform... ...new Machine Learning Engineer jobs in Seattle, WA . data scientist, Staffing- Data & Analytics...Full timeTemporary workLocal area- Seattle YMCA is hiring for an on-site, part-time Kids Zone role in Seattle. You’ll supervise children, run group activities, and support... ...like Parents Night Outs and Family Nights while upholding YMCA policies and safety standards. The position involves engaging with...PolicyPart timeNight shiftWeekend work
$2,500 - $3,200 per month
...Title: Middle School Boys Soccer Coach Location: University Prep - Seattle, WA 98115 Position Type: Part-Time Temporary Salaried Exempt... .../or Athletic Director Is familiar with and complies with all policies and regulations as put forth in the school’s documents Since safety...PolicyTemporary workPart timeWork at officeLocal area$3,377 - $4,279 per month
...Assistant Coach The Northwest School - Seattle, WA 98122 Overview Salary Range $3,... ...collective enthusiasm and rapport # Model excellent character and live the high expectations... ..., Emerald Sound Conference, and WIAA policies and procedures. Promote and work...PolicyMonday to Friday$24.36 per hour
...character, becoming people of wisdom, and modeling grace-filled community.SPU’s Art and... ...Art & Music Program Facilitator supports Seattle Pacific University’s mission by contributing... ...s Statement of Faith and its derivative policies and lifestyle expectations.What You’ll...PolicyFull timePart timeCasual workWork at officeImmediate startAfternoon shift$40 - $70 per unit
BODYROK Seattle is hiring fitness instructors to lead high-energy, music-driven sessions in our expanding Seattle market. You will teach the BODYROK Hybrid Pilates method, motivate clients, and build a thriving class community across our local studios in Belltown and Queen...Local area$4,000 per month
...duties of this position must be conducted in adherence with the policies, rules and regulations of the Northwest Athletic Conference (... ...diverse community. Bellevue College is located just 10 miles east of Seattle where we serve a student population of over 54% students of...PolicyContract workPart timeWork experience placementWork at officeRelocation packageFlexible hours$30 - $75 per hour
...to teach at a high profile company's fitness center located in Seattle, WA. Responsibilities: Develop effective, engaging, and motivating... ...statement Plus One, an Optum company, is committed to a policy of non-discrimination and equal employment opportunity. Qualified...PolicyHourly payMinimum wageFull timeTemporary workWork experience placementLocal area$60 - $80 per unit
...who want to do meaningful work in an environment that expects more and delivers more. Job Description Equinox's locations in Seattle are seeking Studio Cycling Instructors. The position entails the following responsibilities. This is not a complete description, and...Full time- Securitas Security Services USA, Inc. is seeking a training specialist in Seattle, WA. The role involves training employees on compliance and safety practices, developing training schedules, and maintaining records. Candidates should have an Associate degree, one year of...Work at office
$27.5 per hour
...Rooms & Guest Services Operations Location 1400 6th Ave, SEATTLE WA 98101, United States VIEW ON MAP ( Schedule Full Time... ...injuries, and unsafe work conditions to manager. Follow all company policies and procedures, ensure uniform and personal appearance are clean...PolicyHourly payFull timeWork experience placementRemote workFlexible hoursShift work- Summary This on-figure female model will pose for non-recognizable/sell-shot photography used on the Nordstrom website, working in a fast-paced photo studio environment. The role requires adherence to precise body measurements, punctuality, and collaboration with a photography...Flexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Model Policy Trainer, Generalist (Seattle). Be the first to apply!
- ai trainer Seattle, WA
- policy specialist Seattle, WA
- education policy research Seattle, WA
- trade policy jobs Seattle, WA
- vice president public policy Seattle, WA
- cybersecurity policy and compliance analyst Seattle, WA
- public policy Seattle, WA
- policy counsel Seattle, WA
- immigration policy Seattle, WA
- policy assistant Seattle, WA


