AI Model Policy Trainer, Content Risk
Jobtailor
Evaluate user requests and AI model responses involving violence, weapons, threats, and dark fiction in full conversation context Distinguish fictional, educational, historical, and defensive violence from requests seeking real-world uplift or expressing real intent to harm Assess whether model responses provide meaningful real-world capability Distinguish anger, frustration, and dark humor from credible threats or crisis indicators Select defensible classifications for ambiguous cases and write concise policy-based rationales Write and refine adversarial or borderline prompts Identify policy gaps, contradictions, and emerging edge cases and raise them with project leads and policy teams Participate in calibration discussions and update judgments based on stronger reasoning Apply customer policy consistently without substituting personal beliefs Maintain accuracy and attention to detail across repeated evaluations involving graphic material Requirements Strong judgment about violence in fiction, the real world, or among people in distress Ability to distinguish fictional, educational, historical, and defensive violence from real-world uplift or intent to harm Ability to assess meaningful real-world capability in AI model responses Ability to distinguish anger, frustration, or dark humor from credible threats or crisis indicators Ability to write concise rationales citing policy language and conversation details Ability to write and refine adversarial or borderline prompts Ability to identify policy gaps, contradictions, and emerging edge cases Ability to participate in calibration discussions and apply customer policy consistently Accuracy and attention to detail during repetitive, feedback-heavy evaluations Clear and precise written communication Ability to engage carefully, responsibly, and sustainably with graphic material A degree, a clearance, and a technical background are not required Must be authorized to work lawfully in the United States for Handshake Core Competencies Demonstrates strong judgment and analytical skills in evaluating AI model responses related to violence and threats, while maintaining accuracy and attention to detail. Capable of writing concise rationales and engaging with graphic material responsibly. Highest-signal resume keywords Strong Judgment About Violence Ability To Distinguish Fictional And Real-World Violence Ability To Write Concise Rationales Ability To Identify Policy Gaps Clear And Precise Written Communication Soft Skills Attention To Detail Analytical Skills Engagement With Graphic Material Industry Keywords AI Model Evaluation Policy-Based Rationales Calibration Discussions Crisis Indicators Adversarial Prompts #J-18808-Ljbffr Jobtailor
- ..., we started Handshake AI and built the fastest-growing... ...labs currently improve model capabilities with... ...the Role As an AI Model Policy Trainer focused on mental... ...grounded, and sensitive to risk without reinforcing unsupported... ...distressing content Who May Be a Good Fit...PolicyRiskContentImmediate startRelocation
$166k - $258k
....Nordstrom Technology is moving to an AI Native operating model, and we're rebuilding the processes, tools... ...authority who tailors style and content to different audiences — moving comfortably... ..., such as moving faster or reducing risk, into concrete, trackable measures...RiskContentFull time- ...focus on the broad spectrum of Risk Management, Compliance,... ...intersections of assets, processes, policies and people delivering value. ProSidian... ...researched practice | deep content and process expertise |... ...making, and other communication models and tools as part of the agenda...PolicyRiskContentFull timeTemporary workFor contractorsWork at officeLocal areaRemote workFlexible hoursShift work
- ...rationale Classify images under customer policies covering sexual content, violence, hate symbols, real-person... ...Write and refine prompts to test model quality and safety boundaries Identify... ...technical background are not required Prior AI evaluation experience is helpful but...PolicyContentMonday to Friday
$55 - $75 per hour
...In 2025, we started Handshake AI and built the fastest-growing... ...Frontier AI labs currently improve model capabilities with various data... .... About the Role As an AI Policy Generalist, you will turn... ...operations, trust and safety, content moderation, social science, policy...PolicyContent$184.87k - $324.19k
...SAP BTP, Datasphere, SAC, and AI/ML for real-time reporting, predictive... ...; strong track record in risk mitigation, issue resolution,... ...and thought leadership content at publication quality Travel... ...abilities to adhere to company policies, exercise sound judgment, effectively...PolicyRiskContentFull timeLocal area$45 - $55 per hour
...institutions. In 2025, we started Handshake AI and built the fastest-growing AI data... .... Frontier AI labs currently improve model capabilities with various data-intensive... ...why? Does the image violate the customer's content policy, and if so, which category and how severely...PolicyContent- Jobtailor seeks a policy evaluation professional to assess AI model responses involving violence and related content. You will determine classifications in context, distinguishing fiction, education, history, and defensive uses from real-world uplift or harm. Responsibilities...PolicyRiskContent
- ...businesses with the most advanced AI technology:- Combat any kinds of risks/violations issues in E-commerce scenarios... ...detect and control risks/frauds in contents/products/sellers/creators•... ...with strategy team, product managers, policy team and ops team to help define products...PolicyRiskContentWork experience placement
- ...Algorithm team. You will contribute to building advanced AI algorithms that detect and mitigate risks in e-commerce content, products, sellers, and creators. This internship... ...-on learning with collaboration across product, policy, and operations groups. The role is ideal for...PolicyRiskContentInternship
- ...data platforms, and the AI systems that depend on them... ...into reduced risk, controlled access, and... ...threats, cyber-attacks, and policy violationsHelp customers... ...appropriateCreate knowledge base content to capture new learning... ..., identity and access models (EntraID/IAM), and...PolicyRiskContent
- ...Experience Algorithm Team, the AI guardians ensuring the... ...beyond traditional risk control. We are... ...a prosperous, trusted content ecosystem and maintaining... ...RAG, GNN, and Sequence Modeling to solve complex governance... ...interpret complex governance policies. Develop agents that...PolicyRiskContent
$139.8k - $164.5k
...clients modernize governance, risk, and compliance capabilities on... ...oversight, control assurance, policy compliance, audit readiness,... ...management of delivery risks.Use AI and automation tools to... ...integrating external regulatory content, risk data, control libraries,...PolicyRiskContentLocal areaImmediate startFlexible hours- ...Expert (SME) to review AI-generated SEO/marketing... ...expert digital marketing content, evaluating reasoning quality... ..., and compliance/claims risk; fact‑check where needed... ...explanations and model outputs that demonstrate... ...inaccuracies, brand‑risk claims, policy/compliance pitfalls, and...PolicyRiskContentHourly payContract workPart timeFor contractorsImmediate startRemote work
$173k - $259k
...deploy machine learning models that power core... ...driven featuresUtilize AI tools and high velocity... ...bottlenecks, and security risks.Adaptability in learning... ...recommendations, search, content understanding, image generation... ...."Default Together" Policy at Snap: At Snap Inc....PolicyRiskContentFull timeLive inWork at officeLocal area$30 per hour
About mpathic.ai Keeping the human in AI. mpathic is... ...expert annotation, and model evaluation across high-stakes... ...to identify risks related to financial guidance... ...AI‑generated financial content and conversations for accuracy... ...drift, escalation, and policy breakdown over time...PolicyRiskContentPart timeRemote work10 hours per week$169k - $338k
...of program health, technical risks, and strategic trade-offs.Mentor... ..., cloud infrastructure, AI/ML pipelines, or equivalent domains... ...knowledge in implementing Web Content Accessibility Guidelines (WCAG... ...workplace and has a no tolerance policy regarding the use of illegal...PolicyRiskContentFull timeTemporary workPart time$100k - $180k
...participate in security operations, risk assessments, incident... ...of the company. A passion for AI is a must. Responsibilities... ...relates to your role Adhere to policies, guidelines and procedures... ...authentication systems, log management, content filtering, etc Ability to...PolicyRiskContentWork experience placement- ...Artificial Intelligence, risk management, big data,... ...OS behavior, capacity modeling, operational characteristics... ...platforms (e.g., AI/ML workloads) with a focus... ...competitive vacation policies based on employee level... ...financial education and content on a variety of topics...PolicyRiskContentFull timeWork at office
- About mpathic.ai mpathic is keeping humans safe in... ...behavior, communication, policy, education, healthcare,... ...matter. You'll help identify model strengths and weaknesses... ...Identifying emerging risks, behavioral patterns,... ...AI systems and sensitive content Participating in calibration...PolicyRiskContentTemporary work
$141.9k - $190.3k
..., news, and entertainment content, across all media platforms... ...help build services and models enabling efficient ad fills... ...system health and to manage risk.Kindness and pragmatic... ...and using automated tools (AI) while adhering to company policy.Available for On-Call rotations...PolicyRiskContentWork experience placementWork at office- ...scale, and operate Web, API and AI applications globally across... ...maker and business decision maker content, envision a solution demo of... ...-marketIndependently identify risks, drive solutions, and... ...Employment OpportunityIt is the policy of F5 to provide equal employment...PolicyRiskContentFull timeWork at officeLocal areaWeekend workAfternoon shift
$144.8k - $261.45k
...automation, orchestration, and agentic AI systems that reduce mean-time-... ...incidents.Create and maintain policy-enforcement automation,... ...problems, making defensible, risk-based decisions on containment... ...and CI/CD for security content.Experience mentoring other engineers...PolicyRiskContentFull timeTemporary workLocal areaWorldwide$38 per hour
...closely with both people and AI systems — auditing... ...and blockers — escalating risks and gaps to Team... ...cases in both human and model outputsParticipate in calibrations... ...in data annotation, content quality, QA, or related... ...safety, compliance, or policy-driven content evaluationBenefitsPaid...PolicyRiskContentFull timeContract workRemote workVisa sponsorship$175.1k - $236.9k
...technologies, including:- AI-powered anomaly... ...enforcement of access protection policies at scale.- Partner... ...identify gaps, reduce risk, and raise the security... ...of web services, video content protection technologies... ...Experience applying threat modeling or other risk...PolicyRiskContentRemote workFlexible hours$132.1k - $178.8k
...and can enjoy even more content for free with ads.Are... ...new possibilities, take risks, and collaborate with remarkable... ...also contribute to AI enablement efforts,... ...alignment with company policies and regulations- Using... ...- Experience with data modeling, warehousing and building...PolicyRiskContentFlexible hours$188.7k
...analytics, intelligent AI agents, and core product... ...interactions - content engagement, CRM activity... ...need flexible dimensional models. AI agents need fresh,... ...designed compute governance policies, or re-architected... ...screening. To reduce the risk of bias, identifying details...PolicyRiskContentFull timeImmediate startFlexible hours$149.9k - $202.8k
...intersection of data engineering and AI — building systems where... ...pipelines, ownership models, and content lifecycle practices — especially... ...platform — framing the cost, risk, and timeline clearly.About the... ..., and local laws and Company policies. Criminal history may have a...PolicyRiskContentLocal areaFlexible hoursDay shift$193.7k - $262k
...a number of foundation models, you will be working directly... ...requirements, creating content to train the field... ...in Generative AI and key ML powered use... ...objective business metrics and risks.Amazon is an equal opportunity... ...local laws and Company policies. Criminal history may...PolicyRiskContentLocal areaWorldwideFlexible hours$156k - $312k
...management, and architectural risk mitigation across CSI-... ...the organization. Model Walmart’s values and ethical... .... Knowledge of current AI tools for engineering... ...knowledge in implementing Web Content Accessibility Guidelines... ...and has a no tolerance policy regarding the use of...PolicyRiskContentFull timeTemporary workPart time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Model Policy Trainer, Content Risk. Be the first to apply!




