Frontier AI Safety Evaluator & Policy Expert
Obsidian
Obsidian is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across complex policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback. You will review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains, and collaborate with researchers and safety teams to #J-18808-Ljbffr Obsidian
- We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve...PolicyWorldwide
$60 - $70 per hour
...technical talent with leading AI research labs. Headquartered in... ...Jack Dorsey . Position: AI Safety Practitioner Type: Contract... ...Role Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality...PolicyContract workSummer workRemote work- We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing... ..., uncover model weaknesses, and evaluate AI behavior across complex, high... ...behaviours, hallucinations, and policy failures. Evaluate model...Policy
$70 - $84 per hour
...technical talent with leading AI research labs. Headquartered... ...Jack Dorsey . Position AI Safety Red Teamer Type Contract Compensation... ...prompts to stress-test frontier AI models. Identify... ...behaviors, hallucinations, and policy failures. Evaluate model robustness across misinformation...PolicyContract workSummer workRemote work- Mercor is seeking experienced AI Safety Practitioners to assess the safety, quality, and alignment of frontier AI models across complex, policy-sensitive topics. You will evaluate AI-generated responses, apply safety policies, and provide structured feedback to improve...Policy
- ...is seeking a Product Marketing Manager, Research & Safety to shape how the world understands frontier AI and our approach to responsible development. You will... ...enterprise audiences. You will work with Research, Safety, Policy, Communications, and Product teams to lead go-to-...Policy
- ...Team Our Cyber team builds AI systems and products that... ...while improving the safety and reliability of frontier models in security-sensitive... ...engineering, model training, evaluations, safeguards, and deployment... ...Employment Opportunity Policy Statement . Background...PolicyFull time
- Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback. Candidates should have at least 5 years of relevant experience...Remote jobHourly payWork at office
- ...we believe the safest AI is the one that’s already... ...project - human data experts who probe AI models with... ...strengthen customer AI systems Evaluation coverage expands: more... ...customers trust the safety of their AI because you... ...AI red teaming at the frontier of safety Play a direct...Remote work
- Obsidian is seeking expert Evaluators in Biology/environmental science to review and assess AI-generated work products for accuracy and quality. In this remote, hourly role, you will leverage your expertise to provide feedback on documents and presentations, ensuring they...Remote jobHourly pay
- AfterQuery is seeking a Growth Associate in San Francisco to architect and scale the expert supply engine driving frontier AI training. You will own both manual and automated pipelines to grow the expert network from 100K to hundreds of thousands, working with a high-velocity...
$80 - $120 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Position: Government / public administration Evaluator Type: Contract Compensation: $... .... Collaborate with subject matter experts to ensure consistency and quality. Work...Contract workSummer workWork at officeRemote work$90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...genuinely correct replications from those that merely look correct. Evaluate responsive behavior and semantic quality, ensuring proper use of...Contract workSummer workLocal areaRemote work- Welo Data is seeking Data Labeling Associates in California to evaluate AI outputs and ensure cultural context and safety in Arabic datasets. This role requires professional-level proficiency in Portuguese (Brazil), a bachelor's degree, and at least 2 years of experience...
- Synthires is offering a part-time role for PhD-level Chemistry experts to contribute to AI safety and evaluation projects. The work involves applying scientific expertise to understand and improve how AI systems handle specialized chemistry topics, with training provided...Remote jobPart time
$320k
...interpretable, and steerable AI systems. We want AI to... ..., engineers, policy experts, and business leaders... ...systems. About the Team The Frontier Red Team (FRT) is a... ...researching and ensuring safety with self-improving,... ...to elicit and evaluate autonomous AI cyber capabilities...PolicyWork at officeRelocationVisa sponsorshipFlexible hours- ...Francisco is hiring Data Labeling Associates for Project Perseus. This role focuses on evaluating Arabic AI systems, requiring professional proficiency in Portuguese and experience in AI safety. Responsibilities include assessing AI outputs, identifying bias, and...Full time
$80 - $150 per hour
...are partnering with a leading AI research organisation to develop... ...-informed benchmark for evaluating how AI companion chatbots respond... ...standards for this important AI safety initiative. Responsibilities... ...generated conversations and provide expert judgment to calibrate...Hourly payTraineeship10 hours per week- Obsidian is seeking clinicians and researchers to help design a clinician-informed benchmark evaluating how AI chatbots respond to adolescents facing mental health challenges. You will author realistic case scenarios, review clinical realism and ethics, and help calibrate...
- YO AI Labs is seeking a Pharmacovigilance Expert to contribute drug safety expertise to a healthcare AI project. You will review pharmacovigilance documentation, safety... ...The role emphasizes data quality, analytical evaluation of DSURs/PSURs, and compliance with ICH E2F,...Remote jobFor contractorsFlexible hours
$29 - $45 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type: Contract Compensation:...Contract workSummer workRemote work$48 - $62 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Dutch Type: Contract Compensation: $48–$...Contract workSummer workRemote work$48 - $62 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Finnish Type: Contract Compensation: $48...Contract workSummer workRemote work- ...Data is seeking Data Labeling Associates in San Francisco to evaluate AI systems focused on Arabic language nuances. The role includes... ...like gourmet dining and comprehensive medical coverage while contributing to innovative AI safety solutions. #J-18808-Ljbffr Welo Data
- ...Preparedness is a critical Safety Research team at the... ...focused on mitigating AI threats to global security... ...capabilities of frontier AI systems. Mitigation... ...new model capabilities. Evaluate technical trade-offs within... ...fine‑tuning, and policy optimization. Excel at...Policy
- ...Preparedness is a critical Safety Research team at the... ...focused on mitigating AI threats to global security... ...capabilities of frontier AI systems. Mitigation.... ...preparedness capability evaluations—designing new evals grounded... ...Employment Opportunity Policy Statement. Background...PolicyPermanent employmentTemporary work
- ...About the Team The Frontier Systems team at OpenAI builds, launches... ...About OpenAI OpenAI is an AI research and deployment company... ...tool that must be created with safety and human needs at its core, and... ...Equal Employment Opportunity Policy Statement . Background...PolicyFull time
$128.5k - $171.5k
A data-focused AI engineering role inside the Coro unit, delivering... ..., prioritizing reliability, safety, and clear failure modes... ...Create reproducible training and evaluation pipelines with versioning, experiment... ...loops for prompt and policy optimization Design for secure...Policy- OpenAI is seeking a Researcher for Frontier Cybersecurity Risks to design and implement an... ...and close collaboration with risk, policy, product, and engineering teams to ensure... ...enabled safeguards across various surfaces, evaluate trade-offs, and lead testing and red-...Policy
- ...intraoperative needs under physician orders. You will document perioperative care, coordinate with surgeons and anesthesiologists, and perform certain procedures outside the OR according to KP policy, maintaining patient safety and compliance. #J-18808-Ljbffr Kaiser PermanentePolicy
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Frontier AI Safety Evaluator & Policy Expert. Be the first to apply!
- work from home web search evaluator San Francisco, CA
- social media evaluator San Francisco, CA
- evaluator San Francisco, CA
- quality evaluator San Francisco, CA
- education evaluator San Francisco, CA
- ai evaluator San Francisco, CA
- technology expert San Francisco, CA
- subject matter expert San Francisco, CA
- fulfillment expert San Francisco, CA
- guest service support expert San Francisco, CA


