Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Frontier AI Safety Evaluator & Policy Expert

Obsidian

Obsidian is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across complex policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback. You will review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains, and collaborate with researchers and safety teams to #J-18808-Ljbffr Obsidian

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Frontier AI Safety Evaluator & Policy Expert in San Francisco, CA vacancy
  • We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve... 
    Policy
    Worldwide

    Obsidian

    San Francisco, CA
    3 days ago
  • $60 - $70 per hour

     ...technical talent with leading AI research labs. Headquartered in...  ...Jack Dorsey . Position: AI Safety Practitioner Type: Contract...  ...Role Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality... 
    Policy
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    5 days ago
  • We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing...  ..., uncover model weaknesses, and evaluate AI behavior across complex, high...  ...behaviours, hallucinations, and policy failures. Evaluate model... 
    Policy

    Obsidian

    San Francisco, CA
    1 day ago
  • $70 - $84 per hour

     ...technical talent with leading AI research labs. Headquartered...  ...Jack Dorsey . Position AI Safety Red Teamer Type Contract Compensation...  ...prompts to stress-test frontier AI models. Identify...  ...behaviors, hallucinations, and policy failures. Evaluate model robustness across misinformation... 
    Policy
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    2 days ago
  • Mercor is seeking experienced AI Safety Practitioners to assess the safety, quality, and alignment of frontier AI models across complex, policy-sensitive topics. You will evaluate AI-generated responses, apply safety policies, and provide structured feedback to improve... 
    Policy

    Mercor

    San Francisco, CA
    3 days ago
  •  ...is seeking a Product Marketing Manager, Research & Safety to shape how the world understands frontier AI and our approach to responsible development. You will...  ...enterprise audiences. You will work with Research, Safety, Policy, Communications, and Product teams to lead go-to-... 
    Policy

    OpenAI

    San Francisco, CA
    2 days ago
  •  ...Team Our Cyber team builds AI systems and products that...  ...while improving the safety and reliability of frontier models in security-sensitive...  ...engineering, model training, evaluations, safeguards, and deployment...  ...Employment Opportunity Policy Statement . Background... 
    Policy
    Full time

    OpenAI

    San Francisco, CA
    10 hours ago
  • Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback. Candidates should have at least 5 years of relevant experience... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    San Francisco, CA
    4 days ago
  •  ...we believe the safest AI is the one that’s already...  ...project - human data experts who probe AI models with...  ...strengthen customer AI systems Evaluation coverage expands: more...  ...customers trust the safety of their AI because you...  ...AI red teaming at the frontier of safety Play a direct... 
    Remote work

    Obsidian

    San Francisco, CA
    4 days ago
  • Obsidian is seeking expert Evaluators in Biology/environmental science to review and assess AI-generated work products for accuracy and quality. In this remote, hourly role, you will leverage your expertise to provide feedback on documents and presentations, ensuring they... 
    Remote job
    Hourly pay

    Obsidian

    San Francisco, CA
    10 hours ago
  • AfterQuery is seeking a Growth Associate in San Francisco to architect and scale the expert supply engine driving frontier AI training. You will own both manual and automated pipelines to grow the expert network from 100K to hundreds of thousands, working with a high-velocity... 

    David Joseph

    San Francisco, CA
    10 hours ago
  • $80 - $120 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Position: Government / public administration Evaluator Type: Contract Compensation: $...  .... Collaborate with subject matter experts to ensure consistency and quality. Work... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    San Francisco, CA
    3 days ago
  • $90 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...genuinely correct replications from those that merely look correct. Evaluate responsive behavior and semantic quality, ensuring proper use of... 
    Contract work
    Summer work
    Local area
    Remote work

    Mercor

    San Francisco, CA
    8 days ago
  • Welo Data is seeking Data Labeling Associates in California to evaluate AI outputs and ensure cultural context and safety in Arabic datasets. This role requires professional-level proficiency in Portuguese (Brazil), a bachelor's degree, and at least 2 years of experience... 

    Welo Data

    San Francisco, CA
    1 day ago
  • Synthires is offering a part-time role for PhD-level Chemistry experts to contribute to AI safety and evaluation projects. The work involves applying scientific expertise to understand and improve how AI systems handle specialized chemistry topics, with training provided... 
    Remote job
    Part time

    Synthires

    San Francisco, CA
    3 days ago
  • $320k

     ...interpretable, and steerable AI systems. We want AI to...  ..., engineers, policy experts, and business leaders...  ...systems. About the Team The Frontier Red Team (FRT) is a...  ...researching and ensuring safety with self-improving,...  ...to elicit and evaluate autonomous AI cyber capabilities... 
    Policy
    Work at office
    Relocation
    Visa sponsorship
    Flexible hours

    Neura Market

    San Francisco, CA
    4 days ago
  •  ...Francisco is hiring Data Labeling Associates for Project Perseus. This role focuses on evaluating Arabic AI systems, requiring professional proficiency in Portuguese and experience in AI safety. Responsibilities include assessing AI outputs, identifying bias, and... 
    Full time

    Welo Data

    San Francisco, CA
    10 hours ago
  • $80 - $150 per hour

     ...are partnering with a leading AI research organisation to develop...  ...-informed benchmark for evaluating how AI companion chatbots respond...  ...standards for this important AI safety initiative. Responsibilities...  ...generated conversations and provide expert judgment to calibrate... 
    Hourly pay
    Traineeship
    10 hours per week

    Obsidian

    San Francisco, CA
    1 day ago
  • Obsidian is seeking clinicians and researchers to help design a clinician-informed benchmark evaluating how AI chatbots respond to adolescents facing mental health challenges. You will author realistic case scenarios, review clinical realism and ethics, and help calibrate... 

    Obsidian

    San Francisco, CA
    1 day ago
  • YO AI Labs is seeking a Pharmacovigilance Expert to contribute drug safety expertise to a healthcare AI project. You will review pharmacovigilance documentation, safety...  ...The role emphasizes data quality, analytical evaluation of DSURs/PSURs, and compliance with ICH E2F,... 
    Remote job
    For contractors
    Flexible hours

    YO AI Labs

    San Francisco, CA
    4 days ago
  • $29 - $45 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type: Contract Compensation:... 
    Contract work
    Summer work
    Remote work

    Mercor Inc

    San Francisco, CA
    4 days ago
  • $48 - $62 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Dutch Type: Contract Compensation: $48–$... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    9 days ago
  • $48 - $62 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Finnish Type: Contract Compensation: $48... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    15 days ago
  •  ...Data is seeking Data Labeling Associates in San Francisco to evaluate AI systems focused on Arabic language nuances. The role includes...  ...like gourmet dining and comprehensive medical coverage while contributing to innovative AI safety solutions. #J-18808-Ljbffr Welo Data

    Welo Data

    San Francisco, CA
    3 days ago
  •  ...Preparedness is a critical Safety Research team at the...  ...focused on mitigating AI threats to global security...  ...capabilities of frontier AI systems. Mitigation...  ...new model capabilities. Evaluate technical trade-offs within...  ...fine‑tuning, and policy optimization. Excel at... 
    Policy

    United States Digital Space LLC

    San Francisco, CA
    1 day ago
  •  ...Preparedness is a critical Safety Research team at the...  ...focused on mitigating AI threats to global security...  ...capabilities of frontier AI systems. Mitigation....  ...preparedness capability evaluations—designing new evals grounded...  ...Employment Opportunity Policy Statement. Background... 
    Policy
    Permanent employment
    Temporary work

    United States Digital Space LLC

    San Francisco, CA
    1 day ago
  •  ...About the Team The Frontier Systems team at OpenAI builds, launches...  ...About OpenAI OpenAI is an AI research and deployment company...  ...tool that must be created with safety and human needs at its core, and...  ...Equal Employment Opportunity Policy Statement . Background... 
    Policy
    Full time

    OpenAI

    San Francisco, CA
    10 hours ago
  • $128.5k - $171.5k

    A data-focused AI engineering role inside the Coro unit, delivering...  ..., prioritizing reliability, safety, and clear failure modes...  ...Create reproducible training and evaluation pipelines with versioning, experiment...  ...loops for prompt and policy optimization Design for secure... 
    Policy

    Bain & Co.

    San Francisco, CA
    1 day ago
  • OpenAI is seeking a Researcher for Frontier Cybersecurity Risks to design and implement an...  ...and close collaboration with risk, policy, product, and engineering teams to ensure...  ...enabled safeguards across various surfaces, evaluate trade-offs, and lead testing and red-... 
    Policy

    Triwill Group

    San Francisco, CA
    10 hours ago
  •  ...intraoperative needs under physician orders. You will document perioperative care, coordinate with surgeons and anesthesiologists, and perform certain procedures outside the OR according to KP policy, maintaining patient safety and compliance. #J-18808-Ljbffr Kaiser Permanente
    Policy

    Kaiser Permanente

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Frontier AI Safety Evaluator & Policy Expert. Be the first to apply!