Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Safety Specialist - Evaluation Expert

Mercor

We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback. Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality. Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains. Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking. Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. Provide structured feedback to improve model alignment and safety performance. Collaborate with AI researchers and safety teams on ongoing evaluation initiatives. Required Qualifications Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline. 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field. Excellent written English, critical thinking, and analytical reasoning skills. Ability to consistently evaluate nuanced and policy-sensitive scenarios. Preferred Qualifications Experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation. Familiarity with safety policies, content moderation, or evaluation rubric development. Experience reviewing complex, high-risk, or ambiguous content. Why Join? Shape the safety and behaviour of frontier AI models used by millions worldwide. Work on challenging, real-world safety evaluations across nuanced and high-impact domains. Collaborate with leading AI researchers, engineers, and safety teams. #J-18808-Ljbffr Mercor

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the AI Safety Specialist - Evaluation Expert in San Francisco, CA vacancy
  • $60 - $70 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ..., and Jack Dorsey . Position: AI Safety Practitioner Type: Contract Compensation...  ...: Remote Role Responsibilities Evaluate AI-generated responses for safety,... 
    Suggested
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    5 days ago
  • Welo Data in San Francisco seeks a full-time AI Evaluator with professional proficiency in Portuguese (Portugal) and experience in Generative AI safety. The role involves critiquing AI outputs, identifying biases, and refining evaluation frameworks. Candidates should possess... 
    Suggested
    Full time

    Welo Data

    San Francisco, CA
    4 days ago
  • Welo Data is looking for a Data Labeling Associate in San Francisco to evaluate AI systems' handling of Arabic nuances. The role requires professional-level proficiency in Arabic and 2 years of AI safety experience. Responsibilities include critiquing Arabic AI outputs... 
    Suggested

    Welo Data

    San Francisco, CA
    2 days ago
  • $80 - $120 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  .... Position: User/customer research and feedback synthesis Evaluator Type: Contract Compensation: $80–$120/hour Location... 
    Suggested
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    San Francisco, CA
    15 days ago
  • $80 - $120 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Position: Government / public administration Evaluator Type: Contract Compensation: $...  .... Collaborate with subject matter experts to ensure consistency and quality. Work... 
    Suggested
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    San Francisco, CA
    10 days ago
  • $90 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...genuinely correct replications from those that merely look correct. Evaluate responsive behavior and semantic quality, ensuring proper use of... 
    Contract work
    Summer work
    Local area
    Remote work

    Mercor

    San Francisco, CA
    15 days ago
  • OpenAI is seeking a researcher to advance frontier evaluations and environments for safe AGI/ASI. You will help design north star model...  ...products. Collaborate with researchers, engineers, product and safety teams to decide what to measure, how to measure it, and how to... 

    Neura Market

    San Francisco, CA
    3 days ago
  •  ...setting. You will tackle fundamental challenges in building safe AI agents and aligning them with humans, including benchmarking...  ...DPO, or GRPO, plus strong cross-functional communication. Some familiarity with agent evaluation tooling is a plus. #J-18808-Ljbffr Scale

    Scale

    San Francisco, CA
    4 days ago
  •  ...seeking exceptional research engineers to push the boundaries of frontier AI safety, shaping empirical understanding of risk and owning end-to-end threads within this effort. You’ll design evaluations of frontier models against real threat models, develop datasets,... 

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • $70 - $90 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Dorsey . Position: Trainium (NKI) Kernel Expert Type: Contract Compensation: $...  ...: Remote Role Responsibilities Evaluate the quality, correctness, and hardware-appropriateness... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    2 days ago
  • Evaluate the quality, correctness, and hardware-appropriateness of Neuron Kernel Interface (NKI) development tasks used to train and evaluate a frontier AI lab's models. You'll assess CUDA→NKI migration fidelity, Trainium-specific performance-optimization quality, and cross... 

    Mercor

    San Francisco, CA
    2 days ago
  • $293.5k

    General Information Job Title Expert Senior Manager, AI Engineering Job ID 104335 Work...  ...research, model experimentation, and evaluation design to production system...  ...continual improvementBalance performance, safety, responsible AI principles, and cost... 
    Permanent employment
    Full time
    Apprenticeship
    Work at office
    Local area
    Work from home
    Home office
    3 days per week

    Bain & Company

    San Francisco, CA
    2 days ago
  • $50 - $60 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Summers , and Jack Dorsey . Position: Expert Interviewer Type: Contract Compensation...  ...technical depth, communication, and AI-evaluation fit using a provided rubric. Confirm... 
    Contract work
    Summer work
    Remote work
    Weekday work

    Mercor

    San Francisco, CA
    2 days ago
  • $20 - $26 per hour

    Prolific is seeking fluent Kannada speakers to act as evaluators who compare text and voice samples to assess naturalness and authenticity. You will listen to audio clips, rate quality, and flag any mismatches in tone or pronunciation, with emphasis on cultural context... 
    Remote job
    Flexible hours

    Prolific

    San Francisco, CA
    3 days ago
  • $1,750 - $2,150 per month

    Obsidian is looking for experienced cybersecurity professionals to review AI systems' threat detection and vulnerability assessments. Responsibilities include evaluating AI outputs and creating realistic cybersecurity scenarios. Ideal candidates should have over 3 years... 

    Obsidian

    San Francisco, CA
    3 days ago
  • $35 per hour

     ...Mercor, we believe the safest AI is the one that’s already been...  ...for this project - human data experts who probe AI models with adversarial...  ...customer AI systems Evaluation coverage expands: more scenarios...  ...production Mercor customers trust the safety of their AI because you’ve... 
    Remote job

    Obsidian

    San Francisco, CA
    4 days ago
  • $70 - $84 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ..., and Jack Dorsey . Position: AI Safety Red Teamer Type: Contract Compensation...  ...hallucinations, and policy failures. Evaluate model robustness across misinformation, cyber... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    9 days ago
  • YO AI Labs in the United States (Remote) seeks seasoned Adobe Marketing Technology Experts to support an AI training and evaluation project focused on enterprise marketing operations. You will test workflows with Adobe Workfront, AEM, CJA, Analytics and Experience Cloud... 
    Remote job

    YO AI Labs

    San Francisco, CA
    1 day ago
  • $80 - $150 per hour

     ...leading behavioral health consulting firm is seeking Senior Behavioral Health Experts to work part-time and remotely on frontier AI research projects. You will be responsible for designing evaluations and testing AI systems in critical mental health contexts. The ideal... 
    Hourly pay
    Part time
    Remote work

    Aligned Labs

    San Francisco, CA
    2 days ago
  • We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area"... 

    Obsidian

    San Francisco, CA
    3 days ago
  • Mercor is hiring experienced musicians to evaluate generative music AI models, in partnership with a leading AI lab. You will assess AI-generated...  ...Duration: up to 6 months Commitment: flexible. Most experts work around 20 hours per week; there is no cap Application... 
    Immediate start
    Flexible hours

    Obsidian

    San Francisco, CA
    4 days ago
  • Mercor is seeking experienced musicians to evaluate generative music AI models. You will compare AI-generated lyrics with published songs across genres and rate them against detailed quality standards in Greek and English. The role requires native or near-native Greek,... 
    Remote work
    Flexible hours

    Obsidian

    San Francisco, CA
    4 days ago
  • Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) Mercor is partnering with leading AI labs on a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve... 
    Part time
    Immediate start

    Mercor

    San Francisco, CA
    20 hours ago
  • $128.5k - $171.5k

    A data-focused AI engineering role inside the Coro unit, delivering GenAI powered tools...  ...loop controls, prioritizing reliability, safety, and clear failure modes Architect and...  ...relevant Create reproducible training and evaluation pipelines with versioning, experiment tracking... 

    Bain & Co.

    San Francisco, CA
    3 days ago
  •  ...This position involves critiquing Arabic AI outputs for accuracy, identifying biases...  ...degree, and 2+ years of experience in AI safety. The role offers campus benefits such as...  ...fostering a collaborative environment in AI evaluation and improvement. #J-18808-Ljbffr Welo... 

    Welo Data

    San Francisco, CA
    3 days ago
  • # Technical AI Safety SpecialistGet exceptional technical talent working on the most important...  ...2. Work with us>3. Technical AI Safety Specialist## **Who is BlueDot Impact?**We're a...  ...we must solve to make AI go well. Model evaluations are still brittle, jailbreaks are still... 
    Full time
    Work at office
    Immediate start
    Visa sponsorship
    Shift work

    Aisafety

    San Francisco, CA
    20 hours ago
  • A forward-thinking tech company is seeking an AI Trainer specializing in visual and graphic design to evaluate AI outputs. Responsibilities include assessing design quality and providing feedback to ensure high professional standards. Applicants should have formal education... 
    Remote job
    Flexible hours

    Prolific

    San Francisco, CA
    1 day ago
  • $80 - $120 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Position: Humanities / arts / culture Evaluator Type: Contract Compensation: $...  ...performance . Collaborate with subject matter experts to ensure consistency and quality.... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    San Francisco, CA
    4 days ago
  • $50 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Jack Dorsey . Position: Spanish (Spain) Audio Generalist Evaluator Expert Type: Contract Compensation: $50/hour Location:... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    9 days ago
  • Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback. Candidates should have at least 5 years of relevant experience... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Safety Specialist - Evaluation Expert. Be the first to apply!