AI Safety Specialist - Evaluation Expert
Mercor
We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback. Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality. Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains. Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking. Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. Provide structured feedback to improve model alignment and safety performance. Collaborate with AI researchers and safety teams on ongoing evaluation initiatives. Required Qualifications Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline. 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field. Excellent written English, critical thinking, and analytical reasoning skills. Ability to consistently evaluate nuanced and policy-sensitive scenarios. Preferred Qualifications Experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation. Familiarity with safety policies, content moderation, or evaluation rubric development. Experience reviewing complex, high-risk, or ambiguous content. Why Join? Shape the safety and behaviour of frontier AI models used by millions worldwide. Work on challenging, real-world safety evaluations across nuanced and high-impact domains. Collaborate with leading AI researchers, engineers, and safety teams. #J-18808-Ljbffr Mercor
$60 - $70 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ..., and Jack Dorsey . Position: AI Safety Practitioner Type: Contract Compensation... ...: Remote Role Responsibilities Evaluate AI-generated responses for safety,...SuggestedContract workSummer workRemote work- Welo Data in San Francisco seeks a full-time AI Evaluator with professional proficiency in Portuguese (Portugal) and experience in Generative AI safety. The role involves critiquing AI outputs, identifying biases, and refining evaluation frameworks. Candidates should possess...SuggestedFull time
- Welo Data is looking for a Data Labeling Associate in San Francisco to evaluate AI systems' handling of Arabic nuances. The role requires professional-level proficiency in Arabic and 2 years of AI safety experience. Responsibilities include critiquing Arabic AI outputs...Suggested
$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... .... Position: User/customer research and feedback synthesis Evaluator Type: Contract Compensation: $80–$120/hour Location...SuggestedContract workSummer workWork at officeRemote work$80 - $120 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Position: Government / public administration Evaluator Type: Contract Compensation: $... .... Collaborate with subject matter experts to ensure consistency and quality. Work...SuggestedContract workSummer workWork at officeRemote work$90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...genuinely correct replications from those that merely look correct. Evaluate responsive behavior and semantic quality, ensuring proper use of...Contract workSummer workLocal areaRemote work- OpenAI is seeking a researcher to advance frontier evaluations and environments for safe AGI/ASI. You will help design north star model... ...products. Collaborate with researchers, engineers, product and safety teams to decide what to measure, how to measure it, and how to...
- ...setting. You will tackle fundamental challenges in building safe AI agents and aligning them with humans, including benchmarking... ...DPO, or GRPO, plus strong cross-functional communication. Some familiarity with agent evaluation tooling is a plus. #J-18808-Ljbffr Scale
- ...seeking exceptional research engineers to push the boundaries of frontier AI safety, shaping empirical understanding of risk and owning end-to-end threads within this effort. You’ll design evaluations of frontier models against real threat models, develop datasets,...
$70 - $90 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Dorsey . Position: Trainium (NKI) Kernel Expert Type: Contract Compensation: $... ...: Remote Role Responsibilities Evaluate the quality, correctness, and hardware-appropriateness...Contract workSummer workRemote work- Evaluate the quality, correctness, and hardware-appropriateness of Neuron Kernel Interface (NKI) development tasks used to train and evaluate a frontier AI lab's models. You'll assess CUDA→NKI migration fidelity, Trainium-specific performance-optimization quality, and cross...
$293.5k
General Information Job Title Expert Senior Manager, AI Engineering Job ID 104335 Work... ...research, model experimentation, and evaluation design to production system... ...continual improvementBalance performance, safety, responsible AI principles, and cost...Permanent employmentFull timeApprenticeshipWork at officeLocal areaWork from homeHome office3 days per week$50 - $60 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Summers , and Jack Dorsey . Position: Expert Interviewer Type: Contract Compensation... ...technical depth, communication, and AI-evaluation fit using a provided rubric. Confirm...Contract workSummer workRemote workWeekday work$20 - $26 per hour
Prolific is seeking fluent Kannada speakers to act as evaluators who compare text and voice samples to assess naturalness and authenticity. You will listen to audio clips, rate quality, and flag any mismatches in tone or pronunciation, with emphasis on cultural context...Remote jobFlexible hours$1,750 - $2,150 per month
Obsidian is looking for experienced cybersecurity professionals to review AI systems' threat detection and vulnerability assessments. Responsibilities include evaluating AI outputs and creating realistic cybersecurity scenarios. Ideal candidates should have over 3 years...$35 per hour
...Mercor, we believe the safest AI is the one that’s already been... ...for this project - human data experts who probe AI models with adversarial... ...customer AI systems Evaluation coverage expands: more scenarios... ...production Mercor customers trust the safety of their AI because you’ve...Remote job$70 - $84 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ..., and Jack Dorsey . Position: AI Safety Red Teamer Type: Contract Compensation... ...hallucinations, and policy failures. Evaluate model robustness across misinformation, cyber...Contract workSummer workRemote work- YO AI Labs in the United States (Remote) seeks seasoned Adobe Marketing Technology Experts to support an AI training and evaluation project focused on enterprise marketing operations. You will test workflows with Adobe Workfront, AEM, CJA, Analytics and Experience Cloud...Remote job
$80 - $150 per hour
...leading behavioral health consulting firm is seeking Senior Behavioral Health Experts to work part-time and remotely on frontier AI research projects. You will be responsible for designing evaluations and testing AI systems in critical mental health contexts. The ideal...Hourly payPart timeRemote work- We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area"...
- Mercor is hiring experienced musicians to evaluate generative music AI models, in partnership with a leading AI lab. You will assess AI-generated... ...Duration: up to 6 months Commitment: flexible. Most experts work around 20 hours per week; there is no cap Application...Immediate startFlexible hours
- Mercor is seeking experienced musicians to evaluate generative music AI models. You will compare AI-generated lyrics with published songs across genres and rate them against detailed quality standards in Greek and English. The role requires native or near-native Greek,...Remote workFlexible hours
- Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) Mercor is partnering with leading AI labs on a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve...Part timeImmediate start
$128.5k - $171.5k
A data-focused AI engineering role inside the Coro unit, delivering GenAI powered tools... ...loop controls, prioritizing reliability, safety, and clear failure modes Architect and... ...relevant Create reproducible training and evaluation pipelines with versioning, experiment tracking...- ...This position involves critiquing Arabic AI outputs for accuracy, identifying biases... ...degree, and 2+ years of experience in AI safety. The role offers campus benefits such as... ...fostering a collaborative environment in AI evaluation and improvement. #J-18808-Ljbffr Welo...
- # Technical AI Safety SpecialistGet exceptional technical talent working on the most important... ...2. Work with us>3. Technical AI Safety Specialist## **Who is BlueDot Impact?**We're a... ...we must solve to make AI go well. Model evaluations are still brittle, jailbreaks are still...Full timeWork at officeImmediate startVisa sponsorshipShift work
- A forward-thinking tech company is seeking an AI Trainer specializing in visual and graphic design to evaluate AI outputs. Responsibilities include assessing design quality and providing feedback to ensure high professional standards. Applicants should have formal education...Remote jobFlexible hours
$80 - $120 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Position: Humanities / arts / culture Evaluator Type: Contract Compensation: $... ...performance . Collaborate with subject matter experts to ensure consistency and quality....Contract workSummer workWork at officeRemote work$50 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Spanish (Spain) Audio Generalist Evaluator Expert Type: Contract Compensation: $50/hour Location:...Contract workSummer workRemote work- Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback. Candidates should have at least 5 years of relevant experience...Remote jobHourly payWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Specialist - Evaluation Expert. Be the first to apply!
- ai data scientist San Francisco, CA
- ai scientist San Francisco, CA
- subject matter expert San Francisco, CA
- fulfillment expert San Francisco, CA
- guest service support expert San Francisco, CA
- technology expert San Francisco, CA
- director of quality & patient safety San Francisco, CA
- safety assistant San Francisco, CA
- safety sales San Francisco, CA
- patient safety assistant San Francisco, CA


