Remote AI Safety Specialist Evaluations & Alignment
$45 per hourMercor Inc
- Remote job
Mercor is a San Francisco–based company connecting elite AI talent with leading labs. We are seeking an AI Safety Specialist for a remote, contract role that pays $45/hour. The candidate should be based in the United States and possess strong technical skills in AI safety, red teaming, cybersecurity, and related areas. You will evaluate AI-generated content for safety, review safety protocols, and provide concrete feedback to improve data quality and model performance while working #J-18808-Ljbffr Mercor Inc
- Mercor is seeking an AI Safety Practitioner to evaluate AI-generated responses for safety, accuracy, and policy compliance. This remote contract role focuses on applying evaluation rubrics... ...with researchers to improve model alignment. The ideal candidate has 5+ years in...Remote jobContract work
- Obsidian is seeking experienced AI Safety Practitioners to evaluate frontier AI models for safety, quality, and alignment on complex topics. You will assess responses, apply safety policies, and help improve model behavior through structured evaluations and feedback. Responsibilities...Suggested
- Mercor seeks experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations...Suggested
- ...research accelerator is seeking a Geospatial Expert to enhance AI systems through advanced geospatial analysis. This entry-level, remote role involves evaluating geospatial datasets and supporting tasks aligned with crisis management and agriculture. Candidates should...Remote jobContract work
$70 per hour
We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve...Remote jobWorldwide- Rex.zone is seeking an AI Research Scientist to lead applied AI research... ...into measurable experiments in LLM evaluation and RLHF data design. You will evaluate... ...to improve model performance, safety, and usefulness. The role is remote in the United States, with compensation...Remote jobHourly payFlexible hours
- ...Language Alignment & Resource Partner (Arabic) - Freelance AI Trainer Project World Wide - Remote Project Overview We are sourcing independent Language Alignment & Resource... .... The objective of this project is to evaluate, annotate, and validate data to ensure...Remote workHourly payFor contractorsFreelance
$86.8k - $198k
Operational Technology AI EngineerThe... ...analytics. Develop and evaluate machine learning... ...around the unique safety, availability,... ...engineers, cybersecurity specialists, data scientists,... ...during meetings.Remote: If this position... ...facility frequently, in alignment with leadership...Remote workFull timeContract workPart timeWork at officeLocal area$146.37k - $246k
...are helping improve the safety, efficiency and... ...’s data foundation and AI systems. Our team drives... ...reportingLead cross-functional alignment across Finance, BizTech... ..., tool use, and evaluation techniques beyond prompt... ...flexible, employee-led remote model, a professional development...Remote workFull timeFlexible hours- ...an unwavering commitment to safety, partnership, belonging, and... ...engineering lead for Agentic AI delivery across Supply Planning... ...prompt versioning, tool auditing, evaluation frameworks, observability,... ...Own Supply Chain AI metrics alignment cadence — keeping priorities,...Remote workWorldwide
$146.37k - $221.4k
...are helping improve the safety, efficiency and... ...power our HR workflows and AI initiatives.Our People... ...scalability, and reliability.Evaluate and implement emerging... ...flexible, employee-led remote model, a professional... ...support remote work where it aligns with our operational...Remote workFull timeWork experience placementFlexible hours$35 - $42 per hour
Location: USA (Remote) Compensation: $35.00-$42.00 USD per hour(depending... ...Help Shape the Future of AI in Finance, Audit & Risk... ...growing network ofon-call AI Evaluation Specialistssupporting the development... ...for future opportunities that align with theirexpertise. Why Join...Remote workHourly payExtra incomePermanent employmentContract workTemporary workFlexible hours$118.11k - $178.65k
...Improve the safety, efficiency, and sustainability of... ...them into data products, AI agents, and automation,... ...pipelines, and the evaluation harnesses that keep them... ...flexible, employee‑led remote model, a professional development... ...remote work where it aligns with our operational...Remote workFull timeWork experience placementFlexible hours$320k
...founding technical leader for our AI Safety & Security Engineering team.... ..., security partners, and evaluation teams to build consensus through... ....Collaboration: Ability to align engineers and researchers on... ...US, CA, Santa Clara; US, NC, Remote; US, NY, Remote; US, TN, Remote...Remote workFull time$80 per hour
...ethically shape the future of AI. Our platform connects specialists with AI projects from... ...realistic and structured evaluation scenarios for LLM-based... ...to contribute to projects aligned with your skills, on your... ...Take part in a flexible, remote, freelance project that fits...Remote workPart timeFreelanceFlexible hours- ...We are looking for an AI Evaluation Scientist to design and execute evaluation processes that... ...are accurate, reliable, safe, and aligned with mission requirements. This role is... ...relevance, bias, hallucination rate, and safety metrics. Build and maintain automated...Remote work
- ...To support AI development, the part-time AI Evaluation Specialist will review and assess AI-generated outputs for quality and usability while collaborating with teams to refine evaluation standards in a remote contract role. Key responsibilities Review and critically...Remote workContract workPart time
- ...most successful outcomes, we align our capabilities to our customers... .... Job Summary: The AI Program is a strategic cross-business... ...across Bechtel meet quality, safety, and accuracy standards. As AI... ...testing frameworks and evaluation pipelines for AI solutions including...Remote work16 hoursPart timeWork at officeLocal area
$120 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our... ...Compensation: $120–$200/hour Location: Remote Role Responsibilities Evaluate complex technical tasks using deep language expertise...Remote workHourly payWeekly payFull timeContract workFor contractorsSummer work$60 - $90 per hour
...technical talent with leading AI research labs. Headquartered in... ...: Cybersecurity SWE — AI Safety Type: Contract... ...–$90/hour Location: Remote Commitment: 15–40 hours... ...cybersecurity topics. Evaluate and annotate model responses for...Remote workFull timeContract workSummer workImmediate start- ...What Is An Ai Evaluation Specialist? An AI evaluation specialist assesses and tests artificial intelligence... ...meet quality, reliability, and safety requirements before they are released... ...is typically an office-based or remote environment where they spend much of their...Remote workWork at office
$70 per hour
...Compensation: $70 per hour Join a cutting-edge AI research initiative focused on improving... ...thinking and communication skills to evaluate AI-generated responses across a variety... ...challenging tasks. This is a fully remote, contract-based opportunity with flexible...Remote workHourly payWeekly payContract workFor contractorsFlexible hours$30 per hour
...Prolific is seeking Advanced Dutch Speakers in Chicago, IL to train AI models. You will complete AI tasks and assess AI performance in... ...home with competitive rates and direct payment through PayPal. Pass the evaluation, and you can start within 15 minutes. #J-18808-Ljbffr...Remote workWork from home$117k - $145k
...AI Product Engineer The AI Product Engineer is... ...experiences and modules aligned to OKRs and ROI.... ...to guide AI tools and evaluate output. Experience with... ...eligible to be completely remote Fully Paid by Origami... ...insurance, compliance, and safety management. Founded...Remote workFull timeTemporary workWork experience placementFlexible hours- ...Prolific is seeking Product Designers and UX Specialists to join our Expert Network, contributing to the training and evaluation of cutting-edge AI models. This role requires expertise in usability, design systems, and user research, providing essential feedback on design...Remote workWork from homeFlexible hours
- ...firm seeks a Software Engineer to enhance evaluations of frontier AI systems, focusing on security and misuse risk. In this remote role, you'll manage evaluations for new AI... ...experience, data analysis skills, and an alignment with the mission of reducing catastrophic...Remote job
- ...Prolific is seeking talented Product Designers and UX Specialists to join our Expert Network for training and evaluation of cutting-edge AI models. Candidates should have a relevant educational background and at least one year of experience in design-related fields. Responsibilities...Remote workWork from homeFlexible hours
- Handshake, a leading AI-driven education and policy company, seeks an AI Policy Generalist to convert complex customer policies into precise, well-reasoned evaluations of AI model behavior. You will read requests, model outputs, and conversation history, determining the...Remote job
$100 per hour
LEGAL SUBJECT-MATTER EXPERT (AI LEGAL WORKFLOWS) ABOUT PARETO.AI Pareto.AI is a human... ...industry experts to collaborate on AI alignment, safety, and training initiatives. Together, we... ...work product Participants will create and evaluate realistic legal workflows across:...Remote jobWork at officeImmediate start- A leading AI research accelerator is seeking Geospatial Experts to enhance AI systems through evaluations and real-world applications. This entry-level contractor position is fully remote with flexible hours, primarily focusing on geospatial reasoning tasks. Responsibilities...Remote jobFor contractorsFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Remote AI Safety Specialist Evaluations & Alignment. Be the first to apply!
- ai scientist New York, NY
- ai data scientist New York, NY
- senior network engineer remote New York, NY
- business intelligence analyst remote New York, NY
- document specialist remote New York, NY
- chief of staff remote New York, NY
- it director remote New York, NY
- remote legal intern New York, NY
- remote accounts receivable New York, NY
- remote construction estimator New York, NY


