Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Frontier AI Safety Evaluator

Dorado

AIUC is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models on policy-sensitive topics. You will assess AI-generated responses and provide structured feedback to improve model behavior.

Join a collaborative team of researchers and safety engineers to advance rigorous evaluation methodologies and contribute to ongoing safety benchmarking across high-impact domains.

#J-18808-Ljbffr
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Frontier AI Safety Evaluator in United States vacancy
  • $400 per month

    Obsidian is seeking contributors for a Frontier Code Agents project, focused on evaluating AI coding models in fraud and risk engineering. Candidates will use AI coding tools to handle complex tasks and provide technical assessments. The role requires 2+ years of experience... 
    Suggested

    Obsidian

    New York, NY
    3 days ago
  • $55 - $120 per hour

     ...is a non-engineering content-policy evaluation role. Applicants must demonstrate relevant...  ...or threat assessment, trust and safety, content moderation, or closely...  ...educational institutions. Handshake AI works directly with frontier AI lab researchers to create evaluations... 
    Suggested
    Hourly pay
    Full time
    Monday to Friday
    Flexible hours
    Shift work

    Handshake

    Seattle, WA
    15 days ago
  •  ...as a critical component of our nation’s safety and security. Make an impact by using your...  ...Description The Expert-Level Network Evaluator / System Vulnerability Analyst provides advanced...  ...center of everything we do. Growth: AI-powered career tool that identifies... 
    Suggested

    General Dynamics Information Technology

    Maryland, MD
    7 days ago
  • $18 - $22 per hour

     ...Role Overview Help improve the safety of advanced AI systems by applying Hindi language fluency and cultural judgment to how models respond to sensitive topics. This role focuses on evaluating and strengthening Hindi-language interactions; prior AI or machine learning... 
    Suggested
    Remote job
    Hourly pay
    Immediate start

    SaidGig

    Remote
    1 day ago
  •  ...Oncology AI Evaluator is a remote clinical-review track for evaluating AI outputs that touch oncology. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning so the modeling... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  •  ...AI Safety and Red-Team Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  • $18 - $22 per hour

     ...Role Overview Help improve the safety of advanced AI models by applying Thai language fluency and cultural judgment to evaluate how models respond to sensitive topics in Thai. Training is provided, and prior AI or machine learning experience is not required. Key Responsibilities... 
    Remote job
    Hourly pay
    Immediate start

    SaidGig

    Remote
    1 day ago
  •  ...Educational Safety AI Evaluator is a remote evaluation track for reviewing educational safety ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  • $43 - $47 per hour

     ...Role Overview Help evaluate and strengthen how advanced AI systems respond to sensitive topics in Spanish. This role combines Spanish language fluency...  ...or technical content. Background in trust and safety, content moderation, policy evaluation, or adversarial... 
    Remote job
    Hourly pay
    Immediate start

    SaidGig

    Remote
    1 day ago
  • Crisis Evaluator - AllHealth Network At AllHealth Network, compassion, collaboration, and clinical excellence come together to support individuals...  ..., substance use, or behavioral disturbances, with a focus on safety, stabilization, and recovery.As a Crisis Evaluator, you will... 
    Full time
    Temporary work
    Relocation package
    Flexible hours
    Shift work
    Night shift
    Weekday work

    AllHealth Network

    Littleton, CO
    4 days ago
  • $14.5 per hour

    A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical... 
    Hourly pay
    Part time
    Remote work
    Flexible hours

    Welo Data

    United States
    3 days ago
  • $40 - $100 per hour

    About OpenTrain OpenTrain AI is the hiring and contracting organization for this opportunity. OpenTrain is the #1 platform for finding...  ...English-language work About AI Training and Scientific Evaluation AI training is the human side of building modern artificial intelligence... 
    Hourly pay
    Contract work
    Part time
    For contractors
    Remote work

    OpenTrain AI

    Brooklyn, NY
    2 days ago
  • Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional... 
    Remote job
    Work at office

    Obsidian

    New York, NY
    1 day ago
  • $20 per hour

    A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and... 
    Remote job
    Flexible hours

    DataAnnotation

    New York, NY
    1 day ago
  • Dorado is seeking expert Evaluators in program management / implementation planning to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote,... 
    Remote job
    Hourly pay
    Work at office

    Dorado

    New York, NY
    4 days ago
  • Mercor is seeking experienced musicians to evaluate generative music AI models, in partnership with a leading AI lab. You will assess AI-generated music across a wide range of genres and rate it against detailed quality standards, working in Hindi and English. Responsibilities... 

    Obsidian

    New York, NY
    5 days ago
  • $80 - $120 per hour

    Gridnaut Recruiting is hiring a remote Software / AI / IT / data Evaluator contractor (pay $80-$120/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Remote contractor engagement supporting a leading AI lab.; Hands-on technical review of model... 
    Hourly pay
    Contract work
    For contractors
    Remote work

    Gridnaut Recruiting

    Remote
    7 days ago
  • YO AI Labs is seeking an Investment & Finance Expert to work remotely as a contractor. The role focuses on evaluating AI outputs related to valuation, financial modeling, markets, and investments, and crafting expert prompts plus reference answers to improve model reasoning... 
    Remote job
    For contractors

    YO AI Labs

    Austin, TX
    3 days ago
  • Dorado is seeking an AI Language Quality Evaluator fluent in Greek and English for an ongoing, task-based project. This remote freelance role involves reviewing translated and AI-flagged content to judge accuracy, classify issues, and suggest corrected translations. You... 
    Remote job
    For contractors
    Freelance
    Flexible hours

    Dorado

    Brooklyn, NY
    4 days ago
  • Localizationacademy in Redmond, WA is seeking a Multilingual Content Quality Evaluator / Annotator to review and refine AI-translated content onsite. You will identify translation errors, adjust wording for natural, culturally appropriate language, and ensure the output... 

    Localizationacademy

    Redmond, WA
    2 days ago
  • $80 - $120 per hour

    Gridnaut Recruiting is hiring a remote Cybersecurity / IT GRC Evaluator contractor (pay $80-$120/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Remote contractor engagement supporting a leading AI lab.; Remote AI evaluation / expert contributor... 
    Hourly pay
    Contract work
    For contractors
    Remote work
    Flexible hours

    Gridnaut Recruiting

    Remote
    7 days ago
  • Mercor is hiring experienced musicians to evaluate generative music AI models, in partnership with a leading AI lab. You will assess AI-generated music across a wide range of genres and rate it against detailed quality standards, working in Malayalam and English. Key responsibilities... 

    Mercor

    New York, NY
    5 days ago
  •  ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise... 
    Hourly pay
    Work at office
    Remote work

    Mercor Inc

    New York, NY
    4 days ago
  • $18 per hour

     ...Job Description Job Description Title of Position: DUI Evaluator Position Type: Part-time, Non-Exempt Hours: 20-25 per week...  ...Who are we? Founded by concerned citizens in 1953, the Florida Safety Council, a non-profit 501(c)(3) organization, was established... 
    Hourly pay
    Full time
    Part time
    Work at office
    Shift work

    United Safety Council Inc

    Melbourne, FL
    10 days ago
  •  ...Supporting diverse AI data and language projects, the hourly contractor AI Trainer and Evaluator will work remotely to generate content, annotate data, and evaluate AI responses for accuracy and cultural relevance. Key responsibilities Generate high-quality prompts and... 
    Hourly pay
    For contractors
    Remote work

    Virtual Vocations Inc

    United States
    5 days ago
  • $20 per hour

    A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI...  ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with... 
    Hourly pay
    Remote work
    Flexible hours

    SupportFinity

    United States
    1 day ago
  • $400 per month

    About the Role Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments. The work focuses on realistic infrastructure engineering... 

    Mercor

    Miami, FL
    2 days ago
  • YO AI Labs is seeking a Biology Expert to join a remote, contract-based project. You will evaluate biology-related data and craft scientific scenarios to improve AI training datasets. The role emphasizes accuracy, depth, and clear communication within interdisciplinary... 
    Remote job
    Contract work

    YO AI Labs

    Austin, TX
    1 day ago
  • Urgent Hiring u2013 Multilingual Content Quality Evaluator / Annotator | Redmond, WA | Onsite Location: Redmond, WA (100% Onsite u2013 5 Days/Week) Responsibilities Review and clean AI-translated content. Identify and correct translation errors. Ensure translations... 

    Localizationacademy

    Redmond, WA
    2 days ago
  • YO AI Labs seeks an experienced Developer & Infrastructure Expert for remote contract work. You will evaluate AI-powered workflows across DevOps, cloud infrastructure, SRE, and platform engineering, testing AI-generated commands and configurations against real-world engineering... 
    Remote job
    Contract work

    YO AI Labs

    Austin, TX
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Frontier AI Safety Evaluator. Be the first to apply!