Bilingual French Generalist Expert — AI Safety Evaluation (Remote)
$48 - $52 per hourAuraOne Human Data
- Remote job
Bilingual French Generalist Expert — AI Safety Evaluation (Remote) is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Why this role matters
Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.
Responsibilities
- Design adversarial prompts that probe known weakness classes (jailbreak, policy bypass, prompt injection) for Bilingual French Generalist Expert — AI Safety Evaluation (Remote) assignments.
- Document every successful attack with reproduction steps and the policy clause it violated.
- Score model defenses across single-turn and multi-turn conversations.
- Triage emerging attack vectors and route them to the safety team with severity ratings.
- Maintain a personal library of attack patterns and propose new red-team rubrics.
- Calibrate against the broader red-team cohort to keep coverage and severity consistent.
Qualifications
- Demonstrated experience red-teaming AI systems, security research, or adversarial ML work for Bilingual French Generalist Expert — AI Safety Evaluation (Remote) work.
- Strong written communication — your reports become the patch ticket.
- Comfort working in policy-grey areas with clear documentation of what was attempted and why.
- Familiarity with prompt-injection, jailbreak, and policy-bypass taxonomies.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Construct a 5-turn adversarial conversation that bypasses a specific policy clause and write up the patch ticket.
- Score a model's defenses against a known jailbreak pattern across 20 variants.
- Propose a new red-team rubric category after spotting an emerging attack vector.
- Reproduce a failure another reviewer reported and confirm the severity tag.
Nice to have
- Background in offensive security, AppSec, or trust & safety operations.
- Experience publishing or reproducing public adversarial-ML research.
- Multilingual fluency for cross-language attack testing.
Skills
- Adversarial prompting
- Red-team analysis
- Policy taxonomy
- Failure documentation
- French generalist evaluation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
$48–$52 / hr
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$48 - $52 per hour
...Help improve the safety of advanced AI systems by using your French language expertise and cultural... ...topics. Training on the evaluation workflow is provided, and... ...Create expert-level French prompts covering... ...testing. Work Terms Remote, hourly engagement. Immediate...Remote workBilingualFrench language skillsHourly payImmediate start$60 - $100 per hour
...Legal Domain Expert — AI Training & Evaluation is a remote review track for evaluating AI outputs in legal review workflows. Reviewers grade citation accuracy... ...AI-assisted legal-research or compliance tooling. Bilingual experience for cross-jurisdiction matters. Skills...Remote workBilingualFor contractors10 hours per week$28 - $60 per hour
...Evaluate AI-generated music and lyrics across a wide range of genres,... ...and artists. Work Terms Remote, independent-contractor engagement... .... Flexible schedule. Most experts work about 20 hours per week,... ...steps, including a Dutch bilingual competency interview....Remote workBilingualHourly payFor contractorsImmediate startFlexible hours- YO AI Labs is seeking Turkish bilingual experts for a contract-based remote role focusing on language and AI training. You will evaluate Turkish audio samples, assess quality, and provide precise feedback to help improve AI systems. No prior AI experience is required, but...Remote jobBilingualContract work
- ...Bilingual French AI Evaluation Specialist is a remote evaluation track for reviewing french generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare... ..., content moderation, or trust & safety review. Experience with inter-rater agreement...Remote jobBilingualFrench language skillsHourly payFor contractors10 hours per week
$100 - $130 per hour
...Radiology Domain Expert - Diagnostic Imaging AI Review is a remote clinical-review track for evaluating AI outputs that touch radiology. Reviewers... ...adherence; flag patient-safety issues; and document the... ...and its failure modes. Bilingual clinical experience for non...Remote jobBilingualFor contractors10 hours per week$58 - $62 per hour
...Overview Help improve the safety of advanced AI systems by applying... ...technical topics in French. You will assess and... ...provided for the evaluation workflow. Prior AI or... ...Responsibilities Create expert-level French prompts... .... Work Terms Remote hourly engagement,...Remote jobBilingualFrench language skillsHourly payImmediate start$25 - $37 per hour
...Bilingual French AI Evaluation Specialist is a remote French specialist track for evaluating french evaluation outputs against native-speaker standards. Reviewers spot fluency, register, and cultural-context errors that automated checks miss, and write structured rationale...Remote jobBilingualFrench language skillsFor contractors10 hours per week$50 per hour
...Generalist AI Evaluation Specialist is a remote review track for evaluating AI outputs across generalists specialist operations workflows. Reviewers grade... ...-assisted workflow tooling and its failure modes. Bilingual experience for cross-region operations. Skills...Remote jobBilingualFor contractorsWork experience placement10 hours per week- About The Role We're looking for French language experts to help evaluate and improve AI systems trained on French-language content. Your linguistic insight and... ...: Alignerr Type: Hourly Contract Location: Remote Commitment: 15+ hours/week, fully asynchronous French...Remote workFrench language skillsHourly payOngoing contractContract workFreelanceWorldwideFlexible hours
€150 - €160 per hour
...Role Overview Apply your French corporate and commercial law expertise to help evaluate the quality of AI-generated legal answers. You will review responses to practical... ...rubric, and meet deadlines. Work Terms Remote role for candidates located in France. Hourly...Remote workFrench language skillsHourly payFor contractors- YO AI Labs is seeking a Turkish bilingual expert to support a language and AI training project. This remote contractor role focuses on evaluating Turkish audio for nativeness, fluency, pronunciation, and intonation. You will provide clear feedback in English, justify observations...Remote jobBilingualFor contractors
- YO AI Labs is seeking a Korean Language Expert to remotely support AI data quality and localization tasks. You will evaluate, translate, annotate, and refine Korean-language content to help... ...emphasizes strong linguistic judgment, bilingual communication, and attention to...Remote jobBilingual
- A leading AI Data Services company is seeking a bilingual content evaluator to review AI-generated responses and create training content. The ideal candidate will have... ..., and optimizing AI performance. This fully remote role offers flexible hours and requires at least...Remote jobBilingualFlexible hours
- A growing AI Data Services company is seeking a contractor for a fully remote role focusing on reviewing and generating high-quality bilingual training content. You will create and evaluate AI responses, ensuring accuracy and clarity in Hebrew and English. The ideal candidate...Remote jobBilingualFor contractorsFlexible hours
$60 - $65 per hour
...Role Overview Help improve next-generation AI language systems by evaluating Indonesian audio and delivering precise, real-world linguistic feedback. This remote contract role focuses on assessing the naturalness, fluency, cultural fit, and authenticity of AI-generated...Remote workBilingualHourly payContract work- YO AI Labs is seeking Dutch bilingual experts to contribute to a global language and AI training project. You will evaluate AI-generated Dutch speech, assess nativeness and quality, and provide... ...help improve language models. This remote contractor role requires native...Remote jobBilingualFor contractors
- YO AI Labs is seeking Polish bilingual experts to help improve Polish-language understanding for AI training. You will evaluate Polish audio, judge nativeness and fluency, and provide clear feedback in English. This remote contractor role requires strong Polish proficiency...Remote jobBilingualFor contractors
$48 - $52 per hour
...Overview Help improve the safety of advanced AI models by applying Finnish language... ...Responsibilities Create expert-level Finnish prompts... ...behind each assessment. Evaluate and help strengthen how AI models... ...testing. Work Terms Remote, hourly engagement. Immediate...Remote jobBilingualHourly payImmediate start$38 - $42 per hour
...Help improve how advanced AI models respond to sensitive... ...and cultural judgment to evaluate model interactions, identify safety concerns, and support stronger... ...Create expert-level Croatian prompts across... ...testing. Work Terms Remote hourly engagement. Immediate...Remote jobBilingualHourly payImmediate start$48 - $52 per hour
...Overview Help improve the safety of advanced AI systems by applying Chinese... ...fluency and cultural judgment to evaluate how models respond to... ...Responsibilities Develop expert level Chinese prompts covering... ...adversarial testing. Work Terms Remote, with East Asia preferred....Remote jobBilingualHourly payImmediate start$43 - $47 per hour
...Role Overview Help evaluate and strengthen how advanced AI systems respond to sensitive topics... ...Create expert-level Spanish prompts covering... ...Background in trust and safety, content moderation, policy... ...testing. Work Terms Remote, hourly engagement. Immediate...Remote jobBilingualHourly payImmediate start$40 - $44 per hour
...Overview Help improve the safety of advanced AI models by applying Portuguese... ...fluency and cultural judgment to evaluate how they respond to... ...Key Responsibilities Write expert-level Portuguese prompts covering... ...testing. Work Terms Remote, hourly engagement....Remote jobBilingualHourly payImmediate start$18 - $22 per hour
...Overview Help improve the safety of advanced AI models by applying Thai language... ...and cultural judgment to evaluate how models respond to sensitive... ...Responsibilities Create expert level Thai prompts across... ...adversarial testing. Work Terms Remote, with Southeast Asia...Remote jobBilingualHourly payImmediate start$150 - $180 per hour
Prolific in Sacramento is seeking Mental Health Professionals to help train and evaluate advanced AI models. You will review AI-generated responses and engage in various training tasks, earning competitive pay rates up to $150-180/hr. Ideal candidates must hold a verified...Remote jobWork from home$150 - $180 per hour
...Virginia Beach is seeking Mental Health Professionals to train and evaluate AI models. The role involves reviewing AI responses to... ...attention to detail, and a reliable internet connection. Join our Expert Network to influence future AI innovations and work flexibly from...Remote jobWork from home- ...Location: Remote Fluent Language Skills Required: English... ...we believe the safest AI is the one that’s... ...this project - human data experts who probe AI models with... ...customer AI systems Evaluation coverage expands: more... ...Mercor customers trust the safety of their AI because you...Remote work
$140 - $150 per hour
...practise in France. The role supports an AI legal evaluation project benchmarking AI responses to French law questions. You will review AI-... ...reasoning, and provide concise expert feedback. The engagement is part-time, fully remote, with an initial commitment of roughly...Remote jobFrench language skillsPart time$8 - $65 per hour
Prolific is hiring Mental Health Professionals in New York to train and evaluate AI models. As a Domain Expert, you will be responsible for reviewing AI-generated responses, completing tasks related to psychology, and improving AI models based on your expertise. Pay rates...Remote jobHourly payWork from homeFlexible hours$120 per hour
...and technical talent with leading AI research labs. Headquartered in San... ...Jack Dorsey . Position: Expert Software Engineer – Scala / COBOL... ...120–$200/hour Location: Remote Role Responsibilities Evaluate complex technical tasks using deep...Remote workHourly payWeekly payFull timeContract workFor contractorsSummer work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Bilingual French Generalist Expert — AI Safety Evaluation (Remote). Be the first to apply!




