Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Project Harbor Red Teaming - Juno | Remote AI Safety Evaluation

OneForma is looking for native Arabic linguists based in Switzerland , Saudi Arabia or the United Arab Emirates to join an exciting remote AI research project focused on AI safety evaluation and quality assurance.

Participants will review AI-generated responses and perform structured linguistic assessments to identify harmful content, safety concerns, and policy violations. The work involves evaluating large language model (LLM) outputs, classifying potential risks, documenting findings, and validating improvements after issues have been addressed.

This is a flexible, remote opportunity for experienced Arabic language professionals who are interested in helping improve the safety, reliability, and accuracy of next-generation AI systems.

Purpose:

The purpose of this project is to improve the safety and quality of AI chatbot systems by leveraging human expertise to evaluate model responses, identify potential risks, and provide high-quality feedback that supports ongoing AI development.

Main Requirements:

  • Must be a native Arabic speaker based in Switzerland , Saudi Arabia or the United Arab Emirates .
  • Linguistic background or professional language experience is preferred.
  • Must be familiar with Apple products and AppleCare content.
  • Experience evaluating Generative AI (GenAI) models or large language model (LLM) outputs is preferred.
  • Must own a personal MacBook running the latest version of macOS Tahoe .
  • Device must be personal (not shared) and have administrator privileges for software installation.
  • Must confirm whether you have an AppleConnect account.
  • Must be available to receive immediate, ad hoc task assignments.
  • Must demonstrate professionalism, attention to detail, and the ability to follow evaluation guidelines.
  • Strong analytical and written communication skills are required.

Other Important Information:

  • This is a fully remote project.
  • Participants will review AI-generated conversations for safety, quality, and policy compliance.
  • Tasks include identifying harmful or unsafe content, classifying issues by risk level, documenting findings, and validating fixes after remediation.
  • The use of AI tools, automation, or speech-to-text software is strictly prohibited.
  • Completion of onboarding and certification is mandatory before production work begins.
  • Direct communication with the client point of contact will be required throughout the project.

Help shape the future of trustworthy AI by using your language expertise to make AI systems safer, more accurate, and more reliable for users worldwide.

Vacancy posted more than 2 months ago
Similar jobs that could be interesting for youBased on the Project Harbor Red Teaming - Juno | Remote AI Safety Evaluation in United States vacancy
  • Location : Remote Fluent Language Skills Required...  ...we believe the safest AI is the one that’s...  ...We are assembling a red team for this project - human data experts...  ...customer AI systems Evaluation coverage expands: more...  ...customers trust the safety of their AI because you... 
    Remote work
    Project

    Mercor

    New York, NY
    4 days ago
  • $16 - $22 per hour

     ...strengthen conversational AI by testing models...  ...that improves safety and reliability....  ...role focuses on red teaming AI models and agents...  ...consistent evaluations. Create reproducible...  ...Adaptability across projects, task types, and...  ...Work Terms ~ Remote, hourly engagement... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    12 days ago
  • $16 - $22 per hour

     ...strengthen conversational AI by testing models...  ...that improves safety and reliability....  ...role focuses on red teaming AI models and agents...  ...consistent evaluations. Create reproducible...  ...Adaptability across projects, task types, and...  ...Work Terms ~ Remote, hourly engagement... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    14 hours ago
  • $17 - $25 per hour

     ...conversational AI by testing it from...  .... In this remote role, you will...  ...create actionable safety data, and help...  ...higher-sensitivity projects is optional,...  ...attack cases that teams can use to...  ...systems. Expand evaluation coverage by...  ...experience in AI red teaming, adversarial... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    more than 2 months ago
  •  ...strengthen conversational AI systems by probing...  ...high-quality safety data. This text-...  ...Red team conversational AI...  ...and adapt across projects and client needs....  ...Work Terms Remote, hourly position....  ...Impact Expand evaluation coverage across more... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    more than 2 months ago
  • $48 - $62 per hour

     ...Probe conversational AI systems and produce...  ...and improves model safety. This is a text-based red teaming role focused on...  ...comfortable moving across projects and customer...  ...systems. Expand evaluation coverage so more scenarios...  ...Terms Location: Remote. Engagement type... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    14 hours ago
  • $17 - $25 per hour

     ...Overview Join a red team of human data experts...  ...conversational AI models with adversarial...  ...improves model safety. This text-based...  ...higher-sensitivity projects governed by explicit...  ...Expanding evaluation coverage so more scenarios...  ...Work Terms Remote, hourly engagement... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    2 days ago
  • $17 - $25 per hour

     ...of conversational AI, creating reproducible...  ...Red team conversational AI...  ...issues Work on projects that may touch sensitive...  ...systems Expanding evaluation coverage so fewer...  ...confidence in their AI safety because systems have...  ...Work Terms Remote engagement Hourly... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    14 hours ago
  • $48 - $62 per hour

     ...Overview Work as a human red team specialist probing conversational AI to find vulnerabilities,...  ...to move across projects and customers as needs change...  ...customer AI systems. Evaluation coverage increases, reducing...  ...team. Work Terms Remote, text-based work. Participation... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    14 hours ago
  •  ...Contribute expert red-teaming insight to a fast-paced AI safety project focused on tackling safety challenges within a...  ...safety, content moderation, policy evaluation, or adversarial testing Comfort...  ...tight deadlines Work Terms Remote, U.S.-based engagement Per-... 
    Remote work
    Project
    Temporary work

    SaidGig

    United States
    4 days ago
  •  ...Safety Evaluation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure...  ...role-specific routing and review. Final project scope, schedule, and contractor terms are... 
    Remote job
    Project
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  • $20 - $22 per hour

     ...technical talent with leading AI research labs....  ...Dorsey . Position: AI Safety Experts — English & Assamese...  ...$22/hour Location: Remote Role Responsibilities Red team conversational AI models...  ...Adaptability to move across projects and customers.... 
    Remote work
    Project
    Contract work
    Summer work

    Mercor

    San Francisco, CA
    15 days ago
  • $125k - $190k

    About the Role FAR.AI is hiring a Technical Project Manager to be the delivery...  ...frontier AI red‑teaming programmes. You will...  ...frontier labs, AI safety organisations, technical...  ...exposure to AI evaluations, red‑teaming, or AI...  .... Location: Remote globally. We can sponsor... 
    Remote work
    Project
    Full time
    Contract work
    For contractors
    Visa sponsorship
    Shift work

    Aisafety

    Berkeley, CA
    2 days ago
  • $16 - $22 per hour

     ...Help make conversational AI safer by probing models with...  ...and producing actionable red-team data. This text-based remote role focuses on testing AI...  ...Ability to adapt across projects, task types, and customers...  ...testing may miss, expand evaluation coverage, and help reduce... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    3 days ago
  • $17 - $25 per hour

     ...Role Overview Join a red team that probes conversational AI with adversarial inputs to...  ...structure tests and keep evaluations consistent. Document findings...  ...through past projects, roles, or equivalent work...  ...Work Terms Location: Remote. Employment type: hourly... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    2 days ago
  • $48 - $62 per hour

     ...Role Overview Work as a red team human-data expert probing conversational AI models and agents to...  ...comfortable moving across projects and customers Nice-...  ...customer AI systems Evaluation coverage expands,...  ...Work Terms Location: Remote Employment type: Hourly... 
    Remote work
    Project
    Hourly pay
    Flexible hours

    SaidGig

    United States
    14 hours ago
  •  ...Internal Medicine AI Evaluator is a remote clinical-review track for...  ...adherence; flag patient-safety issues; and document...  ...reasoning so the modeling team can close the gap....  ...reasoning, dosing, and red-flag handling on a structured...  ...and review. Final project scope, schedule, and... 
    Remote job
    Project
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    29 days ago
  • $24 - $35 per hour

     ...Role Overview Join a red team committed to finding real...  ...in conversational AI. You will probe models...  ...Adaptable to shifting projects and customer needs....  ...AI systems. Increase evaluation coverage so fewer surprises...  ...Work Terms Location: Remote. Employment type: hourly... 
    Remote work
    Project
    Hourly pay
    Shift work

    SaidGig

    United States
    2 days ago
  • $48 - $62 per hour

     ...Role Overview Red team conversational AI and generate high-quality...  ...higher sensitivity projects is optional and supported...  ...AI systems. Evaluation coverage expands, resulting...  ...confidence in the safety of their AI because...  ...Terms Location: Remote. Employment type:... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    2 days ago
  • $125k - $190k

     ...Aisafety is seeking a Technical Project Manager in Berkeley, California, to lead high-impact AI red-teaming initiatives. You will manage complex multi-party engagements...  ...government and AI stakeholders. The role offers a remote global work possibility with a salary range of... 
    Remote work
    Project

    AISafety

    Berkeley, CA
    3 days ago
  •  ...Secure Code Review AI Evaluator is a remote red-team track for stress-testing AI systems against adversarial...  ...rubric clause it violated so the safety team can patch the gap. Why this role...  ...-specific routing and review. Final project scope, schedule, and contractor... 
    Remote job
    Project
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    3 days ago
  • $17 - $25 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ..., and Jack Dorsey . Position: AI Safety Experts — English & Malay Type: Contract...  ...Compensation: $17-$25/hour Location: Remote Role Responsibilities Red team conversational AI models and agents.... 
    Remote work
    Contract work
    Summer work

    Mercor

    San Francisco, CA
    3 days ago
  • $65 - $75 per hour

     ...Overview Help test advanced AI systems on radiological safety questions where legitimate...  ...-use, or adversarial. Evaluate model responses against a...  ...handling experience. Red-teaming experience is preferred....  ...application. Work Terms Remote, per-task engagement.... 
    Remote work

    SaidGig

    United States
    14 hours ago
  • Mercor is assembling a red team for AI safety. You will review model outputs, simulate adversarial inputs, and surface vulnerabilities across prompts...  ...taxonomies and playbooks to keep testing consistent. Remote position based in the United States; topics include bias, misinformation... 
    Remote job

    aitrainer

    New York, NY
    1 day ago
  • Mercor seeks a remote AI red-teaming specialist fluent in English and Portuguese to assess safety of conversational models. You will simulate jailbreaks, prompt injections, and misuse scenarios while documenting results for customers. You will generate high-quality human... 
    Remote job

    Mercor

    New York, NY
    1 day ago
  • Mercor is building a remote AI safety red team to probe conversational models and surface vulnerabilities in high-sensitivity topics. You will annotate failures, classify risks, and generate reproducible reports for customers to act on. Before exposure to content, topics... 
    Remote job

    Mercor

    New York, NY
    4 days ago
  •  ...Help assess frontier AI systems at the boundary...  ...legitimate radiological safety work and potentially dangerous...  ..., or adversarial. Evaluate AI responses against a...  ...authorized users. Red-teaming experience is preferred...  .... Work Terms Remote, per-task independent contractor... 
    Remote work
    For contractors

    SaidGig

    United States
    15 days ago
  • Obsidian is looking for data experts to join a red team that probes AI models with adversarial inputs. This role requires fluency in both English...  ...AI and cybersecurity, emphasizing the importance of AI safety. Work will be structured, with clear guidelines and wellness... 
    Remote job

    Obsidian

    San Francisco, CA
    4 days ago
  •  ...applied research and AI security company...  ...state‑of‑the‑art red‑team steering across high...  ...security and safety challenges. Role Overview...  .... This is a remote‑first, contract position...  ...privacy, model evaluations, managing...  ...working on multiple projects at once with frequent... 
    Remote work
    Project
    Contract work

    10a Labs

    New York, NY
    2 days ago
  • $16 - $22 per hour

     ...strengthen conversational AI by probing models...  ...high-quality safety data. This text-based remote role focuses on adversarial...  ...attack cases that teams can use to improve...  ...systems. Expand evaluation coverage by identifying...  ...Adaptability across projects, task types, and... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Project Harbor Red Teaming - Juno | Remote AI Safety Evaluation. Be the first to apply!