Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Project Harbor Red Teaming - Juno | Remote AI Safety Evaluation

OneForma is looking for native Arabic linguists based in Switzerland , Saudi Arabia or the United Arab Emirates to join an exciting remote AI research project focused on AI safety evaluation and quality assurance.

Participants will review AI-generated responses and perform structured linguistic assessments to identify harmful content, safety concerns, and policy violations. The work involves evaluating large language model (LLM) outputs, classifying potential risks, documenting findings, and validating improvements after issues have been addressed.

This is a flexible, remote opportunity for experienced Arabic language professionals who are interested in helping improve the safety, reliability, and accuracy of next-generation AI systems.

Purpose:

The purpose of this project is to improve the safety and quality of AI chatbot systems by leveraging human expertise to evaluate model responses, identify potential risks, and provide high-quality feedback that supports ongoing AI development.

Main Requirements:

  • Must be a native Arabic speaker based in Switzerland , Saudi Arabia or the United Arab Emirates .
  • Linguistic background or professional language experience is preferred.
  • Must be familiar with Apple products and AppleCare content.
  • Experience evaluating Generative AI (GenAI) models or large language model (LLM) outputs is preferred.
  • Must own a personal MacBook running the latest version of macOS Tahoe .
  • Device must be personal (not shared) and have administrator privileges for software installation.
  • Must confirm whether you have an AppleConnect account.
  • Must be available to receive immediate, ad hoc task assignments.
  • Must demonstrate professionalism, attention to detail, and the ability to follow evaluation guidelines.
  • Strong analytical and written communication skills are required.

Other Important Information:

  • This is a fully remote project.
  • Participants will review AI-generated conversations for safety, quality, and policy compliance.
  • Tasks include identifying harmful or unsafe content, classifying issues by risk level, documenting findings, and validating fixes after remediation.
  • The use of AI tools, automation, or speech-to-text software is strictly prohibited.
  • Completion of onboarding and certification is mandatory before production work begins.
  • Direct communication with the client point of contact will be required throughout the project.

Help shape the future of trustworthy AI by using your language expertise to make AI systems safer, more accurate, and more reliable for users worldwide.

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the Project Harbor Red Teaming - Juno | Remote AI Safety Evaluation in United States vacancy
  •  ...AI Safety and Red-Team Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios,...  ...specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed... 
    Remote job
    Project
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  • $48 - $62 per hour

     ...conversational AI systems by probing...  ...-generated safety data. This text...  ...higher-sensitivity projects is optional,...  ...conduct consistent evaluations. Create...  ...attack cases that teams can use to...  ...experience in AI red teaming, adversarial...  ...Work Terms Remote, hourly... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    a month ago
  • Location Remote Fluent Language Skills Required English...  ...believe the safest AI is the one that’s...  ...We are assembling a red team for this project - human data experts...  ...customer AI systems Evaluation coverage expands: more...  ...customers trust the safety of their AI because you... 
    Remote work
    Project

    Obsidian

    San Francisco, CA
    1 day ago
  • $26 per hour

     ...: $26/hour Location: Remote Commitment: 10-40 hours...  ...Responsibilities Red-team conversational AI systems using jailbreaks...  ...test cases. Evaluate AI outputs across sensitive...  ...alignment with defined safety guidelines. Requirements...  ...in a remote, project-based environment. #J... 
    Remote job
    Project
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    1 day ago
  • Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate...  .... Remote collaboration across projects and clients is common. #J-18808-Ljbffr... 
    Remote job
    Project

    Mercor Inc

    New York, NY
    3 days ago
  • Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human... 
    Remote job
    Project

    Neon

    New York, NY
    2 days ago
  • Mercor is seeking a remote Red Team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate actionable red team...  ...established taxonomies and playbooks. The role offers multi-project opportunities and requires strong communication of... 
    Remote job
    Project

    Mercor

    San Francisco, CA
    4 days ago
  • Mercor is building a remote red team to probe AI safety. You will test conversational AI models and agents, including jailbreaks, prompt injections...  ...emphasizes clear communication, structured methods, and adaptability across projects and clients. #J-18808-Ljbffr Mercor
    Remote job
    Project

    Mercor

    New York, NY
    2 days ago
  • Mercor is seeking a remote red-team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate high-quality...  .... You’ll document results and collaborate across projects for robust safety testing. This role focuses on reproducible... 
    Remote job
    Project

    Obsidian

    San Francisco, CA
    3 days ago
  • Mercor is building a red team to probe AI models with adversarial inputs, surface vulnerabilities, and generate red team data that strengthens safety for customers. This is a text-based, higher-sensitivity project where participation is optional and guided by clear guidelines... 
    Remote job
    Project

    Mercor

    San Francisco, CA
    9 hours ago
  • Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe...  ...attack data for safer deployments. You’ll work across projects with structured test plans, clear documentation, and reproducible... 
    Remote job
    Project

    Neon

    New York, NY
    2 days ago
  • $20 - $22 per hour

     ...strengthen conversational AI systems by probing them as an adversary. This remote, text-based role...  ...generating high-quality red-team data, and creating reproducible...  ...that improve AI safety and reliability. Key...  ...Comfort adapting across projects and customers. Preferred... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    17 days ago
  • Mercor is seeking an AI Safety Expert to join a remote contract team. You will red team conversational AI, identify jailbreaks, and analyze biases to improve model...  ...and the ability to work asynchronously across projects. #J-18808-Ljbffr United States Digital Space LLC
    Remote job
    Project
    Contract work

    United States Digital Space LLC

    New York, NY
    2 days ago
  • $29 - $45 per hour

     ...Attack and probe conversational AI systems in English and...  ...exposure. Key Responsibilities Red team conversational AI models and...  ...Adaptable, comfortable moving between projects and different customer...  ...adversarial thinking. Work Terms Remote, text-based engagement.... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    a month ago
  • Mercor is building a red-team for AI safety. The work involves reviewing AI outputs touching bias,...  ...and participation in higher-sensitivity projects is optional with clear guidelines and...  ...customers strengthen AI safety. This is a remote role based in the United States, and... 
    Remote job
    Project

    Neon

    New York, NY
    2 days ago
  • Mercor is building a remote AI red team to probe and test conversational models for vulnerabilities, bias, and safety gaps. This role emphasizes reproducible data and actionable findings across projects and customers. You bring prior red teaming experience in AI or cybersecurity... 
    Remote job
    Project

    Obsidian

    New York, NY
    1 day ago
  • $29 - $45 per hour

     ...testing of conversational AI to find safety vulnerabilities and...  ..., human-generated red team data that clients can act on. Work is remote and text-based. The role...  ...moving across varied projects and client needs Nice...  ...AI systems Expand evaluation coverage so fewer unexpected... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    4 days ago
  • $48 - $62 per hour

     ...of conversational AI in both English and...  ...labeled data that helps teams make AI systems...  ...Responsibilities Red team conversational...  ...to move across projects and customer contexts...  ...systems Expanding evaluation coverage so fewer...  ...risks Work Terms Remote engagement, all... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    2 days ago
  • Mercor is assembling a remote red team to test and strengthen AI systems through adversarial inputs. You will annotate failures, classify vulnerabilities...  ...and a structured approach, thriving across different projects and customers while delivering high-quality data and... 
    Remote job
    Project

    Neon

    New York, NY
    2 days ago
  • $250k - $400k

    You'll manage our growing red-team, own delivery for frontier lab customers...  ...Our mission is to automate AI safety , to pave the way for a...  ...we're looking for People and project management : you can hire,...  ...alongside us there. We are open to remote for the right candidate.... 
    Remote work
    Project
    Visa sponsorship
    Shift work

    Trajectory Labs, PBC

    Berkeley, CA
    2 days ago
  •  ...Safety Evaluation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure...  ...role-specific routing and review. Final project scope, schedule, and contractor terms are... 
    Remote job
    Project
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    28 days ago
  • $20 - $22 per hour

     ...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce...  ...in higher-sensitivity projects is optional, supported by clear guidelines...  ...scenarios Work Terms Remote, text based engagement on an hourly... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    4 days ago
  • $20 - $22 per hour

     ...technical talent with leading AI research labs....  ...Dorsey . Position: AI Safety Experts — English & Assamese...  ...$22/hour Location: Remote Role Responsibilities Red team conversational AI models...  ...Adaptability to move across projects and customers.... 
    Remote work
    Project
    Contract work
    Summer work

    Mercor

    New York, NY
    16 days ago
  • $24 - $35 per hour

     ...technical talent with leading AI research labs....  ...Dorsey . Position: AI Safety Experts — English & Thai...  ...$35/hour Location: Remote Role Responsibilities Red team conversational AI models...  ...Participate in higher-sensitivity projects with clear guidelines and... 
    Remote work
    Project
    Contract work
    Summer work

    Mercor

    San Francisco, CA
    8 days ago
  • $65 per hour

    A leading AI consulting firm is seeking an AI Tutor specialized in Coding....  ...This part-time freelance role involves evaluating AI models, creating test cases, and...  ..., this position allows you to work remotely on challenging AI projects that enhance your expertise. Compensation... 
    Remote job
    Project
    Part time
    Freelance
    Flexible hours

    Mindrift

    San Antonio, TX
    2 days ago
  • $26 per hour

    A tech consulting firm is seeking a remote contractor to engage in AI red teaming activities. The ideal candidate will work on evaluating conversational AI systems and generating human data through careful documentation and classification of vulnerabilities. Fluency in... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week
    Flexible hours

    Crossing Hurdles

    New York, NY
    1 day ago
  •  ...AI Red Team Engineer We're looking for an AI Red Team Engineer to break LLM-powered systems...  ...write it up. White Circle is an AI Safety company building the safety, reliability...  ...tool/function calling, and LLM-as-judge evaluation. Familiarity with OWASP LLM Top 10, OWASP... 
    Remote work
    Local area

    Pumpkin Intelligence, Inc.

    United States
    4 days ago
  • $17 - $25 per hour

     ...strengthen conversational AI systems by probing them...  ...producing actionable safety data. This remote role focuses on text-based evaluation of AI outputs and...  ...Key Responsibilities Red team conversational AI models...  ...work across different projects and customer needs.... 
    Remote work
    Project
    Hourly pay

    SaidGig

    United States
    26 days ago
  •  ...Description The Cybersecurity Red Team Analyst - Principal...  ...-functional teams in project testing phases to...  ...~Experience developing AI red team and/or AI threat...  ...capabilities. ~Ability to evaluate 3rd party AI red team...  .... Benefits ~Remote workplace type with flexible... 
    Remote work
    Project
    Full time
    Flexible hours

    Huntington National Bank

    Remote
    2 days ago
  • $125k - $190k

    About the Role FAR.AI is hiring a Technical Project Manager to be the delivery...  ...frontier AI red‑teaming programmes. You will...  ...frontier labs, AI safety organisations, technical...  ...exposure to AI evaluations, red‑teaming, or AI...  .... Location: Remote globally. We can sponsor... 
    Remote work
    Project
    Full time
    Contract work
    For contractors
    Visa sponsorship
    Shift work

    Aisafety

    Berkeley, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Project Harbor Red Teaming - Juno | Remote AI Safety Evaluation. Be the first to apply!