Project Harbor Red Teaming - Juno | Remote AI Safety Evaluation
OneForma is looking for native Arabic linguists based in Switzerland , Saudi Arabia or the United Arab Emirates to join an exciting remote AI research project focused on AI safety evaluation and quality assurance.
Participants will review AI-generated responses and perform structured linguistic assessments to identify harmful content, safety concerns, and policy violations. The work involves evaluating large language model (LLM) outputs, classifying potential risks, documenting findings, and validating improvements after issues have been addressed.
This is a flexible, remote opportunity for experienced Arabic language professionals who are interested in helping improve the safety, reliability, and accuracy of next-generation AI systems.
Purpose:
The purpose of this project is to improve the safety and quality of AI chatbot systems by leveraging human expertise to evaluate model responses, identify potential risks, and provide high-quality feedback that supports ongoing AI development.
Main Requirements:
- Must be a native Arabic speaker based in Switzerland , Saudi Arabia or the United Arab Emirates .
- Linguistic background or professional language experience is preferred.
- Must be familiar with Apple products and AppleCare content.
- Experience evaluating Generative AI (GenAI) models or large language model (LLM) outputs is preferred.
- Must own a personal MacBook running the latest version of macOS Tahoe .
- Device must be personal (not shared) and have administrator privileges for software installation.
- Must confirm whether you have an AppleConnect account.
- Must be available to receive immediate, ad hoc task assignments.
- Must demonstrate professionalism, attention to detail, and the ability to follow evaluation guidelines.
- Strong analytical and written communication skills are required.
Other Important Information:
- This is a fully remote project.
- Participants will review AI-generated conversations for safety, quality, and policy compliance.
- Tasks include identifying harmful or unsafe content, classifying issues by risk level, documenting findings, and validating fixes after remediation.
- The use of AI tools, automation, or speech-to-text software is strictly prohibited.
- Completion of onboarding and certification is mandatory before production work begins.
- Direct communication with the client point of contact will be required throughout the project.
Help shape the future of trustworthy AI by using your language expertise to make AI systems safer, more accurate, and more reliable for users worldwide.
- ...AI Safety and Red-Team Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios,... ...specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed...Remote jobProjectHourly payFor contractors10 hours per week
$48 - $62 per hour
...conversational AI systems by probing... ...-generated safety data. This text... ...higher-sensitivity projects is optional,... ...conduct consistent evaluations. Create... ...attack cases that teams can use to... ...experience in AI red teaming, adversarial... ...Work Terms Remote, hourly...Remote workProjectHourly pay- Location Remote Fluent Language Skills Required English... ...believe the safest AI is the one that’s... ...We are assembling a red team for this project - human data experts... ...customer AI systems Evaluation coverage expands: more... ...customers trust the safety of their AI because you...Remote workProject
$26 per hour
...: $26/hour Location: Remote Commitment: 10-40 hours... ...Responsibilities Red-team conversational AI systems using jailbreaks... ...test cases. Evaluate AI outputs across sensitive... ...alignment with defined safety guidelines. Requirements... ...in a remote, project-based environment. #J...Remote jobProjectHourly payContract work- Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate... .... Remote collaboration across projects and clients is common. #J-18808-Ljbffr...Remote jobProject
- Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human...Remote jobProject
- Mercor is seeking a remote Red Team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate actionable red team... ...established taxonomies and playbooks. The role offers multi-project opportunities and requires strong communication of...Remote jobProject
- Mercor is building a remote red team to probe AI safety. You will test conversational AI models and agents, including jailbreaks, prompt injections... ...emphasizes clear communication, structured methods, and adaptability across projects and clients. #J-18808-Ljbffr MercorRemote jobProject
- Mercor is seeking a remote red-team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate high-quality... .... You’ll document results and collaborate across projects for robust safety testing. This role focuses on reproducible...Remote jobProject
- Mercor is building a red team to probe AI models with adversarial inputs, surface vulnerabilities, and generate red team data that strengthens safety for customers. This is a text-based, higher-sensitivity project where participation is optional and guided by clear guidelines...Remote jobProject
- Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe... ...attack data for safer deployments. You’ll work across projects with structured test plans, clear documentation, and reproducible...Remote jobProject
$20 - $22 per hour
...strengthen conversational AI systems by probing them as an adversary. This remote, text-based role... ...generating high-quality red-team data, and creating reproducible... ...that improve AI safety and reliability. Key... ...Comfort adapting across projects and customers. Preferred...Remote workProjectHourly pay- Mercor is seeking an AI Safety Expert to join a remote contract team. You will red team conversational AI, identify jailbreaks, and analyze biases to improve model... ...and the ability to work asynchronously across projects. #J-18808-Ljbffr United States Digital Space LLCRemote jobProjectContract work
$29 - $45 per hour
...Attack and probe conversational AI systems in English and... ...exposure. Key Responsibilities Red team conversational AI models and... ...Adaptable, comfortable moving between projects and different customer... ...adversarial thinking. Work Terms Remote, text-based engagement....Remote workProjectHourly pay- Mercor is building a red-team for AI safety. The work involves reviewing AI outputs touching bias,... ...and participation in higher-sensitivity projects is optional with clear guidelines and... ...customers strengthen AI safety. This is a remote role based in the United States, and...Remote jobProject
- Mercor is building a remote AI red team to probe and test conversational models for vulnerabilities, bias, and safety gaps. This role emphasizes reproducible data and actionable findings across projects and customers. You bring prior red teaming experience in AI or cybersecurity...Remote jobProject
$29 - $45 per hour
...testing of conversational AI to find safety vulnerabilities and... ..., human-generated red team data that clients can act on. Work is remote and text-based. The role... ...moving across varied projects and client needs Nice... ...AI systems Expand evaluation coverage so fewer unexpected...Remote workProjectHourly pay$48 - $62 per hour
...of conversational AI in both English and... ...labeled data that helps teams make AI systems... ...Responsibilities Red team conversational... ...to move across projects and customer contexts... ...systems Expanding evaluation coverage so fewer... ...risks Work Terms Remote engagement, all...Remote workProjectHourly pay- Mercor is assembling a remote red team to test and strengthen AI systems through adversarial inputs. You will annotate failures, classify vulnerabilities... ...and a structured approach, thriving across different projects and customers while delivering high-quality data and...Remote jobProject
$250k - $400k
You'll manage our growing red-team, own delivery for frontier lab customers... ...Our mission is to automate AI safety , to pave the way for a... ...we're looking for People and project management : you can hire,... ...alongside us there. We are open to remote for the right candidate....Remote workProjectVisa sponsorshipShift work- ...Safety Evaluation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure... ...role-specific routing and review. Final project scope, schedule, and contractor terms are...Remote jobProjectHourly payFor contractors10 hours per week
$20 - $22 per hour
...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce... ...in higher-sensitivity projects is optional, supported by clear guidelines... ...scenarios Work Terms Remote, text based engagement on an hourly...Remote workProjectHourly pay$20 - $22 per hour
...technical talent with leading AI research labs.... ...Dorsey . Position: AI Safety Experts — English & Assamese... ...$22/hour Location: Remote Role Responsibilities Red team conversational AI models... ...Adaptability to move across projects and customers....Remote workProjectContract workSummer work$24 - $35 per hour
...technical talent with leading AI research labs.... ...Dorsey . Position: AI Safety Experts — English & Thai... ...$35/hour Location: Remote Role Responsibilities Red team conversational AI models... ...Participate in higher-sensitivity projects with clear guidelines and...Remote workProjectContract workSummer work$65 per hour
A leading AI consulting firm is seeking an AI Tutor specialized in Coding.... ...This part-time freelance role involves evaluating AI models, creating test cases, and... ..., this position allows you to work remotely on challenging AI projects that enhance your expertise. Compensation...Remote jobProjectPart timeFreelanceFlexible hours$26 per hour
A tech consulting firm is seeking a remote contractor to engage in AI red teaming activities. The ideal candidate will work on evaluating conversational AI systems and generating human data through careful documentation and classification of vulnerabilities. Fluency in...Remote jobHourly payFor contractors10 hours per weekFlexible hours- ...AI Red Team Engineer We're looking for an AI Red Team Engineer to break LLM-powered systems... ...write it up. White Circle is an AI Safety company building the safety, reliability... ...tool/function calling, and LLM-as-judge evaluation. Familiarity with OWASP LLM Top 10, OWASP...Remote workLocal area
$17 - $25 per hour
...strengthen conversational AI systems by probing them... ...producing actionable safety data. This remote role focuses on text-based evaluation of AI outputs and... ...Key Responsibilities Red team conversational AI models... ...work across different projects and customer needs....Remote workProjectHourly pay- ...Description The Cybersecurity Red Team Analyst - Principal... ...-functional teams in project testing phases to... ...~Experience developing AI red team and/or AI threat... ...capabilities. ~Ability to evaluate 3rd party AI red team... .... Benefits ~Remote workplace type with flexible...Remote workProjectFull timeFlexible hours
$125k - $190k
About the Role FAR.AI is hiring a Technical Project Manager to be the delivery... ...frontier AI red‑teaming programmes. You will... ...frontier labs, AI safety organisations, technical... ...exposure to AI evaluations, red‑teaming, or AI... .... Location: Remote globally. We can sponsor...Remote workProjectFull timeContract workFor contractorsVisa sponsorshipShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Project Harbor Red Teaming - Juno | Remote AI Safety Evaluation. Be the first to apply!
- bechtel project United States
- implementation project manager United States
- implementation project manager remote United States
- wounded warrior project United States
- projects work from home United States
- architectural project designer United States
- project integrator United States
- project finance United States
- special projects United States
- project manager agile United States



