AI Safety Expert for English and Thai Red Teaming
$24 - $35 per hourSaidGig
Perform adversarial testing of conversational AI by probing models with crafted inputs, surfacing vulnerabilities, and producing reproducible attack cases and human-labeled data that product and engineering teams can use to harden systems. This is text-based red teaming that can involve sensitive topics such as bias, misinformation, or harmful behaviors. Participation in higher-sensitivity reviews is optional and supported with clear guidelines and wellness resources, and topics will be communicated before exposure.
Key Responsibilities- Red team conversational AI models and agents, including jailbreaks, prompt injection, misuse cases, bias exploitation, and multi-turn manipulation
- Generate high-quality human data: annotate failures, classify vulnerabilities, and flag systemic risks
- Apply structure to testing by following taxonomies, benchmarks, and playbooks to maintain consistency
- Document findings reproducibly by producing reports, datasets, and attack cases that customers or teams can act on
- Prior red teaming experience, such as AI adversarial work, cybersecurity, or socio-technical probing
- Native fluency in English and Thai is required
- Curious and adversarial mindset, with an instinct to push systems to breaking points
- Structured approach, using frameworks or benchmarks rather than ad hoc methods
- Strong communication skills, able to explain risks to technical and non-technical stakeholders
- Adaptable, able to move across projects and customer contexts
- Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attack techniques, and model extraction
- Cybersecurity experience such as penetration testing, exploit development, or reverse engineering
- Socio-technical risk expertise, including harassment and disinformation probing, abuse analysis, and conversational AI testing
- Creative probing skills from psychology, acting, or creative writing to design unconventional adversarial inputs
- Discover vulnerabilities that automated tests miss
- Deliver reproducible artifacts that directly strengthen customer AI systems
- Expand evaluation coverage so more scenarios are tested and fewer surprises arise in production
- Increase customer trust in deployed AI by identifying and mitigating real-world attack vectors
- Remote engagement, all work is text-based
- Hourly employment
- Participation in higher-sensitivity projects is optional and accompanied by clear guidelines and wellness resources; topics will be disclosed before exposure
- $24.00 to $35.00 per hour
- Native fluency in both English and Thai is required
- Work is remote; candidates must be able to work remotely from their location
- ...Role Overview Probe conversational AI systems for safety and misuse vulnerabilities in both English and Norwegian, producing high-quality adversarial test cases,... ...automated tests miss. Key Responsibilities Red team conversational AI models and agents, including jailbreaks...English language skillsHourly payRemote work
- ...Role Overview This role probes conversational AI models to find real-world safety weaknesses and produce reproducible red team data that customers can act on. You will run... ...on Qualifications Native fluency in English and Finnish is required Prior red teaming...English language skillsHourly payRemote work
- ...Role Overview Lead adversarial testing of conversational AI in both English and Malay, producing reproducible red-team data that uncovers vulnerabilities related to bias, misinformation, and harmful behaviors. All tasks are text-based. Participation in higher-sensitivity...English language skillsHourly payRemote work
- ...Overview Probe conversational AI systems with adversarial... ...vulnerabilities, produce reproducible red-team data, and deliver actionable... ...on rigorous, structured safety testing across multiple projects... ...Eligibility Native fluency in both English and Danish is required Role...English language skillsHourly payRemote work
- ...Role Overview Act as a human red teamer who probes conversational AI systems in English and Swedish to uncover safety vulnerabilities, produce high-quality adversarial inputs... ...spoken and written, is required. Prior red teaming experience, such as AI adversarial work,...English language skillsHourly payRemote work
$29 - $45 per hour
...Role Overview Attack and probe conversational AI systems in English and Portuguese to surface vulnerabilities, produce reproducible adversarial... ...signposted before exposure. Key Responsibilities Red team conversational AI models and agents, exploring jailbreaks,...English language skillsHourly payRemote work- ...Role Overview Work as a red team human data expert attacking conversational AI in English and Dutch to surface vulnerabilities, produce reproducible adversarial test... ...-labeled data customers use to improve model safety. The role focuses on text-only interactions and on...English language skillsHourly payRemote work
$20 - $22 per hour
...Probe conversational AI systems in English and Punjabi to uncover vulnerabilities, produce reproducible red team artifacts, and help customers improve model safety. This role focuses on adversarial testing of text outputs and delivering structured findings that engineers...English language skillsHourly payRemote work$20 - $22 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...and Jack Dorsey . Position: AI Safety Experts — English & Assamese Type: Contract... ...Remote Role Responsibilities Red team conversational AI models and agents,...English language skillsContract workSummer workRemote work- ...Language Skills Required English & Punjabi. Native fluency... ...Mercor, we believe the safest AI is the one that’s already... ...us. We are assembling a red team for this project - human data experts who probe AI models with... ...customers trust the safety of their AI because you’ve...English language skillsRemote work
$48 - $62 per hour
...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production. The role... ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and...English language skillsHourly payRemote workFlexible hours$48 - $62 per hour
...Role Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and produce reproducible... ...guidance. Qualifications Native fluency in English and Danish is required. Prior red teaming experience...English language skillsHourly payRemote work$29 - $45 per hour
...performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red team data that clients can act on. Work is remote... ...-technical probing Native fluency in English and Portuguese, with the Portuguese variety...English language skillsHourly payRemote work$17 - $25 per hour
...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers. This text-... ...25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work...English language skillsHourly payRemote work$17 - $25 per hour
...adversarial testing of conversational AI, creating reproducible attack... .... Key Responsibilities Red team conversational AI models and... ...probing Native fluency in English and Indonesian, both required... ...confidence in their AI safety because systems have been thoroughly...English language skillsHourly payRemote work$48 - $62 per hour
...Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents... ...produce human-labeled data that helps teams make AI systems safer. All work is text-... ...exposure. Key Responsibilities Red team conversational AI models and agents...English language skillsHourly payRemote work$20 - $22 per hour
...Role Overview Join a remote red team focused on probing conversational AI to find real-world vulnerabilities and... ...adversarial datasets that improve model safety. You will create and document... ...Eligibility ~ Native fluency in both English and Punjabi is required for this...English language skillsHourly payRemote work$20 - $22 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...and Jack Dorsey . Position: AI Safety Experts — English & Punjabi Type: Contract... ...Remote Role Responsibilities Red team conversational AI models and agents....English language skillsContract workSummer workRemote work$24 - $35 per hour
...AI Safety Experts — English & Thai is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team...English language skillsThai languageRemote jobFor contractors10 hours per week$24 - $35 per hour
...Role Overview Join a red team committed to finding real-world vulnerabilities in conversational AI. You will probe models and agents with adversarial inputs, document... ...Native-level fluency in both English and Thai, able to read, write, and communicate at a...English language skillsThai languageHourly payRemote workShift work- Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human...Remote job
- Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate actionable data and reports to help customers harden their AI. Ideal candidates...Remote job
- ...ActiveFence) is a leading trust, safety, and security company. Just... ...hole into the emerging world of AI and focus on safeguarding these... ...work for the \"best of the best\" red-teamers in the industry. Work... ...As one of Alice's Security Red Team Specialists, you'll focus on...Freelance
- Mercor is seeking a remote Senior Red Team AI Specialist to test conversational agents and AI models against adversarial inputs. You will annotate vulnerabilities, surface systemic risks, and produce reproducible attack cases to help customers strengthen their AI systems...Remote work
- ...Overview Conduct adversarial testing of conversational AI in English and Assamese to find and document safety failures, generate reproducible attack cases, and... ...to you before exposure. Key Responsibilities Red team conversational AI models and agents, including...English language skillsHourly payRemote work
- ...Role Overview This role joins a red team project that probes conversational AI models and agents to uncover vulnerabilities before they reach production... ...cybersecurity, or socio-technical probing. Native fluency in English and Odia is required, including the ability to read...English language skillsHourly payRemote work
$20 - $22 per hour
...adversarial testing of conversational AI in English and Bengali, producing reproducible attack... ..., remote role focused on human-driven red teaming and vulnerability discovery. Key... ...production. Increase client trust in the safety and robustness of their AI through...English language skillsHourly payRemote work- ...Role Overview Probe conversational AI systems to find vulnerabilities and produce reproducible... ...their models. This is a text-based red teaming role focused on exposing bias,... ...-technical probing. Native fluency in English and Vietnamese, with strong written and verbal...English language skillsHourly payRemote work
$35 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ..., and Jack Dorsey . Position AI Safety Experts — English & Thai Type Contract Compensation $24-$35/hour... ...Remote Role Responsibilities Red team conversational AI models and agents to...English language skillsThai languageRemote jobContract workSummer work$26 per hour
...consulting firm is seeking a remote contractor to engage in AI red teaming activities. The ideal candidate will work on evaluating conversational... ...and classification of vulnerabilities. Fluency in English and Spanish is required, along with prior experience in adversarial...Remote jobHourly payFor contractors10 hours per weekFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Expert for English and Thai Red Teaming. Be the first to apply!
- fruit expert United States
- subject matter expert United States
- expert data analyst United States
- guest service support expert United States
- expert systems engineer United States
- technology expert United States
- fulfillment expert United States
- subject matter expert senior United States
- subject matter expert work from home United States
- sql expert United States


