AI Safety Expert for English and Dutch Red Teaming
SaidGig
Work as a red team human data expert attacking conversational AI in English and Dutch to surface vulnerabilities, produce reproducible adversarial test cases, and generate the human-labeled data customers use to improve model safety. The role focuses on text-only interactions and on probing areas such as bias, misinformation, and harmful behavior.
Key Responsibilities- Red team conversational AI models and agents, including jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
- Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
- Apply structure and consistency by following taxonomies, benchmarks, and playbooks during testing.
- Document findings reproducibly, producing reports, datasets, and attack cases customers can act on.
- Work with content that may touch sensitive topics; all tasks are text-based, higher-sensitivity projects are optional, and you will be given clear guidance and access to wellness resources. Topics will be communicated before exposure.
- Prior red teaming experience, such as AI adversarial work, cybersecurity, or socio-technical probing.
- Comfort with adversarial thinking, pushing systems toward failure modes, and explaining risks clearly to both technical and non-technical stakeholders.
- Structured approach to testing, using frameworks or benchmarks rather than ad hoc methods.
- Adaptability to move across projects and customers while maintaining consistent documentation and outputs.
- Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attack experience, and model extraction.
- Cybersecurity skills such as penetration testing, exploit development, or reverse engineering.
- Socio-technical risk expertise, for example harassment, disinformation probing, abuse analysis, or conversational AI testing.
- Creative probing skills, including psychology, acting, or writing for unconventional adversarial approaches.
- Uncovering vulnerabilities that automated tests miss.
- Delivering reproducible artifacts that help strengthen customer AI systems.
- Expanding evaluation coverage so more scenarios are tested and fewer surprises occur in production.
- Enabling customer trust in their AI by proactively finding and documenting adversarial failures.
- Location: Remote.
- Employment type: hourly engagement.
- All tasks are text-based. Participation in higher-sensitivity projects is optional, with pre-notification of sensitive topics and access to clear guidelines and wellness support.
- Hourly rate: 48 - 62 hourly.
- Native fluency in both English and Dutch is required for this role.
$48 - $62 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...and Jack Dorsey . Position: AI Safety Experts — English & Dutch Type: Contract Compensation:... ...Remote Role Responsibilities Red team conversational AI models and agents...English language skillsDutch language skillsContract workSummer workRemote work- ...Role Overview Lead adversarial testing of conversational AI in both English and Malay, producing reproducible red-team data that uncovers vulnerabilities related to bias, misinformation, and harmful behaviors. All tasks are text-based. Participation in higher-sensitivity...English language skillsHourly payRemote work
$24 - $35 per hour
...adversarial testing of conversational AI by probing models with crafted... ...that product and engineering teams can use to harden systems. This is text-based red teaming that can involve sensitive... ...technical probing Native fluency in English and Thai is required Curious...English language skillsHourly payRemote work- ...Role Overview This role probes conversational AI models to find real-world safety weaknesses and produce reproducible red team data that customers can act on. You will run... ...on Qualifications Native fluency in English and Finnish is required Prior red teaming...English language skillsHourly payRemote work
- ...Role Overview Probe conversational AI systems for safety and misuse vulnerabilities in both English and Norwegian, producing high-quality adversarial test cases,... ...automated tests miss. Key Responsibilities Red team conversational AI models and agents, including jailbreaks...English language skillsHourly payRemote work
- ...Overview Probe conversational AI systems with adversarial... ...vulnerabilities, produce reproducible red-team data, and deliver actionable... ...on rigorous, structured safety testing across multiple projects... ...Eligibility Native fluency in both English and Danish is required Role...English language skillsHourly payRemote work
- ...Role Overview Act as a human red teamer who probes conversational AI systems in English and Swedish to uncover safety vulnerabilities, produce high-quality adversarial inputs... ...spoken and written, is required. Prior red teaming experience, such as AI adversarial work,...English language skillsHourly payRemote work
$29 - $45 per hour
...Role Overview Attack and probe conversational AI systems in English and Portuguese to surface vulnerabilities, produce reproducible adversarial... ...signposted before exposure. Key Responsibilities Red team conversational AI models and agents, exploring jailbreaks,...English language skillsHourly payRemote work$20 - $22 per hour
...Probe conversational AI systems in English and Punjabi to uncover vulnerabilities, produce reproducible red team artifacts, and help customers improve model safety. This role focuses on adversarial testing of text outputs and delivering structured findings that engineers...English language skillsHourly payRemote work- Mercor is recruiting AI safety experts to remotely assess and strengthen AI systems. You will conduct red team activities, identify jailbreaks and misuse cases, and produce... ...safety. Ideal candidates fluently speak English and Dutch, have prior red teaming experience in AI...English language skillsDutch language skillsRemote jobContract work
- Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe conversational AI, annotate vulnerabilities, and generate high-quality attack data for safer deployments. You’ll work across...English language skillsDutch language skillsRemote job
$20 - $22 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...and Jack Dorsey . Position: AI Safety Experts — English & Assamese Type: Contract... ...Remote Role Responsibilities Red team conversational AI models and agents,...English language skillsContract workSummer workRemote work- ...Language Skills Required English & Punjabi. Native fluency... ...Mercor, we believe the safest AI is the one that’s already... ...us. We are assembling a red team for this project - human data experts who probe AI models with... ...customers trust the safety of their AI because you’ve...English language skillsRemote work
$48 - $62 per hour
...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production. The role... ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and...English language skillsHourly payRemote workFlexible hours$17 - $25 per hour
...adversarial testing of conversational AI, creating reproducible attack... .... Key Responsibilities Red team conversational AI models and... ...probing Native fluency in English and Indonesian, both required... ...confidence in their AI safety because systems have been thoroughly...English language skillsHourly payRemote work$48 - $62 per hour
...Role Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and produce reproducible... ...guidance. Qualifications Native fluency in English and Danish is required. Prior red teaming experience...English language skillsHourly payRemote work$29 - $45 per hour
...performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red team data that clients can act on. Work is remote... ...-technical probing Native fluency in English and Portuguese, with the Portuguese variety...English language skillsHourly payRemote work$17 - $25 per hour
...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers. This text-... ...25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work...English language skillsHourly payRemote work$48 - $62 per hour
...Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents... ...produce human-labeled data that helps teams make AI systems safer. All work is text-... ...exposure. Key Responsibilities Red team conversational AI models and agents...English language skillsHourly payRemote work$20 - $22 per hour
...Role Overview Join a remote red team focused on probing conversational AI to find real-world vulnerabilities and... ...adversarial datasets that improve model safety. You will create and document... ...Eligibility ~ Native fluency in both English and Punjabi is required for this...English language skillsHourly payRemote work$48 - $62 per hour
...Role Overview Red team conversational AI and generate high-quality human data that exposes vulnerabilities... ...Customers gain confidence in the safety of their AI because adversarial... ...Eligibility Native fluency in English and Dutch is required, both spoken and written...English language skillsDutch language skillsHourly payRemote work$48 - $62 per hour
...AI Safety Experts — English & Dutch is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team...English language skillsDutch language skillsRemote jobFor contractors10 hours per week- ...ActiveFence) is a leading trust, safety, and security company. Just... ...hole into the emerging world of AI and focus on safeguarding these... ...work for the \"best of the best\" red-teamers in the industry. Work... ...As one of Alice's Security Red Team Specialists, you'll focus on...Freelance
- Mercor is seeking a remote Senior Red Team AI Specialist to test conversational agents and AI models against adversarial inputs. You will annotate vulnerabilities, surface systemic risks, and produce reproducible attack cases to help customers strengthen their AI systems...Remote work
- Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human...Remote job
- Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate actionable data and reports to help customers harden their AI. Ideal candidates...Remote job
- ...Overview Conduct adversarial testing of conversational AI in English and Assamese to find and document safety failures, generate reproducible attack cases, and... ...to you before exposure. Key Responsibilities Red team conversational AI models and agents, including...English language skillsHourly payRemote work
- ...Role Overview This role joins a red team project that probes conversational AI models and agents to uncover vulnerabilities before they reach production... ...cybersecurity, or socio-technical probing. Native fluency in English and Odia is required, including the ability to read...English language skillsHourly payRemote work
$20 - $22 per hour
...adversarial testing of conversational AI in English and Bengali, producing reproducible attack... ..., remote role focused on human-driven red teaming and vulnerability discovery. Key... ...production. Increase client trust in the safety and robustness of their AI through...English language skillsHourly payRemote work- ...Role Overview Probe conversational AI systems to find vulnerabilities and produce reproducible... ...their models. This is a text-based red teaming role focused on exposing bias,... ...-technical probing. Native fluency in English and Vietnamese, with strong written and verbal...English language skillsHourly payRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Expert for English and Dutch Red Teaming. Be the first to apply!
- fruit expert United States
- subject matter expert United States
- expert data analyst United States
- guest service support expert United States
- expert systems engineer United States
- technology expert United States
- fulfillment expert United States
- subject matter expert senior United States
- subject matter expert work from home United States
- sql expert United States


