Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Safety Red Teamer, English & Dutch

$48 - $62 per hour

SaidGig

Role Overview

Probe conversational AI systems with adversarial inputs to identify vulnerabilities, generate high-quality safety data, and help make AI systems more robust and trustworthy. This text-based work includes reviewing model outputs that may address sensitive subjects such as bias, misinformation, and harmful behavior.

Key Responsibilities
  • Red team conversational AI models and agents using jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
  • Annotate failures, classify vulnerabilities, and flag systemic risks.
  • Apply testing taxonomies, benchmarks, and playbooks to maintain consistent evaluations.
  • Create reproducible reports, datasets, and attack cases that teams can use to strengthen AI systems.
  • Expand evaluation coverage by testing more scenarios and uncovering vulnerabilities that automated tests can miss.
Qualifications
  • Native fluency in both English and Dutch is required.
  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
  • A naturally adversarial, curious approach to testing systems and finding breaking points.
  • Experience using structured frameworks or benchmarks rather than relying on unstructured testing.
  • Ability to clearly communicate risks to technical and non-technical stakeholders.
  • Adaptability to work across projects and customer contexts.
Preferred Specialties
  • Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attacks, or model extraction.
  • Cybersecurity, including penetration testing, exploit development, or reverse engineering.
  • Socio-technical risk evaluation, including harassment or misinformation probing, abuse analysis, or conversational AI testing.
  • Creative adversarial thinking informed by psychology, acting, or writing.
Work Terms
  • Remote, hourly engagement.
  • All work is text-based.
  • Higher-sensitivity projects are optional and supported by clear guidelines and wellness resources. Topics will be communicated before any content exposure.
Compensation

$48 to $62 per hour.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the AI Safety Red Teamer, English & Dutch in United States vacancy
  • $48 - $62 per hour

     ...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production. The role...  ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and... 
    English language skills
    Hourly pay
    Remote work
    Flexible hours

    SaidGig

    United States
    1 day ago
  • $17 - $25 per hour

     ...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers. This text-...  ...25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $24 - $35 per hour

     ...Role Overview Join a red team committed to finding real-world vulnerabilities in conversational AI. You will probe models and agents with adversarial inputs, document reproducible...  ...Native-level fluency in both English and Thai, able to read, write, and communicate... 
    English language skills
    Hourly pay
    Remote work
    Shift work

    SaidGig

    United States
    1 day ago
  • $48 - $62 per hour

     ...Help strengthen conversational AI systems by probing them with adversarial inputs, identifying vulnerabilities, and producing practical safety data. This remote role focuses on text-based AI red teaming in both English and Danish. Key Responsibilities Test conversational... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    a month ago
  •  ...Fluent Language Skills Required: English & Malay. Native fluency in...  ...Mercor, we believe the safest AI is the one that’s already been...  ...attacked — by us. We are assembling a red team for this project - human...  ...production Mercor customers trust the safety of their AI because you’ve... 
    English language skills
    Remote work

    Mercor

    New York, NY
    4 days ago
  •  ...Help strengthen conversational AI systems by probing them with adversarial...  ..., and creating high-quality safety data. This text-based role...  .... Key Responsibilities Red team conversational AI models and...  ...Qualifications Native fluency in both English and Vietnamese is required.... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    a month ago
  • $17 - $25 per hour

     ...Help strengthen conversational AI by testing it from an adversarial...  ...for weaknesses, create actionable safety data, and help identify risks...  ...Qualifications Native fluency in both English and Indonesian is required. Prior experience in AI red teaming, adversarial AI work,... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    a month ago
  • Mercor seeks a remote AI red-teaming specialist fluent in English and Portuguese to assess safety of conversational models. You will simulate jailbreaks, prompt injections, and misuse scenarios while documenting results for customers. You will generate high-quality human... 
    English language skills
    Remote job

    Mercor

    New York, NY
    1 day ago
  • $48 - $52 per hour

     ...hiring proficient bilingual speakers (English & Dutch) based in Belgium to help make advanced AI models safer. You'll apply your...  ...Experience reviewing, grading, or red-teaming written or technical content. Background in trust and safety, content moderation, policy... 
    English language skills
    Dutch language skills
    Part time
    Immediate start

    Mercor

    San Francisco, CA
    11 days ago
  • $70 - $84 per hour

     ...AI Safety Red Teamer is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the... 
    Remote job
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    12 days ago
  • $61 - $65 per hour

     ...proficient bilingual speakers (English & Dutch) based in the Netherlands to help make advanced AI models safer. You'll apply your...  ...Sound judgment around scientific safety and the responsible handling of...  ...Experience reviewing, grading, or red-teaming technical content. Project... 
    English language skills
    Dutch language skills
    Part time
    Immediate start

    Mercor

    San Francisco, CA
    11 days ago
  • $70 - $84 per hour

    Gridnaut Recruiting is hiring a remote AI Safety Red Teamer contractor (pay $70-$84/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Apply through Gridnaut Recruiting for this remote AI Safety Red Teamer contractor opening.; Remote contractor... 
    Contract work
    For contractors
    Remote work
    Flexible hours

    Gridnaut Recruiting

    Remote
    4 days ago
  • $20 - $22 per hour

     ...Overview Help strengthen conversational AI systems by probing them with...  ...vulnerabilities, and producing actionable safety data. This remote, text-based role...  ...Native fluency in both English and Punjabi is required. Prior red-teaming experience in AI adversarial... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $48 - $62 per hour

     ...Role Overview Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents to surface vulnerabilities...  ...communicated before any exposure. Key Responsibilities Red team conversational AI models and agents, including... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $29 - $45 per hour

     ...performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red team data that clients can act on. Work is...  ...socio-technical probing Native fluency in English and Portuguese, with the Portuguese variety being... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $48 - $62 per hour

     ...Role Overview Probe conversational AI systems and produce reproducible adversarial...  ...vulnerabilities and improves model safety. This is a text-based red teaming role focused on discovering...  ...hour. Eligibility ~ Native fluency in English and Swedish is required.... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $17 - $25 per hour

     ...adversarial testing of conversational AI, creating reproducible attack...  .... Key Responsibilities Red team conversational AI models...  ...probing Native fluency in English and Indonesian, both required...  ...customer confidence in their AI safety because systems have been thoroughly... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $48 - $62 per hour

     ...Role Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and produce reproducible...  ...guidance. Qualifications Native fluency in English and Danish is required. Prior red teaming experience... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $17 - $25 per hour

     ...Role Overview Join a red team of human data experts who probe conversational AI models with adversarial inputs to surface...  ...red team data that improves model safety. This text-based role focuses on...  ...Native fluency in both English and Vietnamese is required Prior... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $48 - $62 per hour

    Gridnaut Recruiting is hiring a remote AI Safety Experts — English & Dutch contractor (pay $48-$62/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Remote contractor engagement supporting a leading AI lab.; Remote AI evaluation / expert contributor... 
    English language skills
    Dutch language skills
    Hourly pay
    Contract work
    For contractors
    Remote work
    Flexible hours

    Gridnaut Recruiting

    Remote
    5 days ago
  • $61 - $65 per hour

    Gridnaut Recruiting is hiring a remote Bilingual Dutch (Netherlands) STEM Expert (PhD) — AI Safety contractor (pay $61-$65/hr). Contribute to frontier AI research...  ...in the Netherlands, plus business-level written English.; Currently based in the Netherlands. This role... 
    English language skills
    Dutch language skills
    Contract work
    For contractors
    Remote work

    Gridnaut Recruiting

    Remote
    5 hours ago
  • $70 - $84 per hour

     ...Role Overview Lead adversarial evaluations of frontier AI models by designing and executing high-impact prompts and tests...  .... Document discovered vulnerabilities and contribute to safety benchmarking and red-teaming reports. Collaborate with AI researchers and... 
    Hourly pay
    Remote work
    Work visa

    SaidGig

    United States
    1 day ago
  •  ...A progressive technology company is seeking a remote AI Red Teamer to assess conversational AI models and identify vulnerabilities through red teaming exercises. Candidates should be fluent in both English and Brazilian Portuguese, with a background in cybersecurity or... 
    English language skills
    Hourly pay
    Remote work

    Crossing Hurdles

    United States
    1 day ago
  • $60 - $90 per hour

     ...technical talent with leading AI research labs. Headquartered in...  ...Position: Cybersecurity SWE — AI Safety Type: Contract...  ...technical reasoning and writing in English . Sound judgment around security...  ...reviewing, grading, or red-teaming technical content.... 
    English language skills
    Full time
    Contract work
    Summer work
    Immediate start
    Remote work

    Mercor

    Remote
    16 hours ago
  • $66.56k - $197.6k

     ...AI Red Teamer (LLM Generalist) Location: Seattle, WA (candidates must reside in the Seattle metro area or be willing to relocate prior to...  ...weaknesses, and unexpected behaviors. Your work directly supports AI safety and model robustness for leading research labs. This is a... 
    Full time
    Contract work
    Relocation

    Handshake

    Seattle, WA
    5 days ago
  •  ...Role Overview Help improve the safety of advanced AI systems by using your Norwegian language fluency...  ...fluency and business-level written English. Bachelor''s degree completed or currently...  .... Experience reviewing, grading, or red-teaming written or technical content.... 
    English language skills
    Hourly pay
    Immediate start
    Remote work

    SaidGig

    Remote
    11 days ago
  •  ...hiring proficient bilingual speakers (English & Vietnamese) to help make advanced AI models safer. You'll apply your...  ...welcome. Experience reviewing, grading, or red-teaming written or technical content. Background in trust and safety, content moderation, policy... 
    English language skills
    Part time
    Immediate start

    Mercor

    San Francisco, CA
    4 days ago
  •  ...institutions. In 2025, we started Handshake AI and built the fastest-growing AI data...  ...largest scale. About the Role As a CBRNE Red Teamer, you will evaluate whether AI models...  ...models for dangerous knowledge gaps in their safety guardrails, testing whether they can be manipulated... 
    Remote job
    Immediate start

    Dorado

    Brooklyn, NY
    3 days ago
  • $26 per hour

     ...Commitment: 10-40 hours/week Role Responsibilities Red-team conversational AI systems using jailbreaks, prompt injections...  ...-risk scenarios in alignment with defined safety guidelines. Requirements Native-level fluency in both English and Spanish. Prior experience in AI red... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    2 days ago
  • Obsidian is looking for data experts to join a red team that probes AI models with adversarial inputs. This role requires fluency in both English and Punjabi, and involves tasks such as...  ..., emphasizing the importance of AI safety. Work will be structured, with clear guidelines... 
    Remote job

    Obsidian

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Safety Red Teamer, English & Dutch. Be the first to apply!