Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Safety Expert for English and Vietnamese

SaidGig

Role Overview

Probe conversational AI systems to find vulnerabilities and produce reproducible adversarial data that customers can use to harden their models. This is a text-based red teaming role focused on exposing bias, misinformation, harmful behaviors, jailbreaks, and other misuse cases, then documenting findings in structured, actionable artifacts. Participation in higher-sensitivity reviews is optional and supported with clear guidelines and wellness resources, and topics will be communicated before any exposure.

Key Responsibilities
  • Red team conversational AI models and agents by designing and executing adversarial interactions, including jailbreaks, prompt injection, misuse cases, bias exploitation, and multi-turn manipulation.
  • Generate high-quality human data: annotate model failures, classify vulnerabilities, and flag systemic risks.
  • Follow established taxonomies, benchmarks, and playbooks to keep testing consistent and reproducible.
  • Document results clearly, delivering reports, datasets, and attack case examples customers can act on.
Qualifications
  • Prior red teaming experience, such as AI adversarial work, cybersecurity testing, or socio-technical probing.
  • Native fluency in English and Vietnamese, with strong written and verbal communication in both languages.
  • Curious and adversarial mindset, with an instinct to push systems to failure points rather than accepting surface behavior.
  • Structured approach to testing, using frameworks and benchmarks rather than ad hoc methods.
  • Able to explain technical risks to both technical and non-technical stakeholders.
  • Adaptable, comfortable moving across projects and customer contexts.
  • Nice-to-have specialties: adversarial machine learning (jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction), cybersecurity skills (penetration testing, exploit development, reverse engineering), socio-technical risk analysis (harassment, disinformation, conversational abuse testing), creative probing skills (psychology, acting, or creative writing for adversarial scenarios).
Work Terms
  • Work location: Remote.
  • Employment type: hourly.
  • All tasks are text-based. Higher-sensitivity assignments are optional, accompanied by clear guidelines and wellness support, and you will be informed of sensitive topics before any exposure.
  • Project assignments may vary by customer and engagement; you should be prepared to move between projects and follow customer-specific playbooks when provided.
Compensation
  • Pay range: 17 - 25 hourly.
Eligibility
  • Native fluency in both English and Vietnamese is required.
  • Remote work is permitted; confirm you can work remotely from your location if selected.
Vacancy posted 18 days ago
Similar jobs that could be interesting for youBased on the AI Safety Expert for English and Vietnamese in United States vacancy
  • $17 - $25 per hour

     ...Overview Join a red team of human data experts who probe conversational AI models with adversarial inputs to...  ...red team data that improves model safety. This text-based role focuses on...  ...Native fluency in both English and Vietnamese is required Prior red teaming experience... 
    Vietnamese language
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    14 hours ago
  • $17 - $25 per hour

     ...AI Safety Experts — English & Vietnamese is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety... 
    Vietnamese language
    English language skills
    Remote job
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    3 days ago
  • $17 - $25 per hour

     ...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our...  ..., Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Vietnamese Type: Contract Compensation: $17–$25/hour Location... 
    Vietnamese language
    English language skills
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    8 days ago
  •  ...Role Overview Conduct adversarial testing of conversational AI in English and Assamese to find and document safety failures, generate reproducible attack cases, and produce high-quality human data that helps customers harden their systems. The work is text-based and... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  •  ...Role Overview Probe conversational AI systems for safety and misuse vulnerabilities in both English and Norwegian, producing high-quality adversarial test cases, annotations, and reproducible reports that help customers reduce bias, misinformation, and harmful behaviors... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  • $24 - $35 per hour

     ...Role Overview Perform adversarial testing of conversational AI by probing models with crafted inputs, surfacing vulnerabilities,...  ...cybersecurity, or socio-technical probing Native fluency in English and Thai is required Curious and adversarial mindset, with... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  •  ...Role Overview This role probes conversational AI models to find real-world safety weaknesses and produce reproducible red team data that customers...  ...reproduce and act on Qualifications Native fluency in English and Finnish is required Prior red teaming experience,... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  •  ...Role Overview Lead adversarial testing of conversational AI in both English and Malay, producing reproducible red-team data that uncovers vulnerabilities related to bias, misinformation, and harmful behaviors. All tasks are text-based. Participation in higher-sensitivity... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    13 days ago
  •  ...This role joins a red team project that probes conversational AI models and agents to uncover vulnerabilities before they reach production...  ...cybersecurity, or socio-technical probing. Native fluency in English and Odia is required, including the ability to read and write... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  • $20 - $22 per hour

     ...Role Overview Lead adversarial testing of conversational AI in English and Bengali, producing reproducible attack cases, datasets, and reports...  ...occur in production. Increase client trust in the safety and robustness of their AI through adversarial verification.... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  •  ...Role Overview Probe conversational AI systems with adversarial inputs to expose vulnerabilities...  ...role focused on rigorous, structured safety testing across multiple projects and...  ...Eligibility Native fluency in both English and Danish is required Role is remote;... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  •  ...Role Overview Act as a human red teamer who probes conversational AI systems in English and Swedish to uncover safety vulnerabilities, produce high-quality adversarial inputs, and deliver reproducible artifacts customers can act on. The work is fully text-based, focuses... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  • $29 - $45 per hour

     ...Role Overview Attack and probe conversational AI systems in English and Portuguese to surface vulnerabilities, produce reproducible adversarial data, and help make deployed models safer. This text-based role focuses on adversarial testing of model outputs that involve... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  • $20 - $22 per hour

     ...Probe conversational AI systems in English and Punjabi to uncover vulnerabilities, produce reproducible red team artifacts, and help customers improve model safety. This role focuses on adversarial testing of text outputs and delivering structured findings that engineers... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  • $29 - $45 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type: Contract Compensation: $29-$45... 
    English language skills
    Contract work
    Summer work
    Remote work

    Mercor Inc

    San Francisco, CA
    3 days ago
  • Mercor is building a remote AI red team to probe and test conversational models for vulnerabilities, bias, and safety gaps. This role emphasizes reproducible data and actionable findings across projects and customers. You bring prior red teaming experience in AI or cybersecurity... 
    Vietnamese language
    English language skills
    Remote job

    Obsidian

    New York, NY
    3 days ago
  • $65 - $70 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Jack Dorsey . Position: Biology Expert (PhD) — AI Safety Type: Contract Compensation:...  ...Strong scientific reasoning and writing in English . Sound judgment around biosecurity... 
    English language skills
    Contract work
    Summer work
    Immediate start
    Remote work

    Mercor

    New York, NY
    a month ago
  • $20 - $22 per hour

     ...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Assamese Type: Contract Compensation: $20–$22/... 
    English language skills
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    3 days ago
  •  ...PhD‑level biologists to help make advanced AI models safer. You'll apply your scientific...  ...on the workflow. Responsibilities Write expert‑level prompts across specialized life‑science...  ...scientific reasoning and writing in English. Sound judgment around biosecurity and the... 
    English language skills
    Part time
    Immediate start

    Obsidian

    New York, NY
    1 day ago
  •  ...Fluent Language Skills Required English & Punjabi. Native fluency in...  ...Mercor, we believe the safest AI is the one that’s already been...  ...for this project - human data experts who probe AI models with adversarial...  ...Mercor customers trust the safety of their AI because you’ve... 
    English language skills
    Remote work

    Obsidian

    San Francisco, CA
    3 days ago
  • $48 - $62 per hour

     ...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production...  ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and adversarial... 
    English language skills
    Hourly pay
    Remote work
    Flexible hours

    SaidGig

    United States
    2 days ago
  • $125k - $175k

     ...fusion, artificial intelligence (AI), machine learning (ML), and...  ...(AR).QinetiQ US’s dedicated experts in defense, aerospace, security...  ...US means being central to the safety and security of the world around...  ..., including Khmer, Thai, Vietnamese, or Arabic.Test Facility Design... 
    Vietnamese language
    Work experience placement
    Interim role
    Remote work
    Overseas

    QinetiQ

    Alexandria, VA
    4 days ago
  • $20 - $22 per hour

     ...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce reproducible attack data...  ...can act on Qualifications Native fluency in both English and Bengali, spoken and written, is required Prior red... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    14 hours ago
  • $20 - $22 per hour

     ...Role Overview Lead adversarial testing of conversational AI in both English and Assamese, probing model behavior to find jailbreaks, misinformation, bias, and other harmful or unsafe outputs. This is a remote, text-based role focused on producing reproducible red team... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    14 hours ago
  • $20 - $22 per hour

     ...Overview This role performs adversarial testing of conversational AI systems, probing models for vulnerabilities and producing...  ...Compensation 20 - 22 hourly Eligibility Native fluency in English and Odia is required Ability to perform remote, text based work... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    14 hours ago
  • $48 - $62 per hour

     ...Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and...  ...adherence to provided guidance. Qualifications Native fluency in English and Danish is required. Prior red teaming experience, such as... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago
  • $29 - $45 per hour

     ...Overview This role performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red...  ...team work, or socio-technical probing Native fluency in English and Portuguese, with the Portuguese variety being global... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    14 hours ago
  • $17 - $25 per hour

     ...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers...  ...: 17 - 25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    14 hours ago
  • $48 - $62 per hour

     ...Role Overview Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents to surface vulnerabilities, generate reproducible attack cases, and produce human-labeled data that helps teams make AI systems safer. All work is text... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    5 days ago
  • $17 - $25 per hour

     ...performs adversarial testing of conversational AI, creating reproducible attack cases,...  ...socio-technical probing Native fluency in English and Indonesian, both required Curious...  ...Increasing customer confidence in their AI safety because systems have been thoroughly... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Safety Expert for English and Vietnamese. Be the first to apply!