AI Safety Expert for English and Vietnamese
SaidGig
Probe conversational AI systems to find vulnerabilities and produce reproducible adversarial data that customers can use to harden their models. This is a text-based red teaming role focused on exposing bias, misinformation, harmful behaviors, jailbreaks, and other misuse cases, then documenting findings in structured, actionable artifacts. Participation in higher-sensitivity reviews is optional and supported with clear guidelines and wellness resources, and topics will be communicated before any exposure.
Key Responsibilities- Red team conversational AI models and agents by designing and executing adversarial interactions, including jailbreaks, prompt injection, misuse cases, bias exploitation, and multi-turn manipulation.
- Generate high-quality human data: annotate model failures, classify vulnerabilities, and flag systemic risks.
- Follow established taxonomies, benchmarks, and playbooks to keep testing consistent and reproducible.
- Document results clearly, delivering reports, datasets, and attack case examples customers can act on.
- Prior red teaming experience, such as AI adversarial work, cybersecurity testing, or socio-technical probing.
- Native fluency in English and Vietnamese, with strong written and verbal communication in both languages.
- Curious and adversarial mindset, with an instinct to push systems to failure points rather than accepting surface behavior.
- Structured approach to testing, using frameworks and benchmarks rather than ad hoc methods.
- Able to explain technical risks to both technical and non-technical stakeholders.
- Adaptable, comfortable moving across projects and customer contexts.
- Nice-to-have specialties: adversarial machine learning (jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction), cybersecurity skills (penetration testing, exploit development, reverse engineering), socio-technical risk analysis (harassment, disinformation, conversational abuse testing), creative probing skills (psychology, acting, or creative writing for adversarial scenarios).
- Work location: Remote.
- Employment type: hourly.
- All tasks are text-based. Higher-sensitivity assignments are optional, accompanied by clear guidelines and wellness support, and you will be informed of sensitive topics before any exposure.
- Project assignments may vary by customer and engagement; you should be prepared to move between projects and follow customer-specific playbooks when provided.
- Pay range: 17 - 25 hourly.
- Native fluency in both English and Vietnamese is required.
- Remote work is permitted; confirm you can work remotely from your location if selected.
$17 - $25 per hour
...Overview Join a red team of human data experts who probe conversational AI models with adversarial inputs to... ...red team data that improves model safety. This text-based role focuses on... ...Native fluency in both English and Vietnamese is required Prior red teaming experience...Vietnamese languageEnglish language skillsHourly payRemote work$17 - $25 per hour
...AI Safety Experts — English & Vietnamese is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety...Vietnamese languageEnglish language skillsRemote jobFor contractors10 hours per week$17 - $25 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our... ..., Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Vietnamese Type: Contract Compensation: $17–$25/hour Location...Vietnamese languageEnglish language skillsContract workSummer workRemote work- ...Role Overview Conduct adversarial testing of conversational AI in English and Assamese to find and document safety failures, generate reproducible attack cases, and produce high-quality human data that helps customers harden their systems. The work is text-based and...English language skillsHourly payRemote work
- ...Role Overview Probe conversational AI systems for safety and misuse vulnerabilities in both English and Norwegian, producing high-quality adversarial test cases, annotations, and reproducible reports that help customers reduce bias, misinformation, and harmful behaviors...English language skillsHourly payRemote work
$24 - $35 per hour
...Role Overview Perform adversarial testing of conversational AI by probing models with crafted inputs, surfacing vulnerabilities,... ...cybersecurity, or socio-technical probing Native fluency in English and Thai is required Curious and adversarial mindset, with...English language skillsHourly payRemote work- ...Role Overview This role probes conversational AI models to find real-world safety weaknesses and produce reproducible red team data that customers... ...reproduce and act on Qualifications Native fluency in English and Finnish is required Prior red teaming experience,...English language skillsHourly payRemote work
- ...Role Overview Lead adversarial testing of conversational AI in both English and Malay, producing reproducible red-team data that uncovers vulnerabilities related to bias, misinformation, and harmful behaviors. All tasks are text-based. Participation in higher-sensitivity...English language skillsHourly payRemote work
- ...This role joins a red team project that probes conversational AI models and agents to uncover vulnerabilities before they reach production... ...cybersecurity, or socio-technical probing. Native fluency in English and Odia is required, including the ability to read and write...English language skillsHourly payRemote work
$20 - $22 per hour
...Role Overview Lead adversarial testing of conversational AI in English and Bengali, producing reproducible attack cases, datasets, and reports... ...occur in production. Increase client trust in the safety and robustness of their AI through adversarial verification....English language skillsHourly payRemote work- ...Role Overview Probe conversational AI systems with adversarial inputs to expose vulnerabilities... ...role focused on rigorous, structured safety testing across multiple projects and... ...Eligibility Native fluency in both English and Danish is required Role is remote;...English language skillsHourly payRemote work
- ...Role Overview Act as a human red teamer who probes conversational AI systems in English and Swedish to uncover safety vulnerabilities, produce high-quality adversarial inputs, and deliver reproducible artifacts customers can act on. The work is fully text-based, focuses...English language skillsHourly payRemote work
$29 - $45 per hour
...Role Overview Attack and probe conversational AI systems in English and Portuguese to surface vulnerabilities, produce reproducible adversarial data, and help make deployed models safer. This text-based role focuses on adversarial testing of model outputs that involve...English language skillsHourly payRemote work$20 - $22 per hour
...Probe conversational AI systems in English and Punjabi to uncover vulnerabilities, produce reproducible red team artifacts, and help customers improve model safety. This role focuses on adversarial testing of text outputs and delivering structured findings that engineers...English language skillsHourly payRemote work$29 - $45 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type: Contract Compensation: $29-$45...English language skillsContract workSummer workRemote work- Mercor is building a remote AI red team to probe and test conversational models for vulnerabilities, bias, and safety gaps. This role emphasizes reproducible data and actionable findings across projects and customers. You bring prior red teaming experience in AI or cybersecurity...Vietnamese languageEnglish language skillsRemote job
$65 - $70 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Jack Dorsey . Position: Biology Expert (PhD) — AI Safety Type: Contract Compensation:... ...Strong scientific reasoning and writing in English . Sound judgment around biosecurity...English language skillsContract workSummer workImmediate startRemote work$20 - $22 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Assamese Type: Contract Compensation: $20–$22/...English language skillsContract workSummer workRemote work- ...PhD‑level biologists to help make advanced AI models safer. You'll apply your scientific... ...on the workflow. Responsibilities Write expert‑level prompts across specialized life‑science... ...scientific reasoning and writing in English. Sound judgment around biosecurity and the...English language skillsPart timeImmediate start
- ...Fluent Language Skills Required English & Punjabi. Native fluency in... ...Mercor, we believe the safest AI is the one that’s already been... ...for this project - human data experts who probe AI models with adversarial... ...Mercor customers trust the safety of their AI because you’ve...English language skillsRemote work
$48 - $62 per hour
...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production... ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and adversarial...English language skillsHourly payRemote workFlexible hours$125k - $175k
...fusion, artificial intelligence (AI), machine learning (ML), and... ...(AR).QinetiQ US’s dedicated experts in defense, aerospace, security... ...US means being central to the safety and security of the world around... ..., including Khmer, Thai, Vietnamese, or Arabic.Test Facility Design...Vietnamese languageWork experience placementInterim roleRemote workOverseas$20 - $22 per hour
...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce reproducible attack data... ...can act on Qualifications Native fluency in both English and Bengali, spoken and written, is required Prior red...English language skillsHourly payRemote work$20 - $22 per hour
...Role Overview Lead adversarial testing of conversational AI in both English and Assamese, probing model behavior to find jailbreaks, misinformation, bias, and other harmful or unsafe outputs. This is a remote, text-based role focused on producing reproducible red team...English language skillsHourly payRemote work$20 - $22 per hour
...Overview This role performs adversarial testing of conversational AI systems, probing models for vulnerabilities and producing... ...Compensation 20 - 22 hourly Eligibility Native fluency in English and Odia is required Ability to perform remote, text based work...English language skillsHourly payRemote work$48 - $62 per hour
...Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and... ...adherence to provided guidance. Qualifications Native fluency in English and Danish is required. Prior red teaming experience, such as...English language skillsHourly payRemote work$29 - $45 per hour
...Overview This role performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red... ...team work, or socio-technical probing Native fluency in English and Portuguese, with the Portuguese variety being global...English language skillsHourly payRemote work$17 - $25 per hour
...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers... ...: 17 - 25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work...English language skillsHourly payRemote work$48 - $62 per hour
...Role Overview Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents to surface vulnerabilities, generate reproducible attack cases, and produce human-labeled data that helps teams make AI systems safer. All work is text...English language skillsHourly payRemote work$17 - $25 per hour
...performs adversarial testing of conversational AI, creating reproducible attack cases,... ...socio-technical probing Native fluency in English and Indonesian, both required Curious... ...Increasing customer confidence in their AI safety because systems have been thoroughly...English language skillsHourly payRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Expert for English and Vietnamese. Be the first to apply!
- fruit expert United States
- subject matter expert United States
- expert data analyst United States
- guest service support expert United States
- expert systems engineer United States
- technology expert United States
- fulfillment expert United States
- subject matter expert senior United States
- subject matter expert work from home United States
- sql expert United States


