AI Safety Expert for English and Bengali
$20 - $22 per hourSaidGig
Lead adversarial testing of conversational AI in English and Bengali, producing reproducible attack cases, datasets, and reports that surface bias, misinformation, and harmful behavior so customers can make their models safer. This is a text-only, remote role focused on human-driven red teaming and vulnerability discovery.
Key Responsibilities- Red team conversational AI models and agents, including jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
- Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
- Follow and apply taxonomies, benchmarks, and playbooks to keep testing consistent and comparable across projects.
- Document attacks and findings reproducibly, producing reports, datasets, and concrete attack cases that clients can act on.
- Optionally participate in higher-sensitivity projects, following provided guidelines and using available wellness resources.
- Work only with text-based content, with topics communicated in advance of exposure.
- Native fluency in English and Bengali is required.
- Prior red teaming experience, such as AI adversarial testing, cybersecurity, or socio-technical probing.
- Curious and adversarial mindset, with an instinct to push systems to breaking points.
- Structured approach to testing, using frameworks or benchmarks rather than ad hoc methods.
- Strong communication skills, able to explain risks to both technical and non-technical stakeholders.
- Adaptable to moving across projects and client contexts.
- Adversarial ML experience, including jailbreak datasets, prompt injection, RLHF/DPO attacks, or model extraction.
- Cybersecurity skills such as penetration testing, exploit development, or reverse engineering.
- Socio-technical risk expertise, including harassment, disinformation, or conversational abuse analysis.
- Creative probing skills, for example psychology, acting, or writing that supports unconventional adversarial thinking.
- Identify vulnerabilities that automated tests miss.
- Deliver reproducible artifacts that materially strengthen client AI systems.
- Expand evaluation coverage so more scenarios are tested and fewer surprises occur in production.
- Increase client trust in the safety and robustness of their AI through adversarial verification.
- Build hands-on experience in human data-driven AI red teaming at the frontier of safety.
- Make a direct impact on the robustness, safety, and trustworthiness of deployed AI systems.
- Remote role, work from your location.
- Employment type: hourly.
- All work is text-based.
- Participation in higher-sensitivity content is optional and supported by clear guidelines and wellness resources; topics will be communicated before exposure.
- Pay rate: $20 to $22 per hour.
- Native fluency in both English and Bengali is mandatory.
- Remote work capability is required.
$20 - $22 per hour
...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce reproducible attack... ...act on Qualifications Native fluency in both English and Bengali, spoken and written, is required Prior red teaming...English language skillsBengaliHourly payRemote work$20 - $22 per hour
...AI Safety Experts — English & Bengali is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety...English language skillsBengaliRemote jobFor contractors10 hours per week$20 - $22 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ..., Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Bengali Type: Contract Compensation: $20–$22/hour...English language skillsBengaliContract workSummer workRemote work- ...Role Overview Conduct adversarial testing of conversational AI in English and Assamese to find and document safety failures, generate reproducible attack cases, and produce high-quality human data that helps customers harden their systems. The work is text-based and...English language skillsHourly payRemote work
- ...Role Overview Probe conversational AI systems for safety and misuse vulnerabilities in both English and Norwegian, producing high-quality adversarial test cases, annotations, and reproducible reports that help customers reduce bias, misinformation, and harmful behaviors...English language skillsHourly payRemote work
- ...Role Overview Lead adversarial testing of conversational AI in both English and Malay, producing reproducible red-team data that uncovers vulnerabilities related to bias, misinformation, and harmful behaviors. All tasks are text-based. Participation in higher-sensitivity...English language skillsHourly payRemote work
$24 - $35 per hour
...Role Overview Perform adversarial testing of conversational AI by probing models with crafted inputs, surfacing vulnerabilities,... ...cybersecurity, or socio-technical probing Native fluency in English and Thai is required Curious and adversarial mindset, with...English language skillsHourly payRemote work- ...Role Overview This role probes conversational AI models to find real-world safety weaknesses and produce reproducible red team data that customers... ...reproduce and act on Qualifications Native fluency in English and Finnish is required Prior red teaming experience,...English language skillsHourly payRemote work
- ...This role joins a red team project that probes conversational AI models and agents to uncover vulnerabilities before they reach production... ...cybersecurity, or socio-technical probing. Native fluency in English and Odia is required, including the ability to read and write...English language skillsHourly payRemote work
- ...Role Overview Probe conversational AI systems with adversarial inputs to expose vulnerabilities... ...role focused on rigorous, structured safety testing across multiple projects and... ...Eligibility Native fluency in both English and Danish is required Role is remote;...English language skillsHourly payRemote work
- ...Role Overview Probe conversational AI systems to find vulnerabilities and produce reproducible adversarial data that customers can... ...cybersecurity testing, or socio-technical probing. Native fluency in English and Vietnamese, with strong written and verbal communication in...English language skillsHourly payRemote work
- ...Role Overview Act as a human red teamer who probes conversational AI systems in English and Swedish to uncover safety vulnerabilities, produce high-quality adversarial inputs, and deliver reproducible artifacts customers can act on. The work is fully text-based, focuses...English language skillsHourly payRemote work
$20 - $22 per hour
...Probe conversational AI systems in English and Punjabi to uncover vulnerabilities, produce reproducible red team artifacts, and help customers improve model safety. This role focuses on adversarial testing of text outputs and delivering structured findings that engineers...English language skillsHourly payRemote work$29 - $45 per hour
...Role Overview Attack and probe conversational AI systems in English and Portuguese to surface vulnerabilities, produce reproducible adversarial data, and help make deployed models safer. This text-based role focuses on adversarial testing of model outputs that involve...English language skillsHourly payRemote work- ...Role Overview Work as a red team human data expert attacking conversational AI in English and Dutch to surface vulnerabilities, produce reproducible adversarial... ...the human-labeled data customers use to improve model safety. The role focuses on text-only interactions and on...English language skillsHourly payRemote work
$29 - $45 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type: Contract Compensation: $29-$45...English language skillsContract workSummer workRemote work$65 - $70 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Jack Dorsey . Position: Biology Expert (PhD) — AI Safety Type: Contract Compensation:... ...Strong scientific reasoning and writing in English . Sound judgment around biosecurity...English language skillsContract workSummer workImmediate startRemote work$20 - $22 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Assamese Type: Contract Compensation: $20–$22/...English language skillsContract workSummer workRemote work- ...Fluent Language Skills Required English & Punjabi. Native fluency in... ...Mercor, we believe the safest AI is the one that’s already been... ...for this project - human data experts who probe AI models with adversarial... ...Mercor customers trust the safety of their AI because you’ve...English language skillsRemote work
- ...PhD‑level biologists to help make advanced AI models safer. You'll apply your scientific... ...on the workflow. Responsibilities Write expert‑level prompts across specialized life‑science... ...scientific reasoning and writing in English. Sound judgment around biosecurity and the...English language skillsPart timeImmediate start
$48 - $62 per hour
...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production... ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and adversarial...English language skillsHourly payRemote workFlexible hours$20 - $22 per hour
...Overview This role performs adversarial testing of conversational AI systems, probing models for vulnerabilities and producing... ...Compensation 20 - 22 hourly Eligibility Native fluency in English and Odia is required Ability to perform remote, text based work...English language skillsHourly payRemote work$20 - $22 per hour
...Role Overview Lead adversarial testing of conversational AI in both English and Assamese, probing model behavior to find jailbreaks, misinformation, bias, and other harmful or unsafe outputs. This is a remote, text-based role focused on producing reproducible red team...English language skillsHourly payRemote work$17 - $25 per hour
...performs adversarial testing of conversational AI, creating reproducible attack cases,... ...socio-technical probing Native fluency in English and Indonesian, both required Curious... ...Increasing customer confidence in their AI safety because systems have been thoroughly...English language skillsHourly payRemote work$17 - $25 per hour
...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers... ...: 17 - 25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work...English language skillsHourly payRemote work$29 - $45 per hour
...Overview This role performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red... ...team work, or socio-technical probing Native fluency in English and Portuguese, with the Portuguese variety being global...English language skillsHourly payRemote work$48 - $62 per hour
...Role Overview Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents to surface vulnerabilities, generate reproducible attack cases, and produce human-labeled data that helps teams make AI systems safer. All work is text...English language skillsHourly payRemote work$48 - $62 per hour
...Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and... ...adherence to provided guidance. Qualifications Native fluency in English and Danish is required. Prior red teaming experience, such as...English language skillsHourly payRemote work$20 - $22 per hour
...red team focused on probing conversational AI to find real-world vulnerabilities and... ...produce adversarial datasets that improve model safety. You will create and document targeted... ...Eligibility ~ Native fluency in both English and Punjabi is required for this role....English language skillsHourly payRemote work$11 - $19 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Dorsey . Position: Music Production Expert - Bengali Type: Contract Compensation:... ...follow detailed written instructions in English. ~2+ years of experience as a music producer...English language skillsBengaliContract workSummer workImmediate startRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Expert for English and Bengali. Be the first to apply!
- fruit expert United States
- subject matter expert United States
- expert data analyst United States
- guest service support expert United States
- expert systems engineer United States
- technology expert United States
- fulfillment expert United States
- subject matter expert senior United States
- subject matter expert work from home United States
- sql expert United States


