AI Safety Expert for English and Assamese
SaidGig
Conduct adversarial testing of conversational AI in English and Assamese to find and document safety failures, generate reproducible attack cases, and produce high-quality human data that helps customers harden their systems. The work is text-based and will include reviewing outputs on sensitive topics such as bias, misinformation, and harmful behaviors. Participation in higher-sensitivity tasks is optional, supported by clear guidelines and wellness resources, and any sensitive topics will be communicated to you before exposure.
Key Responsibilities- Red team conversational AI models and agents, including jailbreaks, prompt injection, misuse cases, bias exploitation, and multi-turn manipulation.
- Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
- Follow structured taxonomies, benchmarks, and playbooks to keep testing consistent and reproducible.
- Document findings clearly, producing reports, datasets, and attack cases that customers can act on.
- Native fluency in English and Assamese, both written and spoken, is required.
- Prior red teaming or adversarial experience, such as AI adversarial work, cybersecurity, or socio-technical probing.
- Ability to work methodically with frameworks and benchmarks rather than ad hoc approaches.
- Strong communication skills, able to explain risks to technical and non-technical stakeholders.
- Curiosity and an adversarial mindset, with adaptability to move across projects and customers.
- Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attack techniques, or model extraction.
- Cybersecurity experience, such as penetration testing, exploit development, or reverse engineering.
- Socio-technical risk analysis, including harassment, disinformation probing, abuse analysis, or conversational AI testing.
- Creative probing skills, including psychology, acting, or writing for unconventional adversarial approaches.
- Uncovering vulnerabilities that automated tests miss.
- Delivering reproducible artifacts that help customers strengthen their AI systems.
- Expanding evaluation coverage so fewer surprises occur in production.
- Helping customers trust their AI because it has been probed like an adversary.
- Location: Remote.
- Employment type: Hourly.
- All work is text-based.
- Higher-sensitivity projects are optional and will be accompanied by guidelines and wellness resources, and topics will be communicated before exposure.
- Pay rate: 20 - 22 hourly.
- Required: Native fluency in both English and Assamese.
$20 - $22 per hour
...Role Overview Lead adversarial testing of conversational AI in both English and Assamese, probing model behavior to find jailbreaks, misinformation, bias, and other harmful or unsafe outputs. This is a remote, text-based role focused on producing reproducible red team...English language skillsHourly payRemote work$20 - $22 per hour
...AI Safety Experts — English & Assamese is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety...English language skillsRemote jobFor contractors10 hours per week$20 - $22 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our... ..., Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Assamese Type: Contract Compensation: $20–$22/hour Location...English language skillsContract workSummer workRemote work- ...Role Overview Lead adversarial testing of conversational AI in both English and Malay, producing reproducible red-team data that uncovers vulnerabilities related to bias, misinformation, and harmful behaviors. All tasks are text-based. Participation in higher-sensitivity...English language skillsHourly payRemote work
- ...Role Overview Probe conversational AI systems for safety and misuse vulnerabilities in both English and Norwegian, producing high-quality adversarial test cases, annotations, and reproducible reports that help customers reduce bias, misinformation, and harmful behaviors...English language skillsHourly payRemote work
$24 - $35 per hour
...Role Overview Perform adversarial testing of conversational AI by probing models with crafted inputs, surfacing vulnerabilities,... ...cybersecurity, or socio-technical probing Native fluency in English and Thai is required Curious and adversarial mindset, with...English language skillsHourly payRemote work- ...Role Overview This role probes conversational AI models to find real-world safety weaknesses and produce reproducible red team data that customers... ...reproduce and act on Qualifications Native fluency in English and Finnish is required Prior red teaming experience,...English language skillsHourly payRemote work
- ...This role joins a red team project that probes conversational AI models and agents to uncover vulnerabilities before they reach production... ...cybersecurity, or socio-technical probing. Native fluency in English and Odia is required, including the ability to read and write...English language skillsHourly payRemote work
- ...Role Overview Probe conversational AI systems with adversarial inputs to expose vulnerabilities... ...role focused on rigorous, structured safety testing across multiple projects and... ...Eligibility Native fluency in both English and Danish is required Role is remote;...English language skillsHourly payRemote work
$20 - $22 per hour
...Role Overview Lead adversarial testing of conversational AI in English and Bengali, producing reproducible attack cases, datasets, and reports... ...occur in production. Increase client trust in the safety and robustness of their AI through adversarial verification....English language skillsHourly payRemote work- ...Role Overview Probe conversational AI systems to find vulnerabilities and produce reproducible adversarial data that customers can... ...cybersecurity testing, or socio-technical probing. Native fluency in English and Vietnamese, with strong written and verbal communication in...English language skillsHourly payRemote work
- ...Role Overview Act as a human red teamer who probes conversational AI systems in English and Swedish to uncover safety vulnerabilities, produce high-quality adversarial inputs, and deliver reproducible artifacts customers can act on. The work is fully text-based, focuses...English language skillsHourly payRemote work
$29 - $45 per hour
...Role Overview Attack and probe conversational AI systems in English and Portuguese to surface vulnerabilities, produce reproducible adversarial data, and help make deployed models safer. This text-based role focuses on adversarial testing of model outputs that involve...English language skillsHourly payRemote work$20 - $22 per hour
...Probe conversational AI systems in English and Punjabi to uncover vulnerabilities, produce reproducible red team artifacts, and help customers improve model safety. This role focuses on adversarial testing of text outputs and delivering structured findings that engineers...English language skillsHourly payRemote work- ...Role Overview Work as a red team human data expert attacking conversational AI in English and Dutch to surface vulnerabilities, produce reproducible adversarial... ...the human-labeled data customers use to improve model safety. The role focuses on text-only interactions and on...English language skillsHourly payRemote work
$29 - $45 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type: Contract Compensation: $29-$45...English language skillsContract workSummer workRemote work$65 - $70 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Jack Dorsey . Position: Biology Expert (PhD) — AI Safety Type: Contract Compensation:... ...Strong scientific reasoning and writing in English . Sound judgment around biosecurity...English language skillsContract workSummer workImmediate startRemote work- ...PhD‑level biologists to help make advanced AI models safer. You'll apply your scientific... ...on the workflow. Responsibilities Write expert‑level prompts across specialized life‑science... ...scientific reasoning and writing in English. Sound judgment around biosecurity and the...English language skillsPart timeImmediate start
- ...Fluent Language Skills Required English & Punjabi. Native fluency in... ...Mercor, we believe the safest AI is the one that’s already been... ...for this project - human data experts who probe AI models with adversarial... ...Mercor customers trust the safety of their AI because you’ve...English language skillsRemote work
$48 - $62 per hour
...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production... ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and adversarial...English language skillsHourly payRemote workFlexible hours$20 - $22 per hour
...Overview This role performs adversarial testing of conversational AI systems, probing models for vulnerabilities and producing... ...Compensation 20 - 22 hourly Eligibility Native fluency in English and Odia is required Ability to perform remote, text based work...English language skillsHourly payRemote work$20 - $22 per hour
...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce reproducible attack data... ...can act on Qualifications Native fluency in both English and Bengali, spoken and written, is required Prior red...English language skillsHourly payRemote work$17 - $25 per hour
...performs adversarial testing of conversational AI, creating reproducible attack cases,... ...socio-technical probing Native fluency in English and Indonesian, both required Curious... ...Increasing customer confidence in their AI safety because systems have been thoroughly...English language skillsHourly payRemote work$48 - $62 per hour
...Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and... ...adherence to provided guidance. Qualifications Native fluency in English and Danish is required. Prior red teaming experience, such as...English language skillsHourly payRemote work$29 - $45 per hour
...Overview This role performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red... ...team work, or socio-technical probing Native fluency in English and Portuguese, with the Portuguese variety being global...English language skillsHourly payRemote work$17 - $25 per hour
...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers... ...: 17 - 25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work...English language skillsHourly payRemote work$48 - $62 per hour
...Role Overview Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents to surface vulnerabilities, generate reproducible attack cases, and produce human-labeled data that helps teams make AI systems safer. All work is text...English language skillsHourly payRemote work$20 - $22 per hour
...red team focused on probing conversational AI to find real-world vulnerabilities and... ...produce adversarial datasets that improve model safety. You will create and document targeted... ...Eligibility ~ Native fluency in both English and Punjabi is required for this role....English language skillsHourly payRemote work$20 - $22 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Punjabi Type: Contract Compensation: $20–$22/...English language skillsContract workSummer workRemote work- ...answers used to train and evaluate advanced AI systems. This remote contractor role... ...question-and-answer assignments that reflect expert-level clinical and biostatistical judgment... ...unambiguous answers in expert-level written English, including comprehensive step-by-step...English language skillsHourly payFor contractorsRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Expert for English and Assamese. Be the first to apply!
- fruit expert United States
- subject matter expert United States
- expert data analyst United States
- guest service support expert United States
- expert systems engineer United States
- technology expert United States
- fulfillment expert United States
- subject matter expert senior United States
- subject matter expert work from home United States
- sql expert United States


