Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Safety Expert for English and Assamese

SaidGig

Role Overview

Conduct adversarial testing of conversational AI in English and Assamese to find and document safety failures, generate reproducible attack cases, and produce high-quality human data that helps customers harden their systems. The work is text-based and will include reviewing outputs on sensitive topics such as bias, misinformation, and harmful behaviors. Participation in higher-sensitivity tasks is optional, supported by clear guidelines and wellness resources, and any sensitive topics will be communicated to you before exposure.

Key Responsibilities
  • Red team conversational AI models and agents, including jailbreaks, prompt injection, misuse cases, bias exploitation, and multi-turn manipulation.
  • Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
  • Follow structured taxonomies, benchmarks, and playbooks to keep testing consistent and reproducible.
  • Document findings clearly, producing reports, datasets, and attack cases that customers can act on.
Qualifications
  • Native fluency in English and Assamese, both written and spoken, is required.
  • Prior red teaming or adversarial experience, such as AI adversarial work, cybersecurity, or socio-technical probing.
  • Ability to work methodically with frameworks and benchmarks rather than ad hoc approaches.
  • Strong communication skills, able to explain risks to technical and non-technical stakeholders.
  • Curiosity and an adversarial mindset, with adaptability to move across projects and customers.
Nice-to-Have Specialties
  • Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attack techniques, or model extraction.
  • Cybersecurity experience, such as penetration testing, exploit development, or reverse engineering.
  • Socio-technical risk analysis, including harassment, disinformation probing, abuse analysis, or conversational AI testing.
  • Creative probing skills, including psychology, acting, or writing for unconventional adversarial approaches.
What Success Looks Like
  • Uncovering vulnerabilities that automated tests miss.
  • Delivering reproducible artifacts that help customers strengthen their AI systems.
  • Expanding evaluation coverage so fewer surprises occur in production.
  • Helping customers trust their AI because it has been probed like an adversary.
Work Terms
  • Location: Remote.
  • Employment type: Hourly.
  • All work is text-based.
  • Higher-sensitivity projects are optional and will be accompanied by guidelines and wellness resources, and topics will be communicated before exposure.
Compensation
  • Pay rate: 20 - 22 hourly.
Eligibility
  • Required: Native fluency in both English and Assamese.
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the AI Safety Expert for English and Assamese in United States vacancy
  • $20 - $22 per hour

     ...Role Overview Lead adversarial testing of conversational AI in both English and Assamese, probing model behavior to find jailbreaks, misinformation, bias, and other harmful or unsafe outputs. This is a remote, text-based role focused on producing reproducible red team... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    13 hours ago
  • $20 - $22 per hour

     ...AI Safety Experts — English & Assamese is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety... 
    English language skills
    Remote job
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    15 days ago
  • $20 - $22 per hour

     ...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our...  ..., Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Assamese Type: Contract Compensation: $20–$22/hour Location... 
    English language skills
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    3 days ago
  •  ...Role Overview Lead adversarial testing of conversational AI in both English and Malay, producing reproducible red-team data that uncovers vulnerabilities related to bias, misinformation, and harmful behaviors. All tasks are text-based. Participation in higher-sensitivity... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    13 days ago
  •  ...Role Overview Probe conversational AI systems for safety and misuse vulnerabilities in both English and Norwegian, producing high-quality adversarial test cases, annotations, and reproducible reports that help customers reduce bias, misinformation, and harmful behaviors... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  • $24 - $35 per hour

     ...Role Overview Perform adversarial testing of conversational AI by probing models with crafted inputs, surfacing vulnerabilities,...  ...cybersecurity, or socio-technical probing Native fluency in English and Thai is required Curious and adversarial mindset, with... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  •  ...Role Overview This role probes conversational AI models to find real-world safety weaknesses and produce reproducible red team data that customers...  ...reproduce and act on Qualifications Native fluency in English and Finnish is required Prior red teaming experience,... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  •  ...This role joins a red team project that probes conversational AI models and agents to uncover vulnerabilities before they reach production...  ...cybersecurity, or socio-technical probing. Native fluency in English and Odia is required, including the ability to read and write... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  •  ...Role Overview Probe conversational AI systems with adversarial inputs to expose vulnerabilities...  ...role focused on rigorous, structured safety testing across multiple projects and...  ...Eligibility Native fluency in both English and Danish is required Role is remote;... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  • $20 - $22 per hour

     ...Role Overview Lead adversarial testing of conversational AI in English and Bengali, producing reproducible attack cases, datasets, and reports...  ...occur in production. Increase client trust in the safety and robustness of their AI through adversarial verification.... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  •  ...Role Overview Probe conversational AI systems to find vulnerabilities and produce reproducible adversarial data that customers can...  ...cybersecurity testing, or socio-technical probing. Native fluency in English and Vietnamese, with strong written and verbal communication in... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  •  ...Role Overview Act as a human red teamer who probes conversational AI systems in English and Swedish to uncover safety vulnerabilities, produce high-quality adversarial inputs, and deliver reproducible artifacts customers can act on. The work is fully text-based, focuses... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  • $29 - $45 per hour

     ...Role Overview Attack and probe conversational AI systems in English and Portuguese to surface vulnerabilities, produce reproducible adversarial data, and help make deployed models safer. This text-based role focuses on adversarial testing of model outputs that involve... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  • $20 - $22 per hour

     ...Probe conversational AI systems in English and Punjabi to uncover vulnerabilities, produce reproducible red team artifacts, and help customers improve model safety. This role focuses on adversarial testing of text outputs and delivering structured findings that engineers... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  •  ...Role Overview Work as a red team human data expert attacking conversational AI in English and Dutch to surface vulnerabilities, produce reproducible adversarial...  ...the human-labeled data customers use to improve model safety. The role focuses on text-only interactions and on... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    16 days ago
  • $29 - $45 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type: Contract Compensation: $29-$45... 
    English language skills
    Contract work
    Summer work
    Remote work

    Mercor Inc

    San Francisco, CA
    3 days ago
  • $65 - $70 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Jack Dorsey . Position: Biology Expert (PhD) — AI Safety Type: Contract Compensation:...  ...Strong scientific reasoning and writing in English . Sound judgment around biosecurity... 
    English language skills
    Contract work
    Summer work
    Immediate start
    Remote work

    Mercor

    New York, NY
    a month ago
  •  ...PhD‑level biologists to help make advanced AI models safer. You'll apply your scientific...  ...on the workflow. Responsibilities Write expert‑level prompts across specialized life‑science...  ...scientific reasoning and writing in English. Sound judgment around biosecurity and the... 
    English language skills
    Part time
    Immediate start

    Obsidian

    New York, NY
    1 day ago
  •  ...Fluent Language Skills Required English & Punjabi. Native fluency in...  ...Mercor, we believe the safest AI is the one that’s already been...  ...for this project - human data experts who probe AI models with adversarial...  ...Mercor customers trust the safety of their AI because you’ve... 
    English language skills
    Remote work

    Obsidian

    San Francisco, CA
    3 days ago
  • $48 - $62 per hour

     ...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production...  ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and adversarial... 
    English language skills
    Hourly pay
    Remote work
    Flexible hours

    SaidGig

    United States
    2 days ago
  • $20 - $22 per hour

     ...Overview This role performs adversarial testing of conversational AI systems, probing models for vulnerabilities and producing...  ...Compensation 20 - 22 hourly Eligibility Native fluency in English and Odia is required Ability to perform remote, text based work... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    13 hours ago
  • $20 - $22 per hour

     ...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce reproducible attack data...  ...can act on Qualifications Native fluency in both English and Bengali, spoken and written, is required Prior red... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    13 hours ago
  • $17 - $25 per hour

     ...performs adversarial testing of conversational AI, creating reproducible attack cases,...  ...socio-technical probing Native fluency in English and Indonesian, both required Curious...  ...Increasing customer confidence in their AI safety because systems have been thoroughly... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago
  • $48 - $62 per hour

     ...Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and...  ...adherence to provided guidance. Qualifications Native fluency in English and Danish is required. Prior red teaming experience, such as... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago
  • $29 - $45 per hour

     ...Overview This role performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red...  ...team work, or socio-technical probing Native fluency in English and Portuguese, with the Portuguese variety being global... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    13 hours ago
  • $17 - $25 per hour

     ...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers...  ...: 17 - 25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    13 hours ago
  • $48 - $62 per hour

     ...Role Overview Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents to surface vulnerabilities, generate reproducible attack cases, and produce human-labeled data that helps teams make AI systems safer. All work is text... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    5 days ago
  • $20 - $22 per hour

     ...red team focused on probing conversational AI to find real-world vulnerabilities and...  ...produce adversarial datasets that improve model safety. You will create and document targeted...  ...Eligibility ~ Native fluency in both English and Punjabi is required for this role.... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    13 hours ago
  • $20 - $22 per hour

     ...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Punjabi Type: Contract Compensation: $20–$22/... 
    English language skills
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    5 days ago
  •  ...answers used to train and evaluate advanced AI systems. This remote contractor role...  ...question-and-answer assignments that reflect expert-level clinical and biostatistical judgment...  ...unambiguous answers in expert-level written English, including comprehensive step-by-step... 
    English language skills
    Hourly pay
    For contractors
    Remote work

    SaidGig

    Indiana
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Safety Expert for English and Assamese. Be the first to apply!