Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Safety Expert for English and Thai Red Teaming

$24 - $35 per hour

SaidGig

Role Overview

Perform adversarial testing of conversational AI by probing models with crafted inputs, surfacing vulnerabilities, and producing reproducible attack cases and human-labeled data that product and engineering teams can use to harden systems. This is text-based red teaming that can involve sensitive topics such as bias, misinformation, or harmful behaviors. Participation in higher-sensitivity reviews is optional and supported with clear guidelines and wellness resources, and topics will be communicated before exposure.

Key Responsibilities
  • Red team conversational AI models and agents, including jailbreaks, prompt injection, misuse cases, bias exploitation, and multi-turn manipulation
  • Generate high-quality human data: annotate failures, classify vulnerabilities, and flag systemic risks
  • Apply structure to testing by following taxonomies, benchmarks, and playbooks to maintain consistency
  • Document findings reproducibly by producing reports, datasets, and attack cases that customers or teams can act on
Qualifications
  • Prior red teaming experience, such as AI adversarial work, cybersecurity, or socio-technical probing
  • Native fluency in English and Thai is required
  • Curious and adversarial mindset, with an instinct to push systems to breaking points
  • Structured approach, using frameworks or benchmarks rather than ad hoc methods
  • Strong communication skills, able to explain risks to technical and non-technical stakeholders
  • Adaptable, able to move across projects and customer contexts
Nice-to-Have Specialties
  • Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attack techniques, and model extraction
  • Cybersecurity experience such as penetration testing, exploit development, or reverse engineering
  • Socio-technical risk expertise, including harassment and disinformation probing, abuse analysis, and conversational AI testing
  • Creative probing skills from psychology, acting, or creative writing to design unconventional adversarial inputs
What Success Looks Like
  • Discover vulnerabilities that automated tests miss
  • Deliver reproducible artifacts that directly strengthen customer AI systems
  • Expand evaluation coverage so more scenarios are tested and fewer surprises arise in production
  • Increase customer trust in deployed AI by identifying and mitigating real-world attack vectors
Work Terms
  • Remote engagement, all work is text-based
  • Hourly employment
  • Participation in higher-sensitivity projects is optional and accompanied by clear guidelines and wellness resources; topics will be disclosed before exposure
Compensation
  • $24.00 to $35.00 per hour
Eligibility
  • Native fluency in both English and Thai is required
  • Work is remote; candidates must be able to work remotely from their location
Vacancy posted 18 days ago
Similar jobs that could be interesting for youBased on the AI Safety Expert for English and Thai Red Teaming in United States vacancy
  •  ...Role Overview Probe conversational AI systems for safety and misuse vulnerabilities in both English and Norwegian, producing high-quality adversarial test cases,...  ...automated tests miss. Key Responsibilities Red team conversational AI models and agents, including jailbreaks... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  •  ...Role Overview This role probes conversational AI models to find real-world safety weaknesses and produce reproducible red team data that customers can act on. You will run...  ...on Qualifications Native fluency in English and Finnish is required Prior red teaming... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  •  ...Role Overview Lead adversarial testing of conversational AI in both English and Malay, producing reproducible red-team data that uncovers vulnerabilities related to bias, misinformation, and harmful behaviors. All tasks are text-based. Participation in higher-sensitivity... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    13 days ago
  •  ...Overview Probe conversational AI systems with adversarial...  ...vulnerabilities, produce reproducible red-team data, and deliver actionable...  ...on rigorous, structured safety testing across multiple projects...  ...Eligibility Native fluency in both English and Danish is required Role... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  •  ...Role Overview Act as a human red teamer who probes conversational AI systems in English and Swedish to uncover safety vulnerabilities, produce high-quality adversarial inputs...  ...spoken and written, is required. Prior red teaming experience, such as AI adversarial work,... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  • $29 - $45 per hour

     ...Role Overview Attack and probe conversational AI systems in English and Portuguese to surface vulnerabilities, produce reproducible adversarial...  ...signposted before exposure. Key Responsibilities Red team conversational AI models and agents, exploring jailbreaks,... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  •  ...Role Overview Work as a red team human data expert attacking conversational AI in English and Dutch to surface vulnerabilities, produce reproducible adversarial test...  ...-labeled data customers use to improve model safety. The role focuses on text-only interactions and on... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    16 days ago
  • $20 - $22 per hour

     ...Probe conversational AI systems in English and Punjabi to uncover vulnerabilities, produce reproducible red team artifacts, and help customers improve model safety. This role focuses on adversarial testing of text outputs and delivering structured findings that engineers... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  • $20 - $22 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...and Jack Dorsey . Position: AI Safety Experts — English & Assamese Type: Contract...  ...Remote Role Responsibilities Red team conversational AI models and agents,... 
    English language skills
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    3 days ago
  •  ...Language Skills Required English & Punjabi. Native fluency...  ...Mercor, we believe the safest AI is the one that’s already...  ...us. We are assembling a red team for this project - human data experts who probe AI models with...  ...customers trust the safety of their AI because you’ve... 
    English language skills
    Remote work

    Obsidian

    San Francisco, CA
    3 days ago
  • $48 - $62 per hour

     ...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production. The role...  ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and... 
    English language skills
    Hourly pay
    Remote work
    Flexible hours

    SaidGig

    United States
    2 days ago
  • $48 - $62 per hour

     ...Role Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and produce reproducible...  ...guidance. Qualifications Native fluency in English and Danish is required. Prior red teaming experience... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago
  • $29 - $45 per hour

     ...performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red team data that clients can act on. Work is remote...  ...-technical probing Native fluency in English and Portuguese, with the Portuguese variety... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    10 hours ago
  • $17 - $25 per hour

     ...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers. This text-...  ...25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    10 hours ago
  • $17 - $25 per hour

     ...adversarial testing of conversational AI, creating reproducible attack...  .... Key Responsibilities Red team conversational AI models and...  ...probing Native fluency in English and Indonesian, both required...  ...confidence in their AI safety because systems have been thoroughly... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago
  • $48 - $62 per hour

     ...Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents...  ...produce human-labeled data that helps teams make AI systems safer. All work is text-...  ...exposure. Key Responsibilities Red team conversational AI models and agents... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    5 days ago
  • $20 - $22 per hour

     ...Role Overview Join a remote red team focused on probing conversational AI to find real-world vulnerabilities and...  ...adversarial datasets that improve model safety. You will create and document...  ...Eligibility ~ Native fluency in both English and Punjabi is required for this... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    3 days ago
  • $20 - $22 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...and Jack Dorsey . Position: AI Safety Experts — English & Punjabi Type: Contract...  ...Remote Role Responsibilities Red team conversational AI models and agents.... 
    English language skills
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    5 days ago
  • $24 - $35 per hour

     ...AI Safety Experts — English & Thai is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team... 
    English language skills
    Thai language
    Remote job
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    3 days ago
  • $24 - $35 per hour

     ...Role Overview Join a red team committed to finding real-world vulnerabilities in conversational AI. You will probe models and agents with adversarial inputs, document...  ...Native-level fluency in both English and Thai, able to read, write, and communicate at a... 
    English language skills
    Thai language
    Hourly pay
    Remote work
    Shift work

    SaidGig

    United States
    10 hours ago
  • Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human... 
    Remote job

    Neon

    New York, NY
    4 days ago
  • Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate actionable data and reports to help customers harden their AI. Ideal candidates... 
    Remote job

    Mercor

    New York, NY
    5 days ago
  •  ...ActiveFence) is a leading trust, safety, and security company. Just...  ...hole into the emerging world of AI and focus on safeguarding these...  ...work for the \"best of the best\" red-teamers in the industry. Work...  ...As one of Alice's Security Red Team Specialists, you'll focus on... 
    Freelance

    Socket.dev

    New York, NY
    4 days ago
  • Mercor is seeking a remote Senior Red Team AI Specialist to test conversational agents and AI models against adversarial inputs. You will annotate vulnerabilities, surface systemic risks, and produce reproducible attack cases to help customers strengthen their AI systems... 
    Remote work

    Neon

    New York, NY
    4 days ago
  •  ...Overview Conduct adversarial testing of conversational AI in English and Assamese to find and document safety failures, generate reproducible attack cases, and...  ...to you before exposure. Key Responsibilities Red team conversational AI models and agents, including... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  •  ...Role Overview This role joins a red team project that probes conversational AI models and agents to uncover vulnerabilities before they reach production...  ...cybersecurity, or socio-technical probing. Native fluency in English and Odia is required, including the ability to read... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  • $20 - $22 per hour

     ...adversarial testing of conversational AI in English and Bengali, producing reproducible attack...  ..., remote role focused on human-driven red teaming and vulnerability discovery. Key...  ...production. Increase client trust in the safety and robustness of their AI through... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    4 days ago
  •  ...Role Overview Probe conversational AI systems to find vulnerabilities and produce reproducible...  ...their models. This is a text-based red teaming role focused on exposing bias,...  ...-technical probing. Native fluency in English and Vietnamese, with strong written and verbal... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 days ago
  • $35 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ..., and Jack Dorsey . Position AI Safety Experts — English & Thai Type Contract Compensation $24-$35/hour...  ...Remote Role Responsibilities Red team conversational AI models and agents to... 
    English language skills
    Thai language
    Remote job
    Contract work
    Summer work

    Mercor

    San Francisco, CA
    1 day ago
  • $26 per hour

     ...consulting firm is seeking a remote contractor to engage in AI red teaming activities. The ideal candidate will work on evaluating conversational...  ...and classification of vulnerabilities. Fluency in English and Spanish is required, along with prior experience in adversarial... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week
    Flexible hours

    Crossing Hurdles

    New York, NY
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Safety Expert for English and Thai Red Teaming. Be the first to apply!