Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Safety Red Teamer, English & Danish

$48 - $62 per hour

SaidGig

Help make conversational AI systems safer by probing them with adversarial inputs, identifying weaknesses, and creating actionable red-team data. This remote, hourly role focuses on text-based evaluations of AI outputs, including optional higher-sensitivity work supported by clear guidance and wellness resources. Key Responsibilities

  • Test conversational AI models and agents for jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
  • Review AI outputs involving sensitive subjects, including bias, misinformation, and harmful behaviors. Topics will be communicated before any content exposure.
  • Annotate model failures, classify vulnerabilities, and identify systemic risks.
  • Use established taxonomies, benchmarks, and playbooks to conduct consistent evaluations.
  • Create reproducible reports, datasets, and attack cases that teams can use to strengthen AI systems.
Qualifications
  • Native fluency in both English and Danish is required.
  • Prior experience in AI red teaming, cybersecurity, or socio-technical risk assessment.
  • Ability to test systems methodically using frameworks or benchmarks and communicate risks clearly to technical and non-technical audiences.
  • Comfort adapting across projects and customer needs.
Preferred Expertise
  • Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attacks, or model extraction.
  • Cybersecurity, including penetration testing, exploit development, or reverse engineering.
  • Socio-technical risk research, including harassment or misinformation testing, abuse analysis, or conversational AI evaluation.
  • Creative adversarial thinking informed by psychology, acting, or writing.
Work Terms
  • Remote, hourly engagement.
Compensation
  • $48 to $62 per hour.
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the AI Safety Red Teamer, English & Danish in United States vacancy
  • Mercor is seeking an AI Safety Expert to remotely assess and fortify AI systems. You will red-team conversational models, identify jailbreaks and biases, and generate...  ...data to help customers mitigate risks. Strong English and Swedish communication is essential, with prior... 
    English language skills
    Remote job
    Contract work
    Flexible hours

    mercor

    San Francisco, CA
    3 days ago
  • $48 - $62 per hour

     ...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production. The role...  ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and... 
    English language skills
    Hourly pay
    Remote work
    Flexible hours

    SaidGig

    United States
    2 days ago
  • $20 - $22 per hour

     ...Role Overview Lead adversarial testing of conversational AI in both English and Assamese, probing model behavior to find jailbreaks, misinformation...  ...a remote, text-based role focused on producing reproducible red team data and actionable reports that help customers... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago
  • $20 - $22 per hour

     ...performs adversarial testing of conversational AI systems, probing models for vulnerabilities and producing reproducible red team data that helps make AI safer. Work is...  ...22 hourly Eligibility Native fluency in English and Odia is required Ability to perform remote... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago
  • $29 - $45 per hour

     ...performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red team data that clients can act on. Work is...  ...socio-technical probing Native fluency in English and Portuguese, with the Portuguese variety being... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $20 - $22 per hour

     ...Role Overview Join a remote red team focused on probing conversational AI to find real-world vulnerabilities and...  ...adversarial datasets that improve model safety. You will create and document...  ...Eligibility ~ Native fluency in both English and Punjabi is required for this... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago
  • $24 - $35 per hour

     ...Role Overview Join a red team committed to finding real-world vulnerabilities in conversational AI. You will probe models and agents with adversarial inputs, document reproducible...  ...Native-level fluency in both English and Thai, able to read, write, and communicate... 
    English language skills
    Hourly pay
    Remote work
    Shift work

    SaidGig

    United States
    1 day ago
  • Intellectt INC is seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high‑risk topics. Join a team... 
    Remote job
    Full time
    Flexible hours

    Intellectt INC

    Warren, MI
    3 days ago
  • We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area"... 

    Obsidian

    San Francisco, CA
    4 days ago
  • Handshake is seeking an AI Red Teamer in Seattle, WA, to stress-test large language models by designing adversarial prompts. The role involves...  ...collaborating with a multidisciplinary team to strengthen AI safety. Ideal candidates should have strong experience with LLMs, an... 

    Handshake

    Seattle, WA
    3 days ago
  • Obsidian is seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design prompts, uncover model weaknesses, and assess safety across high-risk topics. Join a cutting-edge research environment,... 

    Obsidian

    San Francisco, CA
    4 days ago
  • Mercor is seeking an AI Safety Red Teamer to design adversarial prompts and stress-test frontier AI models in a remote contract role. The candidate will identify jailbreaks, unsafe behaviors, hallucinations, and policy failures while evaluating robustness across misinformation... 
    Remote job
    Contract work

    Mercor

    Malta, MT
    3 days ago
  • $17 - $25 per hour

     ...Help strengthen conversational AI systems by probing them with adversarial...  ..., and producing actionable safety data. This remote role focuses...  .... Key Responsibilities Red team conversational AI models...  ...Qualifications Native fluency in both English and Malay is required. Prior... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    21 days ago
  • $17 - $25 per hour

     ...Help strengthen conversational AI systems by probing them with adversarial...  ..., and producing actionable red-team data. This text-based...  ...high-quality human data for AI safety improvements. Use established...  ...Native fluency in both English and Vietnamese is required.... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    26 days ago
  •  ...Fluent Language Skills Required: English & Malay. Native fluency in...  ...Mercor, we believe the safest AI is the one that’s already been...  ...attacked — by us. We are assembling a red team for this project - human...  ...production Mercor customers trust the safety of their AI because you’ve... 
    English language skills
    Remote work

    Mercor Inc

    New York, NY
    3 days ago
  • $29 - $45 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type: Contract...  ...: Remote Role Responsibilities Red team conversational AI models and agents... 
    English language skills
    Contract work
    Summer work
    Remote work

    Mercor Inc

    San Francisco, CA
    1 day ago
  • A technology firm is seeking an AI Red-Teamer for adversarial AI testing. The ideal candidate will be fluent in English and Chinese (Mandarin) and have prior experience in red teaming, cybersecurity, or adversarial testing. The role involves red teaming conversational AI... 
    English language skills
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    1 day ago
  • Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe conversational AI, annotate vulnerabilities, and generate high-quality attack data for safer deployments. You’ll work across... 
    English language skills
    Remote job

    Neon

    New York, NY
    2 days ago
  • Mercor is building a remote AI red team to probe and test conversational models for vulnerabilities, bias, and safety gaps. This role emphasizes reproducible data and actionable findings across projects and customers. You bring prior red teaming experience in AI or cybersecurity... 
    English language skills
    Remote job

    Obsidian

    New York, NY
    1 day ago
  •  ...TLDR: We're looking for an AI Red Team Engineer to break LLM-powered systems responsibly...  .... About us White Circle is an AI Safety company building the safety, reliability,...  ...clear, reproducible bug reports in clear English. Can move fast without perfect requirements... 
    English language skills
    Local area
    Remote work

    White Circle

    United States
    2 days ago
  • $70 - $84 per hour

     ...AI Safety Red Teamer is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the... 
    Remote job
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    21 days ago
  • $20 - $22 per hour

     ...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce reproducible attack data. The role...  ...act on Qualifications Native fluency in both English and Bengali, spoken and written, is required Prior... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago
  • $17 - $25 per hour

     ...adversarial testing of conversational AI, creating reproducible attack...  .... Key Responsibilities Red team conversational AI models...  ...probing Native fluency in English and Indonesian, both required...  ...customer confidence in their AI safety because systems have been thoroughly... 
    Hourly pay
    Remote work

    SaidGig

    United States
    2 days ago
  • $17 - $25 per hour

     ...Role Overview Join a red team of human data experts who probe conversational AI models with adversarial inputs to surface...  ...red team data that improves model safety. This text-based role focuses on...  ...Native fluency in both English and Vietnamese is required Prior... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $48 - $62 per hour

     ...Role Overview Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents to surface vulnerabilities...  ...communicated before any exposure. Key Responsibilities Red team conversational AI models and agents, including... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    18 hours ago
  • $17 - $25 per hour

     ...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers. This text-...  ...25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $70 - $84 per hour

     ...Role Overview Lead adversarial evaluations of frontier AI models by designing and executing high-impact prompts and tests...  .... Document discovered vulnerabilities and contribute to safety benchmarking and red-teaming reports. Collaborate with AI researchers and... 
    Hourly pay
    Remote work
    Work visa

    SaidGig

    United States
    1 day ago
  • Mercor is seeking an AI Safety Red Teamer to stress-test frontier AI models through adversarial prompts and systematic evaluation. You will uncover weaknesses, guide safety benchmarks, and document findings to strengthen model alignment and responsible deployment. The... 
    Remote job
    Hourly pay
    Full time

    Mercor

    Santa Clarita, CA
    3 days ago
  •  ...A progressive technology company is seeking a remote AI Red Teamer to assess conversational AI models and identify vulnerabilities through red teaming exercises. Candidates should be fluent in both English and Brazilian Portuguese, with a background in cybersecurity or... 
    English language skills
    Hourly pay
    Remote work

    Crossing Hurdles

    New York, NY
    1 day ago
  • $48 - $62 per hour

     ...Probe conversational AI systems for weaknesses before they reach...  ...vulnerabilities, create high-quality safety data, and produce practical...  .... Key Responsibilities Red team conversational AI models and...  ...Native fluency in both English and Swedish is required. Prior... 
    English language skills
    Hourly pay
    Remote work

    SaidGig

    United States
    26 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Safety Red Teamer, English & Danish. Be the first to apply!