AI Safety Red Teamer, English & Danish
$48 - $62 per hourSaidGig
Help make conversational AI systems safer by probing them with adversarial inputs, identifying weaknesses, and creating actionable red-team data. This remote, hourly role focuses on text-based evaluations of AI outputs, including optional higher-sensitivity work supported by clear guidance and wellness resources. Key Responsibilities
- Test conversational AI models and agents for jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
- Review AI outputs involving sensitive subjects, including bias, misinformation, and harmful behaviors. Topics will be communicated before any content exposure.
- Annotate model failures, classify vulnerabilities, and identify systemic risks.
- Use established taxonomies, benchmarks, and playbooks to conduct consistent evaluations.
- Create reproducible reports, datasets, and attack cases that teams can use to strengthen AI systems.
- Native fluency in both English and Danish is required.
- Prior experience in AI red teaming, cybersecurity, or socio-technical risk assessment.
- Ability to test systems methodically using frameworks or benchmarks and communicate risks clearly to technical and non-technical audiences.
- Comfort adapting across projects and customer needs.
- Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attacks, or model extraction.
- Cybersecurity, including penetration testing, exploit development, or reverse engineering.
- Socio-technical risk research, including harassment or misinformation testing, abuse analysis, or conversational AI evaluation.
- Creative adversarial thinking informed by psychology, acting, or writing.
- Remote, hourly engagement.
- $48 to $62 per hour.
- Mercor is seeking an AI Safety Expert to remotely assess and fortify AI systems. You will red-team conversational models, identify jailbreaks and biases, and generate... ...data to help customers mitigate risks. Strong English and Swedish communication is essential, with prior...English language skillsRemote jobContract workFlexible hours
$48 - $62 per hour
...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production. The role... ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and...English language skillsHourly payRemote workFlexible hours$20 - $22 per hour
...Role Overview Lead adversarial testing of conversational AI in both English and Assamese, probing model behavior to find jailbreaks, misinformation... ...a remote, text-based role focused on producing reproducible red team data and actionable reports that help customers...English language skillsHourly payRemote work$20 - $22 per hour
...performs adversarial testing of conversational AI systems, probing models for vulnerabilities and producing reproducible red team data that helps make AI safer. Work is... ...22 hourly Eligibility Native fluency in English and Odia is required Ability to perform remote...English language skillsHourly payRemote work$29 - $45 per hour
...performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red team data that clients can act on. Work is... ...socio-technical probing Native fluency in English and Portuguese, with the Portuguese variety being...English language skillsHourly payRemote work$20 - $22 per hour
...Role Overview Join a remote red team focused on probing conversational AI to find real-world vulnerabilities and... ...adversarial datasets that improve model safety. You will create and document... ...Eligibility ~ Native fluency in both English and Punjabi is required for this...English language skillsHourly payRemote work$24 - $35 per hour
...Role Overview Join a red team committed to finding real-world vulnerabilities in conversational AI. You will probe models and agents with adversarial inputs, document reproducible... ...Native-level fluency in both English and Thai, able to read, write, and communicate...English language skillsHourly payRemote workShift work- Intellectt INC is seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high‑risk topics. Join a team...Remote jobFull timeFlexible hours
- We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area"...
- Handshake is seeking an AI Red Teamer in Seattle, WA, to stress-test large language models by designing adversarial prompts. The role involves... ...collaborating with a multidisciplinary team to strengthen AI safety. Ideal candidates should have strong experience with LLMs, an...
- Obsidian is seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design prompts, uncover model weaknesses, and assess safety across high-risk topics. Join a cutting-edge research environment,...
- Mercor is seeking an AI Safety Red Teamer to design adversarial prompts and stress-test frontier AI models in a remote contract role. The candidate will identify jailbreaks, unsafe behaviors, hallucinations, and policy failures while evaluating robustness across misinformation...Remote jobContract work
$17 - $25 per hour
...Help strengthen conversational AI systems by probing them with adversarial... ..., and producing actionable safety data. This remote role focuses... .... Key Responsibilities Red team conversational AI models... ...Qualifications Native fluency in both English and Malay is required. Prior...English language skillsHourly payRemote work$17 - $25 per hour
...Help strengthen conversational AI systems by probing them with adversarial... ..., and producing actionable red-team data. This text-based... ...high-quality human data for AI safety improvements. Use established... ...Native fluency in both English and Vietnamese is required....English language skillsHourly payRemote work- ...Fluent Language Skills Required: English & Malay. Native fluency in... ...Mercor, we believe the safest AI is the one that’s already been... ...attacked — by us. We are assembling a red team for this project - human... ...production Mercor customers trust the safety of their AI because you’ve...English language skillsRemote work
$29 - $45 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type: Contract... ...: Remote Role Responsibilities Red team conversational AI models and agents...English language skillsContract workSummer workRemote work- A technology firm is seeking an AI Red-Teamer for adversarial AI testing. The ideal candidate will be fluent in English and Chinese (Mandarin) and have prior experience in red teaming, cybersecurity, or adversarial testing. The role involves red teaming conversational AI...English language skillsRemote jobHourly payContract work
- Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe conversational AI, annotate vulnerabilities, and generate high-quality attack data for safer deployments. You’ll work across...English language skillsRemote job
- Mercor is building a remote AI red team to probe and test conversational models for vulnerabilities, bias, and safety gaps. This role emphasizes reproducible data and actionable findings across projects and customers. You bring prior red teaming experience in AI or cybersecurity...English language skillsRemote job
- ...TLDR: We're looking for an AI Red Team Engineer to break LLM-powered systems responsibly... .... About us White Circle is an AI Safety company building the safety, reliability,... ...clear, reproducible bug reports in clear English. Can move fast without perfect requirements...English language skillsLocal areaRemote work
$70 - $84 per hour
...AI Safety Red Teamer is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the...Remote jobFor contractors10 hours per week$20 - $22 per hour
...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce reproducible attack data. The role... ...act on Qualifications Native fluency in both English and Bengali, spoken and written, is required Prior...English language skillsHourly payRemote work$17 - $25 per hour
...adversarial testing of conversational AI, creating reproducible attack... .... Key Responsibilities Red team conversational AI models... ...probing Native fluency in English and Indonesian, both required... ...customer confidence in their AI safety because systems have been thoroughly...Hourly payRemote work$17 - $25 per hour
...Role Overview Join a red team of human data experts who probe conversational AI models with adversarial inputs to surface... ...red team data that improves model safety. This text-based role focuses on... ...Native fluency in both English and Vietnamese is required Prior...English language skillsHourly payRemote work$48 - $62 per hour
...Role Overview Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents to surface vulnerabilities... ...communicated before any exposure. Key Responsibilities Red team conversational AI models and agents, including...English language skillsHourly payRemote work$17 - $25 per hour
...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers. This text-... ...25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work...English language skillsHourly payRemote work$70 - $84 per hour
...Role Overview Lead adversarial evaluations of frontier AI models by designing and executing high-impact prompts and tests... .... Document discovered vulnerabilities and contribute to safety benchmarking and red-teaming reports. Collaborate with AI researchers and...Hourly payRemote workWork visa- Mercor is seeking an AI Safety Red Teamer to stress-test frontier AI models through adversarial prompts and systematic evaluation. You will uncover weaknesses, guide safety benchmarks, and document findings to strengthen model alignment and responsible deployment. The...Remote jobHourly payFull time
- ...A progressive technology company is seeking a remote AI Red Teamer to assess conversational AI models and identify vulnerabilities through red teaming exercises. Candidates should be fluent in both English and Brazilian Portuguese, with a background in cybersecurity or...English language skillsHourly payRemote work
$48 - $62 per hour
...Probe conversational AI systems for weaknesses before they reach... ...vulnerabilities, create high-quality safety data, and produce practical... .... Key Responsibilities Red team conversational AI models and... ...Native fluency in both English and Swedish is required. Prior...English language skillsHourly payRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Red Teamer, English & Danish. Be the first to apply!
- english professor United States
- english speaking tourist guide United States
- english language United States
- english full time jobs United States
- english language art teacher United States
- english instructor United States
- english summer camp United States
- english speaking United States
- sales manager in english and russian United States
- english proofreading United States



