AI Safety Red Teamer, English and Norwegian
$48 - $62 per hourSaidGig
Help strengthen conversational AI systems by probing them with adversarial inputs, identifying vulnerabilities, and producing actionable safety data. This text-based remote role focuses on testing AI behavior across areas such as jailbreaks, prompt injection, misuse, bias, misinformation, and harmful behaviors.
Key Responsibilities- Red team conversational AI models and agents through jailbreak attempts, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
- Review AI outputs involving sensitive topics, including bias, misinformation, and harmful behavior.
- Annotate failures, classify vulnerabilities, and flag systemic risks.
- Use established taxonomies, benchmarks, and playbooks to ensure consistent testing.
- Create reproducible reports, datasets, and attack cases that teams can use to improve AI systems.
- Native fluency in both English and Norwegian is required.
- Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
- Ability to test systems methodically using frameworks or benchmarks and communicate risks clearly to technical and non-technical audiences.
- Comfort working across projects and customer contexts.
- Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attacks, or model extraction.
- Cybersecurity, including penetration testing, exploit development, or reverse engineering.
- Socio-technical risk analysis, including harassment or misinformation probing, abuse analysis, or conversational AI testing.
- Creative adversarial thinking informed by psychology, acting, or writing.
- Remote, hourly engagement.
- All work is text-based.
- Participation in higher-sensitivity projects is optional. Topics will be communicated before exposure, with clear guidelines and wellness resources available.
$48 to $62 per hour.
$48 - $62 per hour
...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production. The role... ..., or socio-technical probing Native fluency in English and Finnish, both written and spoken Curious and...English language skillsHourly payRemote workFlexible hours$20 - $22 per hour
...Role Overview Lead adversarial testing of conversational AI in both English and Assamese, probing model behavior to find jailbreaks, misinformation... ...a remote, text-based role focused on producing reproducible red team data and actionable reports that help customers...English language skillsHourly payRemote work$20 - $22 per hour
...performs adversarial testing of conversational AI systems, probing models for vulnerabilities and producing reproducible red team data that helps make AI safer. Work is... ...22 hourly Eligibility Native fluency in English and Odia is required Ability to perform remote...English language skillsHourly payRemote work$48 - $62 per hour
...Role Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and produce reproducible... ...guidance. Qualifications Native fluency in English and Danish is required. Prior red teaming experience...English language skillsHourly payRemote work$29 - $45 per hour
...performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red team data that clients can act on. Work is... ...socio-technical probing Native fluency in English and Portuguese, with the Portuguese variety being...English language skillsHourly payRemote work$20 - $22 per hour
...Role Overview Join a remote red team focused on probing conversational AI to find real-world vulnerabilities and... ...adversarial datasets that improve model safety. You will create and document... ...Eligibility ~ Native fluency in both English and Punjabi is required for this...English language skillsHourly payRemote work$24 - $35 per hour
...Role Overview Join a red team committed to finding real-world vulnerabilities in conversational AI. You will probe models and agents with adversarial inputs, document reproducible... ...Native-level fluency in both English and Thai, able to read, write, and communicate...English language skillsHourly payRemote workShift work- Intellectt INC is seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high‑risk topics. Join a team...Remote jobFull timeFlexible hours
- Handshake is seeking an AI Red Teamer in Seattle, WA, to stress-test large language models by designing adversarial prompts. The role involves... ...collaborating with a multidisciplinary team to strengthen AI safety. Ideal candidates should have strong experience with LLMs, an...
- Mercor is seeking an AI Safety Red Teamer to design adversarial prompts and stress-test frontier AI models in a remote contract role. The candidate will identify jailbreaks, unsafe behaviors, hallucinations, and policy failures while evaluating robustness across misinformation...Remote jobContract work
$17 - $25 per hour
...Help strengthen conversational AI systems by probing them with adversarial... ..., and producing actionable safety data. This remote role focuses... .... Key Responsibilities Red team conversational AI models... ...Qualifications Native fluency in both English and Malay is required. Prior...English language skillsHourly payRemote work$17 - $25 per hour
...Help strengthen conversational AI systems by probing them with adversarial... ..., and producing actionable red-team data. This text-based... ...high-quality human data for AI safety improvements. Use established... ...Native fluency in both English and Vietnamese is required....English language skillsHourly payRemote work- ...Fluent Language Skills Required: English & Malay. Native fluency in... ...Mercor, we believe the safest AI is the one that’s already been... ...attacked — by us. We are assembling a red team for this project - human... ...production Mercor customers trust the safety of their AI because you’ve...English language skillsRemote work
$29 - $45 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type: Contract... ...: Remote Role Responsibilities Red team conversational AI models and agents...English language skillsContract workSummer workRemote work- A technology firm is seeking an AI Red-Teamer for adversarial AI testing. The ideal candidate will be fluent in English and Chinese (Mandarin) and have prior experience in red teaming, cybersecurity, or adversarial testing. The role involves red teaming conversational AI...English language skillsRemote jobHourly payContract work
- Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe conversational AI, annotate vulnerabilities, and generate high-quality attack data for safer deployments. You’ll work across...English language skillsRemote job
$48 - $62 per hour
...AI Safety Experts — English & Norwegian is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety...English language skillsRemote jobFor contractors10 hours per week- Mercor is building a remote AI red team to probe and test conversational models for vulnerabilities, bias, and safety gaps. This role emphasizes reproducible data and actionable findings across projects and customers. You bring prior red teaming experience in AI or cybersecurity...English language skillsRemote job
- ...TLDR: We're looking for an AI Red Team Engineer to break LLM-powered systems responsibly... .... About us White Circle is an AI Safety company building the safety, reliability,... ...clear, reproducible bug reports in clear English. Can move fast without perfect requirements...English language skillsLocal areaRemote work
$20 - $22 per hour
...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce reproducible attack data. The role... ...act on Qualifications Native fluency in both English and Bengali, spoken and written, is required Prior...English language skillsHourly payRemote work$17 - $25 per hour
...adversarial testing of conversational AI, creating reproducible attack... .... Key Responsibilities Red team conversational AI models... ...probing Native fluency in English and Indonesian, both required... ...customer confidence in their AI safety because systems have been thoroughly...Hourly payRemote work$17 - $25 per hour
...Role Overview Join a red team of human data experts who probe conversational AI models with adversarial inputs to surface... ...red team data that improves model safety. This text-based role focuses on... ...Native fluency in both English and Vietnamese is required Prior...English language skillsHourly payRemote work$17 - $25 per hour
...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers. This text-... ...25 hourly. Eligibility Native fluency in both English and Malay is required. Candidates must be able to work...English language skillsHourly payRemote work$70 - $84 per hour
...Role Overview Lead adversarial evaluations of frontier AI models by designing and executing high-impact prompts and tests... .... Document discovered vulnerabilities and contribute to safety benchmarking and red-teaming reports. Collaborate with AI researchers and...Hourly payRemote workWork visa- Mercor is seeking an AI Safety Red Teamer to stress-test frontier AI models through adversarial prompts and systematic evaluation. You will uncover weaknesses, guide safety benchmarks, and document findings to strengthen model alignment and responsible deployment. The...Remote jobHourly payFull time
- ...A progressive technology company is seeking a remote AI Red Teamer to assess conversational AI models and identify vulnerabilities through red teaming exercises. Candidates should be fluent in both English and Brazilian Portuguese, with a background in cybersecurity or...English language skillsHourly payRemote work
$48 - $62 per hour
...Probe conversational AI systems for weaknesses before they reach... ...vulnerabilities, create high-quality safety data, and produce practical... .... Key Responsibilities Red team conversational AI models and... ...Native fluency in both English and Swedish is required. Prior...English language skillsHourly payRemote work- ...known as ActiveFence) is a leading trust, safety, and security company. Just like 'Alice'... ...the rabbit hole into the emerging world of AI and focus on safeguarding these... ...-based work for the \"best of the best\" red-teamers in the industry. Work is on-demand and project...Freelance
$60 - $90 per hour
...technical talent with leading AI research labs. Headquartered in... ...Position: Cybersecurity SWE — AI Safety Type: Contract... ...technical reasoning and writing in English . Sound judgment around security... ...reviewing, grading, or red-teaming technical content....English language skillsFull timeContract workSummer workImmediate startRemote work$250k - $400k
You'll manage our growing red-team, own delivery for frontier lab customers, and build... ...frontier lab researchers and expert red teamers. We're an early-stage startup, so you should... .... About us Our mission is to automate AI safety , to pave the way for a future where the...Remote workVisa sponsorshipShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Red Teamer, English and Norwegian. Be the first to apply!
- english professor United States
- english speaking tourist guide United States
- english language United States
- english full time jobs United States
- english language art teacher United States
- english instructor United States
- english summer camp United States
- english speaking United States
- sales manager in english and russian United States
- english proofreading United States




