AI Safety Red Teamer
$70 - $84 per hourSaidGig
Role Overview
Lead adversarial testing of frontier AI models by designing challenging prompts to surface jailbreaks, hallucinations, unsafe outputs, and policy failures. You will evaluate model behavior on complex, high-risk, and ambiguous "grey-area" topics, document vulnerabilities, and work with researchers to improve model alignment and robustness.
Key Responsibilities- Design adversarial prompts to stress-test frontier AI models.
- Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures.
- Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
- Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
- Collaborate with AI researchers to improve model alignment, robustness, and safety.
- Required : Bachelors degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
- 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
- Strong analytical reasoning, prompt design, and written communication skills.
- Experience designing adversarial prompts or evaluating frontier AI systems.
- Preferred : Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety; familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies; expertise in one or more grey-area domains such as cyber, biosecurity, political content, misinformation, or scientific safety.
- Location: Remote.
- Employment type: Hourly.
70 - 84 hourly.
Why Join- Help secure and strengthen the next generation of frontier AI models.
- Work on cutting-edge adversarial testing alongside leading AI researchers and safety teams.
- Influence how AI systems respond to complex, real-world safety challenges.
Vacancy posted 20 hours ago
Similar jobs that could be interesting for youBased on the AI Safety Red Teamer in United States vacancy
- Handshake is seeking an AI Red Teamer in Seattle, WA, to stress-test large language models by designing adversarial prompts. The role involves... ...collaborating with a multidisciplinary team to strengthen AI safety. Ideal candidates should have strong experience with LLMs, an...Suggested
$70 - $84 per hour
...AI Safety Red Teamer is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the...SuggestedRemote jobFor contractors10 hours per week$70 - $84 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ..., Larry Summers , and Jack Dorsey . Position: AI Safety Red Teamer Type: Contract Compensation: $70–$84/hour Location...SuggestedContract workSummer workRemote work- A technology consulting firm is seeking an independent contractor for an AI Red-Teaming role focused on enhancing AI safety through probing models and generating high-quality datasets. Responsibilities include uncovering vulnerabilities, documenting findings, and contributing...SuggestedRemote jobFor contractors
$17 - $25 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Vietnamese Type:... ...Location: Remote Role Responsibilities Red team conversational AI models and agents...SuggestedContract workSummer workRemote work- Why This Role Exists We believe the safest AI is the one that’s already been attacked —... .... That’s why we’re building a pod of AI Red-Teamers: human data experts who probe AI models... ...driven AI red-teaming at the frontier of safety Play a direct role in making AI systems...Contract workFor contractorsRemote workFlexible hours
- ...known as ActiveFence) is a leading trust, safety, and security company. Just like 'Alice'... ...the rabbit hole into the emerging world of AI and focus on safeguarding these... ...-based work for the \"best of the best\" red-teamers in the industry. Work is on-demand and project...Freelance
- About the Role As an AI Red Teamer, you will stress-test large language models by intentionally trying to break them. Rather than checking... ...weaknesses, and unexpected behaviors. Your work directly supports AI safety and model robustness for leading research labs. This is a...
$29 - $45 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type... ...Location: Remote Role Responsibilities Red team conversational AI models and agents by...Contract workSummer workRemote work$54 - $111 per hour
A leading AI consultancy is seeking an AI Red-Teamer for adversarial AI testing. This remote role involves conducting advanced testing, generating critical reports, and identifying vulnerabilities in AI models. Ideal candidates will have prior experience in testing and...Remote jobHourly pay10 hours per weekFlexible hours- Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human...Remote job
- Mercor is seeking a remote Senior Red Team AI Specialist to test conversational agents and AI models against adversarial inputs. You will annotate vulnerabilities, surface systemic risks, and produce reproducible attack cases to help customers strengthen their AI systems...Remote work
- Mercor is recruiting AI safety experts to remotely assess and strengthen AI systems. You will conduct red team activities, identify jailbreaks and misuse cases, and produce actionable data to help clients improve model safety. Ideal candidates fluently speak English and...Remote jobContract work
- Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate actionable data and reports to help customers harden their AI. Ideal candidates...Remote job
- Obsidian is seeking a remote AI red team expert responsible for probing AI models with adversarial inputs to uncover vulnerabilities. You... ...document findings, and generate high-quality data to improve AI safety. The ideal candidate brings red-teaming experience, a curious...Remote job
- A technology firm is seeking an AI Red-Teamer for adversarial AI testing. The ideal candidate will be fluent in English and Chinese (Mandarin) and have prior experience in red teaming, cybersecurity, or adversarial testing. The role involves red teaming conversational AI...Remote jobHourly payContract work
- Anthropic is seeking a Red Team Engineer to uncover vulnerabilities across our AI product ecosystem. You will simulate adversaries, perform comprehensive adversarial testing, and explore novel abuse vectors in our deployed systems. You will design kill-chain style attacks...Flexible hours
- Mercor is building a fully remote red team to probe AI models for safety, resilience, and ethical considerations. You will work on adversarial inputs, jailbreaks, and vulnerability discovery while producing actionable data and reproducible reports for customers. This role...Remote job
- AuraOne is seeking a remote contractor to perform Chemical Safety Risk Evaluation by designing adversarial prompts, documenting failures... ...least 10 hours per week. Applicants should have experience in red-teaming AI systems, familiarity with jailbreaking and policy-bypass...Remote jobFor contractors10 hours per week
- Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe conversational AI, annotate vulnerabilities, and generate high-quality attack data for safer deployments. You’ll work across...Remote job
- A leading AI research organization is seeking a candidate for an Automated Red Teaming role focused on building scalable, research-driven systems to identify and fix... ...position demands collaboration with various teams to ensure system safety and efficacy. #J-18808-Ljbffr OpenAI
- A leading AI research firm in New York, NY, is looking for a specialist to oversee the red-teaming and adversarial evaluation pipeline for their models. The ideal candidate... ...field and have a deep understanding of LLM safety and adversarial techniques. A strong software...
- Mercor is seeking a remote Red Team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate actionable red team data for safer AI systems. You will annotate failures, classify vulnerabilities, and document risks using established taxonomies...Remote job
$26 per hour
...Remote Commitment: 10-40 hours/week Role Responsibilities Red-team conversational AI systems using jailbreaks, prompt injections, misuse cases,... ...sensitive or high-risk scenarios in alignment with defined safety guidelines. Requirements Native-level fluency in both English...Remote jobHourly payContract work- ...RevolutionBecome a Cybersecurity Senior Advisor - Red Team Lead at Southern California Edison (... ...environments, including understanding of safety and reliability constraints.Demonstrated... ..., and emerging technologies (including AI/ML) to scale Red Team operations and...Remote workRelocation
- ...AI Red Team Engineer We're looking for an AI Red Team Engineer to break LLM-powered systems responsibly, automate the repetitive attacks... ...prove it, script it, and write it up. White Circle is an AI Safety company building the safety, reliability, and optimization layer...Local areaRemote work
- Mercor is building a remote AI red team to probe and test conversational models for vulnerabilities, bias, and safety gaps. This role emphasizes reproducible data and actionable findings across projects and customers. You bring prior red teaming experience in AI or cybersecurity...Remote job
- Mercor is assembling a remote red team to test large language models and AI agents by probing for jailbreaks, prompt injections, and systemic risks. You will work with English and Vietnamese inputs to surface vulnerabilities and produce actionable data for customers. You...Remote job
$30 - $60 per hour
Portland Seed Fund is seeking part-time Red Teaming Experts in Seattle to support AI safety evaluation campaigns. Candidates will design and simulate AI conversations, identifying risks and evaluating performance. This role requires strong analytical skills and creative...Part time$26 per hour
A tech consulting firm is seeking a remote contractor to engage in AI red teaming activities. The ideal candidate will work on evaluating conversational AI systems and generating human data through careful documentation and classification of vulnerabilities. Fluency in...Remote jobHourly payFor contractors10 hours per weekFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Red Teamer. Be the first to apply!
Related searches
- senior safety specialist United States
- product safety specialist United States
- patient safety specialist United States
- safety osha United States
- safety assistant United States
- safety trainer United States
- safety tech United States
- awp safety United States
- safety technician United States
- patient safety assistant United States

