AI Safety Red Teamer [Remote]
$70 - $84 per hourAuraOne Human Data
- Remote job
AI Safety Red Teamer is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Why this role matters
Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.
Responsibilities
- Design adversarial prompts that probe known weakness classes (jailbreak, policy bypass, prompt injection) for AI Safety Red Teamer assignments.
- Document every successful attack with reproduction steps and the policy clause it violated.
- Score model defenses across single-turn and multi-turn conversations.
- Triage emerging attack vectors and route them to the safety team with severity ratings.
- Maintain a personal library of attack patterns and propose new red-team rubrics.
- Calibrate against the broader red-team cohort to keep coverage and severity consistent.
Qualifications
- Demonstrated experience red-teaming AI systems, security research, or adversarial ML work for AI Safety Red Teamer work.
- Strong written communication — your reports become the patch ticket.
- Comfort working in policy-grey areas with clear documentation of what was attempted and why.
- Familiarity with prompt-injection, jailbreak, and policy-bypass taxonomies.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Construct a 5-turn adversarial conversation that bypasses a specific policy clause and write up the patch ticket.
- Score a model's defenses against a known jailbreak pattern across 20 variants.
- Propose a new red-team rubric category after spotting an emerging attack vector.
- Reproduce a failure another reviewer reported and confirm the severity tag.
Nice to have
- Background in offensive security, AppSec, or trust & safety operations.
- Experience publishing or reproducing public adversarial-ML research.
- Multilingual fluency for cross-language attack testing.
Skills
- Adversarial prompting
- Red-team analysis
- Policy taxonomy
- Failure documentation
- AI Safety Red Teamer evaluation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
$70–$84 / hr
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$70 - $84 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Adam D'Angelo , Larry Summers , and Jack Dorsey . Position AI Safety Red Teamer Type Contract Compensation $70-$84/hour Location Remote Role...SuggestedContract workSummer workRemote work- Mercor is seeking an AI Safety Red Teamer to design adversarial prompts and stress-test frontier AI models in a remote contract role. The candidate will identify jailbreaks, unsafe behaviors, hallucinations, and policy failures while evaluating robustness across misinformation...SuggestedRemote jobContract work
- Mercor is seeking an AI Safety Expert to remotely assess and fortify AI systems. You will red-team conversational models, identify jailbreaks and biases, and generate actionable data to help customers mitigate risks. Strong English and Swedish communication is essential...SuggestedRemote jobContract workFlexible hours
$17 - $25 per hour
...Role Overview Join a red team of human data experts who probe conversational AI models with adversarial inputs to surface vulnerabilities and produce high-quality red team data that improves model safety. This text-based role focuses on reviewing AI outputs that touch...SuggestedHourly payRemote work$20 - $22 per hour
...Role Overview This role performs adversarial testing of conversational AI systems, probing models for vulnerabilities and producing reproducible red team data that helps make AI safer. Work is fully text based, and you will review model outputs that may touch on sensitive...SuggestedHourly payRemote work$20 - $22 per hour
...Role Overview Lead adversarial testing of conversational AI in both English and Assamese, probing model behavior to find jailbreaks... ...This is a remote, text-based role focused on producing reproducible red team data and actionable reports that help customers strengthen...Hourly payRemote work$17 - $25 per hour
...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers. This text-based role focuses on testing models for bias, misinformation, harmful behaviors, jailbreaks, and prompt-injection style...Hourly payRemote work$29 - $45 per hour
...Role Overview This role performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red team data that clients can act on. Work is remote and text-based. The role includes reviewing model outputs that may...Hourly payRemote work$20 - $22 per hour
...Role Overview You will join a red team of human experts who probe conversational AI to uncover vulnerabilities and produce reproducible attack data. The role focuses on text based review of model outputs that touch on sensitive topics such as bias, misinformation, and...Hourly payRemote work$17 - $25 per hour
...performs adversarial testing of conversational AI, creating reproducible attack cases,... ...is text-based. Key Responsibilities Red team conversational AI models and agents,... ...Increasing customer confidence in their AI safety because systems have been thoroughly probed...Hourly payRemote work$20 - $22 per hour
...Role Overview Join a remote red team focused on probing conversational AI to find real-world vulnerabilities and produce adversarial datasets that improve model safety. You will create and document targeted attacks, annotate model failures, and help customers reduce bias...Hourly payRemote work$70 - $84 per hour
...Role Overview Lead adversarial evaluations of frontier AI models by designing and executing high-impact prompts and tests... .... Document discovered vulnerabilities and contribute to safety benchmarking and red-teaming reports. Collaborate with AI researchers and...Hourly payRemote workWork visa$250k - $400k
You'll manage our growing red-team, own delivery for frontier lab customers, and build... ...frontier lab researchers and expert red teamers. We're an early-stage startup, so you should... .... About us Our mission is to automate AI safety , to pave the way for a future where the...Remote workVisa sponsorshipShift work$24 - $35 per hour
...Role Overview Help strengthen conversational AI systems by testing them as an adversary would. You will probe models and agents for vulnerabilities, create actionable red-team data, and evaluate text-based outputs involving sensitive topics such as bias, misinformation...Hourly payRemote work$48 - $62 per hour
...Role Overview Probe conversational AI systems for safety and misuse vulnerabilities in both English and Norwegian, producing high-quality adversarial... ...failures automated tests miss. Key Responsibilities Red team conversational AI models and agents, including...Hourly payRemote work- ...Role Overview Probe conversational AI systems with adversarial inputs to expose vulnerabilities, produce reproducible red-team data, and deliver actionable reports that help... ...-based role focused on rigorous, structured safety testing across multiple projects and customers...Hourly payRemote work
- ...Role Exists At Mercor, we believe the safest AI is the one that’s already been attacked — by us. We are assembling a red team for this project - human data experts who... ...surprises in production Mercor customers trust the safety of their AI because you’ve already probed it...Remote work
$48 - $62 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Danish Type: Contract... ...: Remote Role Responsibilities Red team conversational AI models and agents...Contract workSummer workRemote work$20 - $22 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Assamese Type: Contract... ...: Remote Role Responsibilities Red team conversational AI models and agents...Contract workSummer workRemote work$54 - $111 per hour
A leading AI consultancy is seeking an AI Red-Teamer for adversarial AI testing. This remote role involves conducting advanced testing, generating critical reports, and identifying vulnerabilities in AI models. Ideal candidates will have prior experience in testing and...Remote jobHourly pay10 hours per weekFlexible hours- Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate actionable data and reports to help customers harden their AI. Ideal candidates...Remote job
$29 - $45 per hour
Mercor connects elite creative and technical talent with leading AI research labs. The AI Safety Experts — English & Portuguese (global) contract role is remote, offering $29-$45/hour. You will red team conversational AI models, generate high-quality human data, and document...Remote jobContract work- Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human...Remote job
- Mercor is seeking a remote Senior Red Team AI Specialist to test conversational agents and AI models against adversarial inputs. You will annotate vulnerabilities, surface systemic risks, and produce reproducible attack cases to help customers strengthen their AI systems...Remote work
- Obsidian is seeking a remote AI red team expert responsible for probing AI models with adversarial inputs to uncover vulnerabilities. You... ...document findings, and generate high-quality data to improve AI safety. The ideal candidate brings red-teaming experience, a curious...Remote job
- A technology firm is seeking an AI Red-Teamer for adversarial AI testing. The ideal candidate will be fluent in English and Chinese (Mandarin) and have prior experience in red teaming, cybersecurity, or adversarial testing. The role involves red teaming conversational AI...Remote jobHourly payContract work
- ...RevolutionBecome a Cybersecurity Senior Advisor - Red Team Lead at Southern California Edison (... ...environments, including understanding of safety and reliability constraints.Demonstrated... ..., and emerging technologies (including AI/ML) to scale Red Team operations and...Remote workRelocation
- Mercor is seeking a remote red-team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate high-quality data... ...document results and collaborate across projects for robust safety testing. This role focuses on reproducible artifacts,...Remote job
- Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe conversational AI, annotate vulnerabilities, and generate high-quality attack data for safer deployments. You’ll work across...Remote job
- Mercor is building a fully remote red team to probe AI models for safety, resilience, and ethical considerations. You will work on adversarial inputs, jailbreaks, and vulnerability discovery while producing actionable data and reproducible reports for customers. This role...Remote job
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Red Teamer [Remote]. Be the first to apply!
- senior devops engineer remote Remote
- medical coding remote Remote
- remote medical coding supervisor Remote
- remote virtual Remote
- remote contract attorney Remote
- clinical data manager remote Remote
- remote servicenow developer Remote
- administrative assistant remote Remote
- remote animation Remote
- remote video editor Remote



