Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Safety Red Teamer [Remote]

$70 - $84 per hour

AuraOne Human Data

Remote
  • Remote job

AI Safety Red Teamer is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.

Why this role matters

Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.

Responsibilities

  • Design adversarial prompts that probe known weakness classes (jailbreak, policy bypass, prompt injection) for AI Safety Red Teamer assignments.
  • Document every successful attack with reproduction steps and the policy clause it violated.
  • Score model defenses across single-turn and multi-turn conversations.
  • Triage emerging attack vectors and route them to the safety team with severity ratings.
  • Maintain a personal library of attack patterns and propose new red-team rubrics.
  • Calibrate against the broader red-team cohort to keep coverage and severity consistent.

Qualifications

  • Demonstrated experience red-teaming AI systems, security research, or adversarial ML work for AI Safety Red Teamer work.
  • Strong written communication — your reports become the patch ticket.
  • Comfort working in policy-grey areas with clear documentation of what was attempted and why.
  • Familiarity with prompt-injection, jailbreak, and policy-bypass taxonomies.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Construct a 5-turn adversarial conversation that bypasses a specific policy clause and write up the patch ticket.
  • Score a model's defenses against a known jailbreak pattern across 20 variants.
  • Propose a new red-team rubric category after spotting an emerging attack vector.
  • Reproduce a failure another reviewer reported and confirm the severity tag.

Nice to have

  • Background in offensive security, AppSec, or trust & safety operations.
  • Experience publishing or reproducing public adversarial-ML research.
  • Multilingual fluency for cross-language attack testing.

Skills

  • Adversarial prompting
  • Red-team analysis
  • Policy taxonomy
  • Failure documentation
  • AI Safety Red Teamer evaluation

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

$70–$84 / hr

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the AI Safety Red Teamer [Remote] in Remote vacancy
  • AI Trainer Jobs is seeking a Bilingual Korean STEM Expert (PhD) for a remote AI Safety red-team track. You will craft adversarial prompts, document failures with rigor, and map each jailbreak to the violated rubric to help patch gaps. Reviewers simulate attacks, assess... 
    Suggested
    Remote job
    For contractors

    AI Trainer Jobs

    New York, NY
    4 days ago
  • AuraOne is seeking a Bilingual German Generalist Expert — AI Safety for a remote red-team track that stress-tests AI systems against adversarial prompts. Reviewers craft attack scenarios, document failures, and pair each jailbreak with the violated rubric clause so the... 
    Suggested
    Remote job

    AI Trainer Jobs

    New York, NY
    4 days ago
  • $70 - $84 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ..., Larry Summers , and Jack Dorsey . Position: AI Safety Red Teamer Type: Contract Compensation: $70–$84/hour Location... 
    Suggested
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    13 days ago
  • $70 - $84 per hour

    Gridnaut Recruiting is hiring a remote AI Safety Red Teamer contractor (pay $70-$84/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Apply through Gridnaut Recruiting for this remote AI Safety Red Teamer contractor opening.; Remote contractor... 
    Suggested
    Contract work
    For contractors
    Remote work
    Flexible hours

    Gridnaut Recruiting

    Remote
    27 days ago
  • AI Trainer Jobs is seeking a Bilingual Thai STEM Expert (PhD) for a remote AI safety red-team role. Reviewers craft attack scenarios, document failures, and map breaches to rubric clauses to guide patching. This contractor position emphasizes adversarial evaluation, policy... 
    Suggested
    For contractors
    Remote work

    AI Trainer Jobs

    New York, NY
    4 days ago
  • $68 - $72 per hour

    AI Trainer Jobs is seeking a bilingual Japanese STEM expert (PhD) for a remote AI Safety red-team role. You will design jailbreak prompts, document failures, and patch gaps to harden AI systems before deployment. Responsibilities include crafting adversarial scenarios,... 
    Remote job
    Hourly pay
    For contractors

    AI Trainer Jobs

    New York, NY
    4 days ago
  • $48 - $62 per hour

     ...Role Overview Red team conversational AI and generate high-quality human data that exposes vulnerabilities in language models. You will craft adversarial...  ...in production. Customers gain confidence in the safety of their AI because adversarial scenarios have been... 
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $24 - $35 per hour

     ...Role Overview Join a red team committed to finding real-world vulnerabilities in conversational AI. You will probe models and agents with adversarial inputs, document reproducible failures, and generate high-quality human data that helps make AI systems safer for customers... 
    Hourly pay
    Remote work
    Shift work

    SaidGig

    United States
    1 day ago
  • $16 - $22 per hour

     ...Help make conversational AI safer by probing models with adversarial inputs, identifying vulnerabilities, and producing actionable red-team data. This text-based remote role focuses on testing AI behavior across sensitive areas, including bias, misinformation, and harmful... 
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $48 - $62 per hour

     ...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production. The role focuses on adversarial testing, generating reproducible attack cases and datasets, and documenting failures so... 
    Hourly pay
    Remote work
    Flexible hours

    SaidGig

    United States
    1 day ago
  • $20 - $22 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Bengali Type: Contract...  ...: Remote Role Responsibilities Red team conversational AI models and agents... 
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    19 days ago
  • $17 - $25 per hour

     ...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers. This text-based role focuses on testing models for bias, misinformation, harmful behaviors, jailbreaks, and prompt-injection style... 
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $16 - $22 per hour

     ...Role Overview Help make conversational AI systems safer by testing them with adversarial...  ..., and creating high-quality red-team data. This text-based remote role involves...  ...completeness, appropriateness, and potential safety issues. Annotate failures, classify vulnerabilities... 
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $70 - $84 per hour

     ...Role Overview Lead adversarial evaluations of frontier AI models by designing and executing high-impact prompts and tests...  .... Document discovered vulnerabilities and contribute to safety benchmarking and red-teaming reports. Collaborate with AI researchers and... 
    Hourly pay
    Remote work
    Work visa

    SaidGig

    United States
    2 days ago
  • $48 - $62 per hour

     ...Probe conversational AI systems for weaknesses before they reach users. In this remote...  ...uncover vulnerabilities, create high-quality safety data, and produce practical findings that...  ...are provided. Key Responsibilities Red team conversational AI models and agents through... 
    Hourly pay
    Remote work

    SaidGig

    United States
    more than 2 months ago
  • $17 - $25 per hour

     ...Help strengthen conversational AI by testing it from an adversarial perspective. In this...  ...agents for weaknesses, create actionable safety data, and help identify risks before they...  ...Indonesian is required. Prior experience in AI red teaming, adversarial AI work,... 
    Hourly pay
    Remote work

    SaidGig

    United States
    more than 2 months ago
  • $16 - $22 per hour

     ...Help strengthen conversational AI by probing models for weaknesses, documenting risks, and producing high-quality safety data. This text-based remote role focuses on adversarial testing and review of AI outputs, including work involving sensitive subjects such as bias,... 
    Hourly pay
    Remote work

    SaidGig

    United States
    3 days ago
  • $25 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Indonesian Type:...  ...Location: Remote Role Responsibilities Red team conversational AI models and agents... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    8 days ago
  • $48 - $62 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Larry Summers, and Jack Dorsey. Position: AI Safety Experts — English & Danish Type:Contract Compensation...  ...Location:Remote Role Responsibilities Red team conversational AI models and agents... 
    Contract work
    Summer work
    Remote work

    Remote Jobs

    New York, NY
    4 days ago
  •  ...Role Overview Help assess frontier AI systems at the boundary between legitimate radiological safety work and potentially dangerous misuse. You will apply practitioner...  ...reporting across multiple authorized users. Red-teaming experience is preferred. Technical writing... 
    For contractors
    Remote work

    SaidGig

    United States
    23 days ago
  • Mercor is assembling a red team for AI safety. You will review model outputs, simulate adversarial inputs, and surface vulnerabilities across prompts and conversations. You will generate high-quality human data, annotate failures, classify risks, and produce reproducible... 
    Remote job

    aitrainer

    New York, NY
    4 days ago
  • Mercor seeks a remote AI red-teaming specialist fluent in English and Portuguese to assess safety of conversational models. You will simulate jailbreaks, prompt injections, and misuse scenarios while documenting results for customers. You will generate high-quality human... 
    Remote job

    Mercor

    New York, NY
    4 days ago
  • AuraOne is seeking a Bilingual Korean Generalist Expert — AI Safety for a remote red-team track that stress-tests AI systems against adversarial prompts. Reviewers craft attack scenarios, document failure modes, and map each successful jailbreak to the violated rubric... 
    Remote work

    AI Trainer Jobs

    New York, NY
    4 days ago
  • Mercor is building a remote AI safety red team to probe conversational models and surface vulnerabilities in high-sensitivity topics. You will annotate failures, classify risks, and generate reproducible reports for customers to act on. Before exposure to content, topics... 
    Remote job

    Mercor

    New York, NY
    2 days ago
  • AuraOne is seeking a Bilingual Vietnamese STEM Expert (PhD) for an AI Safety remote red-team role. You will design adversarial prompts to test AI systems, document failures with rigorous reproduction steps, and pair each jailbreak with the clause it violated so the safety... 
    Remote job
    For contractors

    AI Trainer Jobs

    New York, NY
    4 days ago
  • AuraOne is seeking a Biology Expert (PhD) for its remote AI Safety red-team track. You design adversarial prompts to probe known weaknesses, document each failure with reproduction steps, and tie it to policy clauses for patching. You will score model defenses across turns... 
    Remote job
    Contract work
    10 hours per week

    AI Trainer Jobs

    New York, NY
    4 days ago
  • AuraOne is seeking a Bilingual Dutch (Belgium) STEM Expert (PhD) to work remotely as a contractor on AI Safety red-teaming. You will craft adversarial prompts, document failures, and pair each jailbreak with the violated rubric clause so the safety team can patch gaps.... 
    Remote job
    Hourly pay
    For contractors

    AI Trainer Jobs

    New York, NY
    4 days ago
  • AuraOne is seeking a bilingual Indonesian STEM expert (PhD) for a remote AI safety red-team role that stress-tests AI systems against adversarial prompts. Reviewers craft attack scenarios, document failures, and pair each successful jailbreak with the rubric clause it violated... 
    Remote job

    AI Trainer Jobs

    New York, NY
    4 days ago
  • AuraOne is seeking a Bilingual Chinese STEM Expert (PhD) for AI Safety, a remote red-team track that stresses testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document failure modes, and pair each successful jailbreak with the violated rubric... 
    For contractors
    Remote work

    AI Trainer Jobs

    New York, NY
    4 days ago
  • Obsidian is looking for data experts to join a red team that probes AI models with adversarial inputs. This role requires fluency in both English...  ...AI and cybersecurity, emphasizing the importance of AI safety. Work will be structured, with clear guidelines and wellness... 
    Remote job

    Obsidian

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Safety Red Teamer [Remote]. Be the first to apply!