Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Safety Red Teamer [Remote]

$70 - $84 per hour

AuraOne Human Data

Remote
  • Remote job

AI Safety Red Teamer is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.

Why this role matters

Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.

Responsibilities

  • Design adversarial prompts that probe known weakness classes (jailbreak, policy bypass, prompt injection) for AI Safety Red Teamer assignments.
  • Document every successful attack with reproduction steps and the policy clause it violated.
  • Score model defenses across single-turn and multi-turn conversations.
  • Triage emerging attack vectors and route them to the safety team with severity ratings.
  • Maintain a personal library of attack patterns and propose new red-team rubrics.
  • Calibrate against the broader red-team cohort to keep coverage and severity consistent.

Qualifications

  • Demonstrated experience red-teaming AI systems, security research, or adversarial ML work for AI Safety Red Teamer work.
  • Strong written communication — your reports become the patch ticket.
  • Comfort working in policy-grey areas with clear documentation of what was attempted and why.
  • Familiarity with prompt-injection, jailbreak, and policy-bypass taxonomies.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Construct a 5-turn adversarial conversation that bypasses a specific policy clause and write up the patch ticket.
  • Score a model's defenses against a known jailbreak pattern across 20 variants.
  • Propose a new red-team rubric category after spotting an emerging attack vector.
  • Reproduce a failure another reviewer reported and confirm the severity tag.

Nice to have

  • Background in offensive security, AppSec, or trust & safety operations.
  • Experience publishing or reproducing public adversarial-ML research.
  • Multilingual fluency for cross-language attack testing.

Skills

  • Adversarial prompting
  • Red-team analysis
  • Policy taxonomy
  • Failure documentation
  • AI Safety Red Teamer evaluation

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

$70–$84 / hr

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 12 days ago
Similar jobs that could be interesting for youBased on the AI Safety Red Teamer [Remote] in Remote vacancy
  • $70 - $84 per hour

    Gridnaut Recruiting is hiring a remote AI Safety Red Teamer contractor (pay $70-$84/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Apply through Gridnaut Recruiting for this remote AI Safety Red Teamer contractor opening.; Remote contractor... 
    Suggested
    Contract work
    For contractors
    Remote work
    Flexible hours

    Gridnaut Recruiting

    Remote
    4 days ago
  • $48 - $62 per hour

     ...Role Overview Work as a red team human-data expert probing conversational AI models and agents to find vulnerabilities before they reach production. The role focuses on adversarial testing, generating reproducible attack cases and datasets, and documenting failures so... 
    Suggested
    Hourly pay
    Remote work
    Flexible hours

    SaidGig

    United States
    1 day ago
  • $17 - $25 per hour

     ...Role Overview Join a red team that probes conversational AI with adversarial inputs to find vulnerabilities before they reach customers. This text-based role focuses on testing models for bias, misinformation, harmful behaviors, jailbreaks, and prompt-injection style... 
    Suggested
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $24 - $35 per hour

     ...Role Overview Join a red team committed to finding real-world vulnerabilities in conversational AI. You will probe models and agents with adversarial inputs, document reproducible failures, and generate high-quality human data that helps make AI systems safer for customers... 
    Suggested
    Hourly pay
    Remote work
    Shift work

    SaidGig

    United States
    1 day ago
  • $48 - $62 per hour

     ...Role Overview Red team conversational AI and generate high-quality human data that exposes vulnerabilities in language models. You will craft adversarial...  ...in production. Customers gain confidence in the safety of their AI because adversarial scenarios have been... 
    Suggested
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $70 - $84 per hour

     ...Role Overview Lead adversarial evaluations of frontier AI models by designing and executing high-impact prompts and tests...  .... Document discovered vulnerabilities and contribute to safety benchmarking and red-teaming reports. Collaborate with AI researchers and... 
    Hourly pay
    Remote work
    Work visa

    SaidGig

    United States
    1 day ago
  •  ...institutions. In 2025, we started Handshake AI and built the fastest-growing AI data...  ...largest scale. About the Role As a CBRNE Red Teamer, you will evaluate whether AI models...  ...models for dangerous knowledge gaps in their safety guardrails, testing whether they can be manipulated... 
    Remote job
    Immediate start

    Dorado

    Brooklyn, NY
    3 days ago
  •  ...Role Exists At Mercor, we believe the safest AI is the one that’s already been attacked — by us. We are assembling a red team for this project - human data experts who...  ...surprises in production Mercor customers trust the safety of their AI because you’ve already probed it... 
    Remote work

    Mercor

    New York, NY
    4 days ago
  • $48 - $62 per hour

     ...Probe conversational AI systems for weaknesses before they reach users. In this remote...  ...uncover vulnerabilities, create high-quality safety data, and produce practical findings that...  ...are provided. Key Responsibilities Red team conversational AI models and agents through... 
    Hourly pay
    Remote work

    SaidGig

    United States
    a month ago
  • $48 - $62 per hour

     ...Help strengthen conversational AI systems by probing them with adversarial inputs, identifying vulnerabilities, and producing practical safety data. This remote role focuses on text-based AI red teaming in both English and Danish. Key Responsibilities Test conversational... 
    Hourly pay
    Remote work

    SaidGig

    United States
    a month ago
  • $17 - $25 per hour

     ...Help strengthen conversational AI by testing it from an adversarial perspective. In this...  ...agents for weaknesses, create actionable safety data, and help identify risks before they...  ...Indonesian is required. Prior experience in AI red teaming, adversarial AI work,... 
    Hourly pay
    Remote work

    SaidGig

    United States
    a month ago
  • $26 per hour

     ...Remote Commitment: 10-40 hours/week Role Responsibilities Red-team conversational AI systems using jailbreaks, prompt injections, misuse cases,...  ...sensitive or high-risk scenarios in alignment with defined safety guidelines. Requirements Native-level fluency in both English... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    2 days ago
  • Mercor seeks a remote AI red-teaming specialist fluent in English and Portuguese to assess safety of conversational models. You will simulate jailbreaks, prompt injections, and misuse scenarios while documenting results for customers. You will generate high-quality human... 
    Remote job

    Mercor

    New York, NY
    6 days ago
  • Mercor is building a remote AI safety red team to probe conversational models and surface vulnerabilities in high-sensitivity topics. You will annotate failures, classify risks, and generate reproducible reports for customers to act on. Before exposure to content, topics... 
    Remote job

    Mercor

    New York, NY
    4 days ago
  • Obsidian is looking for data experts to join a red team that probes AI models with adversarial inputs. This role requires fluency in both English...  ...AI and cybersecurity, emphasizing the importance of AI safety. Work will be structured, with clear guidelines and wellness... 
    Remote job

    Obsidian

    San Francisco, CA
    4 days ago
  • $48 - $62 per hour

     ...Role Overview Probe conversational AI systems and produce reproducible adversarial data that reveals vulnerabilities and improves model safety. This is a text-based red teaming role focused on discovering jailbreaks, prompt injection paths, bias and misinformation failures... 
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $48 - $62 per hour

     ...Role Overview Work as a human red team specialist probing conversational AI to find vulnerabilities, generate adversarial datasets, and produce reproducible reports that help customers harden their systems. This text-based role focuses on identifying failures related... 
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $17 - $25 per hour

     ...performs adversarial testing of conversational AI, creating reproducible attack cases,...  ...is text-based. Key Responsibilities Red team conversational AI models and agents,...  ...Increasing customer confidence in their AI safety because systems have been thoroughly probed... 
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $17 - $25 per hour

     ...Role Overview Join a red team of human data experts who probe conversational AI models with adversarial inputs to surface vulnerabilities and produce high-quality red team data that improves model safety. This text-based role focuses on reviewing AI outputs that touch... 
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $29 - $45 per hour

     ...Role Overview This role performs adversarial testing of conversational AI to find safety vulnerabilities and produce reproducible, human-generated red team data that clients can act on. Work is remote and text-based. The role includes reviewing model outputs that may... 
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $48 - $62 per hour

     ...Role Overview Lead adversarial testing of conversational AI in both English and Norwegian, probing models and agents to surface vulnerabilities...  ...communicated before any exposure. Key Responsibilities Red team conversational AI models and agents, including jailbreaks,... 
    Hourly pay
    Remote work

    SaidGig

    United States
    1 day ago
  • $20 - $22 per hour

     ...Overview Help strengthen conversational AI systems by probing them with adversarial inputs...  ...vulnerabilities, and producing actionable safety data. This remote, text-based role focuses...  ...English and Punjabi is required. Prior red-teaming experience in AI adversarial work,... 
    Hourly pay
    Remote work

    SaidGig

    United States
    18 hours ago
  • $130k - $200k

     ...About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies...  ...platforms. Our adversarial red teaming, model evaluations, and intelligence...  ...with engineers, analysts, red teamers, and subject-matter experts... 
    Remote work

    10a Labs

    United States
    1 day ago
  • $188.5k - $282.7k

     ...We're building SAGE , Rubrik's Semantic AI Governance Engine, which is the first system...  ...at the problems that matter most for AI safety and governance. Nature of the Specialized...  ...-built adversarial agents and automated red-teams whose outputs feed directly into the... 
    Permanent employment
    Full time
    Local area

    Rubrik

    Remote
    1 day ago
  • $128.99k - $185k

     ...customers. Our capabilities range from C5ISR, AI and Big Data, cyber operations and...  ...multi-repo workflows. Familiarity with Red Hat Enterprise Linux (RHEL). Experience...  ...ethics driven organization that puts people's safety and well-being first. Regardless of your... 
    Full time
    Interim role
    Local area
    Remote work
    Worldwide

    Huntington Ingalls Industries

    Remote
    1 day ago
  • $125k - $190k

    About the Role FAR.AI is hiring a Technical Project Manager to be...  ...world's most impactful frontier AI red‑teaming programmes. You will...  ...will work alongside our red‑teamers, researchers, and the rest of the...  ...environment such as frontier labs, AI safety organisations, technical... 
    Full time
    Contract work
    For contractors
    Remote work
    Visa sponsorship
    Shift work

    Aisafety

    Berkeley, CA
    2 days ago
  • $110k - $145k

     ...Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot...  ...the successful candidate to lead and participate directly in red teaming and adversarial evaluation, working in detail with... 
    Permanent employment
    Flexible hours

    Moonshotteam

    Atlanta, GA
    2 days ago
  • $184k - $287.5k

     ...design, build, and deploy secure agentic AI systems on NVIDIA’s accelerated computing...  ...software companies across cybersecurity, AI safety, infrastructure protection, and confidential...  ...-LLM, or NIM Operator.Experience with LLM red-teaming, AI safety evaluation, adversarial... 
    Full time
    Remote work

    Nvidia

    Texas
    2 days ago
  • $48 - $52 per hour

     ...Help improve the safety of advanced AI systems by using your French language expertise and cultural judgment to assess how models respond to sensitive...  ...Qualifications Experience reviewing, grading, or red-teaming written or technical content. Background in trust... 
    Hourly pay
    Immediate start
    Remote work

    SaidGig

    Europe
    11 days ago
  •  ...Role Overview Help improve the safety of advanced AI systems by using your Norwegian language fluency and cultural judgment to assess how models...  ...elsewhere are welcome. Experience reviewing, grading, or red-teaming written or technical content. Background in trust... 
    Hourly pay
    Immediate start
    Remote work

    SaidGig

    Remote
    11 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Safety Red Teamer [Remote]. Be the first to apply!