Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Remote Adversarial ML Specialist AI Safety Red Team

Obsidian

Mercor is building a remote red team for AI safety. You will lead adversarial testing of conversational AI models, focusing on jailbreaks, prompt injections, misuse, and bias exploitation. You will generate high-quality human data, annotate failures, classify vulnerabilities, and document findings in reproducible reports and datasets that help customers strengthen their AI systems. Ideal candidates bring prior red teaming experience, a curious adversarial mindset, structured approaches, and the #J-18808-Ljbffr Obsidian

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Remote Adversarial ML Specialist AI Safety Red Team in San Francisco, CA vacancy
  • Obsidian is seeking a remote AI red team expert responsible for probing AI models with adversarial inputs to uncover vulnerabilities. You will apply structured testing frameworks...  ...and generate high-quality data to improve AI safety. The ideal candidate brings red-teaming... 
    Remote job

    Obsidian

    San Francisco, CA
    2 days ago
  • Mercor is building a remote AI safety red team to probe conversational models and surface vulnerabilities in high-sensitivity topics. You will...  ...resources. You should bring prior red-teaming experience in AI adversarial work or cybersecurity, be curious and adversarial,... 
    Remote job

    Mercor

    New York, NY
    4 days ago
  • Mercor is seeking a remote red-team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate high-quality data for safer AI systems...  ...and collaborate across projects for robust safety testing. This role focuses on reproducible artifacts... 
    Remote job

    Obsidian

    San Francisco, CA
    4 days ago
  • $26 per hour

    A tech consulting firm is seeking a remote contractor to engage in AI red teaming activities. The ideal candidate will work on evaluating conversational...  ...Spanish is required, along with prior experience in adversarial AI testing or cybersecurity. This is a flexible hourly... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week
    Flexible hours

    Crossing Hurdles

    New York, NY
    2 days ago
  • $26 per hour

     ...Compensation: $26/hour Location: Remote Commitment: 10-40 hours/week Role Responsibilities Red-team conversational AI systems using jailbreaks,...  ..., and multi-turn adversarial strategies. Generate high-quality...  ...in alignment with defined safety guidelines. Requirements Native... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    2 days ago
  • Location Remote Fluent Language Skills Required...  ...believe the safest AI is the one that’s...  ...We are assembling a red team for this project -...  ...probe AI models with adversarial inputs, surface...  ...Specialties Adversarial ML: jailbreak datasets...  ...trust the safety of their AI because... 
    Remote work

    Obsidian

    San Francisco, CA
    2 days ago
  • $20 - $22 per hour

     ...talent with leading AI research labs. Headquartered...  .... Position: AI Safety Experts — English &...  ...Location: Remote Role Responsibilities Red team conversational AI models...  ...experience in AI adversarial work ,...  ...Experience with Adversarial ML , Cybersecurity ,... 
    Remote work
    Contract work
    Summer work

    Mercor

    New York, NY
    5 days ago
  • $60 - $90 per hour

     ...technical talent with leading AI research labs. Headquartered...  ...Dorsey . Position: LLM Red Team Specialist — Failure Modes & Edge Cases...  ...$60–$90/hour Location: Remote Commitment: 35 hours/week...  ...frontier AI models on coding, ML , and analysis tasks to... 
    Remote work
    Contract work
    Summer work

    Mercor

    New York, NY
    22 days ago
  • $48 - $62 per hour

     ...strengthen conversational AI systems by probing them with adversarial inputs, identifying...  ...human-generated safety data. This text-...  ...and attack cases that teams can use to improve AI...  ...Prior experience in AI red teaming, adversarial...  .... Work Terms Remote, hourly engagement.... 
    Remote work
    Hourly pay

    SaidGig

    United States
    a month ago
  • $20 - $22 per hour

     ...Overview Help strengthen conversational AI systems by probing them as an adversary. This remote, text-based role focuses on...  ...vulnerabilities, generating high-quality red-team data, and creating reproducible findings that improve AI safety and reliability. Key... 
    Remote work
    Hourly pay

    SaidGig

    United States
    18 days ago
  •  ...AI Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document...  ...clause it violated so the safety team can patch the gap. Why...  ...research, or adversarial ML work for AI Red Team Specialist... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  • $29 - $45 per hour

     ...Attack and probe conversational AI systems in English and...  ...vulnerabilities, produce reproducible adversarial data, and help make deployed...  .... Key Responsibilities Red team conversational AI models and...  ...thinking. Work Terms Remote, text-based engagement.... 
    Remote work
    Hourly pay

    SaidGig

    United States
    a month ago
  • $25 per hour

     ...talent with leading AI research labs. Headquartered...  .... Position: AI Safety Experts — English &...  ...Location: Remote Role Responsibilities Red team conversational AI models...  ...experience in AI adversarial work ,...  ...Experience in Adversarial ML : jailbreak datasets... 
    Remote job
    Contract work
    Summer work

    Mercor

    San Francisco, CA
    9 days ago
  • $24 - $35 per hour

     ...talent with leading AI research labs. Headquartered...  .... Position: AI Safety Experts — English &...  ...Location: Remote Role Responsibilities Red team conversational AI models...  ...experience in AI adversarial work ,...  ...Experience in Adversarial ML , Cybersecurity ,... 
    Remote work
    Contract work
    Summer work

    Mercor

    San Francisco, CA
    9 days ago
  •  ...progressive technology company is seeking a remote AI Red Teamer to assess conversational AI...  ...and identify vulnerabilities through red teaming exercises. Candidates should be fluent...  ...with a background in cybersecurity or adversarial testing. This role involves producing actionable... 
    Remote work
    Hourly pay

    Crossing Hurdles

    New York, NY
    2 days ago
  •  ...is seeking a research-focused professional to join a GenAI red-team within a leading AI lab network. The role involves probing frontier models,...  ...findings for reproducibility. This full-time W-2 position is remote within the United States, requiring about 35 hours per... 
    Remote work
    Full time

    Mercor

    New York, NY
    2 days ago
  •  ...Capability Elicitation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack...  ...clause it violated so the safety team can patch the gap. Why...  ...security research, or adversarial ML work for Capability... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    26 days ago
  • Mercor is building a fully remote red team to probe AI models for safety, resilience, and ethical considerations. You will work on adversarial inputs, jailbreaks, and vulnerability discovery while producing actionable data and reproducible reports for customers. This role... 
    Remote job

    Mercor

    San Francisco, CA
    2 days ago
  • Mercor is seeking a remote Red Team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate actionable red team data for safer AI systems. You will annotate failures, classify vulnerabilities, and document risks using established taxonomies... 
    Remote job

    Mercor

    San Francisco, CA
    5 days ago
  • Mercor is building a remote red-team for AI safety, focusing on adversarial testing of conversational models. You’ll craft attacks, annotate failures, and surface vulnerabilities while following established taxonomies and playbooks. Documentation of results and reproducible... 
    Remote job

    Neon

    New York, NY
    3 days ago
  • Mercor is seeking a remote Senior Red Team AI Specialist to test conversational agents and AI models against adversarial inputs. You will annotate vulnerabilities, surface systemic risks, and produce reproducible attack cases to help customers strengthen their AI systems... 
    Remote work

    Neon

    New York, NY
    3 days ago
  • Mercor is assembling a remote red team to test and strengthen AI systems through adversarial inputs. You will annotate failures, classify vulnerabilities, and surface systemic risks using rigorous taxonomies and playbooks. The role rewards clear communication of risks to... 
    Remote job

    Neon

    New York, NY
    3 days ago
  • Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate actionable data and reports to help customers harden their AI. Ideal candidates... 
    Remote job

    Mercor Inc

    New York, NY
    4 days ago
  • Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human... 
    Remote job

    Neon

    New York, NY
    3 days ago
  • Mercor is building a red team to probe AI models with adversarial inputs, surface vulnerabilities, and generate red team data that strengthens safety for customers. This is a text-based, higher-sensitivity project where participation is optional and guided by clear guidelines... 
    Remote job

    Mercor

    San Francisco, CA
    1 day ago
  • Mercor is assembling a remote red team tasked with probing AI models for vulnerabilities, bias and safety issues. You will generate high-quality data, annotate failures,...  ...and Vietnamese fluency, curiosity, and an adversarial mindset. The project emphasizes safe, guided... 
    Remote job

    Obsidian

    New York, NY
    2 days ago
  • Mercor is building a remote AI red-team program to probe and stress test conversational models for bias, misuse, and safety concerns. The role involves designing adversarial tests, annotating data, and producing reproducible attack datasets to help customers strengthen... 
    Remote job

    Mercor

    New York, NY
    3 days ago
  • Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe conversational AI, annotate vulnerabilities, and generate high-quality attack data for safer deployments. You’ll work across... 
    Remote job

    Neon

    New York, NY
    3 days ago
  • $22 per hour

     ...talent with leading AI research labs. Headquartered...  .... Position: AI Safety Experts — English &...  ...20-$22/hourLocation:Remote Role Responsibilities Red team conversational AI...  ...teaming experience in AI adversarial work, cybersecurity,...  ...in Adversarial ML, Cybersecurity, or socio... 
    Remote job
    Summer work

    United States Digital Space LLC

    New York, NY
    5 days ago
  •  ...AI Red Team Engineer We're looking for an AI Red Team Engineer to break LLM-powered systems...  ...conversations. You'll own hands-on adversarial testing end to end: find the failure, prove...  ...write it up. White Circle is an AI Safety company building the safety,... 
    Remote work
    Local area

    Pumpkin Intelligence, Inc.

    United States
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Remote Adversarial ML Specialist AI Safety Red Team. Be the first to apply!