Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Safety Red Teamer

$70 - $84 per hour

SaidGig

Role Overview

Lead adversarial testing of frontier AI models by designing challenging prompts to surface jailbreaks, hallucinations, unsafe outputs, and policy failures. You will evaluate model behavior on complex, high-risk, and ambiguous "grey-area" topics, document vulnerabilities, and work with researchers to improve model alignment and robustness.

Key Responsibilities
  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.
Qualifications
  • Required : Bachelors degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.
  • Preferred : Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety; familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies; expertise in one or more grey-area domains such as cyber, biosecurity, political content, misinformation, or scientific safety.
Work Terms
  • Location: Remote.
  • Employment type: Hourly.
Compensation

70 - 84 hourly.

Why Join
  • Help secure and strengthen the next generation of frontier AI models.
  • Work on cutting-edge adversarial testing alongside leading AI researchers and safety teams.
  • Influence how AI systems respond to complex, real-world safety challenges.
Vacancy posted 20 hours ago
Similar jobs that could be interesting for youBased on the AI Safety Red Teamer in United States vacancy
  • Handshake is seeking an AI Red Teamer in Seattle, WA, to stress-test large language models by designing adversarial prompts. The role involves...  ...collaborating with a multidisciplinary team to strengthen AI safety. Ideal candidates should have strong experience with LLMs, an... 
    Suggested

    Handshake

    Seattle, WA
    3 days ago
  • $70 - $84 per hour

     ...AI Safety Red Teamer is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the... 
    Suggested
    Remote job
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    11 days ago
  • $70 - $84 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ..., Larry Summers , and Jack Dorsey . Position: AI Safety Red Teamer Type: Contract Compensation: $70–$84/hour Location... 
    Suggested
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    3 days ago
  • A technology consulting firm is seeking an independent contractor for an AI Red-Teaming role focused on enhancing AI safety through probing models and generating high-quality datasets. Responsibilities include uncovering vulnerabilities, documenting findings, and contributing... 
    Suggested
    Remote job
    For contractors

    YO IT Consulting

    San Francisco, CA
    1 day ago
  • $17 - $25 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Vietnamese Type:...  ...Location: Remote Role Responsibilities Red team conversational AI models and agents... 
    Suggested
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    6 days ago
  • Why This Role Exists We believe the safest AI is the one that’s already been attacked —...  .... That’s why we’re building a pod of AI Red-Teamers: human data experts who probe AI models...  ...driven AI red-teaming at the frontier of safety Play a direct role in making AI systems... 
    Contract work
    For contractors
    Remote work
    Flexible hours

    YO IT Consulting

    San Francisco, CA
    1 day ago
  •  ...known as ActiveFence) is a leading trust, safety, and security company. Just like 'Alice'...  ...the rabbit hole into the emerging world of AI and focus on safeguarding these...  ...-based work for the \"best of the best\" red-teamers in the industry. Work is on-demand and project... 
    Freelance

    Socket.dev

    New York, NY
    2 days ago
  • About the Role As an AI Red Teamer, you will stress-test large language models by intentionally trying to break them. Rather than checking...  ...weaknesses, and unexpected behaviors. Your work directly supports AI safety and model robustness for leading research labs. This is a... 

    Cacheflow

    Seattle, WA
    4 days ago
  • $29 - $45 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Summers , and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (global) Type...  ...Location: Remote Role Responsibilities Red team conversational AI models and agents by... 
    Contract work
    Summer work
    Remote work

    Mercor Inc

    San Francisco, CA
    1 day ago
  • $54 - $111 per hour

    A leading AI consultancy is seeking an AI Red-Teamer for adversarial AI testing. This remote role involves conducting advanced testing, generating critical reports, and identifying vulnerabilities in AI models. Ideal candidates will have prior experience in testing and... 
    Remote job
    Hourly pay
    10 hours per week
    Flexible hours

    Crossing Hurdles

    New York, NY
    1 day ago
  • Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human... 
    Remote job

    Neon

    New York, NY
    2 days ago
  • Mercor is seeking a remote Senior Red Team AI Specialist to test conversational agents and AI models against adversarial inputs. You will annotate vulnerabilities, surface systemic risks, and produce reproducible attack cases to help customers strengthen their AI systems... 
    Remote work

    Neon

    New York, NY
    2 days ago
  • Mercor is recruiting AI safety experts to remotely assess and strengthen AI systems. You will conduct red team activities, identify jailbreaks and misuse cases, and produce actionable data to help clients improve model safety. Ideal candidates fluently speak English and... 
    Remote job
    Contract work

    Mercor

    New York, NY
    2 days ago
  • Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate actionable data and reports to help customers harden their AI. Ideal candidates... 
    Remote job

    Mercor

    New York, NY
    3 days ago
  • Obsidian is seeking a remote AI red team expert responsible for probing AI models with adversarial inputs to uncover vulnerabilities. You...  ...document findings, and generate high-quality data to improve AI safety. The ideal candidate brings red-teaming experience, a curious... 
    Remote job

    Obsidian

    San Francisco, CA
    1 day ago
  • A technology firm is seeking an AI Red-Teamer for adversarial AI testing. The ideal candidate will be fluent in English and Chinese (Mandarin) and have prior experience in red teaming, cybersecurity, or adversarial testing. The role involves red teaming conversational AI... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    1 day ago
  • Anthropic is seeking a Red Team Engineer to uncover vulnerabilities across our AI product ecosystem. You will simulate adversaries, perform comprehensive adversarial testing, and explore novel abuse vectors in our deployed systems. You will design kill-chain style attacks... 
    Flexible hours

    Doist

    San Francisco, CA
    10 hours ago
  • Mercor is building a fully remote red team to probe AI models for safety, resilience, and ethical considerations. You will work on adversarial inputs, jailbreaks, and vulnerability discovery while producing actionable data and reproducible reports for customers. This role... 
    Remote job

    Mercor

    San Francisco, CA
    1 day ago
  • AuraOne is seeking a remote contractor to perform Chemical Safety Risk Evaluation by designing adversarial prompts, documenting failures...  ...least 10 hours per week. Applicants should have experience in red-teaming AI systems, familiarity with jailbreaking and policy-bypass... 
    Remote job
    For contractors
    10 hours per week

    AuraOne

    New York, NY
    3 days ago
  • Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe conversational AI, annotate vulnerabilities, and generate high-quality attack data for safer deployments. You’ll work across... 
    Remote job

    Neon

    New York, NY
    2 days ago
  • A leading AI research organization is seeking a candidate for an Automated Red Teaming role focused on building scalable, research-driven systems to identify and fix...  ...position demands collaboration with various teams to ensure system safety and efficacy. #J-18808-Ljbffr OpenAI

    OpenAI

    San Francisco, CA
    2 days ago
  • A leading AI research firm in New York, NY, is looking for a specialist to oversee the red-teaming and adversarial evaluation pipeline for their models. The ideal candidate...  ...field and have a deep understanding of LLM safety and adversarial techniques. A strong software... 

    Reflection AI

    New York, NY
    20 hours ago
  • Mercor is seeking a remote Red Team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate actionable red team data for safer AI systems. You will annotate failures, classify vulnerabilities, and document risks using established taxonomies... 
    Remote job

    Mercor

    San Francisco, CA
    4 days ago
  • $26 per hour

     ...Remote Commitment: 10-40 hours/week Role Responsibilities Red-team conversational AI systems using jailbreaks, prompt injections, misuse cases,...  ...sensitive or high-risk scenarios in alignment with defined safety guidelines. Requirements Native-level fluency in both English... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    1 day ago
  •  ...RevolutionBecome a Cybersecurity Senior Advisor - Red Team Lead at Southern California Edison (...  ...environments, including understanding of safety and reliability constraints.Demonstrated...  ..., and emerging technologies (including AI/ML) to scale Red Team operations and... 
    Remote work
    Relocation

    Edison International

    Rosemead, CA
    1 day ago
  •  ...AI Red Team Engineer We're looking for an AI Red Team Engineer to break LLM-powered systems responsibly, automate the repetitive attacks...  ...prove it, script it, and write it up. White Circle is an AI Safety company building the safety, reliability, and optimization layer... 
    Local area
    Remote work

    White Circle

    United States
    2 days ago
  • Mercor is building a remote AI red team to probe and test conversational models for vulnerabilities, bias, and safety gaps. This role emphasizes reproducible data and actionable findings across projects and customers. You bring prior red teaming experience in AI or cybersecurity... 
    Remote job

    Obsidian

    New York, NY
    1 day ago
  • Mercor is assembling a remote red team to test large language models and AI agents by probing for jailbreaks, prompt injections, and systemic risks. You will work with English and Vietnamese inputs to surface vulnerabilities and produce actionable data for customers. You... 
    Remote job

    Mercor

    New York, NY
    3 days ago
  • $30 - $60 per hour

    Portland Seed Fund is seeking part-time Red Teaming Experts in Seattle to support AI safety evaluation campaigns. Candidates will design and simulate AI conversations, identifying risks and evaluating performance. This role requires strong analytical skills and creative... 
    Part time

    Portland Seed Fund

    Seattle, WA
    3 days ago
  • $26 per hour

    A tech consulting firm is seeking a remote contractor to engage in AI red teaming activities. The ideal candidate will work on evaluating conversational AI systems and generating human data through careful documentation and classification of vulnerabilities. Fluency in... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week
    Flexible hours

    Crossing Hurdles

    New York, NY
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Safety Red Teamer. Be the first to apply!