Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Remote AI Safety Red-Team Evaluator

AuraOne

New York, NY
  • Remote job

AuraOne is seeking a remote contractor to perform Chemical Safety Risk Evaluation by designing adversarial prompts, documenting failures with reproduction steps, and scoring defenses across conversations. The role requires clear documentation, strong written communication, and the ability to work asynchronously at least 10 hours per week. Applicants should have experience in red-teaming AI systems, familiarity with jailbreaking and policy-bypass tactics, and the discipline to maintain rigorous #J-18808-Ljbffr AuraOne

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Remote AI Safety Red-Team Evaluator in New York, NY vacancy
  • As part of AuraOne's safety-focused team, you will design 5‑turn attack scenarios, reproduce reported failures, and help harden AI systems before broader deployment. This is a contract position with US-eligibility suitable for remote workers. #J-18808-Ljbffr AuraOne
    Remote job
    Contract work

    AuraOne

    New York, NY
    1 day ago
  •  ...AI Safety and Red-Team Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    18 days ago
  • Mercor is building a fully remote red team to probe AI models for safety, resilience, and ethical considerations. You will work on adversarial inputs, jailbreaks, and vulnerability discovery while producing actionable data and reproducible reports for customers. This role... 
    Remote job

    Mercor

    San Francisco, CA
    1 day ago
  •  ...IT Consulting is looking for a Red-Teaming Quality Assurance Lead to...  ...quality and consistency across AI red-teaming projects. This remote position involves evaluating AI-generated evaluations and providing...  ...will have a background in AI safety or cybersecurity, strong... 
    Remote job

    YO IT Consulting

    Seattle, WA
    3 days ago
  • YO IT Consulting is looking for a Remote Red-Teaming Quality Assurance Lead. In this contractor role, you will be responsible for ensuring quality across AI red-teaming projects. You will evaluate AI-generated safety evaluations and maintain documentation while managing... 
    Remote job
    For contractors

    YO IT Consulting

    Chicago, IL
    2 days ago
  • YO IT Consulting is hiring a Red-Teaming Quality Assurance Lead to oversee quality and consistency across AI safety evaluation projects. This remote role demands strong expertise in AI safety and red-teaming, alongside excellent communication and organizational skills.... 
    Remote job

    YO IT Consulting

    Atlanta, GA
    1 day ago
  • Mercor is seeking a remote Red Team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate actionable red team data for safer AI systems. You will annotate failures, classify vulnerabilities, and document risks using established taxonomies... 
    Remote job

    Mercor

    San Francisco, CA
    4 days ago
  • Mercor is seeking a remote Senior Red Team AI Specialist to test conversational agents and AI models against adversarial inputs. You will annotate vulnerabilities, surface systemic risks, and produce reproducible attack cases to help customers strengthen their AI systems... 
    Remote work

    Neon

    New York, NY
    2 days ago
  • YO IT Consulting is looking for a Red-Teaming Quality Assurance Lead for a remote contract position. You will oversee quality assurance across AI red-teaming and safety evaluation projects, ensuring consistency and adherence to guidelines. The ideal candidate has strong... 
    Remote job
    Contract work

    YO IT Consulting

    Dallas, TX
    1 day ago
  • Obsidian is seeking a remote AI red team expert responsible for probing AI models with adversarial inputs to uncover vulnerabilities. You will...  ...findings, and generate high-quality data to improve AI safety. The ideal candidate brings red-teaming experience, a curious... 
    Remote job

    Obsidian

    San Francisco, CA
    1 day ago
  • Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate actionable data and reports to help customers harden their AI. Ideal candidates... 
    Remote job

    Mercor

    New York, NY
    3 days ago
  • Mercor is recruiting AI safety experts to remotely assess and strengthen AI systems. You will conduct red team activities, identify jailbreaks and misuse cases, and produce actionable data to help clients improve model safety. Ideal candidates fluently speak English and... 
    Remote job
    Contract work

    Mercor

    New York, NY
    2 days ago
  • Mercor is assembling a remote red team to test and strengthen AI systems through adversarial inputs. You will annotate failures, classify vulnerabilities, and surface systemic risks using rigorous taxonomies and playbooks. The role rewards clear communication of risks to... 
    Remote job

    Neon

    New York, NY
    2 days ago
  • YO IT Consulting is seeking a Red-Teaming Quality Assurance Lead for a remote contract position. In this role, you...  ...quality and trainer performance across AI safety projects and provide feedback to...  .... You will be responsible for evaluating AI outputs, managing quality... 
    Remote job
    Contract work

    YO IT Consulting

    Phoenix, AZ
    1 day ago
  • Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human... 
    Remote job

    Neon

    New York, NY
    2 days ago
  • Mercor is assembling a remote red team to test large language models and AI agents by probing for jailbreaks, prompt injections, and systemic risks. You will work with English and Vietnamese inputs to surface vulnerabilities and produce actionable data for customers. You... 
    Remote job

    Mercor

    New York, NY
    3 days ago
  • YO IT Consulting is looking for a Red-Teaming Quality Assurance Lead to oversee quality across AI safety evaluation projects. You will ensure training data quality through...  ...-teaming. You will manage quality workflows remotely while maintaining documentation and... 
    Remote job

    YO IT Consulting

    Los Angeles, CA
    1 day ago
  • $26 per hour

    A tech consulting firm is seeking a remote contractor to engage in AI red teaming activities. The ideal candidate will work on evaluating conversational AI systems and generating human data through careful documentation and classification of vulnerabilities. Fluency in... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week
    Flexible hours

    Crossing Hurdles

    New York, NY
    1 day ago
  • Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe conversational AI, annotate vulnerabilities, and generate high-quality attack data for safer deployments. You’ll work across... 
    Remote job

    Neon

    New York, NY
    2 days ago
  • Mercor is building a remote AI red team to probe and test conversational models for vulnerabilities, bias, and safety gaps. This role emphasizes reproducible data and actionable findings across projects and customers. You bring prior red teaming experience in AI or cybersecurity... 
    Remote job

    Obsidian

    New York, NY
    1 day ago
  •  ...AI Red Team Engineer We're looking for an AI Red Team Engineer to break LLM-powered systems...  ...write it up. White Circle is an AI Safety company building the safety, reliability...  ...tool/function calling, and LLM-as-judge evaluation. Familiarity with OWASP LLM Top 10, OWASP... 
    Remote work
    Local area

    Pumpkin Intelligence, Inc.

    United States
    4 days ago
  • $26 per hour

     ...Compensation: $26/hour Location: Remote Commitment: 10-40 hours/week Role Responsibilities Red-team conversational AI systems using jailbreaks,...  ...and adversarial test cases. Evaluate AI outputs across sensitive...  ...in alignment with defined safety guidelines. Requirements Native... 
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    1 day ago
  • Location Remote Fluent Language Skills Required English...  ...we believe the safest AI is the one that’s...  ...us. We are assembling a red team for this project - human...  ...strengthen customer AI systems Evaluation coverage expands: more...  ...customers trust the safety of their AI because you... 
    Remote work

    Obsidian

    San Francisco, CA
    1 day ago
  • $29 - $45 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ..., and Jack Dorsey . Position: AI Safety Experts — English & Portuguese (...  ...Compensation: $29-$45/hour Location: Remote Role Responsibilities Red team conversational AI models and agents... 
    Remote work
    Contract work
    Summer work

    Mercor Inc

    San Francisco, CA
    1 day ago
  •  ...Mental Health Clinical AI Evaluator is a remote clinical-review track for evaluating...  ...adherence; flag patient-safety issues; and document the corrected...  ...reasoning so the modeling team can close the gap. Why this...  ...reasoning, dosing, and red-flag handling on a structured... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    9 days ago
  •  ...Cybersecurity Senior Advisor - Red Team Lead at Southern California...  ...environments, including understanding of safety and reliability constraints....  ...technologies (including AI/ML) to scale Red Team...  ...days with the option to work remotely on the remaining days.  Unless... 
    Remote work
    Relocation

    Edison International

    Rosemead, CA
    1 day ago
  • $60 - $70 per hour

     ...technical talent with leading AI research labs....  ...Dorsey . Position: AI Safety Practitioner Type: Contract...  ...–$70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses for...  ...researchers and safety teams on ongoing evaluation initiatives... 
    Remote work
    Contract work
    Summer work

    Mercor

    New York, NY
    6 days ago
  • $145.9k - $234.2k

     ...manufacturing environments with a focus on AI and the attack surface it exposes. You...  ...infrastructure. This role reports to the Red Team Program Lead and works very deeply and closely...  .... #LI-CK1-SummaryLocation: Cambridge, Massachusetts; Remote - US; DigitalType: Full time
    Remote work
    Permanent employment
    Full time
    Work at office
    Work from home

    Moderna Therapeutics

    Cambridge, MA
    2 days ago
  • $48 - $62 per hour

     ...Overview Probe conversational AI systems and produce...  ...vulnerabilities and improves model safety. This is a text-based red teaming role focused on...  ...their AI systems. Expand evaluation coverage so more scenarios...  ...Work Terms Location: Remote. Engagement type: hourly... 
    Remote work
    Hourly pay

    SaidGig

    United States
    7 days ago
  • $22 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Jack Dorsey . Position: AI Safety Experts — English & Malayalam...  ...Compensation: $20–$22/hour Location: Remote Role Responsibilities Red team conversational AI models and... 
    Remote job
    Contract work
    Summer work

    Mercor

    New York, NY
    27 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Remote AI Safety Red-Team Evaluator. Be the first to apply!