Remote Adversarial ML Specialist AI Safety Red Team
Obsidian
- Remote job
Mercor is building a remote red team for AI safety. You will lead adversarial testing of conversational AI models, focusing on jailbreaks, prompt injections, misuse, and bias exploitation. You will generate high-quality human data, annotate failures, classify vulnerabilities, and document findings in reproducible reports and datasets that help customers strengthen their AI systems. Ideal candidates bring prior red teaming experience, a curious adversarial mindset, structured approaches, and the #J-18808-Ljbffr Obsidian
- Obsidian is seeking a remote AI red team expert responsible for probing AI models with adversarial inputs to uncover vulnerabilities. You will apply structured testing frameworks... ...and generate high-quality data to improve AI safety. The ideal candidate brings red-teaming...Remote job
- Mercor is building a remote AI safety red team to probe conversational models and surface vulnerabilities in high-sensitivity topics. You will... ...resources. You should bring prior red-teaming experience in AI adversarial work or cybersecurity, be curious and adversarial,...Remote job
- Mercor is seeking a remote red-team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate high-quality data for safer AI systems... ...and collaborate across projects for robust safety testing. This role focuses on reproducible artifacts...Remote job
$26 per hour
A tech consulting firm is seeking a remote contractor to engage in AI red teaming activities. The ideal candidate will work on evaluating conversational... ...Spanish is required, along with prior experience in adversarial AI testing or cybersecurity. This is a flexible hourly...Remote jobHourly payFor contractors10 hours per weekFlexible hours$26 per hour
...Compensation: $26/hour Location: Remote Commitment: 10-40 hours/week Role Responsibilities Red-team conversational AI systems using jailbreaks,... ..., and multi-turn adversarial strategies. Generate high-quality... ...in alignment with defined safety guidelines. Requirements Native...Remote jobHourly payContract work- Location Remote Fluent Language Skills Required... ...believe the safest AI is the one that’s... ...We are assembling a red team for this project -... ...probe AI models with adversarial inputs, surface... ...Specialties Adversarial ML: jailbreak datasets... ...trust the safety of their AI because...Remote work
$20 - $22 per hour
...talent with leading AI research labs. Headquartered... .... Position: AI Safety Experts — English &... ...Location: Remote Role Responsibilities Red team conversational AI models... ...experience in AI adversarial work ,... ...Experience with Adversarial ML , Cybersecurity ,...Remote workContract workSummer work$60 - $90 per hour
...technical talent with leading AI research labs. Headquartered... ...Dorsey . Position: LLM Red Team Specialist — Failure Modes & Edge Cases... ...$60–$90/hour Location: Remote Commitment: 35 hours/week... ...frontier AI models on coding, ML , and analysis tasks to...Remote workContract workSummer work$48 - $62 per hour
...strengthen conversational AI systems by probing them with adversarial inputs, identifying... ...human-generated safety data. This text-... ...and attack cases that teams can use to improve AI... ...Prior experience in AI red teaming, adversarial... .... Work Terms Remote, hourly engagement....Remote workHourly pay$20 - $22 per hour
...Overview Help strengthen conversational AI systems by probing them as an adversary. This remote, text-based role focuses on... ...vulnerabilities, generating high-quality red-team data, and creating reproducible findings that improve AI safety and reliability. Key...Remote workHourly pay- ...AI Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document... ...clause it violated so the safety team can patch the gap. Why... ...research, or adversarial ML work for AI Red Team Specialist...Remote jobHourly payFor contractors10 hours per week
$29 - $45 per hour
...Attack and probe conversational AI systems in English and... ...vulnerabilities, produce reproducible adversarial data, and help make deployed... .... Key Responsibilities Red team conversational AI models and... ...thinking. Work Terms Remote, text-based engagement....Remote workHourly pay$25 per hour
...talent with leading AI research labs. Headquartered... .... Position: AI Safety Experts — English &... ...Location: Remote Role Responsibilities Red team conversational AI models... ...experience in AI adversarial work ,... ...Experience in Adversarial ML : jailbreak datasets...Remote jobContract workSummer work$24 - $35 per hour
...talent with leading AI research labs. Headquartered... .... Position: AI Safety Experts — English &... ...Location: Remote Role Responsibilities Red team conversational AI models... ...experience in AI adversarial work ,... ...Experience in Adversarial ML , Cybersecurity ,...Remote workContract workSummer work- ...progressive technology company is seeking a remote AI Red Teamer to assess conversational AI... ...and identify vulnerabilities through red teaming exercises. Candidates should be fluent... ...with a background in cybersecurity or adversarial testing. This role involves producing actionable...Remote workHourly pay
- ...is seeking a research-focused professional to join a GenAI red-team within a leading AI lab network. The role involves probing frontier models,... ...findings for reproducibility. This full-time W-2 position is remote within the United States, requiring about 35 hours per...Remote workFull time
- ...Capability Elicitation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack... ...clause it violated so the safety team can patch the gap. Why... ...security research, or adversarial ML work for Capability...Remote jobHourly payFor contractors10 hours per week
- Mercor is building a fully remote red team to probe AI models for safety, resilience, and ethical considerations. You will work on adversarial inputs, jailbreaks, and vulnerability discovery while producing actionable data and reproducible reports for customers. This role...Remote job
- Mercor is seeking a remote Red Team data expert to probe AI models with adversarial inputs, surface vulnerabilities, and generate actionable red team data for safer AI systems. You will annotate failures, classify vulnerabilities, and document risks using established taxonomies...Remote job
- Mercor is building a remote red-team for AI safety, focusing on adversarial testing of conversational models. You’ll craft attacks, annotate failures, and surface vulnerabilities while following established taxonomies and playbooks. Documentation of results and reproducible...Remote job
- Mercor is seeking a remote Senior Red Team AI Specialist to test conversational agents and AI models against adversarial inputs. You will annotate vulnerabilities, surface systemic risks, and produce reproducible attack cases to help customers strengthen their AI systems...Remote work
- Mercor is assembling a remote red team to test and strengthen AI systems through adversarial inputs. You will annotate failures, classify vulnerabilities, and surface systemic risks using rigorous taxonomies and playbooks. The role rewards clear communication of risks to...Remote job
- Mercor is building a remote red team to probe AI models with adversarial inputs. We focus on testing for jailbreaks, bias and safety vulnerabilities in conversational systems. You will generate actionable data and reports to help customers harden their AI. Ideal candidates...Remote job
- Mercor is building a red team for adversarial AI testing, reviewing outputs on sensitive topics with optional involvement in high-sensitivity projects, guided by clear guidelines and wellness resources. You will red team conversational AI models, generate high-quality human...Remote job
- Mercor is building a red team to probe AI models with adversarial inputs, surface vulnerabilities, and generate red team data that strengthens safety for customers. This is a text-based, higher-sensitivity project where participation is optional and guided by clear guidelines...Remote job
- Mercor is assembling a remote red team tasked with probing AI models for vulnerabilities, bias and safety issues. You will generate high-quality data, annotate failures,... ...and Vietnamese fluency, curiosity, and an adversarial mindset. The project emphasizes safe, guided...Remote job
- Mercor is building a remote AI red-team program to probe and stress test conversational models for bias, misuse, and safety concerns. The role involves designing adversarial tests, annotating data, and producing reproducible attack datasets to help customers strengthen...Remote job
- Mercor is building a red team for AI safety. This remote role requires English and Dutch fluency and a proactive, adversarial mindset. You will probe conversational AI, annotate vulnerabilities, and generate high-quality attack data for safer deployments. You’ll work across...Remote job
$22 per hour
...talent with leading AI research labs. Headquartered... .... Position: AI Safety Experts — English &... ...20-$22/hourLocation:Remote Role Responsibilities Red team conversational AI... ...teaming experience in AI adversarial work, cybersecurity,... ...in Adversarial ML, Cybersecurity, or socio...Remote jobSummer work- ...AI Red Team Engineer We're looking for an AI Red Team Engineer to break LLM-powered systems... ...conversations. You'll own hands-on adversarial testing end to end: find the failure, prove... ...write it up. White Circle is an AI Safety company building the safety,...Remote workLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Remote Adversarial ML Specialist AI Safety Red Team. Be the first to apply!
- cost specialist San Francisco, CA
- strategic sourcing specialist San Francisco, CA
- authorization specialist San Francisco, CA
- treasury specialist San Francisco, CA
- workforce management specialist San Francisco, CA
- wellness specialist San Francisco, CA
- helpdesk specialist San Francisco, CA
- program eligibility specialist San Francisco, CA
- vulnerability management specialist San Francisco, CA
- invoice specialist San Francisco, CA


