AI Safety Expert - Red Teaming - AI Trainer
$48 - $62 per hourRemote Jobs
About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey. Position: AI Safety Experts — English & Danish Type:Contract Compensation:$48–$62/hour Location:Remote Role Responsibilities Red team conversational AI models and agents to identify jailbreaks, prompt injections, misuse cases, and bias exploitation. Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks. Apply structure by following taxonomies, benchmarks, and playbooks to maintain consistent testing. Document reproducibly by producing reports, datasets, and attack cases that customers can act on. Work independently and asynchronously to meet deadlines while improving AI model performance. Qualifications Must-Have Fluent in English and Danish. Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing. Strong communication skills to explain risks clearly to both technical and non-technical stakeholders. Preferred Experience in Adversarial ML, Cybersecurity, or socio-technical risk analysis. Skills in creative probing such as psychology, acting, or writing for unconventional adversarial thinking. Resources & Support For details about the interview process and platform information, please check: For any help or support, reach out to:
- hiringmercor
- J-18808-Ljbffr Remote Jobs
$65 - $75 per unit
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Dorsey . Position: Nuclear Engineering & Safeguards Experts for Red Team Type: Contract Compensation: $65–$75 per task Location...SuggestedContract workSummer workRemote workFlexible hours- Mercor is assembling a panel of chemistry and chemical safety experts to red-team frontier AI models and assess misuse potential of technical requests. You will write challenging prompts across benign, dual-use, and adversarial levels, evaluate model responses, and craft...Suggested
- AI Safety Experts — English & Danish is a remote contractor role that tests AI systems using adversarial... ...violated policy clause so the safety team can patch the gap. Adversarial... ...attackers to reproduce, fix, and regress-test failures. #J-18808-Ljbffr AI Trainer JobsSuggestedRemote jobFor contractors
$48 - $62 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ..., and Jack Dorsey . Position: AI Safety Experts — English & Dutch Type: Contract... ...: Remote Role Responsibilities Red team conversational AI models and agents to...SuggestedContract workSummer workRemote work- .... You’ll vibe-code with frontier models, critique the AI against production standards, and red-team for unsafe code and erroneous reasoning before real users... ...that directly shapes next-generation coding agents and safety benchmarks. #J-18808-Ljbffr DataAnnotationSuggestedRemote jobFlexible hours
- AuraOne is seeking a Bilingual Vietnamese STEM Expert (PhD) for an AI Safety remote red-team role. You will design adversarial prompts to test AI systems,... ...emerging attack vectors, ensuring rigorous patch work before models reach customers. #J-18808-Ljbffr AI Trainer JobsRemote jobFor contractors
- AuraOne is seeking a Biology Expert (PhD) for its remote AI Safety red-team track. You design adversarial prompts to probe known weaknesses, document each... ...is remote, US-eligible, with ~10 hours per week, aligning with program needs. #J-18808-Ljbffr AI Trainer JobsRemote jobContract work10 hours per week
- Mercor is building a remote AI safety red team to probe conversational models and surface vulnerabilities in high-sensitivity topics. You will annotate failures, classify risks, and generate reproducible reports for customers to act on. Before exposure to content, topics...Remote job
- Mercor is assembling a red team for AI safety. You will review model outputs, simulate adversarial inputs, and surface vulnerabilities across prompts and conversations. You will generate high-quality human data, annotate failures, classify risks, and produce reproducible...Remote job
- AuraOne is seeking a bilingual Indonesian STEM expert (PhD) for a remote AI safety red-team role that stress-tests AI systems against adversarial prompts... ...defenses in single-turn and multi-turn conversations, and triaging new attack vectors #J-18808-Ljbffr AI Trainer JobsRemote job
- AuraOne is seeking a Chemical Safety & Toxicology Expert for a remote contractor... ...adversarial prompts to stress-test AI systems, document each... ...maintaining a library of red-team rubrics to patch gaps before... ...with US eligibility. #J-18808-Ljbffr AI Trainer JobsRemote jobFor contractors
- AuraOne is seeking a Bilingual Chinese STEM Expert (PhD) for AI Safety, a remote red-team track that stresses testing AI systems against adversarial prompts... ..., score defenses in single- and multi-turn conversations, and triage new attack #J-18808-Ljbffr AI Trainer JobsFor contractorsRemote work
- AuraOne is seeking a Bilingual Korean Generalist Expert — AI Safety for a remote red-team track that stress-tests AI systems against adversarial prompts... ...reproduce failures, and provide rigorous documentation to support model improvements. #J-18808-Ljbffr AI Trainer JobsRemote work
$65 - $75 per unit
...technical talent with leading AI research labs. Headquartered in... ...-turn prompts in radiological safety, labeled as benign, dual-use,... ...Collaborate with radiological safety experts to assess misuse potential in... ...****@*****.*** PS: Our team reviews applications daily....Contract workSummer workRemote work- AuraOne is seeking a Bilingual German STEM Expert (PhD) for a remote AI Safety role focused on adversarial evaluation. You will design prompts, document... ...testing is valued, with flexible remote scheduling and program-defined task volume. #J-18808-Ljbffr AI Trainer JobsRemote jobFor contractorsFlexible hours
- Freelancer - Security Red Teaming Specialist Join to apply for the Freelancer - Security Red Teaming Specialist role at ActiveFence . About The Position As a Red Team Specialist focused on Generative AI Models, you will play a critical role in enhancing the security and...Freelance
- AI Trainer Jobs is seeking a Bilingual Thai STEM Expert (PhD) for a remote AI safety red-team role. Reviewers craft attack scenarios, document failures, and map breaches to rubric clauses to guide patching. This contractor position emphasizes adversarial evaluation, policy...For contractorsRemote work
- Mercor seeks specialists to assemble a panel of energetic materials and propulsion experts for red-teaming frontier AI models. You will craft prompts across three levels and judge model outputs against policy standards, then draft the reference answers with clear technical...
- AI Safety Experts — English & Marathi is a remote contractor track for stress-testing AI systems... ...rubric clause it violated so the safety team can patch the gap. Adversarial... ...provide patches for the modeling team to implement and #J-18808-Ljbffr AI Trainer JobsRemote jobFor contractors
$126.82k - $149.2k
...discover what you excel at—all from Day One.Job DescriptionThe AI Red Team Lead Engineer leads the execution and evolution of offensive... ...applicable to an agreement, such as those related to ethics, safety, or operational procedures.Applicants must be able to comply with...Full timeLocal area3 days per week$70 - $84 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ..., and Jack Dorsey . Position: AI Safety Red Teamer Type: Contract Compensation... ...contribute to safety benchmarking and red-teaming reports. Collaborate with AI researchers...Contract workSummer workRemote work- AI Trainer Jobs is seeking a Capability Elicitation Red Team Specialist for a remote, US-eligible contractor role focused on stress-testing AI systems against adversarial prompts. You will design investigations, document failures with rubric mappings, and help patch gaps...Remote jobFor contractors
$40 - $60 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Position: Healthcare Administration Expert Type: Contract Compensation... ...reach out to: ****@*****.*** PS: Our team reviews applications daily. Please complete...Hourly payWeekly payFull timeContract workFor contractorsSummer workRemote work$65 - $70 per hour
...Biology Expert (PhD) — AI Safety is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team...For contractorsRemote work10 hours per week- A leading technology staffing company is seeking an Offensive Security Specialist to conduct red team operations, specializing in AI systems assessments. You will engage in adversarial simulations, develop defense strategies, and present findings to senior leadership. A...Remote job
- A progressive technology company is seeking a remote AI Red Teamer to assess conversational AI models and identify vulnerabilities through red teaming exercises. Candidates should be fluent in both English and Brazilian Portuguese, with a background in cybersecurity or...Remote jobHourly pay
$70 - $100 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Computational Particle & Nuclear Physics Expert Type: Contract Compensation: $7... ...feedback. Collaborate asynchronously with teams to improve AI model performance . Qualifications...Contract workSummer workRemote work$65 - $75 per hour
...technical talent with leading AI research labs. Headquartered in... ...Hazardous Device Specialist / Public Safety Professional Type:... ...writing, published research, or expert witness experience. Application... ...****@*****.*** PS: Our team reviews applications daily. Please...Contract workSummer workRemote work- ...Exists At Mercor, we believe the safest AI is the one that’s already been attacked — by us. We are assembling a red team for this project - human data experts who probe AI models with adversarial... ...- Mercor customers trust the safety of their AI because you’ve already probed...Part timeRemote work
$20 - $22 per hour
AI Safety Experts — English & Marathi is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document... ...Failure documentation AI Safety Experts English Marathi evaluation #J-18808-Ljbffr AI Trainer JobsFor contractorsRemote work10 hours per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Safety Expert - Red Teaming - AI Trainer. Be the first to apply!
- sql expert New York, NY
- guest service support expert New York, NY
- technology expert New York, NY
- subject matter expert New York, NY
- fulfillment expert New York, NY
- ai trainer New York, NY
- senior safety specialist New York, NY
- warehouse safety New York, NY
- patient safety assistant New York, NY
- safety physician New York, NY




