Secure Code Review AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Secure Code Review AI Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Why this role matters
Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.
Responsibilities
- Design adversarial prompts that probe known weakness classes (jailbreak, policy bypass, prompt injection) for Secure Code Review AI Evaluator assignments.
- Document every successful attack with reproduction steps and the policy clause it violated.
- Score model defenses across single-turn and multi-turn conversations.
- Triage emerging attack vectors and route them to the safety team with severity ratings.
- Maintain a personal library of attack patterns and propose new red-team rubrics.
- Calibrate against the broader red-team cohort to keep coverage and severity consistent.
Qualifications
- Demonstrated experience red-teaming AI systems, security research, or adversarial ML work for Secure Code Review AI Evaluator work.
- Strong written communication — your reports become the patch ticket.
- Comfort working in policy-grey areas with clear documentation of what was attempted and why.
- Familiarity with prompt-injection, jailbreak, and policy-bypass taxonomies.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Construct a 5-turn adversarial conversation that bypasses a specific policy clause and write up the patch ticket.
- Score a model's defenses against a known jailbreak pattern across 20 variants.
- Propose a new red-team rubric category after spotting an emerging attack vector.
- Reproduce a failure another reviewer reported and confirm the severity tag.
Nice to have
- Background in offensive security, AppSec, or trust & safety operations.
- Experience publishing or reproducing public adversarial-ML research.
- Multilingual fluency for cross-language attack testing.
Skills
- Adversarial prompting
- Red-team analysis
- Policy taxonomy
- Failure documentation
- Secure Code Review AI evaluation
- Cybersecurity
- AI evaluation
- Rubric writing
- Expert review
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- ...Senior Software Engineer — AI Coding Evaluator is a remote engineering review track for evaluating production code, debugging traces, and developer-facing... ...failure modes (compile error, runtime crash, off-by-one, security issue) with severity scores. Document recurring...SuggestedRemote jobHourly payFor contractors10 hours per week
$50 - $60 per hour
...and building careers in AI training and data labeling... ...intelligence. Contributors review examples, test model behavior, evaluate responses, and identify... ...working with authentication or security integrations is valuable.... .... People with analytics, coding, language, or other...SuggestedHourly payPart timeFor contractorsRemote workWorldwideFlexible hours$100 per hour
...This Role Actually Is You will assess how AI coding agents behave in real-world scenarios —... ...correctness. What You’ll Be Doing Evaluate AI-generated coding interactions end-to-... ...needing to fully execute or deeply review every line Comfortable giving direct...SuggestedContract workImmediate start- ...Cloud Security AI Evaluator is a remote evaluation track for reviewing cloud security ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team...SuggestedRemote jobHourly payFor contractors10 hours per week
- Dorado is seeking expert Evaluators in program management / implementation planning to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote,...SuggestedRemote jobHourly payWork at office
$100 per hour
...experienced software engineer (SR+) to help evaluate the quality of interactions with modern coding agents such as OpenAI Codex and... ...Is You will assess how AI coding agents behave in real-world... ...needing to fully execute or deeply review every line Comfortable giving...Contract workImmediate start- ...Corporate and Securities Law AI Evaluator is a remote review track for evaluating AI outputs in legal review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can...Remote jobHourly payFor contractors10 hours per week
- ...We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks... ...lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation...Contract work
- ...About The Opportunity We are seeking detail-oriented human reviewers with a strong understanding of their local cultural context to support a range of AI training and evaluation projects . In this role, you will work across diverse task types, including evaluating...Extra incomeFull timeFor contractorsLocal area10 hours per week
- Obsidian is looking for expert Evaluators in Media, journalism, and communications to review and assess AI-generated work products, ensuring accuracy and rigor. This role is remote, offering flexibility, requiring 5+ years of relevant experience and proficiency in Microsoft...Remote jobWork at office
$80 - $120 per hour
...healthcare operations expertise to assess AI-generated documents, spreadsheets, and... ...hourly role focuses on producing reliable evaluations and actionable feedback across a range of... ...work products. Key Responsibilities Review AI-generated healthcare operations materials...Hourly payWork at officeRemote work- ...remote, hourly contractor role supporting AI data and language projects on a project... ...support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy,... ...AI products, with deep expertise in Arabic-native and secure, sovereign solutions....Hourly payFor contractorsRemote workFlexible hours
- ...About The Opportunity We are seeking detail-oriented human reviewers with a strong understanding of their local cultural context to support a range of AI training and evaluation projects . In this role, you will work across diverse task types, including evaluating...Extra incomeFull timeFor contractorsFreelanceLocal area10 hours per week
$14.5 per hour
...AI Web Search Evaluator Unlock the Power of the Internet! Are you curious, tech-savvy, and passionate about improving online search experiences... ...your home! As a Web Search Evaluator, you will: Review and assess internet search results, ensuring users receive...Bi-weekly payHourly payPart timeImmediate startRemote workWork from homeFlexible hours- ...BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance...Remote work
- ...infrastructure engineering backgrounds to test and evaluate AI-assisted workflows across modern... ...is required, but hands-on use of AI coding or productivity agents is highly relevant... ...and platform engineering scenarios. Review workflows involving GitHub, GitLab, JIRA...Temporary work
- Obsidian is hiring expert Evaluators in Healthcare operations for a remote, hourly engagement. You will review AI-generated work products for accuracy, rigor, and domain quality, leveraging your extensive subject-matter expertise in the field. The ideal candidate has over...Remote jobHourly payWork at office
- Mercor is seeking experts in Spreadsheet QA and workbook maintenance to review AI-generated documents, spreadsheets, and slide decks for accuracy and quality. This is a remote, hourly engagement. Ideal candidates have 5+ years in Spreadsheet QA, fluent English, and strong...Remote jobHourly payWork at office
- Mercor is seeking expert Evaluators in Privacy/regulatory compliance to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote, hourly engagement...Remote jobHourly payWork at office
- Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates will leverage their subject matter expertise to assess various documents, spreadsheets, and slide decks....Remote jobHourly payWork at office
- MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers...Remote jobContract workTemporary workImmediate start
- Mercor is seeking expert Evaluators in Clinical / biomedical / pharma to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. This is a remote, hourly engagement. You will apply deep subject-matter...Remote jobHourly pay
- Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Remote jobWork at office
$40 - $100 per hour
About OpenTrain OpenTrain AI is the hiring and contracting organization for this opportunity... ...work About AI Training and Scientific Evaluation AI training is the human side of building... ...artificial intelligence. Specialists review model responses, test reasoning, identify...Hourly payContract workPart timeFor contractorsRemote work$14.5 per hour
Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are high-quality... ...can vary week to week. Some weeks there is more data to review, other weeks less. Start Date ASAP Project Duration 12...Hourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours$24 per hour
Prolific is seeking an AI Trainer with advanced Tamil fluency to evaluate AI models' understanding of the Tamil language's emotional and cultural nuances. Responsibilities include assessing audio clips, reviewing tones for cultural relevancy, and ensuring quality control...Remote jobFlexible hours- Obsidian is seeking expert Evaluators in Humanities / arts / culture to review AI-generated work products for accuracy, rigor, and domain quality. You will apply subject-matter expertise to grade outputs and provide structured feedback. This is a remote, hourly engagement...Remote jobHourly payWork at office
- ...Medical Coding Safety Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning...Remote jobHourly payFor contractors10 hours per week
- Mercor is hiring expert Evaluators in Healthcare operations to review AI-generated outputs (documents, spreadsheets, and slide decks) for accuracy and domain quality. This is a remote, hourly engagement. You will apply deep subject-matter expertise to grade outputs, identify...Remote jobHourly pay
$20 - $80 per hour
...Role Overview Help improve next-generation AI systems by supplying precise, real-world evaluation, annotation, and feedback. This remote contractor role focuses... ...to improve model accuracy and reliability. Review content for relevance, coherence, and factual accuracy...Hourly payContract workFor contractorsRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Secure Code Review AI Evaluator [Remote]. Be the first to apply!






