Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Secure Code Review AI Evaluator [Remote]

AuraOne Human Data

Remote
  • Remote job

Secure Code Review AI Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.

Why this role matters

Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.

Responsibilities

  • Design adversarial prompts that probe known weakness classes (jailbreak, policy bypass, prompt injection) for Secure Code Review AI Evaluator assignments.
  • Document every successful attack with reproduction steps and the policy clause it violated.
  • Score model defenses across single-turn and multi-turn conversations.
  • Triage emerging attack vectors and route them to the safety team with severity ratings.
  • Maintain a personal library of attack patterns and propose new red-team rubrics.
  • Calibrate against the broader red-team cohort to keep coverage and severity consistent.

Qualifications

  • Demonstrated experience red-teaming AI systems, security research, or adversarial ML work for Secure Code Review AI Evaluator work.
  • Strong written communication — your reports become the patch ticket.
  • Comfort working in policy-grey areas with clear documentation of what was attempted and why.
  • Familiarity with prompt-injection, jailbreak, and policy-bypass taxonomies.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Construct a 5-turn adversarial conversation that bypasses a specific policy clause and write up the patch ticket.
  • Score a model's defenses against a known jailbreak pattern across 20 variants.
  • Propose a new red-team rubric category after spotting an emerging attack vector.
  • Reproduce a failure another reviewer reported and confirm the severity tag.

Nice to have

  • Background in offensive security, AppSec, or trust & safety operations.
  • Experience publishing or reproducing public adversarial-ML research.
  • Multilingual fluency for cross-language attack testing.

Skills

  • Adversarial prompting
  • Red-team analysis
  • Policy taxonomy
  • Failure documentation
  • Secure Code Review AI evaluation
  • Cybersecurity
  • AI evaluation
  • Rubric writing
  • Expert review

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 11 days ago
Similar jobs that could be interesting for youBased on the Secure Code Review AI Evaluator [Remote] in Remote vacancy
  •  ...Senior Software Engineer — AI Coding Evaluator is a remote engineering review track for evaluating production code, debugging traces, and developer-facing...  ...failure modes (compile error, runtime crash, off-by-one, security issue) with severity scores. Document recurring... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  • $50 - $60 per hour

     ...and building careers in AI training and data labeling...  ...intelligence. Contributors review examples, test model behavior, evaluate responses, and identify...  ...working with authentication or security integrations is valuable....  .... People with analytics, coding, language, or other... 
    Suggested
    Hourly pay
    Part time
    For contractors
    Remote work
    Worldwide
    Flexible hours

    OpenTrain AI

    Brooklyn, NY
    5 days ago
  • $100 per hour

     ...This Role Actually Is You will assess how AI coding agents behave in real-world scenarios —...  ...correctness. What You’ll Be Doing Evaluate AI-generated coding interactions end-to-...  ...needing to fully execute or deeply review every line Comfortable giving direct... 
    Suggested
    Contract work
    Immediate start

    G2i

    Remote
    a month ago
  •  ...Cloud Security AI Evaluator is a remote evaluation track for reviewing cloud security ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  • Dorado is seeking expert Evaluators in program management / implementation planning to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote,... 
    Suggested
    Remote job
    Hourly pay
    Work at office

    Dorado

    New York, NY
    1 day ago
  • $100 per hour

     ...experienced software engineer (SR+) to help evaluate the quality of interactions with modern coding agents such as OpenAI Codex and...  ...Is You will assess how AI coding agents behave in real-world...  ...needing to fully execute or deeply review every line Comfortable giving... 
    Contract work
    Immediate start

    G2i

    Remote
    a month ago
  •  ...Corporate and Securities Law AI Evaluator is a remote review track for evaluating AI outputs in legal review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    2 days ago
  •  ...We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks...  ...lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation... 
    Contract work

    Mindrift

    Remote
    a month ago
  •  ...About The Opportunity We are seeking detail-oriented human reviewers with a strong understanding of their local cultural context to support a range of AI training and evaluation projects . In this role, you will work across diverse task types, including evaluating... 
    Extra income
    Full time
    For contractors
    Local area
    10 hours per week

    Lilt

    Remote
    a month ago
  • Obsidian is looking for expert Evaluators in Media, journalism, and communications to review and assess AI-generated work products, ensuring accuracy and rigor. This role is remote, offering flexibility, requiring 5+ years of relevant experience and proficiency in Microsoft... 
    Remote job
    Work at office

    Obsidian

    Pennsauken, NJ
    3 days ago
  • $80 - $120 per hour

     ...healthcare operations expertise to assess AI-generated documents, spreadsheets, and...  ...hourly role focuses on producing reliable evaluations and actionable feedback across a range of...  ...work products. Key Responsibilities Review AI-generated healthcare operations materials... 
    Hourly pay
    Work at office
    Remote work

    SaidGig

    United States
    2 days ago
  •  ...remote, hourly contractor role supporting AI data and language projects on a project...  ...support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy,...  ...AI products, with deep expertise in Arabic-native and secure, sovereign solutions.... 
    Hourly pay
    For contractors
    Remote work
    Flexible hours

    CNTXT AI

    Brooklyn, NY
    9 days ago
  •  ...About The Opportunity We are seeking detail-oriented human reviewers with a strong understanding of their local cultural context to support a range of AI training and evaluation projects . In this role, you will work across diverse task types, including evaluating... 
    Extra income
    Full time
    For contractors
    Freelance
    Local area
    10 hours per week

    Lilt

    Remote
    29 days ago
  • $14.5 per hour

     ...AI Web Search Evaluator Unlock the Power of the Internet! Are you curious, tech-savvy, and passionate about improving online search experiences...  ...your home! As a Web Search Evaluator, you will: Review and assess internet search results, ensuring users receive... 
    Bi-weekly pay
    Hourly pay
    Part time
    Immediate start
    Remote work
    Work from home
    Flexible hours

    Welo Data

    United States
    2 days ago
  •  ...BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance... 
    Remote work

    BAM VENTURES LLC

    New York, NY
    2 days ago
  •  ...infrastructure engineering backgrounds to test and evaluate AI-assisted workflows across modern...  ...is required, but hands-on use of AI coding or productivity agents is highly relevant...  ...and platform engineering scenarios. Review workflows involving GitHub, GitLab, JIRA... 
    Temporary work

    Gramian Consulting

    Remote
    21 days ago
  • Obsidian is hiring expert Evaluators in Healthcare operations for a remote, hourly engagement. You will review AI-generated work products for accuracy, rigor, and domain quality, leveraging your extensive subject-matter expertise in the field. The ideal candidate has over... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    San Francisco, CA
    1 day ago
  • Mercor is seeking experts in Spreadsheet QA and workbook maintenance to review AI-generated documents, spreadsheets, and slide decks for accuracy and quality. This is a remote, hourly engagement. Ideal candidates have 5+ years in Spreadsheet QA, fluent English, and strong... 
    Remote job
    Hourly pay
    Work at office

    Mercor

    Miami, FL
    5 days ago
  • Mercor is seeking expert Evaluators in Privacy/regulatory compliance to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote, hourly engagement... 
    Remote job
    Hourly pay
    Work at office

    Mercor

    New York, NY
    2 days ago
  • Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates will leverage their subject matter expertise to assess various documents, spreadsheets, and slide decks.... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    3 days ago
  • MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers... 
    Remote job
    Contract work
    Temporary work
    Immediate start

    MERIT Beauty

    New York, NY
    4 days ago
  • Mercor is seeking expert Evaluators in Clinical / biomedical / pharma to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. This is a remote, hourly engagement. You will apply deep subject-matter... 
    Remote job
    Hourly pay

    Mercor

    New York, NY
    1 day ago
  • Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional... 
    Remote job
    Work at office

    Obsidian

    New York, NY
    3 days ago
  • $40 - $100 per hour

    About OpenTrain OpenTrain AI is the hiring and contracting organization for this opportunity...  ...work About AI Training and Scientific Evaluation AI training is the human side of building...  ...artificial intelligence. Specialists review model responses, test reasoning, identify... 
    Hourly pay
    Contract work
    Part time
    For contractors
    Remote work

    OpenTrain AI

    Brooklyn, NY
    4 days ago
  • $14.5 per hour

    Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are high-quality...  ...can vary week to week. Some weeks there is more data to review, other weeks less. Start Date ASAP Project Duration 12... 
    Hourly pay
    Part time
    Immediate start
    Remote work
    Work from home
    10 hours per week
    Flexible hours

    Welo Data

    New York, NY
    3 days ago
  • $24 per hour

    Prolific is seeking an AI Trainer with advanced Tamil fluency to evaluate AI models' understanding of the Tamil language's emotional and cultural nuances. Responsibilities include assessing audio clips, reviewing tones for cultural relevancy, and ensuring quality control... 
    Remote job
    Flexible hours

    Prolific

    New York, NY
    3 days ago
  • Obsidian is seeking expert Evaluators in Humanities / arts / culture to review AI-generated work products for accuracy, rigor, and domain quality. You will apply subject-matter expertise to grade outputs and provide structured feedback. This is a remote, hourly engagement... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    2 days ago
  •  ...Medical Coding Safety Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    2 days ago
  • Mercor is hiring expert Evaluators in Healthcare operations to review AI-generated outputs (documents, spreadsheets, and slide decks) for accuracy and domain quality. This is a remote, hourly engagement. You will apply deep subject-matter expertise to grade outputs, identify... 
    Remote job
    Hourly pay

    Mercor

    New York, NY
    2 days ago
  • $20 - $80 per hour

     ...Role Overview Help improve next-generation AI systems by supplying precise, real-world evaluation, annotation, and feedback. This remote contractor role focuses...  ...to improve model accuracy and reliability. Review content for relevance, coherence, and factual accuracy... 
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    United States
    a month ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Secure Code Review AI Evaluator [Remote]. Be the first to apply!