Cybersecurity AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Cybersecurity AI Evaluator is a remote evaluation track for reviewing cybersecurity ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn cybersecurity ai evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate cybersecurity ai evaluation model outputs against a versioned rubric and assign severity tags for Cybersecurity AI Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on cybersecurity ai evaluation or adjacent content for Cybersecurity AI Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two cybersecurity ai evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Cybersecurity AI evaluation
- Cybersecurity
- AI evaluation
- Rubric writing
- Expert review
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$80 - $120 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ..., Larry Summers , and Jack Dorsey . Position: Cybersecurity / IT GRC Evaluator Type: Contract Compensation: $80–$120/hour Location...SuggestedContract workSummer workWork at officeRemote work$80 - $120 per hour
...Cybersecurity / IT GRC Evaluator is a remote evaluation track for reviewing cybersecurity / it grc evaluation prompts and responses against AuraOne'... ...modeling team can use to retrain. Why this role matters AI data reviewers help turn cybersecurity / it grc evaluation...SuggestedRemote jobFor contractors10 hours per week$14.5 per hour
A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical...SuggestedHourly payPart timeRemote workFlexible hours- ...Weekday 1 is seeking expert Evaluators to review AI-generated real estate, hospitality and events outputs for accuracy, rigor and domain quality. You will apply deep expertise to grade documents, spreadsheets and slide decks. Requirements include 5+ years in Real estate...SuggestedHourly payWeekly payContract workWork at officeRemote workWeekday work
- ...About the role Swahili Evaluation AI Evaluator is a remote evaluation track for reviewing swahili generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback...SuggestedHourly payFor contractorsRemote work10 hours per week
$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...Remote workFlexible hours- ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers...Contract workTemporary workImmediate startRemote work
- ...Obsidian is hiring expert Evaluators in Real estate, hospitality, and events to review AI-generated work for accuracy, rigor, and domain quality. This remote position requires deep expertise and involves grading outputs like documents and presentations. Applicants must...Work at officeRemote work
$14.5 per hour
...AI Web Search Evaluator As a Web Search Evaluator, you will play a key role in improving the quality of search engine results, ensuring users find the most relevant and useful information. Your work will directly impact the development of AI algorithms, making search...Hourly payPart timeCurrently hiringImmediate startRemote workWork from home10 hours per weekFlexible hours- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Work at officeRemote work
- ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and...Full timeContract workRemote workFlexible hours
$11.5 per hour
...position as an Online Task Contributor. In this role, you will evaluate and provide feedback on content to enhance search engine results... ...11.50 hourly, based on task completion, with a supportive community of contributors involved in AI advancements. #J-18808-Ljbffr...Hourly payPart timeRemote work- ...Alignerr is seeking Creative Writing Evaluators to assess AI-generated stories, essays, poetry, and other writings to shape how AI learns compelling prose. This fully remote, flexible contract roles welcomes avid readers and writers with no publishing credits required...Contract workRemote workFlexible hours
$14.5 per hour
...AI Web Search Evaluator Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years of experience in...Bi-weekly payHourly payPart timeImmediate startRemote workWork from homeFlexible hours- ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise...Hourly payWork at officeRemote work
$150 per hour
...The Work You will help improve AI systems by creating expert-level quantitative finance training data and reviewing AI-generated answers. Your feedback will focus on whether responses are accurate, relevant, and consistent with accepted quantitative finance methods...Hourly payOngoing contractContract workPart timeRemote workFlexible hours- YO AI Labs is seeking Content Writers & Editors to contribute to a customer project focused on text quality and AI evaluation. You will apply your writing and editing expertise to help train and evaluate next-generation AI systems through high-quality input. No prior AI...Remote job
- AuraOne is seeking a Chemicals Safety Risk Evaluator in a remote, US-eligible red-team track to stress-test AI systems against adversarial prompts. Reviewers craft attack scenarios, document failures, and tie each successful jailbreak to the violated policy clause so the...Remote jobFor contractors
$30 per hour
...Location: Remote Commitment: 10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous agent systems... ...stakeholders. Requirements Strong experience in LLM evaluation, AI output analysis, QA/testing, UX research, or similar analytical...Remote jobHourly payContract work- YO AI Labs is seeking an experienced AI Evaluation Specialist to remotely assess AI assistant outputs for an enterprise training project. The contractor will evaluate responses against detailed rubrics, identify issues, and provide actionable feedback to improve AI systems...Remote jobFor contractors
$20 - $30 per hour
A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong...Remote jobHourly pay- YO AI Labs is seeking an AI Evaluation Specialist to remotely assess enterprise AI outputs against detailed rubrics, measuring accuracy, relevance and quality. You will provide concise, actionable feedback to drive improvements. Responsibilities include identifying reasoning...Remote jobFor contractors
- Prolific is seeking fluent Norwegian speakers to act as evaluators for AI training, performing side-by-side assessments of text and voice snippets to judge naturalness and authenticity. You will listen to audio clips and rate how naturally the AI speaks, providing detailed...Remote jobWork from homeFlexible hours
- YO AI Labs is seeking an AI Evaluation Specialist contractor to assess AI outputs against established rubrics, identify gaps, and provide actionable feedback to improve enterprise AI training systems. The role requires strong analytical writing skills, familiarity with...Remote jobFor contractors
- AI Trainer Jobs seeks a remote independent contractor to review AI outputs for fintech operations evaluation. You will assess workflow adherence, tone, and escalation logic, assigning severity tags and documenting next steps for model training. The role requires experience...Remote jobHourly payFor contractorsFlexible hours
- Archangel Health AI is seeking Clinical AI Evaluators to review AI-generated clinical outputs, benchmark diagnostic reasoning, and refine responses to real-world medical queries. You will perform clinical accuracy auditing, RLHF ranking, error and harm identification,...Remote workFlexible hoursShift work
- Mercor is seeking expert Evaluators in Clinical / biomedical / pharma to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. This is a remote, hourly engagement. You will apply deep subject-matter...Remote jobHourly pay
$24 per hour
Prolific is seeking an AI Trainer with advanced Tamil fluency to evaluate AI models' understanding of the Tamil language's emotional and cultural nuances. Responsibilities include assessing audio clips, reviewing tones for cultural relevancy, and ensuring quality control...Remote jobFlexible hours$50 per hour
Prolific seeks Product Designers and UX Specialists to help train AI models using your expertise. You'll evaluate AI-generated designs and ensure usability and accessibility while working from home. Ideal candidates hold a BS, MS, or PhD in a relevant field and have at...Remote jobWork from homeFlexible hours$14.5 per hour
...ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years... ...performance and provide insights on relevance and quality. Evaluate and rate the effectiveness of search engine results to ensure...Hourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Cybersecurity AI Evaluator [Remote]. Be the first to apply!


