Data Annotator - AI Model Evaluation & Labeling [Remote]
AuraOne Human Data
- Remote job
Data Annotator - AI Model Evaluation & Labeling is a remote evaluation track for reviewing data annotator evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn data annotator evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate data annotator evaluation model outputs against a versioned rubric and assign severity tags for Data Annotator - AI Model Evaluation & Labeling assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on data annotator evaluation or adjacent content for Data Annotator - AI Model Evaluation & Labeling work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two data annotator evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Data Annotator evaluation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$20 per hour
...DataAnnotation is committed to creating quality AI. Join our team to help train AI chatbots while... ...chatbots. You will develop complex prompts to test AI models, write high-quality responses to demonstrate excellence, and evaluate different model outputs based on accuracy and...SuggestedHourly payFull timeContract workPart timeFor contractorsSelf employmentFreelanceRemote work$120 per hour
...Role Overview Government Quantitative Professionals evaluate AI model outputs and provide structured, domain-specific feedback to improve... ...government-sector quantitative roles such as Mathematician or Data Scientist. Relevant real-world responsibilities may include...SuggestedHourly payTemporary workPart timeImmediate startRemote workFlexible hours- ...Role Overview Medical professionals evaluate AI-generated medical content and use their clinical and field experience to improve model outputs. No prior AI experience is required. In this role you will assess model responses, create realistic prompts that reflect clinical...SuggestedHourly payPart timeImmediate startRemote workFlexible hours
$238k - $302k
...simulation across 15+ U.S. states. The Large Model Evaluation team is at the nexus of Waymo’s AI ambition . With advancements in Large... ...development and deployment. Build data pipelines for signal discovery, data labeling, feature extraction and metric computation...SuggestedFull timeRemote work$100 per hour
...Aviation professionals apply their operational, safety, and regulatory expertise to evaluate AI-generated content, create realistic prompts, and deliver clear, structured feedback that improves model behavior for aviation tasks and language. Key Responsibilities Develop...SuggestedHourly payTemporary workPart timeRemote workFlexible hours$30 - $90 per hour
...Build and maintain backend services in Go while testing and evaluating alpha-stage AI coding models. This part-time, remote contract role combines hands-on... ...endpoints using a modern Go stack Perform rigorous data validation and implement comprehensive error handling...Hourly payContract workPart timeFor contractorsRemote work$60 per hour
A leading AI development company is seeking quantitative professionals to evaluate AI-generated analyses and develop solutions in various quantitative fields. This fully remote position allows for flexible scheduling and competitive pay up to $60/hour. Candidates should...Remote workFlexible hours$40 per hour
A leading AI training company is seeking a Quantitative Researcher to enhance AI models through rigorous evaluation and problem-solving in mathematics. This remote role involves measuring chatbot progress and assessing performance across various mathematical fields. Ideal...Hourly payRemote workFlexible hours$60 per hour
A leading AI development firm is seeking experienced quantitative professionals to contribute... ...to AI systems. The role involves evaluating AI-generated work and solving technical problems... .... Candidates should have a background in data science, economics, or similar fields,...Hourly payRemote workFlexible hours$85 per hour
...metering, distribution, and substation systems to evaluate AI-generated content and produce expert-level training data. In this contract role you will use your... ...and provide corrections and context that improve model behavior. Key Responsibilities Review AI-generated...Hourly payContract workPart timeFor contractorsRemote workFlexible hours$75 per hour
...Managers apply archival, library, and collections expertise to evaluate and improve AI-generated content related to records, archives, and... ...will create prompts that reflect real workplace tasks, review model outputs for accuracy and relevance, and provide clear, structured...Part timeRemote workFlexible hours$60 - $150 per hour
...Role Overview Provide legal subject-matter expertise to improve and evaluate AI systems, by designing realistic legal tasks, reviewing model outputs, and giving domain-specific feedback that advances frontier AI research. This is an open application to join a Law Expert...Hourly payContract workImmediate startRemote work$85 per hour
...apply hands-on expertise in environmental assessment, GIS analysis, and renewable energy siting to evaluate AI-generated geospatial outputs and develop expert training data that improves AI understanding of environmental workflows and mapping practices. This is a...Hourly payContract workPart timeWork at officeRemote workFlexible hours$80 - $110 per hour
...on the forefront of generative AI by designing and executing... ...and reasoning gaps in advanced models. You will author tasks, produce... ...executable tests where applicable, run evaluations against a target model, and... ..., model evaluation, or data annotation is preferred. Strong...Hourly payPart timeFreelanceRemote work$40 per hour
A data annotation company is seeking professionals in quantitative fields to enhance AI development. This fully remote role allows individuals to set flexible schedules while evaluating AI-generated analyses and solving complex quantitative problems. Candidates should have...Hourly payRemote workFlexible hours$40 per hour
A leading AI development firm in Michigan is seeking experienced quantitative professionals to evaluate AI-generated analyses and contribute to the evolution of AI models. Candidates should have a background in data science, statistics, or similar fields, with at least...Hourly payRemote work$40 per hour
...A forward-thinking AI team is seeking quantitative professionals to evaluate and improve cutting-edge AI systems. The role involves working on AI-generated analyses and providing critical feedback for model enhancement. Candidates should have a strong quantitative background...Hourly payRemote workFlexible hours$60 per hour
...contribute to developing cutting-edge AI systems, while enjoying the... ...advance AI development. AI models are increasingly capable of... ...the-art AI models on tasks like evaluating AI-generated quantitative... ...how these systems reason about data, models, and scientific problems...Hourly payFull timeRemote workFlexible hours$40 per hour
A leading AI development company is seeking experienced quantitative professionals for a remote role. Candidates will evaluate AI-generated quantitative work, solve complex problems, and provide valuable feedback. The ideal candidate has 2+ years of experience in a quantitative...Hourly payFull timeRemote workFlexible hours$40 per hour
...An innovative AI development company is seeking experienced quantitative professionals to contribute to AI advancements. This fully remote role involves evaluating AI-generated analyses and ensuring they are technically accurate and valid in real-world scenarios. Candidates...Hourly payRemote workFlexible hours$40 per hour
A data science team is seeking experienced quantitative professionals to evaluate AI-generated work and contribute to the development of cutting-edge AI systems. This fully remote position offers flexible scheduling and competitive hourly pay starting at $40+. Ideal candidates...Hourly payRemote workFlexible hours- ...Physicians apply clinical judgment and frontline medical experience to evaluate AI-generated medical content, ensuring clinical accuracy, sound... ...planning. Assess clarity, relevance, and safety of model outputs in realistic care scenarios. Provide detailed, constructive...Full timeFor contractorsPrivate practiceRemote workFlexible hours
$40 per hour
...A leading AI development firm is seeking experienced quantitative professionals to evaluate AI-generated analysis and solve complex technical problems. Ideal candidates will have 2+ years in quantitative roles, knowledge of statistical methods, and experience with analytical...Hourly payRemote workFlexible hours$60 per hour
A leading AI development firm is seeking quantitative professionals to join their remote team. The role involves evaluating AI-generated analyses, designing quantitative problems, and providing... ...in quantitative fields like data science or statistics, with proficiency...Remote jobHourly payFlexible hours$60 per hour
A leading data science company is seeking quantitative professionals to evaluate AI-generated analyses and design quantitative problems crucial for advancing AI systems. This fully remote role allows you to work flexibly and choose your projects, with competitive hourly...Remote jobHourly pay$60 per hour
A tech company focused on AI is seeking quantitative professionals to evaluate AI-generated analytical work and help advance AI development. Responsibilities include statistical analysis, predictive modeling, and providing feedback to shape AI systems. The role offers flexible...Remote jobHourly payFlexible hours- A leading AI development firm is looking for experienced quantitative professionals to... ...their team remotely. In this role, you will evaluate AI-generated quantitative analysis and... ...that shapes how these systems reason about data. Candidates should have 2+ years of relevant...Remote jobFlexible hours
$60 per hour
A leading AI development firm seeks experienced quantitative professionals to evaluate and advance AI systems. This fully remote role allows... ...you are driven by rigorous data analysis and are comfortable... ...statistical and predictive modeling, this position is for you. #J...Remote jobFlexible hours$60 per hour
A leading AI development company is seeking quantitative professionals... ...enhance AI systems. You will evaluate AI-generated analyses, solve... ...provide critical feedback for model improvement. This fully remote... ...have strong experiences in data science or related fields and...Remote jobHourly pay$60 per hour
A cutting-edge AI development company is seeking quantitative professionals to evaluate AI-generated work and provide essential feedback. This role allows for remote work... ...2+ years of relevant experience in fields like data science or statistics, alongside coding skills,...Remote jobHourly payFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Annotator - AI Model Evaluation & Labeling [Remote]. Be the first to apply!



