Macroeconomics AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Macroeconomics AI Evaluator is a remote evaluation track for reviewing macroeconomics ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn macroeconomics ai evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate macroeconomics ai evaluation model outputs against a versioned rubric and assign severity tags for Macroeconomics AI Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on macroeconomics ai evaluation or adjacent content for Macroeconomics AI Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two macroeconomics ai evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Macroeconomics AI evaluation
- Quantitative finance and economics
- AI evaluation
- Rubric writing
- Expert review
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...SuggestedRemote workFlexible hours- ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers...SuggestedContract workTemporary workImmediate startRemote work
$14.5 per hour
A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical...SuggestedHourly payPart timeRemote workFlexible hours- ...Turing is seeking graduate students or professionals for a remote role in evaluating AI-generated research reports. Responsibilities include reading, annotating, and scoring reports on a 1-5 scale, alongside providing written justifications. Candidates must possess strong...SuggestedRemote work
$11.5 per hour
...position as an Online Task Contributor. In this role, you will evaluate and provide feedback on content to enhance search engine results... ...at $11.50 hourly, based on task completion, with a supportive community of contributors involved in AI advancements. #J-18808-LjbffrSuggestedHourly payPart timeRemote work- ...BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance...Remote work
- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Work at officeRemote work
$14.5 per hour
...ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years... ...performance and provide insights on relevance and quality. Evaluate and rate the effectiveness of search engine results to ensure...Hourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours$30 per hour
...Location: Remote Commitment: 10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous agent systems... ...stakeholders. Requirements Strong experience in LLM evaluation, AI output analysis, QA/testing, UX research, or similar analytical...Remote jobHourly payContract work- Productive Playhouse seeks a Portugal Portuguese AI Product Evaluator for an upcoming project testing and evaluating leading AI chatbots. You will interact with AI models to assess capability, safety, and helpfulness, informing model development. Open to freelancers outside...Remote jobFor contractorsFreelanceFlexible hours
- Obsidian is hiring expert Evaluators based in New York, United States for a remote role focusing on reviewing AI-generated work products like documents, spreadsheets, and slide decks. The ideal candidate will have over 5 years of experience in BI dashboards and performance...Hourly payWork at officeRemote work
- Alignerr is seeking a Population Health Informaticist to help train and evaluate AI models using population-level health data. This remote hourly contract role emphasizes independent work, rigorous data review, and clear written feedback to improve AI accuracy and relevance...Remote jobHourly payContract work
- Alignerr is seeking a Population Health Informaticist to help train and evaluate health AI on large-scale datasets and public health strategy. You will review AI-generated health insights, assess data-driven metrics, and provide structured feedback that reflects disparities...Remote jobHourly payContract workFlexible hours
$20 - $30 per hour
A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong...Remote jobHourly pay- Get notified about new Github jobs in United States . Staff Software Engineer, Copilot Code Review Principal Software Engineer - GitHub Platform & Enterprise Principal Software Engineer - GitHub Platform & Enterprise Software Engineer, Research Developer Productivity Staff...Remote jobFor contractors
- Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates will leverage their subject matter expertise to assess various documents, spreadsheets, and slide decks....Remote jobHourly payWork at office
- YO AI Labs seeks an experienced Developer & Infrastructure Expert to evaluate AI-powered workflows across software development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands, configurations, and workflows while applying...Remote job
- YO AI Labs is seeking a Turkish Bilingual Expert to support a language and AI training project. This contractor role is remote, with flexible hours to evaluate Turkish audio for nativeness, fluency, pronunciation, and intonation. Provide clear feedback in English, justify...Remote jobFor contractorsFlexible hours
- Dorado is seeking an AI Language Quality Evaluator fluent in Greek and English for an ongoing, task-based project. This remote freelance role involves reviewing translated and AI-flagged content to judge accuracy, classify issues, and suggest corrected translations. You...Remote jobFor contractorsFreelanceFlexible hours
- A virtual AI evaluation firm is seeking individuals to review and evaluate AI-generated responses in therapeutic conversations. The ideal candidate will possess strong written communication and analytical skills, as well as a keen attention to detail for assessing tone...Remote jobImmediate start
- Crossing Hurdles is hiring a remote AI evaluation contractor to perform Vietnamese voice assessments and rate AI responses for quality and accuracy. You will participate in short voice conversations with AI models, follow scenarios, and provide objective feedback to help...Remote jobFor contractorsFlexible hours
- A leading AI research firm is looking for a contractor in internal or emergency medicine to help enhance language models by providing expert feedback based on real-world medical scenarios. Candidates should possess an MD and have relevant experience, including board certification...Remote jobFor contractors
- Turing is seeking detail-oriented AI Analysts based in the United States for a Google Wallet evaluation project. This role allows you to engage with advanced AI tools while contributing to the future of AI. You will evaluate model responses, review output quality, and provide...Remote jobFull timeContract work
- Obsidian is hiring expert Evaluators in Investment analysis / valuation / credit to review AI-generated work products for accuracy and quality. This remote, hourly position requires deep subject-matter expertise and professional fluency in English to provide structured...Hourly payWork at officeRemote work
- Productive Playhouse is building a talent pool of Armenian speakers to test and evaluate leading AI chatbots. This freelance, project-based opportunity lets you choose tasks, set your own hours, and work with other clients as needed. Open to freelancers outside the U.S....Remote jobFreelanceFlexible hours
$40 - $100 per hour
About OpenTrain OpenTrain AI is the hiring and contracting organization for this opportunity. OpenTrain is the #1 platform for finding... ...English-language work About AI Training and Scientific Evaluation AI training is the human side of building modern artificial intelligence...Hourly payContract workPart timeFor contractorsRemote work- About OpenTrain OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and... ...States English-language work 20+ hours per week About AI Response Evaluation AI training is the human side of building artificial...Part timeFor contractorsRemote work
- Productive Playhouse is building an Albanian-speaking evaluator pool for AI chatbot testing. As an AI Evaluator, you’ll interact with AI models to assess capabilities, safety, and usefulness, providing data to improve models. Tasks come in batches with flexible hours and...Remote jobFreelanceFlexible hours
- Cross Border Talents is seeking a US- or Canada-based Generalist Expert to help train and evaluate frontier AI models. This remote, hourly contract role involves reviewing documents, presentations, spreadsheets, and other professional material for quality, clarity, accuracy...Remote jobHourly payContract work
- A leading AI research accelerator is hiring a position focused on contributing to projects that evaluate and enhance AI systems. You will design community service scenarios, write structured explanations, and evaluate AI accuracy. The ideal candidate will have 4+ years...Remote jobFull timeFor contractors
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Macroeconomics AI Evaluator [Remote]. Be the first to apply!

