Physician AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Physician AI Evaluator is a remote clinical-review track for evaluating AI outputs that touch physician. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning so the modeling team can close the gap.
Why this role matters
Physician AI is only useful if a credentialed clinician would sign their name to the output. AuraOne uses licensed and credentialed reviewers to set the bar for clinical reasoning, dosing, and guideline adherence in regulated workflows.
Responsibilities
- Review AI outputs against current physician clinical guidelines and standard of care for Physician AI Evaluator assignments.
- Grade differential reasoning, dosing, and red-flag handling on a structured rubric.
- Flag patient-safety issues with severity tags that match clinical-incident taxonomies.
- Capture the corrected clinical reasoning so the modeling team can train on it.
- Adjudicate disputed cases against published evidence or specialty guidelines.
- Maintain reviewer-quality scores in weekly inter-rater calibration cycles.
Qualifications
- Active or recent license / board certification in physician or an adjacent specialty for Physician AI Evaluator work.
- Hands-on clinical experience reading patient cases and writing structured assessments.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that cites guidelines, evidence, or standard of care.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Grade a model's differential-diagnosis reasoning on a physician case and write a 2-3 sentence assessment.
- Flag a dosing miscalculation with the right severity tag and corrected dose.
- Adjudicate a disputed case between two reviewers using current specialty guidelines.
- Audit a 25-case batch for rubric consistency and report drift to the program lead.
Nice to have
- Prior work as a guideline reviewer, peer reviewer, or clinical educator.
- Familiarity with AI-assisted clinical decision support and its failure modes.
- Bilingual clinical experience for non-English patient cases.
Skills
- Clinical reasoning
- Guideline review
- Patient-safety assessment
- Structured clinical writing
- Physician
- Medicine and healthcare
- AI evaluation
- Rubric writing
- Expert review
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$170 - $190 per hour
About OpenTrain OpenTrain AI is the hiring and contracting organization for this role... ...Build a durable portfolio of AI training and evaluation experience. Create an OpenTrain account... ...recruiting a multilingual Primary Care Physician to support clinical documentation and...SuggestedHourly payPart timeFor contractorsRemote work10 hours per weekFlexible hours- ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers...SuggestedContract workTemporary workImmediate startRemote work
$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...SuggestedRemote workFlexible hours$14.5 per hour
A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical...SuggestedHourly payPart timeRemote workFlexible hours- ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and...SuggestedFull timeContract workRemote workFlexible hours
- ...BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance...Remote work
- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Work at officeRemote work
- Turing is seeking detail-oriented AI Analysts based in the United States for a Google Wallet evaluation project. This role allows you to engage with advanced AI tools while contributing to the future of AI. You will evaluate model responses, review output quality, and provide...Remote jobFull timeContract work
- Mercor is seeking expert Evaluators in Humanities / arts / culture to review AI-generated work products for accuracy, rigor, and domain quality. This remote hourly engagement requires deep subject-matter expertise to grade outputs and provide actionable feedback. Candidates...Remote jobHourly payWork at office
$20 - $26 per hour
Prolific is seeking fluent Kannada speakers to act as evaluators. You will assess how naturally and authentically AI captures Kannada speech, by listening to audio clips and comparing text and voice. This fast-paced project pays $20-26 per hour and may require about one...Remote jobHourly payWork from homeFlexible hours- Mercor is seeking experts in Spreadsheet QA and workbook maintenance to review AI-generated documents, spreadsheets, and slide decks for accuracy and quality. This is a remote, hourly engagement. Ideal candidates have 5+ years in Spreadsheet QA, fluent English, and strong...Remote jobHourly payWork at office
$45 - $55 per hour
This is a non-engineering content-policy evaluation role. Applicants must demonstrate relevant depth in violent fiction or media, military... ...1,600 educational institutions. In 2025, we started Handshake AI and built the fastest-growing AI data business in history. We work...Remote jobMonday to FridayShift work- Alignerr is seeking a Population Health Informaticist for AI training. This remote, hourly contract role requires 10-40 hours per week... ...to improve AI understanding of population health data. You will evaluate AI-generated analyses, identify errors, review data pipelines,...Remote jobHourly payContract work10 hours per weekFlexible hours
- YO AI Labs is seeking Italian bilingual experts to contribute to language and AI training projects by evaluating Italian audio content for nativeness, fluency, and overall linguistic quality. You will assess pronunciation, intonation, and authenticity, then document clear...Remote job
$20 - $26 per hour
Prolific seeks fluent Kannada speakers to act as evaluators for AI language data. You will assess text and voice segments, rate naturalness, and help identify cultural nuances. A PayPal account is required to receive payments for tasks priced at $20-$26 per hour. Responsibilities...Remote jobHourly payWork from homeFlexible hours- Mercor is seeking expert Evaluators in Clinical/biomedical/pharma to review AI-generated documents, spreadsheets, and slide decks for accuracy and domain quality. Remote, hourly engagement, requiring 5+ years of relevant experience and native/professional English fluency...Remote jobHourly payWork at office
$14.5 per hour
Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages...Hourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours- A leading AI research accelerator is hiring a position focused on contributing to projects that evaluate and enhance AI systems. You will design community service scenarios, write structured explanations, and evaluate AI accuracy. The ideal candidate will have 4+ years...Remote jobFull timeFor contractors
$40 - $100 per hour
About OpenTrain OpenTrain AI is the hiring and contracting organization for this opportunity. OpenTrain is the #1 platform for finding... ...English-language work About AI Training and Scientific Evaluation AI training is the human side of building modern artificial intelligence...Hourly payContract workPart timeFor contractorsRemote work- Unknown is seeking a Polish Bilingual Expert (contractor) to evaluate Polish audio content for AI training. The role emphasizes nativeness, fluency, and linguistic quality, with feedback delivered in English. Strong Polish language skills and clear written and verbal English...Remote jobFor contractors
- Productive Playhouse is building a talent pool of Armenian speakers to test and evaluate leading AI chatbots. This freelance, project-based opportunity lets you choose tasks, set your own hours, and work with other clients as needed. Open to freelancers outside the U.S....Remote jobFreelanceFlexible hours
- Dorado is hiring expert Evaluators in Legal/compliance to review AI-generated work products for accuracy and domain quality. You will apply deep subject-matter expertise to grade outputs across documents, spreadsheets, and slide decks. This is a remote, hourly engagement...Remote jobHourly payWork at office
- About OpenTrain OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and... ...States English-language work 20+ hours per week About AI Response Evaluation AI training is the human side of building artificial...Part timeFor contractorsRemote work
$20 per hour
...detail-oriented professionals based in the United States to support AI model improvement through high-quality data review, annotation,... ..., you will work on structured AI data tasks used to train and evaluate the Gemini model. CONTRACT: Short-term freelance/contractor...Hourly payContract workTemporary workFor contractorsFreelanceRemote work- Alignerr is seeking a Population Health Informaticist to help train and evaluate health AI on large-scale datasets and public health strategy. You will review AI-generated health insights, assess data-driven metrics, and provide structured feedback that reflects disparities...Remote jobHourly payContract workFlexible hours
- YO AI Labs is seeking an Investment & Finance Expert to work remotely as a contractor. The role focuses on evaluating AI outputs related to valuation, financial modeling, markets, and investments, and crafting expert prompts plus reference answers to improve model reasoning...Remote jobFor contractors
- Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates will leverage their subject matter expertise to assess various documents, spreadsheets, and slide decks....Remote jobHourly payWork at office
$20 - $30 per hour
A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong...Remote jobHourly pay- A virtual AI evaluation firm is seeking individuals to review and evaluate AI-generated responses in therapeutic conversations. The ideal candidate will possess strong written communication and analytical skills, as well as a keen attention to detail for assessing tone...Remote jobImmediate start
- Dorado is seeking an AI Language Quality Evaluator fluent in Greek and English for an ongoing, task-based project. This remote freelance role involves reviewing translated and AI-flagged content to judge accuracy, classify issues, and suggest corrected translations. You...Remote jobFor contractorsFreelanceFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Physician AI Evaluator [Remote]. Be the first to apply!

