Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Physician AI Evaluator [Remote]

AuraOne Human Data

Remote
  • Remote job

Physician AI Evaluator is a remote clinical-review track for evaluating AI outputs that touch physician. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning so the modeling team can close the gap.

Why this role matters

Physician AI is only useful if a credentialed clinician would sign their name to the output. AuraOne uses licensed and credentialed reviewers to set the bar for clinical reasoning, dosing, and guideline adherence in regulated workflows.

Responsibilities

  • Review AI outputs against current physician clinical guidelines and standard of care for Physician AI Evaluator assignments.
  • Grade differential reasoning, dosing, and red-flag handling on a structured rubric.
  • Flag patient-safety issues with severity tags that match clinical-incident taxonomies.
  • Capture the corrected clinical reasoning so the modeling team can train on it.
  • Adjudicate disputed cases against published evidence or specialty guidelines.
  • Maintain reviewer-quality scores in weekly inter-rater calibration cycles.

Qualifications

  • Active or recent license / board certification in physician or an adjacent specialty for Physician AI Evaluator work.
  • Hands-on clinical experience reading patient cases and writing structured assessments.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that cites guidelines, evidence, or standard of care.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Grade a model's differential-diagnosis reasoning on a physician case and write a 2-3 sentence assessment.
  • Flag a dosing miscalculation with the right severity tag and corrected dose.
  • Adjudicate a disputed case between two reviewers using current specialty guidelines.
  • Audit a 25-case batch for rubric consistency and report drift to the program lead.

Nice to have

  • Prior work as a guideline reviewer, peer reviewer, or clinical educator.
  • Familiarity with AI-assisted clinical decision support and its failure modes.
  • Bilingual clinical experience for non-English patient cases.

Skills

  • Clinical reasoning
  • Guideline review
  • Patient-safety assessment
  • Structured clinical writing
  • Physician
  • Medicine and healthcare
  • AI evaluation
  • Rubric writing
  • Expert review

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Physician AI Evaluator [Remote] in Remote vacancy
  • $170 - $190 per hour

    About OpenTrain OpenTrain AI is the hiring and contracting organization for this role...  ...Build a durable portfolio of AI training and evaluation experience. Create an OpenTrain account...  ...recruiting a multilingual Primary Care Physician to support clinical documentation and... 
    Suggested
    Hourly pay
    Part time
    For contractors
    Remote work
    10 hours per week
    Flexible hours

    OpenTrain AI

    Brooklyn, NY
    1 day ago
  •  ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers... 
    Suggested
    Contract work
    Temporary work
    Immediate start
    Remote work

    MERIT Beauty

    New York, NY
    2 days ago
  • $20 per hour

    A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and... 
    Suggested
    Remote work
    Flexible hours

    DataAnnotation

    United States
    12 hours ago
  • $14.5 per hour

    A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical... 
    Suggested
    Hourly pay
    Part time
    Remote work
    Flexible hours

    Welo Data

    United States
    2 days ago
  •  ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and... 
    Suggested
    Full time
    Contract work
    Remote work
    Flexible hours

    Virtual Vocations Inc

    United States
    1 day ago
  •  ...BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance... 
    Remote work

    BAM VENTURES LLC

    New York, NY
    4 days ago
  •  ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional... 
    Work at office
    Remote work

    Obsidian

    New York, NY
    4 days ago
  • Turing is seeking detail-oriented AI Analysts based in the United States for a Google Wallet evaluation project. This role allows you to engage with advanced AI tools while contributing to the future of AI. You will evaluate model responses, review output quality, and provide... 
    Remote job
    Full time
    Contract work

    Turing

    Seattle, WA
    4 days ago
  • Mercor is seeking expert Evaluators in Humanities / arts / culture to review AI-generated work products for accuracy, rigor, and domain quality. This remote hourly engagement requires deep subject-matter expertise to grade outputs and provide actionable feedback. Candidates... 
    Remote job
    Hourly pay
    Work at office

    Mercor Inc

    San Francisco, CA
    12 hours ago
  • $20 - $26 per hour

    Prolific is seeking fluent Kannada speakers to act as evaluators. You will assess how naturally and authentically AI captures Kannada speech, by listening to audio clips and comparing text and voice. This fast-paced project pays $20-26 per hour and may require about one... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    Dallas, TX
    12 hours ago
  • Mercor is seeking experts in Spreadsheet QA and workbook maintenance to review AI-generated documents, spreadsheets, and slide decks for accuracy and quality. This is a remote, hourly engagement. Ideal candidates have 5+ years in Spreadsheet QA, fluent English, and strong... 
    Remote job
    Hourly pay
    Work at office

    Mercor

    Miami, FL
    2 days ago
  • $45 - $55 per hour

    This is a non-engineering content-policy evaluation role. Applicants must demonstrate relevant depth in violent fiction or media, military...  ...1,600 educational institutions. In 2025, we started Handshake AI and built the fastest-growing AI data business in history. We work... 
    Remote job
    Monday to Friday
    Shift work

    Apply

    Brooklyn, NY
    1 day ago
  • Alignerr is seeking a Population Health Informaticist for AI training. This remote, hourly contract role requires 10-40 hours per week...  ...to improve AI understanding of population health data. You will evaluate AI-generated analyses, identify errors, review data pipelines,... 
    Remote job
    Hourly pay
    Contract work
    10 hours per week
    Flexible hours

    Alignerr

    Charlotte, AR
    4 days ago
  • YO AI Labs is seeking Italian bilingual experts to contribute to language and AI training projects by evaluating Italian audio content for nativeness, fluency, and overall linguistic quality. You will assess pronunciation, intonation, and authenticity, then document clear... 
    Remote job

    YO AI Labs

    Dallas, TX
    3 days ago
  • $20 - $26 per hour

    Prolific seeks fluent Kannada speakers to act as evaluators for AI language data. You will assess text and voice segments, rate naturalness, and help identify cultural nuances. A PayPal account is required to receive payments for tasks priced at $20-$26 per hour. Responsibilities... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    Phoenix, AZ
    12 hours ago
  • Mercor is seeking expert Evaluators in Clinical/biomedical/pharma to review AI-generated documents, spreadsheets, and slide decks for accuracy and domain quality. Remote, hourly engagement, requiring 5+ years of relevant experience and native/professional English fluency... 
    Remote job
    Hourly pay
    Work at office

    Mercor

    New York, NY
    2 days ago
  • $14.5 per hour

    Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages... 
    Hourly pay
    Part time
    Immediate start
    Remote work
    Work from home
    10 hours per week
    Flexible hours

    Welo Data

    New York, NY
    5 days ago
  • A leading AI research accelerator is hiring a position focused on contributing to projects that evaluate and enhance AI systems. You will design community service scenarios, write structured explanations, and evaluate AI accuracy. The ideal candidate will have 4+ years... 
    Remote job
    Full time
    For contractors

    Turing

    New York, NY
    3 days ago
  • $40 - $100 per hour

    About OpenTrain OpenTrain AI is the hiring and contracting organization for this opportunity. OpenTrain is the #1 platform for finding...  ...English-language work About AI Training and Scientific Evaluation AI training is the human side of building modern artificial intelligence... 
    Hourly pay
    Contract work
    Part time
    For contractors
    Remote work

    OpenTrain AI

    Brooklyn, NY
    1 day ago
  • Unknown is seeking a Polish Bilingual Expert (contractor) to evaluate Polish audio content for AI training. The role emphasizes nativeness, fluency, and linguistic quality, with feedback delivered in English. Strong Polish language skills and clear written and verbal English... 
    Remote job
    For contractors

    YO AI Labs

    Annapolis, MD
    3 days ago
  • Productive Playhouse is building a talent pool of Armenian speakers to test and evaluate leading AI chatbots. This freelance, project-based opportunity lets you choose tasks, set your own hours, and work with other clients as needed. Open to freelancers outside the U.S.... 
    Remote job
    Freelance
    Flexible hours

    Productive Playhouse

    New York, NY
    3 days ago
  • Dorado is hiring expert Evaluators in Legal/compliance to review AI-generated work products for accuracy and domain quality. You will apply deep subject-matter expertise to grade outputs across documents, spreadsheets, and slide decks. This is a remote, hourly engagement... 
    Remote job
    Hourly pay
    Work at office

    Dorado

    New York, NY
    3 days ago
  • About OpenTrain OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and...  ...States English-language work 20+ hours per week About AI Response Evaluation AI training is the human side of building artificial... 
    Part time
    For contractors
    Remote work

    OpenTrain AI

    Brooklyn, NY
    12 hours ago
  • $20 per hour

     ...detail-oriented professionals based in the United States to support AI model improvement through high-quality data review, annotation,...  ..., you will work on structured AI data tasks used to train and evaluate the Gemini model. CONTRACT: Short-term freelance/contractor... 
    Hourly pay
    Contract work
    Temporary work
    For contractors
    Freelance
    Remote work

    Gramian Consulting Group

    Florida, NY
    1 day ago
  • Alignerr is seeking a Population Health Informaticist to help train and evaluate health AI on large-scale datasets and public health strategy. You will review AI-generated health insights, assess data-driven metrics, and provide structured feedback that reflects disparities... 
    Remote job
    Hourly pay
    Contract work
    Flexible hours

    Alignerr

    Sheffield, TX
    12 hours ago
  • YO AI Labs is seeking an Investment & Finance Expert to work remotely as a contractor. The role focuses on evaluating AI outputs related to valuation, financial modeling, markets, and investments, and crafting expert prompts plus reference answers to improve model reasoning... 
    Remote job
    For contractors

    YO AI Labs

    Austin, TX
    2 days ago
  • Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates will leverage their subject matter expertise to assess various documents, spreadsheets, and slide decks.... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    New York, NY
    12 hours ago
  • $20 - $30 per hour

    A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong... 
    Remote job
    Hourly pay

    Crossing Hurdles

    New York, NY
    3 days ago
  • A virtual AI evaluation firm is seeking individuals to review and evaluate AI-generated responses in therapeutic conversations. The ideal candidate will possess strong written communication and analytical skills, as well as a keen attention to detail for assessing tone... 
    Remote job
    Immediate start

    Crossing Hurdles

    New York, NY
    3 days ago
  • Dorado is seeking an AI Language Quality Evaluator fluent in Greek and English for an ongoing, task-based project. This remote freelance role involves reviewing translated and AI-flagged content to judge accuracy, classify issues, and suggest corrected translations. You... 
    Remote job
    For contractors
    Freelance
    Flexible hours

    Dorado

    Brooklyn, NY
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Physician AI Evaluator [Remote]. Be the first to apply!