Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Remote Medical Evaluation Specialist for AI Benchmarks

YO AI Labs

Phoenix, NY
  • Remote job

YO AI Labs seeks Medical Evaluation Specialists, including medical students, residents, physicians, or biomedical professionals, to contribute clinical expertise to evaluating next-generation AI systems. You will create and validate difficult medical questions and answers to test clinical reasoning, synthesize evidence from primary literature and guidelines, and document rationale with citations in a remote contractor role. #J-18808-Ljbffr YO AI Labs

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Remote Medical Evaluation Specialist for AI Benchmarks in Phoenix, NY vacancy
  • YO AI Labs is seeking Medical Evaluation Specialists to contribute clinical expertise for evaluating and improving next-generation AI systems. You will create...  ...skills, and clinical reasoning are essential for success in a remote contractor role. #J-18808-Ljbffr YO AI Labs
    Remote job
    For contractors

    YO AI Labs

    San Jose, CA
    2 days ago
  • YO AI Labs is seeking Medical Evaluation Specialists, including medical students, residents, physicians, and biomedical professionals, to contribute clinical expertise to a project evaluating next-generation AI systems. You will create and validate high-difficulty medical... 
    Remote job

    YO AI Labs

    Boston, MA
    2 days ago
  • YO AI Labs seeks Medical Evaluation Specialists (medical students, residents, physicians, and biomedical professionals) to create and validate high-difficulty medical questions and answers for testing AI systems. You will craft items that challenge clinical reasoning and... 
    Remote job

    YO AI Labs

    Chicago, IL
    2 days ago
  • YO AI Labs seeks Medical Evaluation Specialists to create and validate high-difficulty medical question-and-answer pairs for AI systems. You will synthesize...  ..., and document rationales with citations. This remote contractor role emphasizes accuracy, clarity, and defensible... 
    Remote job
    For contractors

    YO AI Labs

    Houston, TX
    2 days ago
  • YO AI Labs seeks Medical Evaluation Specialists, including medical students, residents, physicians, and biomedical professionals, to contribute clinical expertise to evaluating and improving next‑generation AI systems. In this contractor role, you will create and validate... 
    Remote job
    For contractors

    YO AI Labs

    Seattle, WA
    2 days ago
  • YO AI Labs is seeking an experienced Producer (Film/TV/Digital/Live) for a remote contractor role to design and evaluate realistic production benchmarks. You will craft evaluation tasks, call sheets, budgets, and schedules, advancing multi-constraint scenarios while ensuring... 
    Remote job
    For contractors

    YO AI Labs

    Chicago, IL
    4 days ago
  • YO AI Labs is seeking experienced Producers across film, television...  ...live events to support an AI evaluation project. You will create realistic...  ...evaluate outputs, and develop benchmarks. No prior AI experience is required, and work is remote. Key responsibilities include... 
    Remote job

    YO AI Labs

    San Francisco, CA
    4 days ago
  •  ...States Digital Space LLC offers a remote contract role focusing on fine-...  ...down complex problems for evaluation tasks. You will collaborate with LLM researchers on benchmarks spanning undergraduate to PhD topics, enabling cutting-edge AI projects while working independently... 
    Remote job
    Contract work

    United States Digital Space LLC

    New York, NY
    2 days ago
  • $80 - $100 per hour

     ...Senior Python Developer (AI Evaluation & Benchmarking) An enterprise client is seeking experienced Senior Python Developers to help build the...  ...software platforms. Additional Information Fully remote contract opportunity. Compensation ranges from $80–$100... 
    Remote work
    Hourly pay
    Weekly pay
    Contract work
    10 hours per week

    Lifted, an Upwork Company™

    United States
    4 days ago
  •  ...Join a pioneering AI initiative focused on building the next generation of evaluation benchmarks for frontier AI models. We are seeking experienced QA and Test Engineers...  ...ensure benchmark integrity. This is a fully remote, full-time engagement requiring approximately... 
    Remote work
    Full time
    Contract work
    For contractors
    Flexible hours

    Weekday

    Remote
    26 days ago
  •  ...Role Overview Evaluate and benchmark the coding abilities of frontier AI models by reviewing AI-generated solutions, validating...  ...type: Contractor assignment, no medical or paid leave provided....  ...is next week. Work location: Remote, United States only. Perks: fully... 
    Remote work
    Contract work
    For contractors

    SaidGig

    United States
    28 days ago
  • $201.3k - $352.3k

     ...meaningful work. Today, ServiceNow is the AI control tower for business reinvention...  ...Engineering Manager, Agentic & GenAI Benchmarking and Evaluations to establish and lead AI evaluation...  ...flexibility and trust. Work personas (flexible, remote, or required in office) are categories... 
    Remote work
    Work experience placement
    Work at office
    Immediate start
    Flexible hours
    Shift work

    ServiceNow

    Santa Clara, CA
    3 days ago
  • Mercor seeks expert medical and health science professionals...  ...content for an AI research initiative. You...  ...and medicine domains, evaluate solution quality, and help...  ...gold-standard benchmarks used to advance AI capabilities...  ...capabilities. This is a fully remote, asynchronous... 
    Remote job
    10 hours per week

    Mercor

    Nashville, TN
    5 days ago
  •  ...We are seeking expert medical and health science professionals...  ...content for an AI research initiative. You...  ...and medicine domains, evaluate solution quality, and help...  ...gold-standard benchmarks used to advance AI capabilities...  ...Asynchronous, fully remote work #J-18808-Ljbffr Mercor
    Remote work

    Mercor

    Nashville, TN
    5 days ago
  • YO AI Labs seeks Medical Evaluation Specialists, including medical students, residents, physicians, and biomedical professionals, to contribute clinical expertise to evaluating and improving AI systems. You will create and validate high-difficulty medical questions and... 
    Remote job

    YO AI Labs

    Wisconsin
    2 days ago
  • YO AI Labs seeks Medical Evaluation Specialists to craft and validate challenging medical questions for AI evaluation. Remote collaboration with physicians, students, and biomedical professionals to test AI systems on complex clinical reasoning and evidence interpretation... 
    Remote job
    Contract work

    YO AI Labs

    New York, NY
    3 days ago
  • 24-MAG LLC is offering a part-time remote consulting opportunity for board-certified physicians with deep expertise in a defined therapeutic...  ...trials and drug development. Selected clinicians will create evaluation rubrics and interpret endpoints to assess impact on prescribing... 
    Remote job
    Part time

    24-Mag Llc

    New York, NY
    3 days ago
  •  ...Job Description Job Title: Medical Evaluation Specialist Role Type: Contractor Location: Remote Job Overview We are seeking...  ...and improving next-generation AI systems. In this role, you...  ...that help establish rigorous benchmarks for medical AI evaluation.... 
    Remote job
    For contractors

    YO AI Labs

    Boston, MA
    6 days ago
  • $25 - $35 per hour

     ...expertise with solid knowledge of medical terminology, regulations, and...  ...information into appropriate evaluation and management, diagnostic,...  ...Workplace Type*This is a fully remote position. *Application Deadline...  ...of Artificial Intelligence (AI):* We may use Artificial Intelligence... 
    Remote work
    Contract work
    Temporary work

    TEKsystems

    Saint Paul, MN
    4 days ago
  • YO AI Labs seeks Medical Evaluation Specialists to contribute clinical expertise to evaluating next-generation AI medical systems. You will create and validate challenging medical questions and answers designed to test advanced clinical reasoning and interpretation of... 
    Remote job

    YO AI Labs

    Raleigh, NC
    2 days ago
  • Medical Specialist (Fluent in Arabic) - Freelance AI Trainer Project World Wide - Remote Are you a medical professional fluent in Arabic and eager to shape the future of AI?...  ...suggest improvements to prompt engineering and evaluation metrics. Challenge advanced language... 
    Remote work
    Hourly pay
    Contract work
    For contractors
    Freelance

    Meridial

    New York, NY
    2 days ago
  • Video & Audio Annotation and AI Prompt Evaluation - English RWS Group is looking for Data Specialists to help train a broad range of...  ...looking for freelance, part-time, remote, work-from-home jobs where you...  ...contribute to building a new benchmark dataset that will evaluate a... 
    Remote work
    Part time
    Freelance
    Immediate start
    Work from home
    Flexible hours

    RWS Group

    San Antonio, TX
    4 days ago
  • $146.2k - $261.4k

     ...Description RAND's Center on AI, Security, and...  ...will build systems to evaluate how AI models perform...  ...may include developing benchmarks for fully autonomous operations...  ...Program (CNODP), Remote Interactive Operator...  ...Lead at either the specialist or expert level of experience... 
    Remote work
    Fixed term contract
    Work experience placement
    Work from home

    RAND

    Santa Monica, CA
    5 days ago
  •  ...Prolific is seeking Product Designers and UX Specialists to join our Expert Network, contributing to the training and evaluation of cutting-edge AI models. This role requires expertise in usability, design systems, and user research, providing essential feedback on design... 
    Remote work
    Work from home
    Flexible hours

    Prolific

    Tucson, AZ
    3 days ago
  • $60 - $75 per hour

     ...Applied Biology Benchmark Specialist - AI Evaluation is a remote review track for evaluating AI outputs across biology reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method... 
    Remote job
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    17 days ago
  •  ...Join a pioneering AI initiative focused on building the next generation of evaluation benchmarks for frontier AI models. We are seeking experienced QA and Test Engineers...  ...ensure benchmark integrity. This is a fully remote, full-time engagement requiring approximately... 
    Remote work
    Full time
    Contract work
    For contractors
    Flexible hours

    Weekday

    Remote
    26 days ago
  •  ...Prolific is seeking talented Product Designers and UX Specialists to join our Expert Network for training and evaluation of cutting-edge AI models. Candidates should have a relevant educational background and at least one year of experience in design-related fields. Responsibilities... 
    Remote work
    Work from home
    Flexible hours

    Prolific

    Arizona City, AZ
    4 days ago
  • $50 per hour

     ...Role Overview Evaluate, benchmark, and help improve the coding capabilities of advanced AI models by assessing AI-generated solutions...  ...type, contractor assignment, no medical or paid leave provided....  ...next week. Work location, remote, candidates must be located in... 
    Remote work
    Contract work
    For contractors

    SaidGig

    United States
    1 day ago
  •  ...-MAG LLC is seeking an experienced QA/test engineer to design robust benchmark test cases, review complex tasks, and debug Python environments for frontier AI evaluation work. The role is fully remote within the United States and requires strong attention to detail and... 
    Remote job

    24-MAG LLC

    New York, NY
    4 days ago
  • $30 per hour

    Prolific is seeking Advanced Dutch Speakers in Chicago, IL to train AI models. You will complete AI tasks and assess AI performance in...  ...competitive rates and direct payment through PayPal. Pass the evaluation, and you can start within 15 minutes. #J-18808-Ljbffr Prolific
    Remote job
    Work from home

    Prolific

    Chicago, IL
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Remote Medical Evaluation Specialist for AI Benchmarks. Be the first to apply!