Remote Medical Evaluation Specialist for AI Benchmarks
YO AI Labs
Phoenix, NY
- Remote job
YO AI Labs seeks Medical Evaluation Specialists, including medical students, residents, physicians, or biomedical professionals, to contribute clinical expertise to evaluating next-generation AI systems. You will create and validate difficult medical questions and answers to test clinical reasoning, synthesize evidence from primary literature and guidelines, and document rationale with citations in a remote contractor role. #J-18808-Ljbffr YO AI Labs
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Remote Medical Evaluation Specialist for AI Benchmarks in Phoenix, NY vacancy
- YO AI Labs is seeking Medical Evaluation Specialists to contribute clinical expertise for evaluating and improving next-generation AI systems. You will create... ...skills, and clinical reasoning are essential for success in a remote contractor role. #J-18808-Ljbffr YO AI LabsRemote jobFor contractors
- YO AI Labs is seeking Medical Evaluation Specialists, including medical students, residents, physicians, and biomedical professionals, to contribute clinical expertise to a project evaluating next-generation AI systems. You will create and validate high-difficulty medical...Remote job
- YO AI Labs seeks Medical Evaluation Specialists (medical students, residents, physicians, and biomedical professionals) to create and validate high-difficulty medical questions and answers for testing AI systems. You will craft items that challenge clinical reasoning and...Remote job
- YO AI Labs seeks Medical Evaluation Specialists to create and validate high-difficulty medical question-and-answer pairs for AI systems. You will synthesize... ..., and document rationales with citations. This remote contractor role emphasizes accuracy, clarity, and defensible...Remote jobFor contractors
- YO AI Labs seeks Medical Evaluation Specialists, including medical students, residents, physicians, and biomedical professionals, to contribute clinical expertise to evaluating and improving next‑generation AI systems. In this contractor role, you will create and validate...Remote jobFor contractors
- YO AI Labs is seeking an experienced Producer (Film/TV/Digital/Live) for a remote contractor role to design and evaluate realistic production benchmarks. You will craft evaluation tasks, call sheets, budgets, and schedules, advancing multi-constraint scenarios while ensuring...Remote jobFor contractors
- YO AI Labs is seeking experienced Producers across film, television... ...live events to support an AI evaluation project. You will create realistic... ...evaluate outputs, and develop benchmarks. No prior AI experience is required, and work is remote. Key responsibilities include...Remote job
- ...States Digital Space LLC offers a remote contract role focusing on fine-... ...down complex problems for evaluation tasks. You will collaborate with LLM researchers on benchmarks spanning undergraduate to PhD topics, enabling cutting-edge AI projects while working independently...Remote jobContract work
$80 - $100 per hour
...Senior Python Developer (AI Evaluation & Benchmarking) An enterprise client is seeking experienced Senior Python Developers to help build the... ...software platforms. Additional Information Fully remote contract opportunity. Compensation ranges from $80–$100...Remote workHourly payWeekly payContract work10 hours per week- ...Join a pioneering AI initiative focused on building the next generation of evaluation benchmarks for frontier AI models. We are seeking experienced QA and Test Engineers... ...ensure benchmark integrity. This is a fully remote, full-time engagement requiring approximately...Remote workFull timeContract workFor contractorsFlexible hours
- ...Role Overview Evaluate and benchmark the coding abilities of frontier AI models by reviewing AI-generated solutions, validating... ...type: Contractor assignment, no medical or paid leave provided.... ...is next week. Work location: Remote, United States only. Perks: fully...Remote workContract workFor contractors
$201.3k - $352.3k
...meaningful work. Today, ServiceNow is the AI control tower for business reinvention... ...Engineering Manager, Agentic & GenAI Benchmarking and Evaluations to establish and lead AI evaluation... ...flexibility and trust. Work personas (flexible, remote, or required in office) are categories...Remote workWork experience placementWork at officeImmediate startFlexible hoursShift work- Mercor seeks expert medical and health science professionals... ...content for an AI research initiative. You... ...and medicine domains, evaluate solution quality, and help... ...gold-standard benchmarks used to advance AI capabilities... ...capabilities. This is a fully remote, asynchronous...Remote job10 hours per week
- ...We are seeking expert medical and health science professionals... ...content for an AI research initiative. You... ...and medicine domains, evaluate solution quality, and help... ...gold-standard benchmarks used to advance AI capabilities... ...Asynchronous, fully remote work #J-18808-Ljbffr MercorRemote work
- YO AI Labs seeks Medical Evaluation Specialists, including medical students, residents, physicians, and biomedical professionals, to contribute clinical expertise to evaluating and improving AI systems. You will create and validate high-difficulty medical questions and...Remote job
- YO AI Labs seeks Medical Evaluation Specialists to craft and validate challenging medical questions for AI evaluation. Remote collaboration with physicians, students, and biomedical professionals to test AI systems on complex clinical reasoning and evidence interpretation...Remote jobContract work
- 24-MAG LLC is offering a part-time remote consulting opportunity for board-certified physicians with deep expertise in a defined therapeutic... ...trials and drug development. Selected clinicians will create evaluation rubrics and interpret endpoints to assess impact on prescribing...Remote jobPart time
- ...Job Description Job Title: Medical Evaluation Specialist Role Type: Contractor Location: Remote Job Overview We are seeking... ...and improving next-generation AI systems. In this role, you... ...that help establish rigorous benchmarks for medical AI evaluation....Remote jobFor contractors
$25 - $35 per hour
...expertise with solid knowledge of medical terminology, regulations, and... ...information into appropriate evaluation and management, diagnostic,... ...Workplace Type*This is a fully remote position. *Application Deadline... ...of Artificial Intelligence (AI):* We may use Artificial Intelligence...Remote workContract workTemporary work- YO AI Labs seeks Medical Evaluation Specialists to contribute clinical expertise to evaluating next-generation AI medical systems. You will create and validate challenging medical questions and answers designed to test advanced clinical reasoning and interpretation of...Remote job
- Medical Specialist (Fluent in Arabic) - Freelance AI Trainer Project World Wide - Remote Are you a medical professional fluent in Arabic and eager to shape the future of AI?... ...suggest improvements to prompt engineering and evaluation metrics. Challenge advanced language...Remote workHourly payContract workFor contractorsFreelance
- Video & Audio Annotation and AI Prompt Evaluation - English RWS Group is looking for Data Specialists to help train a broad range of... ...looking for freelance, part-time, remote, work-from-home jobs where you... ...contribute to building a new benchmark dataset that will evaluate a...Remote workPart timeFreelanceImmediate startWork from homeFlexible hours
$146.2k - $261.4k
...Description RAND's Center on AI, Security, and... ...will build systems to evaluate how AI models perform... ...may include developing benchmarks for fully autonomous operations... ...Program (CNODP), Remote Interactive Operator... ...Lead at either the specialist or expert level of experience...Remote workFixed term contractWork experience placementWork from home- ...Prolific is seeking Product Designers and UX Specialists to join our Expert Network, contributing to the training and evaluation of cutting-edge AI models. This role requires expertise in usability, design systems, and user research, providing essential feedback on design...Remote workWork from homeFlexible hours
$60 - $75 per hour
...Applied Biology Benchmark Specialist - AI Evaluation is a remote review track for evaluating AI outputs across biology reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method...Remote jobFor contractors10 hours per week- ...Join a pioneering AI initiative focused on building the next generation of evaluation benchmarks for frontier AI models. We are seeking experienced QA and Test Engineers... ...ensure benchmark integrity. This is a fully remote, full-time engagement requiring approximately...Remote workFull timeContract workFor contractorsFlexible hours
- ...Prolific is seeking talented Product Designers and UX Specialists to join our Expert Network for training and evaluation of cutting-edge AI models. Candidates should have a relevant educational background and at least one year of experience in design-related fields. Responsibilities...Remote workWork from homeFlexible hours
$50 per hour
...Role Overview Evaluate, benchmark, and help improve the coding capabilities of advanced AI models by assessing AI-generated solutions... ...type, contractor assignment, no medical or paid leave provided.... ...next week. Work location, remote, candidates must be located in...Remote workContract workFor contractors- ...-MAG LLC is seeking an experienced QA/test engineer to design robust benchmark test cases, review complex tasks, and debug Python environments for frontier AI evaluation work. The role is fully remote within the United States and requires strong attention to detail and...Remote job
$30 per hour
Prolific is seeking Advanced Dutch Speakers in Chicago, IL to train AI models. You will complete AI tasks and assess AI performance in... ...competitive rates and direct payment through PayPal. Pass the evaluation, and you can start within 15 minutes. #J-18808-Ljbffr ProlificRemote jobWork from home
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Remote Medical Evaluation Specialist for AI Benchmarks. Be the first to apply!
Related searches
- remote no experience Phoenix, NY
- remote medical coder (no experience in coding) Phoenix, NY
- implementation project manager remote Phoenix, NY
- remote work no experience Phoenix, NY
- remote contract attorney Phoenix, NY
- remote legal writer Phoenix, NY
- remote data entry no experience Phoenix, NY
- fully remote Phoenix, NY
- remote tasks Phoenix, NY
- remote work Phoenix, NY




