Dermatology MD - AI Evaluation Specialist
Obsidian
About the Role Mercor is hiring board-certified Dermatologists for non-clinical work developing and evaluating clinical AI systems. You will apply your diagnostic expertise to image interpretation, annotation, and evaluation work that helps define what a clinically sound dermatological assessment looks like. This is a non-clinical role — no direct patient care, and no responsibility for live diagnosis. This is a shared expert pool. After onboarding you may be matched to any of several concurrent workstreams based on your subspecialty, availability, and interest. You are not committing to a single project, and you may move between streams as priorities shift. What you may work on Work varies by workstream and may include: Case interpretation with written reasoning — working through clinical cases that pair dermatological images with patient history, recording your diagnostic reasoning step by step rather than only the final impression. Structured image annotation — describing lesions against a defined schema: morphology, colour, configuration, distribution, and body site. Differential and severity assessment — recording a ranked differential with your reasoning, and grading severity or progression from images. Output review — judging AI-generated dermatological assessments against clinical standards, flagging hallucinated findings, missed findings, and unsupported certainty. Grading criteria and guideline authoring — defining what separates an excellent assessment from a merely acceptable one, so judgment can be applied consistently at scale. Difficult-case writing — constructing cases that probe the limits of current model reasoning, including presentations that vary across skin tones. Task length varies by stream, from roughly 10 minutes for structured annotation up to an hour for full case authoring. You will get a specific throughput target for whichever stream you are matched to. Required qualifications MD or DO with completed dermatology residency Board certification — ABD, or the equivalent certifying body in your jurisdiction — or board-eligible with residency complete Active, unrestricted medical license in your country of practice 3+ years post-residency clinical dermatology experience, practising or previously practising Strong ability to identify and describe dermatological conditions in precise medical terminology, across the full range of skin tones Written and spoken English fluency Minimum 15 hours per week , with the ability to concentrate hours when a stream is time-boxed Preferred qualifications Fellowship or focused practice in dermatopathology, pediatric dermatology, surgical dermatology, or complex medical dermatology U.S. licensure and board certification Prior medical image annotation, AI evaluation, or dataset curation experience Teaching, rubric design, or resident assessment experience Published research or sustained technical writing (please link a sample) Why this work Dermatology is one of the most image-dependent specialties in medicine, and one where model performance still varies sharply across skin tones and rarer presentations. The reference standard these systems get graded against is written by dermatologists — here you would be writing it. #J-18808-Ljbffr Obsidian
- ...and technical talent with leading AI research labs. Headquartered in... .... Position: Dermatologist (MD) — Clinical Image Interpretation & AI Evaluation Type: Contract Compensation... ...Responsibilities Evaluate and annotate dermatological images using precise medical...SuggestedContract workSummer workRemote work
- ...certified Dermatologists for non-clinical work developing and evaluating clinical AI systems, applying diagnostic expertise to image... ...evaluation to define what constitutes a clinically sound dermatological assessment. This is a shared expert pool with onboarding that...SuggestedFlexible hoursShift work
$121k - $206k
...available in our New York City office or with our AI Engineering squad in Harbor Point, Baltimore, MD. In either location, you will join a collaborative... ..., including architecture, implementation, testing, evaluation, deployment, observability, and production support....SuggestedFull timeWork at officeLocal areaRemote work3 days per week- ...A leading AI organization in Australia is seeking individuals with strong writing and analytical skills to evaluate and improve AI outputs. The ideal candidate must possess the ability to assess emotional nuances and detail while adhering to structured guidelines. Responsibilities...SuggestedImmediate start
$97k - $165k
AI Software Engineer- T. Rowe Price Labs (NY or MD) Apply ( locations New York, NY Baltimore, MD time type Full time posted on Posted 3 Days Ago job... ...engineering, tool use, retrieval-based systems, evaluation frameworks, observability, and guardrails. Champion...SuggestedFull timeWork at officeLocal areaRemote work3 days per week- Welo Data is hiring Data Labeling Associates in New York City for Project Perseus. The role involves evaluating Arabic language AI outputs and ensuring AI safety, requiring professional proficiency in Arabic and experience in writing and AI safety. You’ll critique models...
- ....zone is seeking a Remote Data Labeling Specialist to work from anywhere within the United... ...will label and review multi-modal data for AI training, including text, images, audio,... ...segmentation, content safety labeling, and RLHF-style evaluation tasks. #J-18808-Ljbffr REXRemote job
$11 - $30.65 per hour
Meridial is seeking contractors to evaluate advanced agentic audio models by simulating realistic customer service interactions across multiple domains. You will contribute to developing diverse datasets and assess model performance using various metrics. The role requires...Remote jobHourly payFor contractors- YO AI Labs seeks a Creative & Design Specialist to evaluate AI-generated outputs and provide high-quality human feedback. You will work on graphic design, UI/UX, motion graphics, and multimedia tasks, applying professional standards and contributes to training AI models...Remote jobWorldwideFlexible hours
- Obsidian is collaborating with AI labs to find experienced health insurance professionals to enhance AI systems related to coverage... ...assess AI performance on health insurance tasks. The role includes evaluating AI outputs, creating health insurance scenarios, and providing...
- About the Opportunity A leading AI research organization is seeking advanced LLM power users with strong experience using MCP and (... ...connectors for real-world personal life tasks. This project focuses on evaluating how well AI systems handle personalized, multi-step life tasks...Trial period
- ...Job TitleAI Evaluation EngineerLocationHybrid / RemoteEmployment TypeFull-timeJob SummaryWe are seeking an AI Evaluation Engineer to design, implement, and maintain evaluation frameworks for AI and machine learning systems, with a focus on Large Language Models (LLMs)...
$80 per hour
...human intelligence to ethically shape the future of AI. What We Do The Mindrift platform connects specialists with AI projects from major tech innovators. Our... ...for someone who can design realistic and structured evaluation scenarios for LLM-based agents. You'll create test...Part timeFreelanceRemote workFlexible hours- YO AI Labs is seeking an experienced Management Consultant to support AI training and evaluation projects. This contractor role emphasizes structured problem-solving, business analysis, and consulting expertise to assess and improve AI model performance. Remote work enables...Remote jobFor contractors
$30 - $50 per hour
A tech company is seeking an AI Researcher to support end-to-end research for modern AI systems. This remote role involves designing experiments, defining evaluation protocols, and improving evaluation rigor for large language models. Key responsibilities include developing...Remote jobHourly pay$400 per month
Obsidian is seeking contributors for a Frontier Code Agents project, focused on evaluating AI coding models in fraud and risk engineering. Candidates will use AI coding tools to handle complex tasks and provide technical assessments. The role requires 2+ years of experience...$20 - $160 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Position: Generalist Annotator — Health AI Conversation Quality Evaluation Type: Contract Compensation: $20–$160/hour Location...Contract workSummer workRemote work$50 - $190 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Remote Commitment: 20+ hours/week Role Responsibilities Evaluate AI systems on complex personal workflows, including personal...Hourly payContract workFor contractorsSummer workRemote workTrial period$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... .... Position: User/customer research and feedback synthesis Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...and Jack Dorsey . Position: Biology / environmental science Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Data analysis / quantitative readouts Evaluator Type: Contract Compensation: $80–$120/hour...Contract workSummer workWork at officeRemote work- ...has about 60 attorneys, 75 staff and over 50 staff attorneys. About the Team and the Role The Junior AI Specialist will support the firm’s efforts to evaluate, test, and responsibly adopt generative AI tools in support of legal work. Based in our New York office,...Full timeContract workWork at officeWork from homeWorldwide
$197.3k - $225.1k
AI Engineer 4 (Gen AI Platform Services: Agentic AI, Agent Guardrails, Agent Evaluation, Agent Memory) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning...Full timePart timeLocal area- Data & AI Solutions Consultant Databricks SFO -CA -/ Remote Duration - 8-12 months Core Skills Needed 12+ years of full-stack engineering... ...Generation (RAG), AI agents, prompt engineering, model evaluation, and enterprise AI application patterns. Proficiency with AI-assisted...Remote work
- ...time role, requiring strong analytical skills and independent work. The position involves creating historically relevant prompts, evaluating AI outputs, and contributing to research initiatives. Ideal candidates will hold a PhD in History or a related discipline and...Part timeRemote workFlexible hours
- ...Remote Job Summary We are seeking experienced Senior Software Engineers to support an AI training project by creating reinforcement learning environments that evaluate AI models on complex software engineering tasks using Model Context Protocol (MCP) tools....Remote jobFor contractors
- Meridial is seeking a Danish Voice Acting Specialist to support AI training by providing recorded speech samples and expert feedback. You will... ...recording setup. You will work collaboratively with a team to evaluate AI output and improve voice design. The position offers a...Remote jobHourly pay
- Feedinkoo is seeking a Copywriting & Content Subject Matter Expert to provide expert evaluation of AI-generated content and create high-quality written material. This remote contractor position requires strong copywriting skills and the ability to give precise feedback...Remote jobContract workFor contractors
- ...individual to support governance, oversight, and responsible adoption of AI technologies across the organization, collaborating with... ..., compliance, product, and FinOps teams. The role focuses on evaluating AI solutions, maintaining governance standards, monitoring usage...
$50 per hour
A leading AI training company is seeking experienced software engineers to help train generative AI models remotely. The role involves crafting questions related to computer science and evaluating AI-generated code. This freelance opportunity offers flexible hours and...Remote jobHourly payFreelanceFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Dermatology MD - AI Evaluation Specialist. Be the first to apply!
- ai scientist New York, NY
- ai data scientist New York, NY
- dermatology medical assistant New York, NY
- dermatology New York, NY
- dermatology manager New York, NY
- dermatology associates New York, NY
- schweiger dermatology New York, NY
- lpn dermatology New York, NY
- research fellow dermatology New York, NY
- physician assistant dermatology New York, NY



