Applied Data Scientist, Health AI Evaluation & Datasets
Innodata
Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale. We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers.
Scope of the Role:
Healthcare is one of the highest-stakes domains for generative AI. Clinical accuracy, patient safety, regulatory compliance, health equity, auditability, and workflow fit are the bar for shipping anything real. Innodata partners with foundation model labs, medical AI startups, payers, providers, pharma, and digital health companies building LLMs, multimodal systems, and AI agents for healthcare and life sciences.
As an Applied Data Scientist, Health AI Evaluation & Datasets , you own the design, measurement quality, and clinical validity of datasets used to train, fine-tune, and evaluate health-domain models. You bring clinical or biomedical fluency and data science rigor: you can read a clinical guideline, payer policy, medical literature artifact, or patient communication workflow; translate it into a measurable dataset and evaluation plan; and defend the methodology to sophisticated clinical, data science, and ML stakeholders.
You will work in a tight pod with a Technical Solutions Architect, Applied Research Scientist, AI/ML Research Engineer, and Language Data Scientists. Your role is to make sure the data, rubrics, review workflows, and measurement evidence are clinically realistic, statistically defensible, compliant, and useful for evaluation and post-training.
What You’ll Own:
- Translate customer goals — such as improving differential diagnosis, evaluating a clinical note summarizer, testing a RAG-based medical literature assistant, or creating preference data for patient-facing chatbots — into dataset specifications, taxonomies, rubrics, sampling plans, and acceptance criteria.
- Make multimodal health AI a core focus: design training and evaluation datasets across clinical text, medical images, waveforms, structured EHR data, claims, trial data, medical literature, patient communications, payer policies, drug information, and other clinical artifacts, as well as use cases such as clinical reasoning, medical QA, note summarization, medical coding, patient communication, utilization management, and literature synthesis.
- Design evaluations for retrieval-augmented and source-grounded health AI systems, including evidence citation, faithfulness, contraindication handling, guideline adherence, source freshness, and failure modes caused by incomplete, conflicting, or stale context.
- Define sampling strategies, label schemas, inter-annotator agreement targets, adjudication workflows, SME review patterns, and quality thresholds in partnership with Language Data Scientists, clinicians, biomedical experts, and quality teams.
- Build statistical and ML checks that make healthcare datasets trustworthy: stratified sampling across specialties and patient subgroups, bias and representation analysis, leakage detection, distribution shift checks, uncertainty estimates, reliability metrics, and subgroup performance analysis.
- Partner with Applied Research Scientists and AI/ML Research Engineers to instrument datasets into evaluation and post-training pipelines, including rubric-grounded LLM-as-judge prompts, regression suites, model comparison workflows, experiment tracking, and model-improvement feedback loops.
- Evaluate health AI behavior beyond surface accuracy: calibration, hallucination on safety-critical content, refusal appropriateness, robustness under ambiguity, equity across patient subgroups, and safe handoff in agentic or workflow-integrated systems. Reason concretely about clinical workflow fit: where outputs enter care delivery, what evidence a clinician or reviewer would need to trust them, when uncertainty must be surfaced, and how patient-facing, clinician-facing, payer, pharma, and operational use cases differ in risk.
- Own data quality from source intake through delivery, including de-identified clinical text, medical literature, synthetic cases, structured records, client policies, and knowledge bases, with attention to PHI/PII handling, provenance, audit trails, versioning, and compliance documentation.
- Stay current on the health AI landscape — regulatory developments such as FDA guidance on AI/ML-enabled medical devices and EU AI Act health provisions , benchmark releases such as MedQA , MedMCQA , and HealthBench , and emerging clinical evaluation methodology.
- Support customer discovery and proposal work by scoping dataset programs, sizing annotation and SME review effort, identifying regulatory or data-access constraints, and explaining methodology choices to client clinical and ML leadership.
- Contribute to Innodata internal IP: reusable health-domain taxonomies, evaluation rubrics, golden datasets, clinical review playbooks, dataset quality checks, and methodology templates.
You’ll Thrive in This Role If You Have:
- 5+ years of data science experience , including at least 2+ years with healthcare, clinical, biomedical, payer, provider, pharma, life sciences, or comparable regulated health data .
- Working knowledge of healthcare data and standards: EHR structure , clinical documentation conventions, ICD-10 , CPT , SNOMED CT , LOINC , RxNorm , and at least passing familiarity with FHIR , HL7 , or equivalent interoperability concepts.
- Hands-on experience designing ML datasets, not just consuming them: writing annotation guidelines, sizing cohorts, setting quality thresholds, designing QA checks, and shipping data that downstream teams can train or evaluate on.
- Familiarity with LLM-based health AI workflows, including prompt design, rubric-based evaluation, retrieval-augmented generation, LLM-as-judge methods, model comparison, and the limitations of automated evaluation in clinical contexts.
- Strong Python and SQL ; comfort with pandas , scikit-learn , statsmodels or equivalent tools; and working familiarity with modern LLM tooling such as Hugging Face , evaluation frameworks, prompt development tools, or model APIs.
- Statistical literacy across sampling design, bias and fairness analysis, inter-annotator agreement metrics ( Cohen or Fleiss kappa , Krippendorff alpha ), confidence intervals, significance testing where appropriate, error analysis, and the ability to push back when a number is being over-interpreted.
- Solid grasp of healthcare privacy, compliance, and governance: HIPAA , de-identification standards ( Safe Harbor and Expert Determination ), practical mechanics of working with PHI safely, auditability, access control, and documentation fit for high-stakes or regulated AI programs.
- Ability to work credibly with clinicians, biomedical SMEs, research scientists, engineers, technical solutions teams, annotators, and customer stakeholders.
- A bias toward clinical realism: you would rather build a smaller dataset that reflects what clinicians, reviewers, patients, or care teams actually see than a larger dataset that looks impressive on paper but fails in practice.
- Degree in a relevant field such as biostatistics, epidemiology, computational biology, health informatics, computer science with a health focus, statistics, a clinical degree with quantitative training, or equivalent demonstrated experience.
- Clinical credentials are not required, but candidates must be able to work credibly with clinicians, biomedical SMEs, and health AI customers; candidates with MD , RN , PharmD , MPH , PhD , or health informatics backgrounds are especially encouraged.
- ...is one of the highest-stakes domains for generative AI. Numerical accuracy, regulatory compliance, model risk... ..., and AI agents for financial workflows. As an Applied Data Scientist, Financial AI Evaluation & Datasets , you own the design, measurement quality, and...SuggestedFull timeShift work
- ...LLM training, post-training, and evaluation systems. As an AI/ML Research Engineer, LLM Training... ...You will work closely with Language Data Scientists, Applied Research Scientists, data... ...R&D efforts, including benchmark datasets, evaluation frameworks, and reusable...SuggestedFull time
$111k - $160k
...and Summary Senior Data Scientist Overview The... ...Intelligence (AI) and Machine Learning... ...Strong experience applying machine learning... ...analyzing large-scale datasets • Hands-on... ...machine learning evaluation • Ability to identify... ...any medical or health information in this...SuggestedFull timeWorldwide$91k - $140k
...Title and Summary Data Scientist II Overview... ...Artificial Intelligence (AI) and Machine... ..., and performance evaluation • Support model... ...• Experience in applying data science and machine... ...Exposure to large datasets and interest in... ...include any medical or health information in...SuggestedFull timeWorldwide$127k - $203k
...Title and Summary Lead Data Scientist - R&D Overview... ...functional teams, including AI/ML engineering,... ...validation, and performance evaluation. • Experience... ...maintainable code and applying modern software engineering... ...any medical or health information in this email...SuggestedFull timeWorldwide$96.33k - $160.55k
...Req ID: 375278 NTT DATA strives to hire exceptional, innovative... ...forward-thinking organization, apply now. We are currently... ...semi-structured, and streaming datasets. · Enable enterprise reporting... ...are one of the world's leading AI and digital infrastructure providers...Work experience placementWork at officeRemote workFlexible hours$127k - $203k
...Title and Summary Lead Data Scientist Overview:... ...for developing advanced AI and machine learning solutions... ...: • Design, build, evaluate, enhance, and monitor... ..., complex transaction datasets. • Experience with machine... ...any medical or health information in this email...Full timeWork at officeWorldwide3 days per week$200k - $230k
...THE ROLE At SQUIRE, trusted data turns information into action... ...reporting and automation to the AI capabilities shaping our next... ...depth to effectively evaluate architecture Strong ML and... ...t let this list stop you from applying - If you are interested in the...Full timeFor contractorsLocal areaRemote workWorldwide$200k - $300k
...learning and generative AI systems that power... .... This role blends applied machine learning,... ...stakeholders to evaluate buy vs. build decisions... ..., covering data ingestion, feature... ...DDRNet, RFTM with datasets like COCO and Cityscapes... ...You may enroll in health, life, disability and...Full timeWork at office$111k - $160k
...potential. Title and Summary Sr. Data Scientist Overview: Services... ...cyber attacks • Leverage AI models to provide automated transaction... ...attention • Coursework in Applied Mathematics or Statistics •... ...not include any medical or health information in this email....Full timeWorldwideFlexible hours- ...provides a progressive data-driven footprint where... ...structured and unstructured datasets, implement reliable... ...technical advisor to Data Scientists, Engineering Managers,... ...as Computer Science, Applied Mathematics, Statistics... ...coverage options and a Health Care Spending Account....Full timeRemote workWork from homeFlexible hours
$154k - $247k
...Title and Summary Director, Data Scientist Overview: The... ...deployment and performance evaluation. • Strong programming proficiency... ...experience working with large-scale datasets and distributed data... ...not include any medical or health information in this email. The...Full timeWork at officeWorldwide3 days per week- ...Robots & Pencils is an applied AI engineering firm building the next frontier of business architecture. We design and ship AI co-workers... ...build things that actually ship. We’re looking for a Staff Data Engineer to join a multi-disciplinary engineering team building...Full time
$138k - $221k
...Title and Summary Principal AI Engineer Overview: Mastercard... ...help teams build, deploy, evaluate, and operate AI-powered applications... ..., event-driven workflows, and data-intensive workloads. • Build... ...Do not include any medical or health information in this email. The...Full timeWorldwide$142k - $192k
...Senior Research Manager - Data Scientist Job Description: Senior Research... ...capability and focuses on applying advanced quantitative... ...techniques to primary research datasets and partner closely with ARC... ...not include any medical or health information in this email. The...Full timeWorldwide- ...graph, and behavioral data You will build and scale... ...pipelines and training datasets from proprietary and... ...monitor model and data health, and help define retraining... ...define requirements, evaluate tradeoffs, and... ...Proficient in using AI-powered developer tools...Full time
$139k - $181.5k
...world’s first truly converged, AI-native workspace that unifies... ...extreme autonomy to tackle unique data engineering challenges and... ...next to cross-functional data scientists, analytics engineers, and product... ..., including elite group health, dental, and vision insurance...Permanent employmentFull timeTemporary workWork at officeRemote workFlexible hoursShift work- Data Engineer (Data Designing, ETL, ELT, Microsoft SQL Server, Oracle, Snowflake, AWS, Azure, data integration, Data Movement, on premise... ...of issues. ·Improve on number of rollbacks, errors, average time to commit. ================================ Apply for this jobFull timeTemporary workRemote work
- INTEGRITYOne Partners - Data Engineer (Analytics) Location: Washington, D.C. 100% Onsite... ...security, we want to hear from you. Apply now and join our team of dedicated professionals... ...that positively impact our country's health, safety, and security. Our team of dedicated...Full timeMonday to Friday
- ...reaching over 390 million unique visitors each month. We are a data driven company that leverages our data to empower our decisions.... ...software development experience ~ Experience working with large datasets (terabyte scale and growing) and familiarity with various...Full time
$91k - $140k
...security controls and procedures and recommend improvements. · Evaluate appropriate tools for supporting the security operations... ...or assistance you are requesting. Do not include any medical or health information in this email. The Reasonable Accommodations team will...Full timeWork experience placementWorldwide$111k - $160k
...Title and Summary Senior AI Engineer Overview The Security Solutions Data Science team is... ...profile, enabling richer datasets and unlocking new opportunities... ...• Openness to learn and apply new technologies,... ...include any medical or health information in this email...Full timeWorldwide- ...infrastructure. We govern how AI agents reason about,... ...-leverage role for an applied practitioner ready to... ..., device state data, and deep system documentation. Production Evaluations: Construct and expand... ...Comprehensive corporate health insurance provisions....Full timeRemote workFlexible hours
$164.49k - $197.39k
...Grafana Cloud's actually useful AI, organizations can see,... ...and act on all their disparate data to move at the speed of their... ...experimentation. This is an applied product data science role. You... ...design experiments, establish evaluation methodology, and define the scientific...Full timeLocal areaRemote workFlexible hours$154k - $247k
...simple, smart, and accessible. Using secure data and networks, partnerships and passion,... ...onboarding and ongoing consumption. • Apply strong cryptographic knowledge to support,... ...requesting. Do not include any medical or health information in this email. The Reasonable...Full timeWorldwide- ...industrial services POSITION TITLE: Data Center Sales Engineer COMPENSATION: Competitive... ...DUTIES OR RESPONSIBILITIES: Ensuring Health and Safety is the number one priority by... ...NOTE: We try to respond to everyone who applies for our jobs, but in periods of high...Full timeFor contractorsLocal areaRemote work
$136k - $204k
...Design and build end-to-end AI systems integrating ML... ..., APIs, and enterprise data Translate business... ...Work closely with Data Scientists and ML Engineers to... ...decision workflows Apply agent-based workflows and... ...Insurance: You may enroll in health, life, disability and...Full timeWork at office- ...Staff Engineering Technician - Virtual Design & Data Solutions JOB-10046603 Anticipated Start Date July 13, 2026... ...engineering drawings and models under general supervision. This role applies foundational civil engineering principles to develop accurate, high...Full timeContract workWork experience placementLocal area
$127k - $203k
...governments realize their greatest potential. Title and Summary Lead Data Scientist - Financial Crime Overview: Within Financial Crime... ...assistance you are requesting. Do not include any medical or health information in this email. The Reasonable Accommodations team...Full timeWork at officeWorldwide3 days per week- ...The Role We are looking for a highly skilled AI Engineer with 7+ years of experience in software engineering, with a heavy focus... ...) in multi-turn interactions and external API interfacing. 3. Evaluation and Optimization Familiarity with evaluation frameworks (e.g....Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Applied Data Scientist, Health AI Evaluation & Datasets. Be the first to apply!









