Data Scientist for AI Model Evaluation [Remote]
$100 per hourSaidGig
- Remote job
Apply deep domain knowledge to help train and evaluate next-generation AI systems by reviewing, refining, annotating, and validating AI-generated outputs. This part-time, contract role focuses on ensuring accuracy, clarity, and relevance of model outputs through rubric-based evaluation, prompt refinement, fact checking, and high-quality technical writing.
Key Responsibilities- Review, edit, and refine AI-generated content and data outputs for accuracy, clarity, and domain relevance according to project rubrics.
- Develop and optimize prompts to guide AI models toward desired outputs, using professional writing and technical documentation skills.
- Conduct rubric-based evaluations of AI model performance, providing structured feedback and actionable suggestions.
- Annotate data, perform fact checking, and participate in quality assurance to maintain analytic standards.
- Perform independent research to validate facts and improve content quality.
- Interpret and summarize complex datasets, findings, or analyses into clear reports and technical summaries.
- Collaborate asynchronously with project leads and other domain experts to share insights and best practices.
- Minimum 3 years of professional experience in Data Science, Machine Learning, Applied AI, Statistics, Quantitative Analytics, or Data Analytics.
- Proven experience producing or reviewing research papers, analytical reports, experiment summaries, notebooks, or technical documentation.
- Strong analytical reasoning, critical thinking, and meticulous attention to detail for written and quantitative deliverables.
- Advanced proficiency in professional writing, report writing, and business or technical communication.
- Required skills include critical thinking, analytical reasoning, attention to detail, quality assurance, written communication, technical documentation, prompt authoring and refinement, AI output evaluation, and fact checking.
- Experience with data annotation, content review, rubric-based evaluation, or professional editing is highly desirable.
- Background in prompt engineering, AI output evaluation, fact checking, or RLHF (Reinforcement Learning from Human Feedback) is advantageous but not required.
- Advanced degrees such as a Master, JD, MBA, or PhD are preferred but not mandatory.
- Engagement type: Independent contractor, part-time.
- Location: Remote, work performed asynchronously with project teams.
- Project-based contributions to a customer project focused on advancing AI technology; specific hours are not prescribed and will depend on assignment and project needs.
- Pay range: $100 to $200 per hour.
- No prior AI employment is required; strong domain expertise and experience producing or reviewing technical or research deliverables is the primary qualification.
- Candidates join after being identified and vetted through the platform''s AI-driven screening process; passing that vetting is required to participate in projects.
- Participation is as an independent contractor; applicants should be able to engage under contractor terms and provide the necessary professional-level deliverables remotely.
$40 - $65 per hour
...conversations and task scenarios that probe frontier language models, then evaluate and document model behavior so engineering teams can... ...strong organizational structure. Preferred experience with AI human data environments such as RLHF, SFT, evaluations, annotation,...SuggestedRemote jobHourly payFor contractorsImmediate start- ...of the highest-stakes domains for generative AI. Numerical accuracy, regulatory compliance, model risk management, auditability, and customer... ...agents for financial workflows. As an Applied Data Scientist, Financial AI Evaluation & Datasets , you own the design, measurement...SuggestedFull time
- ...of the highest-stakes domains for generative AI. Numerical accuracy, regulatory compliance, model risk management, auditability, and customer... ...for financial workflows. As an Applied Data Scientist, Financial AI Evaluation & Datasets , you own the design, measurement...SuggestedFull timeShift work
- ...Nasdaq: INOD) is a global data engineering company. We... ...Intelligence (AI) are inextricably linked... ...by providing the data, evaluation frameworks, and human expertise... ...with foundation model labs, medical AI startups... ...As an Applied Data Scientist, Health AI Evaluation &...SuggestedFull timeShift work
- Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the... ...artificial intelligence by providing the data, evaluation frameworks, and human expertise required to...SuggestedFull time
$100 per hour
...financial domain expertise to improve AI-driven financial applications... ..., annotating, and refining model outputs and prompts. You will... .... Develop, refine, and evaluate prompts related to financial analysis... ...annotate complex financial data, reports, and model outputs...Remote jobHourly payPart timeFor contractors$100 per hour
...Role Overview Evaluate and optimize AI-generated outputs for a customer-facing project by applying your... ..., clarity, and business alignment of model outputs through detailed review, prompt... ...assessment of AI outputs. Annotate data, interpret findings, and perform fact-checking...Remote jobHourly payPart timeFor contractors$100 per hour
...Help train and refine advanced AI systems by applying deep software engineering expertise to evaluate, edit, and produce high‑quality... ...role focuses on improving how models learn and reason by providing precise... ...technical claims and ensure data-driven accuracy. Interpret...Remote jobHourly payContract workPart timeFor contractors$20 - $60 per hour
...Help improve how large language models create, understand, and modify Office... ...business scenarios, produce and evaluate complex .xlsx, .docx, and .pptx... ...and provide detailed preference data and feedback that guide model improvements. No prior AI experience is required. Key...Remote jobHourly payFor contractorsWork at office- ...in LLM training, post-training, and evaluation systems. As an AI/ML Research Engineer, LLM Training &... ...the technical foundations that power model improvement for foundation model... ...You will work closely with Language Data Scientists, Applied Research Scientists, data engineers...Full time
- ...in LLM training, post-training, and evaluation systems. As an AI/ML Research Engineer, LLM Training &... ...the technical foundations that power model improvement for foundation model... ...You will work closely with Language Data Scientists, Applied Research Scientists, data engineers...Full time
$91k - $140k
...greatest potential. Title and Summary Data Scientist II Overview The Security... ...Artificial Intelligence (AI) and Machine Learning (ML) models that power Mastercard’s Identity and... ...testing, benchmarking, and performance evaluation • Support model monitoring, benchmarking...Full timeWorldwide$111k - $160k
...potential. Title and Summary Senior Data Scientist Overview The Security... ...developing Artificial Intelligence (AI) and Machine Learning (ML) models that support Mastercard’s Identity... ...experimentation, and machine learning evaluation • Ability to identify...Full timeWorldwide$154k - $247k
...Title and Summary Principal Data Scientist Overview: Mastercard is... ...development and delivery of advanced AI and machine learning... ...Impact • Build and scale AI/ML models that enrich merchant data,... ...prototypes and proofs-of-concept to evaluate new methods before scaling...Full timeWorldwide$20 - $40 per hour
...used to train next-generation AI systems. This role contributes... ...reasoning, helping improve how models learn and perform. No prior AI... ..., privilege, and data privacy obligations in legal practice... ...will be used for training and evaluating AI systems. Materials must...Remote jobHourly payContract workFor contractors$127k - $203k
...potential. Title and Summary Lead Data Scientist - R&D Overview We are... ..., scalable machine learning models, predictive algorithms, and... ...-functional teams, including AI/ML engineering, product, and... ...validation, and performance evaluation. • Experience working with...Full timeWorldwide$127k - $203k
...potential. Title and Summary Lead Data Scientist Overview: We work to connect and... ...Machine Learning, Deep Learning, and other AI techniques. Responsibilities: 1.... ..., data cleaning, preparation, modeling, and evaluation 5. Deploy, document, maintain and monitor...Full timeWorldwide$6 - $8 per hour
...directly train next-generation AI systems. As a remote... ...corrections that improve how models learn, reason, and perform. No... ...correct ambiguities or errors in data samples. Document annotation... ...experience required, applicants are evaluated on domain expertise, annotation...Remote jobHourly payFor contractorsWork at office$127k - $203k
...potential. Title and Summary Lead Data Scientist Overview: Services... ...for developing advanced AI and machine learning solutions... ...transaction data and machine learning models focused on merchant risk,... ...Role: • Design, build, evaluate, enhance, and monitor machine...Full timeWork at officeWorldwide3 days per week$20 - $60 per hour
...document expertise to a project that trains next-generation AI systems. You will design realistic Fortune 500 style scenarios and interact iteratively with an advanced language model to create, edit, and evaluate Office Open XML files, with a focus on .pptx deliverables....Remote jobHourly payContract workFor contractorsWork at office$140 per hour
...subject matter expertise to help train next-generation AI systems by reviewing and annotating legal materials, evaluating AI-generated legal text, and delivering concise,... ...and emerging issues in criminal law to improve model learning. Summarize legal information clearly...Remote jobHourly payFor contractors$80 - $160 per hour
...Your contributions will be used to train and evaluate next generation AI systems, providing high quality, domain-grounded data and judgments. Prior experience in AI is not... ...producing annotated explanations suitable for model training. Compensation ~ Pay range, $80....Remote jobHourly payFor contractors$140 per hour
...review, and refine real-world legal content that trains and evaluates next-generation AI systems. This contract role focuses on transforming... ...domain. Draft and refine realistic scenarios, queries, and model responses to simulate authentic civil rights legal situations...Remote jobHourly payContract workFor contractorsVisa sponsorship$80 - $90 per hour
...Overview Apply PhD-level biology expertise to evaluate, create, and refine biology content used to train and evaluate next-generation AI systems. This remote contractor role... ..., and scientific rigor of AI training data and model outputs. Prior experience in AI is not required...Remote jobHourly payFor contractors$20 - $40 per hour
...trains and improves next-generation AI systems in the life sciences.... ...accurate scientific communication, data annotation, and content development to help models learn, reason, and perform. No... ...educational content. Author, evaluate, and edit laboratory reports, research...Remote jobHourly payFor contractors$80 - $160 per hour
...many-body systems, with a focus on problems such as the PXP model, Rydberg blockade, quantum many-body scars, and constrained... ...analyses that support research-level benchmarks and help train and evaluate advanced AI systems. Key Responsibilities Implement and analyze...Remote jobHourly payFor contractors$60 - $80 per hour
...leverages your expertise in molecular biology to advance AI technology with authentic, expert-driven data. As a Molecular Biology Expert, you will play a... ...molecular biology concepts, techniques, and findings. Evaluate AI-generated outputs for scientific validity, clarity...Remote jobHourly payFor contractors$111k - $160k
...governments realize their greatest potential. Title and Summary Sr. Data Scientist Overview: Services within Mastercard is responsible... ...global banking networks to detect cyber attacks • Leverage AI models to provide automated transaction risk decisions All About...Full timeWorldwideFlexible hours- ...and machine learning for the social good sector. We're looking for a foundational member of our data team to architect the data models and intelligence layer that enable AI agents to automate business processes across our products. This is a greenfield opportunity to work...Full time
$25 - $40 per hour
...personal training expertise to design, evaluate, and provide clear written and... ...content that helps train next-generation AI systems. You will shape how models learn, reason, and perform by... ...the key contribution. micro1 is an AI data lab that converts subject matter expertise...Remote jobHourly payFor contractors
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Scientist for AI Model Evaluation [Remote]. Be the first to apply!
- junior data scientist remote Canada
- healthcare data scientist Canada
- python data scientist Canada
- energy data scientist Canada
- senior data scientist Canada
- data scientist Canada
- data scientist (hedge fund) Canada
- python data scientist (contract) Canada
- entry level data scientist remote Canada
- data visualization Canada



