Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Data Scientist for AI Model Evaluation [Remote]

$100 per hour

SaidGig

Remote
  • Remote job
Role Overview

Apply deep domain knowledge to help train and evaluate next-generation AI systems by reviewing, refining, annotating, and validating AI-generated outputs. This part-time, contract role focuses on ensuring accuracy, clarity, and relevance of model outputs through rubric-based evaluation, prompt refinement, fact checking, and high-quality technical writing.

Key Responsibilities
  • Review, edit, and refine AI-generated content and data outputs for accuracy, clarity, and domain relevance according to project rubrics.
  • Develop and optimize prompts to guide AI models toward desired outputs, using professional writing and technical documentation skills.
  • Conduct rubric-based evaluations of AI model performance, providing structured feedback and actionable suggestions.
  • Annotate data, perform fact checking, and participate in quality assurance to maintain analytic standards.
  • Perform independent research to validate facts and improve content quality.
  • Interpret and summarize complex datasets, findings, or analyses into clear reports and technical summaries.
  • Collaborate asynchronously with project leads and other domain experts to share insights and best practices.
Qualifications
  • Minimum 3 years of professional experience in Data Science, Machine Learning, Applied AI, Statistics, Quantitative Analytics, or Data Analytics.
  • Proven experience producing or reviewing research papers, analytical reports, experiment summaries, notebooks, or technical documentation.
  • Strong analytical reasoning, critical thinking, and meticulous attention to detail for written and quantitative deliverables.
  • Advanced proficiency in professional writing, report writing, and business or technical communication.
  • Required skills include critical thinking, analytical reasoning, attention to detail, quality assurance, written communication, technical documentation, prompt authoring and refinement, AI output evaluation, and fact checking.
  • Experience with data annotation, content review, rubric-based evaluation, or professional editing is highly desirable.
  • Background in prompt engineering, AI output evaluation, fact checking, or RLHF (Reinforcement Learning from Human Feedback) is advantageous but not required.
  • Advanced degrees such as a Master, JD, MBA, or PhD are preferred but not mandatory.
Work Terms
  • Engagement type: Independent contractor, part-time.
  • Location: Remote, work performed asynchronously with project teams.
  • Project-based contributions to a customer project focused on advancing AI technology; specific hours are not prescribed and will depend on assignment and project needs.
Compensation
  • Pay range: $100 to $200 per hour.
Eligibility and Application Process
  • No prior AI employment is required; strong domain expertise and experience producing or reviewing technical or research deliverables is the primary qualification.
  • Candidates join after being identified and vetted through the platform''s AI-driven screening process; passing that vetting is required to participate in projects.
  • Participation is as an independent contractor; applicants should be able to engage under contractor terms and provide the necessary professional-level deliverables remotely.
Vacancy posted 9 hours ago
Similar jobs that could be interesting for youBased on the Data Scientist for AI Model Evaluation [Remote] in Remote vacancy
  • $60 - $90 per hour

     ...Role Overview Design and execute realistic, research-grade data analysis tasks that serve as ground-truth references for evaluation of frontier generative AI models. You will create one-to-two day, end-to-end analysis challenges that include data cleaning, statistical... 
    Suggested
    Hourly pay
    Full time
    Part time
    Work experience placement
    Freelance
    Remote work

    SaidGig

    United States
    15 hours ago
  • A leading AI development firm is looking for experienced quantitative professionals to...  ...their team remotely. In this role, you will evaluate AI-generated quantitative analysis and...  ...that shapes how these systems reason about data. Candidates should have 2+ years of relevant... 
    Suggested
    Remote work
    Flexible hours

    DataAnnotation

    Madison, WI
    3 days ago
  • $60 per hour

    A tech company focused on AI is seeking quantitative professionals to evaluate AI-generated analytical work and help advance AI development. Responsibilities include statistical analysis, predictive modeling, and providing feedback to shape AI systems. The role offers flexible... 
    Suggested
    Hourly pay
    Remote work
    Flexible hours

    DataAnnotation

    Helena, MT
    4 days ago
  • $60 per hour

    A leading AI development company is seeking quantitative professionals to evaluate AI-generated analyses and develop solutions in various quantitative fields. This fully remote position allows for flexible scheduling and competitive pay up to $60/hour. Candidates should... 
    Suggested
    Remote work
    Flexible hours

    DataAnnotation

    Providence, RI
    3 days ago
  • $60 per hour

    A leading data analysis firm is seeking experienced quantitative professionals to join their remote team. In this role, you'll evaluate AI-generated quantitative analysis, design problem-solving tasks for AI training, and provide insightful feedback on AI systems. Candidates... 
    Suggested
    Remote work

    DataAnnotation

    Nevada, IA
    3 days ago
  • $60 per hour

    A leading AI development firm is seeking experienced quantitative professionals to evaluate AI-generated quantitative work and contribute to AI systems. This fully remote role offers flexibility in your schedule while allowing you to apply your quantitative skills in real... 
    Hourly pay
    Remote work

    DataAnnotation

    Iowa, LA
    3 days ago
  • $60 per hour

    A leading AI development firm is seeking experienced quantitative professionals to contribute...  ...to AI systems. The role involves evaluating AI-generated work and solving technical problems...  .... Candidates should have a background in data science, economics, or similar fields,... 
    Hourly pay
    Remote work
    Flexible hours

    DataAnnotation

    Nashville, TN
    3 days ago
  • $60 per hour

    A leading data science company is seeking quantitative professionals to evaluate AI-generated analyses and design quantitative problems crucial for advancing AI systems. This fully remote role allows you to work flexibly and choose your projects, with competitive hourly... 
    Hourly pay
    Remote work

    DataAnnotation

    Columbia, SC
    3 days ago
  • $60 per hour

    A leading AI development firm seeks experienced quantitative professionals to evaluate and advance AI systems. This fully remote role allows...  ...you are driven by rigorous data analysis and are comfortable...  ...statistical and predictive modeling, this position is for you. #J... 
    Remote work
    Flexible hours

    DataAnnotation

    Honolulu, HI
    4 days ago
  • $20 per hour

     ...DataAnnotation is committed to creating quality AI. Join our team to help train AI chatbots while...  ...chatbots. You will develop complex prompts to test AI models, write high-quality responses to demonstrate excellence, and evaluate different model outputs based on accuracy and... 
    Hourly pay
    Full time
    Contract work
    Part time
    For contractors
    Self employment
    Freelance
    Remote work

    DataAnnotation

    Wyoming, OH
    3 days ago
  • $60 per hour

    A leading AI development company is looking for quantitative professionals to evaluate AI-generated work and design quantitative problems to improve AI systems. The role offers fully remote work options and a flexible schedule. Candidates should have at least 2 years of... 
    Hourly pay
    Remote work
    Flexible hours

    DataAnnotation

    New York, NY
    3 days ago
  • $60 per hour

    A cutting-edge AI development company is seeking quantitative professionals to evaluate AI-generated work and provide essential feedback. This role allows for remote work...  ...2+ years of relevant experience in fields like data science or statistics, alongside coding skills,... 
    Hourly pay
    Remote work
    Flexible hours

    DataAnnotation

    Florida, NY
    2 days ago
  • $60 per hour

    A leading AI development firm is seeking quantitative professionals to join their remote team. The role involves evaluating AI-generated analyses, designing quantitative problems, and providing...  ...in quantitative fields like data science or statistics, with proficiency... 
    Hourly pay
    Remote work
    Flexible hours

    DataAnnotation

    Salt Lake City, UT
    3 days ago
  • $60 per hour

    A leading AI development company is seeking quantitative professionals...  ...enhance AI systems. You will evaluate AI-generated analyses, solve...  ...provide critical feedback for model improvement. This fully remote...  ...have strong experiences in data science or related fields and... 
    Hourly pay
    Remote work

    DataAnnotation

    Hartford, CT
    3 days ago
  •  ...We are hiring a Senior Solutions Engineer to help shape and scale our AI Data Solutions, working with leading AI labs, frontier model developers, and enterprise AI teams on complex data, evaluation, and model development workflows. This is a senior, customer-facing role... 
    Full time
    For contractors

    L10n People Ltd

    Remote
    5 days ago
  • $60 per hour

    A cutting-edge AI company is looking for experienced quantitative professionals to evaluate AI-generated analyses and solve complex problems. This fully remote role allows...  ...hour. Ideal candidates will have a background in data science or related fields and be comfortable... 
    Remote job
    Hourly pay
    Flexible hours

    DataAnnotation

    New York, NY
    2 days ago
  • $40 per hour

    A leading AI training company is seeking a Quantitative Researcher to enhance AI models through rigorous evaluation and problem-solving in mathematics. This remote role involves measuring chatbot progress and assessing performance across various mathematical fields. Ideal... 
    Remote job
    Hourly pay
    Flexible hours

    DataAnnotation

    Jackson, MS
    2 days ago
  • $238k - $302k

     ...billions in simulation across 15+ U.S. states. The Large Model Evaluation team is at the nexus of Waymo’s AI ambition . With advancements in Large Language...  ...inform model development and deployment.  Build data pipelines for signal discovery, data labeling, feature... 
    Full time
    Remote work

    Waymo

    San Francisco, CA
    15 hours ago
  • $100 per hour

     ...deep domain knowledge to help train and evaluate next-generation AI systems by reviewing, refining,...  ...ensuring accuracy, clarity, and relevance of model outputs through rubric-based...  ..., and refine AI-generated content and data outputs for accuracy, clarity, and domain... 
    Remote job
    Hourly pay
    Contract work
    Part time
    For contractors

    SaidGig

    Canada
    9 hours ago
  •  ...Data Annotator - AI Model Evaluation & Labeling is a remote evaluation track for reviewing data annotator evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    7 days ago
  • $80 per hour

     ...Role Overview Lead hands-on evaluations of frontier AI coding agents by applying them to realistic data engineering workflows, assessing model-produced ETL, data warehouse, analytics, and distributed system implementations, and surfacing bugs, scalability limits, and... 
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    1 day ago
  • A leading AI research company is seeking a Quantitative Researcher to join their team. The role involves training AI models, solving advanced mathematics problems, and ensuring the quality of AI outputs. Fluency in English and expertise in various mathematical fields are... 
    Remote job
    Hourly pay
    For contractors

    DataAnnotation

    Boston, MA
    3 days ago
  • $85 per hour

     ...Role Overview Help evaluate and improve frontier AI coding models by completing realistic machine learning engineering tasks and assessing model-generated implementations. You will work with cutting-edge coding agents to surface bugs, failure modes, and tradeoffs in... 
    Hourly pay
    Remote work

    SaidGig

    United States
    15 hours ago
  • $60 - $90 per hour

     ...Author and run rigorous, multi-step machine learning evaluation tasks for a leading generative AI research team. You will take high-level research ideas...  ...experiments, and analyze results to determine where frontier models succeed or fail. Typical tasks require one to two days... 
    Hourly pay
    Full time
    Freelance
    Remote work

    SaidGig

    United States
    1 day ago
  •  ...Medical professionals apply clinical and workplace expertise to evaluate AI-generated content in their specialty, assess field-specific materials, and provide clear, structured feedback that improves model performance on medical tasks and language. This hourly, temporary... 
    Hourly pay
    Temporary work
    Part time
    Remote work
    Flexible hours

    SaidGig

    United States
    15 hours ago
  • $40 - $65 per hour

     ...conversations and task scenarios that probe frontier language models, then evaluate and document model behavior so engineering teams can...  ...strong organizational structure. Preferred experience with AI human data environments such as RLHF, SFT, evaluations, annotation,... 
    Remote job
    Hourly pay
    For contractors
    Immediate start

    SaidGig

    Frontier County, NE
    1 day ago
  • $150k - $175k

    Role Description As an Applied Data Scientist, Financial AI Evaluation & Datasets, you own the design, measurement quality, and domain validity of the...  ...evaluate, and monitor financial-domain LLMs, vision-language models, multimodal document models, and AI agents. ~... 
    Full time

    Innodata Inc.

    Remote
    1 day ago
  • $150k - $175k

    Role Description As an Applied Data Scientist, Health AI Evaluation & Datasets, you own the design, measurement quality, and clinical validity of datasets...  ...used to train, fine-tune, and evaluate health-domain models. You bring clinical or biomedical fluency and data... 
    Full time

    Innodata Inc.

    Remote
    1 day ago
  • $148k - $184k

    Role Description As a Data Scientist on our team, you will help define how we measure and trust...  ...ll work at the intersection of applied AI evaluation, analytics engineering, and clinical...  ...requirements, then research and build the models and reporting that meet them. This is a... 
    Full time

    Paradigm Health

    Remote
    5 days ago
  • $65 per hour

     ...Role Overview Cybersecurity experts apply security engineering and offensive security skills to design domain-specific prompts, evaluate AI model responses, and help improve the safety, accuracy, and resilience of large language models. This role offers hands-on... 
    Hourly pay
    Part time
    Remote work
    Flexible hours

    SaidGig

    United States
    15 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Data Scientist for AI Model Evaluation [Remote]. Be the first to apply!