Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Data Scientist for AI Model Evaluation

$100 - $150 per hour

SaidGig

Role Overview

Join a talent network of experienced data scientists who may contribute to future projects evaluating how effectively AI systems perform real-world data science work for leading AI research organizations.

Key Responsibilities

  • Create precise, task-specific grading criteria for data science deliverables, including exploratory analyses, statistical modeling, machine learning pipelines, experimentation and A/B test write-ups, feature engineering, and technical reports or notebooks.
  • Evaluate AI-generated or human-created work using established criteria.
  • Provide detailed written explanations supporting evaluations and scores.
  • Apply consistent, evidence-based judgment to produce reproducible, defensible assessments.
  • Incorporate structured feedback from senior reviewers and revise submitted work accordingly.

Responsibilities will vary by project.

Qualifications

  • At least 1 year of professional data science experience.
  • Experience at a leading technology, research, or quantitative firm, such as a top FAANG company, AI lab, top-tier quantitative fund, or equivalent organization.
  • Strong Python and SQL skills, plus expertise in statistical modeling, machine learning, experimentation, causal inference, and turning messy real-world data into rigorous analyses.
  • Exceptional written communication skills and the ability to explain technical findings clearly.
  • A detail-oriented, consistent approach to evaluating complex work.
  • Comfort receiving feedback and calibrating judgment to established standards.

Work Terms

  • Remote, hourly engagement.
  • There is no immediate project opening. Qualified applicants may be contacted as relevant opportunities become available.

Compensation

$100 to $150 per hour.

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the Data Scientist for AI Model Evaluation in United States vacancy
  •  ...Apply your data science and quantitative expertise to improve how AI models reason through statistics, machine learning, experimentation...  .... Key Responsibilities Evaluate AI model outputs for data...  ...1 year of experience as a data scientist or quantitative professional at... 
    Suggested
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    a month ago
  • $100 - $150 per hour

     ...Role Overview Apply your data science expertise to evaluate AI-generated slides, spreadsheets, and documents for real-world quality and usability. You will assess outputs against professional standards and deliver clear feedback that improves their accuracy and presentation... 
    Suggested
    Hourly pay
    Work at office
    Remote work

    SaidGig

    United States
    3 days ago
  •  ...for providing independent assurance and evaluating the company's risk management,...  ...actions until completion.We are looking for data scientists and AI developers who will power our mission...  ...Proficiency in frameworks for auditing models, including criteria like robustness, fairness... 
    Suggested

    TikTok

    Los Angeles, CA
    12 hours ago
  • $60 - $90 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation... 
    Suggested
    Full time
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    a month ago
  • $147.5k - $211k

     ...days per week Business UnitTechnology, Data, AI and Ventures (TDAV)Within the Tech,...  ...OverviewThe Corporate Vice President, Data Scientist - Model Validation and AI Governance will play...  ...challenging model methodologies, evaluation approaches, controls, and monitoring strategies... 
    Suggested
    Local area
    3 days per week

    New York Life Insurance Company

    New York, NY
    12 hours ago
  • $100 per hour

     ...Apply data science expertise to help improve next-generation AI systems through rigorous content review, prompt development, model evaluation, and data-quality work. This remote, part-time contract role is designed for specialists who produce or review research papers,... 
    Hourly pay
    Contract work
    Part time
    For contractors
    Remote work

    SaidGig

    United States
    a month ago
  • $184k - $287.5k

     ...is redefining what is possible with AI, and the Relational Foundation Model team is helping lead that transformation...  ...models: you will design, build, and evaluate novel Transformer and graph neural...  ...that generalize across diverse data schemas. Your work will power meaningful... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $110.7k - $218.3k

     ...performance.Work you'll doAs a Data Science Analytics and...  ...validating quantitative workforce models that translate...  ...Collaborating with Senior Data Scientists and AI engineers to integrate scenario...  ...feature engineering, model evaluation, and performance tuning techniquesExperience... 
    Civilian Contractor
    Full time
    Local area

    Deloitte

    Rosslyn, VA
    3 days ago
  • $123.3k - $140.7k

     ...Overview Senior Associate, Data Scientist - Model Risk Audit Data is at the center of everything...  ...latest emerging NLP and Generative AI technologies. If you love a fast-paced...  ..., from design through training, evaluation, validation, and implementation Flex... 
    Full time
    Part time
    Local area
    Flexible hours

    Capital One

    McLean, VA
    more than 2 months ago
  • $204k - $216k

     ...Job Description Sapience AI is the collective intelligence...  ...research role focused on the models at the foundation of collective...  ...platform: designing experiments, evaluating models, adapting them to the...  ...Foundational Model Research Data Scientist does that. You run the... 

    Sapience AI Corporation

    San Francisco, CA
    8 days ago
  •  ...is building the foundation for physical AI — a unified platform that combines high-...  ...to build the infrastructure that powers data, models, and AI runtime across Dexmate’s Physical...  ...model infrastructure for training, evaluation, versioning, deployment, and serving.... 
    Full time

    Dexmate

    Fremont, CA
    a month ago
  •  ...Aerospace VTI Aerospace builds AI-powered perception and pilot...  ...a Senior Software Engineer – Evaluation, you will design and implement...  ...(ASR), and small language model (SLM) systems. You will develop...  ...closely with machine learning and data engineering teams to evaluate... 

    VTI Aerospace

    Seattle, WA
    a month ago
  • $110k - $220k

     ...Opportunity: Walmart’s Supply Chain AI Lab & Innovation Factory is...  ...improve through rigorous evaluation and model post-training. This role...  ...is not an analytics-focused data science role. It is a deeply...  .... As a Principal Data Scientist in this space, you are a hands... 
    Full time
    Contract work
    Temporary work
    Part time

    Walmart

    Bentonville, AR
    1 day ago
  • $70 - $90 per hour

     ...Role Overview Help evaluate Neuron Kernel Interface development tasks that support the training and evaluation of advanced AI models. You will assess kernel quality, numerical correctness...  ...topology, and FP32, BF16, FP8, and INT8 data types. Experience benchmarking... 
    Hourly pay
    Remote work

    SaidGig

    Remote
    a month ago
  • $70 - $90 per hour

     ...Review GPU and accelerator kernel development tasks that support the training and evaluation of advanced AI models. This role focuses on assessing task quality, numerical correctness, completeness, fair performance benchmarking, appropriate scope, and whether kernels... 
    Hourly pay
    Remote work

    SaidGig

    Remote
    a month ago
  • $136.44k - $265.11k

     ...rebuilding biotech for the AI era.When a breakthrough...  ...structured scientific data and AI are built into...  ...for biotech R&D. Scientists use Benchling to design...  ...and run AI agents and models directly in their workflows...  ...ll build the datasets, evaluations, and systems that help... 
    Work at office
    Local area
    Monday to Friday
    Shift work

    Benchling

    San Francisco, CA
    4 days ago
  • Python Infrastructure Engineer - Model Evaluation (AI Training) About the Role What if your Python expertise could directly shape how the world...  ...Python Infrastructure Engineer to design and build the data pipelines, annotation tooling, and evaluation systems that leading... 
    Hourly pay
    Ongoing contract
    Contract work
    Freelance
    Remote work
    Flexible hours

    Alignerr

    Seattle, WA
    4 days ago
  • General Motors, through Embodied AI, seeks a Senior Engineer to measure and visualize AV model performance. You will design and implement evaluation workflows, collaborate across Data, Infra and validation teams, and connect offline evaluation with real-world behavior.... 
    Remote job

    General Motors

    Austin, TX
    1 day ago
  • Data Scientist - AV Metrics & Evaluation Analytics Austin, Texas About the team Our team is dedicated to advancing...  ...with machine learning (e.g., model evaluation, feature engineering). Familiarity...  ...the essential functions of a job, please email ****@*****.***.ai. Avride
    Remote work
    Relocation

    Avride

    Austin, TX
    1 day ago
  • $20 per hour

    SupportFinity™ in Maine is seeking an Editorial Proofreader to join their team focused on training AI models. The position involves evaluating AI chatbot outputs and improving model quality through expert editing and writing skills. This flexible role allows you to work... 
    Remote job
    Hourly pay
    Flexible hours

    SupportFinity™

    Montgomery, AL
    4 days ago
  •  ...TX / Seattle, WA - Onsite On Contract AI Evaluation Model Risk Lead Job Responsibilities Own...  ...Engineering, Artificial Intelligence, Data Science, or a Related Field (Preferred)...  ...AI models. (Preferred) Certified Data Scientist (CDS): Endorses skills in data science... 
    Contract work
    Work experience placement

    Amaze Systems

    Bellevue, WA
    3 days ago
  • Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant... 
    Remote job
    Flexible hours

    Prolific

    Charlotte, NC
    3 days ago
  • $20 per hour

    SupportFinity™ is seeking an Editorial Proofreader to evaluate AI models and improve their quality through expert writing and editing skills. This role can be part‑time or full‑time, allowing for a flexible schedule and project selection. Applicants must be fluent in English... 
    Remote job
    Hourly pay
    Full time
    Part time
    Flexible hours

    SupportFinity™

    Sioux Falls, SD
    4 days ago
  • SupportFinity™ is looking for an Editorial Proofreader to join our team to train AI models. In this role, you will measure AI chatbot progress, evaluate logic, and solve problems to enhance model quality. Applicants should have a strong command of English and experience... 
    Remote job
    Hourly pay
    Full time
    Part time
    Flexible hours

    SupportFinity™

    Columbia, SC
    4 days ago
  • $60 per hour

     ...in Arizona, is seeking Biology Experts and Life Science Professionals to join their Expert Network. This role involves evaluating and training AI models with real scientific expertise. Successful candidates will review AI-generated content for accuracy and assist in fact... 
    Hourly pay

    Prolific

    Arizona City, AZ
    3 days ago
  • $30 - $35 per hour

    Model Evaluation Operator (Swing Shift) Milpitas, CA Why RoboForce RoboForce is an AI robotics company developing Physical AI-powered Robo-Labor for dull, dirty, and dangerous...  ...will also collect high-quality operational data and performance feedback that directly contribute... 
    Hourly pay
    Monday to Friday
    Shift work
    Afternoon shift

    roboforce

    Milpitas, CA
    4 days ago
  •  ...is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities... 
    Remote job
    Hourly pay
    Flexible hours

    Prolific

    Jacksonville, FL
    3 days ago
  • $20 per hour

    SupportFinity™ is looking for an Editorial Proofreader to join our team for AI model training. In this remote role, you'll evaluate AI chatbots and enhance model quality. Candidates should have fluency in English and strong editing skills. This position can be full‑time... 
    Remote job
    Hourly pay
    Full time
    Contract work
    Part time

    SupportFinity™

    New York, NY
    3 days ago
  • $100 per hour

     ...Role Overview Apply your finance expertise to help improve AI models across complex financial problem-solving areas, including capital...  ...prior AI experience is required. Key Responsibilities Evaluate language models in finance domains where performance needs improvement... 
    Hourly pay
    Contract work
    Remote work
    10 hours per week
    Flexible hours

    SaidGig

    United States
    more than 2 months ago
  •  ...Role Overview Use your investment and finance expertise to evaluate and improve AI model performance on financial reasoning, valuation, markets, and real-world investment scenarios. Key Responsibilities Assess AI model outputs on valuation, financial modeling, markets... 
    For contractors
    Remote work

    SaidGig

    United States
    29 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Data Scientist for AI Model Evaluation. Be the first to apply!