Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Evaluation Scientist: Metrics & Insights (Remote)

$50 - $70 per hour

United States Digital Space LLC

New York, NY
  • Remote job

United States Digital Space LLC is seeking a Data Scientist for AI Evaluation Analytics to work as a remote contractor on long-term projects. You will develop evaluation metrics, validate datasets, analyze signals, and apply statistical methods to measure AI system performance. This role emphasizes clear communication of insights and collaboration with engineering teams; pay ranges $50–$70 per hour and 100% remote. #J-18808-Ljbffr United States Digital Space LLC

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the AI Evaluation Scientist: Metrics & Insights (Remote) in New York, NY vacancy
  •  ...We are looking for an AI Evaluation Scientist to design and execute evaluation processes that ensure...  ...product teams to develop evaluation metrics, build test harnesses, analyze model behavior...  ...qualitative and quantitative insight into evaluation reports. Assist in... 
    Remote work

    Convergenz

    United States
    5 days ago
  •  ...Intelligence is the open platform for evaluating how AI models perform in the real...  ...variety of Machine Learning Scientist to help advance how we...  ...to productionize research insights and iterating with product teams...  ...dimensions Develop new metrics, methodologies, and evaluation... 
    Suggested
    Permanent employment
    Work at office

    Arena

    San Francisco, CA
    4 days ago
  •  ...technology by developing the metrics and evaluation frameworks that...  ...a skilled Data Scientist to join our team and...  ...based on data‑driven insights. Build scalable data...  ...relocation sponsorship, and remote work options are not...  ...please email ****@*****.***.ai. #J-18808-Ljbffr... 
    Remote work
    Relocation

    Avride

    Austin, TX
    4 days ago
  •  ...To support AI development, the part-time AI Evaluation Specialist will review and assess AI-generated outputs for...  ...refine evaluation standards in a remote contract role. Key responsibilities...  ...various content types Provide insightful feedback to inform model improvements... 
    Remote work
    Contract work
    Part time

    Virtual Vocations Inc

    United States
    5 days ago
  •  ...Join a pioneering AI initiative focused on building next-generation evaluation benchmarks for frontier AI models. We are seeking...  ...AI systems. This is a fully remote, full-time engagement requiring...  ...grading criteria. Share insights and recommendations with cross-functional... 
    Remote work
    Full time
    Contract work
    For contractors
    Flexible hours

    Weekday

    Remote
    13 days ago
  •  ...growing data-driven company is looking for an AI Data Specialist in the United States to...  ..., and supply chain data into actionable insights. The ideal candidate should have 4-8...  ...operational performance. This role supports remote work from any state. #J-18808-Ljbffr Matcha
    Remote work

    Matcha

    Los Angeles, CA
    2 days ago
  • $60 - $85 per hour

     ...About the job Remote | Licensed Chemical Engineer & AI Evaluation Specialist - $60-$85/hour We are sharing a specialised part-time consulting opportunity...  ...relevant reference materials, and provide practical insight into real-world chemical engineering workflows, standards... 
    Remote work
    Hourly pay
    Full time
    Contract work
    Part time
    10 hours per week
    Flexible hours

    24-MAG LLC

    New York, NY
    23 days ago
  • $175 per hour

     ...technical talent with leading AI research labs. Headquartered in...  ...Report Quality & Strategic Insights Expert Type: Contract...  ...: $175/hour Location: Remote Duration: 1 month initially...  ...and storytelling. Develop evaluation rubrics to distinguish exceptional... 
    Remote work
    Contract work
    Summer work

    Mercor

    Dallas, TX
    more than 2 months ago
  •  ...AI Trainer & Evaluator In this role, you'll apply your expertise to help train next-generation...  ...responses using well-defined rubrics and metrics across diverse subject areas. #...  ...independently and collaboratively in a fully remote environment. Preferred... 
    Remote work

    micro1

    United States
    2 days ago
  •  ...French Audio Evaluations Specialist - Freelance AI Trainer Project World Wide - Remote Project Overview We are sourcing independent Audio Evaluation Specialists...  ...across standardized qualitative and quantitative metrics, focusing strictly on task completion accuracy... 
    Remote work
    Hourly pay
    For contractors
    Freelance

    Meridial

    United States
    3 days ago
  • $30 - $50 per hour

    A tech company is seeking an AI Researcher to support end-to-end research for modern AI systems. This remote role involves designing experiments, defining evaluation protocols, and improving evaluation rigor for large language models. Key responsibilities include developing... 
    Remote job
    Hourly pay

    Rex.zone

    New York, NY
    4 days ago
  • A leading AI research platform is seeking an AI Trainer with expertise...  ...Design. The role involves evaluating designs, ensuring they meet...  ...standards, and providing valuable insights for AI development. Candidates...  ...pay, flexible hours, and remote work options. #J-18808-Ljbffr... 
    Remote job
    Flexible hours

    Prolific

    Los Angeles, CA
    1 day ago
  • Rex.zone is seeking an AI Research Scientist to lead applied AI research projects for US-based customers...  ...into measurable experiments in LLM evaluation and RLHF data design. You will...  ..., safety, and usefulness. The role is remote in the United States, with compensation... 
    Remote job
    Hourly pay
    Flexible hours

    AIToolboard

    New York, NY
    10 hours ago
  •  ...Mercor is hiring for a STEM Computational Scientific Software & Evaluation Design position in Astrophysics & Cosmology. This role is remote and requires candidates to design computational problems and evaluate AI systems for research purposes. A graduate degree in a... 
    Remote work
    Part time

    Mercor

    New York, NY
    1 day ago
  • $105k - $145k

    Steampunk is seeking an AI Evaluation Scientist in McLean, Virginia, to design evaluation processes for AI systems ensuring accuracy and reliability. Responsibilities include implementing evaluation frameworks and performing error analysis. Ideal candidates will have a... 

    Steampunk

    Mc Lean, VA
    1 day ago
  • This AI Research Scientist will lead the design and build biological...  ...to derive actionable insights from a broad range of...  ...project work. Fully remote applicants will not...  ...and develop rigorous evaluation frameworks and...  ...including alignment metrics, retrieval precision,... 
    Remote work
    3 days per week

    Q-state Biosciences

    Cambridge, MA
    3 days ago
  • $80 per hour

     ...ethically shape the future of AI. What We Do The Mindrift...  ...design realistic and structured evaluation scenarios for LLM‑based agents...  .... Understanding of scoring metrics (precision, recall, coverage,...  .... Take part in a flexible, remote, freelance project that fits... 
    Remote work
    Temporary work
    Part time
    Freelance
    Flexible hours

    Mindrift

    Wisconsin
    2 days ago
  •  ...integrating modern AI and LLMs into the...  ...which translates insights from traditional ML...  ...As a Staff AI Scientist , you will own the...  ...tuning, retrieval, and evaluation. You will be part...  .... This is a US Remote role. What You...  ..., and proxy metrics that responsibly accelerate... 
    Remote work
    Full time
    Flexible hours

    Oura

    Remote
    5 days ago
  • $130k - $162.5k

     ...advanced data analytics and AI to cybersecurity, we...  ...AI team. We are AI/ML scientists and engineers with deep...  ..., Optum Health, Optum Insight, and Optum Rx. In addition...  ...flexibility to work remotely * from anywhere within...  ...strategies, priorities, and metrics for technical progress... 
    Remote work
    Minimum wage
    Full time
    Work experience placement
    Work at office
    Local area

    UnitedHealth Group

    Bellevue, WA
    1 day ago
  •  ...position of STEM Computational Scientific Software & Evaluation Design in Astrophysics & Cosmology. This position is remote with a commitment of 15-20 hours per week....  ...include designing computational problems, evaluating AI systems, and developing strategies for data... 
    Remote job
    Contract work

    Mercor

    New York, NY
    4 days ago
  • $211k - $290.5k

     ...the right tools and insights, we believe that...  ...Faire’s user facing AI bet within the...  ...Senior Applied AI/ML Scientist on the Compass...  ...quality through data, evaluation, and modeling,...  ...suites, LLM-as-judge metrics, and quality criteria...  ...to work remotely up to 4 weeks per... 
    Remote work
    Work experience placement
    Work at office
    Local area
    Immediate start
    Monday to Friday
    Flexible hours
    3 days per week

    Faire

    San Francisco, CA
    1 day ago
  • $77.6k - $176k

    AI and ML Data ScientistThe Opportunity: As an Agentic...  ...AI Engineer and Data Scientist for military...  ...generation pipelines, evaluation frameworks, prompt strategies...  ...effectively to accelerate insight, improve analytic tradecraft...  ...on during meetings.Remote: If this position is... 
    Remote work
    Full time
    Contract work
    Part time
    Work at office
    Local area

    Booz Allen Hamilton

    McLean, VA
    1 day ago
  • A leading research accelerator is seeking a Geospatial Expert to enhance AI systems through advanced geospatial analysis. This entry-level, remote role involves evaluating geospatial datasets and supporting tasks aligned with crisis management and agriculture. Candidates... 
    Remote work
    Contract work
    Temporary work

    Turing

    Washington DC
    1 day ago
  • $20 per hour

    A leading AI training company is seeking remote workers to assist in training AI chatbots. Responsibilities include developing conversations, writing responses, and evaluating AI performance. The ideal candidates are analytical and detail-oriented with strong communication... 
    Remote work
    Hourly pay
    Flexible hours

    DataAnnotation

    Madison, WI
    3 days ago
  • $20 per hour

    A growing AI development company is seeking individuals for a remote position to help train AI chatbots. You will create diverse conversations, provide high-quality responses, and evaluate AI model performance. This role is perfect for those with experience in managing... 
    Remote work
    Hourly pay
    Flexible hours

    DataAnnotation

    Columbia, SC
    3 days ago
  • $20 per hour

    A leading AI training company is seeking analytical, detail-oriented individuals to remotely teach AI chatbots. Responsibilities include developing prompts, writing high-quality responses, and evaluating different AI models. The role is ideal for those with experience in... 
    Remote work
    Hourly pay
    Full time
    Part time
    Flexible hours

    DataAnnotation

    Providence, RI
    1 day ago
  • $25 - $30 per hour

     ...independent contractors in the Town of Vermont, Wisconsin to help train AI chatbots. Ideal candidates are proactive individuals who can...  ...include creating complex prompts, writing responses, and evaluating AI models. Candidates are required to have strong writing and research... 
    Remote work
    Hourly pay
    For contractors
    Flexible hours

    DataAnnotation

    Vermont
    1 day ago
  • $20 per hour

    A tech firm specializing in AI seeks analytical and detail-oriented individuals for remote work in training AI chatbots. The role involves developing diverse conversation...  ...prompts, writing high-quality answers, and evaluating different AI models. Ideal candidates should have... 
    Remote work
    For contractors
    Freelance
    Flexible hours

    DataAnnotation

    Lansing, MI
    1 day ago
  • $175k - $275k

    Full-Time in Austin, TX Remote (any location) -...  ...$275k Applied Data Scientist, LLM Evaluation Introduction At Driver...  ...layer for employees and AI agents alike to use...  ...it. Define quality metrics and build evaluation...  ...team to turn evaluation insights into shipped improvements... 
    Remote job
    Full time
    Flexible hours

    Driverai

    Austin, TX
    4 days ago
  • $70 per hour

     ...Compensation: $70 per hour Join a cutting-edge AI research initiative focused on improving...  ...thinking and communication skills to evaluate AI-generated responses across a variety...  ...challenging tasks. This is a fully remote, contract-based opportunity with flexible... 
    Remote work
    Hourly pay
    Weekly pay
    Contract work
    For contractors
    Flexible hours

    Weekday AI

    United States
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Evaluation Scientist: Metrics & Insights (Remote). Be the first to apply!