Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Remote Math Expert for AI Benchmarking & Evaluation

$100 per hour
Temporary

Turing

Turing is seeking PhD-level mathematicians to design challenging AI evaluation problems and craft rigorous, step-by-step solutions. You will assess AI reasoning, ensure clarity of feedback, and collaborate with researchers to build robust benchmarks across math topics from undergraduate to PhD level.

This fully remote contract role offers flexible hours and a $100/hour pay rate. Candidates should have a PhD in mathematics or related fields, strong reasoning, and communication skills.

#J-18808-Ljbffr
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Remote Math Expert for AI Benchmarking & Evaluation in Lineville, IA vacancy
  • $50 per hour

     ...focuses on improving and evaluating large language models...  ...advances into reliable AI systems for enterprise...  ...Contribute to new evaluation benchmarks spanning curricula from...  ...effectively in a remote setting. Self-motivated...  ...and solve complex math problems using a structured... 
    Remote work
    Contract work
    For contractors
    Freelance

    SaidGig

    United States
    1 day ago
  •  ...Remote Mathematics Expert (AI/LLM) - 34877 Remote Mathematics Expert (AI/LLM) - 3...  ...Design and solve challenging math problems to probe the...  ...align problem types with evaluation goals, particularly in areas...  ...to defining new evaluation benchmarks based on Mathematics curricula... 
    Remote work
    Hourly pay
    Contract work
    Part time
    For contractors
    Freelance
    Internship
    Work from home
    Worldwide
    Afternoon shift

    Turing Inc

    New York, NY
    4 days ago
  • $60 - $80 per hour

     ...technical talent with leading AI research labs....  ...our investors include Benchmark , General Catalyst ,...  ...0/hour Location: Remote Role Responsibilities...  ...real retail practice. Evaluate AI model outputs...  ...with other subject matter experts to ensure consistency and... 
    Remote work
    Contract work
    Summer work
    Weekday work

    Mercor

    New York, NY
    6 days ago
  •  ...to help train next-generation AI systems. Your work will shape...  ...prompts. # Participate in remote collaboration, contributing via...  ...the rigorous training and evaluation of advanced AI models....  ...producing data-annotation tasks, benchmark questions, technical reports,... 
    Remote work
    Temporary work

    AquSag Technologies

    Middletown, OH
    5 days ago
  • $60 - $75 per hour

    Role Description Join an advanced AI research initiative focused on improving how...  ...professionals to design high-quality benchmark tasks that evaluate AI performance across software...  ...structured outputs. This is a fully remote, independent contractor opportunity with... 
    Remote work
    Weekly pay
    Contract work
    Part time
    For contractors
    Flexible hours

    Weekday AI

    Remote
    a month ago
  • $50 per hour

    A leading AI research accelerator is looking for remote PhD candidates in Chemistry, Chemical Engineering, or related fields...  ...design advanced chemistry problems to evaluate AI performance and collaborate with researchers on benchmarks. This role offers flexible hours and a... 
    Remote work
    Hourly pay
    Flexible hours

    Turing

    New York, NY
    2 days ago
  • AuraOne is seeking a Ukrainian Law Expert for the Ukrainian Law Expert — Multiple-Choice AI Benchmark Review (Remote). This remote review track evaluates AI outputs in legal review workflows, with tasks like verifying citations and flagging policy-adherence gaps. As a... 
    Remote job
    For contractors
    10 hours per week
    Flexible hours

    AuraOne

    New York, NY
    1 day ago
  • $60 - $70 per hour

     ...and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel...  ...Compensation: $60–$70/hour Location: Remote Role Responsibilities Evaluate AI-generated responses for safety... 
    Remote work
    Contract work
    Summer work

    Mercor

    New York, NY
    8 days ago
  • $90 per hour

     ...technical talent with leading AI research labs. Headquartered in...  ..., our investors include Benchmark , General Catalyst , Peter...  ...Compensation: $90/hour Location: Remote Role Responsibilities...  ...that merely look correct. Evaluate responsive behavior and semantic... 
    Remote work
    Contract work
    Summer work
    Local area

    Mercor

    San Francisco, CA
    2 days ago
  •  ...Role Overview Provide expert human judgment on commercial drug launches by creating and critiquing evaluation rubrics, building or reviewing launch curves, and assessing the...  ...total over a 1 to 2 week pilot period. Remote work, candidates must be US-based or have deep... 
    Remote work
    Hourly pay

    SaidGig

    Remote
    12 days ago
  • $80 per hour

     ...inventory planning, and supply chain operations to create expert training data and evaluate AI-generated responses for accuracy and relevance. This is...  ...type, hourly contract work, part time. Fully remote, flexible and asynchronous schedule, no minimum weekly hour... 
    Remote work
    Hourly pay
    Contract work
    Part time
    Flexible hours

    SaidGig

    United States
    more than 2 months ago
  • $150 - $180 per hour

     ...Virginia Beach is seeking Mental Health Professionals to train and evaluate AI models. The role involves reviewing AI responses to...  ...attention to detail, and a reliable internet connection. Join our Expert Network to influence future AI innovations and work flexibly from... 
    Remote job
    Work from home

    Prolific

    Virginia Beach, VA
    3 days ago
  •  ...About Us Join an innovative AI initiative dedicated to improving...  ...their real-world expertise by evaluating AI-generated content,...  ...This is a flexible, fully remote, part-time opportunity requiring...  ...multidisciplinary team of subject matter experts to improve AI performance... 
    Remote work
    Part time
    10 hours per week
    Flexible hours

    Weekday

    Remote
    20 days ago
  • $150 - $180 per hour

    Prolific in Sacramento is seeking Mental Health Professionals to help train and evaluate advanced AI models. You will review AI-generated responses and engage in various training tasks, earning competitive pay rates up to $150-180/hr. Ideal candidates must hold a verified... 
    Remote job
    Work from home

    Prolific

    Sacramento, CA
    3 days ago
  • $8 - $65 per hour

    Prolific is hiring Mental Health Professionals in New York to train and evaluate AI models. As a Domain Expert, you will be responsible for reviewing AI-generated responses, completing tasks related to psychology, and improving AI models based on your expertise. Pay rates... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    New York, NY
    1 day ago
  • Prolific in New York, NY, is seeking Chemistry Experts and Chemical Engineers to join its Expert Network. In this role, you will help train and evaluate AI models using your chemical expertise. Duties include evaluating AI-generated responses for accuracy and validating... 
    Remote job
    Work from home
    Flexible hours

    Prolific

    New York, NY
    1 day ago
  •  ...Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant... 
    Remote work
    Flexible hours

    Prolific

    Charlotte, NC
    1 day ago
  •  ...A leading AI research firm is seeking Expert Prompt Curators to design challenging prompts for evaluating advanced AI models. The role requires advanced knowledge in diverse fields and offers flexible hours, remote work, and a competitive hourly wage. Ideal candidates... 
    Remote work
    Hourly pay
    Temporary work
    Flexible hours

    CloudDevs

    New York, NY
    4 days ago
  • $60 per hour

    Prolific in Seattle, WA is searching for Chemistry Experts and Chemical Engineers to join our Expert Network to train AI models with your expertise. The role involves evaluating AI-generated chemistry, fact-checking chemical reactions, and auditing technical documentation... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    Seattle, WA
    5 days ago
  • $50 - $70 per hour

     ...Role Overview Help improve frontier AI models by evaluating the quality of real-world professional materials and AI-generated work. You will...  ...Guidelines and context will be provided. Work Terms Remote, hourly engagement. Applicants must be based in the United... 
    Remote work
    Hourly pay

    SaidGig

    Canada
    2 days ago
  • $65 per hour

     ...security expertise to design domain-specific prompts and evaluate large language model outputs for AI research projects, improving model behavior, safety,...  ..., or Hashcat. Ability to work independently in a remote, asynchronous setting and to document findings clearly... 
    Remote work
    Part time
    Flexible hours

    SaidGig

    United States
    more than 2 months ago
  • A leading AI research accelerator is seeking an Associate to leverage expertise in internal...  .... This role involves designing and evaluating clinical scenarios to enhance AI diagnostics...  ...skills. The position is fully remote, requiring a commitment of 20-40 hours per... 
    Remote job

    Turing

    Seattle, WA
    3 days ago
  • Prolific is seeking Mental Health Professionals in Indianapolis, Indiana, to assist in training and evaluating AI models. Ideal candidates will have a verified status as a Mental Health Professional, an understanding of psychological theory, and the ability to focus on... 
    Remote job
    Hourly pay
    Flexible hours

    Prolific

    Indianapolis, IN
    3 days ago
  • $8 - $65 per hour

     ...Mental Health Professionals in Houston, Texas, to train and evaluate cutting-edge AI models. The role offers flexible hours and competitive pay...  ...This is an opportunity to influence AI development and work remotely. Join Prolific in just 15 minutes after passing the... 
    Remote job
    Hourly pay
    Flexible hours

    Prolific

    Houston, TX
    3 days ago
  • $8 - $65 per unit

    Prolific is seeking Mental Health Professionals to train and evaluate AI models. In this role, you will review AI responses, analyze psychology...  ...pay rates range from $8 to $65 per task, with flexible hours and remote work options available. #J-18808-Ljbffr Prolific
    Remote job
    Flexible hours

    Prolific

    Chicago, IL
    3 days ago
  • $60 per hour

    Prolific seeks Chemistry Experts and Chemical Engineers to join our Expert Network and evaluate AI models using your chemical expertise. Successful candidates will be invited to assess and review AI-generated chemistry tasks and ensure their accuracy. Compensation can reach... 
    Remote job
    Hourly pay
    Flexible hours

    Prolific

    San Francisco, CA
    3 days ago
  • $60 per hour

    Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    San Jose, CA
    5 days ago
  • $8 - $65 per hour

     ...for Mental Health Professionals to help train and evaluate AI models in Phoenix, Arizona. As a Domain Expert participant, you'll review AI-generated psychological...  ...completed task, with flexible hours that allow for remote work. Applicants need verified professional status... 
    Remote job
    Flexible hours

    Prolific

    Phoenix, AZ
    3 days ago
  • $8 - $65 per hour

    Prolific is seeking Mental Health Professionals to train and evaluate AI models from home. Responsibilities include reviewing AI responses and enhancing model performance using psychological expertise. Ideal candidates will have verified professional status, a solid understanding... 
    Remote job
    Flexible hours

    Prolific

    El Paso, TX
    3 days ago
  • Prolific is seeking Mental Health Professionals to train and evaluate AI models. Candidates will review psychological scenarios, analyze...  ...health professionals and must complete a quick skills assessment to join Prolific’s Domain Expert Network. #J-18808-Ljbffr Prolific
    Remote job
    Flexible hours

    Prolific

    San Antonio, TX
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Remote Math Expert for AI Benchmarking & Evaluation. Be the first to apply!