Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Mathematician for AI Model Evaluation

$50 per hour

SaidGig

Shape and evaluate advanced AI systems through rigorous mathematical reasoning, problem-solving, computational work, and clear written explanations. This remote contract role focuses on creating and assessing challenging mathematics tasks across undergraduate through Ph.D.-level subject areas. Key Responsibilities

  • Design original, challenging mathematics problems that test large language model reasoning in multi-step, abstract, and proof-based settings.
  • Solve problems independently and produce detailed, logically structured solutions with clear justifications.
  • Review model-generated solutions, identify mathematical errors and missing arguments, and provide precise feedback, annotations, and corrections.
  • Help define mathematics evaluation benchmarks spanning early undergraduate through Ph.D.-level curricula.
  • Design precise, closed-ended computational prompts, write reliable Python solutions, validate numerical answers, and provide clear rationales using approved scientific libraries.
  • Complete theorem-prover work in Lean, including translating mathematical problems and proofs into formal language and verifying that formal proofs compile correctly.
  • Break down complex mathematical concepts into clear explanations using simple language, visuals, and examples.
Qualifications
  • Strong mathematical foundation at engineering entrance-exam and graduate-program levels.
  • Research, analytical, creative, and lateral-thinking skills.
  • Ability to solve complex mathematics problems with a structured, logical approach.
  • Ability to provide constructive feedback and detailed annotations.
  • Excellent structured communication and collaboration skills for a remote environment.
  • Self-motivated, able to work independently, and able to work efficiently.
  • Desktop or laptop with a reliable internet connection.
Eligibility
  • Candidates pursuing a Master’s, Ph.D., or postdoctoral degree in Mathematics, Applied Mathematics, Statistics, or a related field are eligible and encouraged to apply.
Work Terms
  • Fully remote contract, independent-contractor assignment.
  • Choose a commitment of 20, 30, or 40 hours per week.
  • Commit at least 4 hours per day, with a minimum of 20 hours per week.
  • Maintain 4 hours of overlap with Pacific Time.
  • This contractor engagement does not include medical coverage or paid leave.
Vacancy posted 15 days ago
Similar jobs that could be interesting for youBased on the Mathematician for AI Model Evaluation in United States vacancy
  •  ...reasoning and computational problem solving to improve and evaluate large language models. You will design rigorous math problems, produce clear, logically...  ...How this work supports customers Accelerate frontier AI research by contributing high quality data and evaluation... 
    Suggested
    Contract work
    For contractors
    Freelance
    Remote work

    SaidGig

    United States
    a month ago
  • $60 - $80 per hour

     ...mathematics experts considered for future contract opportunities with AI labs and companies. This is an open application, not a posting...  ...by applying mathematics knowledge to real-world tasks, model evaluation, and domain-specific feedback. Key Responsibilities... 
    Suggested
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    more than 2 months ago
  • $60 - $80 per hour

     ...Role Overview Provide high-level mathematical expertise to support AI research and product development. Mathematicians in this expert network train and evaluate mathematical models, design realistic problem tasks and deliverables, and give domain-specific feedback that... 
    Suggested
    Hourly pay
    Contract work
    Immediate start
    Remote work

    SaidGig

    United States
    2 days ago
  • $100 - $150 per hour

     ...data scientists who will be considered for future projects evaluating how well AI systems perform real-world data science tasks. Members of this...  ...criteria, assess AI or human-produced analyses and models, document decisions in writing, and iterate on evaluations with... 
    Suggested
    Hourly pay
    Immediate start
    Remote work

    SaidGig

    United States
    a month ago
  • $60 - $90 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation... 
    Suggested
    Full time
    Contract work
    Summer work
    Remote work

    Mercor

    Remote
    1 day ago
  • A leading AI development firm is seeking a Mathematician (PhD) to join their remote AI training project. In this role, you will evaluate AI-generated mathematical responses, ensuring accuracy and clarity. To qualify, you must have a PhD in Mathematics/Statistics, significant... 
    Remote job
    Weekly pay
    Flexible hours

    SME Careers

    Cambridge, MA
    2 days ago
  • $136.44k - $265.11k

    We are rebuilding biotech for the AI era.When a breakthrough is delayed, the world waits...  ...structured data, and run AI agents and models directly in their workflows. Over 200,000...  ...our work here.You’ll build the datasets, evaluations, and systems that help close that gap.... 
    Work at office
    Local area
    Monday to Friday
    Shift work

    Benchling

    San Francisco, CA
    4 days ago
  • $20 per hour

    SupportFinity™ in Maine is seeking an Editorial Proofreader to join their team focused on training AI models. The position involves evaluating AI chatbot outputs and improving model quality through expert editing and writing skills. This flexible role allows you to work... 
    Remote job
    Hourly pay
    Flexible hours

    SupportFinity™

    Montgomery, AL
    3 days ago
  • $70 - $80 per hour

     ...Role Overview Apply advanced drug safety expertise to help improve AI systems through rigorous, real-world evaluation of pharmacovigilance documentation and data. This remote contract role focuses on the quality, accuracy, and regulatory alignment of complex safety reports... 
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    a month ago
  • $20 - $36 per hour

     ...Role Overview Evaluate generative music AI across a wide range of genres, applying your knowledge of Hungarian music and lyrics to detailed quality standards. You will work in both Hungarian and English to help assess the quality, originality, and naturalness of AI-generated... 
    Hourly pay
    For contractors
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    4 days ago
  • $70 - $90 per hour

     ...Role Overview Help evaluate Neuron Kernel Interface development tasks that support the training and evaluation of advanced AI models. You will assess kernel quality, numerical correctness, CUDA-to-NKI migration fidelity, and whether implementations are well suited to... 
    Hourly pay
    Remote work

    SaidGig

    Remote
    21 days ago
  • $100 per hour

     ...expertise to improve the performance of large language models on finance tasks. You will work with AI researchers to identify model weaknesses in areas...  ...focused on advanced AI systems. Key Responsibilities Evaluate LLM performance in finance areas where models... 
    Hourly pay
    Contract work
    For contractors
    Freelance
    Remote work
    10 hours per week
    Flexible hours

    SaidGig

    United States
    more than 2 months ago
  • $60 per hour

    Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    San Jose, CA
    4 days ago
  • Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant... 
    Remote job
    Flexible hours

    Prolific

    Charlotte, NC
    3 days ago
  • $30 - $35 per hour

    Milpitas, CA Why RoboForce RoboForce is an AI robotics company developing Physical AI-powered Robo-Labor for dull, dirty, and dangerous...  ...real-world deployment and scalability. We are looking for a Model Evaluation Operator- AI Robotics (Contractor) to help evaluate, validate,... 
    Hourly pay
    For contractors
    Monday to Friday
    Shift work
    Afternoon shift

    RoboForce

    Milpitas, CA
    2 days ago
  • SupportFinity™ is looking for an Editorial Proofreader to join our team to train AI models. In this role, you will measure AI chatbot progress, evaluate logic, and solve problems to enhance model quality. Applicants should have a strong command of English and experience... 
    Remote job
    Hourly pay
    Full time
    Part time
    Flexible hours

    SupportFinity™

    Columbia, SC
    4 days ago
  • RoboForce in Milpitas, CA is seeking a Model Evaluation Operator for AI robotics on a contractor basis. You will execute closed‑loop evaluation workflows, manipulate robots, document performance, and help improve AI models through real‑world testing. The role requires... 
    Contract work
    For contractors
    Shift work
    Afternoon shift

    RoboForce

    Milpitas, CA
    2 days ago
  • $20 per hour

    SupportFinity™ is seeking an Editorial Proofreader to evaluate AI models and improve their quality through expert writing and editing skills. This role can be part‑time or full‑time, allowing for a flexible schedule and project selection. Applicants must be fluent in English... 
    Remote job
    Hourly pay
    Full time
    Part time
    Flexible hours

    SupportFinity™

    Sioux Falls, SD
    3 days ago
  • $60 per hour

     ...and Life Science Professionals to join their Expert Network to evaluate AI-generated science. This role allows you to work from home with...  ...offers a competitive pay rate of up to $60 per hour for reviewing model responses, validating technical claims, and critiquing... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    Dallas, TX
    3 days ago
  • $20 per hour

    SupportFinity™ is looking for an Editorial Proofreader to join our team for AI model training. In this remote role, you'll evaluate AI chatbots and enhance model quality. Candidates should have fluency in English and strong editing skills. This position can be full‑time... 
    Remote job
    Hourly pay
    Full time
    Contract work
    Part time

    SupportFinity™

    New York, NY
    3 days ago
  •  ...is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities... 
    Remote job
    Hourly pay
    Flexible hours

    Prolific

    Jacksonville, FL
    3 days ago
  •  ...Role Overview Use your investment and finance expertise to evaluate and improve AI model performance on financial reasoning, valuation, markets, and real-world investment scenarios. Key Responsibilities Assess AI model outputs on valuation, financial modeling, markets... 
    For contractors
    Remote work

    SaidGig

    United States
    19 days ago
  • $60 - $80 per hour

     ...expertise to help develop advanced large language models. In this role, you will bring practical brand, growth, and campaign judgment to AI training data, partnering with research...  ...reasoning quality. Develop and improve evaluation guidelines and scoring rubrics for... 
    Hourly pay
    Weekday work

    SaidGig

    United States
    a month ago
  • $100 - $150 per hour

     ...Overview Apply senior legal judgment to help a leading AI research organization improve how advanced AI models reason about real-world legal work. You will work...  ...define high-quality legal tasks, standards, and evaluations. This role is designed for a practicing legal... 
    Hourly pay
    Full time
    Freelance
    Internship
    Live in
    Relocation
    Relocation package

    SaidGig

    California
    a month ago
  • $70 - $90 per hour

     ...Review GPU and accelerator kernel development tasks that support the training and evaluation of advanced AI models. This role focuses on assessing task quality, numerical correctness, completeness, fair performance benchmarking, appropriate scope, and whether kernels... 
    Hourly pay
    Remote work

    SaidGig

    Remote
    21 days ago
  • $60 per hour

     ...Chemical Engineers to join their Expert Network. Participants will evaluate AI-generated chemistry through tasks that assess factual accuracy...  ...reasoning, enabling cutting-edge advancements in AI models. The position requires a strong educational background in Chemistry... 
    Hourly pay

    Prolific

    Arizona City, AZ
    2 days ago
  • SME Careers is seeking biologists to contribute to an AI training project that involves reviewing AI-generated responses and providing...  ...hold a MS or PhD in a relevant field and have experience in evaluating complex biology content. Strong communication skills and proficient... 
    Immediate start

    SME Careers

    New York, NY
    2 days ago
  • Zebra Technologies is seeking an AI Quality Analyst to ensure performance, safety, and reliability of cutting-edge AI/ML models. Design evaluation strategies, identify edge cases, bias sources, and provide actionable insights to drive model improvements across the development... 

    RXinsider LTD.

    Lincolnshire, IL
    4 days ago
  • $110 per hour

     ...Apply to join a physician talent network supporting AI labs and companies with medical expertise. This is an open application...  ...projects by bringing real-world clinical expertise to model development and evaluation. Key Responsibilities Train and evaluate AI models... 
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    a month ago
  • $60 - $80 per hour

     ...Overview Apply deep insurance expertise to help develop advanced large language models by bringing real-world underwriting, claims, and risk-assessment judgment to AI training and evaluation work. Key Responsibilities Partner with research and engineering teams to... 
    Hourly pay
    Weekday work

    SaidGig

    United States
    more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Mathematician for AI Model Evaluation. Be the first to apply!