Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Mathematics PhD - AI Evaluation Expert

$70 per hour

Mercor

Job Description

Job Description

About the job

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .

Position: Mathematics PhD Coding Experts
Type: Contract
Compensation: $70/hour
Location: Remote
Duration: 6 weeks
Commitment: 20+ hours/week

Role Responsibilities

  • Source material from published papers, Kaggle datasets , open-source repositories, or self-designed scenarios.
  • Write scientific prompts based on sourced inputs.
  • Build grading criteria to define correct answers.
  • Calibrate tasks against frontier models, ensuring tasks ship only when strong models fail more often than succeed.
  • Work independently and asynchronously to meet deadlines while improving AI model performance .

Qualifications

Must-Have

  • PhD in mathematics, applied mathematics, computational mathematics, or a closely related field.
  • Depth in at least two subdomains: numerical linear algebra, computational mechanics, computational finance.
  • Working proficiency in Python for scientific computing.
  • Comfortable with Git/GitHub and running code in Docker .

Preferred

  • Publications in peer-reviewed journals.
  • Prior scientific software or research engineering experience.

Start Date

  • Immediate

Interview Process

  • Upload your resume and application form.
  • A 25-minute conversational interview covering your background, experience, and motivations.
  • Follow up within a few days with next steps and onboarding.

Application Process (Takes 20–30 mins to complete)

  • Upload resume
  • AI interview based on your resume
  • Submit form

Resources & Support

PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Mathematics PhD - AI Evaluation Expert in San Francisco, CA vacancy
  • Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) Mercor is partnering with leading AI labs on a new benchmark for scientific...  ...in at least two subdomains (with a coding focus) Mathematics — numerical linear algebra, computational mechanics... 
    For phd
    Part time
    Immediate start

    Mercor

    San Francisco, CA
    2 days ago
  • Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) for a new benchmark in scientific computing, partnering with leading...  ...hours per week, start date immediate. Required: PhD in mathematics or related field, depth in two subdomains, Python... 
    For phd
    Part time
    Immediate start

    Mercor

    San Francisco, CA
    2 days ago
  • $50 per hour

     ...Remote contract for PhDs in Mathematics, Statistics, or related...  ...edge projects with top AI labs while earning $50+...  ...rigorous logic. Evaluate AI outputs for accuracy...  ...undergraduate to PhD-level math topics. Requirements...  ...: Shortlisted experts complete an evaluation... 
    For phd
    Contract work
    Remote work
    Flexible hours

    Turing

    San Francisco, CA
    5 days ago
  • $80 - $150 per hour

     ...are partnering with a leading AI research organisation to develop...  ...-informed benchmark for evaluating how AI companion chatbots respond...  ...generated conversations and provide expert judgment to calibrate...  ...Required Qualifications MD, DO, PhD, PsyD, or equivalent qualification... 
    For phd
    Hourly pay
    Traineeship
    10 hours per week

    Obsidian

    San Francisco, CA
    5 days ago
  •  ...tasks. You will contribute to creating, evaluating, and refining AI-generated presentations across core...  ...figures, and more. Candidates should hold a PhD with 3+ years of active research,...  ...strong publication record, and demonstrate expert PowerPoint skills, with superb written... 
    For phd
    Remote job

    Ethos

    San Francisco, CA
    2 days ago
  • Computational Statistics and Applied Mathematics Expert About the Project We're...  ...to test how well advanced AI systems can solve hard scientific...  ...-level expertise (MS or PhD required; PhD preferred, or MS...  ...Familiarity with benchmark or evaluation design Background in scientific... 
    For phd
    Remote work

    Obsidian

    San Francisco, CA
    2 days ago
  • $110 per hour

     ...technical talent with leading AI research labs. Headquartered in...  ...Jack Dorsey . Position: Mathematics Research Collaborator (Part-time...  ...Review and evaluate research papers in mathematics...  ...an active research role as a PhD candidate , postdoctoral researcher... 
    For phd
    Remote job
    Contract work
    Part time
    Summer work
    Immediate start

    Mercor

    San Francisco, CA
    15 days ago
  •  ...Job Description Job Title: AI Consulting Domain Remote Job...  ...experienced AI Consulting Domain Experts to contribute their...  ...systems. In this role, you will evaluate, review, and refine AI-generated...  ...degree (Master's, MBA, JD, or PhD) is a plus. ~ Candidates with... 
    For phd
    Remote job
    Part time
    For contractors

    YO AI Labs

    San Francisco, CA
    19 days ago
  • $70 - $90 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Computational Statistics and Applied Mathematics Expert (R, Python, and Matlab/Scilab) Type:...  ...packages. Familiarity with benchmark or evaluation design. Background in scientific... 
    For phd
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    9 days ago
  • $145 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...Jack Dorsey . Position: ML Research PhD Experts (ICML / NeurIPS / ICLR Publications) Type...  ...Location: Remote Role Responsibilities Evaluate the accuracy and depth of AI-generated... 
    For phd
    Remote job
    Contract work
    Summer work

    Mercor

    San Francisco, CA
    17 days ago
  • $75 - $100 per hour

    Join to apply for the Math PhD - Expert Trainer role at Handshake 1 week...  ...in Math to join our AI research community. This program...  ...with real-world expertise and evaluate where they excel or fail Get...  ...by 2x Get notified about new Mathematics Specialist jobs in San Francisco... 
    For phd
    Contract work
    Part time
    Summer work
    Freelance
    H1b
    Remote work
    Visa sponsorship
    10 hours per week
    Flexible hours

    Handshake

    San Francisco, CA
    3 days ago
  • $164.5k - $219k

     ...Information Job Title Expert Senior Manager, Data...  ...to work with major AI ecosystem partners through...  ...to prompt design, evaluation and output quality assuranceDevelop...  ...to explain and discuss mathematical and machine learning...  ..., or PhD in a technical fieldBackground... 
    For phd
    Permanent employment
    Full time
    Apprenticeship
    Local area
    Home office
    3 days per week

    Bain & Company

    San Francisco, CA
    1 day ago
  • $70 - $100 per hour

     ...technical talent with leading AI research labs. Headquartered...  ...Bayesian Statistics and Applied Mathematics Expert Type: Contract...  ...relevant STEM field ( MS, PhD , or equivalent research experience...  ...Familiarity with benchmark or evaluation design. Background in... 
    For phd
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    27 days ago
  • $80 - $150 per hour

     ...leading behavioral health consulting firm is seeking Senior Behavioral Health Experts to work part-time and remotely on frontier AI research projects. You will be responsible for designing evaluations and testing AI systems in critical mental health contexts. The ideal... 
    Hourly pay
    Part time
    Remote work

    Aligned Labs

    San Francisco, CA
    4 days ago
  • Mercor is seeking experienced musicians to evaluate generative music AI models. You will compare AI-generated lyrics with published songs across genres and rate them against detailed quality standards in Greek and English. The role requires native or near-native Greek,... 
    Remote work
    Flexible hours

    Obsidian

    San Francisco, CA
    18 hours ago
  • $60 per hour

    Prolific seeks Chemistry Experts and Chemical Engineers to join our Expert Network and evaluate AI models using your chemical expertise. Successful candidates will be invited to assess and review AI-generated chemistry tasks and ensure their accuracy. Compensation can reach... 
    Remote job
    Hourly pay
    Flexible hours

    Prolific

    San Francisco, CA
    2 days ago
  • YO AI Labs in the United States (Remote) seeks seasoned Adobe Marketing Technology Experts to support an AI training and evaluation project focused on enterprise marketing operations. You will test workflows with Adobe Workfront, AEM, CJA, Analytics and Experience Cloud... 
    Remote job

    YO AI Labs

    San Francisco, CA
    3 days ago
  • $80 - $120 per hour

    A leading AI consulting firm is searching for a Senior Mathematics Expert to join their remote team. The ideal candidate will possess a PhD in Mathematics and have a rich experience in solving complex...  ...AI models can't solve and evaluating these models. This position offers... 
    For phd
    Remote job

    Aligned Labs

    San Francisco, CA
    4 days ago
  • A leading AI evaluation firm based in San Francisco seeks a Machine Learning Scientist to foster understanding of AI model performance. You'...  ...while collaborating across teams. Applicants should possess a PhD in a relevant field and hands-on experience with large-scale models... 
    For phd

    Arena Intelligence, Inc.

    San Francisco, CA
    5 days ago
  • $170k - $216k

     ...high-scale, mission-critical automation and evaluation frameworks that establish the "ultimate...  ...experience ~2+ years of experience in industrial AI applications involving the creation,...  ...efficient code We Prefer MS or PhD in Computer Science, Robotics, similar technical... 
    For phd
    Full time
    Remote work

    Waymo

    San Francisco, CA
    1 day ago
  • Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback. Candidates should have at least 5 years of relevant experience... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    San Francisco, CA
    3 days ago
  • $50 per hour

    A leading AI research firm is seeking PhDs in Mathematics or related fields for a fully remote contract role. The successful candidate will design advanced math problems to test AI performance and evaluate outputs for accuracy. Strong mathematical reasoning, problem-solving... 
    For phd
    Remote job
    Contract work
    Flexible hours

    Turing

    San Francisco, CA
    4 days ago
  • Cincinnatus LLC is seeking a Marketing SME to join a leading GenAI team focused on evaluating AI outputs against rubrics and strengthening brand strategy in AI training data. We require 8+ years of marketing experience with top-tier brands, plus hands-on evaluation of LLM... 
    Weekday work

    Mercor

    San Francisco, CA
    2 days ago
  • Obsidian is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across complex policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured... 

    Obsidian

    San Francisco, CA
    5 days ago
  •  ...looking for a talented researcher to work on improving AI tutoring systems. You will shape experiments, evaluate tutoring quality, and turn pedagogical principles...  ...potential for continuation. Ideal candidates have a PhD or equivalent experience in CS, ML, or related fields... 
    For phd
    Contract work
    Summer work

    Heyaristotle

    San Francisco, CA
    4 days ago
  • Obsidian is seeking expert Evaluators in Biology/environmental science to review and assess AI-generated work products for accuracy and quality. In this remote, hourly role, you will leverage your expertise to provide feedback on documents and presentations, ensuring they... 
    Remote job
    Hourly pay

    Obsidian

    San Francisco, CA
    4 days ago
  • $213k - $263k

     ...and state-of-the-art Generative AI to create a training ground...  ...the Waymo Driver. The Simulator Evaluation team faces the ultimate data challenge: How do you mathematically prove that a virtual world is...  ...Bachelor's, Master's, or PhD in computer science, machine learning... 
    For phd
    Full time
    Remote work

    Waymo

    San Francisco, CA
    1 day ago
  • $50 per hour

    A leading AI research accelerator is seeking remote PhD candidates in Mathematics or related fields to design math problems and evaluate AI performance. The role involves collaboration with researchers and offers flexible hours at a pay rate of $50+/hour. Ideal candidates... 
    For phd
    Remote job
    Hourly pay
    Flexible hours

    Turing

    San Francisco, CA
    3 days ago
  • $160k - $250k

     ...What Granica does Granica is an AI research and systems company building...  ...new model architectures, evaluate on live datasets, and publish results...  ...efficient learning. What you’ll bring PhD in Machine Learning, Statistics, Applied Mathematics, or a related field with... 
    For phd
    Flexible hours

    Granica

    San Francisco, CA
    4 days ago
  • Mercor is seeking a remote Physics PhD to tackle complex physics problems and ensure the accuracy of AI-generated solutions. The role involves independently solving physics challenges and collaborating with AI research teams. Ideal candidates will have a PhD in Physics... 
    For phd
    Remote job
    Immediate start

    Mercor

    San Francisco, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Mathematics PhD - AI Evaluation Expert. Be the first to apply!