Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Theoretical Physicist - AI Reasoning Evaluator

Obsidian

Obsidian is collaborating with a leading AI research group to engage experts with advanced training in physics. The role involves solving complex physics problems and reviewing AI-generated proofs. Ideal candidates should hold a PhD in Physics from a top program and possess strong problem-solving skills. This is a project-based opportunity, requiring approximately 15 hours of commitment per week. #J-18808-Ljbffr Obsidian

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Theoretical Physicist - AI Reasoning Evaluator in San Francisco, CA vacancy
  •  ...accomplished biology and biophysics researchers to contribute to AI systems for scientific reasoning. This position is placed at a leading AI lab as part of...  ...workforce in San Francisco. You will review and evaluate papers for rigor, author and review challenging problems... 
    Suggested
    Part time

    Mercor

    San Francisco, CA
    5 days ago
  • $70 - $180 per hour

     ...a Physics PhD to solve challenging physics problems and review AI-generated proofs for clarity and correctness. This remote position...  ...and a strong ability in problem-solving and articulation of reasoning. Applicants can apply immediately and the application process is... 
    Suggested
    Remote job
    Hourly pay
    Immediate start

    Mercor

    San Francisco, CA
    1 day ago
  • Synthires is seeking a PhD-level expert to contribute to advanced AI research and evaluation projects in San Francisco. The role centers on applying...  ...and craft high-quality reference solutions that push the reasoning abilities of next-generation systems. Ideal candidates... 
    Suggested
    Part time

    Synthires

    San Francisco, CA
    3 days ago
  • Tabula is an AI-first therapeutics company focusing on phage-bacteria interactions. We are hiring a Research Scientist to build computational...  ...with wet lab scientists. The role emphasizes quantitative reasoning, rigorous experimentation, and a fast feedback loop from model... 
    Suggested

    Pantera Capital

    San Francisco, CA
    1 day ago
  •  ...building a Large Physics foundation Model to reason with physical laws and predict and alter...  ...experts to advance physics-informed AI across fluid dynamics, thermodynamics, and...  ...principles into engineering requirements, design evaluations for physical coherence, and help the... 
    Suggested

    Causal Labs

    San Francisco, CA
    2 days ago
  • $50 - $75 per hour

    A leading tech company based in Australia is seeking an AI Model Evaluator on a contract basis. The role involves evaluating AI-generated responses...  ...or Accounting, possess strong communication and critical reasoning skills, and have relevant professional experience. This... 
    Hourly pay
    Contract work

    Mercor

    San Francisco, CA
    4 days ago
  • Mercor is seeking experienced Clinical Law Professors and Clinic Directors to evaluate AI-generated legal reasoning in civil legal services. Experts will blend doctrinal knowledge with practical supervision to assess AI analyses. Applicants should hold a JD with active... 
    Remote job

    Mercor

    San Francisco, CA
    2 days ago
  • $80 - $120 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Jack Dorsey . Position: Data analysis / quantitative readouts Evaluator Type: Contract Compensation: $80–$120/hour... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    San Francisco, CA
    4 days ago
  • $80 - $120 per hour

    Mercor, based in San Francisco, is seeking a Cybersecurity / IT GRC Evaluator to evaluate AI-generated artifacts and provide structured feedback. Candidates should have over 5 years of relevant experience and fluency in English. This remote position offers a competitive... 
    Remote job
    Hourly pay
    Work at office

    Mercor Inc

    San Francisco, CA
    5 days ago
  • Obsidian is looking for expert Evaluators in Finance operations/audit support to review AI-generated work products for accuracy and quality. This remote hourly position requires a minimum of 5 years in finance and fluency in English. Your role will involve evaluating outputs... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    San Francisco, CA
    2 days ago
  • $80 - $120 per hour

    Mercor is seeking a User/Customer Research and Feedback Synthesis Evaluator to evaluate AI-generated artifacts using specific quality rubrics. The role demands deep subject-matter expertise in user research and involves providing feedback to improve AI performance. The... 
    Remote job
    Contract work
    Work at office
    Flexible hours

    Mercor

    San Francisco, CA
    2 days ago
  • Obsidian is hiring expert Evaluators in Investment analysis / valuation / credit to review AI-generated work products for accuracy and quality. This remote, hourly position requires deep subject-matter expertise and professional fluency in English to provide structured... 
    Hourly pay
    Work at office
    Remote work

    Obsidian

    San Francisco, CA
    2 days ago
  • Obsidian is seeking a Spanish Audio Generalist Evaluator Expert to contribute to a high-impact audio AI research project. You will handle transcription, annotation, and evaluation tasks to help train and benchmark advanced language models. The ideal candidate should have... 
    Part time
    10 hours per week

    Obsidian

    San Francisco, CA
    3 days ago
  • Welo Data is seeking Data Labeling Associates in California to evaluate AI outputs and ensure cultural context and safety in Arabic datasets. This role requires professional-level proficiency in Portuguese (Brazil), a bachelor's degree, and at least 2 years of experience... 

    Welo Data

    San Francisco, CA
    4 days ago
  • Synthires is offering a part-time role for PhD-level Chemistry experts to contribute to AI safety and evaluation projects. The work involves applying scientific expertise to understand and improve how AI systems handle specialized chemistry topics, with training provided... 
    Remote job
    Part time

    Synthires

    San Francisco, CA
    19 hours ago
  • Obsidian is hiring expert Evaluators in Real estate, hospitality, and events to review AI-generated work for accuracy, rigor, and domain quality. This remote position requires deep expertise and involves grading outputs like documents and presentations. Applicants must... 
    Remote job
    Work at office

    Obsidian

    San Francisco, CA
    2 days ago
  •  ...hiring experienced music producers and audio engineers to evaluate generative music AI models, in partnership with a leading AI lab. You will...  ...qualified applicants without regard to protected characteristics and provide reasonable accommodations #J-18808-Ljbffr Obsidian
    Immediate start

    Obsidian

    San Francisco, CA
    2 days ago
  • $80 - $120 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Dorsey . Position: Incident management / reliability / SRE Evaluator Type: Contract Compensation: $80–$120/hour Location... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    San Francisco, CA
    6 days ago
  • Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback. Candidates should have at least 5 years of relevant experience... 
    Remote job
    Hourly pay
    Work at office

    Obsidian

    San Francisco, CA
    2 days ago
  • Welo Data in San Francisco is hiring Data Labeling Associates for Project Perseus. This role focuses on evaluating Arabic AI systems, requiring professional proficiency in Portuguese and experience in AI safety. Responsibilities include assessing AI outputs, identifying... 
    Full time

    Welo Data

    San Francisco, CA
    3 days ago
  • Mercor is hiring experienced music professionals to evaluate generative music AI models, in partnership with a leading AI lab. You will assess AI-generated music across a wide range of genres and rate it against detailed quality standards, working in Thai and English.... 
    Flexible hours

    Mercor

    San Francisco, CA
    3 days ago
  • Mercor is hiring experienced musicians to evaluate generative musical AI models in partnership with a leading AI lab. You will assess model outputs across lyrics, voice generation, and other standards, using your bilingual language skills. Ideal candidates have 3+ years... 
    Part time
    Immediate start
    10 hours per week

    Mercor Inc

    San Francisco, CA
    5 days ago
  • Mercor is seeking experienced musicians to evaluate generative musical AI models in collaboration with a leading AI lab. You will assess model outputs across different categories of music in your bilingual language and contribute to structured taxonomy annotations. Ideal... 
    Part time
    Immediate start
    10 hours per week

    Mercor

    San Francisco, CA
    5 days ago
  • $85 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    a month ago
  • Welo Data is seeking Data Labeling Associates in San Francisco to evaluate AI systems focused on Arabic language nuances. The role includes model evaluation, identifying biases in datasets, and ensuring quality across global AI workflows. Candidates should have native French... 

    Welo Data

    San Francisco, CA
    1 day ago
  • Causal is building a Large Physics foundation Model to predict how physical systems evolve, using multimodal sensor data and simulations to learn verifiable cause and effect. You will design architectures and training recipes to turn heterogeneous observations into accurate...

    Causal

    San Francisco, CA
    5 days ago
  •  ...define grading criteria for pre‑sales deliverables and to score AI‑generated and human work samples with detailed written justifications...  ...across technical discovery plans, demos, proofs of concept, and evaluation plans. The role emphasizes clear written communication,... 

    Obsidian

    San Francisco, CA
    2 days ago
  • About the roleWe're building a high-quality evaluation dataset for CNC manufacturing and are looking for experienced CNC machinists to...  ...approved or rejected for a production run, and clearly explain your reasoning. What you'll do Contribute to text-only, objective, verifiable... 

    Obsidian

    San Francisco, CA
    2 days ago
  • $120 - $175 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...Review sample solutions and explain pass/fail outcomes with detailed reasoning. Apply real production-floor judgment to determine the... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    4 days ago
  • Mercor is hiring experienced music producers and audio engineers to evaluate generative music AI models, partnering with a leading AI lab. You will analyze AI-generated music across genres and rate it against detailed quality standards, working in Telugu and English. Responsibilities... 

    Mercor

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Theoretical Physicist - AI Reasoning Evaluator. Be the first to apply!