Theoretical Physicist - AI Reasoning Evaluator
Obsidian
Obsidian is collaborating with a leading AI research group to engage experts with advanced training in physics. The role involves solving complex physics problems and reviewing AI-generated proofs. Ideal candidates should hold a PhD in Physics from a top program and possess strong problem-solving skills. This is a project-based opportunity, requiring approximately 15 hours of commitment per week. #J-18808-Ljbffr Obsidian
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Theoretical Physicist - AI Reasoning Evaluator in San Francisco, CA vacancy
- ...accomplished biology and biophysics researchers to contribute to AI systems for scientific reasoning. This position is placed at a leading AI lab as part of... ...workforce in San Francisco. You will review and evaluate papers for rigor, author and review challenging problems...SuggestedPart time
$70 - $180 per hour
...a Physics PhD to solve challenging physics problems and review AI-generated proofs for clarity and correctness. This remote position... ...and a strong ability in problem-solving and articulation of reasoning. Applicants can apply immediately and the application process is...SuggestedRemote jobHourly payImmediate start- Synthires is seeking a PhD-level expert to contribute to advanced AI research and evaluation projects in San Francisco. The role centers on applying... ...and craft high-quality reference solutions that push the reasoning abilities of next-generation systems. Ideal candidates...SuggestedPart time
- ...building a Large Physics foundation Model to reason with physical laws and predict and alter... ...experts to advance physics-informed AI across fluid dynamics, thermodynamics, and... ...principles into engineering requirements, design evaluations for physical coherence, and help the...Suggested
- Tabula is an AI-first therapeutics company focusing on phage-bacteria interactions. We are hiring a Research Scientist to build computational... ...with wet lab scientists. The role emphasizes quantitative reasoning, rigorous experimentation, and a fast feedback loop from model...Suggested
$50 - $75 per hour
A leading tech company based in Australia is seeking an AI Model Evaluator on a contract basis. The role involves evaluating AI-generated responses... ...or Accounting, possess strong communication and critical reasoning skills, and have relevant professional experience. This...Hourly payContract work- Mercor is seeking experienced Clinical Law Professors and Clinic Directors to evaluate AI-generated legal reasoning in civil legal services. Experts will blend doctrinal knowledge with practical supervision to assess AI analyses. Applicants should hold a JD with active...Remote job
- Synthires is offering a part-time role for PhD-level Chemistry experts to contribute to AI safety and evaluation projects. The work involves applying scientific expertise to understand and improve how AI systems handle specialized chemistry topics, with training provided...Remote jobPart time
$80 - $120 per hour
Mercor, based in San Francisco, is seeking a Cybersecurity / IT GRC Evaluator to evaluate AI-generated artifacts and provide structured feedback. Candidates should have over 5 years of relevant experience and fluency in English. This remote position offers a competitive...Remote jobHourly payWork at office- Obsidian is looking for expert Evaluators in Finance operations/audit support to review AI-generated work products for accuracy and quality. This remote hourly position requires a minimum of 5 years in finance and fluency in English. Your role will involve evaluating outputs...Remote jobHourly payWork at office
$80 - $120 per hour
Mercor is seeking a User/Customer Research and Feedback Synthesis Evaluator to evaluate AI-generated artifacts using specific quality rubrics. The role demands deep subject-matter expertise in user research and involves providing feedback to improve AI performance. The...Remote jobContract workWork at officeFlexible hours- Obsidian is seeking a Spanish Audio Generalist Evaluator Expert to contribute to a high-impact audio AI research project. You will handle transcription, annotation, and evaluation tasks to help train and benchmark advanced language models. The ideal candidate should have...Part time10 hours per week
- Welo Data is seeking Data Labeling Associates in California to evaluate AI outputs and ensure cultural context and safety in Arabic datasets. This role requires professional-level proficiency in Portuguese (Brazil), a bachelor's degree, and at least 2 years of experience...
- Obsidian is hiring expert Evaluators in Investment analysis / valuation / credit to review AI-generated work products for accuracy and quality. This remote, hourly position requires deep subject-matter expertise and professional fluency in English to provide structured...Hourly payWork at officeRemote work
- Obsidian is hiring expert Evaluators in Real estate, hospitality, and events to review AI-generated work for accuracy, rigor, and domain quality. This remote position requires deep expertise and involves grading outputs like documents and presentations. Applicants must...Remote jobWork at office
- ...hiring experienced music producers and audio engineers to evaluate generative music AI models, in partnership with a leading AI lab. You will... ...qualified applicants without regard to protected characteristics and provide reasonable accommodations #J-18808-Ljbffr ObsidianImmediate start
- Welo Data in San Francisco is hiring Data Labeling Associates for Project Perseus. This role focuses on evaluating Arabic AI systems, requiring professional proficiency in Portuguese and experience in AI safety. Responsibilities include assessing AI outputs, identifying...Full time
- Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback. Candidates should have at least 5 years of relevant experience...Remote jobHourly payWork at office
- Mercor is seeking experienced musicians to evaluate generative musical AI models in collaboration with a leading AI lab. You will assess model outputs across different categories of music in your bilingual language and contribute to structured taxonomy annotations. Ideal...Part timeImmediate start10 hours per week
- Mercor is hiring experienced music professionals to evaluate generative music AI models, in partnership with a leading AI lab. You will assess AI-generated music across a wide range of genres and rate it against detailed quality standards, working in Thai and English....Flexible hours
- Mercor is hiring experienced musicians to evaluate generative musical AI models in partnership with a leading AI lab. You will assess model outputs across lyrics, voice generation, and other standards, using your bilingual language skills. Ideal candidates have 3+ years...Part timeImmediate start10 hours per week
- Welo Data is seeking Data Labeling Associates in San Francisco to evaluate AI systems focused on Arabic language nuances. The role includes model evaluation, identifying biases in datasets, and ensuring quality across global AI workflows. Candidates should have native French...
- Causal is building a Large Physics foundation Model to predict how physical systems evolve, using multimodal sensor data and simulations to learn verifiable cause and effect. You will design architectures and training recipes to turn heterogeneous observations into accurate...
- ...define grading criteria for pre‑sales deliverables and to score AI‑generated and human work samples with detailed written justifications... ...across technical discovery plans, demos, proofs of concept, and evaluation plans. The role emphasizes clear written communication,...
$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Data analysis / quantitative readouts Evaluator Type: Contract Compensation: $80–$120/hour...Contract workSummer workWork at officeRemote work- About the roleWe're building a high-quality evaluation dataset for CNC manufacturing and are looking for experienced CNC machinists to... ...approved or rejected for a production run, and clearly explain your reasoning. What you'll do Contribute to text-only, objective, verifiable...
$120 - $175 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Review sample solutions and explain pass/fail outcomes with detailed reasoning. Apply real production-floor judgment to determine the...Contract workSummer workRemote work- Mercor is hiring experienced music producers and audio engineers to evaluate generative music AI models, partnering with a leading AI lab. You will analyze AI-generated music across genres and rate it against detailed quality standards, working in Telugu and English. Responsibilities...
- Mercor is hiring experienced musicians to evaluate generative musical AI models in partnership with a leading AI lab. You will assess model outputs across different categories of music in your bilingual language. Ideal candidates have 3+ years in music production or audio...Part timeImmediate start10 hours per week
$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Dorsey . Position: Incident management / reliability / SRE Evaluator Type: Contract Compensation: $80–$120/hour Location...Contract workSummer workWork at officeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Theoretical Physicist - AI Reasoning Evaluator. Be the first to apply!
Related searches
- health physicist San Francisco, CA
- medical physicist San Francisco, CA
- petrophysicist San Francisco, CA
- engineering physicist San Francisco, CA
- nuclear physicist San Francisco, CA
- physicist San Francisco, CA
- work from home web search evaluator San Francisco, CA
- evaluator San Francisco, CA
- ai evaluator San Francisco, CA
- quality evaluator San Francisco, CA

