Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Politics Domain Expert for AI Evaluation & Benchmarks

Turing Global India

Turing Global India is seeking a Politics Domain Expert to evaluate and improve Large Language Models. You will design challenging prompts across political science, governance, elections, public policy and international relations, and assess factual accuracy and reasoning. The role requires 3+ years in politics or public policy, strong English writing, and keen attention to detail. This is a contractor assignment for 8 weeks, 40 hours per week with PST overlap. #J-18808-Ljbffr Turing Global India

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Politics Domain Expert for AI Evaluation & Benchmarks in New York, NY vacancy
  • $80 - $120 per hour

     ...technical talent with leading AI research labs. Headquartered in...  ..., our investors include Benchmark , General Catalyst , Peter...  ...Biology / environmental science Evaluator Type: Contract Compensation...  ...-generated artifacts against domain-specific quality rubrics.... 
    Suggested
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    New York, NY
    8 days ago
  •  ...are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and...  ...involving misinformation, political persuasion, self-harm, violence...  ..., and other sensitive domains. Apply and refine...  ...RLHF, SFT, and AI safety benchmarking. Identify unsafe outputs... 
    Suggested
    Worldwide

    Obsidian

    New York, NY
    1 day ago
  • Mercor is seeking a Music Audio Expert - German for a remote, project-based engagement. You will evaluate AI output lyrics and voice generation, score training data quality, and write music in the domain language listed in the title. Applicants should have 3+ years as music... 
    Suggested
    Remote job
    10 hours per week

    United States Digital Space LLC

    New York, NY
    1 day ago
  •  ...insurance, compliance, and legal professionals to join our on-call AI Evaluation Specialists network. This remote, project-based role involves...  ...thinking, and providing high-quality feedback across multiple domains. Start date is July 13, 2026, with 0-20 hours per week and... 
    Suggested
    Remote job
    Flexible hours

    Dorado

    New York, NY
    3 days ago
  • 1. Overview Join a leading AI lab's cutting-edge GenAI team and...  ...talented Retail subject-matter experts (SMEs) with deep domain expertise and hands-on experience evaluating AI model outputs against...  ...disability, genetic information, political views or activity, or any other... 
    Suggested
    Contract work
    Weekday work

    Obsidian

    New York, NY
    5 days ago
  • $8 - $65 per hour

    Prolific is hiring Mental Health Professionals in New York to train and evaluate AI models. As a Domain Expert, you will be responsible for reviewing AI-generated responses, completing tasks related to psychology, and improving AI models based on your expertise. Pay rates... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    New York, NY
    1 day ago
  •  ...expertise. The role focuses on authoring and verifying challenging chemistry questions and establishing high-quality evaluation standards for AI benchmarks. Ideal candidates hold a PhD or are pursuing one, with strong knowledge of reaction mechanisms, materials chemistry... 
    Remote job
    Part time

    24-Mag Llc

    New York, NY
    2 days ago
  • $60 - $80 per hour

     ...technical talent with leading AI research labs....  ...our investors include Benchmark , General Catalyst ,...  ...Design challenging, domain-relevant retail tasks and...  ...real retail practice. Evaluate AI model outputs against...  ...other subject matter experts to ensure consistency and... 
    Contract work
    Summer work
    Remote work
    Weekday work

    Mercor

    New York, NY
    17 days ago
  •  ...is seeking advanced Mathematics and Statistics experts to support the training and evaluation of state-of-the-art AI systems. We need subject-matter experts who...  ...problems in your area of expertise, contribute to benchmarking efforts, and help push the #J-18808-Ljbffr... 
    Remote job
    Contract work
    Immediate start
    Flexible hours

    BAM Ventures

    New York, NY
    5 days ago
  • $50 per hour

    A leading AI research accelerator is looking for remote PhD candidates in Chemistry, Chemical Engineering...  .... You will design advanced chemistry problems to evaluate AI performance and collaborate with researchers on benchmarks. This role offers flexible hours and a pay rate... 
    Remote job
    Hourly pay
    Flexible hours

    Turing

    New York, NY
    5 days ago
  • $80 - $120 per hour

     ...technical talent with leading AI research labs. Headquartered in...  ..., our investors include Benchmark , General Catalyst , Peter...  ...Position: Education / school Evaluator Type: Contract Compensation...  ...-generated artifacts against domain-specific quality rubrics.... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    New York, NY
    19 days ago
  • $80 - $120 per hour

     ...technical talent with leading AI research labs. Headquartered in...  ..., our investors include Benchmark , General Catalyst , Peter...  ...Position: Software / AI / IT / data Evaluator Type: Contract...  ...-generated artifacts against domain-specific quality rubrics. Identify... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    New York, NY
    17 days ago
  • $80 - $120 per hour

     ...technical talent with leading AI research labs. Headquartered in...  ..., our investors include Benchmark , General Catalyst , Peter...  ...Position: Education / school Evaluator Type: Contract Compensation...  ...-generated artifacts against domain-specific quality rubrics.... 
    Contract work
    Summer work
    Work at office
    Remote work

    Mercor

    New York, NY
    13 days ago
  • $80 - $150 per hour

    Prolific in New York, NY, is seeking Medical Doctors to join our Expert Network for training and evaluating AI models. As a Domain Expert participant, you will review and rate AI-generated clinical responses, providing crucial feedback for AI development. You must hold... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    New York, NY
    2 days ago
  • $60 - $70 per hour

     ...technical talent with leading AI research labs....  ...our investors include Benchmark , General Catalyst ,...  ...Role Responsibilities Evaluate AI-generated responses...  ...involving misinformation, political persuasion, self-harm,...  ..., and other sensitive domains. Apply and refine... 
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    19 days ago
  • $100 - $150 per hour

     ...technical talent with leading AI research labs....  ...our investors include Benchmark , General Catalyst ,...  ...Cybersecurity Labeling Expert Type: Contract...  ...intent and harm across domains like scaled data exfiltration...  ...and exploits . Evaluate POC exploit... 
    Hourly pay
    Weekly pay
    Full time
    Contract work
    For contractors
    Summer work
    Remote work

    Mercor

    New York, NY
    1 day ago
  • $30 per hour

    About Prolific Prolific is not just another player in the AI space - we are building the biggest pool of quality...  ...seeking Professional Graphic and Visual Designers to act as Domain Experts for a high-level AI evaluation project. AI models are evolving beyond simple image... 
    Remote job
    Work from home
    Flexible hours

    Prolific

    New York, NY
    5 days ago
  • YO AI Labs seeks experienced AI Software Engineering Domain Experts to evaluate and refine AI-generated software engineering content, enhancing accuracy and reasoning of AI models. You will review content and develop prompts to steer outputs, collaborating with remote... 
    Remote job
    Part time

    YO AI Labs

    New York, NY
    2 days ago
  • $30 per hour

    Prolific is seeking fluent Hindi speakers to join their Expert Network, helping to train and evaluate AI models with real legal expertise. Responsibilities include analyzing and writing tasks in Hindi, judging AI’s performance, and aiding in the improvement of AI models... 
    Remote job
    Hourly pay
    Flexible hours

    Prolific

    New York, NY
    5 days ago
  • Join a leading AI lab's cutting-edge GenAI team...  ...insurance and actuarial domain expert to work directly with...  ...like, and build the benchmarks that show whether the...  ...challenging insurance tasks and evaluation sets, and help build...  ...genetic information, political views or activity, or... 
    Full time
    Contract work
    Part time
    Freelance
    Live in
    Relocation
    Relocation package

    Obsidian

    New York, NY
    2 days ago
  • A tech company focusing on AI research is looking for experienced Krita users for a flexible, project-based contract opportunity. This role allows you to earn while evaluating AI-generated content related to digital painting and concept art. Candidates should have at least... 
    Contract work
    Remote work
    Flexible hours

    Handshake

    New York, NY
    3 days ago
  •  ...are seeking experienced AI Safety Red Teamers to identify...  ...model weaknesses, and evaluate AI behavior across...  ...cyber, biosecurity, fraud, political content, and other sensitive domains. Document...  ...and contribute to safety benchmarking and red-teaming reports.... 

    Obsidian

    New York, NY
    2 days ago
  • A leading research services provider is seeking a Chemistry Expert (PhD) to evaluate complex chemistry problems and review AI-generated outputs for accuracy. This remote, hourly contract role requires deep subject-matter expertise and excellent communication skills. Candidates... 
    Remote job
    Hourly pay
    Contract work
    Flexible hours

    Crossing Hurdles

    New York, NY
    4 days ago
  • $73 per hour

    A leading AI innovation firm is seeking individuals for a flexible, project-based role focusing on creating complex tasks for evaluating AI performance. Ideal candidates should possess a postgraduate degree and relevant industry experience, especially in economics. This... 
    Remote work
    Flexible hours

    Mindrift

    New York, NY
    5 days ago
  • Prolific in New York, NY, is seeking Chemistry Experts and Chemical Engineers to join its Expert Network. In this role, you will help train and evaluate AI models using your chemical expertise. Duties include evaluating AI-generated responses for accuracy and validating... 
    Remote job
    Work from home
    Flexible hours

    Prolific

    New York, NY
    1 day ago
  • Mercor is seeking experienced musicians to evaluate generative music AI models, collaborating with a leading AI lab. You will assess AI-generated music and lyrics in Telugu and English, applying detailed quality standards across genres. Responsibilities include comparing... 
    Part time
    Immediate start

    Obsidian

    New York, NY
    5 days ago
  • Obsidian is seeking a Spanish (Mexico) Audio Generalist Evaluator Expert for a high-impact audio AI research project. The role focuses on transcription, annotation, and evaluation tasks to help train advanced language models. Candidates need to be fluent in Spanish (Mexico... 
    Temporary work
    10 hours per week

    Obsidian

    New York, NY
    2 days ago
  • $120 per hour

    Prolific is seeking Licensed Pharmacists to evaluate and train AI models related to healthcare. You will earn competitive pay rates (up to $120 per hour) while conducting assessments for tasks requiring focused attention on complex medical data. Ideal candidates have a... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    New York, NY
    4 days ago
  • A leading AI research firm is seeking Expert Prompt Curators to design challenging prompts for evaluating advanced AI models. The role requires advanced knowledge in diverse fields and offers flexible hours, remote work, and a competitive hourly wage. Ideal candidates will... 
    Remote job
    Hourly pay
    Temporary work
    Flexible hours

    CloudDevs

    New York, NY
    4 days ago
  •  ...hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) Mercor is partnering with leading AI labs on a new benchmark for scientific computing. You will author...  ...today's frontier models cannot solve. Domains — depth required in at least two subdomains... 
    Part time
    Immediate start

    Obsidian

    New York, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Politics Domain Expert for AI Evaluation & Benchmarks. Be the first to apply!