Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Evaluation Specialist - Movies & TV Domain

eDataBae

About the Opportunity

We are looking for experienced entertainment professionals and subject matter experts to support AI evaluation initiatives focused on Movies and TV content.

The role involves creating challenging evaluation scenarios, reviewing AI-generated responses, and providing expert feedback to improve the accuracy and quality of AI systems.

Key Responsibilities

• Create domain-specific prompts related to movies, television, streaming, and entertainment topics
• Evaluate AI-generated responses for accuracy, reasoning, completeness, and relevance
• Identify factual errors, inconsistencies, outdated information, and knowledge gaps
• Compare multiple AI responses and provide structured evaluations
• Provide evidence-based feedback using reliable references
• Support the development of evaluation datasets and testing scenarios

Required Skills & Experience

• Strong knowledge of movies, television, streaming platforms, actors, directors, genres, awards, and entertainment history
• Excellent written English and analytical skills
• Strong attention to detail and ability to assess factual accuracy
• Ability to research, analyze, and provide objective feedback
• Experience creating written content, research material, or evaluations preferred

Preferred Qualifications

• Master's degree or advanced academic background related to film, media studies, entertainment, or related fields
• Professional, research, teaching, or industry experience in movies/TV domain
• Experience with Generative AI, LLMs, prompt engineering, or AI evaluation is a plus
• Experience working independently on research-based assignments

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the AI Evaluation Specialist - Movies & TV Domain in United States vacancy
  •  ...professionals with strong expertise in the video game industry to support AI evaluation initiatives. The role involves creating gaming-related...  ...of AI systems. Key Responsibilities • Create domain-specific prompts related to video games, game development, esports... 
    Suggested

    eDataBae

    United States
    5 days ago
  •  ...AI Evaluation Specialist Role Type: Contractor Location: Remote (US, CA, UK, IE, AU, NZ) Micro1 is engaging AI Evaluation Specialists...  ...-world input. No prior experience in AI is required — your domain knowledge is what matters. Scope of Work # Evaluate... 
    Suggested
    For contractors
    Remote work

    micro1

    United States
    1 day ago
  •  ...Domain Architect- AI/ML, Senior Specialist Apply ( locations Malvern, PA Philadelphia, PA Charlotte, NC time type Full time posted...  .... Establish best practices for prompt engineering, evaluation, testing, observability, model lifecycle management,... 
    Suggested
    Full time

    Vanguard

    Philadelphia, PA
    5 days ago
  • $150k - $250k

     ...About Distyl AI Distyl is an applied AI technology company partnering with the world...  ...At Distyl, we build AI systems using Evaluation-Driven Development —an approach where evaluation...  ...how system quality is measured in each domain, ensuring that evaluation signals reflect... 
    Suggested
    Full time
    Work at office
    Flexible hours
    3 days per week

    Distyl Ai

    Remote
    17 hours ago
  • $35 - $120 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San...  ...quality through structured review. Evaluate the accuracy and depth of AI-generated content...  .... Live Code Review Session . Domain Expert Interview . You're paid $200... 
    Suggested
    Full time
    Contract work
    Summer work
    Remote work

    Mercor

    Remote
    17 hours ago
  •  ...A leading AI organization in Australia is seeking individuals with strong writing and analytical skills to evaluate and improve AI outputs. The ideal candidate must possess the ability to assess emotional nuances and detail while adhering to structured guidelines. Responsibilities... 
    Immediate start

    Crossing Hurdles

    New York, NY
    5 days ago
  • $120 per hour

     ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors...  ...hour Location: Remote Role Responsibilities Evaluate complex technical tasks using deep language expertise in Scala... 
    Hourly pay
    Weekly pay
    Full time
    Contract work
    For contractors
    Summer work
    Remote work

    Mercor

    Remote
    17 hours ago
  •  ...About this role Role Title: AI Software Engineering Domain Expert Role Type: Contractor (Part Time) Location: Remote micro1 is engaging...  ...practices in software engineering. Create, refine, and evaluate prompts and instructions designed to guide AI models in producing... 
    Part time
    For contractors
    Remote work

    micro1

    Remote
    2 days ago
  •  ...Job Position: Lead AI Engineer with Retirement OR Wealth Domain Location: Boston, MA or Windsor, CT Job Type: Contract Job...  ...strategies, hybrid search, re-ranking, and rigorous evaluation (RAGAS, custom eval frameworks, or equivalent).... 
    Contract work

    Apptad Inc

    Remote
    2 days ago
  •  ...Join a fast-paced AI evaluation initiative supporting one of the world's leading AI research organizations. We are seeking detail-oriented professionals to evaluate AI-generated outputs by applying structured grading rubrics with precision and consistency. This is... 
    Temporary work
    Immediate start

    Weekday

    Remote
    a month ago
  •  ...Join a pioneering AI initiative focused on building next-generation evaluation benchmarks for frontier AI models. We are seeking analytical and technically skilled professionals to identify where advanced AI systems fail in subtle, real-world scenarios. Working in a red... 
    Full time
    Contract work
    For contractors
    Remote work
    Flexible hours

    Weekday

    Remote
    a month ago
  • $60 per hour

     ...seeking contributors for a part-time QA project focused on autonomous AI agents. This flexible remote opportunity requires strong...  ...familiarity with structured data formats. Candidates will review evaluation tasks, identify inconsistencies, and help define expected AI behaviors... 
    Part time
    Remote work
    Flexible hours

    Mind Rift

    Kansas City, MO
    3 days ago
  • $25 - $30 per hour

     ...Bilingual Simplified Chinese AI Evaluation Specialist is a remote Chinese specialist track for evaluating chinese evaluation outputs against native-speaker standards. Reviewers spot fluency, register, and cultural-context errors that automated checks miss, and write structured... 
    For contractors
    Remote work
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  • $141.02k - $204.53k

     ...your future. Responsibilities As the Senior AI/ML Engineer - Validation & Evaluation within AI Validation & Monitoring (AVM), you will provide...  ...-kit-learn, Keras, etc. Knowledge of the healthcare domain, including clinical workflows, electronic health... 
    Full time
    Work at office
    Remote work
    Flexible hours
    Weekend work

    Mayo Clinic

    Rochester, MN
    4 days ago
  • $40 - $85 per hour

     ...in over 190 countries enjoying TV series, films and games across...  ...support provided) Domain expertise in one or more of the...  ...prompt engineering, alignment, evaluation, text generation, embeddings...  ...panel/choice modeling Agentic AI: LLM agents, tool use, retrieval... 
    Hourly pay
    Full time
    Internship
    Immediate start
    Relocation package
    Flexible hours

    Netflix

    Los Gatos, CA
    2 days ago
  • $116.04k - $168.29k

     .... Responsibilities AI/ML Engineers at AI Validation...  ...-factors and user-experience specialists, data scientists, IT, architecture...  .../ML Engineer - Validation & Evaluation, with the functional...  ...Knowledge of the healthcare domain, including clinical workflows... 
    Full time
    Remote work
    Flexible hours
    Weekend work

    Mayo Clinic

    Rochester, MN
    4 days ago
  • $49 - $98 per hour

     ...Bilingual Japanese AI Evaluation Specialist is a remote evaluation track for reviewing japanese generalist evaluation prompts and responses against...  ...inter-rater agreement metrics and calibration cycles. Domain expertise that lets you spot subject-matter errors... 
    Remote job
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  •  ...for local news.   We are seeking an AI Engineer to design, develop, and deploy...  ...integrating Snowflake data with LLMs. Evaluate and fine-tune foundation models via AWS Bedrock...  ...come from a TEGNA email address with a domain of tegna.com or one of our affiliate... 
    Full time
    Temporary work
    Part time
    Local area

    Tegna

    United States
    17 hours ago
  • $163.28k - $236.75k

     ...your future. Responsibilities As the Principal AI/ML Engineer - Validation & Evaluation Governance within AI Validation & Monitoring (AVM), you...  ...MLOps communities. In-depth knowledge of healthcare domain, including clinical workflows, electronic health... 
    Full time
    Interim role
    Work at office
    Remote work
    Flexible hours
    Weekend work

    Mayo Clinic

    Rochester, MN
    5 days ago
  •  ...experienced frontend and full-stack software developers across eligible global regions to support leading AI labs in training frontier models on frontend code evaluation. This role focuses on leveraging your web development expertise to evaluate and grade AI-generated... 
    Temporary work
    For contractors
    Remote work

    Mercor

    Remote
    2 days ago
  •  ...of technical experts in LLM training, post-training, and evaluation systems. As an AI/ML Research Engineer, LLM Training & Evaluation , you will...  ...related methods) ~ RLHF / RLAIF-style workflows ~ task- or domain-adaptation of foundation models Strong programming... 
    Full time

    Innodata

    Remote
    a month ago
  • $20 - $80 per hour

     ...Role Overview Help improve next-generation AI systems by supplying precise, real-world evaluation, annotation, and feedback. This remote contractor role focuses on how AI models learn, reason, and perform across diverse subject areas. Key Responsibilities Evaluate... 
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    United States
    a month ago
  • $36 - $72 per hour

     ...power 25 million job seekers, 1 million+ employers, and 1,600 educational institutions. Handshake AI works directly with frontier AI lab researchers to create evaluations, publish benchmarks, and improve AI models through human expertise. Role Details Location:... 
    Hourly pay
    Full time
    Monday to Friday
    Flexible hours

    Handshake

    Seattle, WA
    6 days ago
  •  ...Employment Type: Project-based | Contract  We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists with strong proficiency in English. In this role, you will support AI/ML projects by annotating,... 
    Contract work
    Remote work
    Work from home
    Monday to Friday
    Day shift

    iMerit Technologies

    San Jose, CA
    a month ago
  •  ...Summary This is a fully remote, hourly contractor role supporting AI data and language projects on a project-based, flexible hour...  ..., and other content to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy, reasoning quality... 
    Hourly pay
    For contractors
    Remote work
    Flexible hours

    CNTXT AI

    Brooklyn, NY
    8 days ago
  • $20 - $80 per hour

     ...expertise to help train next-generation AI systems. Your work will shape how models...  ...real-world input. Key Responsibilities: Evaluate and score AI-generated responses using...  ...finance to marketing, healthcare, and legal domains. Required Skills and Qualifications:... 
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    17 hours ago
  •  ...We are looking for an AI Evaluation Scientist to design and execute evaluation processes that ensure our predictive and generative AI systems are accurate, reliable, safe, and aligned with mission requirements. This role is essential for establishing trust in AI solutions... 
    Remote work

    Convergenz

    United States
    1 day ago
  • $229.9k - $262.4k

    Senior Lead AI Engineer (SDK's: Gen AI Evaluation and MCP) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized... 
    Full time
    Part time
    Local area

    Capital One

    San Francisco, CA
    2 days ago
  • $100k - $150k

     ...Generative AI Specialist- Remote Bright Vision Technologies is a technology consulting and...  ...of building reusable design patterns, evaluation frameworks, and developer tooling that...  ...over prompting. Exposure to product domains such as customer support, coding assistants... 
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    United States
    3 days ago
  • $155k - $180k

     ...Senior AI Engineer Remote, United States | Remote | $155,000 to $180,000 + Equity +...  ...company powering the convergence of linear TV and digital video advertising. This enables...  ...live customers, operate in a rigorous evaluation-driven culture, and help VideoAmp safely scale... 
    Summer holiday
    Work at office
    Remote work
    Flexible hours

    VideoAmp Careers Website

    United States
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Evaluation Specialist - Movies & TV Domain. Be the first to apply!