Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Evaluation Specialist

Avance Consulting

AI Evaluation Specialist

Client is engaging AI Evaluation Specialists to assess and elevate the quality of AI assistant outputs for an enterprise AI training initiative. In this role, you'll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and perform through high-quality, real-world input. No prior experience in AI is required — your domain knowledge is what matters.

Scope of Work

  1. Evaluate AI-generated outputs against detailed rubrics and defined quality standards, focusing on accuracy, relevance, and adherence to guidelines.
  2. Apply consistent, impartial judgment across a high volume of examples, ensuring a fair and reliable assessment process.
  3. Identify reasoning gaps, tool-use failures, or logic errors in AI assistant responses, providing actionable feedback for iterative improvement.
  4. Produce clear, concise written feedback on both strengths and areas for improvement, directly influencing model refinement and AI adoption practices.
  5. Participate in discussions regarding rubric interpretation and evolving quality standards, contributing to process optimization and best practices.
  6. Maintain meticulous documentation of evaluations and recommendations, ensuring transparency and traceability in assessment workflows.

Preferred Qualifications

  1. Experience in grading, quality assurance, editorial review, assessment, annotation, or similar fields demanding careful analysis and detailed feedback.
  2. Advanced, daily use of AI assistants (such as ChatGPT, Claude, or similar) as an essential work and productivity tool.
  3. Demonstrated ability to synthesize complex information and communicate findings effectively in writing.
  4. Background in process improvement, rubric development, or operational quality assessment in an enterprise or educational context.
  5. Strong critical thinking skills with a focus on consistency, integrity, and fairness in evaluations.
  6. Comfort working independently on large volumes of similar examples while maintaining high attention to detail.
  7. Collaborative mindset for sharing insights, discussing ambiguous cases, and refining evaluation criteria as models evolve.

About the company

Avance Consulting is a global leader in innovative talent solutions for diverse industries. Since 2007, we have been serving nearly a third of Fortune 500 companies and most of the top 100 technology companies worldwide.

Our team of 700+ experts offers Engineering Services, Information Technology, Digital and Executive Search solutions from our offices in North America, UK, Europe and APAC.

We understand the local markets and cultures and tailor our services to meet the needs of our clients. Our clients value our flexibility and agility, which enable us to deliver high-quality results at any scale across the globe.

This allows us to work with numerous industries at varying scale across the globe while ensuring the excellent service standard that our clients have come to expect.

At Avance you can learn new skills and technologies, have more advancement opportunities and career growth, being a part of our productive atmosphere, and a positive reinforcement culture.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the AI Evaluation Specialist in United States vacancy
  • $70 per hour

     ...clients Compensation: $70 per hour Join a cutting-edge AI research initiative focused on improving the quality, accuracy,...  ...with exceptional critical thinking and communication skills to evaluate AI-generated responses across a variety of topics. In this role... 
    Suggested
    Hourly pay
    Weekly pay
    Contract work
    For contractors
    Remote work
    Flexible hours

    Weekday AI

    United States
    16 hours ago
  •  ...What Is An Ai Evaluation Specialist? An AI evaluation specialist assesses and tests artificial intelligence systems to ensure they perform accurately, reliably, and safely. They measure how well AI models complete tasks such as answering questions, generating content... 
    Suggested
    Work at office
    Remote work

    Valence

    United States
    21 hours ago
  • $80 per hour

     ...part-time opportunity focused on quality assurance for autonomous AI agents. You will analyze complex systems, review tasks for logic...  ...and detail-oriented skills, with experience in policy evaluation or logic puzzles preferred. Compensation can reach up to $80/hour... 
    Suggested
    Part time
    Remote work
    Flexible hours

    Mind Rift

    United States
    3 days ago
  • A leading research accelerator is seeking a Geospatial Expert to enhance AI systems through advanced geospatial analysis. This entry-level, remote role involves evaluating geospatial datasets and supporting tasks aligned with crisis management and agriculture. Candidates... 
    Suggested
    Remote job
    Contract work

    Turing

    Seattle, WA
    2 days ago
  • $25 - $30 per hour

     ...Bilingual German AI Evaluation Specialist is a remote German specialist track for evaluating german evaluation outputs against native-speaker standards. Reviewers spot fluency, register, and cultural-context errors that automated checks miss, and write structured rationale... 
    Suggested
    For contractors
    Remote work
    10 hours per week

    AuraOne Human Data

    Remote
    a month ago
  •  ...Join a pioneering AI initiative focused on building next-generation evaluation benchmarks for frontier AI models. We are seeking analytical and technically skilled professionals to identify where advanced AI systems fail in subtle, real-world scenarios. Working in a red... 
    Full time
    Contract work
    For contractors
    Remote work
    Flexible hours

    Weekday

    Remote
    a month ago
  •  ...Join a fast-paced AI evaluation initiative supporting one of the world's leading AI research organizations. We are seeking detail-oriented professionals to evaluate AI-generated outputs by applying structured grading rubrics with precision and consistency. This is... 
    Temporary work
    Immediate start

    Weekday

    Remote
    a month ago
  •  ...About the Opportunity We are looking for experienced entertainment professionals and subject matter experts to support AI evaluation initiatives focused on Movies and TV content. The role involves creating challenging evaluation scenarios, reviewing AI-generated responses... 

    eDataBae

    United States
    9 days ago
  •  ...About the Opportunity We are looking for experienced professionals with strong expertise in the video game industry to support AI evaluation initiatives. The role involves creating gaming-related evaluation scenarios, reviewing AI-generated responses, and providing... 

    eDataBae

    United States
    9 days ago
  •  ...Supporting diverse AI data and language projects, the hourly contractor AI Trainer and Evaluator will work remotely to generate content, annotate data, and evaluate AI responses for accuracy and cultural relevance. Key responsibilities Generate high-quality prompts and... 
    Hourly pay
    For contractors
    Remote work

    Virtual Vocations Inc

    United States
    16 hours ago
  • $20 per hour

    A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI...  ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with... 
    Hourly pay
    Remote work
    Flexible hours

    SupportFinity

    United States
    3 days ago
  • $80 per hour

    A technology firm is seeking QAs for autonomous AI agents to validate and improve task structures within a new project. Candidates...  ...analytical thinkers with strong attention to detail and experience in evaluating scenarios. This flexible, project-based role offers competitive... 
    Remote job
    Flexible hours

    Mindrift

    Providence, RI
    4 days ago
  • Welo Data in San Francisco seeks a full-time AI Evaluator with professional proficiency in Portuguese (Portugal) and experience in Generative AI safety. The role involves critiquing AI outputs, identifying biases, and refining evaluation frameworks. Candidates should possess... 
    Full time

    Welo Data

    San Francisco, CA
    2 days ago
  • $60 per hour

     ...seeking contributors for a part-time QA project focused on autonomous AI agents. This flexible remote opportunity requires strong...  ...familiarity with structured data formats. Candidates will review evaluation tasks, identify inconsistencies, and help define expected AI behaviors... 
    Remote job
    Part time
    Flexible hours

    Mindrift

    Kansas City, MO
    4 days ago
  •  ...experienced frontend and full-stack software developers across eligible global regions to support leading AI labs in training frontier models on frontend code evaluation. This role focuses on leveraging your web development expertise to evaluate and grade AI-generated... 
    Temporary work
    For contractors
    Remote work

    Mercor

    Remote
    6 days ago
  • $35 - $42 per hour

     ...week based on project availability Help Shape the Future of AI in Finance, Audit & Risk Volga Partners is seeking experienced...  ...legal professionals to join our growing network of on-call AI Evaluation Specialists supporting the development, training, and evaluation of next-... 
    Hourly pay
    Extra income
    Permanent employment
    Full time
    Contract work
    Temporary work
    Remote work
    Flexible hours

    Volga Partners

    Remote
    3 days ago
  • $150k - $250k

    About Distyl AI Distyl is an applied AI technology company partnering with the world’s most ambitious institutions to rearchitect critical...  ...0s. What We Are Looking ForAt Distyl, we build AI systems using Evaluation-Driven Development—an approach where evaluation is not an... 
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    4 days ago
  • Overview In this role, you will evaluate AI-generated responses and provide structured written feedback. This is a great opportunity for sharp, analytical thinkers to contribute to high-impact AI research projects. Basic Qualifications Bachelor's degree from a top-500 globally... 

    Obsidian

    Boston, MA
    1 day ago
  • $11 - $30.65 per hour

    Meridial is seeking contractors to evaluate advanced agentic audio models by simulating realistic customer service interactions across multiple domains. You will contribute to developing diverse datasets and assess model performance using various metrics. The role requires... 
    Remote job
    Hourly pay
    For contractors

    Meridial

    New York, NY
    1 day ago
  • $60 per hour

    Prolific is looking for Biology Experts and Life Science Professionals in Las Vegas, NV to evaluate AI-generated science. Successful candidates will work flexibly from home, earning up to $60 per hour for paid tasks. Responsibilities include reviewing scientific data accuracy... 
    Remote job
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    Las Vegas, NV
    4 days ago
  • Obsidian is collaborating with AI labs to find experienced health insurance professionals to enhance AI systems related to coverage...  ...assess AI performance on health insurance tasks. The role includes evaluating AI outputs, creating health insurance scenarios, and providing... 

    Obsidian

    New York, NY
    5 days ago
  • Prolific is seeking talented Product Designers and UX Specialists to join our Expert Network for training and evaluation of cutting-edge AI models. Candidates should have a relevant educational background and at least one year of experience in design-related fields. Responsibilities... 
    Remote job
    Work from home
    Flexible hours

    Prolific

    Arizona City, AZ
    3 days ago
  • Prolific is seeking Product Designers and UX Specialists to join our Expert Network, contributing to the training and evaluation of cutting-edge AI models. This role requires expertise in usability, design systems, and user research, providing essential feedback on design... 
    Remote job
    Work from home
    Flexible hours

    Prolific

    Tucson, AZ
    3 days ago
  • $20 - $80 per hour

     ...Role Overview Help improve next-generation AI systems by supplying precise, real-world evaluation, annotation, and feedback. This remote contractor role focuses on how AI models learn, reason, and perform across diverse subject areas. Key Responsibilities Evaluate... 
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    United States
    a month ago
  • $30 per hour

    Prolific is seeking Advanced Dutch Speakers in Chicago, IL to train AI models. You will complete AI tasks and assess AI performance in...  ...competitive rates and direct payment through PayPal. Pass the evaluation, and you can start within 15 minutes. #J-18808-Ljbffr Prolific
    Remote job
    Work from home

    Prolific

    Chicago, IL
    4 days ago
  •  ...Employment Type: Project-based | Contract  We are looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation and Transcription Specialists with strong proficiency in English. In this role, you will support AI/ML projects by annotating,... 
    Contract work
    Remote work
    Work from home
    Monday to Friday
    Day shift

    iMerit Technologies

    San Jose, CA
    more than 2 months ago
  •  ...Summary This is a fully remote, hourly contractor role supporting AI data and language projects on a project-based, flexible hour...  ..., and other content to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy, reasoning quality... 
    Hourly pay
    For contractors
    Remote work
    Flexible hours

    CNTXT AI

    Brooklyn, NY
    12 days ago
  • $147.76k - $240.11k

     ..., we are building a better world, so we can all enjoy living in it.Job Summary:Join the AI Engineering team of Cat Digital and take charge of leading a team dedicated to evaluating and validating our advanced generative AI solutions—including intelligent agents, digital... 
    Full time
    Part time
    Flexible hours

    Caterpillar

    Westminster, CO
    4 days ago
  •  ...Job TitleAI Evaluation EngineerLocationHybrid / RemoteEmployment TypeFull-timeJob SummaryWe are seeking an AI Evaluation Engineer to design, implement, and maintain evaluation frameworks for AI and machine learning systems, with a focus on Large Language Models (LLMs)... 

    Ova Technologies

    New York, NY
    16 hours ago
  • $105k - $145k

    OverviewWe are looking for an AI Evaluation Scientistto design and execute evaluation processes that ensure our predictive and generative AI systems are accurate, reliable, safe, and aligned with mission requirements. This role is essential for establishing trust in AI... 
    Local area

    Steampunk

    McLean, VA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Evaluation Specialist. Be the first to apply!