Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Insurance Expert for AI Model Evaluation

$60 - $80 per hour

SaidGig

Role Overview

Bring real-world underwriting, claims, and risk assessment expertise into GenAI model development by designing insurance-specific tasks, producing practiced solutions, and rigorously evaluating large language model outputs against structured rubrics. This role combines deep domain judgment with hands-on model evaluation to improve the correctness and reasoning quality of training data. The position is W-2 employment through Cincinnatus LLC, with placement on a leading AI lab team.

Key Responsibilities
  • Work with research and engineering teams to close knowledge gaps in underwriting, claims, and risk-assessment reasoning.
  • Design challenging, domain-relevant insurance tasks that reflect real underwriting and claims practice.
  • Write accurate, well-reasoned solutions for those tasks grounded in real-world practice.
  • Evaluate AI model outputs against structured rubrics, providing clear written feedback on correctness, judgment, and reasoning quality.
  • Develop and refine evaluation guidelines and scoring rubrics specific to insurance tasks.
  • Collaborate with other subject matter experts to ensure consistency and accuracy in training data.
Qualifications
  • At least 8 years of professional experience in insurance, such as underwriting, claims, actuarial work, or risk management, at a recognized organization (examples include AIG, Chubb, Allstate, Progressive, MetLife, Marsh McLennan, or equivalent).
  • Prior hands-on experience evaluating LLM or AI model outputs against rubrics or structured scoring criteria, this is mandatory, please describe this experience in your application.
  • Demonstrable career progression, for example moving from Underwriter to Senior Underwriter to VP of Underwriting or similar trajectory.
  • Strong verbal and written communication skills, problem-solving ability, and interpersonal skills for cross-functional collaboration.
  • Ability to engage reliably for at least 35 hours per week during weekdays.
Work Terms
  • Employment type, W-2 employee of Cincinnatus LLC, placed as part of an extended workforce at a leading AI lab.
  • Location, United States.
  • Schedule, minimum commitment of 35 hours per week, weekdays required.
  • This is an hourly role, employment relationship and workplace administration managed by Cincinnatus LLC, including payroll and benefits administration.
Compensation
  • Pay range, 60 to 80 hourly.
Eligibility
  • Candidates must be able to work in the United States under W-2 employment with Cincinnatus LLC.
  • Cincinnatus LLC is an equal employment opportunity employer, and does not discriminate based on legally protected characteristics.
Application Process

When you apply, include your resume and a written description of your prior experience evaluating LLM or AI model outputs against rubrics or structured scoring criteria. Be prepared to confirm your weekday availability for at least 35 hours per week. Applications that explicitly describe the required LLM evaluation experience will be prioritized.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Insurance Expert for AI Model Evaluation in United States vacancy
  • $65 per hour

     ...Cybersecurity professionals apply offensive and defensive security expertise to design domain-specific prompts and evaluate large language model outputs for AI research projects, improving model behavior, safety, and relevance in security-related scenarios. Key... 
    Suggested
    Part time
    Remote work
    Flexible hours

    SaidGig

    United States
    1 day ago
  • $60 - $80 per hour

     ...building foundational large language models, applying real-world...  ...operations judgment to design tasks, evaluate model outputs, and guide...  ...retail practice. Evaluate AI model outputs against structured...  ...Collaborate with other subject matter experts to ensure consistency and... 
    Suggested
    Hourly pay
    Contract work
    Weekday work

    SaidGig

    United States
    2 days ago
  • $140 per hour

     ...matter expertise to improve how next-generation AI systems learn, reason, and perform. As an AI Domain Expert you will evaluate AI outputs, create challenging prompts,...  ...feedback that helps train and validate advanced models. This is a remote, part-time contractor role... 
    Suggested
    Remote job
    Hourly pay
    Part time
    For contractors
    Visa sponsorship
    Work visa
    Free visa

    SaidGig

    Remote
    5 hours ago
  •  ...Prolific is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia.... 
    Suggested
    Hourly pay
    Remote work
    Flexible hours

    Prolific

    Jacksonville, FL
    4 days ago
  • $60 per hour

     ...Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing... 
    Suggested
    Hourly pay
    Remote work
    Work from home
    Flexible hours

    Prolific

    San Jose, CA
    5 days ago
  • $60 per hour

     ...Prolific is seeking Biology Experts and Life Science Professionals to join their Expert Network to evaluate AI-generated science. This role allows you to work from home with...  ...pay rate of up to $60 per hour for reviewing model responses, validating technical claims, and critiquing... 
    Hourly pay
    Remote work
    Work from home
    Flexible hours

    Prolific

    Dallas, TX
    5 days ago
  •  ...Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant... 
    Remote work
    Flexible hours

    Prolific

    Charlotte, NC
    5 days ago
  • $105 per hour

     ...extensive experience in insurance verification and...  ...management to enhance AI tools aimed at automating...  ...crucial role in shaping AI models that improve accuracy and...  ...service delivery. Evaluate and annotate AI-generated...  ...a management role. ~ Expert knowledge of EDI 270/27... 

    SaidGig

    United States
    1 day ago
  • $50 - $101 per hour

     ...expertise to help train next-generation AI systems. In this remote, contractor role...  ...and nutrition content that improves how models learn and reason. No prior AI experience...  ...general wellness guidance. Create and evaluate sample fitness programs and nutrition plans... 
    Hourly pay
    For contractors
    Remote work

    SaidGig

    Indiana
    1 day ago
  • $80 - $160 per hour

     ...Role Overview Model stochastic bacterial population dynamics to derive asymptotic growth rates, analyze effects of growth-rate switching...  ...clear, reproducible methodological documentation to train and evaluate AI systems. Key Responsibilities Analyze and model... 
    Remote job
    Hourly pay
    For contractors

    SaidGig

    Remote
    5 hours ago
  • $80 - $160 per hour

     ...expertise to a research-focused project that benchmarks and models quantum optical systems. The work centers on cascaded...  ...grade explanations and analyses that can be used to train and evaluate next-generation AI systems, no prior AI experience required.... 
    Remote job
    Hourly pay
    For contractors
    Work at office

    SaidGig

    Remote
    4 days ago
  •  ...Medical professionals apply clinical and workplace expertise to evaluate AI-generated content in their specialty, assess field-specific materials, and provide clear, structured feedback that improves model performance on medical tasks and language. This hourly, temporary... 
    Hourly pay
    Temporary work
    Part time
    Remote work
    Flexible hours

    SaidGig

    United States
    5 hours ago
  • $85 per hour

     ...world aerodynamics engineering problems intended to challenge the most capable AI models. You will create original free-response questions grounded in industry scenarios, produce complete expert-level solutions, and verify problem difficulty by testing each question... 
    Hourly pay
    Remote work

    SaidGig

    United States
    16 hours ago
  • $40 - $65 per hour

     ...scenarios that probe frontier language models, then evaluate and document model behavior so engineering...  .... Preferred experience with AI human data environments such as RLHF, SFT...  ...This engagement is output-based, with experts paid per completed task that meets project... 
    Remote job
    Hourly pay
    For contractors
    Immediate start

    SaidGig

    Frontier County, NE
    5 hours ago
  • $40 per hour

    A healthcare technology company in the United States seeks medical experts to evaluate AI chatbots. Responsibilities include presenting healthcare problems to AI and assessing their responses for correctness. Candidates must be fluent in English and possess a current or... 
    Hourly pay
    Remote work

    DataAnnotation

    Helena, MT
    1 day ago
  • $60 - $80 per hour

     ...expertise to a GenAI team building foundational AI models. You will create realistic marketing tasks, evaluate model outputs against structured rubrics, and advise...  ...tasks. Collaborate with other subject matter experts to ensure consistency and accuracy in training... 
    Hourly pay
    Weekday work

    SaidGig

    United States
    5 hours ago
  • $75 per hour

     ...Role Overview Finance professionals apply their financial analysis, modeling, and advisory expertise to evaluate AI-generated financial content, identify errors, and provide guidance that improves AI understanding of financial concepts, quantitative reasoning, industry... 
    Full time
    Part time
    For contractors
    Bank staff
    Remote work
    Flexible hours

    SaidGig

    United States
    1 day ago
  • $30 - $90 per hour

     ...Overview Work as a remote Go developer contributing to the evaluation and training of next-generation AI coding tools in confidential alpha stages. This part-...  ...performance. Test and evaluate alpha AI coding models in Cursor, running focused testing sessions.... 
    Hourly pay
    Contract work
    Part time
    Remote work

    SaidGig

    United States
    1 day ago
  • $110 per hour

     ...Overview Provide clinical expertise to help train, evaluate, and shape medical AI systems by joining a Physician Expert Network. This is an open application to be...  ...Key Responsibilities Train and evaluate AI models in medical and clinical contexts. Create tasks... 
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    5 hours ago
  • $60 - $75 per hour

    Role Description Join an advanced AI research initiative focused on improving how next...  ...design high-quality benchmark tasks that evaluate AI performance across software...  ...help measure and improve the ability of AI models to interpret technical information, follow... 
    Weekly pay
    Contract work
    Part time
    For contractors
    Remote work
    Flexible hours

    Weekday AI

    Remote
    28 days ago
  • $60 - $150 per hour

     ...Role Overview Join a Law Expert Network to provide legal subject matter expertise that helps train and evaluate AI systems and inform frontier AI research. This is an open application...  ...Support training and evaluation of AI models on legal topics. Create tasks, prompts,... 
    Hourly pay
    Contract work
    Immediate start
    Remote work

    SaidGig

    United States
    1 day ago
  • $85 per hour

     ...Role Overview Help evaluate and improve frontier AI coding models by completing realistic machine learning engineering tasks and assessing model-generated implementations. You will work with cutting-edge coding agents to surface bugs, failure modes, and tradeoffs in... 
    Hourly pay
    Remote work

    SaidGig

    United States
    5 hours ago
  • $80 - $110 per hour

     ...Contribute subject-matter expertise to the development and evaluation of next-generation AI systems that must reason about pure and applied...  ...of a working research mathematician. You will help ensure models understand and produce correct, rigorous mathematics across... 
    Hourly pay
    Part time
    Immediate start
    Remote work

    SaidGig

    United States
    1 day ago
  • $80 - $110 per hour

     ...Role Overview Contribute frontier research expertise to the development and evaluation of AI systems that reason about real research chemistry. This role supports model training and assessment across organic synthesis, reaction mechanism, catalysis, structural and physical... 
    Hourly pay
    Part time
    Immediate start
    Remote work

    SaidGig

    United States
    1 day ago
  • $60 - $90 per hour

     ...analysis tasks that serve as ground-truth references for evaluation of frontier generative AI models. You will create one-to-two day, end-to-end analysis...  ...Collaborate with researchers and other subject-matter experts to align evaluation standards and maintain consistency... 
    Hourly pay
    Full time
    Part time
    Work experience placement
    Freelance
    Remote work

    SaidGig

    United States
    5 hours ago
  • $224k - $356.5k

     ...people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts...  ...computing. As a Senior / Principal Deep Learning Engineer — Model Evaluation & AI Systems, you will play a meaningful role in crafting the... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $75 per hour

     ...Role Overview Evaluate and improve how medical AI systems reason about real clinical problems. In this remote, contract role you will partner with research...  ..., then work with engineers and researchers to improve model performance. Key Responsibilities Design systematic... 
    Contract work
    Remote work
    Flexible hours

    SaidGig

    United States
    2 days ago
  • $50 per hour

     ...Role Overview Work on fine-tuning large language models by designing, solving, and explaining challenging Biology problems. You will...  ...step-by-step solutions that probe LLM limitations, help define evaluation benchmarks across undergraduate to PhD level curricula, and collaborate... 
    Contract work
    For contractors
    Freelance
    Remote work

    SaidGig

    United States
    2 days ago
  •  ...diagnostic expertise to review and improve AI-generated medical content, assessing...  ...detailed feedback and corrections to help models better represent medical knowledge and clinical...  ...making. Follow written guidelines for evaluations, maintain professional judgment, and... 
    Full time
    For contractors
    Private practice
    Remote work
    Flexible hours

    SaidGig

    United States
    5 hours ago
  • $60 per hour

     ...and contribute to developing cutting-edge AI systems, while enjoying the flexibility...  ...professionals to help advance AI development. AI models are increasingly capable of performing...  ...state-of-the-art AI models on tasks like evaluating AI-generated quantitative analysis,... 
    Hourly pay
    Full time
    Remote work
    Flexible hours

    DataAnnotation

    Kansas City, MO
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Insurance Expert for AI Model Evaluation. Be the first to apply!