Insurance Expert for AI Model Evaluation
$60 - $80 per hourSaidGig
Bring real-world underwriting, claims, and risk assessment expertise into GenAI model development by designing insurance-specific tasks, producing practiced solutions, and rigorously evaluating large language model outputs against structured rubrics. This role combines deep domain judgment with hands-on model evaluation to improve the correctness and reasoning quality of training data. The position is W-2 employment through Cincinnatus LLC, with placement on a leading AI lab team.
Key Responsibilities- Work with research and engineering teams to close knowledge gaps in underwriting, claims, and risk-assessment reasoning.
- Design challenging, domain-relevant insurance tasks that reflect real underwriting and claims practice.
- Write accurate, well-reasoned solutions for those tasks grounded in real-world practice.
- Evaluate AI model outputs against structured rubrics, providing clear written feedback on correctness, judgment, and reasoning quality.
- Develop and refine evaluation guidelines and scoring rubrics specific to insurance tasks.
- Collaborate with other subject matter experts to ensure consistency and accuracy in training data.
- At least 8 years of professional experience in insurance, such as underwriting, claims, actuarial work, or risk management, at a recognized organization (examples include AIG, Chubb, Allstate, Progressive, MetLife, Marsh McLennan, or equivalent).
- Prior hands-on experience evaluating LLM or AI model outputs against rubrics or structured scoring criteria, this is mandatory, please describe this experience in your application.
- Demonstrable career progression, for example moving from Underwriter to Senior Underwriter to VP of Underwriting or similar trajectory.
- Strong verbal and written communication skills, problem-solving ability, and interpersonal skills for cross-functional collaboration.
- Ability to engage reliably for at least 35 hours per week during weekdays.
- Employment type, W-2 employee of Cincinnatus LLC, placed as part of an extended workforce at a leading AI lab.
- Location, United States.
- Schedule, minimum commitment of 35 hours per week, weekdays required.
- This is an hourly role, employment relationship and workplace administration managed by Cincinnatus LLC, including payroll and benefits administration.
- Pay range, 60 to 80 hourly.
- Candidates must be able to work in the United States under W-2 employment with Cincinnatus LLC.
- Cincinnatus LLC is an equal employment opportunity employer, and does not discriminate based on legally protected characteristics.
When you apply, include your resume and a written description of your prior experience evaluating LLM or AI model outputs against rubrics or structured scoring criteria. Be prepared to confirm your weekday availability for at least 35 hours per week. Applications that explicitly describe the required LLM evaluation experience will be prioritized.
$65 per hour
...Cybersecurity professionals apply offensive and defensive security expertise to design domain-specific prompts and evaluate large language model outputs for AI research projects, improving model behavior, safety, and relevance in security-related scenarios. Key...SuggestedPart timeRemote workFlexible hours$60 - $80 per hour
...building foundational large language models, applying real-world... ...operations judgment to design tasks, evaluate model outputs, and guide... ...retail practice. Evaluate AI model outputs against structured... ...Collaborate with other subject matter experts to ensure consistency and...SuggestedHourly payContract workWeekday work$140 per hour
...matter expertise to improve how next-generation AI systems learn, reason, and perform. As an AI Domain Expert you will evaluate AI outputs, create challenging prompts,... ...feedback that helps train and validate advanced models. This is a remote, part-time contractor role...SuggestedRemote jobHourly payPart timeFor contractorsVisa sponsorshipWork visaFree visa- ...Prolific is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia....SuggestedHourly payRemote workFlexible hours
$60 per hour
...Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing...SuggestedHourly payRemote workWork from homeFlexible hours$60 per hour
...Prolific is seeking Biology Experts and Life Science Professionals to join their Expert Network to evaluate AI-generated science. This role allows you to work from home with... ...pay rate of up to $60 per hour for reviewing model responses, validating technical claims, and critiquing...Hourly payRemote workWork from homeFlexible hours- ...Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant...Remote workFlexible hours
$105 per hour
...extensive experience in insurance verification and... ...management to enhance AI tools aimed at automating... ...crucial role in shaping AI models that improve accuracy and... ...service delivery. Evaluate and annotate AI-generated... ...a management role. ~ Expert knowledge of EDI 270/27...$50 - $101 per hour
...expertise to help train next-generation AI systems. In this remote, contractor role... ...and nutrition content that improves how models learn and reason. No prior AI experience... ...general wellness guidance. Create and evaluate sample fitness programs and nutrition plans...Hourly payFor contractorsRemote work$80 - $160 per hour
...Role Overview Model stochastic bacterial population dynamics to derive asymptotic growth rates, analyze effects of growth-rate switching... ...clear, reproducible methodological documentation to train and evaluate AI systems. Key Responsibilities Analyze and model...Remote jobHourly payFor contractors$80 - $160 per hour
...expertise to a research-focused project that benchmarks and models quantum optical systems. The work centers on cascaded... ...grade explanations and analyses that can be used to train and evaluate next-generation AI systems, no prior AI experience required....Remote jobHourly payFor contractorsWork at office- ...Medical professionals apply clinical and workplace expertise to evaluate AI-generated content in their specialty, assess field-specific materials, and provide clear, structured feedback that improves model performance on medical tasks and language. This hourly, temporary...Hourly payTemporary workPart timeRemote workFlexible hours
$85 per hour
...world aerodynamics engineering problems intended to challenge the most capable AI models. You will create original free-response questions grounded in industry scenarios, produce complete expert-level solutions, and verify problem difficulty by testing each question...Hourly payRemote work$40 - $65 per hour
...scenarios that probe frontier language models, then evaluate and document model behavior so engineering... .... Preferred experience with AI human data environments such as RLHF, SFT... ...This engagement is output-based, with experts paid per completed task that meets project...Remote jobHourly payFor contractorsImmediate start$40 per hour
A healthcare technology company in the United States seeks medical experts to evaluate AI chatbots. Responsibilities include presenting healthcare problems to AI and assessing their responses for correctness. Candidates must be fluent in English and possess a current or...Hourly payRemote work$60 - $80 per hour
...expertise to a GenAI team building foundational AI models. You will create realistic marketing tasks, evaluate model outputs against structured rubrics, and advise... ...tasks. Collaborate with other subject matter experts to ensure consistency and accuracy in training...Hourly payWeekday work$75 per hour
...Role Overview Finance professionals apply their financial analysis, modeling, and advisory expertise to evaluate AI-generated financial content, identify errors, and provide guidance that improves AI understanding of financial concepts, quantitative reasoning, industry...Full timePart timeFor contractorsBank staffRemote workFlexible hours$30 - $90 per hour
...Overview Work as a remote Go developer contributing to the evaluation and training of next-generation AI coding tools in confidential alpha stages. This part-... ...performance. Test and evaluate alpha AI coding models in Cursor, running focused testing sessions....Hourly payContract workPart timeRemote work$110 per hour
...Overview Provide clinical expertise to help train, evaluate, and shape medical AI systems by joining a Physician Expert Network. This is an open application to be... ...Key Responsibilities Train and evaluate AI models in medical and clinical contexts. Create tasks...Hourly payContract workRemote work$60 - $75 per hour
Role Description Join an advanced AI research initiative focused on improving how next... ...design high-quality benchmark tasks that evaluate AI performance across software... ...help measure and improve the ability of AI models to interpret technical information, follow...Weekly payContract workPart timeFor contractorsRemote workFlexible hours$60 - $150 per hour
...Role Overview Join a Law Expert Network to provide legal subject matter expertise that helps train and evaluate AI systems and inform frontier AI research. This is an open application... ...Support training and evaluation of AI models on legal topics. Create tasks, prompts,...Hourly payContract workImmediate startRemote work$85 per hour
...Role Overview Help evaluate and improve frontier AI coding models by completing realistic machine learning engineering tasks and assessing model-generated implementations. You will work with cutting-edge coding agents to surface bugs, failure modes, and tradeoffs in...Hourly payRemote work$80 - $110 per hour
...Contribute subject-matter expertise to the development and evaluation of next-generation AI systems that must reason about pure and applied... ...of a working research mathematician. You will help ensure models understand and produce correct, rigorous mathematics across...Hourly payPart timeImmediate startRemote work$80 - $110 per hour
...Role Overview Contribute frontier research expertise to the development and evaluation of AI systems that reason about real research chemistry. This role supports model training and assessment across organic synthesis, reaction mechanism, catalysis, structural and physical...Hourly payPart timeImmediate startRemote work$60 - $90 per hour
...analysis tasks that serve as ground-truth references for evaluation of frontier generative AI models. You will create one-to-two day, end-to-end analysis... ...Collaborate with researchers and other subject-matter experts to align evaluation standards and maintain consistency...Hourly payFull timePart timeWork experience placementFreelanceRemote work$224k - $356.5k
...people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts... ...computing. As a Senior / Principal Deep Learning Engineer — Model Evaluation & AI Systems, you will play a meaningful role in crafting the...Full time$75 per hour
...Role Overview Evaluate and improve how medical AI systems reason about real clinical problems. In this remote, contract role you will partner with research... ..., then work with engineers and researchers to improve model performance. Key Responsibilities Design systematic...Contract workRemote workFlexible hours$50 per hour
...Role Overview Work on fine-tuning large language models by designing, solving, and explaining challenging Biology problems. You will... ...step-by-step solutions that probe LLM limitations, help define evaluation benchmarks across undergraduate to PhD level curricula, and collaborate...Contract workFor contractorsFreelanceRemote work- ...diagnostic expertise to review and improve AI-generated medical content, assessing... ...detailed feedback and corrections to help models better represent medical knowledge and clinical... ...making. Follow written guidelines for evaluations, maintain professional judgment, and...Full timeFor contractorsPrivate practiceRemote workFlexible hours
$60 per hour
...and contribute to developing cutting-edge AI systems, while enjoying the flexibility... ...professionals to help advance AI development. AI models are increasingly capable of performing... ...state-of-the-art AI models on tasks like evaluating AI-generated quantitative analysis,...Hourly payFull timeRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Insurance Expert for AI Model Evaluation. Be the first to apply!
- expert data analyst United States
- subject matter expert senior United States
- technology expert United States
- expert systems engineer United States
- subject matter expert United States
- guest service support expert United States
- sql expert United States
- subject matter expert work from home United States
- fulfillment expert United States
- customer service representative insurance United States



