Lawyer for AI Model Evaluation
$70 - $110 per hourSaidGig
Role Overview
Apply your legal practice experience to evaluate and improve AI-generated legal content and workflows. You will assess work grounded in litigation, legal drafting, and client advisory practice.
Key Responsibilities
- Evaluate AI-generated legal memoranda, contracts, case analyses, and other legal materials for quality and accuracy.
- Compare AI responses for legal reasoning, application of precedent, and drafting rigor.
- Review scenarios involving litigation strategy, statutory interpretation, and client advisory decisions, then provide expert feedback.
Qualifications
- At least 4 years of experience as a Lawyer, Attorney, Counsel, General Counsel, Prosecutor, or in a closely related legal practice role.
- Experience representing clients in criminal or civil proceedings, drafting legal documents, or advising on legal transactions.
- Strong written English communication skills.
Work Terms
- Remote, hourly engagement.
- Immediate start.
- Commitment of at least 20 hours per week.
- Expected project length of approximately 4 to 6 weeks.
Compensation
- Pay range of $70 to $110 per hour.
- Initial work is paid on a task basis after approval of the first task. Completing that task within the required timeline qualifies you for hourly payment for the remainder of the project.
$60 - $150 per hour
...connects experienced legal professionals with AI labs and companies for short term... ...legal subject matter expertise to train and evaluate AI models, create realistic tasks and deliverables... ...This is an open application to join the Lawyer talent network, not a single, immediate...SuggestedHourly payContract workTemporary workImmediate startRemote work$100 - $150 per hour
...Role Overview Help shape how AI systems understand and apply legal principles by designing evaluation rubrics, performing complex legal research and document analysis... ...research and draft complex memoranda to guide AI model training and evaluation. Analyze large...SuggestedHourly payContract workRemote work$208k - $300k
...Machine Learning Engineer - Model Evaluations, Public Sector The Public Sector ML team at Scale deploys advanced AI systems—including LLMs, agentic models, and multimodal pipelines—into mission-critical government environments. We build evaluation frameworks that ensure...SuggestedFull time$60 - $90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation...SuggestedFull timeContract workSummer workRemote work$224k - $356.5k
...people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts... ...computing. As a Senior / Principal Deep Learning Engineer — Model Evaluation & AI Systems, you will play a meaningful role in crafting the...SuggestedFull time- Cincinnatus LLC is placing finance SMEs at a leading AI lab to critically evaluate AI model outputs for financial tasks. The role focuses on rigorous, rubric-based assessment of model performance and constructing finance-focused evaluation frameworks. Candidates should...Weekday work
$20 - $60 per hour
...Help train next-generation AI systems by creating rigorous, real-world evaluations that test how well advanced models learn, reason, and perform. This remote contract opportunity is open to recent graduates, advanced-degree holders, and professionals from any background...Hourly payContract workFor contractorsRemote work- ...Evaluate generative music AI across a broad range of genres, applying your knowledge of Indonesian music and lyrics to detailed quality standards. You will work in both Indonesian and English to help assess lyric quality and authenticity. Key Responsibilities Compare...Hourly payImmediate startRemote workFlexible hours
$42 - $78 per hour
...Evaluate generative music and lyrics in Norwegian and English, helping assess output across a broad range of genres against detailed quality standards. Key Responsibilities Compare AI-generated lyrics with published songs to identify similarities. Rate lyrics for...Hourly payFor contractorsImmediate startRemote workFlexible hours$70 - $110 per hour
...Role Overview Help a leading AI research team improve how advanced AI models reason about real clinical work. In this hybrid, full-time role, you will... ...define high-quality clinical tasks, model answers, and evaluation standards alongside research and program management...Hourly payFull timeFreelanceLive inRelocationRelocation package- ...Overview Apply advanced physics knowledge to help improve and evaluate large language models. You will design rigorous problems, produce clear reasoning... ...into accessible explanations while contributing to AI research projects. Key Responsibilities Design and solve...For contractorsFreelanceRemote work
$65 - $90 per hour
...Apply deep financial judgment to help improve foundational AI models. In this role, you will create and evaluate finance-focused work that strengthens AI systems'' reasoning, analysis, and decision-making capabilities. Role Overview This is a W-2 employment opportunity...Hourly payWeekday work- ...reasoning and computational problem solving to improve and evaluate large language models. You will design rigorous math problems, produce clear, logically... ...How this work supports customers Accelerate frontier AI research by contributing high quality data and evaluation...Contract workFor contractorsFreelanceRemote work
$13 - $54 per hour
...Evaluate AI-generated music and lyrics across a wide range of genres, using your knowledge of the Spanish (Mexico) music scene to assess quality against detailed standards. This remote role involves working in both Spanish (Mexico) and English. Key Responsibilities...Hourly payImmediate startRemote workFlexible hours$70 - $110 per hour
...Help advance frontier AI systems by bringing rigorous materials science and engineering judgment to the evaluation, design, and improvement of technical knowledge work. You will... ...reasoning looks like in practice and ensure model outputs can withstand technical scrutiny....Hourly payFull timeLive inRelocationRelocation package$60 - $80 per hour
...Help shape the training and evaluation of foundational large language models by applying real-world expertise in brand strategy, growth marketing, and campaign... .... This role brings rigorous marketing judgment to AI tasks, model assessments, and training data for a leading...Hourly payWeekday work$35 - $62 per hour
...Apply your Japanese music expertise to evaluate AI-generated music and lyrics across a wide range of genres. You will assess outputs against detailed quality standards in both Japanese and English. Key Responsibilities Compare AI-generated lyrics with published songs...Hourly payFor contractorsImmediate startRemote workFlexible hours$18 - $42 per hour
...Role Overview Evaluate generative music AI outputs across a wide range of genres, applying your knowledge of Portuguese-language music and lyrics to detailed quality standards. This role combines critical listening with lyric analysis in Portuguese and English. Key...Hourly payImmediate startRemote workFlexible hours$18 per hour
...Role Overview Evaluate AI-generated music and lyrics across a range of genres, applying detailed quality standards in both Thai and English... ...of Thai music, language, and lyrical expression to help assess model outputs. Key Responsibilities Compare AI-generated lyrics...Hourly payFor contractorsImmediate startRemote workFlexible hours- ...is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities...Remote jobHourly payFlexible hours
$60 per hour
...and Life Science Professionals to join their Expert Network to evaluate AI-generated science. This role allows you to work from home with... ...offers a competitive pay rate of up to $60 per hour for reviewing model responses, validating technical claims, and critiquing...Remote jobHourly payWork from homeFlexible hours$60 per hour
...in Arizona, is seeking Biology Experts and Life Science Professionals to join their Expert Network. This role involves evaluating and training AI models with real scientific expertise. Successful candidates will review AI-generated content for accuracy and assist in fact...Hourly pay- Dorado is seeking an experienced Investment Banking SME to support the development, evaluation, and improvement of advanced AI models in finance. You will assess AI-generated analyses for accuracy, reasoned judgments, and data integrity, guiding model refinements and prompts...Remote job
$60 per hour
Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing...Remote jobHourly payWork from homeFlexible hours- Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant...Remote jobFlexible hours
$60 - $90 per hour
...Role Overview Help advance frontier AI research by creating rigorous, real-world data analysis evaluations for generative AI models. You will design and complete complex analytical tasks that mirror practical research work, then use your reference analyses to assess where...Hourly payFull timeFreelanceRemote work$80 - $135 per hour
...09.26574v3), a frontier research-level physics benchmark. The role produces fully human-verified reference data used to evaluate large language model performance on frontier physics reasoning. Work includes solving CritPt research-level problems end-to-end, auditing other...Hourly payRemote work10 hours per week$100 per hour
...Role Overview Apply deep domain expertise to train and evaluate next-generation AI systems by producing, refining, and validating high-quality,... ...data. This part-time contractor role focuses on improving model outputs through careful content review, prompt refinement,...Hourly payPart timeFor contractorsRemote work$100 - $150 per hour
...Role Overview Evaluate how well AI systems perform real-world technical sales work by defining excellence and judging completed work samples. Rather than producing sales deliverables yourself, you will create task-specific grading rubrics and assign scores to AI-generated...Hourly payRemote work$60 - $90 per hour
...Author and run rigorous, multi-step machine learning evaluation tasks for a leading generative AI research team. You will take high-level research ideas... ...experiments, and analyze results to determine where frontier models succeed or fail. Typical tasks require one to two days...Hourly payFull timeFreelanceRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lawyer for AI Model Evaluation. Be the first to apply!


