Tax Advisor, CPA for AI Model Evaluation
$80 - $110 per hourSaidGig
Role Overview
Apply your tax preparation and advisory expertise to help improve next generation AI systems through accurate, practical evaluation and feedback. This part time contract role supports a customer project focused on tax preparation and advisory services. No prior AI experience is required.
Key Responsibilities
- Prepare and review personal Form 1040 and small business tax returns, including Schedule C, S Corporation Form 1120-S, and Partnership Form 1065 filings, for accuracy and compliance with current tax law.
- Advise on tax optimization strategies and explain complex tax concepts clearly in writing and conversation.
- Assess AI generated tax advice and content against provided rubrics, then deliver concise written feedback that improves model performance.
- Verify documentation and assist with questions about filing status, deadlines, and required forms.
- Help develop and refine evaluation rubrics and grading criteria for tax related outputs.
- Stay current on relevant federal and state tax code changes for client cases and project deliverables.
Qualifications
- Active CPA license in good standing and eligible to practice in the United States.
- 5+ years of professional tax preparation or advisory experience.
- Extensive experience preparing and filing both personal and small business income tax returns.
- Current knowledge of federal income tax regulations and income tax requirements for at least one state.
- Exceptional written and verbal communication skills, including the ability to make technical tax topics clear for varied audiences.
Preferred Qualifications
- Experience evaluating AI models, red teaming, or providing feedback on machine generated outputs, such as RLHF, is helpful but not required.
- Experience drafting rubrics, assessment guidelines, or scoring criteria for complex subject matter.
- Experience with professional tax software, including ProSeries, Lacerte, or Drake.
Work Terms
- Part time independent contractor engagement.
- Remote position available only to candidates based in the United States.
Compensation
- $80 to $110 per hour.
$60 - $90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation...SuggestedFull timeContract workSummer workRemote work$60 per hour
...Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing...SuggestedHourly payRemote workWork from homeFlexible hours$35 - $62 per hour
...Apply your Japanese music expertise to evaluate AI-generated music and lyrics across a wide range of genres. You will assess outputs against detailed quality standards in both Japanese and English. Key Responsibilities Compare AI-generated lyrics with published songs...SuggestedHourly payFor contractorsImmediate startRemote workFlexible hours$17 - $42 per hour
...Evaluate AI-generated music and lyrics in Hebrew and English, applying your knowledge of the Hebrew music scene and detailed quality standards across a wide range of genres. Key Responsibilities Compare AI-generated lyrics with published songs to identify similarities...SuggestedHourly payFor contractorsImmediate startRemote workFlexible hours$50 - $70 per hour
...Role Overview Help improve frontier AI models by evaluating the quality of real-world professional materials and AI-generated work. You will apply careful judgment across documents, presentations, spreadsheets, and other written content, providing feedback that helps...SuggestedHourly payRemote work$13 - $54 per hour
...Evaluate AI-generated music and lyrics across a wide range of genres, using your knowledge of the Spanish (Mexico) music scene to assess quality against detailed standards. This remote role involves working in both Spanish (Mexico) and English. Key Responsibilities...Hourly payImmediate startRemote workFlexible hours- ...reasoning and computational problem solving to improve and evaluate large language models. You will design rigorous math problems, produce clear, logically... ...How this work supports customers Accelerate frontier AI research by contributing high quality data and evaluation...Contract workFor contractorsFreelanceRemote work
$75 per hour
...Join a project focused on evaluating AI models in the architecture domain, specifically in visual document understanding and instruction-following. This role involves authoring complex, grounded tasks that include a clear ground-truth output and an objective rubric....Hourly payRemote work$18 per hour
...Role Overview Evaluate AI-generated music and lyrics across a range of genres, applying detailed quality standards in both Thai and English... ...of Thai music, language, and lyrical expression to help assess model outputs. Key Responsibilities Compare AI-generated lyrics...Hourly payFor contractorsImmediate startRemote workFlexible hours$18 - $42 per hour
...Role Overview Evaluate generative music AI outputs across a wide range of genres, applying your knowledge of Portuguese-language music and lyrics to detailed quality standards. This role combines critical listening with lyric analysis in Portuguese and English. Key...Hourly payImmediate startRemote workFlexible hours- ...Overview Apply advanced physics knowledge to help improve and evaluate large language models. You will design rigorous problems, produce clear reasoning... ...into accessible explanations while contributing to AI research projects. Key Responsibilities Design and solve...For contractorsFreelanceRemote work
- ...Evaluate generative music AI across a broad range of genres, applying your knowledge of Indonesian music and lyrics to detailed quality standards. You will work in both Indonesian and English to help assess lyric quality and authenticity. Key Responsibilities Compare...Hourly payImmediate startRemote workFlexible hours
$20 - $60 per hour
...Help train next-generation AI systems by creating rigorous, real-world evaluations that test how well advanced models learn, reason, and perform. This remote contract opportunity is open to recent graduates, advanced-degree holders, and professionals from any background...Hourly payContract workFor contractorsRemote work$70 - $110 per hour
...Role Overview Help a leading AI research team improve how advanced AI models reason about real clinical work. In this hybrid, full-time role, you will... ...define high-quality clinical tasks, model answers, and evaluation standards alongside research and program management...Hourly payFull timeFreelanceLive inRelocationRelocation package$70 - $110 per hour
...Role Overview Apply your legal practice experience to evaluate and improve AI-generated legal content and workflows. You will assess work grounded in litigation, legal drafting, and client advisory practice. Key Responsibilities Evaluate AI-generated legal memoranda...Hourly payImmediate startRemote work- Dorado is seeking an experienced Investment Banking SME to support the development, evaluation, and improvement of advanced AI models in finance. You will assess AI-generated analyses for accuracy, reasoned judgments, and data integrity, guiding model refinements and prompts...Remote job
- ...is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities...Remote jobHourly payFlexible hours
$60 per hour
...and Life Science Professionals to join their Expert Network to evaluate AI-generated science. This role allows you to work from home with... ...offers a competitive pay rate of up to $60 per hour for reviewing model responses, validating technical claims, and critiquing...Remote jobHourly payWork from homeFlexible hours- Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant...Remote jobFlexible hours
$40 per hour
A technology company in Massachusetts is seeking an R&D Biologist to join their team to train AI models by evaluating chatbot outputs against complex biology questions. Ideal candidates will hold advanced qualifications in biology or biochemistry. This position allows full...Hourly payFull timePart timeRemote work- ...Help improve large language models by creating rigorous Biology problems, evaluating model reasoning, and translating complex concepts into clear explanations.... ...paid leave. Opportunity to contribute to advanced AI projects involving leading large language model organizations...Contract workFor contractorsRemote work
$60 - $90 per hour
...Author and run rigorous, multi-step machine learning evaluation tasks for a leading generative AI research team. You will take high-level research ideas... ...experiments, and analyze results to determine where frontier models succeed or fail. Typical tasks require one to two days...Hourly payFull timeFreelanceRemote work$60 - $90 per hour
...Role Overview Help advance frontier AI research by creating rigorous, real-world data analysis evaluations for generative AI models. You will design and complete complex analytical tasks that mirror practical research work, then use your reference analyses to assess where...Hourly payFull timeFreelanceRemote work$80 - $135 per hour
...09.26574v3), a frontier research-level physics benchmark. The role produces fully human-verified reference data used to evaluate large language model performance on frontier physics reasoning. Work includes solving CritPt research-level problems end-to-end, auditing other...Hourly payRemote work10 hours per week$100 per hour
...Role Overview Apply deep domain expertise to train and evaluate next-generation AI systems by producing, refining, and validating high-quality,... ...data. This part-time contractor role focuses on improving model outputs through careful content review, prompt refinement,...Hourly payPart timeFor contractorsRemote work$100 - $150 per hour
...Role Overview Evaluate how well AI systems perform real-world technical sales work by defining excellence and judging completed work samples. Rather than producing sales deliverables yourself, you will create task-specific grading rubrics and assign scores to AI-generated...Hourly payRemote work$100 per hour
...the performance of large language models on finance tasks. You will work with AI researchers to identify model weaknesses... .... Key Responsibilities Evaluate LLM performance in finance areas where... ...: CFA (Level I, II, or III), CA, CPA, or an MBA in Finance. Work Terms...Hourly payContract workFor contractorsFreelanceRemote work10 hours per weekFlexible hours$11 - $19 per hour
...Role Overview Evaluate AI-generated music in Tamil and English across a wide range of genres, judging musicality, creativity, adherence... ...performance, and production quality. You will listen critically, compare model outputs, annotate musical characteristics, and check lyrics and...Hourly payImmediate startRemote workFlexible hours$100 per hour
A leading technology firm is seeking finance experts to enhance AI models. Responsibilities include evaluating performance in capital markets and creating assessment rubrics. Candidates should have 2+ years in finance fields like investment banking and possess strong financial...Remote jobHourly pay10 hours per week- Dorado is seeking a Physics Specialist to contribute deep scientific expertise to AI model evaluation. You will craft and assess challenging physics problems, probe model reasoning at the frontier, and help identify where models fail under rigorous scientific scrutiny....Remote job
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Tax Advisor, CPA for AI Model Evaluation. Be the first to apply!



