Marketing Specialist for AI Model Evaluation
$60 - $80 per hourSaidGig
Role Overview
Apply your marketing expertise to help develop advanced large language models. In this role, you will bring practical brand, growth, and campaign judgment to AI training data, partnering with research and engineering teams to improve model performance on marketing-focused tasks.
Key Responsibilities
- Help research and engineering teams address knowledge gaps in brand strategy, growth marketing, and campaign reasoning.
- Create challenging, relevant marketing tasks and produce accurate, well-reasoned solutions based on real-world marketing practice.
- Assess AI model outputs using structured rubrics and provide clear written feedback on correctness, judgment, and reasoning quality.
- Develop and improve evaluation guidelines and scoring rubrics for marketing tasks.
- Work with other subject-matter experts to maintain consistent, accurate training data.
Qualifications
- At least 8 years of dedicated professional marketing experience, such as brand strategy, growth marketing, or performance marketing, at a recognized top-tier organization comparable to P&G, Unilever, Nike, Ogilvy, WPP, or Omnicom.
- Hands-on experience evaluating LLM or AI model outputs against rubrics or structured scoring criteria. This is required and should be described in your application.
- Demonstrated career progression, such as Marketing Manager to Senior Manager to Director or VP of Marketing.
- Strong verbal and written communication, problem-solving, and interpersonal skills.
Work Terms
- Hourly W-2 employment, with placement on the extended workforce of a leading AI lab.
- Reliable weekday availability of at least 35 hours per week is required.
- Location: United States.
Compensation
$60 to $80 per hour.
Equal Employment Opportunity
Employment decisions are made without discrimination based on race, religion, color, national origin, sex, including pregnancy, childbirth, reproductive health decisions, or related medical conditions, sexual orientation, gender identity or expression, age, protected veteran status, disability, genetic information, political views or activity, or any other legally protected characteristic.
$60 - $80 per hour
...Role Overview Contribute marketing subject-matter expertise to a GenAI team building foundational AI models. You will create realistic marketing tasks, evaluate model outputs against structured rubrics, and advise research and engineering teams on brand strategy, growth...SuggestedHourly payWeekday work$60 - $80 per hour
...Apply your marketing expertise to help train and evaluate AI systems through future remote contract projects aligned... ...active role. Qualified marketing specialists may be considered for relevant opportunities... ...Train and evaluate AI models in marketing-related work. Create...SuggestedHourly payContract workRemote work$60 - $90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation...SuggestedFull timeContract workSummer workRemote work- Cincinnatus LLC is recruiting for an Insurance Subject‑Matter Expert to join a leading AI lab's GenAI team in New York. This W-2 role involves evaluating insurance tasks and guiding model development for high‑quality underwriting judgments, with placement at the client...SuggestedWeekday work
- Cincinnatus LLC is placing finance SMEs at a leading AI lab to critically evaluate AI model outputs for financial tasks. The role focuses on rigorous, rubric-based assessment of model performance and constructing finance-focused evaluation frameworks. Candidates should...SuggestedWeekday work
- ...Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant...Remote workFlexible hours
- ...is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities...Hourly payRemote workFlexible hours
$136.44k - $265.11k
...are rebuilding biotech for the AI era.When a breakthrough is... ...structured data, and run AI agents and models directly in their workflows.... ....You’ll build the datasets, evaluations, and systems that help close... ...commuter, and more.Benchling takes a market-based approach to pay. The...Work at officeLocal areaMonday to FridayShift work$70 - $90 per hour
...Role Overview Help evaluate Neuron Kernel Interface development tasks that support the training and evaluation of advanced AI models. You will assess kernel quality, numerical correctness, CUDA-to-NKI migration fidelity, and whether implementations are well suited to...Hourly payRemote work$100 per hour
...the performance of large language models on finance tasks. You will work with AI researchers to identify model weaknesses in areas such as capital markets, portfolio management, trading, investment... .... Key Responsibilities Evaluate LLM performance in finance areas where...Hourly payContract workFor contractorsFreelanceRemote work10 hours per weekFlexible hours$28 - $60 per hour
...Evaluate AI-generated music and lyrics across a wide range of genres, applying your knowledge of the Dutch music scene and strong editorial judgment to detailed quality standards. Key Responsibilities Assess AI-generated music and rate it against detailed quality criteria...Hourly payFor contractorsImmediate startRemote workFlexible hours- ...Role Overview Apply research-grade expertise to help evaluate and improve AI reasoning across technical and humanities disciplines. This remote contractor role supports AI-model training through rigorous analysis, high-quality feedback, and clearly articulated academic...Hourly payFor contractorsRemote work
$70 - $80 per hour
...drug safety expertise to help improve next-generation AI systems through high-quality evaluations, safety-report analysis, and structured feedback. This... ...equivalent. Experience in development-stage and post-marketing safety reporting, QPPV or deputy QPPV roles, safety...Hourly payFor contractorsRemote work$70 - $90 per hour
...Review GPU and accelerator kernel development tasks that support the training and evaluation of advanced AI models. This role focuses on assessing task quality, numerical correctness, completeness, fair performance benchmarking, appropriate scope, and whether kernels...Hourly payRemote work$15 per hour
...Evaluate AI-generated music and lyrics in Malayalam and English, helping assess outputs across a broad range of genres against detailed quality standards. Key Responsibilities Compare AI-generated lyrics with published songs to identify similarities. Rate lyrics...Hourly payImmediate startRemote workFlexible hours$20 - $60 per hour
...Role Overview Help train and evaluate next-generation AI systems by creating rigorous, real-world assessments that test how advanced models learn, reason, and perform. This remote contract opportunity is open to recent graduates and other researchers and writers with...Hourly payContract workFor contractorsRemote work- Verita AI is seeking an Applied AI Researcher to work with clients on model evaluation and data strategy. You will assess model performance, identify failure modes, and design data-driven solutions, collaborating with operations and engineering to implement scalable data...
- Xperteez Technology seeks a data science and ML-focused specialist to evaluate AI model outputs across statistics, ML, and quantitative problems. You will assess quality, identify errors, and provide actionable feedback to improve model capability. Responsibilities include...
- Job Description - Member of Technical Staff (Language Model Evaluations) Location: San Francisco (preferred), Sydney, Melbourne, Brisbane About... ...Analysis Artificial Analysis is the leading independent AI benchmarking company. We support labs, engineers and enterprises...
- We are seeking an expert to evaluate and improve our AI models through comprehensive testing and analysis. You will be responsible for designing evaluation frameworks, conducting model assessments, and providing actionable insights for model improvement. Key Responsibilities...
$11 - $19 per hour
...Evaluate AI-generated music and lyrics across a broad range of genres, applying your Bengali music expertise to help assess quality, originality, and natural expression. Key Responsibilities Compare AI-generated lyrics with published songs to identify similarities....Hourly payImmediate startRemote workFlexible hours$100 - $150 per hour
...Shape how advanced AI models handle real-world legal work by applying senior-level legal judgment to the design, review, and evaluation of legal knowledge tasks. You will work closely with... ...researchers and adjacent-domain specialists to calibrate consistent standards and...Hourly payFull timeInternshipLive inRelocationRelocation package$50 - $75 per hour
A leading tech company based in Australia is seeking an AI Model Evaluator on a contract basis. The role involves evaluating AI-generated responses, writing prompts, and providing justifications based on specific criteria. Ideal candidates will hold a Master's degree in...Hourly payContract work- Google DeepMind seeks a Senior Product Manager embedded in Gemini research and model training. You will read evaluations, analyze model outputs, and make judgment calls on quality alongside researchers, translating user needs into product priorities and feedback loops for...
- Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code). You will author original, executable research problems that today's frontier models cannot solve. The role emphasizes material sourcing, prompt design, and robust grading criteria against...Part timeImmediate start
$45 - $55 per hour
...institutions. In 2025, we started Handshake AI and built the fastest-growing AI data... ...frontier AI lab researchers to create evaluations, publish benchmarks, and push the boundary... ...advancement. Frontier AI labs currently improve model capabilities with various data-intensive...- ...reasoning and computational problem solving to improve and evaluate large language models. You will design rigorous math problems, produce clear, logically... ...How this work supports customers Accelerate frontier AI research by contributing high quality data and evaluation...Contract workFor contractorsFreelanceRemote work
$60 - $80 per hour
...Overview Join a Mathematician Expert Network that connects mathematicians with AI labs and companies to shape and evaluate cutting-edge AI in mathematics. Experts contribute domain expertise to model training and evaluation, create real-world tasks and deliverables, and...Hourly payContract workImmediate startRemote work$110 per hour
...Apply your clinical expertise to help shape and evaluate advanced AI systems through remote, hourly contract opportunities aligned with your... ...needs arise. Key Responsibilities Train and evaluate AI models used in medicine. Create tasks and deliverables based on...Hourly payContract workRemote work$100 per hour
...and business expertise to improve AI-generated content for a customer-facing evaluation and optimization project. You... ...documentation, strategy reports, market analyses, business cases, and executive... ...and refine large language model prompts using structured problem-...Hourly payPart timeFor contractorsRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Marketing Specialist for AI Model Evaluation. Be the first to apply!
- email marketing strategist United States
- event marketing specialist United States
- market development representative United States
- marketing automation consultant United States
- marketing proposal coordinator United States
- marketing proposal specialist United States
- salesforce marketing cloud specialist United States
- entry level marketing representative United States
- marketing representative United States
- technical marketing specialist United States




