Professional Gamer for AI Model Evaluation
$10 - $14 per hourSaidGig
Apply your PC gaming expertise to generate high-quality gameplay data that helps train and evaluate next-generation AI systems. In this contractor role you will play curated PC titles, capture synchronized gameplay video and telemetry, annotate events, and perform quality assurance to ensure the dataset is accurate and actionable. No prior AI experience is required, your domain knowledge and gaming skill are the primary qualifications.
Key Responsibilities- Play selected PC games while recording per-frame synchronized gameplay video and capturing keyboard and mouse telemetry.
- Collect and submit detailed engine-level signals and in-game event data according to project guidelines.
- Annotate gameplay scenarios and events with high accuracy and contextual detail.
- Perform hands-on quality assurance to identify and document gameplay issues, bugs, or data anomalies.
- Maintain consistent participation, completing daily gaming sessions and submitting data on schedule.
- Provide written and verbal progress updates and feedback to project managers and technical stakeholders.
- Follow all project protocols for data security, privacy, and compliance.
- Required skills: gaming, quality assurance, annotation.
- Demonstrated expertise in PC gaming, ideally with competitive or professional experience.
- Experience using gameplay recording tools, telemetry capture, or similar technical setups.
- Background in QA testing or detailed annotation within gaming or related fields is preferred.
- Strong written and verbal communication skills, with an emphasis on clarity and proactive updates.
- Familiarity with game mechanics, strategy, and meta across multiple PC titles.
- Ability to provide actionable feedback on user experience, bugs, and gameplay flow.
- Self-motivated, reliable, and comfortable working independently in a remote contractor engagement.
- Role type: Contractor.
- Location: Remote.
- Daily commitment: consistent participation with gaming sessions of at least 2 to 3 hours per day.
- Timely submission of recorded video, telemetry, engine signals, annotations, and QA reports as specified by project guidelines.
- Strict adherence to project protocols for data security, privacy, and compliance.
$18.00 per hour.
How to ApplyIf you are interested, follow the application instructions on this listing to indicate your interest and submit any requested information. The role requires the skills and commitments described above, and selected contractors will receive further onboarding details and project guidelines upon engagement.
About the Project PartnerThe work supports an AI data lab that transforms subject matter expertise into high-quality training data and evaluations for frontier AI models. Your gameplay and annotations will directly influence how models learn, reason, and perform.
$65 - $90 per hour
...deep financial judgment to help improve foundational AI models. In this role, you will create and evaluate finance-focused work that strengthens AI systems''... ...Qualifications At least 8 years of dedicated professional finance experience, such as investment banking,...SuggestedHourly payWeekday work$60 - $80 per hour
...building foundational large language models by applying deep insurance domain expertise to create, evaluate, and refine training data. You... ...claims practice. Evaluate AI model outputs against... ...Qualifications At least 8 years of professional experience in insurance, such...SuggestedHourly payWeekday work$60 - $100 per hour
...Role Overview Help advance frontier AI models by bringing senior insurance and actuarial judgment to the evaluation of real-world insurance work. You will work directly... ...and program management team, translating professional standards into tasks, solutions, benchmarks,...SuggestedHourly payFull timeLive inRelocationRelocation package$110 - $150 per hour
...Role Overview Drive how frontier AI models reason about real-world private equity and... ...that appear plausible but would not pass professional review. Author high-quality instruction... ...Design challenging finance tasks and evaluation sets, and help create finance-specific...SuggestedHourly payFull timeLive inRelocationRelocation package$60 - $80 per hour
...Apply deep retail domain expertise to help build and evaluate foundational generative AI models. You will design realistic retail tasks, produce authoritative... ...Qualifications Eight or more years of focused professional experience in retail, such as merchandising, category...SuggestedHourly payWeekday work$60 - $90 per hour
...Help shape how AI systems evaluate and preserve the timing, emotion, micro-expressions, and physical nuance of human and character performances... ...-time role supports the development of performance-transfer models by defining high-quality evaluation standards and helping...GamesHourly payPart time$60 - $90 per hour
...expertise to shape how a performance transfer model is evaluated, ensuring actor timing, emotional... ...model outputs for a leading generative AI research effort. Key... ...New York, NY. Candidacy requires the professional credits and experience listed above; roles...GamesHourly payPart timeFreelance$208k - $300k
...Machine Learning Engineer - Model Evaluations, Public Sector The Public Sector ML team at Scale deploys advanced AI systems—including LLMs, agentic models, and multimodal... ...collect, retain and use personal data for our professional business purposes, including notifying...Full time$60 per hour
...Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing...Hourly payRemote workWork from homeFlexible hours$70 - $110 per hour
...Help advance frontier AI systems by bringing rigorous materials... ...engineering judgment to the evaluation, design, and improvement of... ...like in practice and ensure model outputs can withstand technical... ...strongly preferred. Hands-on professional use of large language models...Hourly payFull timeLive inRelocationRelocation package$60 - $80 per hour
...Help shape the training and evaluation of foundational large language models by applying real-world expertise in brand... ...rigorous marketing judgment to AI tasks, model assessments, and training... ...-reasoned solutions grounded in professional marketing practice. Assess AI model...Hourly payWeekday work$60 per hour
Prolific, located in Arizona, is seeking Biology Experts and Life Science Professionals to join their Expert Network. This role involves evaluating and training AI models with real scientific expertise. Successful candidates will review AI-generated content for accuracy...Hourly pay$20 - $60 per hour
...Help train next-generation AI systems by creating rigorous, real-world evaluations that test how well advanced models learn, reason, and perform. This remote contract opportunity... ...graduates, advanced-degree holders, and professionals from any background with strong research...Hourly payContract workFor contractorsRemote work$100 per hour
...finance expertise to help improve AI-driven financial applications... .... Develop, refine, and evaluate prompts related to financial... ...financial data, reports, and model outputs using detailed... ...robustness, and adherence to professional financial standards. Conduct...Hourly payPart timeFor contractorsRemote work$70 - $110 per hour
...Role Overview Help a leading AI research team improve how advanced AI models reason about real clinical work. In... ...clinical tasks, model answers, and evaluation standards alongside research and... ...clinical decisions. Hands-on professional use of large language models and...Hourly payFull timeFreelanceLive inRelocationRelocation package- Prolific is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities...Remote jobHourly payFlexible hours
$60 per hour
Prolific is seeking Biology Experts and Life Science Professionals to join their Expert Network to evaluate AI-generated science. This role allows you to work... ...competitive pay rate of up to $60 per hour for reviewing model responses, validating technical claims, and...Remote jobHourly payWork from homeFlexible hours- Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant...Remote jobFlexible hours
$100 per hour
...Apply deep domain expertise to train and evaluate next-generation AI systems by producing, refining, and... ...contractor role focuses on improving model outputs through careful content... ...prompts to guide model behavior, using professional writing and technical documentation skills...Hourly payPart timeFor contractorsRemote work$100 - $150 per hour
...Role Overview Evaluate how well AI systems perform real-world technical sales work by defining excellence and judging completed work samples... ...Process Submit your resume or a summary of relevant professional experience to begin the review. Qualified candidates may be...Hourly payRemote work- A leading AI company is seeking a legal professional for a contractor role focused on evaluating AI model outputs in legal contexts. Candidates must hold a Juris Doctor (J.D.) and have more than 3 years of experience in law. The role involves reviewing complex legal hypotheticals...For contractors10 hours per week
$65 - $105 per hour
...Overview Help advance frontier AI systems by applying deep engineering... ...benchmarks used to assess how models reason about real software and... ...engineering tasks that reflect professional practice. Design rigorous engineering evaluation sets and contribute to engineering...Hourly payFull timeFreelanceInternshipLive inRelocationRelocation package$65 - $105 per hour
...Help improve how frontier AI models reason about real-world life sciences research. In... ...will apply deep scientific judgment to evaluate research tasks and model outputs, define... ...experience using large language models in professional work and the ability to distinguish...Hourly payFull timeLive inRelocationRelocation package$100 per hour
...the performance of large language models on finance tasks. You will work with AI researchers to identify model... ...systems. Key Responsibilities Evaluate LLM performance in finance areas... ...Qualifications Minimum 2 years of professional experience in one or more of the...Hourly payContract workFor contractorsFreelanceRemote work10 hours per weekFlexible hours$60 - $90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation...Full timeContract workSummer workRemote work- OpenTrain AI seeks a Video Game AI Evaluation Expert on a contractor, part-time basis. You will design prompts for game development, esports, platforms, and communities; assess AI responses for accuracy, completeness and nuance; and develop evaluation datasets with clear...GamesRemote jobPart timeFor contractors
$400 per month
...Mercor is partnering with a leading AI research lab to support a Frontier... ...project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments... ...strengths and weaknesses. Apply professional engineering judgment to realistic...- ...Opportunity Deepgram is looking for a Senior Software Engineer - Model Evaluation & AI Systems to join the team responsible for validating the... ...related field, or equivalent experience. ~5+ years of professional software or QA engineering experience, with a track record...Full time
$400 per month
...Mercor is partnering with a leading AI research lab to support a Frontier... ...Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical... ...their strengths and weaknesses. Apply professional engineering judgment to realistic...$224k - $356.5k
...people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts... ...computing. As a Senior / Principal Deep Learning Engineer — Model Evaluation & AI Systems, you will play a meaningful role in crafting the...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Professional Gamer for AI Model Evaluation. Be the first to apply!




