Generalist for AI Training and Evaluation
$40 - $70 per hourSaidGig
Help evaluate and improve advanced AI systems by applying careful human judgment to everyday questions and general reasoning. This standing opportunity is for generalists interested in AI training and evaluation work; it is not tied to a specific active project and remains open as new projects begin. Role Overview
Projects can begin on short notice. Qualified generalists are maintained in a talent pool and invited to relevant project-specific opportunities when they become available.
Key Responsibilities- Write prompts and reference answers for everyday questions with objectively correct answers, including the context a strong response must consider.
- Assess AI responses for correctness, instruction-following, completeness, and unnecessary padding.
- Identify responses that are confidently incorrect, internally contradictory, or answer a different question than the one asked.
- Careful reading and sound judgment gained through a professional or academic background in any field.
- Clear written communication and the ability to explain reasoning precisely.
- Comfort working with ambiguous tasks and flagging unclear instructions.
- No specific credential is required; the key requirement is the ability to identify and clearly explain why an answer is incorrect.
- Remote, hourly engagements.
- This pool is intended for people whose experience does not fit a single specialist area or spans several areas.
- Project invitations depend on current client demand and may arrive within a week or several months.
- This listing itself does not make hiring decisions or issue acceptance or rejection outcomes. Any hiring decision is made through the specific project opportunity.
- Each project invitation identifies its rate, expected hours, and client.
Most contracts in this field have paid $40 to $70 per hour, with rates set by project scope and depth.
EligibilityComplete registration once to be eligible for matching across all networks for which you qualify. Completing the domain-expert interview and joining each relevant network may improve the likelihood of being matched.
Application Process- Submit a resume and confirm your work location.
- Complete a short AI interview, approximately 20 minutes.
- Your background is verified before you are added to the generalist pool for future matching.
$50 per hour
...Generalist AI Evaluation Specialist is a remote review track for evaluating AI outputs across generalists specialist operations workflows. Reviewers... ...and document the right next step so the modeling team can train on it. Why this role matters Generalists specialist...TrainingRemote jobFor contractorsWork experience placement10 hours per week$20 - $30 per hour
...Role Overview Evaluate images to train next-generation AI systems by applying careful visual review and clear written reasoning. You will support a client image assessment project as a remote contractor, producing high-quality, high-volume image assessments that shape...TrainingHourly payFor contractorsRemote work$60 per hour
...Senior Lua Developer (Roblox) – AI Code EvaluationAn enterprise client is seeking experienced Lua Developers to support AI model training through technical evaluations, scripting tasks, and code quality reviews. This opportunity is ideal for developers who have strong...TrainingWeekly payContract workFreelanceRemote work10 hours per week$14.5 per hour
...AI Web Search Evaluator Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant,... ...brings together a curated global community of over 500,000 AI training and domain experts...and we'd like you to join us! Job...TrainingBi-weekly payHourly payPart timeImmediate startRemote workWork from homeFlexible hours$14.5 per hour
...Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are... ...brings together a curated global community of over 500,000 AI training and domain experts. As a Web Search Evaluator you will...TrainingHourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours$70 - $120 per hour
...Role Overview Apply your data judgment to help evaluate and improve advanced AI systems. Your professional standards and subject-matter expertise... ...listing for analytics and data practitioners interested in AI training and evaluation work, rather than a guarantee of an...TrainingHourly payImmediate startRemote work- ...About VTI Aerospace VTI Aerospace builds AI-powered perception and pilot assist... ...Role As a Senior Software Engineer – Evaluation, you will design and implement systems that... ...improvements to data collection and model training, helping bring AI-powered aviation tools...Training
$229.9k - $262.4k
Senior Lead AI Engineer (SDK's: Gen AI Evaluation and MCP) Overview: At Capital One, we are creating responsible and reliable AI systems, changing... ...support AI software components including foundation model training, large language model inference, similarity search,...TrainingFull timePart timeLocal area$150 - $180 per hour
...Prolific in Virginia Beach is seeking Mental Health Professionals to train and evaluate AI models. The role involves reviewing AI responses to psychological scenarios and completing related tasks, offering competitive pay up to $150-180/hr. You must have a valid mental...TrainingRemote workWork from home- ...Prolific is seeking Mental Health Professionals in Nashville, TN, to help train and evaluate AI models. Candidates will review AI-generated responses, complete training tasks, and provide insights based on their understanding of psychological theory. Successful participants...TrainingHourly payRemote workWork from homeFlexible hours
- ...extract actionable insights. Develop and evaluate novel algorithms and methodologies for... ...data processing, feature engineering, model training, and performance optimization.... .... We may use artificial intelligence (AI) tools to support parts of the hiring process...TrainingTemporary workRelocation package
$141.02k - $204.53k
...your future. Responsibilities As the Senior AI/ML Engineer - Validation & Evaluation within AI Validation & Monitoring (AVM), you will provide... ..., automation-bias risk, patient-facing behavior, training, accessibility, guardrails, task boundaries, safe refusal...TrainingFull timeWork at officeRemote workFlexible hoursWeekend work$8 - $65 per hour
...Prolific is looking for Mental Health Professionals in Houston, Texas, to train and evaluate cutting-edge AI models. The role offers flexible hours and competitive pay rates ranging from $8 to $65 per hour for AI tasks. Candidates must have a verified status as a Mental...TrainingHourly payRemote workFlexible hours- ...AI Evaluation Specialist Role Type: Contractor Location: Remote (US, CA, UK, IE, AU, NZ) Micro1 is engaging AI Evaluation Specialists... ...the quality of AI assistant outputs for an enterprise AI training initiative. In this role, you'll apply your expertise to help...TrainingFor contractorsRemote work
$20 - $80 per hour
...Role Overview Help improve next-generation AI systems by supplying precise, real-world evaluation, annotation, and feedback. This remote contractor role focuses... ...Preferred Qualifications Experience with AI training, machine learning data annotation, human-in-the-loop...TrainingHourly payContract workFor contractorsRemote work$164.52k - $246.78k
...building a connected, end-to-end Enterprise AI engine - uniting data foundations, AI... ...Immunology, Infectious Disease) can access, evaluate, and mobilize the right external... ...be as comfortable interrogating a model's training methodology and validation evidence as they...TrainingHourly payContract workTemporary workWork at officeFlexible hours3 days per week$70 - $90 per hour
...Role Overview Help evaluate Neuron Kernel Interface development tasks that support the training and evaluation of advanced AI models. You will assess kernel quality, numerical correctness, CUDA-to-NKI migration fidelity, and whether implementations are well suited to...TrainingHourly payRemote work- ...shape one of the world's most widely used AI assistants, powered by our next-... ...macOS, watchOS, and visionOS. Our Speech Evaluation team sits at the center of Apple's ASR, TTS... ...datasets for machine learning evaluation or training. Working knowledge of statistics as applied...Training
- Rex.zone is seeking a remote, full-time Senior AI Data Annotation role supporting training-data quality for modern AI/ML systems. You will deliver high-precision human feedback across RLHF, LLM evaluation, prompt evaluation, and QA evaluation to drive measurable model performance...TrainingRemote jobFull time
$60 - $100 per hour
Gridnaut Recruiting is hiring a remote Legal Domain Expert — AI Training & Evaluation contractor (pay $60-$100/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: JD from accredited U.S. law school; 5+ years post-qualification legal practice;...TrainingFor contractorsRemote workRelocation$350k
...Engineer The mission of Thinking Machines is to build AI that extends human will and judgment. We are training frontier models with Inkling, developing Tinker to... ...build it. About the Role We're looking for generalist infrastructure and systems engineers to help build...TrainingVisa sponsorshipWork visaRelocation packageFlexible hours$70 per hour
...Position: Video related professional for AI training Type: Contract Compensation: $20 - $70/hour Location: Remote Commitment: 10-40 hrs... ..., annotate, and assess video datasets for model training and evaluation. Review and critique AI-generated video outputs to ensure...TrainingContract workRemote work- Rex.zone is hiring a Senior Data Annotator to produce and review high-quality training data for AI systems across NLP, computer vision, and multimodal evaluation. This is a Remote, full-time role focused on annotation guidelines compliance, content safety labeling, and...TrainingRemote jobFull time
- ...About Turing Turing is one of the world’s fastest-growing AI companies, accelerating the advancement and deployment of powerful... ...language models (LLMs) through high-quality human feedback, evaluation, and training data. Role Overview We are seeking Board Game...TrainingContract workTemporary workFor contractorsFreelanceRemote work
$127k - $223k
...Description Job Description Waabi, founded by AI visionary Raquel Urtasun, is the leader... ...way. To learn more visit: The Evaluation Algorithms team is responsible for building... ...discovery of interesting scenarios for training and evaluation. - Develop and maintain...TrainingFull timeWork at officeWork from homeFlexible hours$144.7k - $221.4k
...experiences. About the Organization The Evaluation team builds and evolves the evaluation... ..., treating road testing, data mining, training, and metrics as first-class use cases in... ...systems. Experience leveraging AI-assisted development and analytics tools...TrainingFull timeWork at officeLocal areaWork from homeRelocationRelocation packageFlexible hours- ...About Turing Turing is one of the world’s fastest-growing AI companies, accelerating the advancement and deployment of powerful... ...language models (LLMs) through high-quality human feedback, evaluation, and training data. Role Overview We are looking for experienced...TrainingContract workTemporary workFor contractorsFreelanceRemote work
- ...Design and run capability, uplift, and safety evaluations for cyber-relevant model risks.... ...an equivalent combination of education, training, and experience. Preferred qualifications... ...security or security-research experience, AI security benchmark development, adversarial...TrainingFull timeWork at officeVisa sponsorshipFlexible hours
$80 - $150 per hour
...Prolific is seeking Medical Doctors in Las Vegas to help train and evaluate AI models. Successful candidates will join as Domain Expert participants, reviewing AI-generated responses and providing professional insights. The role offers flexible hours and pay rates between...TrainingHourly payRemote workFlexible hours$229.9k - $262.4k
...Overview Senior Manager, AI Engineer (Gen AI Platform Services: Agentic AI, Guardrails, Evaluation) Overview : At Capital One, we are creating responsible... ...software components including foundation model training, large language model inference, similarity...TrainingFull timePart timeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Generalist for AI Training and Evaluation. Be the first to apply!



