People Ops Evaluator - Frontier Model Review (Remote)
AI Trainer Jobs
- Remote job
AI Trainer Jobs is seeking a remote contractor to evaluate people ops / recruiting prompts and responses against a evolving quality rubric. You will compare model outputs, label edge cases, and provide structured feedback for retraining. Responsibilities include evaluating rubric-compliance, selecting stronger answers with rationale, and annotating issues like hallucinations or unsafe content. Strong written reasoning and 10+ hours weekly async availability are required. #J-18808-Ljbffr AI Trainer Jobs
- Mercor is seeking expert Evaluators in People ops / recruiting to review AI-generated work products for accuracy, rigor, and domain quality. You will apply... ...subject-matter expertise to grade outputs. This is a remote, hourly engagement. You will evaluate artifacts...Remote jobHourly pay
- AI Trainer Jobs is seeking an Illustration Quality Evaluator for a remote contractor role. Review illustration quality evaluation prompts and responses against a published rubric, compare paired outputs, and provide structured feedback to support retraining efforts. You...Remote jobPart timeFor contractors
- ...Frontier Model Misuse Risk Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety...Remote jobHourly payFor contractors10 hours per week
- ...Annotator—Product Management & Marketing to remotely review evaluation prompts and responses against the company... ..., and provide structured feedback the modeling team can use to retrain. As a contractor, you will assess frontier model outputs, judge benchmark failures,...Remote jobFor contractors
- Surgical Planning Safety Evaluator is a remote evaluation track for reviewing surgical planning safety evaluation prompts... ...kind of structured feedback the modeling team can use to retrain. AI data... ...cases for AuraOne Human Data. Review frontier model outputs. #J-18808-Ljbffr...Remote job
- Receipt and Invoice Understanding Model Evaluator is a remote evaluation track for reviewing receipt and invoice understanding model evaluation prompts and responses... ...the modeling team can use to retrain. Category: Frontier Model Evaluation · Pay: Hourly rate confirmed...Remote workHourly payFor contractors10 hours per week
$80 per hour
...Type: Contract Compensation: $80/hour Location: Remote Role Responsibilities Use frontier AI coding agents to complete and evaluate complex data engineering tasks. Review model-generated implementations involving ETL pipelines , data warehouses...Remote workContract workSummer work- AuraOne is seeking a remote Generalist for Data Annotation to evaluate prompts and responses against a quality rubric. You’ll help turn outputs into... ..., and regression cases for Human Data, reviewing frontier model outputs and calibrating evaluators. Responsibilities...Remote job
$50 per hour
Combinatorics Model Evaluator is a remote review track for evaluating AI outputs across combinatorics model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method...Remote workHourly payFor contractors- AuraOne is seeking a Latin Bilingual Expert for a remote evaluation track. Reviewers compare paired outputs to a quality rubric, label edge cases, and generate structured feedback to retrain models. Responsibilities include evaluating outputs against rubrics, tagging issues...Remote workFor contractors10 hours per week
$295k
...edge foundation AI models and end-to-end... ...training and deploying frontier models for... ...us!Role Overview:Evaluation is critical to making... ...spent dozens of hours reviewing complex data and LLM... ...role can be based remotely or from one of our... ...exceptional people regardless of locations...Remote workFull timeWork at officeLocal areaHome office$160k - $327k
Grade 12- People and Organization DirectorBusiness Consulting Director - 22202150Who we... ...and constructive feedback through regular evaluations.• Work-Life Balance Support• Strong and... ...roles. The starting pay range for this remote role is $160,000.00-327,000.00. This range...Remote workTemporary workFlexible hours- ...Licensed Pharmacists to assist in AI model training and evaluation from a home office. Successful candidates... ...tests to assess suitability. Experts review AI-driven pharmaceutical scenarios,... ...and rationales to the training data. Remote work and flexible hours are offered...Remote workHome officeFlexible hours
$295k
...cutting-edge foundation AI models and end-to-end products... ...training and deploying frontier models for enterprises... ...London. We embrace a remote-friendly environment,... ...about hiring exceptional people regardless of locations... ...our recruiters may review or consider.Beware of Scams...Remote workFull timeWork at officeLocal areaHome office$219k - $351k
...dedicated to empowering people to be their true... ...business. As models scale past what any... ...Flexible Work policy; remote/hybrid option... ...access patterns of frontier open-weight models... ...Lead architecture reviews and deep-dive design... ...candidate is evaluated fairly and holistically...Remote workWork at officeFlexible hours$85 per hour
...Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-generated implementations involving cloud platforms ,...Remote workContract workSummer work$295k
...cutting-edge foundation AI models and end-to-end products... ...training and deploying frontier models for enterprises... ...London. We embrace a remote-friendly environment,... ...about hiring exceptional people regardless of locations... ...our recruiters may review or consider.Beware of Scams...Remote workFull timeWork at officeLocal areaHome office- ...Jobs is seeking a Grocery Shopper Expert contractor for remote work in the US. You will review AI outputs across grocery shopper operations, grade workflow... ..., and stakeholder fit, and document next steps for model training. Experience with grocery shopper operations and...Remote jobFor contractors10 hours per week
$45 - $70 per hour
...Recruiting is hiring a remote Behavioral Health Expert, AI Safety and Model Evaluation contractor (pay $45-$70/hr). Contribute to frontier AI research and evaluation... ...conversations.; Review conversations on relationships... ...3+ years supporting people in a mental health setting...Remote workTemporary workPart timeFor contractors- AI Trainer Jobs seeks a skilled reviewer for the Competitive Programmer track. This remote contractor role evaluates production code, debugging traces, and AI outputs for correctness... ...reproduce failures and explain fixes so the modeling team can target gaps. Engineering model...Remote jobFor contractors
- AI Trainer Jobs seeks a Research Physics Expert for a remote review track evaluating AI outputs across physics reasoning, calculations, and research... ...key results, and document the correct method so the modeling team can train on it. Graduate-level physics training or...Remote jobFor contractors10 hours per week
- ...Frontier Model Misuse Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair... ...this role matters Adversarial evaluation is how AuraOne hardens AI...Remote jobHourly payFor contractors10 hours per week
- FitOn is looking for a licensed Nurse Practitioner to serve as a Clinical Chart Reviewer. This remote role requires conducting patient evaluations and reviewing clinical documentation. The ideal candidate will have a strong focus on ensuring effective and evidence-based...Remote job
$15 - $20 per hour
...Contract Compensation: $15-$20/hour Location: Remote Role Responsibilities Conduct fact-... ...external tools. Generate high-quality human evaluation data by identifying response strengths,... ...tone, and completeness of responses. Ensure model responses align with expected...Remote workContract workSummer work- ...Medical Document OCR Model Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning...Remote jobHourly payFor contractors10 hours per week
- AuraOne is seeking a Combinatorics Model Evaluator for a remote review track evaluating AI outputs across combinatorics model research, reasoning, and workflows. Reviewers grade derivations and reproduce results to train the modeling team. This contractor role offers hourly...Remote jobHourly payFor contractors
- ...Policy Preference Reward Model Evaluator is a remote review track for evaluating AI outputs in policy review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can...Remote jobHourly payFor contractors10 hours per week
- ...Slide Deck Understanding Model Evaluator is a remote evaluation track for reviewing slide deck understanding model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback...Remote jobHourly payFor contractors10 hours per week
- ...Quant Finance Reasoning Model Evaluator is a remote review track for evaluating AI outputs across finance and risk workflows. Reviewers grade calculations, narrative reasoning, and policy adherence; flag compliance and reconciliation issues; and document the correct treatment...Remote jobHourly payFor contractorsWork experience placement10 hours per week
- ...engagement to train AI systems using project management materials. You will review model outputs against guidelines, explain decisions clearly, and provide evidence-based judgments. The role involves evaluating sources, maintaining consistency across tasks, and delivering...Remote jobContract workPart time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to People Ops Evaluator - Frontier Model Review (Remote). Be the first to apply!
- ai evaluator New York, NY
- education evaluator New York, NY
- work from home web search evaluator New York, NY
- evaluator New York, NY
- program evaluator New York, NY
- work from home social media evaluator New York, NY
- clinical evaluator New York, NY
- social media evaluator New York, NY
- quality evaluator New York, NY
- revenue manager remote New York, NY


