Combinatorics Model Evaluator [Remote]
AuraOne Human Data
- Remote job
Combinatorics Model Evaluator is a remote review track for evaluating AI outputs across combinatorics model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method so the modeling team can train on it.
Why this role matters
Combinatorics Model research review models live or die on whether their derivations actually hold up under scrutiny. AuraOne uses scientific specialists to grade outputs the way a peer reviewer would — checking assumptions, reproducing key steps, and capturing the right method alongside the wrong one.
Responsibilities
- Review AI outputs against current combinatorics model research review methods, conventions, and prior work for Combinatorics Model Evaluator assignments.
- Reproduce or sanity-check key derivations, calculations, or experimental claims.
- Flag dimensional, methodological, and citation errors with structured severity tags.
- Capture the corrected reasoning or worked example so the modeling team can train on it.
- Adjudicate disputed answers against textbooks, papers, or community standards.
- Maintain reviewer-quality scores in inter-rater calibration cycles.
Qualifications
- Graduate-level training or equivalent applied experience in combinatorics model research review or a closely related field for Combinatorics Model Evaluator work.
- Hands-on experience publishing, teaching, or advising on the topic at a professional level.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that cites methods, papers, or worked examples.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Reproduce a combinatorics model research review derivation from a model output and flag any algebraic or dimensional errors.
- Grade a model's literature summary against the cited papers and rate the citation quality.
- Adjudicate a disputed answer between two reviewers using textbook methods.
- Audit a 25-row batch for rubric consistency and report drift to the program lead.
Nice to have
- PhD, postdoc, or industry research experience in the topic area.
- Prior work reviewing AI-assisted research tooling and its failure modes.
- Multilingual fluency for non-English papers and corpora.
Skills
- Scientific reasoning
- Method validation
- Citation review
- Quantitative analysis
- Combinatorics Model research review
- Formal reasoning
- Proof review
- Combinatorics
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$20 per hour
...public sources and external tools. Generate high-quality human evaluation data by identifying response strengths, areas for improvement,... ...quality, clarity, tone, and completeness of responses. Ensure model responses align with expected conversational behavior and system...SuggestedRemote jobContract workPart timeSummer work$65 per hour
Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience...SuggestedHourly paySelf employmentWork from homeFlexible hours- ...Medical Document OCR Model Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Helpfulness Ranking Reward Model Evaluator is a remote evaluation track for reviewing helpfulness ranking reward model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Refusal Preference Reward Model Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Policy Preference Reward Model Evaluator is a remote review track for evaluating AI outputs in policy review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can...Remote jobHourly payFor contractors10 hours per week
$52k - $57k
...Job Description Job Description POSITION TITLE: Program Evaluator REPORTS TO: Director of Quality BROAD FUNCTION: Collects, analyzes, and reports on data to evaluate effectiveness of program services. I. CORE VALUES: # CULTURAL PROFICIENCY: Articulates...Full timeSummer workLocal area- Job Description Our client in MI is looking for Model-N admins . experience in integrating with Model N and configuring Model N. Additional InformationFull time
- ...Mental Health Evaluator The Mental Health Evaluator is a critical role dedicated to delivering comprehensive, clinically appropriate, culturally competent, and trauma-informed patient behavioral health interventions in the hospital Emergency Department. Responsibilities...
- ...Luxury Brand Evaluator Turn your passion for luxury into a career opportunity. Explore the world of premium brands and make a lasting impact in fashion, beauty, jewelry, or automobiles. Join CXG, the global leader in customer experience, and work alongside iconic names...Remote workWorldwideFlexible hours
$79.4k - $119.1k
...This position represents Auto Spec Control department in a development team environment as Project Lead on a mix of complex Full Model Change (FMC) and Minor Model Change (MMC) developments. Creates, promotes, and manages critical milestones for successful package delivery...Full timeTemporary workWork experience placementWork at officeRemote workRelocation package$14.5 per hour
A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical...Hourly payPart timeRemote workFlexible hours- A leading digital solutions provider is seeking a Personalized Ads Evaluator for a part-time position. This entry-level role involves reviewing online advertisements to assess their relevance to search terms, providing detailed feedback on content and cultural context....Part timeRemote work
$14.5 per hour
...AI Web Search Evaluator Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years of experience in...Hourly payPart timeCurrently hiringImmediate startRemote workWork from home10 hours per weekFlexible hours- A leading AI research accelerator is seeking a contractor to evaluate North American teen humor. The ideal candidate will review and rate humorous content popular among teens and explain its appeal. This role offers flexible, project-based work with competitive pay and...For contractorsFreelanceRemote workFlexible hours
$80 - $120 per hour
...This role is for one of our clients Compensation: $80 - $120 per hour We are hiring expert Evaluators in Special education / IEP to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality...Hourly payContract workFor contractorsWork at officeRemote work$80 - $120 per hour
...Special Education Iep Evaluator This role is for one of our clients. Compensation: $80 - $120 per hour. We are hiring expert Evaluators in special education / IEP to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy...Hourly payRemote work- ...wholesale partners, and sells online through its website, which serves customers across the United States and globally at WOMEN'S FIT MODEL: This is a part-time opportunity to be a part of the design process and to influence the comfort and fit of our clothing. You will...Part timeFlexible hours
- ...Family Model Provider Type: Family Model Provider (TN) - Independent Contractor If you are a positive and personable individual looking for a satisfying and fun opportunity to make a real difference in the lives of people with intellectual, developmental disabilities...Full timeContract workFor contractorsLive inWork from home
- ...a caring, compassionate, and reliable person? Here is an opportunity to make a lasting impact on someone’s life – become an Family Model Provider !! This is a chance to help make a difference by opening your home to an individual with developmental and/or physical disabilities...Daily paid
- ...talent in a variety of ways. We are pleased that you are interested in career opportunities offered at MICA. General purpose: Models will pose for art students for the purpose of studying the human figure. Summary of Essential Functions: Model poses may include...
- ...programs and deliver value to our business. At Applied Intuition, You Will: Conduct research on pretraining world-action foundation model with various world modalities including vision and physics associated with ego actions, serving the purpose for both robot action...Full timeFor contractorsFor subcontractorCasual workInternshipWork at officeImmediate startRemote workDay shift
$170k - $216k
...engineers like you to (1) develop methods for efficiently and continuously learning from large scale real-world data, to (2) develop models and model training at scale, to (3) analyze real-world behavior and develop systems for handling the complexities of interacting...Full timeRemote work$175k - $215k
...are currently focusing on include reinforcement learning, learning from demonstration, generative modeling, Bayesian inference, hierarchical learning, and robust evaluation. In this hybrid role, you will report to a Senior Research Scientist. You will:...Full timeRemote work$228.7k - $343.1k
...and screens for financial crime at enormous scale, and one bad model can mean millions in credit losses, suspicious activity that goes... ...that lets a lean team validate at scale, so you critically evaluate what it produces and own the evaluation that confirms its output...Remote jobFull timeLocal areaShift work- ...building next-generation AI infrastructure to power the world’s most demanding AI workloads. We provide high-performance GPU compute and Model API services, enabling AI companies, research labs, and enterprises to train and deploy cutting-edge models at scale. Our platform...Remote jobFull timeFlexible hours
$60 - $90 per hour
...Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation: $60–$90/hour Location: Remote Commitment:...Full timeContract workSummer workRemote work- ...communities thrive. Schedule: M-F, 8am- 5pm . This is a remote position, open to Missouri Residents only. The Project Evaluator-GLS plays a key role in evaluating the Garrett Lee Smith (GLS) Program by overseeing data collection, analysis, and reporting. This...Remote work
$22 - $25 per hour
...REVOLVEcareers or #lifeatrevolve. Are you ready to set the standard for Premium apparel? Main purpose of the position: The Fit Model will try on samples of clothes and express his or her expert knowledge on the fit of the garment, movement, fabric hand and feel,...Hourly payFlexible hours$298k - $368k
...engineers like you to (1) develop methods for efficiently and continuously learning from large scale real-world data, to (2) develop models and model training at scale, to (3) analyze real-world behavior and develop systems for handling the complexities of interacting...Full timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Combinatorics Model Evaluator [Remote]. Be the first to apply!




