Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Operations Research Model Prompt Evaluator [Remote]

$60 - $80 per hour

AuraOne Human Data

Remote
  • Remote job

Operations Research Model Prompt Evaluator is a remote review track for evaluating AI outputs across operations research model prompt research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method so the modeling team can train on it.

Why this role matters

Operations Research Model Prompt research review models live or die on whether their derivations actually hold up under scrutiny. AuraOne uses scientific specialists to grade outputs the way a peer reviewer would — checking assumptions, reproducing key steps, and capturing the right method alongside the wrong one.

Responsibilities

  • Review AI outputs against current operations research model prompt research review methods, conventions, and prior work for Operations Research Model Prompt Evaluator assignments.
  • Reproduce or sanity-check key derivations, calculations, or experimental claims.
  • Flag dimensional, methodological, and citation errors with structured severity tags.
  • Capture the corrected reasoning or worked example so the modeling team can train on it.
  • Adjudicate disputed answers against textbooks, papers, or community standards.
  • Maintain reviewer-quality scores in inter-rater calibration cycles.

Qualifications

  • Graduate-level training or equivalent applied experience in operations research model prompt research review or a closely related field for Operations Research Model Prompt Evaluator work.
  • Hands-on experience publishing, teaching, or advising on the topic at a professional level.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that cites methods, papers, or worked examples.
  • Reliable async availability for at least 10 hours per week.

Example tasks

  • Reproduce a operations research model prompt research review derivation from a model output and flag any algebraic or dimensional errors.
  • Grade a model's literature summary against the cited papers and rate the citation quality.
  • Adjudicate a disputed answer between two reviewers using textbook methods.
  • Audit a 25-row batch for rubric consistency and report drift to the program lead.

Nice to have

  • PhD, postdoc, or industry research experience in the topic area.
  • Prior work reviewing AI-assisted research tooling and its failure modes.
  • Multilingual fluency for non-English papers and corpora.

Skills

  • Scientific reasoning
  • Method validation
  • Citation review
  • Quantitative analysis
  • Operations Research Model Prompt research review

Work model

Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.

Compensation

$60–$80 / hr

Application process

Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.

Vacancy posted 20 hours ago
Similar jobs that could be interesting for youBased on the Operations Research Model Prompt Evaluator [Remote] in Remote vacancy
  •  ...topline and streamline their operations. We specialize in providing tailored...  ...: Linkedin Job Title: Evaluator - Political Science Location:...  ...question posed in the user prompt. The goal is to assess quality...  ...Stay updated with the latest research, guidelines, and advancements... 
    Operations
    Contract work
    Remote work
    Worldwide

    Straive

    New York, NY
    1 day ago
  •  ...Role Overview Help evaluate frontier AI systems’ ability...  ...scenarios, assess model responses, and define...  ...challenging single-turn prompts in radiological safety...  ...is preferred. Operational health physics experience...  ...Technical writing, published research, or expert witness... 
    Operations
    Remote work

    SaidGig

    United States
    24 days ago
  •  ...Refusal Preference Reward Model Evaluator is a remote red-team track for...  ...systems against adversarial prompts. Reviewers craft attack scenarios...  ...AI systems, security research, or adversarial ML work for...  ..., AppSec, or trust & safety operations. Experience publishing or... 
    Operations
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    20 hours ago
  •  ...network and facility operations. Bitdeer also offers advanced...  ...product built on a model is bounded by what it...  ...This role builds the evaluation and decision systems...  ...consumption. You will research and prototype adaptive...  ...calling, context windows, prompt caching, reasoning... 
    Operations
    Full time

    Bitdeer

    Singapore
    11 days ago
  •  ...Annotator—Product Management & Marketing to remotely review evaluation prompts and responses against the company's quality rubric. You will...  ...outputs, label edge cases, and provide structured feedback the modeling team can use to retrain. As a contractor, you will assess... 
    Suggested
    Remote job
    For contractors

    AI Trainer Jobs

    New York, NY
    5 days ago
  • AI Trainer Jobs is seeking an Illustration Quality Evaluator for a remote contractor role. Review illustration quality evaluation prompts and responses against a published rubric, compare paired outputs, and provide structured feedback to support retraining efforts. You... 
    Remote job
    Part time
    For contractors

    AI Trainer Jobs

    New York, NY
    5 days ago
  • AI Trainer Jobs is seeking a remote contractor to evaluate people ops / recruiting prompts and responses against a evolving quality rubric. You will compare model outputs, label edge cases, and provide structured feedback for retraining. Responsibilities include evaluating... 
    Remote job
    For contractors

    AI Trainer Jobs

    New York, NY
    5 days ago
  • Surgical Planning Safety Evaluator is a remote evaluation track for reviewing surgical planning safety evaluation prompts and responses against AuraOne's quality rubric. Reviewers...  ...write the kind of structured feedback the modeling team can use to retrain. AI data... 
    Remote job

    AI Trainer Jobs

    New York, NY
    5 days ago
  • Receipt and Invoice Understanding Model Evaluator is a remote evaluation track for reviewing receipt and invoice understanding model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind... 
    Hourly pay
    For contractors
    Remote work
    10 hours per week

    AI Trainer Jobs

    New York, NY
    5 days ago
  •  ...Certified Evaluator The Certified Evaluator delivers...  ...mobile appointments, researches and evaluates merchandise...  ...details, and promptly communicate changes or...  ...Accurately record brand, model, serial number, condition...  ...needs or damage. Daily Operations and Team... 
    Operations
    Minimum wage
    Temporary work
    Work at office
    Local area
    Relocation

    Pawn America

    Burnsville, MN
    3 days ago
  • $65 - $70 per hour

     ...and Safety Policy AI Evaluator is a remote red-team track...  ...against adversarial prompts. Reviewers craft...  ...how AuraOne hardens AI models before they ship to customers...  ...AI systems, security research, or adversarial ML...  ...AppSec, or trust & safety operations. Experience publishing... 
    Operations
    Hourly pay
    For contractors
    Remote work
    10 hours per week

    AI Trainer Jobs

    New York, NY
    5 days ago
  • $173k

     ...ManagementCompany: CitiCitibank, N.A. seeks a Model Validation 2nd LOD Lead Analyst for its...  ...processes, aiming to increase operational efficiency for the organization. Developed...  ...Transformer Architectures, RAG Pipelines, Prompt Engineering, Fine-Tuning (LoRA, RLHF, PEFT... 
    Operations
    Full time
    Remote work

    Citigroup

    Wilmington, DE
    2 days ago
  •  ...profitable growth of the company. The Fit Model has an integral role in the product...  ...patterns, and retrieve or ship packages. Operate within Agenda. Meet all assigned dates and...  ...Fitting Schedules and Business Meetings promptly. Collaborate Effectively with Colleagues.... 
    Operations
    Contract work
    Part time
    Casual work
    Shift work

    Lilly Pulitzer

    King of Prussia, PA
    3 days ago
  • $40 - $100 per hour

     ...About AI Training and Scientific Evaluation AI training is the human side...  .... Specialists review model responses, test reasoning, identify...  ...-by-step solutions Evaluate prompts, model answers, and technical...  ...scenarios, calculations, and research-style reasoning tasks. Focus... 
    Hourly pay
    Contract work
    Part time
    For contractors
    Remote work

    OpenTrain AI

    Brooklyn, NY
    4 days ago
  • $90 - $120 per hour

     ...Create benchmark-quality responses for future model outputs and document the context and...  ...educational statistics materials, such as research articles, curricula, industry reports, or...  ...each week. Availability to begin promptly is expected. Selected candidates should be... 
    Hourly pay
    For contractors
    Remote work

    SaidGig

    United States
    more than 2 months ago
  • $125k - $150k

     ...ECS is seeking an AI Model Engineer to work in a...  ...with a track record of evaluating, experimenting with, and...  ...technologies into operational use. This role requires...  ...execution of cutting-edge research programs designed to...  ...agentic workflow, and prompt engineeringExperience... 
    Operations
    Contract work
    Work at office
    Remote work

    ECS Federal

    Fairfax, VA
    3 days ago
  •  ...AI Labs seeks PhD and academic experts to support AI research projects remotely. You will evaluate model outputs across technical and humanities topics,...  ...and depth. As a contractor, you will create reference prompts, critiques, and quality benchmarks, and contribute to... 
    Remote job
    For contractors

    YO AI Labs

    Austin, TX
    5 days ago
  • $50 per hour

    Combinatorics Model Evaluator is a remote review track for evaluating AI outputs across combinatorics model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method... 
    Hourly pay
    For contractors
    Remote work

    AI Trainer Jobs

    New York, NY
    5 days ago
  • $15 - $20 per hour

     ...creative and technical talent with leading AI research labs. Headquartered in San Francisco,...  ...tools. Generate high-quality human evaluation data by identifying response strengths,...  ..., and completeness of responses. Ensure model responses align with expected conversational... 
    Contract work
    Summer work
    Remote work

    Remote Jobs

    New York, NY
    5 days ago
  •  ...Multimodal Hallucination Model Evaluator is a remote evaluation track for reviewing multimodal hallucination model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback... 
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    15 days ago
  • YO AI Labs is seeking Humanities Evaluation Specialists for a remote contract to support an AI training project. You will research, analyze, and craft challenging humanities questions...  ...experience is required. You will design prompts that require deep interpretation and... 
    Remote job
    Contract work

    YO AI Labs

    New York, NY
    3 days ago
  •  ...a PhD & Academic Expert to support AI research projects remotely. You will apply your subject-matter expertise to evaluate and improve model responses across technical and humanities...  ...disciplines. You will develop expert prompts, reference answers, and critiques, while... 
    Remote job

    YO AI Labs

    New York, NY
    4 days ago
  • $50 - $100 per hour

     ...engineering and problem-solving expertise to code generation and model evaluation work that helps improve how next-generation AI systems learn,...  ..., with preference for experts who are available to begin promptly. Participation is subject to selection for the project.... 
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    United States
    24 days ago
  • $37.5 per hour

     ...originals. In this role you will be responsible for evaluating advertising copy generated by a large language model (LLM) to ensure that it meets their benchmark in...  ...expertise in strategy, design, execution and operations unlocks business value through a range of... 
    Operations
    Contract work
    Temporary work
    Remote work

    TEKsystems

    New York, NY
    17 hours ago
  •  ...computational problem solving to improve and evaluate large language models. You will design rigorous math...  ...customers Accelerate frontier AI research by contributing high quality data and...  ...topics. Design exact, closed ended prompts and produce reliable Python solutions... 
    Contract work
    For contractors
    Freelance
    Remote work

    SaidGig

    United States
    more than 2 months ago
  • AuraOne is seeking a Latin Bilingual Expert for a remote evaluation track. Reviewers compare paired outputs to a quality rubric, label edge cases, and generate structured feedback to retrain models. Responsibilities include evaluating outputs against rubrics, tagging issues... 
    For contractors
    Remote work
    10 hours per week

    AI Trainer Jobs

    New York, NY
    5 days ago
  •  ...Research Interns at Applied Intuition Applied Intuition, Inc. is powering the future...  ...three core areas: tools and infrastructure, operating systems, and autonomy. Eighteen of the...  ...on pretraining world-action foundation model with various world modalities including... 
    Full time
    For contractors
    For subcontractor
    Casual work
    Internship
    Work at office
    Immediate start
    Remote work
    Day shift

    Applied Intuition

    Sunnyvale, CA
    2 days ago
  •  ...Role Overview Help evaluate how advanced AI systems handle sensitive...  ...misuse potential, ensuring models remain useful for routine professional...  ...challenging single-turn prompts from your domain and classify...  ...writing ability. Published research, prior technical writing, or... 
    For contractors
    Remote work

    SaidGig

    United States
    24 days ago
  • $65 - $75 per unit

     ...analytical chemistry expertise to evaluate how AI systems handle...  ...could enable harm, ensuring models provide useful answers when appropriate...  ...single-turn chemistry prompts classified as benign, dual-use...  ...technical writing ability. Published research, prior technical writing, or... 
    Remote work

    SaidGig

    United States
    24 days ago
  • $16 - $20 per hour

     ...Research Evaluator Position The research team at Penn State Ross and Carol Nese College of Nursing is hiring part-time research evaluators for projects focused on dementia care in assisted living settings. The research evaluator will assist with in-person recruitment... 
    Hourly pay
    Part time
    Summer work

    Penn State University

    State College, PA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Operations Research Model Prompt Evaluator [Remote]. Be the first to apply!