Quant Finance Reasoning Model Evaluator [Remote]
AuraOne Human Data
- Remote job
Quant Finance Reasoning Model Evaluator is a remote review track for evaluating AI outputs across finance and risk workflows. Reviewers grade calculations, narrative reasoning, and policy adherence; flag compliance and reconciliation issues; and document the correct treatment so the modeling team can train on it.
Why this role matters
Finance and risk workflows leave no room for fuzzy reasoning. AuraOne uses experienced finance-and-risk professionals to grade AI outputs the way an audit reviewer would — with citations, severity tags, and the corrected calculation alongside the original.
Responsibilities
- Review AI outputs against current finance and risk standards, regulations, and firm policy for Quant Finance Reasoning Model Evaluator assignments.
- Re-perform calculations and flag rounding, classification, or treatment errors.
- Tag compliance, reconciliation, and disclosure issues with structured severity scores.
- Capture the corrected workpaper or narrative so the modeling team can train on it.
- Adjudicate disputed treatments against published standards or firm guidance.
- Maintain reviewer-quality scores in inter-rater calibration cycles.
Qualifications
- Direct working experience with finance and risk on real engagements, transactions, or audits for Quant Finance Reasoning Model Evaluator work.
- A CPA, CFA, or ACA. A comparable credential or equivalent applied experience also works.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that cites standards, regulations, or firm policy.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Re-perform a finance and risk calculation produced by a model and flag any rounding or classification errors.
- Grade a model's narrative reasoning on a complex transaction and rate the citation quality.
- Adjudicate a disputed treatment between two reviewers using current standards.
- Audit a 25-row batch for rubric consistency and report drift to the program lead.
Nice to have
- Big Four, bulge-bracket, or in-house controller / risk experience.
- Familiarity with AI-assisted close, audit, or risk tooling and its failure modes.
- Bilingual experience for cross-jurisdiction reviews.
Skills
- Financial analysis
- Audit-grade review
- Regulatory standards
- Risk assessment
- Finance and risk
- Formal reasoning
- Proof review
- Quantitative analysis
- Quant
- Finance
- Reasoning
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$50 - $60 per hour
...remote team in the United States. This role involves training AI models, measuring their progress, and evaluating outputs to enhance their quality. Ideal candidates should possess expert-level financial reasoning and proficiency in financial analysis. This is a flexible...SuggestedRemote jobHourly payFlexible hours- ...Geometry Reasoning Model Evaluator is a remote evaluation track for reviewing geometry reasoning model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling...SuggestedRemote jobHourly payFor contractors10 hours per week
$50 - $60 per hour
A technology company specializing in AI and finance is seeking a Wealth Advisor to help train AI models. In this independent contract role, you will measure the effectiveness of AI chatbots by solving complex financial problems. The ideal candidate should be fluent in English...SuggestedRemote jobHourly payContract work$20 per hour
...external tools. Generate high-quality human evaluation data by identifying response strengths, areas... ...improvement, and factual inaccuracies. Assess reasoning quality, clarity, tone, and completeness of responses. Ensure model responses align with expected conversational...SuggestedRemote jobContract workPart timeSummer work- ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and...SuggestedFull timeContract workRemote workFlexible hours
- Affirm is looking for an intelligent, driven professional to join our Bank Model Risk Management (MRM) team in a remote-friendly capacity. You will validate credit/fraud models, develop automated monitoring in Python, and work with developers to ensure robust, compliant...Remote job
- Affirm is seeking a seasoned professional for the Bank Model Risk Management (MRM) team. You will validate credit/fraud models, build automated Python monitoring, and drive remediation with developers to ensure robust, compliant models. You will partner with Audit, Controls...Remote job
- Affirm is seeking an experienced professional to join the Bank Model Risk Management team. You will validate credit/fraud models, develop automated monitoring in Python, and work with cross-functional teams to ensure robust, compliant risk management. You will collaborate...Remote job
- Affirm is seeking an experienced Bank Model Risk Manager to lead independent validations of credit and fraud models. You will build automated Python monitoring, work with model developers to remediate issues, and partner with Audit and Compliance to satisfy regulatory requests...Work at officeRemote work
- ...remote-first company reinventing credit to make it more honest and friendly. Affirm seeks an experienced professional to join the Bank Model Risk Management team to identify, quantify, monitor and report on model risk across credit and fraud domains. You will validate...Remote job
- ...Labs is seeking an Investment & Finance Expert to work remotely as a contractor. The role focuses on evaluating AI outputs related to valuation, financial modeling, markets, and investments, and crafting... ...answers to improve model reasoning. The ideal candidate brings experience...Remote jobFor contractors
- Affirm is seeking an experienced professional for Bank Model Risk Management (MRM) to build and oversee risk frameworks for credit and fraud models. You will validate models, monitor performance, and collaborate cross-functionally to ensure mathematical robustness and regulatory...Remote work
$100 per hour
...Role Overview Apply your finance expertise to improve AI-driven... ...rigorous, real-world analysis, evaluation, and feedback. This remote, part... ...financial data, reports, and model outputs using detailed... ...critical thinking, analytical reasoning, attention to detail, and quality...Hourly payContract workPart timeFor contractorsRemote work- ...Role Overview Use your investment and finance expertise to evaluate and improve AI model performance on financial reasoning, valuation, markets, and real-world investment scenarios. Key Responsibilities Assess AI model outputs on valuation, financial modeling, markets...For contractorsRemote work
$100 per hour
...improve the performance of large language models on finance tasks. You will work with AI... ...AI systems. Key Responsibilities Evaluate LLM performance in finance areas where... ...Portfolio Management, Research, Trading, Quant, Investment Banking, Private Equity, Venture...Hourly payContract workFor contractorsFreelanceRemote work10 hours per weekFlexible hours$96k - $181k
...Analytics Associate, you will be at the forefront of validating models for Market Risk, IRRBB (including NII, EVE, Deposit modeling),... ...or limited in their ability to apply on this site may request reasonable accommodations by emailing ****@*****.***. #LI-Remote...Work at officeRemote workFlexible hours$85k - $110k
Officer, Model Risk Management page is loaded## Officer, Model Risk Managementlocations:... ...community and commercial banking, specialty finance and wealth management services through... ...validation process to discuss justification and reasoning behind validation and review findings.*...Temporary workRemote workFlexible hours- ...Proof Verification Model Evaluator is a remote review track for evaluating AI outputs across proof verification model research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the...Remote jobHourly payFor contractors10 hours per week
- ...Medical Document OCR Model Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers grade differential reasoning, dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical...Remote jobHourly payFor contractors10 hours per week
- ...Policy Preference Reward Model Evaluator is a remote review track for evaluating AI outputs in policy review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can...Remote jobHourly payFor contractors10 hours per week
- ...training data that teaches AI systems to reason through complex ASC judgments. This role... ...deliverables into clear scenarios and evaluations so AI can reach standards accurate, defensible... .... Bachelor''s degree in Accounting, Finance, or a related field. Strong written...Hourly payRemote work
- Mercor is hiring Legal Experts to evaluate AI-generated responses for employment and labor law scenarios. This fully remote, hourly contract... ...week. You will assess accuracy, provide feedback to improve model behavior and participate in calibration sessions. Requirements...Remote jobHourly payContract workFlexible hours
- ...next generation of AI systems reason through accounting concepts,... ...challenge advanced language models on accountant-specific topics... ...Bachelor’s degree in Accounting, Finance, or a closely related field... ...training, annotation, or evaluating AI-generated technical content...Remote jobFor contractors
- ...Job Description: The Model Risk Manager will be responsible for performing and managing... ...services, primarily for credit risk, finance, and treasury management models (asset liability... ...business and organizational needs. A reasonable estimate of the current range is 106,600...Local areaRemote work
$40 - $95 per hour
...knowledge that improves how AI understands, reasons about, and processes financial... ...compliance, and accounting best practices. Evaluate financial statements prepared by others... ...Bachelor''s degree or higher in Accounting, Finance, Economics, or a related field. In-...Hourly payContract workRemote work- CapitexAI is seeking a Finance Domain Expert to join our GenAI team in the Bay Area. You will help shape benchmarks and evaluation criteria to improve model reasoning in finance, translating professional financial judgment into concrete criteria and golden solutions. Work...Remote jobFlexible hours
- ...Tax Law experts to join projects that fine-tune AI models. You will apply your U.S. tax expertise to evaluate AI responses on income tax, corporate tax, and... ...provisions. This freelance role requires strong legal reasoning and precise writing, with a flexible, fully remote...Remote jobFreelanceFlexible hours
$91k - $169k
...& Strategy Specialist (Fraud Model Analyst) within PNC's Technology... ...model controls.• Leads in evaluating identified model risks, defects... ...with the line of business, Finance, and Risk partners to assess... ...extent required to provide needed reasonable accommodations. At PNC we...Full timeTemporary workPart timeWork experience placementWork at officeRemote work$135k - $165k
...community and commercial banking, specialty finance and wealth management services through... ...is seeking a highly motivated Model Risk Vice President to join our Model Risk... ...clustering, etc.). Familiarity with performance evaluation metrics and techniques for LLMs.PhD or...Full timeTemporary workRemote workFlexible hours$80 - $120 per hour
...slide decks for accuracy, rigor, and professional quality. Role Overview You will evaluate AI-generated work products using domain-specific quality rubrics and deliver well-reasoned assessments grounded in deep subject-matter expertise. Key Responsibilities Review...Hourly payWork at officeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Quant Finance Reasoning Model Evaluator [Remote]. Be the first to apply!


