Refusal Preference Reward Model Evaluator [Remote]
AuraOne Human Data
- Remote job
Refusal Preference Reward Model Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Why this role matters
Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.
Responsibilities
- Design adversarial prompts that probe known weakness classes (jailbreak, policy bypass, prompt injection) for Refusal Preference Reward Model Evaluator assignments.
- Document every successful attack with reproduction steps and the policy clause it violated.
- Score model defenses across single-turn and multi-turn conversations.
- Triage emerging attack vectors and route them to the safety team with severity ratings.
- Maintain a personal library of attack patterns and propose new red-team rubrics.
- Calibrate against the broader red-team cohort to keep coverage and severity consistent.
Qualifications
- Demonstrated experience red-teaming AI systems, security research, or adversarial ML work for Refusal Preference Reward Model Evaluator work.
- Strong written communication — your reports become the patch ticket.
- Comfort working in policy-grey areas with clear documentation of what was attempted and why.
- Familiarity with prompt-injection, jailbreak, and policy-bypass taxonomies.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Construct a 5-turn adversarial conversation that bypasses a specific policy clause and write up the patch ticket.
- Score a model's defenses against a known jailbreak pattern across 20 variants.
- Propose a new red-team rubric category after spotting an emerging attack vector.
- Reproduce a failure another reviewer reported and confirm the severity tag.
Nice to have
- Background in offensive security, AppSec, or trust & safety operations.
- Experience publishing or reproducing public adversarial-ML research.
- Multilingual fluency for cross-language attack testing.
Skills
- Adversarial prompting
- Red-team analysis
- Policy taxonomy
- Failure documentation
- Refusal Preference Reward Model evaluation
- Preference ranking
- RLHF
- Rater calibration
- Refusal
- Preference
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- ...Preference Dataset QA Reward Model Evaluator is a remote evaluation track for reviewing preference dataset qa reward model evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Policy Preference Reward Model Evaluator is a remote review track for evaluating AI outputs in policy review workflows. Reviewers grade citation accuracy, statutory reasoning, and policy adherence; flag risk; and document the corrected analysis so the modeling team can...SuggestedRemote jobHourly payFor contractors10 hours per week
$15 - $20 per hour
...tools. ~Generate high-quality human evaluation data by identifying response strengths,... ...and completeness of responses. ~Ensure model responses align with expected conversational... ...structured analytical thinking. ~Preferred: ~Prior experience with RLHF, model evaluation...SuggestedPart timeSummer work$60 per hour
...professionals to help advance AI development. AI models are increasingly capable of performing... ...-of-the-art AI models on tasks like evaluating AI-generated quantitative analysis,... ...bachelor's degree in a quantitative field is preferred (Statistics, Computer Science,...SuggestedHourly payFull timeRemote workFlexible hours$40 per hour
...quantitative professionals to work remotely. In this role, you'll evaluate AI-generated quantitative work and solve technical problems... ...field, coding experience, and a bachelor's degree preferred. Enjoy a flexible schedule with competitive pay starting at $4...SuggestedHourly payRemote workFlexible hours$40 per hour
...their team. This remote role involves training AI models by posing complex mathematical problems, evaluating outputs, and assessing the model's performance. Candidates... ...of mathematics. A relevant Master's or PhD is preferred but not required. The position offers the...Hourly payRemote work$20 per hour
...external tools. Generate high-quality human evaluation data by identifying response strengths,... ..., and completeness of responses. Ensure model responses align with expected... ...requiring structured analytical thinking Preferred Experience with RLHF, model evaluation,...Remote jobContract workPart timeSummer work$40 per hour
...is seeking a Research Scientist (Chemistry) to enhance AI models by evaluating their performance with complex chemistry questions. The role... ...Ideal candidates will have expertise in chemistry, with a preference for those with advanced degrees. This independent contract...Hourly payContract workRemote workFlexible hours$85 per hour
...Responsibilities ~Use frontier AI coding agents to complete and evaluate complex engineering tasks. ~Review model-generated mobile application code for correctness,... ...implementations and architectural decisions. ~Preferred: Experience shipping production mobile applications...Contract workPart timeSummer workRemote work$60 per hour
...is seeking quantitative professionals to evaluate AI-generated analytical work and help... ...include statistical analysis, predictive modeling, and providing feedback to shape AI systems... ...and 2+ years of relevant experience are preferred. Join the team to make a significant...Hourly payRemote workFlexible hours$20 per hour
...technology company specializing in AI is seeking a Copywriter to evaluate and improve AI model outputs. This role is remote, offering both full-time and... ...writing and editing skills. A Bachelor’s degree is preferred but not required, making this a great entry-level...Remote jobHourly payFull timeContract workPart time$60 per hour
...data science company is seeking quantitative professionals to evaluate AI-generated analyses and design quantitative problems crucial... ...Fluency in English and a background in fields like data science, statistics, or economics is preferred. #J-18808-Ljbffr DataAnnotationRemote jobHourly pay$60 per hour
...development company is looking for quantitative professionals to evaluate AI-generated work and design quantitative problems to improve... ...experience in a quantitative field and a bachelor's degree is preferred. The job pays up to $60 per hour and is available to those...Remote jobHourly payFlexible hours$60 per hour
...development company is seeking quantitative professionals to evaluate AI-generated work and provide essential feedback. This role allows... ...coding skills, and a degree in a quantitative discipline is preferred. This position offers an opportunity to shape AI systems...Remote jobHourly payFlexible hours$40 per hour
A leading AI development firm in Michigan is seeking experienced quantitative professionals to evaluate AI-generated analyses and contribute to the evolution of AI models. Candidates should have a background in data science, statistics, or similar fields, with at least...Hourly payRemote work$40 per hour
...professionals in quantitative fields to enhance AI development. This fully remote role allows individuals to set flexible schedules while evaluating AI-generated analyses and solving complex quantitative problems. Candidates should have 2+ years' experience in relevant areas...Hourly payRemote workFlexible hours$40 per hour
...forward-thinking AI team is seeking quantitative professionals to evaluate and improve cutting-edge AI systems. The role involves working on AI-generated analyses and providing critical feedback for model enhancement. Candidates should have a strong quantitative...Hourly payRemote workFlexible hours$40 per hour
...development company is seeking experienced quantitative professionals to contribute to AI advancements. This fully remote role involves evaluating AI-generated analyses and ensuring they are technically accurate and valid in real-world scenarios. Candidates should have over 2...Hourly payRemote workFlexible hours$40 per hour
A data science team is seeking experienced quantitative professionals to evaluate AI-generated work and contribute to the development of cutting-edge AI systems. This fully remote position offers flexible scheduling and competitive hourly pay starting at $40+. Ideal candidates...Hourly payRemote workFlexible hours$40 per hour
...A leading AI development firm is seeking experienced quantitative professionals to evaluate AI-generated analysis and solve complex technical problems. Ideal candidates will have 2+ years in quantitative roles, knowledge of statistical methods, and experience with analytical...Hourly payRemote workFlexible hours$40 per hour
A leading AI development company is seeking experienced quantitative professionals for a remote role. Candidates will evaluate AI-generated quantitative work, solve complex problems, and provide valuable feedback. The ideal candidate has 2+ years of experience in a quantitative...Hourly payFull timeRemote workFlexible hours$40 per hour
...solutions company is seeking experienced quantitative professionals to evaluate AI-generated analyses and contribute to the development of... ...statistics and strong coding skills. Join us to directly impact the future of AI analytics and model reasoning. #J-18808-Ljbffr...Hourly payRemote workFlexible hours$40 per hour
A leading AI development firm is seeking experienced quantitative professionals to evaluate AI-generated quantitative work and provide critical feedback. This role offers the flexibility of remote work, allowing you to set your own schedule while focusing on impactful...Hourly payRemote work$40 per hour
...A leading AI development firm is looking for experienced quantitative professionals to evaluate AI-generated work and design problems for AI training. This fully remote position allows for a flexible schedule, offering competitive pay starting at $40+ per hour. Ideal...Hourly payRemote workFlexible hours$40 per hour
...A leading AI development firm is seeking experienced quantitative professionals to join their team remotely. The role involves evaluating AI-generated quantitative work, providing insights, and shaping the future of AI systems. Candidates should have over two years of...Hourly payRemote work$40 per hour
...A leading AI development firm is seeking experienced quantitative professionals to evaluate AI-generated quantitative analysis and provide impactful feedback. This fully remote role allows for flexible scheduling and competitive pay starting at $40 per hour. Candidates...Hourly payRemote workFlexible hours$40 per hour
...A leading AI company in the United States is seeking experienced quantitative professionals to evaluate and validate AI-generated analytical work. This fully remote position allows you to set your own schedule, with competitive hourly pay starting at $40 USD. Responsibilities...Hourly payRemote work$40 per hour
A leading AI development company is seeking experienced quantitative professionals to evaluate AI-generated analyses and design quantitative problems for AI training. This fully remote role offers flexibility in project selection and scheduling, with competitive pay starting...Hourly payRemote work$40 per hour
...A leading AI development firm is seeking experienced quantitative professionals to join their remote team. You will evaluate AI-generated quantitative analysis and solve complex problems to ensure technical accuracy. The ideal candidate should have at least 2 years of...Hourly payRemote workFlexible hours$40 per hour
...development company is seeking experienced quantitative professionals to evaluate and validate AI systems. The role is fully remote, offering... ...evaluating AI-generated work and designing problems for model training, contributing to shaping the future of AI systems. #J-...Hourly payRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Refusal Preference Reward Model Evaluator [Remote]. Be the first to apply!


