Rust Programming AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Rust Programming AI Evaluator is a remote evaluation track for reviewing rust programming ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn rust programming ai evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate rust programming ai evaluation model outputs against a versioned rubric and assign severity tags for Rust Programming AI Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on rust programming ai evaluation or adjacent content for Rust Programming AI Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two rust programming ai evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Rust Programming AI evaluation
- Software engineering and computer use
- AI evaluation
- Rubric writing
- Expert review
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$100 - $150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...00–$150/hour Location: Remote Role Responsibilities Evaluate AI-generated artifacts for usability in a software engineering...SuggestedContract workSummer workWork at officeRemote work$14.5 per hour
A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical...SuggestedHourly payPart timeRemote workFlexible hours$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...SuggestedRemote workFlexible hours- ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and...SuggestedFull timeContract workRemote workFlexible hours
- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...SuggestedWork at officeRemote work
- ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers...Contract workTemporary workImmediate startRemote work
$14.5 per hour
...diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data... ...! Job Description As a Web Search Evaluator, you will play a key role in improving the... ...paid sick time and employee assistance programs. Benefits Following eligibility requirements...Part timeCurrently hiringImmediate startRemote workWork from home10 hours per weekFlexible hours$14.5 per hour
...AI Web Search Evaluator Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years of experience in...Bi-weekly payHourly payPart timeImmediate startRemote workWork from homeFlexible hours- ...Senior Software Engineer — AI Coding Evaluator is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit test the model should...Remote jobHourly payFor contractors10 hours per week
- ...annotations from vendors Build necessary tooling for managing evaluation datasets Collaborate with model development teams to... ...Ensure consistency and accuracy in measurements, involving programming and auditing tasks Collaborate with model development teams...Remote job
- ...About the role We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise...Hourly payWork at officeRemote work
- Obsidian is hiring expert Evaluators in Healthcare operations for a remote, hourly engagement. You will review AI-generated work products for accuracy, rigor, and domain quality, leveraging your extensive subject-matter expertise in the field. The ideal candidate has over...Remote jobHourly payWork at office
- Mercor is hiring expert Evaluators in General Sales / GTM to review AI-generated outputs for accuracy and quality. This remote hourly role requires applying deep domain knowledge to assess documents, spreadsheets, and slide decks. Ideal candidates have 5+ years in General...Remote jobHourly payWork at office
$14.5 per hour
...ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data leverages over 25 years... ...performance and provide insights on relevance and quality. Evaluate and rate the effectiveness of search engine results to ensure...Hourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours$85 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-...Contract workSummer workRemote work$20 per hour
A leading AI development company in the United States is seeking detail-oriented individuals for remote opportunities in training AI... ...include developing prompts, writing high-quality responses, and evaluating AI outputs. The ideal candidates are fluent in English with...Hourly payRemote workFlexible hours- ...Prolific is seeking talented Product Designers and UX Specialists to join our Expert Network for training and evaluation of cutting-edge AI models. Candidates should have a relevant educational background and at least one year of experience in design-related fields. Responsibilities...Remote workWork from homeFlexible hours
$20 - $80 per hour
...Role Overview Help improve next-generation AI systems by supplying precise, real-world evaluation, annotation, and feedback. This remote contractor role focuses on how AI models learn, reason, and perform across diverse subject areas. Key Responsibilities Evaluate...Hourly payContract workFor contractorsRemote work$20 - $160 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Position: Generalist Annotator — Health AI Conversation Quality Evaluation Type: Contract Compensation: $20–$160/hour Location...Contract workSummer workRemote work- ...Summary This is a fully remote, hourly contractor role supporting AI data and language projects on a project-based, flexible hour... ..., and other content to support AI training datasets. LLM evaluation: reviewing AI-generated responses for accuracy, reasoning quality...Hourly payFor contractorsRemote workFlexible hours
- ...Prolific is seeking Product Designers and UX Specialists to join our Expert Network, contributing to the training and evaluation of cutting-edge AI models. This role requires expertise in usability, design systems, and user research, providing essential feedback on design...Remote workWork from homeFlexible hours
- YO AI Labs is seeking a PhD and academic expert to support AI research projects remotely. You will apply subject-matter expertise to evaluate and improve AI model responses across technical and humanities disciplines. You will design expert prompts, reference answers,...Remote job
$70 - $110 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...weeks Commitment: 20+ hours/week Role Responsibilities Evaluate AI-generated financial plans , budgets, and forecasts for quality...Hourly payContract workSummer workWork at officeImmediate startRemote work$20 per hour
A leading AI technology firm is seeking a Graphic Design Expert to evaluate graphic design elements for AI model training. This part-time contract role offers remote work flexibility and requires expertise in design principles. Ideal candidates should have formal training...Contract workPart timeRemote work$50 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Remote Commitment: 20–40 hours/week Role Responsibilities Evaluate the effectiveness of AI systems in handling personalized, real...Contract workSummer workRemote work- Mercor is hiring expert Evaluators in Media, journalism, and communications to review AI-generated work products for accuracy, rigor, and domain quality. This is a remote, hourly engagement. Role requires 5+ years in media/journalism/communications, native/professional...Remote jobHourly payWork at office
$30 - $50 per hour
...A leading AI training company is seeking a Remote Annotator to support human-in-the-loop AI training workflows for large language... ...models. This role involves reviewing labeled datasets, performing evaluations for helpfulness and safety, and ensuring quality in training...Hourly payRemote work$80 - $150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...quickly on work. Work independently and asynchronously to evaluate and improve AI model performance . Qualifications Must-Have...Contract workSummer workRemote work$90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...genuinely correct replications from those that merely look correct. Evaluate responsive behavior and semantic quality, ensuring proper use of...Contract workSummer workLocal areaRemote work$50 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...: $50/hour Location: Remote Role Responsibilities Evaluate the quality and accuracy of LLM-generated English text across...Contract workSummer workRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Rust Programming AI Evaluator [Remote]. Be the first to apply!


