Rust Programming AI Evaluator [Remote]
AuraOne Human Data
- Remote job
Rust Programming AI Evaluator is a remote evaluation track for reviewing rust programming ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn rust programming ai evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate rust programming ai evaluation model outputs against a versioned rubric and assign severity tags for Rust Programming AI Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on rust programming ai evaluation or adjacent content for Rust Programming AI Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two rust programming ai evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Rust Programming AI evaluation
- Software engineering and computer use
- AI evaluation
- Rubric writing
- Expert review
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
$70 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Working proficiency in Python , R , or another relevant programming language for scientific computing. Comfortable with Git/GitHub...SuggestedContract workSummer workImmediate startRemote work$100 per hour
...looking for a highly experienced software engineer (SR+) to help evaluate the quality of interactions with modern coding agents such as... ...whether the model thinks like a great engineer. You will assess how AI coding agents behave in real-world scenarios — focusing on:...SuggestedTemporary workImmediate start$85 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-...SuggestedContract workSummer workRemote work- ...leading tech firm is seeking a Software Engineering, Data Science, and Systems Design Expert with expertise in C#. The role involves evaluating AI-generated coding responses for accuracy and completeness. Candidates should possess strong experience in software engineering,...SuggestedPart timeRemote workFlexible hours
- Mercor is hiring experienced music producers and audio engineers to evaluate generative music AI models. You will assess AI-generated music across genres and rate it against detailed quality standards, working in Hindi and English. Responsibilities include head-to-head...SuggestedRemote work
$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...Remote workFlexible hours$11.5 per hour
...position as an Online Task Contributor. In this role, you will evaluate and provide feedback on content to enhance search engine results... ...at $11.50 hourly, based on task completion, with a supportive community of contributors involved in AI advancements. #J-18808-LjbffrHourly payPart timeRemote work$14.5 per hour
...Join to apply for the AI Web Search Evaluator role at Welo Data Welo Data works with technology companies to provide datasets that are high-quality, ethically sourced, relevant, diverse, and scalable to supercharge their AI models. As a Welocalize brand, Welo Data...Hourly payPart timeImmediate startRemote workWork from home10 hours per weekFlexible hours- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Work at officeRemote work
- ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and...Full timeContract workRemote workFlexible hours
$80 - $120 per hour
...Cybersecurity / IT GRC Evaluator $80-$120 per hour Hourly contract Remote About the Role We are hiring expert Evaluators in Cybersecurity / IT GRC to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor...Hourly payContract workFor contractorsWork at officeRemote work- ...Senior Software Engineer — AI Coding Evaluator is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit test the model should...Remote jobHourly payFor contractors10 hours per week
- ...experienced professionals from software engineering, DevOps, SRE, cloud, platform, or infrastructure engineering backgrounds to test and evaluate AI-assisted workflows across modern developer and workplace tools. The role focuses on assessing technical accuracy, workflow...Contract work
- ...annotations from vendors Build necessary tooling for managing evaluation datasets Collaborate with model development teams to... ...Ensure consistency and accuracy in measurements, involving programming and auditing tasks Collaborate with model development teams...Remote job
- A virtual AI evaluation firm is seeking individuals to review and evaluate AI-generated responses in therapeutic conversations. The ideal candidate will possess strong written communication and analytical skills, as well as a keen attention to detail for assessing tone...Remote jobImmediate start
- About the role We are hiring expert Evaluators in Compliance / regulatory response with financial-services AI to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject‑...Hourly payWork at officeRemote work
- A leading AI research firm is looking for a contractor in internal or emergency medicine to help enhance language models by providing expert feedback based on real-world medical scenarios. Candidates should possess an MD and have relevant experience, including board certification...Remote jobFor contractors
$100 - $125 per hour
...hour fast-track onboarding required Key Responsibilities Translate real-world insurance workflows into structured tasks for AI systems Evaluate AI-generated outputs for accuracy, logical reasoning, and business relevance Work on use cases including underwriting,...Remote jobHourly payFor contractorsFreelanceWork at officeImmediate startFlexible hours- Mercor is seeking expert Evaluators in Humanities / arts / culture to review AI-generated work products for accuracy, rigor, and domain quality. This remote hourly engagement requires deep subject-matter expertise to grade outputs and provide actionable feedback. Candidates...Remote jobHourly payWork at office
- Dorado is seeking Speech AI Evaluation Specialists based in Thailand to help improve Vietnamese-language AI content. This is a remote, freelance role with Bangkok as the working location and a flexible, part-time schedule. You will participate in short voice conversations...Remote jobPart timeFreelanceImmediate start10 hours per weekFlexible hours
- YO AI Labs is seeking Dutch bilingual experts to remotely evaluate AI-generated Dutch speech for nativeness and quality, providing detailed written feedback and documented findings. The role requires native Dutch, strong English, and high attention to detail to ensure linguistic...Remote job
- Turing is seeking an experienced professional to review AI training issues and research environments for frontier AI labs. The ideal candidate will have a Master's/PhD or 4+ years experience in relevant engineering fields. This role emphasizes feedback skills, structured...Remote jobContract workFor contractors
$8 per hour
Dorado is seeking a Speech AI Evaluation Specialist to help improve AI-generated content in Thai or Chinese Simplified. This freelance, part-time role is remote from Thailand, offering 10+ hours weekly with an immediate start and a rate of 8 USD/hour. You will engage in...Remote jobPart timeFreelanceImmediate start- Dorado is seeking an AI Language Quality Evaluator fluent in Greek and English for an ongoing, task-based project. This remote freelance role involves reviewing translated and AI-flagged content to judge accuracy, classify issues, and suggest corrected translations. You...Remote jobFor contractorsFreelanceFlexible hours
- Obsidian is hiring expert Evaluators in Investment analysis / valuation / credit to review AI-generated work products for accuracy and quality. This remote, hourly position requires deep subject-matter expertise and professional fluency in English to provide structured...Hourly payWork at officeRemote work
- Turing is seeking detail-oriented AI Analysts based in the United States for a Google Wallet evaluation project. This role allows you to engage with advanced AI tools while contributing to the future of AI. You will evaluate model responses, review output quality, and provide...Remote jobFull timeContract work
- Rex.zone is seeking a Senior AI data annotator to perform data labeling and evaluation for NLP tasks, RLHF assessments, and prompt QA to improve training data quality and model performance. This is a US-based remote, full-time role aligned with Miami talent demand. You...Remote jobFull time
- Dorado is seeking a Speech AI Evaluation Specialist to support the improvement of AI-generated content in Vietnamese (USA). This is a freelance, remote, part-time role starting immediately. You will participate in short voice conversations with AI models, follow prompts...Remote jobPart timeFreelanceImmediate startFlexible hours
- Obsidian is seeking expert Evaluators in Legal/compliance to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote, hourly engagement. You will...Remote jobHourly pay
$56k - $84k
...mission is to serve the people of our communities. The Quality Evaluator is responsible for designing, implementing, and overseeing evaluation... ...vision, life, and short-term disability insurance Retirement program with generous company default contribution and match Generous...Temporary workWork at officeLocal areaRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Rust Programming AI Evaluator [Remote]. Be the first to apply!



