Embodied Navigation Task Evaluator [Remote]
AuraOne Human Data
- Remote job
Embodied Navigation Task Evaluator is a remote evaluation track for reviewing embodied navigation task evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn embodied navigation task evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate embodied navigation task evaluation model outputs against a versioned rubric and assign severity tags for Embodied Navigation Task Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on embodied navigation task evaluation or adjacent content for Embodied Navigation Task Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two embodied navigation task evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Embodied Navigation Task evaluation
- Robotics evaluation
- Embodied AI
- Safety review
- Embodied
- Navigation
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- ...Robot Policy Compliance Task Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft... ...Policy Compliance Task evaluation Robotics evaluation Embodied AI Safety review Robot Policy Compliance Work...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Surgical Robotics Task Evaluator is a remote review track for evaluating robotics review outputs against real-world physical-AI constraints... ...Motion safety Incident taxonomy Robotics review Embodied AI Safety review Surgical Robotics Work model...SuggestedRemote jobHourly payFor contractors10 hours per week
- ...Physical Affordance Task Evaluator is a remote review track for evaluating AI outputs across physics reasoning, calculations, and research... ...Quantitative analysis Physics Robotics evaluation Embodied AI Safety review Physical Affordance Work model...SuggestedRemote jobHourly payFor contractors10 hours per week
$60 per hour
...team, you'll work closely with state-of-the-art AI models on tasks like evaluating AI-generated security content, solving technical security... ...cyber operations. ~ Some coding experience required; comfort navigating and patching a codebase is key. ~ Fluency in English (...SuggestedHourly payFull timeRemote workFlexible hours- Search Quality Evaluator (AI Training) What if your everyday curiosity and sharp judgment... ...explanations to support your ratings Work through task-based assignments independently at your... ...Comfortable using search engines and navigating the web with confidence Able to assess...SuggestedHourly payOngoing contractContract workFreelanceRemote workFlexible hours
$20 per hour
...is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’... .... The position offers flexible hours and pays $20+ USD/hr for general projects and $40+ for design-focused tasks. #J-18808-LjbffrRemote workFlexible hours$20 per hour
...company specializing in AI is looking for a Digital Web Designer to evaluate AI-generated designs and help train models for better aesthetic... ...pay rates start at $20/hr for general projects and $40/hr for design-focused tasks, with possible bonuses. #J-18808-LjbffrRemote work$20 per hour
...company focused on AI and design is seeking a Digital Designer to evaluate AI-generated designs and enhance AI's understanding of user-... ...from $20/hr for general projects and $40/hr for design-specific tasks. Candidates should possess strong design backgrounds and...Remote work$11.5 per hour
...A leading educational institution is offering a part-time remote position as an Online Task Contributor. In this role, you will evaluate and provide feedback on content to enhance search engine results and quality. No prior experience is needed, but proficiency in English...Hourly payPart timeRemote work$20 per hour
...Candidates should possess a strong design background and fluency in English. Compensation starts at $20+ USD/hr for general AI projects and $40+ USD/hr for design-focused tasks. This is a remote, independent contract position for applicants in the United States. #J-18808-LjbffrContract workRemote work- ...Prolific is seeking fluent Thai speakers to act as evaluators for AI language models. You will assess how naturally Thai is spoken in AI... ...ensuring cultural and contextual accuracy. This is a remote, paid task with flexible hours and competitive pay. Applicants should have...Remote workFlexible hours
$30 - $45 per hour
...Web Browsing Evaluator remote $30 - $45/hour pay Required Skills Attention to Detail Data Annotation Web Browsing... ...perform through high-quality, real-world input. Scope of Work Navigate web pages according to detailed prompts and project...For contractorsRemote work$20 per hour
...outputs and help refine design quality. Applicants should have a background in UI/UX and be fluent in English. This is a remote position with project flexibility and competitive pay starting at $20/hr for general projects and $40/hr for design-focused tasks. #J-18808-LjbffrRemote work$60 per hour
...That's where you come in. As a member of DataAnnotation's team, you'll work closely with state-of-the-art AI models on tasks like evaluating AI-generated quantitative analysis, solving technical problems, and providing feedback that directly shapes how these systems...Hourly payFull timeRemote workFlexible hours$25 - $30 per hour
...analytical and detail-oriented independent contractors in New York to help train AI chatbots. You will manage various tasks, write high-quality responses, and evaluate AI models based on guidelines. This role offers flexibility as you can choose projects and work from home,...Hourly payContract workFor contractorsRemote workWork from home$5,083 per month
...will begin on February 11, 2026. This position specializes in evaluation functions under the general supervision of the University Registrar... ...and degree audit programs to assure timely completion of tasks. Accurately enter and update student information in administrative...Work at officeRemote work$20 per hour
...firm in the United States is seeking a Digital Web Designer to evaluate AI-generated designs and enhance the model's understanding of aesthetics... ...and involves both general AI training and design-focused tasks, with compensation starting at $20/hr. Candidates should have a...Remote job$20 per hour
A leading data annotation firm is seeking a Web Designer to evaluate AI-generated designs and enhance their understanding of design principles... ...at $20/hr for general projects and $40/hr for design-specific tasks, with bonus opportunities for exceptional work. #J-18808-Ljbffr...Remote job$20 per hour
...looking for an Experience Designer to improve AI model outputs by evaluating design work, including interfaces and visuals. This role offers... ..., with compensation starting at $20+ USD/hr for general AI tasks and $40+ USD/hr for design-focused work. Applicants must be fluent...Remote jobContract workFlexible hours$70 per hour
...Location: Remote Commitment: 10-40 hrs/week Role Responsibilities Evaluate and critique the performance and accuracy of AI-generated... ...Contribute user-level expertise by simulating real-life scenarios or tasks to gauge model utility and effectiveness. Requirements Have...Contract workRemote work$60 per hour
...development firm seeks experienced quantitative professionals to evaluate AI-generated quantitative analysis and contribute to cutting-... ...systems. This fully remote role offers a flexible schedule, with tasks including solving technical problems and providing impactful...Remote workFlexible hours$125 per hour
...therapist to join our team in providing outpatient psychotherapy and evaluations for ADHD and Autism. This role is ideal for a clinician who... ...and clinical interviews under supervision Support clients navigating co-occurring concerns such as anxiety, depression, trauma,...Hourly payFull timePart timeRemote workFlexible hours$60 per hour
A leading AI research company is seeking quantitative professionals to evaluate AI-generated work and provide valuable feedback on analytical tasks. This fully remote position offers a flexible schedule and competitive pay up to $60 USD/hour. Ideal candidates will have...Remote jobFlexible hours$20 per hour
...Candidates should have a background in design and fluency in English. The position offers flexibility to choose projects and a pay rate starting at $20+ USD/hr for general tasks, with higher rates for design-focused projects. Independent contractors only. #J-18808-LjbffrFor contractorsRemote work$40 per hour
...training company is seeking a Postdoctoral Physics Associate to evaluate AI chatbot outputs and enhance model quality. This remote... ...assess the performance of AI models through complex problem-solving tasks. Flexible project choices and hourly pay starting at $40+ are offered...Hourly payRemote workFlexible hours$30 per hour
...Speakers in Chicago, IL to train AI models. You will complete AI tasks and assess AI performance in Dutch, with pay at $30/hr. Ideal... ...home with competitive rates and direct payment through PayPal. Pass the evaluation, and you can start within 15 minutes. #J-18808-Ljbffr...Remote workWork from home$60 per hour
...company is seeking proficient programmers to advance AI development in a fully remote setting. As part of this team, you'll engage in tasks like solving coding problems and building applications, with a competitive pay rate of up to $60/hour based on performance....Remote workFlexible hours$75k - $110k
...Opportunity to specialize in unique rural and forestland appraisals Stable, respected firm with over 70 years of industry experience A Task and Duty Guide for this position is available upon request. Pay and benefits are commensurate with skills and experience....Temporary workRemote work$19 per hour
...is building realistic, high-fidelity simulated environments to evaluate and train AI models on real-world procurement workflows for a leading... ..., distribution, packaging, receiving, and quality assurance tasks within the Mobile Services Team.This posi... Show more Full-...Hourly payPermanent employmentFull timeTemporary workCasual workSeasonal workWork at officeRemote workWork from homeMonday to Friday$30 per hour
...A prominent AI data company is seeking Advanced Mandarin Speakers for evaluating AI models and completing training tasks. This remote role offers competitive pay of $30/hr for tasks requiring one hour of focused work. Candidates should be fluent in Mandarin with the ability...Remote workWork from homeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Embodied Navigation Task Evaluator [Remote]. Be the first to apply!

