Autonomous Driving Edge Case Task Evaluator [Remote]
AuraOne Human Data
- Remote job
Autonomous Driving Edge Case Task Evaluator is a remote evaluation track for reviewing autonomous driving edge case task evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Why this role matters
AI data reviewers help turn autonomous driving edge case task evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Responsibilities
- Evaluate autonomous driving edge case task evaluation model outputs against a versioned rubric and assign severity tags for Autonomous Driving Edge Case Task Evaluator assignments.
- Compare paired responses and pick the stronger answer with a written rationale.
- Label hallucinations, instruction-following failures, and unsafe content with structured tags.
- Capture ambiguous prompts and route them back to the program team for rubric updates.
- Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
- Document recurring failure modes so the modeling team can target them in the next training run.
Qualifications
- Prior evaluation, annotation, or human-rater experience on autonomous driving edge case task evaluation or adjacent content for Autonomous Driving Edge Case Task Evaluator work.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that names the issue and the rubric clause being applied.
- Strong attention to detail and the ability to flag when a prompt itself is the problem.
- Reliable async availability for at least 10 hours per week.
Example tasks
- Compare two autonomous driving edge case task evaluation model responses to the same prompt and pick the stronger one with rationale.
- Tag an unsafe response with the correct policy category and severity.
- Audit a 50-row batch for rubric consistency and report drift to the program lead.
- Propose a rubric clarification after spotting a recurring failure mode.
Nice to have
- Background in linguistics, content moderation, or trust & safety review.
- Experience with inter-rater agreement metrics and calibration cycles.
- Domain expertise that lets you spot subject-matter errors automated checks miss.
Skills
- Model output evaluation
- Rubric-based annotation
- Severity tagging
- Inter-rater calibration
- Autonomous Driving Edge Case Task evaluation
- Robotics evaluation
- Embodied AI
- Safety review
- Autonomous
- Driving
- Edge
Work model
Remote — US-eligible. Remote · Independent specialist contractor. Employment type: CONTRACTOR. Applicants must be authorized to work from US.
Compensation
Hourly rate confirmed after the interview process.
Application process
Apply through AuraOne's specialist intake for role-specific routing and review. Final project scope, schedule, and contractor terms are confirmed before placement.
- ...Surgical Robotics Task Evaluator is a remote review track for evaluating robotics review outputs... ...train on it. Adjudicate disputed cases against published standards or firm operating... ..., supervising, or evaluating robotic, autonomous, or teleoperated systems for Surgical...SuggestedRemote jobHourly payFor contractors10 hours per week
$50 per hour
...project by transcribing, annotating, and evaluating Hindi audio and video to help train and... ...quality responses. Identify linguistic edge cases such as colloquialisms, grammatical... ...improve model performance on Hindi audio tasks. Participate in quality assurance and...SuggestedHourly payTemporary workRemote work10 hours per week$30 per hour
...Commitment: 10-40 hours/week Role Responsibilities Evaluate outputs from large language models and autonomous agent systems using defined rubrics and quality... ...benchmarking criteria consistently while identifying edge cases and recurring failure patterns. Provide structured...SuggestedRemote jobHourly payContract work$85 per hour
...Responsibilities ~Use frontier AI coding agents to complete and evaluate complex engineering tasks. ~Review model-generated mobile application code for... ..., maintainability, and performance. ~Identify bugs, edge cases, and failure modes in model outputs. ~Compare outputs...SuggestedContract workPart timeSummer workRemote work$65k - $75k
Piper Companies is seeking a Remote Utilization Management/Review Nurse to evaluate member cases, ensuring medical necessity and appropriate healthcare service utilization. The role requires clinical expertise and collaboration with providers to make high-quality care decisions...SuggestedRemote job$60 per hour
...the DataAnnotation team and contribute to developing cutting-edge AI systems, while enjoying the flexibility of remote work... ...team, you'll work closely with state-of-the-art AI models on tasks like evaluating AI-generated quantitative analysis, solving technical problems...Hourly payFull timeRemote workFlexible hours$60 per hour
...the DataAnnotation team and contribute to developing cutting-edge AI systems, while enjoying the flexibility of remote work... ...team, you'll work closely with state-of-the-art AI models on tasks like evaluating AI-generated security content, solving technical security problems...Remote jobHourly payFull timeFlexible hours$60 per hour
...development firm seeks experienced quantitative professionals to evaluate AI-generated quantitative analysis and contribute to cutting-edge AI systems. This fully remote role offers a flexible schedule, with tasks including solving technical problems and providing impactful...Remote workFlexible hours- ...Travel Research Task Evaluator is a remote review track for evaluating AI outputs across travel research task research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct...Remote jobHourly payFor contractors10 hours per week
- ...Sensitive Advice Boundary Risk Evaluator is a remote evaluation track for reviewing sensitive... ...Reviewers compare paired outputs, label edge cases, and write the kind of structured... ...at least 10 hours per week. Example tasks Compare two sensitive advice boundary...Remote jobHourly payFor contractors10 hours per week
$80 - $120 per hour
...Incident management / reliability / SRE Evaluator is a remote engineering review track for... ...actually compiles, passes tests, and handles edge cases. AuraOne pairs experienced engineers... ...experience is a plus. Example tasks Reproduce a generated engineering solution...Remote jobFor contractors10 hours per week$50 - $60 per hour
...work Push the models with complex, real-world scenarios and edge cases to see where their reasoning holds up – and where it doesn’t.... ...Responsibilities Give AI chatbots diverse and complex problems and evaluate their outputs Evaluate the quality produced by AI models...Hourly payFull timeContract workPart timeWork experience placementRemote workFlexible hours- ...looking for detail-oriented Image Quality Evaluator for a multilingual AI data Annotation... ...productivity. Report data quality issues, edge cases, and annotation inconsistencies to the... ..., annotation, or linguistic tasks is preferred Access to a laptop/desktop...Contract workRemote workWork from homeMonday to FridayDay shift
$55k
...Statutes, Administrative rules and tax court cases pertaining to property tax Working... ...ability to make progress on multiple duties and tasks simultaneously and work in high-pressure... ...applications suchas Gmail, Sheets, Docs, and Drive. Abilities Ability to clear a...Temporary workWork at officeRemote workFlexible hours$20 per hour
...firm in the United States is seeking a Digital Web Designer to evaluate AI-generated designs and enhance the model's understanding of aesthetics... ...and involves both general AI training and design-focused tasks, with compensation starting at $20/hr. Candidates should have a...Remote work$20 per hour
A leading data annotation firm is seeking a Web Designer to evaluate AI-generated designs and enhance their understanding of design principles... ...at $20/hr for general projects and $40/hr for design-specific tasks, with bonus opportunities for exceptional work. #J-18808-LjbffrRemote work$20 per hour
...looking for an Experience Designer to improve AI model outputs by evaluating design work, including interfaces and visuals. This role offers... ..., with compensation starting at $20+ USD/hr for general AI tasks and $40+ USD/hr for design-focused work. Applicants must be fluent...Contract workRemote workFlexible hours$20 per hour
...Candidates should possess a strong design background and fluency in English. Compensation starts at $20+ USD/hr for general AI projects and $40+ USD/hr for design-focused tasks. This is a remote, independent contract position for applicants in the United States. #J-18808-LjbffrContract workRemote work$14.5 per hour
...improving online search experiences? Join our team as a Web Search Evaluator and help shape the future of search engines from the comfort... ...from home anywhere in the U.S. while contributing to a cutting-edge project. Project Details: Pay Rate: $14.50 per hour...Bi-weekly payHourly payPart timeImmediate startRemote workWork from homeFlexible hours$20 per hour
...is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’... .... The position offers flexible hours and pays $20+ USD/hr for general projects and $40+ for design-focused tasks. #J-18808-LjbffrRemote workFlexible hours- ...This role will focus on AI-related projects that involve prompt evaluation and multimedia content understanding. Ideal candidates will... ...reviewers or evaluators. Enjoy rapid payments, access to cutting-edge technology, and the possibility to work on innovative AI developments...FreelanceRemote work
$11.5 per hour
...A leading educational institution is offering a part-time remote position as an Online Task Contributor. In this role, you will evaluate and provide feedback on content to enhance search engine results and quality. No prior experience is needed, but proficiency in English...Hourly payPart timeRemote work- ...leading tech organization is seeking candidates aged 18 to 19 to evaluate North American teen humor through structured rating and... ...0 hours per week with part-time overlap with PST. Join a cutting-edge project and contribute to advanced AI training. #J-18808-LjbffrPart timeFor contractorsFreelanceRemote workFlexible hours
- ...remote freelance support in AI-related projects focusing on prompt evaluation, video understanding, and text review. The ideal candidates... ...reviewers. This role offers the opportunity to work on cutting-edge AI developments and receive rapid payments without invoicing. #J...FreelanceRemote work
- ...The Mental Health Evaluator (MHE) is a critical role dedicated to delivering comprehensive, clinically appropriate, culturally competent, and trauma-informed patient behavioral health interventions in the hospital Emergency Department. Responsibilities include conducting...
$25 - $30 per hour
...analytical and detail-oriented independent contractors in New York to help train AI chatbots. You will manage various tasks, write high-quality responses, and evaluate AI models based on guidelines. This role offers flexibility as you can choose projects and work from home,...Hourly payContract workFor contractorsRemote workWork from home$40 per hour
...development company is seeking experienced quantitative professionals to evaluate AI-generated work and help shape future AI systems. This fully... ...skills, and a strong education in a relevant field. Join the team to contribute to cutting-edge AI development. #J-18808-LjbffrHourly payRemote workFlexible hours$40 per hour
...A forward-thinking AI team is seeking quantitative professionals to evaluate and improve cutting-edge AI systems. The role involves working on AI-generated analyses and providing critical feedback for model enhancement. Candidates should have a strong quantitative background...Hourly payRemote workFlexible hours$40 per hour
...A forward-thinking AI solutions company is seeking experienced quantitative professionals to evaluate AI-generated analyses and contribute to the development of cutting-edge AI systems. This fully remote role offers flexible scheduling and competitive pay starting at...Hourly payRemote workFlexible hours$55 per hour
...seeking to hire a part-time (15-20 hrs. per week) fee-for-service evaluator. Founded in 1985, the Forensic Services Team has a long-... ...Requirements/Description Evaluators carry at least two evaluation cases at a time with the goal of completing a minimum of twelve...Hourly payPart timeRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Autonomous Driving Edge Case Task Evaluator [Remote]. Be the first to apply!





