Government Evaluator - Domain Expert - AI Trainer
$80 - $120 per hourMercor
Job Description
Job Description
About the job
Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .
Position: Government / public administration Evaluator
Type: Contract
Compensation: $80–$120/hour
Location: Remote
Role Responsibilities
- Evaluate AI-generated artifacts against domain-specific quality rubrics.
- Identify factual, aesthetic, and presentation errors in documents, spreadsheets, and slide decks.
- Provide clear, structured written feedback to improve AI model outputs.
- Collaborate with subject matter experts to ensure consistency and quality.
- Work independently and asynchronously to meet deadlines while enhancing AI model performance.
Qualifications
Must-Have
- 5+ years of relevant professional experience in Government / public administration.
- Native or professional fluency in English .
- Highly proficient in Microsoft Office and Google Workspace , especially Slides .
Preferred
- Master's or higher from a reputable institution.
Application Process (Takes 20–30 mins to complete)
- Upload resume
- AI interview based on your resume
- Submit form
Resources & Support
- For details about the interview process and platform information, please check:
- For any help or support, reach out to: View email address on ziprecruiter.com
PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.
#hiringmercor- Obsidian is seeking expert Evaluators in Biology/environmental science to review and assess AI-generated work products for accuracy and quality. In this remote, hourly role... ...and presentations, ensuring they meet domain-specific standards. A minimum of 5 years of relevant...SuggestedRemote jobHourly pay
$90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...genuinely correct replications from those that merely look correct. Evaluate responsive behavior and semantic quality, ensuring proper use of...SuggestedContract workSummer workLocal areaRemote work$60 - $70 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Location: Remote Role Responsibilities Evaluate AI-generated responses for safety,... ...cyber, biosecurity, and other sensitive domains. Apply and refine evaluation rubrics for...SuggestedContract workSummer workRemote work$80 - $120 per hour
...technical talent with leading AI research labs. Headquartered in... .../ ROI / revenue economics Evaluator Type: Contract Compensation... ...-generated artifacts against domain-specific quality rubrics.... ...Collaborate with subject matter experts to ensure consistency and...SuggestedContract workSummer workWork at officeRemote work$60 per hour
...technical talent with leading AI research labs. Headquartered in... ...and objective rubrics. Evaluate AI models on visual document... ...-following in the Education domain. Develop and refine evaluation... ...Collaborate with subject matter experts to ensure task relevance and...SuggestedContract workSummer workRemote work$120 - $175 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Larry Summers , and Jack Dorsey . Position: CNC Machining Expert Type: Contract Compensation: $120–$175/hour Location...Contract workSummer workRemote work- Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback. Candidates should have at least 5 years of relevant experience...Remote jobHourly payWork at office
- Mercor is hiring experienced musicians to evaluate generative music AI models, in partnership with a leading AI lab. You will assess AI-generated music across genres and rate it against detailed quality standards, working in Hindi and English. Bring a background in songwriting...
- About the roleWe're building a high-quality evaluation dataset for CNC manufacturing and are looking for experienced CNC machinists to help author and validate grading rubrics for CNC machining work. You'll bring real production-floor judgment to determine whether a machining...
- ...Clinical Medicine Domain Expert Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels... ...practiced. Design challenging clinical tasks and evaluation sets, and help build medicine-specific skills and tools...Full timeContract workPart timeFreelanceLive inRelocationRelocation package
$50 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Spanish (Spain) Audio Generalist Evaluator Expert Type: Contract Compensation: $50/hour Location:...Contract workSummer workRemote work$100 - $150 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Jack Dorsey . Position: B2B Sales Expert Type: Contract Compensation: $... ...asynchronously to meet deadlines while improving evaluation processes. Qualifications Must-Have...Contract workSummer workRemote work$50 - $60 per hour
...technical talent with leading AI research labs.... ...Jack Dorsey . Position: Government & Public Policy Expert Type: Contract Compensation... ..., RFP/solicitation, and evaluation criteria. Develop... ...help us understand your sub-domain strengths. Qualified candidates...Contract workSummer workLocal areaRemote work$15 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Jack Dorsey . Position: Music & Lyrics Expert - Malayalam Type: Contract Compensation... ...+ hours/week Role Responsibilities Evaluate AI-generated music across various genres...Contract workSummer workImmediate startRemote workFlexible hours$120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...'Angelo , Larry Summers , and Jack Dorsey . Position: Domain Expert – Legal (Lexis+ Research) Type: Contract Compensation:...Remote jobContract workSummer work$80 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... .... Position: STEM PhDs and Technical Domain Experts Type: Contract Compensation:... ...developing difficult problems in your domain. Evaluate and improve AI model performance...Remote jobContract workSummer work$400 per month
...About the Role Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments. The work focuses on realistic infrastructure engineering...- 1. Role Overview Mercor is partnering with a leading AI research organization to engage experienced sales engineers for a project focused on evaluating how well AI systems perform real-world technical sales work. Rather than producing deliverables yourself, you will define...
- ...s leading research accelerator for frontier AI labs and a trusted partner for global enterprises... ...are seeking a highly qualified Video Games Domain Reviewer to support the quality assurance of Large Language Model (LLM) evaluation projects. In this role, you will review...Contract workFor contractors
- ...Care is building the leading AI-native platform for family-led... ...helps health plans and government partners better understand, verify... ...internal platforms. Design evaluation frameworks, datasets, metrics... ...and operations-heavy domains. Experience with human-in-...Full time
$220.5k - $245k
...data in advertising technology domains (ad delivery, ranking,... ...with incomplete information. Expert‑level SQL and Python . Strong... ...experimentation. Working with Data AI tools to establish greater self... ...Voluntary Self Identification For government reporting purposes (EEO-1), we...Full timeTemporary workLocal area$120 per hour
...Prolific, located in San Francisco, is looking for Licensed Pharmacists to train and evaluate AI models. You will guide these systems by reviewing AI-generated pharmaceutical scenarios and providing your expertise. With competitive pay rates of up to $120 USD per hour...Hourly payRemote workWork from homeFlexible hours$225k - $320k
...organizations, state and local and governments, federal agencies, and... ...unprecedented speed and accuracy. Our AI-enabled platform turns siloed... ...intelligence is applied, evaluated, and operationalized across... ...data platforms.Background in domains where trust, explainability,...Local area- ...Care is building the leading AI-native platform for family-led... ...helps health plans and government partners better understand, verify... ..., how its behavior should be evaluated, where deterministic controls... ...Care's operational and clinical domains and use that understanding to...Full timeWork experience placementRelocation
$207.48k - $385.32k
...across Commercial, Medical, and Government Affairs (CMG) to make... ...leveraging data, analytics, and AI/ML to enable fast, targeted actions... ...trusted, objective advisor and expert, recommending critical... ...Marketing, and Medical Affairs domains. Operating with a high degree...Full timeLocal areaImmediate start3 days per week$30 per hour
A leading platform for AI training is seeking advanced Japanese speakers to join as Domain Expert participants. You will complete various AI training tasks and evaluate AI responses in Japanese. Compensation is competitive at $30 per hour, with flexible working hours from...Remote jobHourly payWork from homeFlexible hours$40 per hour
Prolific is seeking AI Trainers with advanced SQL development skills to train and evaluate AI models. The role requires strong attention to detail, ability to work... .... Successful candidates will join Prolific as Domain Expert participants, earning approximately $40/hour for...Remote jobFlexible hours$30 per hour
AI Trainer - Advanced Japanese Fluency Prolific is building the biggest pool of quality human... ...Japanese speakers to help train and evaluate cutting‑edge AI models. If you have the... ...will be invited to join Prolific as a Domain Expert participant, where you’ll get paid to train...Remote jobHourly paySelf employmentWork from homeFlexible hours- Cincinnatus LLC is hiring an experienced counsel-level lawyer to work with a leading AI lab's research and program management teams, shaping how frontier models reason about legal work. This is a full-time W-2 employment position with the opportunity to be placed at a...Full timeRelocationRelocation package
- CGS Federal (Contact Government Services) seeks a Senior eDiscovery Analytics Lead to apply legal experience in support of a large federal agency. Lead processing, analytics, and production tasks using Relativity, while coordinating with attorneys and clients to meet ESI...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Government Evaluator - Domain Expert - AI Trainer. Be the first to apply!
- work from home web search evaluator San Francisco, CA
- social media evaluator San Francisco, CA
- evaluator San Francisco, CA
- quality evaluator San Francisco, CA
- education evaluator San Francisco, CA
- ai evaluator San Francisco, CA
- technology expert San Francisco, CA
- subject matter expert San Francisco, CA
- fulfillment expert San Francisco, CA
- guest service support expert San Francisco, CA



