Czech Language AI Safety Evaluator [Remote]
$43 - $47 per hourSaidGig
- Remote job
Role Overview
Help improve how advanced AI models respond to sensitive topics in Czech. In this remote, hourly role, you will use Czech language fluency and cultural judgment to evaluate model behavior and support safer, more reliable interactions. Training is provided, and prior AI or machine learning experience is not required.
Key Responsibilities
- Create expert-level Czech prompts covering a range of sensitive subject areas.
- Use structured guidelines to classify prompts and conversations.
- Identify adversarial wording and escalation patterns, then document the reasoning behind your assessments.
- Evaluate and strengthen AI model handling of sensitive Czech-language content using linguistic and cultural expertise.
Qualifications
- Native or near-native Czech fluency and business-level written English.
- Bachelor''s degree completed or currently in progress.
- Strong written reasoning skills and close attention to detail.
- Sound judgment when working with sensitive and dual-use information.
Preferred Background
- Experience reviewing, grading, or red-teaming written or technical content.
- Experience in trust and safety, content moderation, policy evaluation, or adversarial testing.
Work Terms
- Remote hourly engagement.
- Immediate start date.
- Applicants located in the Czech Republic or elsewhere in Eastern Europe are preferred, though candidates based in other locations are welcome.
Compensation
$43 to $47 per hour.
$43 - $47 per hour
...Overview Help improve how advanced AI models respond to sensitive topics in Czech. In this remote, hourly role, you will use Czech language fluency and cultural judgment to evaluate model behavior and support... .... Experience in trust and safety, content moderation, policy...Czech language skillsRemote jobHourly payImmediate start$93.6k - $114.4k
...applications and next steps. Our partner is looking for an AI Safety Policy Evaluator, Violence & Threats based in United States. This role... ...produce concise, well-supported rationales referencing policy language and relevant conversation details. Develop and refine...SuggestedHourly payFull timeRemote workMonday to Friday$45 - $55 per hour
...non-engineering content-policy evaluation role. Applicants must... ...threat assessment, trust and safety, content moderation, or closely... ...In 2025, we started Handshake AI and built the fastest-growing... ...concise rationales that cite policy language and conversation details Write...SuggestedRemote jobMonday to FridayShift work- ...Role Overview Use your scientific expertise and Czech language skills to help improve the safety of advanced AI systems handling specialized technical subjects.... ...prompts covering specialized scientific topics. Evaluate and annotate model responses for scientific...Czech language skillsHourly payImmediate startRemote work
$65 per hour
Prolific is seeking registered nurses in Las Vegas, NV, to assist in training AI models. You will review AI responses, rate their accuracy and safety, and write feedback to enhance AI learning. Candidates must be verified registered nurses, have recent clinical experience...SuggestedHourly paySelf employmentWork from homeFlexible hours- ...Educational Safety AI Evaluator is a remote evaluation track for reviewing educational safety ai evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling...Remote jobHourly payFor contractors10 hours per week
$18 - $22 per hour
...Role Overview Help improve the safety of advanced AI models by applying Thai language fluency and cultural judgment to evaluate how models respond to sensitive topics in Thai. Training is provided, and prior AI or machine learning experience is not required. Key Responsibilities...Remote jobHourly payImmediate start- ...AI Safety and Red-Team Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios... ...-ML research. Multilingual fluency for cross-language attack testing. Skills Adversarial prompting Red...Remote jobHourly payFor contractors10 hours per week
- ...Mental Health Clinical AI Evaluator is a remote clinical-review track for evaluating AI outputs that touch clinical review. Reviewers grade... ..., dosing logic, and guideline adherence; flag patient-safety issues; and document the corrected clinical reasoning so the modeling...Remote jobHourly payFor contractors10 hours per week
$43 - $47 per hour
...Role Overview Help evaluate and strengthen how advanced AI systems respond to sensitive topics in Spanish. This role combines Spanish language fluency, cultural judgment, and careful written... ...content. Background in trust and safety, content moderation, policy...Remote jobHourly payImmediate start$48 - $52 per hour
...Role Overview Help improve the safety of advanced AI systems by applying Chinese language fluency and cultural judgment to evaluate how models respond to sensitive topics. Training is provided, and prior AI or machine learning experience is not required. Key Responsibilities...Remote jobHourly payImmediate start- ...building the tools, components, and safety nets that the entire... ...resilient services. Leverage AI-powered tools and coding agents... ...; open to leveraging other languages when appropriate. Solid understanding... ...Hybrid work mode (Prague, Czech Republic / Nicosia, Cyprus),...Czech language skillsFull time
- ...Seeking a full-time Remote AI Research Evaluator with a PhD in Quantitative Finance to assess and enhance AI models' capabilities in financial reasoning and quantitative analysis through flexible, contract-based work. Key responsibilities Assessing the factuality and...Full timeContract workRemote workFlexible hours
- ...BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance...Remote work
$20 per hour
A tech company specializing in AI is hiring a Digital Web Designer. In this remote role, you will evaluate AI-generated designs and provide feedback to enhance the model’s understanding of aesthetics. An ideal candidate will have a strong background in UI/UX design and...Remote workFlexible hours- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Work at officeRemote work
$14.5 per hour
A technology company is seeking an AI Web Search Evaluator to enhance the quality of search engine results. This flexible, remote, part-time role focuses on analyzing search performance and providing feedback to improve algorithms. Ideal candidates will have strong analytical...Hourly payPart timeRemote workFlexible hours$20 - $26 per hour
Prolific is seeking fluent Kannada speakers to act as evaluators. You will assess how naturally and authentically AI captures Kannada speech, by listening to audio clips and comparing text and voice. This fast-paced project pays $20-26 per hour and may require about one...Remote jobHourly payWork from homeFlexible hours- Obsidian is looking for expert Evaluators in Finance operations/audit support to review AI-generated work products for accuracy and quality. This remote hourly position requires a minimum of 5 years in finance and fluency in English. Your role will involve evaluating outputs...Remote jobHourly payWork at office
- Handshake is seeking an AI Image Evaluator to assess prompts and generated images for quality, accuracy, and policy compliance. You will weigh visual details, compare outputs, and explain your reasoning to help train evaluation models. Remote US role with flexible scheduling...Remote jobFlexible hours
- Mercor is seeking experts in Spreadsheet QA and workbook maintenance to review AI-generated documents, spreadsheets, and slide decks for accuracy and quality. This is a remote, hourly engagement. Ideal candidates have 5+ years in Spreadsheet QA, fluent English, and strong...Remote jobHourly payWork at office
- Alignerr is seeking a Population Health Informaticist to help train and evaluate health AI on large-scale datasets and public health strategy. You will review AI-generated health insights, assess data-driven metrics, and provide structured feedback that reflects disparities...Remote jobHourly payContract workFlexible hours
- YO AI Labs is seeking a Turkish Bilingual Expert to support a language and AI training project. This contractor role is remote, with flexible hours to evaluate Turkish audio for nativeness, fluency, pronunciation, and intonation. Provide clear feedback in English, justify...Remote jobFor contractorsFlexible hours
- YO AI Labs is seeking an Investment & Finance Expert to work remotely as a contractor. The role focuses on evaluating AI outputs related to valuation, financial modeling, markets, and investments, and crafting expert prompts plus reference answers to improve model reasoning...Remote jobFor contractors
- MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers...Remote jobContract workTemporary workImmediate start
$20 - $30 per hour
A remote-focused technology firm is looking for an individual proficient in evaluating large language model outputs. The role involves assessing AI systems, reviewing workflows, and providing actionable feedback to enhance product quality. Ideal candidates will have strong...Remote jobHourly pay- Get notified about new Github jobs in United States . Staff Software Engineer, Copilot Code Review Principal Software Engineer - GitHub Platform & Enterprise Principal Software Engineer - GitHub Platform & Enterprise Software Engineer, Research Developer Productivity Staff...Remote jobFor contractors
- YO AI Labs seeks an experienced Developer & Infrastructure Expert to evaluate AI-powered workflows across software development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands, configurations, and workflows while applying...Remote job
- Obsidian is seeking expert Evaluators for a remote, hourly role focused on reviewing AI-generated work products for quality and accuracy. The successful candidates will leverage their subject matter expertise to assess various documents, spreadsheets, and slide decks....Remote jobHourly payWork at office
- Productive Playhouse is building an Albanian-speaking evaluator pool for AI chatbot testing. As an AI Evaluator, you’ll interact with AI models to assess capabilities, safety, and usefulness, providing data to improve models. Tasks come in batches with flexible hours and...Remote jobFreelanceFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Czech Language AI Safety Evaluator [Remote]. Be the first to apply!



