Frontier AI Safety Evaluator & Alignment Specialist
Obsidian
Obsidian is seeking experienced AI Safety Practitioners to evaluate frontier AI models for safety, quality, and alignment on complex topics. You will assess responses, apply safety policies, and help improve model behavior through structured evaluations and feedback. Responsibilities include reviewing misinformation, political persuasion, self-harm, violence, cyber, and biosecurity content while refining rubrics for RLHF and SFT. Collaboration with researchers and safety teams is required. #J-18808-Ljbffr Obsidian
- Mercor is seeking experienced AI Safety Practitioners to evaluate frontier AI models for safety, alignment, and policy compliance across nuanced topics. You will assess AI-generated responses for quality, identify unsafe outputs, and apply robust evaluation rubrics for...Suggested
- Obsidian seeks experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across grey-area topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and...Suggested
- Mercor is seeking experienced AI Safety Practitioners to assess the safety, quality, and alignment of frontier AI models across complex, policy-sensitive topics. You will evaluate AI-generated responses, apply safety policies, and provide structured feedback to improve...Suggested
- Mercor is recruiting experienced musicians to evaluate generative music AI models in partnership with a leading AI lab. You will assess AI-generated music across genres and rate it against detailed quality standards, using Punjabi and English. Ideal candidates have strong...SuggestedHourly payImmediate start
- ...Welo Data is seeking Data Labeling Associates in California to evaluate AI outputs and ensure cultural context and safety in Arabic datasets. This role requires professional-level proficiency in Portuguese (Brazil), a bachelor's degree, and at least 2 years of experience...Suggested
- ...Francisco is hiring Data Labeling Associates for Project Perseus. This role focuses on evaluating Arabic AI systems, requiring professional proficiency in Portuguese and experience in AI safety. Responsibilities include assessing AI outputs, identifying bias, and...Full time
$162.7k - $220.2k
The AWS Data and AI Strategic Partner GTM team supports the world... .... The Senior WW Partner GTM Specialist for Agentic AI will be the driving... ...with AWS service teams to align partner requirements with product... ...their data to advancing the frontier of GenAI, we champion emerging...Local areaWorldwideFlexible hours- Mercor is building a remote red team for AI safety. You will lead adversarial testing of conversational AI models, focusing on jailbreaks, prompt injections, misuse, and bias exploitation. You will generate high-quality human data, annotate failures, classify vulnerabilities...Remote job
- Obsidian evaluates Neuron Kernel Interface (NKI) development tasks used for frontier AI model training and evaluation. The role focuses on assessing CUDA→NKI migration fidelity, Trainium-specific performance optimization quality, and cross-platform numerical-correctness...
$94k - $121k
...Description Job Description Job Title: Senior Safety Specialist Location: San Francisco, CA... ...~ Purpose Driven : ESG focused work aligned with strong values ~ Better Together... ...Performance : By people, strengthened through AI and technology that enhance...Work at office$179k - $242.2k
...help define the future of Go to Market (GTM) at AWS for generative AI (GenAI)? You will be part of the core worldwide GenAI Foundation... ...adoption of our services and solutions with lighthouse Frontier AI model builders across segments and industry verticals. You will...Local areaWorldwideFlexible hours- ...We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will... ...adversarial testing, implementing protective measures, and aligning AI behavior with ethical principles. Responsibilities...
$100k - $120k
...opening for a Reality Capture Specialist Manager in our California... ...models that elevate accuracy and safety throughout the lifecycle of... ...the Reality Capture Team to align priorities and workflowsTrain... ...equipmentApply expert technical skills to evaluate building conditions and...Work at office- ...into clear requirements and evaluation criteria. Partner with Sales... ...offboarding processes that balance safety and revenue needs. Partner... ...to identify high value AI and agent use cases, provide... ...functional partners and drive alignment. Experience partnering with Data...Immediate start
$188k - $275k
...Description CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers,... ...of our cloud platform. Field Engineering aligns closely with internal and customer... ...development. About the role: As a Specialist Field Engineer - Compute Infrastructure...Permanent employmentFull timeContract workTemporary workCasual workWork at officeFlexible hours- Obsidian is looking for expert Evaluators in Finance operations/audit support to review AI-generated work products for accuracy and quality. This remote hourly position requires a minimum of 5 years in finance and fluency in English. Your role will involve evaluating outputs...Remote jobHourly payWork at office
- Obsidian is hiring expert Evaluators in Investment analysis / valuation / credit to review AI-generated work products for accuracy and quality. This remote, hourly position requires deep subject-matter expertise and professional fluency in English to provide structured...Hourly payWork at officeRemote work
- Mercor is seeking expert Evaluators in Humanities / arts / culture to review AI-generated work products for accuracy, rigor, and domain quality. This remote hourly engagement requires deep subject-matter expertise to grade outputs and provide actionable feedback. Candidates...Remote jobHourly payWork at office
- YO AI Labs is seeking a Turkish Bilingual Expert to support a language and AI training project. This contractor role is remote, with flexible hours to evaluate Turkish audio for nativeness, fluency, pronunciation, and intonation. Provide clear feedback in English, justify...Remote jobFor contractorsFlexible hours
- Obsidian is seeking expert Evaluators in Compliance/regulatory response with financial-services AI to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This...Remote jobHourly payFlexible hours
- Obsidian is hiring expert Evaluators in Real estate, hospitality, and events to review AI-generated work for accuracy, rigor, and domain quality. This remote position requires deep expertise and involves grading outputs like documents and presentations. Applicants must...Remote jobWork at office
- Obsidian is seeking expert Evaluators in Legal/compliance to review AI-generated documents, spreadsheets, and slide decks for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. This is a remote, hourly engagement. You will...Remote jobHourly pay
- Mercor is seeking expert Evaluators in Legal/compliance to review AI-generated documents, spreadsheets, and slide decks for accuracy and quality. This is a remote, hourly engagement requiring deep subject-matter expertise and careful judgment. You will assess artifacts...Remote jobHourly payWork at office
$302k - $335k
...business leaders to translate frontier research into real-... ...bring the benefits of AI to organizations... ...roleWe are hiring a Senior Specialist Seller to join the Strategic... ...-facing environments, aligning multiple stakeholders,... ...must be created with safety and human needs at its...Work at officeLocal areaRemote workRelocation packageFlexible hours3 days per week$76k - $136.8k
...discover, create and connect. The Trust & Safety (T&S) team at TikTok helps ensure that... ...animals. Demonstrate operational excellence in evaluating risk, threats, and user privacy,... ...global teams to address emerging risks and align on response strategies. This role requires...Temporary workWork at officeLocal areaWeekend work$100k - $185.6k
...transforming society.The PositionThe Senior EHS Specialist, Associate Biosafety Officer is a hands-... ...is to mitigate risks and ensure the safety of our people, the environment, and our... ...renewals to ensure all documentation aligns with regulatory requirements.Provide expert...Full timeLocal area- Obsidian is seeking expert Evaluators in FP&A / corporate finance to assess AI-generated work products for accuracy and quality. This role entails deep expertise to grade outputs and provide structured feedback. Candidates should have at least 5 years of relevant experience...Remote jobHourly payWork at office
$400 per month
About the Role Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments. The work focuses on realistic infrastructure engineering...- Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. You will help evaluate frontier AI coding models by performing infrastructure engineering tasks and reviewing model-generated implementations on cloud platforms, Kubernetes, and...
$50 - $75 per hour
A leading tech company based in Australia is seeking an AI Model Evaluator on a contract basis. The role involves evaluating AI-generated responses, writing prompts, and providing justifications based on specific criteria. Ideal candidates will hold a Master's degree in...Hourly payContract work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Frontier AI Safety Evaluator & Alignment Specialist. Be the first to apply!
- quality evaluator San Francisco, CA
- work from home web search evaluator San Francisco, CA
- evaluator San Francisco, CA
- ai evaluator San Francisco, CA
- education evaluator San Francisco, CA
- social media evaluator San Francisco, CA
- ai scientist San Francisco, CA
- ai data scientist San Francisco, CA
- safety analyst San Francisco, CA
- safety attendant San Francisco, CA


