Frontier AI Safety Evaluator & Policy Auditor
Obsidian
Obsidian is seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback. Responsibilities include evaluating responses for safety, factual accuracy, and policy compliance; reviewing high-risk content; and collaborating with AI researchers on ongoing evaluation initiatives. #J-18808-Ljbffr Obsidian
- AIUC is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models on policy-sensitive topics. You will assess AI-generated responses and provide structured feedback to improve model behavior. Join a collaborative team...Policy
- Mercor seeks experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations...Policy
- We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve...PolicyWorldwide
$400 per month
Obsidian is seeking contributors for a Frontier Code Agents project, focused on evaluating AI coding models in fraud and risk engineering. Candidates will use AI coding tools to handle complex tasks and provide technical assessments. The role requires 2+ years of experience...Suggested- ...lead governance research and risk modeling for the Frontier Safety Framework (FSF). You will drive research, model AI severe risks, and represent the program in... ...processes. The role emphasizes collaboration with policy teams, publication of safety reports, and advancing...Policy
- ...lead governance research and risk modeling for the Frontier Safety Framework (FSF). This role develops risk assessments for frontier AI and informs safeguards and deployment decisions. You will collaborate with policy and governance teams, publish safety reports, and guide...Policy
$119k - $299.93k
...cybersecurity measures, data and AI systems, and their associated... ...assurance projects focused on evaluating clients' digital environments,... ...Certified Information Systems Auditor (CISA) certificationWhat Sets... ...forth within the following policy: more about how we work: only...PolicyFull timeH1b$100k - $110k
...currently looking for a Senior IT Auditor to support our Internal Audit... ...with initial guidanceLeverage AI tools effectively to support... ...peers and market benchmarks evaluated for the scope and responsibilities... ...Media Affirmative Action policy statement.SummaryLocation: New...PolicyFull time$99k - $232k
...their internal audit functions, leveraging AI and other risk technologies to address a... ...except as set forth within the following policy: more about how we work: only those... ...collaborating closely with team members. We evaluate these factors thoughtfully to establish a...PolicyFull timeH1b$55.07k - $69.67k
...Design and Construction, Division of Safety and Site Support seeks a Safety Auditor. The selected candidate will be... ..., and enforcing DDC safety policies and procedures in alignment with... ...equivalency verification from an approved evaluation service. A list of providers (...PolicyPermanent employmentFull timeFor contractorsWork experience placementH1bLocal areaVisa sponsorship$99k - $252.45k
...cybersecurity measures, data, and AI systems. Within our Assurance... ...- Guiding teams in the evaluation of digital environments, including... ...Certified Information Systems Auditor (CISA) certificationWhat Sets... ...set forth within the following policy: more about how we work: only...PolicyFull timeH1b- ...York Department of Design and Construction, Division of Safety and Site Support, seeks a Safety Auditor to oversee safety programs on construction sites. You... ...safety audits, identify hazards, and enforce safety policies in alignment with OSHA, DOT, DOB, and MUTCD....Policy
- AuraOne is seeking an Expert Project Manager for a remote review track to evaluate AI outputs across program management workflows. Reviewers assess workflow accuracy, policy adherence, and stakeholder fit while flagging operational risk and documenting the right next step...PolicyRemote job
$62k - $100k
...continuously invest in innovative ideas, such as AI-enabled insights and technology-powered... ...of security areas such as Auditing, Policy, Database Security, Firewall Design and... ...committed to a merit-based hiring process, evaluating all candidates consistently using...PolicyLocal areaWorldwide$250k
...Audit (IA) team as the Chief Auditor for Artificial Intelligence. This... ...be at the front‑line of the AI revolution, providing critical... ...development lifecycle. Identify, evaluate, and incorporate coverage for... ...at Citi. View Citi’s EEO Policy Statement and the Know Your Rights...Policy$113.2k - $164.05k
...risk assessment, we’re advancing AI to move from insight to action... ...and Institute of Internal Auditors standards Translate technical... ...issue closure, including evaluating action plans, independently reviewing... ...here to view our full EEO policy statement . Click here for more...PolicyFull timeWork at officeWorldwide$116.72k - $175k
DescriptionClinical Revenue Auditor-CDM Patient Financial Services-Corporate-Full-Time-Days- Hybrid.The Clinical Revenue Auditor for the... ...privacy and security standards.Conform to the established policies/ procedures/ processes/ Standards of Behavior.Performs other duties...PolicyFull timeTraineeshipLocal area- Mercor partners with a leading AI lab to train frontier models on insurance reasoning data. We seek Multiline P&C Underwriters to design risk scenarios, evaluate model outputs, and shape AI reasoning about insurability, pricing, and coverage terms. Candidates should have...
$115k - $150k
...capabilities are used responsibly, ethically, and effectively.As Lead Auditor - Data & AI, you will deliver independent assurance over some of the... ..., Asia, Europe, and the Middle East. As part of our New Frontier strategy, MetLife is building an AI-enabled, people-centered...Full timeTemporary workWork at officeLocal areaRelocation package3 days per week$365k
...reliable, interpretable, and steerable AI systems. We want AI to be safe and... ...committed researchers, engineers, policy experts, and business leaders working... ...to alignment, interpretability, and safety, each operating at the frontier of AI development. As a Technical Program...PolicyWork at officeFlexible hoursShift work- ...Agent Robustness to advance safe and aligned AI agents. You will contribute to evaluating risks, building testing harnesses, and prototyping... ...techniques, and publishing results to shape policy and industry practices in frontier AI. #J-18808-Ljbffr United States Digital...Policy
- We're hiring AI‑Native Auditors Early‑career and already living in AI tools? We'll hand you our platform and let you out‑produce an entire... ...with reviewers and partners through sign‑off Push the product frontier — you'll be one of our heaviest users, and your feedback...
$230k - $325k
...Data Scientist, SafetyOpenAI's Safety teams work to ensure our products are safe, trusted, and resilient as frontier AI systems scale globally. We tackle some of the company'... ...operating at the intersection of product, safety, policy, and research.As a Data Scientist, Safety,...Policy$119.8k - $234.7k
...MicrosoftOverviewAt Microsoft AI, we are on a mission to... ...’s most capable AI frontier models, pushing the... ...preparedness, multi-agent safety, traditional frontier... ...governance, including technical evaluation and mitigation efforts,... ...Responsible AI, Public Policy, red teaming, and...PolicyOngoing contractWork at officeLocal area- ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers...Contract workTemporary workImmediate startRemote work
$405k
...interpretable, and steerable AI systems. We want AI to be safe... ...committed researchers, engineers, policy experts, and business leaders... ..., how reliably we can run safety experiments, and how effectively... ..., reliable infrastructure and frontier capabilities can go hand in hand...PolicyFull timeInternshipWork at officeRemote workVisa sponsorshipFlexible hours- We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing... ..., uncover model weaknesses, and evaluate AI behavior across complex, high... ...behaviours, hallucinations, and policy failures. Evaluate model...Policy
- ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional...Work at officeRemote work
- ...BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance...Remote work
$400 per month
About the Role Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments. The work focuses on realistic data engineering workflows...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Frontier AI Safety Evaluator & Policy Auditor. Be the first to apply!
- quality evaluator New York, NY
- clinical evaluator New York, NY
- work from home web search evaluator New York, NY
- evaluator New York, NY
- program evaluator New York, NY
- ai evaluator New York, NY
- education evaluator New York, NY
- social media evaluator New York, NY
- information system auditor New York, NY
- lease auditor New York, NY



