Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Frontier AI Safety Evaluator & Policy Auditor

Obsidian

Obsidian is seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback. Responsibilities include evaluating responses for safety, factual accuracy, and policy compliance; reviewing high-risk content; and collaborating with AI researchers on ongoing evaluation initiatives. #J-18808-Ljbffr Obsidian

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Frontier AI Safety Evaluator & Policy Auditor in New York, NY vacancy
  • AIUC is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models on policy-sensitive topics. You will assess AI-generated responses and provide structured feedback to improve model behavior. Join a collaborative team... 
    Policy

    Dorado

    New York, NY
    22 hours ago
  • Mercor seeks experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations... 
    Policy

    Mercor

    New York, NY
    1 day ago
  • We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve... 
    Policy
    Worldwide

    Obsidian

    New York, NY
    4 days ago
  • $400 per month

    Obsidian is seeking contributors for a Frontier Code Agents project, focused on evaluating AI coding models in fraud and risk engineering. Candidates will use AI coding tools to handle complex tasks and provide technical assessments. The role requires 2+ years of experience... 
    Suggested

    Obsidian

    New York, NY
    22 hours ago
  •  ...lead governance research and risk modeling for the Frontier Safety Framework (FSF). You will drive research, model AI severe risks, and represent the program in...  ...processes. The role emphasizes collaboration with policy teams, publication of safety reports, and advancing... 
    Policy

    Google

    New York, NY
    2 days ago
  •  ...lead governance research and risk modeling for the Frontier Safety Framework (FSF). This role develops risk assessments for frontier AI and informs safeguards and deployment decisions. You will collaborate with policy and governance teams, publish safety reports, and guide... 
    Policy

    Google DeepMind

    New York, NY
    22 hours ago
  • $119k - $299.93k

     ...cybersecurity measures, data and AI systems, and their associated...  ...assurance projects focused on evaluating clients' digital environments,...  ...Certified Information Systems Auditor (CISA) certificationWhat Sets...  ...forth within the following policy: more about how we work: only... 
    Policy
    Full time
    H1b

    PwC

    New York, NY
    1 day ago
  • $100k - $110k

     ...currently looking for a Senior IT Auditor to support our Internal Audit...  ...with initial guidanceLeverage AI tools effectively to support...  ...peers and market benchmarks evaluated for the scope and responsibilities...  ...Media Affirmative Action policy statement.SummaryLocation: New... 
    Policy
    Full time

    OUTFRONT Media

    New York, NY
    3 days ago
  • $99k - $232k

     ...their internal audit functions, leveraging AI and other risk technologies to address a...  ...except as set forth within the following policy: more about how we work: only those...  ...collaborating closely with team members. We evaluate these factors thoughtfully to establish a... 
    Policy
    Full time
    H1b

    PwC

    New York, NY
    5 days ago
  • $55.07k - $69.67k

     ...Design and Construction, Division of Safety and Site Support seeks a Safety Auditor. The selected candidate will be...  ..., and enforcing DDC safety policies and procedures in alignment with...  ...equivalency verification from an approved evaluation service. A list of providers (... 
    Policy
    Permanent employment
    Full time
    For contractors
    Work experience placement
    H1b
    Local area
    Visa sponsorship

    City of New York

    New York, NY
    3 days ago
  • $99k - $252.45k

     ...cybersecurity measures, data, and AI systems. Within our Assurance...  ...- Guiding teams in the evaluation of digital environments, including...  ...Certified Information Systems Auditor (CISA) certificationWhat Sets...  ...set forth within the following policy: more about how we work: only... 
    Policy
    Full time
    H1b

    PwC

    New York, NY
    1 day ago
  •  ...York Department of Design and Construction, Division of Safety and Site Support, seeks a Safety Auditor to oversee safety programs on construction sites. You...  ...safety audits, identify hazards, and enforce safety policies in alignment with OSHA, DOT, DOB, and MUTCD.... 
    Policy

    City of New York

    New York, NY
    3 days ago
  • AuraOne is seeking an Expert Project Manager for a remote review track to evaluate AI outputs across program management workflows. Reviewers assess workflow accuracy, policy adherence, and stakeholder fit while flagging operational risk and documenting the right next step... 
    Policy
    Remote job

    AuraOne

    New York, NY
    2 days ago
  • $62k - $100k

     ...continuously invest in innovative ideas, such as AI-enabled insights and technology-powered...  ...of security areas such as Auditing, Policy, Database Security, Firewall Design and...  ...committed to a merit-based hiring process, evaluating all candidates consistently using... 
    Policy
    Local area
    Worldwide

    Crowe

    New York, NY
    2 days ago
  • $250k

     ...Audit (IA) team as the Chief Auditor for Artificial Intelligence. This...  ...be at the front‑line of the AI revolution, providing critical...  ...development lifecycle. Identify, evaluate, and incorporate coverage for...  ...at Citi. View Citi’s EEO Policy Statement and the Know Your Rights... 
    Policy

    3M HEALTHCARE

    New York, NY
    3 days ago
  • $113.2k - $164.05k

     ...risk assessment, we’re advancing AI to move from insight to action...  ...and Institute of Internal Auditors standards Translate technical...  ...issue closure, including evaluating action plans, independently reviewing...  ...here to view our full EEO policy statement . Click here for more... 
    Policy
    Full time
    Work at office
    Worldwide

    Moody's

    New York, NY
    9 days ago
  • $116.72k - $175k

    DescriptionClinical Revenue Auditor-CDM Patient Financial Services-Corporate-Full-Time-Days- Hybrid.The Clinical Revenue Auditor for the...  ...privacy and security standards.Conform to the established policies/ procedures/ processes/ Standards of Behavior.Performs other duties... 
    Policy
    Full time
    Traineeship
    Local area

    Mount Sinai Health System

    New York, NY
    5 days ago
  • Mercor partners with a leading AI lab to train frontier models on insurance reasoning data. We seek Multiline P&C Underwriters to design risk scenarios, evaluate model outputs, and shape AI reasoning about insurability, pricing, and coverage terms. Candidates should have... 

    Mercor

    New York, NY
    1 day ago
  • $115k - $150k

     ...capabilities are used responsibly, ethically, and effectively.As Lead Auditor - Data & AI, you will deliver independent assurance over some of the...  ..., Asia, Europe, and the Middle East. As part of our New Frontier strategy, MetLife is building an AI-enabled, people-centered... 
    Full time
    Temporary work
    Work at office
    Local area
    Relocation package
    3 days per week

    Metropolitan Life Insurance Company

    New York, NY
    3 days ago
  • $365k

     ...reliable, interpretable, and steerable AI systems. We want AI to be safe and...  ...committed researchers, engineers, policy experts, and business leaders working...  ...to alignment, interpretability, and safety, each operating at the frontier of AI development. As a Technical Program... 
    Policy
    Work at office
    Flexible hours
    Shift work

    Anthropic

    New York, NY
    22 hours ago
  •  ...Agent Robustness to advance safe and aligned AI agents. You will contribute to evaluating risks, building testing harnesses, and prototyping...  ...techniques, and publishing results to shape policy and industry practices in frontier AI. #J-18808-Ljbffr United States Digital... 
    Policy

    United States Digital Space LLC

    New York, NY
    4 days ago
  • We're hiring AI‑Native Auditors Early‑career and already living in AI tools? We'll hand you our platform and let you out‑produce an entire...  ...with reviewers and partners through sign‑off Push the product frontier — you'll be one of our heaviest users, and your feedback... 

    Modus

    New York, NY
    3 days ago
  • $230k - $325k

     ...Data Scientist, SafetyOpenAI's Safety teams work to ensure our products are safe, trusted, and resilient as frontier AI systems scale globally. We tackle some of the company'...  ...operating at the intersection of product, safety, policy, and research.As a Data Scientist, Safety,... 
    Policy

    OpenAI

    New York, NY
    1 day ago
  • $119.8k - $234.7k

     ...MicrosoftOverviewAt Microsoft AI, we are on a mission to...  ...’s most capable AI frontier models, pushing the...  ...preparedness, multi-agent safety, traditional frontier...  ...governance, including technical evaluation and mitigation efforts,...  ...Responsible AI, Public Policy, red teaming, and... 
    Policy
    Ongoing contract
    Work at office
    Local area

    Microsoft

    New York, NY
    1 day ago
  •  ...MERIT Beauty is seeking Turkish-speaking annotators for a contract role evaluating AI-generated content. You will review content for coherence, consistency, and alignment with real-world expectations in Turkish. The ideal candidates are native or fluent Turkish speakers... 
    Contract work
    Temporary work
    Immediate start
    Remote work

    MERIT Beauty

    New York, NY
    22 hours ago
  • $405k

     ...interpretable, and steerable AI systems. We want AI to be safe...  ...committed researchers, engineers, policy experts, and business leaders...  ..., how reliably we can run safety experiments, and how effectively...  ..., reliable infrastructure and frontier capabilities can go hand in hand... 
    Policy
    Full time
    Internship
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours

    Anthropic

    New York, NY
    22 hours ago
  • We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing...  ..., uncover model weaknesses, and evaluate AI behavior across complex, high...  ...behaviours, hallucinations, and policy failures. Evaluate model... 
    Policy

    Obsidian

    New York, NY
    4 days ago
  •  ...Obsidian is hiring expert Evaluators in real estate, hospitality, and events to review AI-generated work products for accuracy and quality. This is a remote position requiring strong expertise in the relevant domains. Ideal candidates should have over 5 years of professional... 
    Work at office
    Remote work

    Obsidian

    New York, NY
    2 days ago
  •  ...BAM Ventures is seeking Swedish-speaking remote annotators to evaluate AI-generated content, ensuring that it's coherent and aligns with real-world expectations. Your role will involve reviewing outputs, identifying deviations, and providing structured feedback to enhance... 
    Remote work

    BAM VENTURES LLC

    New York, NY
    2 days ago
  • $400 per month

    About the Role Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments. The work focuses on realistic data engineering workflows... 

    Mercor Inc

    New York, NY
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Frontier AI Safety Evaluator & Policy Auditor. Be the first to apply!