Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Safeguards Enforcement Analyst, Cyber Harm

$285k - $330k
Full-time

Anthropic

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

About the role

As an Enforcement Analyst, you will be responsible for reviewing content and executing enforcement actions across our products and services, with a focus on detecting and mitigating attempts to misuse Anthropic's AI systems for malicious cyber operations. Your initial focus will center on reviewing flagged activity related to cyberattacks, malware development, and offensive exploitation; however, this position may later expand to include broader areas of enforcement.

Safety is core to our mission, and you'll help uphold policy enforcement so that our users can safely interact with and build on top of our products in a harmless, helpful, and honest way.

Important context for this role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a violent, technical, or psychologically disturbing nature. This role may require responding to escalations during weekends and holidays.

Key responsibilities

  • Review flagged content and accounts to make accurate, well-documented enforcement decisions in line with our usage policies
  • Detect and mitigate potential misuse of AI systems to facilitate cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations
  • Triage and escalate novel, ambiguous, or high-severity cases to appropriate stakeholders
  • Provide detailed feedback to the Safeguards policy design team on policy gaps surfaced through real enforcement scenarios
  • Partner with Engineering and Data Science teams by surfacing detection model errors and quality signals from review to improve precision and recall
  • Maintain high accuracy and consistency standards across review queues
  • Keep up to date with emerging AI policy enforcement best practices, threat actor tactics, and the evolving cyber threat landscape, using these to inform enforcement decisions

Minimum Qualifications

  • Experience in cybersecurity, including knowledge of offensive techniques, exploit development, malware analysis, or vulnerability research
  • Experience performing content review, abuse investigations, or policy enforcement at volume
  • Proficiency in SQL and/or Python for data analysis and threat detection
  • Experience identifying emerging risks and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams
  • Experience working with generative AI products, including writing effective prompts for content review and enforcement

Preferred qualifications

  • Experience in trust & safety, abuse investigations, cybersecurity investigations, or threat intelligence in a technology or AI company
  • Experience with large language models and an understanding of how AI technology could be misused for cyber operations
  • Experience operating within abuse monitoring programs or enforcement review systems
  • Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space
  • Experience working with government agencies, regulated environments, or information sharing communities

The annual compensation range for this role is listed below.

For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role.

Annual Salary:

$285,000—$330,000 USD

Logistics

Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience

Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience

Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position

Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.

Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.

We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.

Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings.

How we're different

We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills.

The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.

Come work with us!

Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Safeguards Enforcement Analyst, Cyber Harm in San Francisco, CA vacancy
  • $245k - $285k

    # Safeguards Enforcement Analyst, Bio HarmsSafeguards (Trust & Safety) Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington...  ...## About the roleAs an Enforcement Analyst focused on Bio Harms, you will play a critical role in protecting against the... 
    Suggested
    Remote job
    Work at office
    Visa sponsorship
    Flexible hours

    Applied Methods Ltd

    San Francisco, CA
    3 days ago
  • $245k - $285k

     ...to build beneficial AI systems. About the role As an Enforcement Analyst focused on Bio Harms, you will protect against the misuse of AI systems for...  ...potential violations, and continuously strengthening safeguards. The work sits at the intersection of biosecurity threat... 
    Suggested
    Remote work
    Visa sponsorship

    Anthropic

    San Francisco, CA
    3 days ago
  • Anthropic in San Francisco is seeking a Safeguards Enforcement Analyst to build and run enforcement workflows that keep our products safe. The role focuses on detecting and mitigating potential harm, with an initial emphasis on preventing recidivism: banned actors evading... 
    Suggested

    Anthropic

    San Francisco, CA
    4 days ago
  • $245k - $285k

    # Safeguards Enforcement Analyst, Age-Appropriate DesignSafeguards (Trust & Safety) Remote-Friendly, United States; San Francisco, CA | New York City...  ...safe, with a focus on detecting and mitigating potential harm. Your initial focus will be on how Anthropic handles age.... 
    Suggested
    Remote job
    Work at office
    Visa sponsorship
    Flexible hours

    Applied Methods Ltd

    San Francisco, CA
    3 days ago
  • $245k - $285k

    # Safeguards Enforcement Analyst, Child SafetySafeguards (Trust & Safety) Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington...  ...'s mission. Child safety is one of our highest-priority harm areas, and the person in this role will have a direct and... 
    Suggested
    Remote job
    Work at office
    Visa sponsorship
    Flexible hours

    Applied Methods Ltd

    San Francisco, CA
    3 days ago
  • Anthropic is hiring an Enforcement Analyst focused on Bio Harms in the United States. You will protect against misuse of AI systems, enforce Usage Policies, and strengthen safeguards. You will read model interactions and decide if activity is benign research or harmful... 

    Anthropic

    San Francisco, CA
    3 days ago
  • Doist is seeking a Safeguards Enforcement Analyst on the account abuse team to build and execute enforcement workflows that keep our products safe, focusing on detecting and mitigating potential harm. You will drive enforcement areas including access controls and identity... 

    Doist

    San Francisco, CA
    1 day ago
  • $285k - $330k

     ...together to build beneficial AI systems. About The Role As a Safeguards Enforcement Analyst on the account abuse team, you’ll build and execute...  ...safe, with a focus on detecting and mitigating potential harm. Your focus will be driving a number of enforcement areas... 
    For contractors
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    2 days ago
  • $245k - $285k

     ...our users and for society as a whole. About the role As a Safeguards Enforcement Analyst on the account abuse team, you will build and execute enforcement...  ...safe, focusing on detecting and mitigating potential harm. Your initial focus will be on account compromise,... 
    Visa sponsorship

    Doist

    San Francisco, CA
    1 day ago
  • $245k - $285k

     ...beneficial AI systems. About the role As a Safeguards Analyst on the User Well-being team, you will...  ...covers a broad set of interconnected harms including suicide, self-harm,...  ...into broader areas of user well-being enforcement over time. Safety is core to our mission... 
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    2 days ago
  • $245k - $285k

    About the Role As a Safeguards Enforcement Analyst on the account abuse team, you will build and execute enforcement workflows that keep our products safe, focusing on detecting and mitigating potential harm. Your initial focus will be on recidivism: ensuring that banned... 
    For contractors
    Visa sponsorship

    Anthropic

    San Francisco, CA
    4 days ago
  • $245k - $285k

     ...together to build beneficial AI systems. About the Role As a Safeguards Enforcement Analyst on the account abuse team, you’ll build and execute...  ...safe, with a focus on detecting and mitigating potential harm. Your initial focus will be standing up fraud & scams enforcement... 
    For contractors
    Visa sponsorship

    Anthropic

    San Francisco, CA
    4 days ago
  • Anthropic is seeking an Enforcement Analyst focused on Chem & Explosives Harms to protect against AI misuse and strengthen safeguards. You will enforce Usage Policy, investigate potential violations, and improve monitoring workflows across Policy, Threat Intelligence, Data... 
    Remote job

    Neura Market

    San Francisco, CA
    4 days ago
  • $285k - $330k

    Safeguards Enforcement Analyst, Integrity & Authenticity Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC About...  .... Your work will span a broad and interconnected set of harm areas: AI‑enabled influence operations and disinformation... 
    Remote work
    Visa sponsorship
    Weekend work

    Anthropic

    San Francisco, CA
    13 hours ago
  • Anthropic is seeking a Safeguards Enforcement Analyst to build and run enforcement workflows that keep our products safe, focusing on detecting and mitigating potential harm. You will own detection, revocation, user notification, remediation, and restoration criteria for... 

    Doist

    San Francisco, CA
    1 day ago
  • Anthropic seeks an Enforcement Analyst focused on Radiological and Nuclear Harms. You will enforce Usage Policy, read real model interactions, and decide if activity...  ..., Data Science, and Engineering to strengthen safeguards. The role requires experience in Trust & Safety,... 

    Applied Methods Ltd

    San Francisco, CA
    3 days ago
  • $230k - $270k

     ...Safeguards Enforcement Analyst, Safety Evaluations Remote-Friendly (Travel-Required) | San Francisco, CA | Washington, DC; San Francisco, CA | New York City, NY About Anthropic Anthropic's mission is to create reliable, interpretable, and steerable AI systems. We want... 
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours
    Shift work

    Anthropic

    San Francisco, CA
    4 days ago
  • Anthropic is seeking a Safeguards Enforcement Analyst to design scalable enforcement workflows and partner with engineering and data science teams to improve policy enforcement against AI-enabled influence operations and disinformation. The role involves reviewing content... 
    Remote work

    Anthropic

    San Francisco, CA
    13 hours ago
  • Anthropic is seeking a Safeguards Enforcement Analyst in San Francisco to ensure our models meet safety standards. This role is remote-friendly but requires occasional travel. You will collaborate with various teams to manage evaluations and drive improvements. The ideal... 
    Remote work

    Anthropic

    San Francisco, CA
    1 day ago
  • $245k - $285k

     ...build beneficial AI systems. About the role As a Safeguards Enforcement Analyst on the account abuse team, you'll build and execute enforcement...  ...safe, with a focus on detecting and mitigating potential harm. Your initial focus will be recidivism: a ban that an... 
    Full time
    For contractors
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    2 days ago
  • $285k - $330k

     ...build beneficial AI systems. About the role As a Safeguards Enforcement Analyst focused on Violence & Extremism, you will be responsible for...  ...to misuse Anthropic's AI systems to facilitate real-world harm, including weapons and dangerous technology, critical... 
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    2 days ago
  • $190k - $285k

     ...About the role We're looking for an analyst to support the team's cyber product policy work: the usage policy language, help-center and enforcement guidance, and launch policy notes that...  ...that enforcement decisions and safeguards match what they say; maintain a running... 
    Cyber
    Work at office
    Visa sponsorship
    Flexible hours
    Shift work

    Anthropic

    San Francisco, CA
    2 days ago
  • Anthropic is seeking a Safeguards Enforcement Analyst focused on Bio Harms to protect against misuse of AI systems. You will enforce Usage Policies, monitor enforcement workflows, and investigate potential violations with emphasis on biosecurity and dual-use concerns. The... 
    Remote job

    Applied Methods Ltd

    San Francisco, CA
    3 days ago
  • Anthropic is seeking a Safeguards Enforcement Analyst on the account abuse team to build enforcement workflows that keep our products safe, with a focus on detecting and mitigating potential harm. You’ll own the policy layer: what we ask users for, when, on what grounds... 

    Anthropic

    San Francisco, CA
    2 days ago
  • Doist is seeking a Safeguards Enforcement Analyst on the account abuse team to build enforcement workflows that detect and mitigate harm, with an initial focus on recidivism and preventing re-registration by banned actors. You will own detection of returning actors, link... 

    Doist

    San Francisco, CA
    1 day ago
  • $189k - $210k

     ...Safety & Risk Operations teams safeguard our products, users, and the...  ...Safety, and Risk Operations analysts who have subject matter expertise...  ...the following areas: policy enforcement and content moderation, fraud...  ...-priority cases across all harm and risk areas, ensuring... 
    Work at office
    Relocation package
    Shift work
    Night shift
    Rotating shift
    Weekend work

    OpenAI

    San Francisco, CA
    2 days ago
  • $252k - $280k

    ABOUT THE TEAM Critical Harm Operations sits within User Safety & Risk Operations and builds enforcement systems for Frontier Risk and Material Harm that are accurate, fast, defensible, and built to scale. The Cyber vertical turns policy into reviewer standards, calibrated... 
    Cyber

    OpenAI

    San Francisco, CA
    4 days ago
  •  ...sectors. About the Role As an Agentic Risk Analyst, you will shape OpenAI’s operating...  ...abuse patterns, and potential downstream harms across products, deployment environments,...  ...in trust and safety, integrity, security, cyber threat intelligence, AI safety, product risk... 
    Cyber
    Shift work

    Neura Market

    San Francisco, CA
    13 hours ago
  • $295k

     ...developed and deployed. We build evaluations, safeguards, and safety frameworks that help our...  ...-end mitigation stack to reduce severe cyber misuse across OpenAI's products. This...  ...collaboration to ensure safeguards are enforceable, scalable, and effective. You'll contribute... 
    Cyber

    Slope

    San Francisco, CA
    2 days ago
  • OpenAI is seeking a Cyber Operations Strategist to lead Critical Harm Operations from our San Francisco site. You will shape the operating model, design escalation paths, and drive durable improvements across reviews, vendors, and automation. Expect a senior IC role focused... 
    Cyber

    Triwill Group

    San Francisco, CA
    13 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Safeguards Enforcement Analyst, Cyber Harm. Be the first to apply!