Safeguards Enforcement Analyst, Cyber HarmActive$285K

The opportunity

As an Enforcement Analyst, you will be responsible for reviewing content and executing enforcement actions across our products and services, with a focus on detecting and mitigating attempts to misuse Anthropic's AI systems for malicious cyber operations. Your initial focus will…

What you'll do

  • Review flagged content and accounts to make accurate, well-documented: enforcement decisions in line with our usage policies

  • Detect and mitigate potential misuse of AI systems to facilitate: cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations

  • Triage and escalate novel, ambiguous, or high-severity cases to appropriate stakeholders

  • Provide detailed feedback to the Safeguards policy design team on policy gaps: surfaced through real enforcement scenarios

  • Partner with Engineering and Data Science teams by surfacing detection model: errors and quality signals from review to improve precision and recall

  • Maintain high accuracy and consistency standards across review queues

What they're looking for

  • Experience in trust & safety, abuse investigations, cybersecurity: investigations, or threat intelligence in a technology or AI company
  • Experience with large language models and an understanding of how AI: technology could be misused for cyber operations
  • Experience operating within abuse monitoring programs or enforcement review systems
  • Understanding of the challenges involved in implementing product policies at: scale, including in the content moderation space
  • Experience working with government agencies, regulated environments, or information sharing communities