The opportunity
As an Enforcement Analyst, you will be responsible for reviewing content and executing enforcement actions across our products and services, with a focus on detecting and mitigating attempts to misuse Anthropic's AI systems for malicious cyber operations. Your initial focus will…
What you'll do
Review flagged content and accounts to make accurate, well-documented: enforcement decisions in line with our usage policies
Detect and mitigate potential misuse of AI systems to facilitate: cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations
Triage and escalate novel, ambiguous, or high-severity cases to appropriate stakeholders
Provide detailed feedback to the Safeguards policy design team on policy gaps: surfaced through real enforcement scenarios
Partner with Engineering and Data Science teams by surfacing detection model: errors and quality signals from review to improve precision and recall
Maintain high accuracy and consistency standards across review queues
What they're looking for
- Experience in trust & safety, abuse investigations, cybersecurity: investigations, or threat intelligence in a technology or AI company
- Experience with large language models and an understanding of how AI: technology could be misused for cyber operations
- Experience operating within abuse monitoring programs or enforcement review systems
- Understanding of the challenges involved in implementing product policies at: scale, including in the content moderation space
- Experience working with government agencies, regulated environments, or information sharing communities