The opportunity
As a Safeguards Enforcement Analyst on the user well-being team, you'll build and execute enforcement workflows that keep our products safe, with a focus on detecting and mitigating potential harm. Your initial focus will be on how Anthropic handles age.
What you'll do
Design and architect automated enforcement systems and review workflows that: scale effectively while maintaining high accuracy
Partner with Engineering and Data Science teams to optimize detection models: for policy violations and automated enforcement systems
Review flagged content to drive enforcement and policy improvements
Enforce usage policies with a focus on detecting and mitigating potential harmful use of AI systems
Work with Legal, Public Policy, and Privacy stakeholders to keep our age: assurance approach proportionate, privacy-preserving, and responsive to an evolving regulatory landscape
Support the Safeguards policy design team by providing detailed feedback on: policy gaps based on real enforcement scenarios
What they're looking for
- Experience building or operating age-gating flows, age estimation signals,: underage account detection, or appeals workflows
- Experience advising or partnering with third-party platforms on deploying safely to younger users
- Experience working with or evaluating third-party age verification providers
- A deep interest in AI safety and responsible technology development
- Experience writing effective prompts for generative AI systems in a content review or enforcement context