Safeguards Enforcement Analyst, User Well-beingPosted today$245K

Remote · WorldwideTechnology

The opportunity

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole.

What they're looking for

  • Keep up to date with emerging AI policy and external research on AI's: relationship to mental health, and use these to inform our decision-making and workflows
  • Experience in trust & safety, product policy, content moderation, or a: related field, with direct exposure to mental health, suicide and self-harm, or related well-being harm areas
  • Experience designing or running experiments, evaluations, or measurement: studies to determine whether an intervention worked
  • Experience translating policy definitions into measurable form: the rubrics, review guidelines, or classification criteria, whether applied by human reviewers or automated systems