Threat Intel Manager, Model Exploitation & FraudActive$375K

Hybrid · San Francisco, CATechnology

The opportunity

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole.

What they're looking for

  • Expand the team's coverage into fraud and scams, building the detection and: investigation playbooks from the ground up
  • Own the external engagement program for the area, including regular: intelligence sharing with U.S. government partners and industry peers, ensuring the investigators driving the work are visible in those channels
  • Anticipate how resellers, proxies, and third-party platforms change the abuse: surface, and shape coverage accordingly
  • Work with policy, enforcement, and engineering to convert findings into bans,: product mitigations, and safety-by-design improvements