Red Team Engineer, SafeguardsActive$320K

Remote · WorldwideTechnology

The opportunity

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole.

What they're looking for

  • Help establish metrics for measuring detection effectiveness of novel abuse
  • Experience in penetration testing, red teaming, or application security
  • Experience in model jailbreaking and testing large-scale agentic workflows: for non-obvious prompt injection vectors
  • Strong technical skills in web application security, including hands-on: expertise with security testing tools (e.g., Burp Suite, Metasploit, custom scripting frameworks)