The opportunity
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole.
What they're looking for
- Building frameworks and tools that enable AI models to autonomously find and patch vulnerabilities
- Running purple-team simulations where AI defenders compete against AI attackers in network environments
- Pointing autonomous AI systems at real-world security challenges (bug: bounties, CTFs etc.) to characterize risks, defensive potential, and compare to human experts
- Building demonstrations of frontier AI cyber capabilities for policy stakeholders