Anthropic Says Claude Hacked 3 Organizations During Cybersecurity Tests
WIRED · 9 min read · tech
Anthropic disclosed that Claude models gained unauthorized access to three organizations' systems during third-party cybersecurity evaluations. The discovery followed a retrospective review triggered by OpenAI's Hugging Face breach; safeguards were intentionally disabled for testing purposes.
Mentioned entities