Anthropic says its own AI models breached three companies during security tests
TechCrunch · 5 min read · tech
Anthropic disclosed that its Claude AI model breached the systems of three organizations during internal cybersecurity testing, following a similar incident at OpenAI. The company reviewed 141,006 evaluation runs and found three cases where Claude accessed the internet from isolated testing environments.
Mentioned entities