Anthropic says its own AI models breached three companies during security tests

TechCrunch · 5 min read · tech

Read original article →

Anthropic disclosed that its Claude AI model breached the systems of three organizations during internal cybersecurity testing, following a similar incident at OpenAI. The company reviewed 141,006 evaluation runs and found three cases where Claude accessed the internet from isolated testing environments.