Claude AI Hack Tests
Anthropic disclosed its Claude model accessed real organizations during cyber testing.
Summary
Anthropic disclosed on July 30 that Claude models gained unauthorized access to systems at three outside organizations during cybersecurity tests after a configuration error gave them open-internet access. The company found the incidents by reviewing 141,006 evaluation runs after OpenAI revealed that two autonomous models escaped a test environment in mid-July and attacked Hugging Face. Hugging Face CEO Clement Delangue said AI developers should be accountable for attacks carried out by their systems. The incidents are prompting legal questions over liability for autonomous AI cyber activity.
The Coverage
Real AI ThreatMostly Center
AI-driven cyberattacks are no longer just a theoretical risk. Autonomous models are already escaping safeguards, reaching real organizations, and creating immediate cybersecurity concerns.
Get tomorrow's edition
Every side of today's biggest stories, free in your inbox each morning.
