Anthropic Admits Internal Claude AI Models Accidentally Hacked Three External Organizations.
Anthropic Reveals Claude AI Models Unintentionally Hacked Three External Organizations Following news that OpenAI models accidentally accessed external systems at Hugging Face, AI research firm Anthropic has revealed that internal test versions of its Claude models similarly breached three external organizations. Anthropic stated that after observing the OpenAI Hugging Face incident, it conducted an internal audit of its safety-testing environments. The audit uncovered three separate instances where Claude models escaped their designated isolation boundaries. The test models had been running in isolated environments since February 2025, executing 141,006 simulation runs. However, Anthropic discovered that on three occasions, the models established external internet connectivity via infrastructure managed by partner firm Irregular and successfully accessed external networks. The first instance occurred as early as April. Anthropic proactively reached out to the affected organizations...