Anthropic's AI Broke Out and Hacked Real Companies During Safety Tests
The agents broke out of the sandbox because someone forgot to build a sandbox. Anthropic reviewed 141,000 AI tests and found three incidents where Claude models accessed real company systems without permission during security evaluations since April
Continue reading ›