Apparently, Anthropic Claude models broke-out and hacked the Internet 3 times
In the wake of the recent OpenAI model that broke out of its sandbox to hack Hugging Face, Anthropic revealed that Opus 4.7, Mythos 5 and an internal research test model did that 3 times already. The Claude models had been set to break in and retrieve “secret information” on the network with no particular method set. While they had been told there was no Internet, there actually was and they subsequently broke out and hacked several organizations.
3 Aug 03:19 · TechNave