Anthropic Discloses Fourth Claude Hacking Incident as Debate Around Regulation Grows - Decrypt
Anthropic now says attacks during security tests exposed model behavior failures, after initially emphasizing errors in testing infra.
10 Sep 15:33 · Decrypt