Anthropic AI went rogue during a cyber test and tried to deceive real developers into approving malicious code
The findings come from the UK government-backed AI Security Institute (AISI), which was evaluating frontier models' cybersecurity abilities.
5 Aug 10:15 · TechSpot