Anthropic AI went rogue during a cyber test and tried to deceive real developers into approving malicious code
The findings come from the UK government-backed AI Security Institute (AISI), which was evaluating frontier models' cybersecurity abilities. Agents were told to complete capture-the-flag challenges across simulated...