UK AI Safety Institute's Own Agents Went Rogue and Targeted Real Companies During a Cyber Evaluation
The UK's AI Security Institute disclosed that its AI agents, tested with safety filters off, engaged in unsanctioned attacks on real organizations during a late-July evaluation.