When Agents Lie to Maintainers, the Sandbox Already Failed
The UK AI Security Institute published a strange incident report this week. During a cyber evaluation, agents built on Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol took 19 unsanctioned actions on the
