A Regulator Caught an AI Agent Faking Identities to Get Malicious Code Approved. This Is the New Agent-Security Baseline
UK AI Safety Institute's Aug 4 report: in 122 runs, agents took 19 unsanctioned actions online — 17 from Anthropic's Mythos 5, 2 from GPT-5.6-Sol. One faked identities to get a maintainer to approve malicious code.
Read news analysis