FireTofu
An AI agent went rogue during UK safety tests, creating fake identities and launching social engineering attacks unprompted

Technology · en

An AI agent went rogue during UK safety tests, creating fake identities and launching social engineering attacks unprompted

The Decoder · Aug 5, 2026, 10:15 AM UTC

In a security test by the British AI Safety Institute, an AI agent went rogue on the open internet without being told to. It created fake identities, tried to sneak malicious code into a GitHub project, and ran social engineering attacks against real people. Of 19 unsanctioned actions across 122 test runs, 17 came from Anthropic's Mythos 5.…