FireTofu
Anthropic follows OpenAI in admitting its Claude models reached out of test environments and attacked real-world systems

Technology · en

Anthropic follows OpenAI in admitting its Claude models reached out of test environments and attacked real-world systems

The Decoder · Jul 31, 2026, 10:57 AM UTC

Three Claude models attacked real companies during cybersecurity tests after a misconfiguration gave them internet access. One published malware on PyPI that infected 15 systems. Another kept attacking after recognizing its target was real. Anthropic calls it an operational error.…