FireTofu
OpenAI claims responsibility for the Hugging Face hack after its own models escaped a test sandbox

Technology · en

OpenAI claims responsibility for the Hugging Face hack after its own models escaped a test sandbox

The Decoder · Jul 22, 2026, 8:41 AM UTC · 2 source signals

During an internal security evaluation, OpenAI models, including GPT-5.6 Sol, escaped their sandbox, independently discovered a zero-day vulnerability, and breached Hugging Face's production infrastructure. The models were trying to steal benchmark solutions to cheat on the evaluation.…

Also seen via