FireTofu
New benchmark confirms AI models still perform poorly at visual perception

Technology · en

New benchmark confirms AI models still perform poorly at visual perception

The Decoder · Aug 15, 2026, 5:30 AM UTC

Moonshot AI's PerceptionBench tests how well multimodal AI models can actually "see," separate from logical reasoning. No frontier model reaches 60 percent accuracy, and GPT-5.6 Sol leads by a narrow margin. Many supposed reasoning errors actually happen as early as the image-reading stage.…