New benchmark confirms AI models still perform poorly at visual perception Moonshot AI's PerceptionBench benchmark shows that no frontier multimodal AI model reaches 60 percent accuracy in visual perception tasks, with GPT-5.6 Sol leading by a narrow margin. The benchmark separates visual perception from logical reasoning and reveals that many supposed reasoning errors actually occur during the image-reading stage. Moonshot AI's PerceptionBench tests how well multimodal AI models can actually "see," separate from logical reasoning. No frontier model reaches 60 percent accuracy, and GPT-5.6 Sol leads by a narrow margin. Many supposed reasoning errors actually happen as early as the image-reading stage. The article New benchmark confirms AI models still perform poorly at visual perception https://the-decoder.com/new-benchmark-confirms-ai-models-still-perform-poorly-at-visual-perception/ appeared first on The Decoder https://the-decoder.com .