cd /news/artificial-intelligence/new-benchmark-confirms-ai-models-sti… · home topics artificial-intelligence article
[ARTICLE · art-97692] src=the-decoder.com ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

New benchmark confirms AI models still perform poorly at visual perception

Moonshot AI's PerceptionBench benchmark shows that no frontier multimodal AI model reaches 60 percent accuracy in visual perception tasks, with GPT-5.6 Sol leading by a narrow margin. The benchmark separates visual perception from logical reasoning and reveals that many supposed reasoning errors actually occur during the image-reading stage.

read1 min views1 publishedAug 15, 2026

Moonshot AI's PerceptionBench tests how well multimodal AI models can actually "see," separate from logical reasoning. No frontier model reaches 60 percent accuracy, and GPT-5.6 Sol leads by a narrow margin. Many supposed reasoning errors actually happen as early as the image-reading stage.

The article New benchmark confirms AI models still perform poorly at visual perception appeared first on The Decoder.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @moonshot ai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/new-benchmark-confir…] indexed:0 read:1min 2026-08-15 ·