cd/entity/SimpleQA· home entities SimpleQA
grep -l @simpleqa /news/*.json | wc -l → 6

SimpleQA

mentions 6 type Organization feed RSS

// recent coverage 6 mentions

19:04
2026-08-16
w4g1.dev
artificial-intelligence

Models Are Getting Dumber on Purpose

Reasoning models are deliberately trading world knowledge for reasoning skill, with GLM-5.2 scoring 99.2% on AIME 2026 using about 40 billion active parameters per token, while factual recall remains …

18:30
2026-07-09
thedeepview.com
artificial-intelligence

Why Google AI Overviews expose AI's biggest problem

Google's AI Overviews, used by over two billion monthly users, are accurate about 90% of the time, according to a New York Times analysis by startup Oumi. The 10% error rate translates to tens of mill…

13:00
2026-06-18
dev.to
large-language-models

Ninety-one percent accurate is not what it sounds like

An analysis by Oumi of Google's AI Overviews found that while accuracy improved from 85% on Gemini 2 to 91% on Gemini 3 on the SimpleQA benchmark, the rate of ungrounded claims among correct answers i…

// co-occurs with top 8 entities
// topics top 6 topics