cd/entity/ARC-AGI-1· home entities ARC-AGI-1
grep -l @arc-agi-1 /news/*.json | wc -l → 15

ARC-AGI-1

mentions 15 type Organization feed RSS

// recent coverage 15 mentions

04:50
2026-09-11
arxiv.org
machine-learning

Thinking with Looped Flows

Researchers Ayhan Suleymanzade and co-authors submitted a paper to arXiv on 10 September 2026 proposing "looped flows," a training method that uses local denoising objectives with progressively decrea…

16:14
2026-09-02
promptcube3.com
artificial-intelligence

7M parameter models are outperforming GPT on ARC and it's kind

Samsung's Tiny Recursive Model (TRM), with only 7 million parameters and two layers, outperforms frontier models like GPT-4 on the ARC-AGI-2 benchmark, suggesting that architecture and recursive logic…

21:05
2026-09-01
arcprize.org
artificial-intelligence

Claude Fable 5.1 results on ARC-AGI

Anthropic's Claude Fable 5.1 scores 97.5% on ARC-AGI-1 Semi-Private at $1.40 per task and 90.0% on ARC-AGI-2 Semi-Private at $4.49 per task at max effort, according to results published by Anthropic o…

09:52
2026-09-01
mvakde.github.io
artificial-intelligence

44% on ARC-AGI-1 in 67 cents

A developer trained a small transformer from scratch in 1.5 hours on an NVIDIA RTX 5090 for 67 cents of compute, scoring 44% on the ARC-AGI-1 benchmark and 7% on ARC-2, matching the performance of TRM…

21:12
2026-08-25
codesota.com
ai-research

Open SOTA Registry

CodeSOTA, an open data terminal for reinforcement learning environments and state-of-the-art models, has launched a registry tracking 9,102 results, 163 models, 371 datasets, and 9 capability areas, w…

08:59
2026-08-22
arxiv.org
artificial-intelligence

BDH-CQ: In-Context Learning with Recurrent Latent Reasoning

Researchers introduced BDH-CQ, a reasoning model combining in-context learning with recurrent latent reasoning, achieving 29.5% pass@2 on the ARC-AGI-1 evaluation set with a 150M-parameter configurati…

19:27
2026-08-20
arcprize.org
artificial-intelligence

Gemini 3.7 Flash scores on ARC-AGI

Google's Gemini 3.7 Flash scored 95.5% on ARC-AGI-1 Semi-Private at $0.12 per task and 84.6% on ARC-AGI-2 Semi-Private at $0.25 per task at high effort, according to results published by the ARC Prize…

12:31
2026-08-15
promptcube3.com
artificial-intelligence

BDH-CQ hits 29.5% on ARC-AGI-1 with only 150M parameters

BDH-CQ, a 150M-parameter model, achieved a 29.5% pass@2 score on the ARC-AGI-1 benchmark at a cost of $0.00070 per task, demonstrating that recurrent latent state reasoning can outperform larger model…

04:33
2026-08-15
promptcube3.com
artificial-intelligence

A 150M parameter model hitting 29.5% on ARC-AGI-1 is insane

A 150-million-parameter recurrent latent reasoning model has achieved a 29.5% score on the ARC-AGI-1 benchmark, a result that challenges the cost-to-accuracy frontier typically associated with much la…

11:30
2026-08-11
thedeepview.com
artificial-intelligence

Pathway breakthrough challenges AI economics

Pathway, a startup founded by Zuzanna Stamirowska, unveiled a 150-million parameter reasoning model, BDH-CQ, claiming it achieves comparable performance to leading frontier models at a fraction of the…

// co-occurs with top 8 entities
// topics top 6 topics