cd/entity/ARC-AGI-2· home entities ARC-AGI-2
grep -l @arc-agi-2 /news/*.json | wc -l → 14

ARC-AGI-2

mentions 14 type Organization feed RSS

// recent coverage 14 mentions

19:53
2026-09-03
twitter.com
artificial-intelligence

GPT-6 Astra Achieves SOTA on ARC-AGI

OpenAI's GPT-6 Astra achieved state-of-the-art results on the ARC-AGI benchmark, scoring 63% on ARC-AGI-3 and 99% via a new provider adapter harness, surpassing human performance on 96% of ARC-AGI-3 l…

16:14
2026-09-02
promptcube3.com
artificial-intelligence

7M parameter models are outperforming GPT on ARC and it's kind

Samsung's Tiny Recursive Model (TRM), with only 7 million parameters and two layers, outperforms frontier models like GPT-4 on the ARC-AGI-2 benchmark, suggesting that architecture and recursive logic…

13:43
2026-09-02
cryptobriefing.com
artificial-intelligence

Fable 5.1 achieves 32% lower cost per task on ARC-AGI benchmarks

Anthropic released Claude Fable 5.1 on September 1, achieving 90% coverage on the ARC-AGI-2 benchmark at 32% lower cost per task than its predecessor, Fable 5, which cost $5.45 per task. The model als…

21:05
2026-09-01
arcprize.org
artificial-intelligence

Claude Fable 5.1 results on ARC-AGI

Anthropic's Claude Fable 5.1 scores 97.5% on ARC-AGI-1 Semi-Private at $1.40 per task and 90.0% on ARC-AGI-2 Semi-Private at $4.49 per task at max effort, according to results published by Anthropic o…

06:07
2026-08-27
arxiv.org
artificial-intelligence

Metaⁿ: Recursive Self-Improvement Through Emergent Depth

Researchers introduced Meta^n, a recursive self-improvement system that keeps its meta-operation fixed and recurses on its input, achieving superior performance over prior self-improving agents on all…

19:27
2026-08-20
arcprize.org
artificial-intelligence

Gemini 3.7 Flash scores on ARC-AGI

Google's Gemini 3.7 Flash scored 95.5% on ARC-AGI-1 Semi-Private at $0.12 per task and 84.6% on ARC-AGI-2 Semi-Private at $0.25 per task at high effort, according to results published by the ARC Prize…

18:26
2026-08-13
machinebrief.com
artificial-intelligence

It’s On: The 2026 ARC-AGI Prize Is Part Of Vanguard AI Research

The 2026 ARC-AGI Prize, now part of Vanguard AI Research, will award $1 million to the first team that achieves an 85% accuracy score on the ARC-AGI-2 benchmark, a test designed to measure fluid intel…

04:00
2026-07-23
machinebrief.com
artificial-intelligence

PoTRE: Test-Time Reasoning inspired by Cognitive Heterogeneity

Researchers introduced PoTRE (Poly-Topological Reasoning Ensembles), a heterogeneous framework that decouples inference into four agents to improve complex reasoning in large language models. PoTRE ac…

09:17
2026-07-15
verantyx.ai
artificial-intelligence

Language vs. Vectors

A project called ARC-AGI-2 has achieved a 20.7% score using zero neural networks, focusing on LLM-free symbolic reasoning. The initiative also includes innovative iOS games controlled by mouth movemen…

22:35
2026-06-04
eitanturok.github.io
artificial-intelligence

A one-parameter model that gets 100% on ARC-AGI-2

A researcher built a model with a single parameter that achieved 100% accuracy on the ARC-AGI-2 benchmark, a million-dollar challenge designed to test reasoning models. The model used chaos theory and…

// co-occurs with top 8 entities
// topics top 6 topics