cd/entity/ARC-AGI· home entities ARC-AGI
grep -l @arc-agi /news/*.json | wc -l → 12

ARC-AGI

mentions 12 type Organization feed RSS

// recent coverage 12 mentions

02:44
2026-08-12
promptcube3.com
artificial-intelligence

Pathway's 150M model just hit 29.

Pathway's 150M parameter model achieved a score of 29 on the ARC-AGI benchmark, demonstrating that architectural innovations like recurrent memory and latent reasoning can rival much larger models. Th…

00:00
2026-08-01
mindstudio.ai
artificial-intelligence

DeepSeek V4 Flash: The Cheapest Frontier-Level Open Model Yet

DeepSeek V4 Flash, a post-trained update to DeepSeek's existing V4 Flash preview model with 284 billion parameters, jumped from 7% to 54% on the DeepSweep agentic coding benchmark, rivaling larger mod…

18:20
2026-07-25
promptcube3.com
artificial-intelligence

Claude 3 Opus and the ARC-AGI Benchmark

An analysis of Claude 3 Opus suggests the model may be 'benchmaxxing'—optimizing for the ARC-AGI benchmark rather than demonstrating genuine reasoning, according to a post on the site. The ARC benchma…

00:34
2026-07-23
poetiq.ai
artificial-intelligence

Benchmarks Are Dead (For Us)

Poetiq claims its Recursive Self-Improvement (RSI) Metasystem has autonomously set state-of-the-art results on six diverse benchmarks, including outperforming Muse Spark 1.1 within 48 hours of publica…

03:43
2026-07-22
pub.towardsai.net
artificial-intelligence

The Next AI Breakthrough May Not Be a Bigger Model

OpenAI's o3 model scored 87.5% on the ARC-AGI benchmark in December 2024, up from GPT-4o's 5%, by using 5.5 billion tokens of inference-time compute instead of scaling model parameters. The benchmark,…

18:17
2026-07-21
khola.blog
artificial-intelligence

The Top-Down Bet Needs A Bottom-Up Audit

A mid-2026 audit of top-down AI-assisted software engineering shows agents absorbing implementation work on schedule, with SWE-bench Verified scores rising from 1.96% in October 2023 to near saturatio…

04:00
2026-07-21
arxiv.org
artificial-intelligence

Quantizing Recursive Reasoning Models

A new study shows that 4-bit quantization causes catastrophic accuracy collapse in recursive reasoning models, dropping Sudoku exact-solution accuracy from 84.1% to 0.0%, but per-block scaling using M…

14:05
2026-07-17
ycrootaccess.com
artificial-intelligence

World Models: An intuitive introduction

Y Combinator General Partner Ankit Gupta and Visiting Partner Francois Chaubard discussed world models as a promising solution to sample efficiency, one of AI's biggest open problems, in a Decoded epi…

00:00
2026-06-30
evalevalai.com
artificial-intelligence

When AI Benchmarks Stop Measuring Progress

Nearly half of 60 popular AI benchmarks studied in a new ICML paper show high levels of saturation, meaning they can no longer reliably distinguish between leading models. The researchers from the pap…

21:54
2026-06-15
whenwill.ai
artificial-intelligence

Show HN: When Will AI? – A timeline of top AI predictions

A new website, 'When Will AI?', compiles a timeline of top AI predictions from labs, reports, markets, and experts, sorted by significance, covering years from 2026 onward. The predictions include mil…

22:33
2026-05-01
thealgorithmicbridge.com
artificial-intelligence

Weekly Top Picks #120

The AI industry has spent $725 billion in a high-stakes gamble on returns, while the U.S. government moves to nationalize AI models as strategic resources. A Chinese court ruled companies cannot offlo…

// co-occurs with top 8 entities
// topics top 6 topics