cd/entity/ARC Prize· home entities ARC Prize
grep -l @arc prize /news/*.json | wc -l → 18

ARC Prize

mentions 18 type Person feed RSS

// recent coverage 18 mentions

10:22
2026-09-08
techstrong.ai
artificial-intelligence

Is It AGI, or Is It Memorex?

OpenAI's Astra has intensified the debate over whether AGI has arrived, with OpenAI president Greg Brockman personally believing AGI has been achieved, but benchmark results vary dramatically by confi…

23:57
2026-09-06
kobaran.com
artificial-intelligence

GPT-6 Astra’s Wild Score Swing Reframes the Race Toward AGI

OpenAI's GPT-6 Astra scored 62.7% on the ARC-AGI-3 benchmark under the Standard harness and 99.9% under the Provider Adapter harness, with identical weights and problems, according to ARC Prize. The P…

19:53
2026-09-03
twitter.com
artificial-intelligence

GPT-6 Astra Achieves SOTA on ARC-AGI

OpenAI's GPT-6 Astra achieved state-of-the-art results on the ARC-AGI benchmark, scoring 63% on ARC-AGI-3 and 99% via a new provider adapter harness, surpassing human performance on 96% of ARC-AGI-3 l…

00:00
2026-09-03
arcprize.org
artificial-intelligence

OpenAI's GPT-6 Astra on ARC-AGI-3

OpenAI's GPT-6 Astra scored 62.7% for $26K on ARC-AGI-3 Semi-Private with the Standard harness and 99.9% for $19K with the Provider Adapter harness, surpassing the human baseline in action efficiency …

16:34
2026-08-24
promptcube3.com
artificial-intelligence

Stop trusting raw benchmark scores without looking at the harness

ARC-AGI-3 benchmark scores vary by up to 70 points depending on the testing harness, with NVIDIA reporting Claude Opus 5 at 100.00% in late August versus the official ARC Prize verified score of 30.16…

00:54
2026-08-22
startupfortune.com
artificial-intelligence

Nvidia Shows a Better AI Harness Beat a Smarter Model on ARC-AGI-3

Nvidia reported a 100.00 RHAE score on the ARC-AGI-3 public set using its Agentic Variation Operators (AVO) system, up from Claude Opus 5's 30.16% published score, by wrapping the same model family in…

00:00
2026-07-26
runagentrun.co.uk
artificial-intelligence

Opus 5 nearly quadruples the ARC-AGI-3 record

Anthropic's Claude Opus 5 scored 30.2% on the ARC-AGI-3 benchmark, nearly quadrupling the previous record of 7.8% set by OpenAI's GPT-5.6 Sol, according to the ARC Prize team. The result marks the mos…

06:31
2026-07-25
arcprize.org
artificial-intelligence

ARC-AGI Leaderboard

The ARC-AGI-3 leaderboard, released by the ARC Prize team, ranks AI systems on their ability to adapt to novel interactive environments, measuring performance against cost-per-task. The leaderboard sh…

// co-occurs with top 8 entities
// topics top 6 topics