cd /news/artificial-intelligence/openai-claims-gpt-5-6-sol-beats-opus… · home topics artificial-intelligence article
[ARTICLE · art-79866] src=the-decoder.com ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 but only with its own custom test harness

OpenAI claims its GPT-5.6 Sol model scored 38.3 percent on the ARC-AGI-3 benchmark, surpassing Anthropic's Opus 5 at 30.2 percent, but only when using OpenAI's own API with retained reasoning and context compaction. In the official test environment, GPT-5.6 Sol managed just 7.8 percent, while Opus 5 achieved its score without such aids.

read1 min views1 publishedJul 30, 2026

OpenAI counters Anthropic's ARC-AGI-3 record: GPT-5.6 Sol scores 38.3 percent, but only through its own API with retained reasoning and context compaction. In the official test environment, the model managed just 7.8 percent. Opus 5 hit its 30.2 percent without such aids.

The article OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 but only with its own custom test harness appeared first on The Decoder.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/openai-claims-gpt-5-…] indexed:0 read:1min 2026-07-30 ·