cd/entity/Opus 5· home entities Opus 5
grep -l @opus 5 /news/*.json | wc -l → 135

Opus 5

mentions 135 type Person page 2/7 feed RSS

// recent coverage 135 mentions

16:05
2026-08-27
lockstep.greg.technology
artificial-intelligence

Lockstep – Can a language model run a logic circuit in its head?

Opus 5 topped a new benchmark called Lockstep, scoring highest among 10 leading LLMs tested on their ability to run logic circuits in their heads, with GPT 5.5 second and Kimi K3 and DeepSeek V4 Pro t…

10:42
2026-08-27
twitter.com
ai-safety

Fable nuked my dev machine

Sebastien Guillemot reported on X that Fable, an AI agent, deleted his entire dev machine after Claude, the underlying model, ran `rm -rf` on his home directory while testing a sandbox it was building…

16:32
2026-08-26
buntinglabs.com
artificial-intelligence

SurveyorBench

SurveyorBench, a new benchmark from an unnamed AI mapping company, shows that leading LLMs fail at multimodal CAD tasks for civil engineering, with the strongest models scoring between 0% and 4% on th…

17:42
2026-08-25
kevinmahoney.co.uk
artificial-intelligence

AI Review Loops Don't Always Stabilise

A developer's test of AI review loops found that defect counts increase with each review-fix iteration, contrary to the assumption that such loops stabilize code. The test, conducted by KMahoney using…

09:35
2026-08-25
pran.sh
artificial-intelligence

Highest AI benchmarks ≠ adoption stonks

Anthropic's most expensive model, Fable 5, has seen spending plateau at about 11% of total spend on the company's tools more than two months after launch, while its smaller Opus 5 has already surpasse…

00:46
2026-08-25
cio.com
artificial-intelligence

Companies not as keen on Anthropic’s best AI model

Payment data from Ramp shows that Anthropic's most advanced AI model, Fable 5, accounts for only 11% of total enterprise spending on Anthropic models two months after launch, breaking the previous tre…

00:14
2026-08-25
github.com
ai-tools

Poka-Yoke: Mistake-Proofing Claude Code Skill for Software

Poka-yoke, a mistake-proofing skill set for AI coding assistants, improves the rate at which models identify design constraints from 42% to 81%, according to benchmarks from developer rainmanjam. Acro…

23:07
2026-08-24
thedeepview.com
artificial-intelligence

Why OpenAI is resetting frontier AI prices

OpenAI cut the price of its top-tier model GPT-5.6 Sol by more than 20% for the next three months, lowering input costs to $4 per million tokens and output to $20 per million tokens, undercutting Anth…

21:54
2026-08-24
tokenstead.ai
artificial-intelligence

Grok 4.6

XAI released Grok 4.6, a proprietary frontier model built with Cursor, featuring 500K context, a knowledge cutoff of Feb 2026, and text and image input with text output. Priced at $2 per 1M input toke…

21:54
2026-08-24
tokenstead.ai
artificial-intelligence

ox-alpha

An anonymous reasoning model called 'stealth/ox-alpha' has been available free on OpenRouter since Aug 20, 2026, with undisclosed operators but community fingerprinting suggesting it is an unreleased …

13:04
2026-08-24
machinebrief.com
artificial-intelligence

Anthropic's Flagship Models Went Dark for Three Hours - and

Anthropic's flagship models Fable 5, Mythos 5, Opus 5, and Opus 4.8 experienced elevated error rates for roughly three hours starting around midnight ET on August 24, taking down Claude.ai, the API, C…

07:44
2026-08-24
snipvote.com
artificial-intelligence

Anthropic Opus 5 adoption lags behind cheaper models in July 2026

Anthropic's Opus 5 model accounted for only $1 million in corporate spend in July 2026, according to billing data from 70,000 Ramp customers, as cheaper models gained favor. The data highlights a tren…

06:43
2026-08-24
ainexusdaily.vercel.app
artificial-intelligence

What Changed in AI in the Last 90 Days (Quick Round-up)

OpenAI released GPT-5.6 in three tiers (Sol, Terra, Luna) after a government review, with the fastest tier reportedly hitting 750 tokens/sec on Cerebras hardware and a new 'Ultra' mode for maximum rea…

00:00
2026-08-24
olafalders.com
ai-tools

Pinning Claude Code

Developer Olaf Alders pinned his Claude Code to Opus 4.8, citing that Opus 5 feels worse in some ways, and shared a script that forces environment variables via Claude's config to disable auto-updates…

18:35
2026-08-23
dbreunig.com
artificial-intelligence

Fable & The End of the Free Lunch

Anthropic's Fable model, released the same week as GLM 5.2, costs roughly nine times more than GLM 5.2, prompting agentic coders to adopt cheaper alternatives for routine coding tasks. The high price …

13:04
2026-08-23
machinebrief.com
artificial-intelligence

Anthropic's Protein Models Just Got Verified in a Wet Lab -

Anthropic published wet-lab-validated protein design results showing Claude-designed binders succeeded against 14 of 15 tested targets, with a 22 to 35 percent success rate versus an industry baseline…

23:13
2026-08-22
byteiota.com
ai-agents

Claude Managed Agents: Budgets, Advisors, Geo, Skills

Anthropic released four new controls for Claude Managed Agents on August 7, including session budgets, advisor models, GitHub-hosted skills, and geo controls, to address governance issues that Gartner…

← prev page 2 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics