cd/entity/Opus 5· home entities Opus 5
grep -l @opus 5 /news/*.json | wc -l → 135

Opus 5

mentions 135 type Person page 3/7 feed RSS

// recent coverage 135 mentions

22:14
2026-08-22
primeintellect.ai
artificial-intelligence

NanoGPT Speedrun Frontier

A benchmark of 153 autonomous runs across 18 frontier models on the nanoGPT optimizer speedrun shows Fable 5, run via claude-code at high effort, achieved the best validated result of 2,726 tokens wit…

16:12
2026-08-21
promptcube3.com
artificial-intelligence

Claude's 20-block cache lookback silently kills agent loops —

Anthropic's Claude prompt caching silently fails in agent loops when the conversation exceeds 20 content blocks between cache breakpoints, causing cache reads to drop to zero and triggering full-prefi…

09:06
2026-08-21
aikido.dev
artificial-intelligence

We burned 11.7bn tokens to find the best cyber AI model

Aikido Security burned 11.7 billion tokens to benchmark 10 AI models on rediscovering 32 fresh vulnerabilities, finding that DeepSeek V4 Pro 0813 found the most (28 of 32 across three runs) and that o…

07:09
2026-08-21
benhoyt.com
artificial-intelligence

Updating a side project with AI in 275 commits

Ben Hoyt, a manager at Canonical, spent his two-week mid-year break upgrading his side project Gifty Weddings from a wedding gift registry to a wedding website builder, using AI tools Claude Code with…

01:09
2026-08-21
zmuda.dev
large-language-models

Hotdog Bench

A user's informal test comparing ChatGPT and Claude responses to a humorous gif found ChatGPT's reply more concise, while Claude's response was criticized as verbose and flowery. The user, who calls t…

08:56
2026-08-20
hjerpbakk.com
ai-tools

Two ways to make Opus 5 concise

Claude Code version 2.1.237 introduces a built-in Concise output style that trims verbose responses, replacing the need for custom ASD-STE100-based styles. In tests, the Concise style reduced a 663-wo…

05:11
2026-08-20
webenclave.com
artificial-intelligence

Claude Revived My Microsoft Band 2

A developer revived a Microsoft Band 2 that was stuck in setup mode by using Claude Code on Opus 5 and the open-source msband-py library, which speaks the Band's USB protocol without a phone. The Band…

22:00
2026-08-19
blog.quarkslab.com
artificial-intelligence

Defeating AI-Assisted Reverse Engineering (or at Least Trying To)

Rémy Salim's team at Quarkslab spent two weeks testing whether LLM-assisted reverse engineering defeats obfuscation, handing Claude Code agents a series of hardened AArch64 binaries with one prompt: r…

17:23
2026-08-19
dev.to
artificial-intelligence

Opus 5: Review bottleneck

Anthropic's Opus 5 model improves self-verification but shifts the bottleneck to code review, as developers face larger diffs and longer review times. Telemetry from Faros AI and LinearB shows median …

21:11
2026-08-18
bloop.monster
artificial-intelligence

ASCII-art bake-off – 7 models × 3 prompts

A benchmark comparing 8 AI models on ASCII-art generation across 3 prompts found Kimi K3 fastest at 14 seconds for the first prompt, while Fable 5 took 353 seconds and cost $1.229, and Opus 5 complete…

22:07
2026-08-17
byteiota.com
artificial-intelligence

Meta Muse Code vs Claude Code: What Devs Must Know

Meta launched Muse Code on August 5, an AI coding agent built on Muse Spark 1.2 that uses parallel sub-agents in isolated git worktrees, a feature not offered by Claude Code or Codex CLI. However, the…

12:16
2026-08-16
primeintellect.ai
artificial-intelligence

Measuring Autonomous AI Research

A new public experiment by Prime Intellect ran 153 autonomous runs on the nanoGPT optimizer speedrun across 18 frontier models, finding that Claude Fable 5 and Opus 5 dramatically outperformed others,…

20:10
2026-08-15
github.com
artificial-intelligence

Ctok: Reconstructed Claude Tokenizer

Ctok, an unofficial open-source library, reconstructs Anthropic's Claude tokenizer offline, reporting exact token counts for 1,664,940 v3 and 1,722,961 v4.7 texts with zero under-counts. The library s…

00:58
2026-08-15
benchling.com
artificial-intelligence

Can LLMs work in the wet lab?

Benchling released BenchBench-Protocol, a benchmark built from thousands of real-world experiments, showing that Anthropic's Opus 5 leads at 59.2%, followed by OpenAI's GPT 5.6 at 47.1%, and open-sour…

00:10
2026-08-15
byteiota.com
ai-ethics

The AI Trust Deficit: How Every Major Lab Is Breaking It

Anthropic apologized after secretly routing paying Fable 5 customers to the cheaper Opus 4.8 model while billing them for the premium tier, a practice that sparked developer backlash and was fixed onl…

← prev page 3 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics