cd/entity/Sonnet 5· home› entities› Sonnet 5
grep -l @sonnet 5 /news/*.json | wc -l → 105

Sonnet 5

mentions 105 type Person page 2/6 feed RSS

// recent coverage 105 mentions

00:00
2026-09-01
vercel.com
artificial-intelligence

Claude Fable 5.1 now available on AI Gateway

Anthropic's Claude Fable 5.1 is now available on Vercel's AI Gateway, featuring improvements for long, multi-stage tasks like agentic coding and research, with cybersecurity and biology safety classif…

13:50
2026-08-29
motherduck.com
artificial-intelligence

Agentic SQL for Free: Qwen3.8 27B and DuckDB

Qwen 3.8 27B, running locally on a 16GB RAM MacBook Pro, outperformed OpenAI's GPT 5.6 Luna Max on the DABstep benchmark at over 17 times lower cost, with electricity costs under $0.50 versus over $8.…

15:30
2026-08-28
securitylabs.datadoghq.com
artificial-intelligence

Putting models to the secure coding test: Plan vs. default mode

Datadog's new research series testing coding agents' secure code generation found that running models in plan mode versus default mode had no significant security impact across Sonnet 5, Composer 2.5,…

00:00
2026-08-28
motherduck.com
artificial-intelligence

Agentic SQL for Free with Qwen3.8 27B and DuckDB

Qwen 3.8 27B, an open-weight model, outperformed OpenAI's GPT 5.6 Luna Max on the DABstep benchmark while running locally on a laptop, costing under $0.50 in electricity versus over $8 for Luna Max, a…

01:57
2026-08-26
dev.to
artificial-intelligence

The AI Exam Author Was Never Wrong. I Still Can't Use Its Exam.

An engineer who previously created a 29-question exam for an LLM and made five errors in the process tested whether an AI could write a better exam. The AI author (Sonnet 5) generated 50 questions wit…

00:14
2026-08-25
github.com
ai-tools

Poka-Yoke: Mistake-Proofing Claude Code Skill for Software

Poka-yoke, a mistake-proofing skill set for AI coding assistants, improves the rate at which models identify design constraints from 42% to 81%, according to benchmarks from developer rainmanjam. Acro…

00:00
2026-08-24
rubyonrails.org
ai-research

Agents on Rails: lemans goes open source

Rails has open-sourced lemans, the Ruby-based harness behind its Agents on Rails benchmark, and released new scores for four models, including Sonnet 5, Terra, and Qwen 3.8-27B. Qwen 3.8-27B scored 48…

23:13
2026-08-22
byteiota.com
ai-agents

Claude Managed Agents: Budgets, Advisors, Geo, Skills

Anthropic released four new controls for Claude Managed Agents on August 7, including session budgets, advisor models, GitHub-hosted skills, and geo controls, to address governance issues that Gartner…

16:12
2026-08-21
promptcube3.com
artificial-intelligence

Claude's 20-block cache lookback silently kills agent loops —

Anthropic's Claude prompt caching silently fails in agent loops when the conversation exceeds 20 content blocks between cache breakpoints, causing cache reads to drop to zero and triggering full-prefi…

01:09
2026-08-21
zmuda.dev
large-language-models

Hotdog Bench

A user's informal test comparing ChatGPT and Claude responses to a humorous gif found ChatGPT's reply more concise, while Claude's response was criticized as verbose and flowery. The user, who calls t…

21:11
2026-08-18
bloop.monster
artificial-intelligence

ASCII-art bake-off – 7 models × 3 prompts

A benchmark comparing 8 AI models on ASCII-art generation across 3 prompts found Kimi K3 fastest at 14 seconds for the first prompt, while Fable 5 took 353 seconds and cost $1.229, and Opus 5 complete…

20:46
2026-08-18
stuckinalocalminima.com
artificial-intelligence

Caveman Saves Tokens by Doing Less, Not Just Saying Less

Caveman, an open-source skill for coding agents, reduces output tokens by 65% on standalone answers but only cuts token usage by 18% in Claude Code and 3.9% in Codex CLI across 60 tasks, while also re…

19:28
2026-08-18
momo5502.com
artificial-intelligence

200B Tokens Later: A Month of Letting AI Agents Decompile MW2

A four-week effort using Anthropic's Claude Code CLI with Sonnet 5 agents has decompiled 34% of Call of Duty: Modern Warfare 2 (2009), producing nearly 7,000 commits and consuming 199.8 billion tokens…

00:00
2026-08-16
digitalapplied.com
developer-tools

Building on a Coding Agent That Ships Every Day

Between August 12 and 14, 2026, Anthropic's Claude Code published four tagged releases (v2.1.229, v2.1.231, v2.1.232, v2.1.233) in roughly 50 hours, with default behavior changes that can silently bre…

20:10
2026-08-15
github.com
artificial-intelligence

Ctok: Reconstructed Claude Tokenizer

Ctok, an unofficial open-source library, reconstructs Anthropic's Claude tokenizer offline, reporting exact token counts for 1,664,940 v3 and 1,722,961 v4.7 texts with zero under-counts. The library s…

← prev page 2 / 6 next →
// co-occurs with top 8 entities
// topics top 6 topics