cd/entity/Sonnet 5· home entities Sonnet 5
grep -l @sonnet 5 /news/*.json | wc -l → 64

Sonnet 5

mentions 64 type Person page 1/4 feed RSS

// recent coverage 64 mentions

00:00
2026-08-14
digitalapplied.com
artificial-intelligence

Budgeting AI Agents on Intro Pricing Built to Double

Google's Gemini 3.7 Flash and 3.6 Flash models carry introductory pricing of $0.75 per million input tokens and $3.75 per million output through December 31, 2026, after which the posted standard rate…

07:00
2026-08-11
vercel.com
artificial-intelligence

Everything hackable will get hacked

Vercel warns that near-frontier open-weight AI models like Kimi K3, which ranks highest on Vercel's DeepSec Bench among open-weight models and roughly matches Sonnet 5 while outperforming Opus 4.8, ca…

00:08
2026-08-11
sourcefeed.dev
artificial-intelligence

Your Model's Knowledge Cutoff Is a Marketing Date

Probing by Shrivu Shankar found that frontier models' actual knowledge can lag their stated cutoffs by months, with Claude Opus 4.7, 4.8, Sonnet 5, and Fable 5 all losing signal around late December 2…

13:58
2026-08-05
androidauthority.com
artificial-intelligence

Is Claude down for you? Here’s what’s going on

Anthropic confirmed that multiple Claude AI models, including Mythos 5, Fable 5, Opus 5, and Sonnet 5, are experiencing degraded performance due to an outage that began around 3:00 AM ET on August 5, …

09:54
2026-08-04
letsdatascience.com
artificial-intelligence

Databricks Benchmarks Coding Agents on Its Own Codebase

Databricks published an internal coding-agent benchmark on July 8 using tasks from its multi-million-line codebase, reporting that model choice, task difficulty, token use, and the agent harness affec…

10:21
2026-08-03
github.com
artificial-intelligence

Claude gen-5 models show significant regression in BullshitBench

Anthropic's Claude generation 5 models (Sonnet 5, Opus 5, Fable 5) show a measurable quality regression on the BullshitBench dataset, engaging with nonsense prompts instead of rejecting them at a high…

23:31
2026-08-01
pub.towardsai.net
artificial-intelligence

The Sonnet 5 Price is Not What You Think It Is

Anthropic launched Sonnet 5 with a promotional rate of $2 per million input tokens and $10 per million output tokens that expires August 31, 2026, after which prices rise to $3 and $15 per million tok…

17:23
2026-07-31
schneier.com
ai-safety

Anthropic’s Opus 5 Is Better at Resisting Prompt Injection

Anthropic's Claude Opus 5 reduced prompt injection attack success from 5.5% to 2.0% within 15 attempts compared to Opus 4.8, making it the most robust model on the IPI benchmark, outperforming all non…

07:00
2026-07-31
supabase.com
ai-agents

Introducing Supabase Evals

Supabase has open-sourced supabase/evals, a benchmark and framework for testing AI agents like Claude Code, Codex, and OpenCode on real Supabase tasks, with results published on supabase.com/evals. In…

15:26
2026-07-30
lesswrong.com
large-language-models

Testing LLMs on Undergraduate Music Theory

A test of five modern LLMs on undergraduate music theory found that GPT 5.6 Sol scored a perfect 100%, while older models like Claude Sonnet 4 scored 0% and GPT 4.1 scored 16%, indicating LLMs have su…

09:08
2026-07-29
sourcefeed.dev
large-language-models

The Viral Claude Context Window Guide Was Written by a Bot

A viral Dev.to tutorial titled "How Claude's Context Window Actually Works: A Deep Dive" was generated by LLaMA 3.3 70B and published without human editing, according to its own disclosure. The tutori…

01:47
2026-07-29
schneier.com
large-language-models

Measuring LLMs’ Ability to Perform Cryptanalysis

A new benchmark, CryptanalysisBench, shows that large language models can perform mathematical cryptanalysis, with Anthropic's Claude Opus 4.8 among five frontier models breaking 65-86% of Tier 1 sche…

page 1 / 4 next →
// co-occurs with top 8 entities
// topics top 6 topics