cd/entity/Sonnet· home entities Sonnet
grep -l @sonnet /news/*.json | wc -l → 137

Sonnet

mentions 137 type Organization page 3/7 feed RSS

// recent coverage 137 mentions

11:42
2026-07-16
blog.getcassis.com
large-language-models

Do you put rules or examples in your LLM context?

A study by Cassis on LLM context optimization for data analytics found that adding a full example library raised Haiku's accuracy from 70% to 82% and Sonnet's from 81% to 93%, while removing rules and…

06:49
2026-07-16
gist.github.com
ai-agents

Fable Orchestrator

A developer introduced Fable Orchestrator, a multi-agent system that coordinates specialized AI executors—including models like Sonnet, Opus, and Codex gpt-5.6-sol—for software development tasks. The …

12:30
2026-07-15
andywidjaja.com
artificial-intelligence

The $110/month self-improving pipeline

A developer has built an open-source system called autoloop that autonomously triages, implements, and tests GitHub issues using Claude AI, achieving 27 autonomous merges in two weeks with a 97% succe…

04:25
2026-07-15
machinebrief.com
artificial-intelligence

EG-VAR: Setting a New Standard in AI Reasoning

EG-VAR, a Lean 4-based architecture for AI reasoning, achieved a flawless 120 out of 120 on TableBench numerical reasoning tasks and maintained 100% source fidelity during counterfactual stress tests,…

22:10
2026-07-12
gist.github.com
developer-tools

Setup 2025 - https://www.youtube.com/watch?v=6M7LgYkxS4g

A developer shared a starter configuration for Claude Code, a tool that uses AI to assist with coding tasks. The configuration includes a curated allowlist of safe commands, integration with the open-…

00:09
2026-07-12
byteiota.com
ai-agents

Amazon Project Moonraker: What Developers Need to Know

Amazon is developing Project Moonraker, an upgrade to Alexa that would enable multi-step task execution from a single command, according to leaked internal documents. The project, which relies on Anth…

00:00
2026-07-09
quesma.com
artificial-intelligence

Tokenflation: When “Hi” triggers 33 tool calls

A new phenomenon called 'tokenflation' shows AI coding agents using excessive tool calls and reasoning for trivial inputs like 'Hi', costing developers time and money. Benchmarking 14 models revealed …

18:00
2026-07-08
lightningjar.com
large-language-models

The Benchmark Said No | barkup-bench Studies L and M

The barkup-bench series found that LLM agents editing structured trees fail when conversation history is removed and when models must locate nodes without stable IDs. Study M showed stateless sessions…

13:05
2026-07-08
lightningjar.com
large-language-models

Stable IDs Are All You Need | barkup-bench

Researchers at barkup-bench ran seven pre-registered studies with over 13,000 model runs to determine the most reliable way for LLM agents to edit structured data. They found that giving every node a …

06:28
2026-07-08
github.com
ai-tools

Fable Advisor

Claude Code's Fable Advisor plugin implements an architect pattern that routes implementation tasks to the cheapest adequate AI model, using Fable 5 for judgment and Sonnet for routine coding, reducin…

← prev page 3 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics