cd/entity/Opus 5· home entities Opus 5
grep -l @opus 5 /news/*.json | wc -l → 135

Opus 5

mentions 135 type Person page 1/7 feed RSS

// recent coverage 135 mentions

11:04
2026-09-03
dev.to
ai-agents

Superpowers vs Plain Old Debugger in Explyt

A developer's controlled comparison of two AI debugging workflows found that an agent using the Superpowers systematic-debugging skill missed a real bug despite following a structured method, while Ex…

07:09
2026-09-03
byteiota.com
large-language-models

Claude Fable 5.1: The Cache Cut That Changes Agent Costs

Anthropic shipped Claude Fable 5.1 on September 1, cutting cache read prices by 75% from $1.00 to $0.25 per million tokens, which reduces the real cost of highly agentic workloads by up to 45%. The mo…

23:01
2026-09-02
antithetical-labs.com
artificial-intelligence

Ontologies catch untrustworthy LLM claims

A new evaluation method using typed ontologies caught 100% of citation defects in LLM-generated research memos that a standard LLM judge missed, according to an experiment by antithetical-labs. Removi…

15:12
2026-09-02
twitter.com
artificial-intelligence

Gemini 3.8 Flash benchmark scores

Google has released Gemini 3.8 Flash, which according to a post by X user leo 🐾 provides approximately Opus 5 performance at a much lower cost while being super fast. The user notes that real-world pe…

00:00
2026-09-02
mindstudio.ai
artificial-intelligence

Fable 5.1 vs GPT-5.6 vs GLM 5.3: Which Model Actually Wins?

Anthropic's Fable 5.1 tops Artificial Analysis's composite intelligence index with a score of 66, ahead of GPT-5.6 Soul at 61, but independent cost-per-task benchmarking shows it costs $3.69 per task …

23:57
2026-09-01
simonwillison.net
artificial-intelligence

Claude Fable 5.1 made me a really nice animated pelican

Anthropic released Claude Fable (and Mythos) 5.1, which scores 52.6% on the new Terminal-Bench-Science 0.1 benchmark, up from 24.7% for Fable 5, 29.0% for Opus 5, and 22.4% for GPT-5.6 Sol. In a hands…

21:16
2026-09-01
byteiota.com
artificial-intelligence

Claude Fable 5.1: 75% Cheaper Cache, Stronger Agents

Anthropic released Claude Fable 5.1 on September 1, cutting cache read costs by 75% from $1.00 to $0.25 per million tokens, and also released Claude Mythos 5.1 with restricted safety guardrails for cy…

19:39
2026-09-01
techcrunch.com
artificial-intelligence

Anthropic’s new Fable release is cheaper, less restrictive

Anthropic released Fable 5.1 and Mythos 5.1 on Tuesday, with the unrestricted Fable version now available on cloud platforms and via the Anthropic API, while Mythos remains limited to registered partn…

00:00
2026-09-01
vercel.com
artificial-intelligence

Claude Fable 5.1 now available on AI Gateway

Anthropic's Claude Fable 5.1 is now available on Vercel's AI Gateway, featuring improvements for long, multi-stage tasks like agentic coding and research, with cybersecurity and biology safety classif…

21:22
2026-08-29
danilo.segan.org
artificial-intelligence

Postscriptum on LLMs: pelicans on bicycles with a twist

Danilo Segan tested frontier LLMs' ability to generate native PostScript images of a pelican riding a bicycle, finding that models like Gemini 3.7 Flash and Opus 5 required multiple iterations and err…

17:53
2026-08-28
promptcube3.com
ai-safety

Claude Code's Auto Mode is failing its own safety claims

An investigation has found that Anthropic's Claude Code Auto Mode, which uses a safety classifier to autonomously execute commands, is vulnerable to indirect prompt injection attacks, achieving a 60-8…

15:33
2026-08-28
github.com
artificial-intelligence

GLM-5.3-Flash on Apple Silicon

WARP, an embeddable inference engine written in C, now runs the full 2.78-trillion-parameter Kimi K3 model on a 64 GB MacBook Pro at about 0.6 tokens per second, and the 313-billion-parameter GLM-5.3-…

page 1 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics