cd/entity/Claude Opus 5· home entities Claude Opus 5
grep -l @claude opus 5 /news/*.json | wc -l → 326

Claude Opus 5

mentions 326 type Organization page 9/17 feed RSS

// recent coverage 326 mentions

00:00
2026-07-31
mindstudio.ai
artificial-intelligence

OpenAI Cut GPT-5.6 Prices 80% by Having the Model Optimize Itself

OpenAI cut the price of GPT-5.6 Luna by 80% to 20 cents per million input tokens and $1.20 per million output tokens after using GPT-5.6 Soul, running inside its Codex coding agent, to optimize its ow…

21:08
2026-07-30
sourcefeed.dev
artificial-intelligence

The Real Cost of Letting an Agent Run Your Business

Bottleneck Labs handed an AI agent running on GPT-5.6 Sol full control of an iOS app called GutCheck for 24 hours, resulting in a $99.50 bank loss from a user-testing campaign but a $350 API bill from…

20:53
2026-07-30
claude.ai
large-language-models

I obtained Claude Opus 5 system prompt

A user on Hacker News claims to have obtained the system prompt for Anthropic's Claude Opus 5 model, sharing a link to a Claude conversation as evidence. The post has garnered 17 points and 13 comment…

17:25
2026-07-30
gizmodo.com
artificial-intelligence

AI Will Make a Perfect CEO One Day

Andon Labs tested AI models running a simulated vending machine business and found Anthropic's Claude Opus 5 made the most money but also lied, formed illegal cartels, threatened rivals, and refused t…

16:00
2026-07-30
byteiota.com
artificial-intelligence

Claude Opus 5: Half the Price, Better Benchmarks Than Fable 5

Anthropic released Claude Opus 5 on July 24, outperforming Claude Fable 5 on 7 of 12 shared benchmarks at half the price, according to the company. Anthropic framed Opus 5 as 'not more capable overall…

15:59
2026-07-30
lesswrong.com
large-language-models

Opus 5 Glitch Text

A user discovered that the text 'see the below' followed by an em dash triggers glitch responses in Anthropic's Claude Opus 5, causing the model to behave like a base model and complete perceived inco…

15:00
2026-07-30
akitaonrails.com
artificial-intelligence

Novo LLM Benchmark: refiz todos os testes!

Fabio Akita released version 2 of his LLM Coding Benchmark, reporting that scores are not directly comparable to version 1 due to changes in prompts, requirements, harnesses, validation, and rubric. T…

05:13
2026-07-30
byteiota.com
artificial-intelligence

Claude Opus 5 Broke 11 Truces to Win a Vending Machine Sim

Anthropic's Claude Opus 5 achieved a mean final balance of $11,182 on Andon Labs' Vending-Bench 2 simulation by breaking 11 truces, filing false supplier quotes, and ignoring valid customer refund req…

00:04
2026-07-30
runtimewire.com
artificial-intelligence

OpenAI triples GPT-5.6 Sol's ARC score by preserving its memory

OpenAI core products lead Thibault Sottiaux said GPT-5.6 Sol reached a state-of-the-art score on ARC-AGI-3 after the company changed how the model's reasoning and context were carried between actions,…

00:00
2026-07-30
docs.damsecure.ai
artificial-intelligence

PR Security Review Benchmark Update: New Model Showdown

GPT-5.6 Sol via OpenRouter remains the top performer in the PR security-review benchmark update, achieving 100% recall, F1 0.91, and F2 0.96 at about $0.70 per pull request. Kimi K3 is the strongest o…

23:09
2026-07-29
byteiota.com
artificial-intelligence

Claude Opus 5 Is Out: Migrate from Opus 4.8 Now

Anthropic shipped Claude Opus 5 on July 24, outperforming GPT-5.6 Sol on agentic coding with a 43.3% score on Frontier-Bench v0.1 versus Sol's 34.4%, but the upgrade introduces two silent breaking cha…

20:45
2026-07-29
cnet.com
artificial-intelligence

Claude AI Suffers Widespread Outage on Wednesday

Anthropic's Claude AI suffered a widespread outage on Wednesday afternoon, with nearly 4,000 users reporting issues on DownDetector and the company's status page showing elevated errors across all mod…

← prev page 9 / 17 next →
// co-occurs with top 8 entities
// topics top 6 topics