Composer 2 Technical Report
Anthropic's Composer 2, a specialized model for agentic software engineering, achieves 61.3 on CursorBench, 61.7 on Terminal-Bench, and 73.7 on SWE-bench Multilingual, marking a major accuracy improve…
Anthropic's Composer 2, a specialized model for agentic software engineering, achieves 61.3 on CursorBench, 61.7 on Terminal-Bench, and 73.7 on SWE-bench Multilingual, marking a major accuracy improve…
Anthropic released Claude Opus 5 on July 24 at the same price as Opus 4.8 ($5/$25 per million tokens), closing most of the performance gap with the more expensive Fable 5 ($10/$50). Opus 5 scores 79.2…
Anthropic launched Claude Opus 5, a model that matches or beats its flagship Fable 5 at half the cost, topping third-party leaderboards. Meanwhile, Microsoft, Meta, and Nvidia led a 20-company coaliti…
Z.ai's open-weight GLM 5.2 completed an agentic coding task in 17 minutes at $2.76, while Anthropic's Claude Fable 5 finished in 9 minutes at over $10, with comparable output quality. The open model c…