06:50
2026-07-08
developer.microsoft.com
large-language-models
Not all model upgrades are upgrades
Newer AI models with cheaper per-token pricing can lead to higher costs and worse output, according to tests comparing Claude Sonnet 4.6 and Sonnet 5 on 150 agent tasks. Sonnet 5 consumed up to 12x moβ¦