21:07
2026-09-02
frontierswe.com
artificial-intelligence
Fable 5.1 way ahead on revised FrontierSWE long horizon benchmarks
Anthropic released FrontierSWE v2, a benchmark with 21 new ultra-long horizon tasks totaling 34, and reported that its Claude Fable 5.1 model leads the ranking, followed by GPT-5.6 and GLM-5.3. The reβ¦