15:08
2026-08-04
github.com
artificial-intelligence
Benchmarking Fable, Sol, and Kimi K3 on SlopCodeBench
In a benchmark run on Thursday, Fable and Sol tied at 33.3% (10/30 strict checkpoint passes) on SlopCodeBench, with Expo and Kimi K3 following at 26.7% and 23.3%, respectively. The evaluation, conductβ¦