I Shrank a 428B Model From 855GB to 128GB and It Still Beats GPT-5.5 at Coding
A 428-billion-parameter open-weight model achieved 59.0% on SWE-Bench Pro, outperforming GPT-5.5's 58.6%, after being compressed from 855GB to 128GB. The model's efficiency and performance mark a significant advance in A…