13:29
2026-10-08
github.com
large-language-models
Sub-1-Bit LLM Compression via Latent Factorization
Banseok Lee and Youngmin Kim released the official implementation of LittleBit (NeurIPS 2025) and LittleBit-2 (ICML 2026), which compress large language models into the sub-1-bit regime down to 0.1 bi…