Escha-W2: 2-Bit Quantization That Shrinks a 27B Model to 10GB
Escha Labs Inc. released Escha-W2, a 2-bit quantized build of Qwen3.8-27B that compresses the 27-billion-parameter model to 10.15GB, enabling 128k context on a single 24GB GPU while matching FP8 quali…