09:00
2026-10-03
aiflash.com
large-language-models
LessThink-Qwen3-4B: the same model, with far less thinking [P]
A developer posting as stey1r released LessThink-Qwen3-4B, a post-trained version of Qwen3-4B that uses 44% fewer tokens on reasoning while retaining the base model's knowledge and answer style, with …