LessThink-Qwen3-4B: the same model, with far less thinking [P] A developer posting as stey1r released LessThink-Qwen3-4B, a post-trained version of Qwen3-4B that uses 44% fewer tokens on reasoning while retaining the base model's knowledge and answer style, with the full pipeline run on a single GPU. The model is available at https://5ivatej.com/lessthink/. I post-trained Qwen3-4B to spend 44% fewer tokens on reasoning, keeping its knowledge and answer style. The whole pipeline ran on one GPU. folks, you can check it out on : https://5ivatej.com/lessthink/ submitted by /u/stey1r