# LessThink-Qwen3-4B: the same model, with far less thinking [P]

> Source: <https://aiflash.com/news/130262/>
> Published: 2026-10-03 09:00:45+00:00

I post-trained Qwen3-4B to spend 44% fewer tokens on reasoning, keeping its knowledge and answer style. The whole pipeline ran on one GPU. folks, you can check it out on : https://5ivatej.com/lessthink/ submitted by /u/stey1r
