Spent the weekend turning small LLMs into decision models, with some good results and a lot of learnings along the way. Topping the decision index leaderbard for their categories https://huggingface.co/spaces/multimodalart/jev-decision-ind...
Sharing two of them: a sub-1B model and a 4B model, both Qwen3.5 fine-tunes. When evaluated on the decision Index shared last week, the 0.8b tops the sub-1B category, 45.7% above the best other Qwen3.5-0.8B fine-tune on the leaderboard. and the 4b comes in second in the 3–6B class.
Most of work was data calibration, finding external datasets and readapting them towards this scenario, so i expect a lot of improvements and work like this to come from community and push these numbers even higher. Even more if we get qwen 3.8 releases for these model categories.
Comments URL: [https://news.ycombinator.com/item?id=49967455](https://news.ycombinator.com/item?id=49967455)
Points: 1