- SignalDesk3小时前
Original Summary
Spent the weekend turning small LLMs into decision models, with some good results and a lot of learnings along the way. Topping the decision index leaderbard for their categories <a href="https://huggingface.co/spaces/multimodalart/jev-decision-index" rel="nofollow">https://huggingface.co/spaces/multimodalart/jev-decision-ind...</a><p>Sharing two of them: a sub-1B model and a 4B model, both Qwen3.5 fine-tunes. When evaluated on the decision Index shared last week, the 0.8b tops the sub-1B category, 45.7% above the best other Qwen3.5-0.8B fine-tune on the leaderboard. and the 4b comes in second in the 3–6B class.<p>Most of work was data calibration, finding external datasets and readapting them towards this scenario, so i expect a lot of improvements and work like this to come from community and push these numbers even higher. Even more if we get qwen 3.8 releases for these model categories.
- 情报分类:技术学习与提效
- 分类依据:内容涉及技术、AI、软件工具或工程实践
- 信息来源:Hacker News 新项目
- 发布时间:2026/10/6 01:02:16
- 暂无回复