RCreddit.com·
LessThink-Qwen3-4B: the same model, with far less thinking [P]
A developer has post-trained the Qwen3-4B model, creating "LessThink-Qwen3-4B" which significantly reduces token usage for reasoning by 44% while maintaining its original knowledge and answer style. This entire process was achieved using a single GPU. The developer invites interested individuals to explore this new model further on their website.
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
发布当时偏移:UTC+02026年9月30日 07:19 UTC
收录当时偏移:UTC+02026年9月30日 14:00 UTC
- 发布
- 2026年9月30日 07:19
- 收录
- 2026年9月30日 14:00
- 来源类型
- 开发者社区
- 档位
- 社区
- 信源状态
- 正常
档位是按信源手工设定的编辑判断,不是逐条打分。
I post-trained Qwen3-4B to spend 44% fewer tokens on reasoning, keeping its knowledge and answer style. The whole pipeline ran on one GPU.
folks, you can check it out on: https://5ivatej.com/lessthink/