3.8-27B has ruined 3.5/3.6-35B’s for me. It’s just *absurdly* superior.
一位开发者指出,3.8-27B 模型“荒谬地优于”3.5/3.6-35B 模型,包括 vanilla、kat、Ornith/tiel 和 nex-2 版本。3.8-27B 与这些旧模型之间的差距,远大于 5.3(flash 和 regular)与 3.8-27B 之间的差距。这种优越性在五个重复项目中得到了验证,这些项目涵盖了工作流设计、数据管道、结果分析和数据在线发布等环节。
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
发布当时偏移:UTC+02026年9月12日 10:16 UTC
收录当时偏移:UTC+02026年9月12日 22:01 UTC
- 发布
- 2026年9月12日 10:16
- 收录
- 2026年9月12日 22:01
- 来源类型
- 开发者社区
- 档位
- 社区
- 信源状态
- 正常
档位是按信源手工设定的编辑判断,不是逐条打分。
注册后,可在设置中选择默认翻译语言。
Applied science work, from workflow design, data pipeline, results analysis, article/reports writing and data publishing online. 5 projects I did in the past replicated from start to finish.
3x to 4x more total wall time. Yes, HUGE toll on how much you can do in a day if this was the only model you could use in your laptop.
But oh my…the quality of that thing. The stupid level of attention to detail. I have the Z.ai api, so I can compare it with 5.3 and 5.3-flash:
The gap between 5.3 (flash and regular) and 3.8-27B is much less, smaller when not plain tiny, than the gap between 3.8-27B and any of the 3.5/3.6-35B-A3B (vanilla, kat, Ornith/tiel, nex-2).
But it’s also spending 22 to 33% less tokens (effort =medium) and less ram footprint, so you get more done without hitting limits,compaction, etc.
So yeah, guess I’ll sip more tea, play the piano, whatever. Let that fat bottom Qwen work.