跳到正文
RCreddit.com·

3.8-27B has ruined 3.5/3.6-35B’s for me. It’s just *absurdly* superior.

AI 摘要

一位开发者指出,3.8-27B 模型“荒谬地优于”3.5/3.6-35B 模型,包括 vanilla、kat、Ornith/tiel 和 nex-2 版本。3.8-27B 与这些旧模型之间的差距,远大于 5.3(flash 和 regular)与 3.8-27B 之间的差距。这种优越性在五个重复项目中得到了验证,这些项目涵盖了工作流设计、数据管道、结果分析和数据在线发布等环节。

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

发布当时偏移:UTC+02026年9月12日 10:16 UTC

收录当时偏移:UTC+02026年9月12日 22:01 UTC

发布
2026年9月12日 10:16
收录
2026年9月12日 22:01
来源类型
开发者社区
档位
社区
信源状态
正常

档位是按信源手工设定的编辑判断,不是逐条打分。

正文

注册后,可在设置中选择默认翻译语言。

Applied science work, from workflow design, data pipeline, results analysis, article/reports writing and data publishing online. 5 projects I did in the past replicated from start to finish.

3x to 4x more total wall time. Yes, HUGE toll on how much you can do in a day if this was the only model you could use in your laptop.

But oh my…the quality of that thing. The stupid level of attention to detail. I have the Z.ai api, so I can compare it with 5.3 and 5.3-flash:

The gap between 5.3 (flash and regular) and 3.8-27B is much less, smaller when not plain tiny, than the gap between 3.8-27B and any of the 3.5/3.6-35B-A3B (vanilla, kat, Ornith/tiel, nex-2).

But it’s also spending 22 to 33% less tokens (effort =medium) and less ram footprint, so you get more done without hitting limits,compaction, etc.

So yeah, guess I’ll sip more tea, play the piano, whatever. Let that fat bottom Qwen work.

来源·reddit.com