跳到正文
RCreddit.com·

Qwen 3.8 Next Flash is really really REALLY verbose..

AI 摘要

一位用户从 Qwen 3.6 27b 升级到 Qwen 3.8 Next Flash 后反映,新模型过于冗长,处理单轮编码请求需要长达 13 分钟。尽管在直接任务中输出质量尚可,但该模型在需要决策时表现不佳,生成充斥着行业术语和行话的回复。用户不愿降低“思考级别”,担心这会影响模型质量,并指出 Qwen 3.8 27b 的性能受此设置影响显著。

时间与来源
发布
2026年9月7日 09:31
来源类型
开发者社区
档位
社区
信源状态
正常
档位是按信源手工设定的编辑判断,不是逐条打分。

时间以 UTC 显示

更多信息
首次发现2026年9月7日 13:00时区UTC · UTC+0
正文

Long time user of 3.6 27b, switched over to Next Flash since it's a logical step up even from 3.8 27b. It's soooo verbose, i'm talking 13 minutes of thinking time on single turn coding requests at approximately 150 tokens per second tg and 7000 tokens per second pp. It's honestly kind of painful to use since I look back and it's still thinking, then when I go to check the output it's decent most of the time but if the task requires ANY decision making, it turns into alphabet soup where it's just buzzwords and jargon that nobody actually uses in the SWE space.

The runtime is actually shorter for me if I BYOK it to VSCode, but for pi.dev it's often takes 1 hour!

Before anyone tells me to lower the thinking level, I don't want to do that given the chance it makes the model worse. There's no solid benchmarks for how the model performs at different thinking levels yet, but looking towards 3.8 27b, it seems to affect the quality of the output quite a bit.

来源·reddit.com·RSS 全文