跳到正文
RCreddit.com·
暂不在当前实时榜单

600tok/s single request on qwen3.6 35ba3b with Ninfer on an RTX Pro 6000. Anybody remember that Comcast ad "stupid fast"?

AI 摘要

A user achieved 600 tokens/second on a single request using the qwen3.6 35ba3b model with Ninfer on an RTX Pro 6000. While acknowledging it's not the most intelligent model, they find it suitable for "read+find" or coding tasks where brute-force processing is acceptable. Despite potentially using 20 times more tokens, its speed surpasses many other local models, making it a "stupid fast" and enjoyable tool.

为什么是这条

This report highlights a specific model, qwen3.6 35ba3b, achieving 600 tokens/second, a speed described as "stupid fast" compared to many other local models.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月18日 07:00 UTC

收录
2026年9月18日 07:00
来源类型
开发者社区

讨论趋势

→ 平稳
最近 24 小时与此前 24 小时的快照均值对比 · 7 天曲线

百分比基于采集到的讨论信号,不代表新增评论数或独立参与人数。曲线仅用于同一话题在不同时段的比较。

本站未收录正文。

前往源站阅读 →
来源·reddit.com