跳到正文
RCreddit.com·
暂不在当前实时榜单

Gufo performance .... 70tps Qwen 3.8 27b but you need to read the fine print.

AI 摘要

A user tested Gufo 0.4.0 with Qwen3.8 27B UD-Q4_K_XL and a DFlash2 Q4_K_M draft model, achieving 70.56 tok/s for a single user and 123 tok/s with eight users on a Strix Halo device. The setup used Gufo's Podman image and benchmark script with specific settings like greedy decoding and 128 output tokens. The user plans to further evaluate output quality.

为什么是这条

This report provides a first-hand, detailed account of Gufo 0.4.0's performance with specific Qwen models, unlike other discussions that only cite headline figures.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年10月2日 03:00 UTC

收录
2026年10月2日 03:00
来源类型
开发者社区

本站未收录正文。

前往源站阅读 →
来源·reddit.com