跳到正文
RCreddit.com·
暂不在当前实时榜单

This draft model is OP on 16 GB cards for Qwen 3.8 27b

AI 摘要

A user reported that a draft model, specifically HermiHg/Qwen3.8-27B-DFlash2-Q2_K_S-MIX-GGUF, performs exceptionally well on 16 GB cards. When paired with ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF (IQ3_XXS with 128k context), it achieved an average token generation speed of about 60 tokens per second on an RX 9070 XT graphics card. Further testing by another individual on a different card is also mentioned.

为什么是这条

This report uniquely details specific performance metrics for the HermiHg/Qwen3.8-27B-DFlash2-Q2_K_S-MIX-GGUF draft model on 16 GB cards, unlike general discussions of model efficiency.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月13日 06:01 UTC

收录
2026年9月13日 06:01
来源类型
开发者社区

本站未收录正文。

前往源站阅读 →
来源·reddit.com