返回
RCreddit.com
12
·2天前·RSS
暂不在当前实时榜单

Experience report - Qwen 3.8 Flash Next on memory rich, GPU poor setup

查看原文
LlamaQwen模型发布

热度趋势

趋势数据积累中

百分比基于当前可用热度信号,而非评论数或独立用户人数。

推荐理由

Llama 相关模型动态已经出现,适合跟踪能力变化、生态影响和后续可用性。

AI 摘要

A user shared an experience report on running Qwen 3.8 Flash Next on a memory-rich, GPU-poor setup. The setup utilized llama-server with a Qwen3.6-35B-A3B model, specifying --ctx-size 131744, --batch-size 1024, and --flash-attn on. The user also configured --n-gpu-layers 999 and --n-cpu-moe 26, seeking comparisons with other constrained setups.