跳到正文
RCreddit.com·
暂不在当前实时榜单

Running Qwen 3.8 next on 16vram+32ram - A useful/fun post for the gpu poors

AI 摘要

A Reddit user shared their experience running Qwen 3.8 Next on a system with 16GB VRAM and 32GB RAM, a challenging setup for the large model. They noted that the model's components, including the N-gram/PLE Embedding (~29.48 GB) and MoE Routed Experts (~34.89 GB), require significant memory. Running the model with mmap enabled in llama.cpp resulted in slow speeds of ~2 tok/sec, making it effectively useless. The user suggests that systems with 64GB RAM would benefit more from an optimized llama.cpp fork.

为什么是这条

This report details a specific, challenging setup for Qwen 3.8 Next on limited hardware, unlike other guides that assume more robust systems.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月14日 12:00 UTC

收录
2026年9月14日 12:00
来源类型
开发者社区

本站未收录正文。

前往源站阅读 →
来源·reddit.com