跳到正文
RCreddit.com·
暂不在当前实时榜单

I made a custom llama.cpp build optimized for 7900xtx (one or two). for qwen 3.8 next and 27B. includes optimizations for PciE x4 and tensor parallel. read inside! (no AI slop)

AI 摘要

A developer created a custom llama.cpp build optimized for AMD 7900xtx GPUs, specifically for Qwen 3.8 next and 27B models. This build includes optimizations for PCIe x4 and tensor parallel processing. Key improvements include a +58.88% Flash increase from lazy PLE/load path, +19.95% from GPU MoE expert cache, and +14.32% Flash from MoE MMQ sizing RDNA3, demonstrating significant performance gains for these specific hardware and model configurations.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月8日 22:00 UTC

收录
2026年9月8日 22:00
来源类型
开发者社区
正文

本站未收录正文。

前往源站阅读 →
来源·reddit.com