返回
查看原文
查看原文
RCreddit.com·

Expert expansion with llama.cpp

AI 摘要

一位开发者利用Glm 5.3 flash,创建了一个llama.cpp的自定义分支,旨在支持MOE模型的Expert扩展。该开发者报告称,这个名为“moex-expansion”的新版本在金属硬件上的表现优于他们之前的DS4版本。目前,他们正在寻求来自其他平台和不同模型的反馈,以进一步评估其性能和稳定性。

时间与来源
发布
2026年9月6日 18:27
来源类型
开发者社区
档位
社区
信源状态
正常
档位是按信源手工设定的编辑判断,不是逐条打分。

时间以 UTC 显示

更多信息
首次发现2026年9月6日 23:00时区UTC · UTC+0
正文

With the help of Glm 5.3 flash I built a custom branch of llama.cpp in order to support Expert expansion with MOE models, I've tested only on metal and It works better than my DS4 version, i need feedback from other platforms, and different models.