跳到正文
RCreddit.com·
暂不在当前实时榜单

ggml-cpu: tiled mul_mat for k-quants by jbooth · Pull Request #27851 · ggml-org/llama.cpp

AI 摘要

A recent Pull Request #27851 to ggml-org/llama.cpp by jbooth introduces a significant improvement for CPU prompt processing. This update, titled "ggml-cpu: tiled mul_mat for k-quants," aims to achieve 3-7x faster CPU mul_mat operations. The enhancement is attributed to the use of VNNI, with the developer noting its minimal complexity. This development is expected to boost the performance of llama.cpp on CPUs.

为什么是这条

This pull request is the first to claim a 3-7x speedup for CPU mul_mat operations in llama.cpp, unlike previous updates that offered more modest gains.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月26日 20:00 UTC

收录
2026年9月26日 20:00
来源类型
开发者社区

本站未收录正文。

前往源站阅读 →
来源·reddit.com