返回
RCreddit.com
14
·14小时前·RSS
暂不在当前实时榜单

Android Studios native Gemma 4 runs on llama.cpp

查看原文
Llama模型发布

热度趋势

新上榜
最近 24 小时与此前 24 小时对比 · 7 天曲线

百分比基于当前可用热度信号,而非评论数或独立用户人数。

推荐理由

Llama 相关模型动态已经出现,适合跟踪能力变化、生态影响和后续可用性。

AI 摘要

Android Studios' native Gemma 4, running on llama.cpp, is speculated to be the Vulkan and QAT versions. It supports multi-GPU configurations and the 31B model boasts a maximum context length of 128k. When fully loaded, Gemma 4 utilizes 34 GB of VRAM. Currently, there are no visible options to adjust the context length or display PP/TG speed within the application.