RCreddit.com
14
·14小时前·RSS
暂不在当前实时榜单
Android Studios native Gemma 4 runs on llama.cpp
热度趋势
新上榜
百分比基于当前可用热度信号,而非评论数或独立用户人数。
Llama 相关模型动态已经出现,适合跟踪能力变化、生态影响和后续可用性。
Android Studios' native Gemma 4, running on llama.cpp, is speculated to be the Vulkan and QAT versions. It supports multi-GPU configurations and the 31B model boasts a maximum context length of 128k. When fully loaded, Gemma 4 utilizes 34 GB of VRAM. Currently, there are no visible options to adjust the context length or display PP/TG speed within the application.