跳到正文
RCreddit.com·
暂不在当前实时榜单

Hoping for Optimized Smarter Upcoming Models .... Like DeepSeek-V4.1-Flash( KVCache + Engram) in Small/Medium/Big sizes

AI 摘要

The DeepSeek-V4.1-Flash model, incorporating KVCache and Engram, is anticipated to offer optimized performance across small, medium, and large sizes. Upcoming models like Qwen4.0-27B-Q8, Muse-Glimmer-2-30B-Q8, and Gemma-5-31B are projected to significantly reduce the total GB required for a 256K KVCache F16, with some models needing as little as 1 GB. This suggests that 32GB VRAM may be sufficient for these advanced models, with future inventions potentially reducing requirements to 24GB.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月13日 21:01 UTC

收录
2026年9月13日 21:01
来源类型
开发者社区

本站未收录正文。

前往源站阅读 →
来源·reddit.com