跳到正文
RCreddit.com·
暂不在当前实时榜单

Transformers vs RNNs vs SSMs: Where Does Memory Actually Live? [D]

AI 摘要

The discussion explores where memory resides in AI architectures like RNNs, Transformers, and SSMs, viewing them through the lens of working memory. It questions whether memory is a compact recurrent state, a growing KV cache, or integrated within the network itself. An example, BDH (Dragon Hatchling), uses linear attention and a low-rank GPU implementation, where recurrent attention state is an N × D matrix. This approach suggests a synaptic interpretation of working memory, aligning it with learned connectivity, though fixed-size states still have finite information capacity.

为什么是这条

This report uniquely frames the comparison of RNNs, Transformers, and SSMs by focusing on where memory actually lives, unlike typical architectural horse races.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年10月6日 22:00 UTC

收录
2026年10月6日 22:00
来源类型
开发者社区

本站未收录正文。

前往源站阅读 →
来源·reddit.com