跳到正文
HNHacker News·
暂不在当前实时榜单

Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)

AI 摘要

A new paradigm called Cache-to-Cache (C2C) enables direct semantic communication between Large Language Models (LLMs), addressing limitations of text-based communication. C2C projects and fuses the KV-cache of source and target models using a neural network, allowing direct semantic transfer and avoiding explicit intermediate text generation. Experiments show C2C achieves 6.4-14.2% higher average accuracy than individual models and outperforms text communication by 3.1-5.4%, with a 2.5x speedup in latency. This method leverages rich semantic information for improved performance and efficiency.

为什么是这条

This paper introduces Cache-to-Cache, a novel paradigm for direct LLM semantic communication, unlike previous methods that rely on text, achieving 2.5x speedup and higher accuracy.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月18日 21:00 UTC

收录
2026年9月18日 21:00
来源类型
研究

讨论趋势

暂无对比
最近 24 小时与此前 24 小时的快照均值对比 · 7 天曲线

百分比基于采集到的讨论信号,不代表新增评论数或独立参与人数。曲线仅用于同一话题在不同时段的比较。

本站未收录正文。

前往源站阅读 →
来源·Hacker News·arxiv.org