1
Out of Sight, Still in Mind: Token Compression for Omni-LLMs
多模态大模型也能「压缩记忆」?这篇论文提出在不牺牲理解能力的前提下,高效压缩视觉与文本token,让全模态LLM更轻更快。
arXiv:2607.21179v1 Announce Type: new Abstract: The goal of this paper is to reduce the input token cost of Omni-modal large language models (Omni-LLM…