DASH-KV: Accelerating Long-Context LLM Inference via Asymmetric KV Cache Hashing
DASH-KV:通过非对称KV缓存哈希加速长上下文LLM推理
机构 * School of Information and Software Engineering, University of Electronic Science and Technology of China(电子科技大学信息与软件工程学院) ; School of Computer Science and Engineering, University of Electronic Science and Technology of China(电子科技大学计算机科学与工程学院) ; Bangladesh University of Engineering and Technology(孟加拉工程与技术大学) ; The Hong Kong Polytechnic University(香港理工大学) ; Kyung Hee University, School of Computing(庆熙大学计算机学院)
AI总结 DASH-KV通过非对称深度哈希将注意力机制近似为最近邻搜索,降低计算复杂度至线性,提升长上下文LLM推理效率并保持性能。
Comments Accepted by ACL 2026 (Findings)