VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching
机构 * School of Computer Science, University of Sydney(悉尼大学计算机科学学院) ; John Hopcropt Center for Computer Science, Shanghai Jiao Tong University(上海交通大学约翰·霍普克罗夫特计算机科学中心)
专题命中 VLA模型 :vision-language-action(title,abstract);VLA(title,abstract);分类 cs.RO、cs.CV、cs.LG
Comments Accepted to NeurIPS 2025