Steering Vision-Language-Action Models as Anti-Exploration: A Test-Time Scaling Approach
引导视觉-语言-动作模型作为反探索:一种测试时间缩放方法
机构 * Institute of Artificial Intelligence, China Telecom(中国电信人工智能研究院) ; University of Science and Technology of China(中国科学技术大学) ; Tsinghua University(清华大学) ; The Hong Kong University of Science and Technology(香港科学与技术大学)
AI总结 本文提出TACO框架,通过测试时间缩放方法在视觉-语言-动作模型中引入反探索机制,提升推理稳定性和任务成功率。
Comments The first two authors contributed equally. Yang Zhang leads the whole project