Reasoning-VLA: A Fast and General Vision-Language-Action Reasoning Model for Autonomous Driving
Reasoning-VLA: 一种用于自动驾驶的快速且通用的视觉-语言-动作推理模型
机构 * Lanzhou University(兰州大学) ; National University of Singapore(新加坡国立大学) ; University of Science and Technology of China(中国科学技术大学) ; Tsinghua University(清华大学) ; University of New South Wales(新南威尔士大学)
专题命中 其他自动驾驶 :autonomous driving(title,abstract);分类 cs.RO、cs.CV
AI总结 Reasoning-VLA通过引入可学习的动作查询和链式思维推理的数据格式,实现了在自动驾驶中快速且通用的视觉-语言-动作推理。