AIR-VLA: Vision-Language-Action Systems for Aerial Manipulation
AIR-VLA:面向空中操作的视觉-语言-动作系统
机构 * The Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; School of Automation, Beijing Institute of Technology(北京理工大学自动化学院) ; School of Information and Intelligent Engineering, University of Sanya(三亚大学信息与智能工程学院) ; School of Mechanical and Vehicle Engineering, Hunan University(湖南大学机械与车辆工程学院) ; State Key Lab of Intelligent Transportation Systems, School of Transportation Science and Engineering, Beihang University(北京航空航天大学交通科学与工程学院)
专题命中 空间理解 :spatial understanding(abstract);分类 cs.RO
AI总结 AIR-VLA提出首个针对空中操作的视觉-语言-动作系统,通过构建仿真环境和多模态数据集,评估主流模型并揭示其在无人机移动、机械臂控制和高层规划中的能力和限制。