Vision-Language-Action in Robotics: A Survey of Datasets, Benchmarks, and Data Engines
机器人中的视觉-语言-动作:数据集、基准和数据引擎的综述
Ziyao Wang, Bingying Wang, Hanrong Zhang, Tingting Du, Tianyang Chen, Guoheng Sun, Yexiao He, Zheyu Shen, Wanghao Ye, Ang Li
机构
*
University of Maryland, College Park(马里兰大学学院公园分校)
;
University of Utah(犹他大学)
;
Northeastern University(东北大学)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
机构
*
MoE Key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(上海交通大学人工智能研究院教育部人工智能重点实验室)
;
Central Research Institute, Huawei(华为中央研究院)
HarvestFlex: Strawberry Harvesting via Vision-Language-Action Policy Adaptation in the Wild
HarvestFlex: 通过视觉-语言-动作策略适应实现草莓采摘
Ziyang Zhao, Shuheng Wang, Zhonghua Miao, Ya Xiong
机构
*
The Intelligent Equipment Research Center, Beijing Academy of Agriculture and Forestry Sciences(北京农业与林业科学研究院智能装备研究所)
;
The School of Mechanical Electrical Engineering and Automation, Shanghai University(上海大学机械电子工程与自动化学院)
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
The Grainger College of Engineering, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校格拉inger工程学院)
;
CAS Center for Excellence in Brain Science and Intelligence Technology(中国科学院脑科学与智能技术卓越创新中心)
;
Joint Laboratory of Intelligence Science and Technology, Institute of Systems Engineering, Macau University of Science and Technology(澳门科技大学系统工程学院智能科学与技术联合实验室)