Evaluating point-light biological motion in multimodal large language models
Akila Kadambi, Marco Iacoboni, Lisa Aziz-Zadeh, Srini Narayanan
机构
*
Psychiatry and Biobehavioral Sciences, UCLA(乌尔拉克大学精神病学与生物行为科学系)
;
Brain and Creativity Institute, USC(美国大学脑与创造力研究所)
;
Google DeepMind, Zurich(谷歌深度Mind瑞士分公司)
专题命中
GUI与屏幕智能体
:multimodal large language model(title);分类 cs.CV、cs.AI
机构
*
Beijing Institute of Technology(北京理工大学)
;
State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,BIGAI)
;
DataCanvas
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Shenzhen MSU-BIT University(深圳MSU-BIT大学)
IA-VLA: Input Augmentation for Vision-Language-Action models in settings with semantically complex tasks
Eric Hannus, Miika Malin, Tran Nguyen Le, Ville Kyrki
机构
*
Intelligent Robotics Group at the Department of Electrical Engineering and Automation, School of Electrical Engineering, Aalto University(Aalto大学电气工程学院电气工程与自动化系智能机器人组)
;
Biomimetics and Intelligent Systems Group at the Faculty of Information Technology and Electrical Engineering, University of Oulu(奥卢大学信息科技与电气工程学院仿生学与智能系统组)
;
Section of Mechanical Technology at the Department of Engineering Technology and Didactics, Technical University of Denmark(丹麦技术大学工程技术与教学系机械技术部门)
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Microsoft Research(微软研究院)
;
Nanjing University(南京大学)
;
Central South University(中南大学)
;
Zhejiang University(浙江大学)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)