Omni-MMSI: Toward Identity-attributed Social Interaction Understanding
Omni-MMSI:迈向基于身份的社会互动理解
Xinpeng Li, Bolin Lai, Hardy Chen, Shijian Deng, Cihang Xie, Yuyin Zhou, James Matthew Rehg, Yapeng Tian
机构
*
University of Texas at Dallas(德克萨斯大学达拉斯分校)
;
Georgia Institute of Technology(佐治亚理工学院)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
机构
*
Computer Science, Taibah University, Saudi Arabia(塔布大学计算机科学系,沙特阿拉伯)
;
La Sapienza, University of Rome, Rome, Italy(罗马大学拉·萨皮恩扎分校,意大利)
;
CISPA Helmholtz Center for Information Security, Saarbrücken, Germany(信息安全赫尔姆霍茨研究中心,德国萨尔布吕肯)
;
University of Cagliari, Cagliari, Italy(卡利亚里大学,意大利卡利亚里)
Listening with the Eyes: Benchmarking Egocentric Co-Speech Grounding across Space and Time
用眼睛倾听:跨时空的自体视觉共指基准测试
Weijie Zhou, Xuantang Xiong, Zhenlin Hu, Xiaomeng Zhu, Chaoyang Zhao, Honghui Dong, Zhengyou Zhang, Ming Tang, Jinqiao Wang
机构
*
Beijing Jiaotong University(北京交通大学)
;
Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences (CASIA)(基础模型研究中心、自动化研究所、中国科学院(CASIA))
;
Tencent Robotics X(腾讯机器人X)
;
Department of Computer Science and Engineering, The Hong Kong University of Science and Technology (HKUST)(计算机科学与工程系、香港科学与技术大学(HKUST))
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学深圳学院)