How Far Are We from Generating Missing Modalities with Foundation Models?
我们距离用基础模型生成缺失模态还有多远?
Guanzhou Ke, Bo Wang, Guoqing Chao, Weiming Hu, Shengfeng He
机构
*
Institute of Data Science and Intelligent Decision Support, Beijing Jiaotong University(数据科学与智能决策支持研究所,北京交通大学)
;
School of Computing and Information Systems, Singapore Management University(计算与信息系统学院,新加坡管理大学)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
;
School of Computer Science and Technology, Harbin Institute of Technology(计算机科学与技术学院,哈尔滨工业大学)
专题命中
多模态评测
:multimodal(abstract);multimodal foundation model(abstract);分类 cs.CV、cs.CL、cs.MM
机构
*
University of Manchester(曼彻斯特大学)
;
Queen Mary University of London(伦敦大学玛丽女王学院)
;
Hongkong University of Science and Technology(香港科学与技术大学)
;
Nanjing University(南京大学)
;
Dartmouth College(达特茅斯学院)
机构
*
Research Center for Social Computing and Interactive Robotics, Harbin Institute of Technology(社会计算与交互机器人研究中心,哈尔滨工业大学)
;
Faculty of Computing, Harbin Institute of Technology(计算学院,哈尔滨工业大学)
;
School of Computer Science and Engineering, Central South University(计算机科学与工程学院,中南大学)
;
Chinese University of Hong Kong(香港中文大学)
;
MMLab, The Chinese University of Hong Kong(香港中文大学MMLab)
CommentsAccepted by ECCV 2024. A comprehensive and hierarchical 3D reasoning grounding benchmark in the era of foundation models. Project page: https://zcmax.github.io/projects/ScanReason