Reconstructing Content with Collaborative Attention for Universal Multimodal Representation Learning
通过协同注意力重建内容以提升多模态嵌入质量
Jiahan Chen, Da Li, Hengran Zhang, Yinqiong Cai, Lixin Su, Jiafeng Guo, Daiting Shi, Dawei Yin, Keping Bi
机构
*
State Key Laboratory of AI Safety(人工智能安全国家重点实验室)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Baidu Inc.(百度公司)
UnsOcc: 3D Semantic Occupancy Prediction in Unstructured Scene via Rendering Fusion
UnsOcc:非结构化场景下基于渲染融合的3D语义占用预测
Ye Wu, Ruiqi Song, Baiyong Ding, Nanxin Zeng, Junjie Cheng, Yunfeng Ai
机构
*
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Waytous Inc.(Waytous公司)
VOIC: Visible-Occluded Integrated Guidance for 3D Semantic Scene Completion
VOIC:可见-遮挡联合引导的3D语义场景补全
Zaidao Han, Risa Higashita, Jiang Liu
机构
*
Research Institute of Trustworthy Autonomous Systems, Southern University of Science and Technology(可信自主系统研究院,南方科技大学)
;
Department of Computer Science and Engineering, Southern University of Science and Technology(计算机科学与工程系,南方科技大学)
;
School of Computer Science, University of Nottingham Ningbo China(宁波大学计算机学院)
;
Department of Electronic and Information Engineering, Changchun University(电子与信息工程学院,长春大学)
机构
*
Informatics Institute, University of Amsterdam, Amsterdam, The Netherlands(阿姆斯特丹大学信息学院)
;
Department of Computer Science, University College London(伦敦大学学院计算机科学系)
;
University of Thessaly, Volos, Greece(塞萨洛尼基大学)
;
Department of Electronic and Electrical Engineering, Trinity College Dublin, Dublin, Ireland(都柏林信任学院电子与电气工程系)
;
School of Information Technology, Halmstad University, Halmstad, Sweden(哈姆斯塔德大学信息科技学院)
;
Amazon AGI, Seattle, USA(亚马逊人工智能研究部)