The Role of Visual Modality in Multimodal Mathematical Reasoning: Challenges and Insights
Yufang Liu, Yao Du, Tao Ji, Jianing Wang, Yang Liu, Yuanbin Wu, Aimin Zhou, Mengdi Zhang, Xunliang Cai
机构
*
School of Computer Science and Technology, East China Normal University(东华师范大学计算机科学与技术学院)
;
Meituan Inc.(美团公司)
;
Fudan University(复旦大学)
;
Pazhou Laboratory(Pazhou实验室)
MUDI: A Multimodal Biomedical Dataset for Understanding Pharmacodynamic Drug-Drug Interactions
Tung-Lam Ngo, Ba-Hoang Tran, Duy-Cat Can, Trung-Hieu Do, Oliver Y. Chén, Hoang-Quynh Le
机构
*
VNU University of Engineering and Technology (VNU-UET)(越南工程技术大学)
;
Lausanne University Hospital (CHUV)(洛桑大学医院)
;
University of Lausanne (UNIL)(洛桑大学)
;
Hanoi Medical University(河内医学院)
;
National Geriatric Hospital, Hanoi(河内国立老年医院)
机构
*
School of Data Science, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)数据科学学院)
;
Tencent(腾讯)
;
Renmin University of China(中国人民大学)
;
Johns Hopkins University(约翰霍普金斯大学)
Large Language Model-Driven Distributed Integrated Multimodal Sensing and Semantic Communications
Yubo Peng, Luping Xiang, Bingxin Zhang, Kun Yang
机构
*
State Key Laboratory of Novel Software Technology, Nanjing University(新型软件技术国家重点实验室,南京大学)
;
School of Intelligent Software and Engineering, Nanjing University (Suzhou Campus)(智能软件与工程学院,南京大学(苏州校区))
Reefknot: A Comprehensive Benchmark for Relation Hallucination Evaluation, Analysis and Mitigation in Multimodal Large Language Models
Kening Zheng, Junkai Chen, Yibo Yan, Xin Zou, Xuming Hu
机构
*
Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
Guangxi Zhuang Autonomous Region Big Data Research Institute(广西壮族自治区大数据研究院)
机构
*
State Key Laboratory for Novel Software Technology(新型软件技术国家重点实验室)
;
School of Intelligence Science and Technology(智能科学与技术学院)
;
State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室)
;
XMU(厦门大学)
;
HKU(香港大学)
;
PKU(北京大学)
;
CUHK(香港中文大学)
;
ECNU(华东师范大学)
;
CASIA(中国科学院自动化研究所)
机构
*
School of Computer Science and Engineering, Central South University, China(中南大学计算机科学与工程学院)
;
National University of Singapore, Singapore(新加坡国立大学)
机构
*
School of Electronics and Information Engineering, Harbin Institute of Technology(电子信息工程学院,哈尔滨工业大学)
;
Faculty of Computing, Harbin Institute of Technology(计算机学院,哈尔滨工业大学)
;
Peng Cheng Laboratory, Shenzhen(鹏城实验室,深圳)
MDIT-Bench: Evaluating the Dual-Implicit Toxicity in Large Multimodal Models
Bohan Jin, Shuhan Qi, Kehai Chen, Xinyi Guo, Xuan Wang
机构
*
Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))
;
Guangdong Provincial Key Laboratory of Novel Security Intelligence Technologies(广东省新型安全智能技术重点实验室)
;
University of Barcelona(巴塞罗那大学)
机构
*
Artificial Intelligence Innovation and Incubation Institute, Fudan University(复旦大学人工智能创新与孵化院)
;
Shanghai Academy of Artificial Intelligence for Science(上海人工智能科学研究院)
;
Huashan Hospital, Fudan University(复旦大学华山医院)
;
Human Phenome Institute, Fudan University(复旦大学人类表型研究院)
机构
*
Lappeenranta-Lahti University of Technology LUT(拉佩兰塔-拉赫蒂技术大学)
;
University of Oulu(奥卢大学)
;
Southeast University(东南大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Brno University of Technology(布拉格技术大学)
BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset
Jiuhai Chen, Zhiyang Xu, Xichen Pan, Yushi Hu, Can Qin, Tom Goldstein, Lifu Huang, Tianyi Zhou, Saining Xie, Silvio Savarese, Le Xue, Caiming Xiong, Ran Xu
机构
*
Salesforce Research(Salesforce 研究院)
;
University of Maryland(马里兰大学)
;
Virginia Tech(弗吉尼亚理工大学)
;
New York University(纽约大学)
;
University of Washington(华盛顿大学)
;
UC Davis(加州大学戴维斯分校)
Multi-Modal Explainable Medical AI Assistant for Trustworthy Human-AI Collaboration
Honglong Yang, Shanshan Song, Yi Qin, Lehan Wang, Haonan Wang, Xinpeng Ding, Qixiang Zhang, Bodong Du, Xiaomeng Li
机构
*
Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(电子与计算机工程系,香港科学与技术大学)
;
Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(计算机科学与工程系,香港科学与技术大学)
Twins-PainViT: Towards a Modality-Agnostic Vision Transformer Framework for Multimodal Automatic Pain Assessment using Facial Videos and fNIRS
Stefanos Gkikas, Manolis Tsiknakis
机构
*
Computational Biomedicine Laboratory of the Foundation for Research and Technology (FORTH)(希腊国家研究与技术基金会计算生物医学实验室)
;
Department of Electrical & Computer Engineering Hellenic Mediterranean University(希腊地中海大学电气与计算机工程系)