DRKF: Decoupled Representations with Knowledge Fusion for Multimodal Emotion Recognition
Peiyuan Jiang, Yao Liu, Qiao Liu, Zongshun Zhang, Jiaye Yang, Lu Liu, Daibing Yao
机构
*
School of Computer Science and Engineering, University of Electronic Science and Technology of China(电子科技大学计算机科学与工程学院)
;
School of Information and Software Engineering, University of Electronic Science and Technology of China(电子科技大学信息与软件工程学院)
AMMNet: An Asymmetric Multi-Modal Network for Remote Sensing Semantic Segmentation
Hui Ye, Haodong Chen, Zeke Zexi Hu, Xiaoming Chen, Yuk Ying Chung
机构
*
School of Computer Science, The University of Sydney(计算机科学学院,悉尼大学)
;
School of Computer and Artificial Intelligence, Beijing Technology and Business University(计算机与人工智能学院,北京技术与商业大学)
DUNIA: Pixel-Sized Embeddings via Cross-Modal Alignment for Earth Observation Applications
Ibrahim Fayad, Max Zimmer, Martin Schwartz, Fabian Gieseke, Philippe Ciais, Gabriel Belouze, Sarah Brood, Aurelien De Truchis, Alexandre d'Aspremont
机构
*
Laboratoire des Sciences du Climat et de l’Environnement, LSCE/IPSL, France
;
Department for AI in Society, Science
;
Technology, Zuse Institute Berlin, Germany
;
Department of Information Systems, University of Münster, Germany
;
Department of Computer Science, CNRS, INRIA \& École Normale Supérieure, Paris 75230, France
Boosting Multimodal Learning via Disentangled Gradient Learning
Shicai Wei, Chunbo Luo, Yang Luo
机构
*
The Laboratory of Intelligent Collaborative Computing of UESTC(UESTC智能协同计算实验室)
;
The School of Information and Communication Engineering of UESTC(UESTC信息与通信工程学院)
Multi-omic Prognosis of Alzheimer's Disease with Asymmetric Cross-Modal Cross-Attention Network
Yang Ming, Jiang Shi Zhong, Zhou Su Juan
机构
*
College of Medical Information Engineering, Guangdong Pharmaceutical University, Guangzhou, Guangdong 510006, China(医学信息工程学院,广东药科大学,广州,广东510006,中国)
Multi-level Mixture of Experts for Multimodal Entity Linking
Zhiwei Hu, Víctor Gutiérrez-Basulto, Zhiliang Xiang, Ru Li, Jeff Z. Pan
机构
*
School of Computer
;
Information Technology Shanxi University Taiyuan China
;
School of Computer Science
;
Informatics Cardiff University Cardiff UK
;
ILCC, School of Informatics University of Edinburgh Edinburgh UK
;
Information Technology Shanxi University
;
Informatics Cardiff University
;
ILCC, School of Informatics University of Edinburgh
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs
Tao Zhang, Shiqing Wei, Shihao Chen, Wenling Yu, Muying Luo, Shunping Ji
机构
*
School of Remote Sensing and Information Engineering(遥感与信息工程学院)
;
College of Oceanography and Space Informatics(海洋学与空间信息学院)
;
School of Surveying and Geoinformation Engineering(测绘与地理信息工程学院)
机构
*
Hong Kong Baptist University(香港 Baptist 大学)
;
Sichuan University(四川大学)
;
Shanghai AI Lab(上海人工智能实验室)
;
Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))
DALR: Dual-level Alignment Learning for Multimodal Sentence Representation Learning
Kang He, Yuzhe Ding, Haining Wang, Fei Li, Chong Teng, Donghong Ji
机构
*
Key Laboratory of Aerospace Information Security and Trusted Computing, Ministry of Education, School of Cyber Science and Engineering, Wuhan University(航天信息安全部门与可信计算重点实验室,教育部,网络安全科学与工程学院,武汉大学)
CommentsAccepted in the 2025 IEEE International Geoscience and Remote Sensing Symposium (IGARSS 2025), scheduled for 3 - 8 August 2025 in Brisbane, Australia
UniFuse: A Unified All-in-One Framework for Multi-Modal Medical Image Fusion Under Diverse Degradations and Misalignments
Dayong Su, Yafei Zhang, Huafeng Li, Jinxing Li, Yu Liu
机构
*
Kunming University of Science and Technology(昆明理工大学)
;
Harbin Institute of Technology at Shenzhen(哈尔滨工业大学深圳研究院)
;
Hefei University of Technology(合肥工业大学)
机构
*
Shanghai AI Laboratory(上海人工智能实验室)
;
Zhejiang University(浙江大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究所)
;
Nanjing University(南京大学)
;
Shanghai Innovation Institute(上海创新研究院)
Tracing Intricate Cues in Dialogue: Joint Graph Structure and Sentiment Dynamics for Multimodal Emotion Recognition
Jiang Li, Xiaoping Wang, Zhigang Zeng
机构
*
School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院)
;
Institute of Artificial Intelligence, Huazhong University of Science and Technology(华中科技大学人工智能研究院)
;
Hubei Key Laboratory of Brain-Inspired Intelligent Systems, Huazhong University of Science and Technology(华中科技大学脑启发智能系统省重点实验室)
;
Key Laboratory of Image Processing and Intelligent Control (Huazhong University of Science and Technology), Ministry of Education(图像处理与智能控制重点实验室(华中科技大学))
EAGLE: Efficient Alignment of Generalized Latent Embeddings for Multimodal Survival Prediction with Interpretable Attribution Analysis
Aakash Tripathi, Asim Waqas, Matthew B. Schabath, Yasin Yilmaz, Ghulam Rasool
机构
*
Dept. of Machine Learning Moffitt Cancer Center(机器学习系莫菲特癌症中心)
;
Dept. of Cancer Epidemiology Moffitt Cancer Center(癌症流行病学系莫菲特癌症中心)
;
Dept. of Electrical Engineering University of South Florida(电气工程系佛罗里达州立大学)
Multimodal Machine Learning in Mental Health: A Survey of Data, Algorithms, and Challenges
Zahraa Al Sahili, Ioannis Patras, Matthew Purver
机构
*
Queen Mary University of London United Kingdom
;
Queen Mary University of London \& Jo z ef Stefan Institute United Kingdom \& Slovenia
;
Queen Mary University of London
;
Queen Mary University of London \& Jo z ef Stefan Institute