When Fusion Fails: Corruption-Aware Rebalanced Fusion for Multi-Modal Medical Image Segmentation
当融合失败时:面向多模态医学图像分割的腐败感知再平衡融合
Yuchen Pei, Xiaoyu Hu, Yixiong Zou, Dingwen Hu, Hui Chu, Yutao Ma, Shijun Qiu, Gang Li
机构
*
Central China Normal University(华中师范大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
Guangzhou University of Chinese Medicine(广州中医药大学)
;
The First Affiliated Hospital of Guangzhou University of Chinese Medicine(广州中医药大学第一附属医院)
;
University of North Carolina, Chapel Hill(北卡罗来纳大学教堂山分校)
Comments19 pages, 10 figures, 4 tables. Accepted at ACM Multimedia 2026 (MM '26). This arXiv version includes supplementary appendices not included in the conference proceedings version
Comments7 pages, 2 figures, 3 tables. Accepted to the 34th ACM International Conference on Multimedia (MM '26). Ranked 1st in the 3rd Micro-Action Analysis Grand Challenge at ACM MM 2026
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Hunan University(湖南大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
机构
*
Southeast University(东南大学)
;
Shenzhen Loop Area Institute(深圳河套学院)
;
Kobe University(神户大学)
;
Beijing Institute of Technology(北京理工大学)
;
Nanjing Medical University(南京医科大学)
;
The University of Osaka(大阪大学)
Comments5 pages, 1 figure, 3 tables. To appear in the Proceedings of the 34th ACM International Conference on Multimedia (MM '26), November 10-14, 2026, Rio de Janeiro, Brazil. Zhaojie Luo and Junkun Wang contributed equally
Empowering VLMs for Few-Shot Multimodal Time Series Classification via Tailored Agentic Reasoning
通过定制代理推理增强VLMs在少样本多模态时间序列分类中的能力
Lin Li, Jiawei Huang, Qihao Quan, Dan Li, Boxin Li, Xiao Zhang, Erli Meng, Wenjie Feng, Jian Lou, See-Kiong Ng
机构
*
Sun Yat-sen University(中山大学)
;
Xiaomi Corporation(小米公司)
;
University of Science and Technology of China(中国科学技术大学)
;
National University of Singapore(新加坡国立大学)
机构
*
School of Cyber Science and Engineering, Wuhan University(武汉大学网络空间安全学院)
;
School of Mathematics and Statistics, Wuhan University(武汉大学数学与统计学院)
;
School of Synthetic Biology and Biomanufacturing, Tianjin University(天津大学合成生物学与生物制造学院)
CF-VLA: Efficient Coarse-to-Fine Action Generation for Vision-Language-Action Policies
CF-VLA:面向视觉-语言-动作策略的高效粗到细动作生成
Fan Du, Feng Yan, Jianxiong Wu, Xinrun Xu, Weiye Zhang, Weinong Wang, Yu Guo, Bin Qian, Zhihai He, Fei Wang, Heng Yang
机构
*
Southern University of Science and Technology(南方科技大学)
;
Xi’an Jiaotong University(西安交通大学)
;
United Nova Technology(联合Nova技术)
;
University of Science and Technology of China(中国科学技术大学)
Comments18 pages. Accepted at ACM Multimedia 2026. Updated to the final camera-ready version with supplementary material, reproducibility clarifications, and a public source-code repository. Main results and conclusions remain unchanged
机构
*
Nanjing University of Science and Technology(南京理工大学)
;
University of Science and Technology of China(中国科学技术大学)
;
National University of Singapore(新加坡国立大学)
;
Meituan(美团)
Comments12 pages, 4 figures, 11 tables. Accepted at ACM Multimedia 2026 (MM '26), Rio de Janeiro, Brazil. This version includes the supplementary material as Appendix A-E