Frequency-domain Multi-modal Fusion for Language-guided Medical Image Segmentation
机构 * School of Computer Science and Technology, Anhui University(安徽大学计算机科学与技术学院) ; NLPR, MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院) ; School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) ; School of Artificial Intelligence, Anhui University(安徽大学人工智能学院)
专题命中 VLA模型 :action model(abstract);分类 cs.CV
Comments Accepted by MICCAI 2025