Cross4D-JEPA: Dense Cross-modal Correspondence Distillation for 4D Point Cloud Representation Learning
Cross4D-JEPA: 用于4D点云表示学习的密集跨模态对应蒸馏
Trung Thanh Nguyen, Hai Nguyen-Truong, Tu Vo, Hoang M. Truong, Tuan-Anh Vu
机构
*
Nagoya University(名古屋大学)
;
Northeastern University(东北大学)
;
KC Machine Learning Lab(KC机器学习实验室)
;
University of Science, Vietnam National University Ho Chi Minh City(胡志明市越南国立大学理科大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
机构
*
Zhejiang University(浙江大学)
;
DAMO Academy, Alibaba Group(阿里巴巴集团达摩院)
;
Hupan Lab(华平实验室)
;
Huazhong University of Science and Technology(华中科技大学)
;
East China Normal University(华东师范大学)
;
Shanghai Jiao Tong University(上海交通大学)
An approach with Visual and Tabular Mamba to multimodal medical data using Mixed Fusion
一种基于视觉和表格Mamba的混合融合多模态医疗数据方法
Matheus B. Rocha, Gustavo B. Dettogni, Renato A. Krohling
机构
*
Labcin - Nature Inspired Computing Lab, Federal University of Esp\'irito Santo, Vit\'oria, Brazil PPGI - Graduate Program in Computer Science, Federal University of Esp\'irito Santo, Vit\'oria, Brazil
Plug-and-Adapt: Multimodal Coreference Resolution at First Sight with a Pretrained Alignment Model
即插即适应:基于预训练对齐模型的首眼多模态指代消解
Jinghan Wu, Jing Li, Ivor W. Tsang, Xuetao Zhang
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi'an Jiaotong University(西安交通大学人工智能与机器人研究所人机混合增强智能全国重点实验室)
;
Centre for Frontier AI Research and Institute of High-Performance Computing, Agency for Science, Technology and Research (A*STAR)(新加坡科技研究局前沿人工智能研究中心与高性能计算研究所)
Information-Theoretic Decomposition for Multimodal Interaction Learning
多模态交互学习的信息论分解
Zequn Yang, Yake Wei, Haotian Ni, Zhihao Xu, Di Hu
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China, Beijing(中国人民大学人工智能学院,北京)
;
Beijing Key Laboratory of Research on Large Models(北京大模型研究关键实验室)
;
Engineering Research Center of Next-Generation Intelligent Search(下一代智能搜索与推荐工程研究中心)
;
Beihang University, Beijing(北航,北京)
;
Gaotu Techedu Inc.(高图科技有限公司)
机构
*
Department of Computer Science and Technology, Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)计算机科学与技术学院)
;
School of Intelligence Science and Engineering, College of Artificial Intelligence, Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)人工智能学院智能科学与工程学院)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
Guangdong Key Laboratory of Biomedical Measurements and Ultrasound Imaging, School of Biomedical Engineering, Shenzhen University Medical School, Shenzhen University(深圳大学医学部生物医学工程学院广东省生物医学测量与超声成像重点实验室)
;
Department of Radiology, The People’s Hospital of Guangxi Zhuang Autonomous Region, Guangxi Academy of Medical Sciences(广西壮族自治区人民医院放射科,广西医学科学院)
;
Shenzhen Sixth People’s Hospital (Nanshan Hospital), Huazhong University of Science and Technology Union Shenzhen Hospital(华中科技大学协和深圳医院(深圳市第六人民医院))
;
School of Basic Medical Sciences, Shenzhen University(深圳大学基础医学院)
;
Egypt-Japan University of Science and Technology (E-JUST)(埃及日本科技大学)
;
School of Biomedical Engineering, National-Regional Key Technology Engineering Laboratory for Medical Ultrasound, Guangdong Key Laboratory for Biomedical Measurements and Ultrasound Imaging, Shenzhen University Medical School(深圳大学医学部生物医学工程学院,国家地方联合医学超声关键技术工程实验室,广东省生物医学测量与超声成像重点实验室)
When Two Tracers Disagree: An Investigation of Multimodal Fusion for Clinical PET/CT Segmentation
当两种示踪剂意见不合时:临床PET/CT分割的多模态融合研究
Jack A. Johnson, Bartłomiej W. Papież
机构
*
University of Oxford(牛津大学)
;
Nuffield Department of Medicine(纳菲尔德医学系)
;
Department of Oncology(肿瘤学系)
;
Big Data Institute(大数据研究所)
;
Nuffield Department of Population Health(纳菲尔德人口健康系)
Comments10 pages (8 pages main content and 2 pages of references), 2 figures, 2 tables, accepted to MICCAI 2026 Cancer Prevention, Detection, and IntervenTion (CaPTion) Workshop
机构
*
Center for AI and Data Science (CAIDAS), Julius-Maximilians-Universität Würzburg(维尔茨堡大学人工智能与数据科学中心)
;
Institute for Computational Imaging and AI in Medicine (CompAI), Technical University of Munich(慕尼黑工业大学计算成像与医学人工智能研究所)
;
Pattern Recognition Lab, Friedrich-Alexander Universität Erlangen-Nürnberg(埃尔朗根-纽伦堡大学模式识别实验室)
;
Department Artificial Intelligence in Biomedical Engineering (AIBE), Friedrich-Alexander-Universität Erlangen-Nürnberg(埃尔朗根-纽伦堡大学生物医学工程人工智能系)
GeoUniPR: A Geometry-Consistent Unified Framework for Cross-Modal Place Recognition
GeoUniPR:用于跨模态地点识别的几何一致性统一框架
Wonbong Kim, Jiatong Xiao, Rui Li, Xufei Wang, Qiwen Gu, Junqiao Zhao, Chen Ye, Guang Chen
机构
*
School of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院)
;
Shanghai Research Institute for Intelligent Autonomous System, Tongji University(同济大学上海智能自主系统研究院)
;
Shanghai Innovation Institute(上海创新研究院)
AI-driven Multimodal Representation Learning for Latent Mediation Structure Discovery of Socioeconomic Disadvantage, Psychosocial Factors, and Cardiometabolic Multimorbidity: Insights from the All of Us Research Program
人工智能驱动的多模态表征学习用于发现社会经济劣势、心理社会因素与心脏代谢共病的潜在中介结构:来自All of Us研究计划的见解
机构
*
The City College of New York(纽约城市学院)
;
INSAIT
;
Sofia University “St. Kliment Ohridski”(圣克莱门特奥赫里德斯基索非亚大学)
;
Julius-Maximilians-Universität Würzburg(维尔茨堡大学)
;
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所人工智能安全国家重点实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Shenzhen University of Advanced Technology(深圳先进技术研究院)