机构
*
Center for Artificial Intelligence and Robotics, Hong Kong Institute of Science & Innovation, Chinese Academy of Sciences, Hong Kong, China(人工智能与机器人中心,香港科学创新研究所,中国科学院,香港,中国)
;
State Key Laboratory of Mathematical Sciences, Academy of Mathematics and Systems Science, Chinese Academy of Sciences, Beijing, China(数学科学国家重点实验室,数学与系统科学研究院,中国科学院,北京,中国)
;
University of Chinese Academy of Sciences, Beijing, China(中国科学院大学,北京,中国)
;
Division of Electronic Engineering, Faculty of Engineering, The Chinese University of Hong Kong, Hong Kong, China(电子工程系,工程学院,香港中文大学,香港,中国)
;
Accident and Emergency Medicine Academic Unit, The Chinese University of Hong Kong, Hong Kong, China(急诊医学学术单位,香港中文大学,香港,中国)
;
Xiangya Hospital Central South University, Changsha, China(湘雅医院,中南大学,长沙,中国)
;
Hunan Frontline Medical Technology Co., Ltd, Changsha, China(湖南前线医疗技术有限公司,长沙,中国)
;
Qilu Hospital of Shandong University, Jinan, China(山东大学齐鲁医院,济南,中国)
;
Zhongshan Hospital of Fudan University, Shanghai, China(复旦大学中山医院,上海,中国)
;
Shanghai Geriatric Medical Center, Shanghai, China(上海老年医学中心,上海,中国)
Leveraging Generic Foundation Models for Multimodal Surgical Data Analysis
Simon Pezold, Jérôme A. Kurylec, Jan S. Liechti, Beat P. Müller, Joël L. Lavanchy
机构
*
Department of Biomedical Engineering, University of Basel, Allschwil, Switzerland(巴塞尔大学生物医学工程系)
;
Clarunis – University Digestive Health Care Center Basel, Basel, Switzerland(Clarunis – 巴塞尔大学消化健康医疗中心)
专题命中
领域大模型
:foundation model(title,abstract)
Comments13 pages, 3 figures; accepted at ML-CDS @ MICCAI 2025, Daejeon, Republic of Korea
MM-DINOv2: Adapting Foundation Models for Multi-Modal Medical Image Analysis
Daniel Scholz, Ayhan Can Erdur, Viktoria Ehm, Anke Meyer-Baese, Jan C. Peeken, Daniel Rueckert, Benedikt Wiestler
机构
*
Chair for AI for Image-Guided Diagnosis and Therapy, Technical University of Munich (TUM)(人工智能辅助影像诊断与治疗研究所,慕尼黑技术大学)
;
TUM University Hospital, Munich, Germany(慕尼黑技术大学医院)
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)
;
Chair for AI in Healthcare and Medicine, Technical University of Munich (TUM)(人工智能在医疗与健康领域研究所,慕尼黑技术大学)
;
Department of Radiation Oncology, TUM University Hospital, Munich, Germany(放射肿瘤科,慕尼黑技术大学医院)
;
Chair for Computer Vision and Artificial Intelligence, Technical University of Munich (TUM)(计算机视觉与人工智能研究所,慕尼黑技术大学)
;
Department of Scientific Computing, Florida State University(科学计算系,佛罗里达州立大学)
;
Deutsches Konsortium für Translationale Krebsforschung (DKTK), Partner Site Munich(德国转化癌症研究联盟(DKTK)慕尼黑分部)
;
Institute of Radiation Medicine (IRM), Department of Radiation Sciences (DRS), Helmholtz Center Munich(放射医学研究所(IRM),辐射科学部门(DRS),海德堡中心慕尼黑)
;
Institute for Advanced Study, Technical University of Munich (TUM)(高级研究所,慕尼黑技术大学)
VLSM-Ensemble: Ensembling CLIP-based Vision-Language Models for Enhanced Medical Image Segmentation
Julia Dietlmeier, Oluwabukola Grace Adegboro, Vayangi Ganepola, Claudia Mazo, Noel E. O'Connor
机构
*
Insight Research Ireland Centre for Data Analytics, DCU, Dublin, Ireland(爱尔兰洞察研究爱尔兰数据分析中心,都柏林大学,都柏林)
;
Research Ireland Centre for Research Training in Machine Learning, DCU, Dublin, Ireland(爱尔兰研究爱尔兰机器学习研究培训中心,都柏林大学,都柏林)
;
School of Computing, Dublin City University, Dublin, Ireland(计算学院,都柏林城市大学,都柏林)
专题命中
领域大模型
:language model(title,abstract)
CommentsMedical Imaging with Deep Learning (MIDL 2025) short paper
Generalist versus Specialist Vision Foundation Models for Ocular Disease and Oculomics
Yukun Zhou, Paul Nderitu, Jocelyn Hui Lin Goh, Justin Engelmann, Siegfried K. Wagner, Anran Ran, Hongyang Jiang, Lie Ju, Ke Zou, Sahana Srinivasan, Hyunmin Kim, Takahiro Ninomiya, Zheyuan Wang, Gabriel Dawei Yang, Eden Ruffell, Dominic Williamson, Rui Santos, Gabor Mark Somfai, Carol Y. Cheung, Tien Yin Wong, Daniel C. Alexander, Yih Chung Tham, Pearse A. Keane
机构
*
College of Information Science and Engineering, Northeastern University, Shenyang, China(信息科学与工程学院,东北大学,沈阳,中国)
;
School of Mechanical Engineering and Automation, Northeastern University, Shenyang, China(机械工程与自动化学院,东北大学,沈阳,中国)
;
Enterprise AI, Midea AI Innovation Center, Foshan, China(企业人工智能,美的人工智能创新中心,佛山,中国)
;
Foshan Graduate School of innovation, Northeastern University, Foshan, China(佛山创新研究生学院,东北大学,佛山,中国)
机构
*
Vanderbilt University(范德比尔特大学)
;
Weill Cornell Medicine(韦尔·科恩医学中心)
;
Vanderbilt University Medical Center(范德比尔特大学医学中心)
;
UT MD Anderson Cancer Center(德克萨斯大学MD安德森癌症中心)
机构
*
National Engineering Laboratory for Integrated Aero-Space-Ground-Ocean Big Data Application Technology(集成空天地海大数据应用技术国家工程实验室)
;
Northwestern Polytechnical University(西北工业大学)
;
Huiying Medical Technology Company Ltd.(慧影医疗技术有限公司)
;
The School of Computer Science(计算机学院)
;
The University of Sydney(悉尼大学)
;
Department of Computer Science and Engineering(计算机科学与工程系)
;
The Chinese University of Hong Kong(香港中文大学)
Your other Left! Vision-Language Models Fail to Identify Relative Positions in Medical Images
Daniel Wolf, Heiko Hillenhagen, Billurvan Taskin, Alex Bäuerle, Meinrad Beer, Michael Götz, Timo Ropinski
机构
*
Visual Computing Group, Institute of Media Informatics, Ulm University, Germany(媒体信息研究所视觉计算组,乌尔姆大学,德国)
;
Diagnostic and Interventional Radiology, Ulm University Medical Center, Germany(乌尔姆大学医学中心诊断与介入放射学)
;
Axiom Bio, USA(Axiom Bio公司,美国)
专题命中
领域大模型
:language model(title,abstract)
CommentsAccepted at the International Conference on Medical Image Computing and Computer Assisted Intervention (MICCAI) 2025
Zero-shot Shape Classification of Nanoparticles in SEM Images using Vision Foundation Models
Freida Barnatan, Emunah Goldstein, Einav Kalimian, Orchen Madar, Avi Huri, David Zitoun, Ya'akov Mandelbaum, Moshe Amitay
机构
*
Department of Bioinformatics, Jerusalem College of Technology(生物信息学系,耶路撒冷技术学院)
;
Bar-Ilan University(巴伊兰大学)
;
Jerusalem College of Technology(耶路撒冷技术学院)
;
Bar-Ilan Institute of Nanotechnology and Advanced Material(巴伊兰纳米技术与先进材料研究所)
;
Department of Electrical Engineering / Department of Electrooptical Engineering and Applied Physics(电气工程系 / 电光学工程与应用物理学系)
Predicting EGFR Mutation in LUAD from Histopathological Whole-Slide Images Using Pretrained Foundation Model and Transfer Learning: An Indian Cohort Study
Learning from SAM: Harnessing a Foundation Model for Sim2Real Adaptation by Regularization
Mayara E. Bonani, Max Schwarz, Sven Behnke
机构
*
Autonomous Intelligent Systems, University of Bonn, Germany(博恩大学自主智能系统)
;
Lamarr Institute for Machine Learning and AI, Germany(机器学习与人工智能拉马尔研究所)
;
Center for Robotics, University of Bonn, Germany(机器人中心)
专题命中
领域大模型
:foundation model(title,abstract)
CommentsAccepted for IEEE International Conference on Automation Science and Engineering (CASE) 2025. M. E. Bonani and M. Schwarz contributed equally
Mind the Modality Gap: Towards a Remote Sensing Vision-Language Model via Cross-modal Alignment
Angelos Zavras, Dimitrios Michail, Begüm Demir, Ioannis Papoutsis
机构
*
organization= Orion Lab, National Observatory of Athens \& National Technical University of Athens , country= Greece
;
organization= Department of Informatics \& Telematics, Harokopio University of Athens , country= Greece
;
organization= Faculty of Electrical Engineering
;
organization= BIFOLD - Berlin Institute for the Foundations of Learning
CommentsAccepted at the ISPRS Journal of Photogrammetry and Remote Sensing. Our code implementation and weights for all experiments are publicly available at https://github.com/Orion-AI-Lab/MindTheModalityGap