MedVL-SAM2: A unified 3D medical vision-language model for multimodal reasoning and prompt-driven segmentation
MedVL-SAM2:一种统一的3D医学视觉-语言模型,用于多模态推理和基于提示的分割
Yang Xing, Jiong Wu, Savas Ozdemir, Ying Zhang, Yang Yang, Wei Shao, Kuang Gong
机构
*
Department of Biomedical Engineering, University of Florida(佛罗里达大学生物医学工程系)
;
Department of Radiology, University of Florida(佛罗里达大学放射学系)
;
Research Computing, University of Florida(佛罗里达大学研究计算中心)
;
Department of Medicine, University of Florida(佛罗里达大学医学系)
;
Department of Radiology, UC San Francisco(旧金山大学放射学系)
SVII-3D: Advancing Roadside Infrastructure Inventory with Decimeter-level 3D Localization and Comprehension from Sparse Street Imagery
SVII-3D:利用厘米级3D定位与稀疏街道影像的综合理解,推进道路基础设施库存建设
Chong Liu, Luxuan Fu, Yang Jia, Zhen Dong, Bisheng Yang
机构
*
State Key Laboratory of Information Engineering in Surveying, Mapping and Remote Sensing (LIESMARS), Wuhan University, Wuhan 430079, China(信息工程测绘遥感国家重点实验室(LIESMARS),武汉大学)
;
Research Institute Ltd, Chengdu 610000, China(四川省公路规划设计研究有限公司)