机构
*
School of Mechanical Engineering, University of Science and Technology Beijing(北京科技大学机械工程学院)
;
The Laboratory for Computational Sensing and Robotics, Johns Hopkins University(约翰霍普金斯大学计算传感与机器人实验室)
Comments9 pages, 2 figures, system description paper for the CHiPSAL 2026 shared task at LREC 2026
Journal refProceedings of the Second Workshop on Challenges in Processing South Asian Languages (CHiPSAL 2026) @ LREC 2026, pages 275-283, Palma, Mallorca, Spain, 16 May 2026. ELRA Language Resources Association (ELRA). ISBN 978-2-493814-66-1
A report-grounded vision-language foundation model for colonoscopy from 280000 routine reports
基于28万份常规报告的肠镜报告驱动的视觉-语言基础模型
Jia Yu, Yan Zhu, Yili He, Zilong Wang, Xinyang Jiang, Peiyao Fu, Ruijie Yang, Tianyi Chen, Siyuan Li, Zhihua Wang, Fei Wu, Quanlin Li, Xian Yang, Pinghong Zhou, Shuo Wang
机构
*
Digital Medical Research Center, School of Basic Medical Sciences, Fudan University(复旦大学基础医学院数字医学研究中心)
;
Shanghai Collaborative Innovation Center of Endoscopy(上海内镜诊疗协同创新中心)
;
Zhejiang University(浙江大学)
;
Shanghai Institute for Advanced Study of Zhejiang University(浙江大学上海高等研究院)
;
Alliance Manchester Business School, The University of Manchester(曼彻斯特大学联盟曼彻斯特商学院)
;
Data Science Institute, Imperial College London(伦敦帝国理工学院数据科学研究所)
;
Microsoft Research Asia(微软亚洲研究院)
Localization-Infused Vision-Language Semantic Fusion for Text-Guided Medical Image Segmentation
用于文本引导医学图像分割的定位注入视觉语言语义融合
Songyue Han, Mingye Zou, Shuchang Ye, Lei Bi, Mingyuan Meng
机构
*
Air Force Engineering University(空军工程大学)
;
Harbin Institute of Technology(哈尔滨工业大学)
;
University of Sydney(悉尼大学)
;
Institute of Translational Medicine, Shanghai Jiao Tong University(上海交通大学转化医学研究院)
;
Zhongguancun Institute of Artificial Intelligence(中关村人工智能研究院)