机构
*
School of Biomedical Engineering, Tsinghua University(清华大学生物医学工程学院)
;
Department of Radiology, Mianyang Central Hospital(绵阳市中心医院放射科)
;
Center for Biomedical Imaging Research, Tsinghua University(清华大学生物医学成像研究中心)
From Adaptation to Generalization: Adaptive Visual Prompting for Medical Image Segmentation
从适应到泛化:面向医学图像分割的自适应视觉提示
Evren Çetinkaya, Sangmin Lee, Jung Uk Kim, Hong Joo Lee, Nassir Navab
机构
*
Technical University of Munich(慕尼黑技术大学)
;
Korea University(韩国大学)
;
Kyung Hee University(庆熙大学)
;
Seoul National University of Science and Technology(首尔科学技术大学)
机构
*
School of Instrument Science and Engineering, Southeast University, State Key Lab of Comprehensive PNT Network and Equipment Technology, Key Lab of Micro-Inertial Instrument and Advanced Navigation Technology, MOE(仪器科学与工程学院,东南大学,综合PNT网络与设备技术国家重点实验室,微惯性仪器与先进导航技术重点实验室,教育部)
;
Department of Computer Science and Engineering, University of Bologna(计算机科学与工程系,博洛尼亚大学)
;
College of Computer Science and Software Engineering, Hohai University(计算机科学与软件工程学院,河海大学)
;
Beijing Institute of Technology(北京理工大学)
Towards Responsible Multimodal Medical Reasoning via Context-Aligned Vision-Language Models
通过上下文对齐的视觉-语言模型实现负责任的多模态医疗推理
Sumra Khan, Sagar Chhabriya, Aizan Zafar, Sheeraz Arif, Amgad Muneer, Anas Zafar, Shaina Raza, Rizwan Qureshi
机构
*
Salim Habib University(萨利姆·哈比卜大学)
;
Institute of Business Administration Sukkur(苏库尔工商管理学院)
;
University of Central Florida(中佛罗里达大学)
;
The University of Texas MD Anderson Cancer Center(德克萨斯大学MD安德森癌症中心)
;
Toronto Metropolitan University(多伦多都会大学)
;
Vector Institute(向量研究所)
RVLM: Recursive Vision-Language Models with Adaptive Depth
具有自适应深度的递归视觉-语言模型
Nicanor Mayumu, Zeenath Khan, Melodena Stephens, Patrick Mukala, Farhad Oroumchian
机构
*
Department of Computer Science(计算机科学系)
;
University of Wollongong in Dubai(迪拜沃林戈大学)
;
Dubai Knowledge Park(迪拜知识园区)
;
Mohammed Bin Rashid School of Government(穆罕默德·本·拉希德政府学院)
机构
*
Department of Biomedical Engineering, The Chinese University of Hong Kong(生物医学工程系,香港中文大学)
;
Department of Radiation Oncology, Columbia University Irving Medical Center and Data Science Institute, Columbia University(放射肿瘤学系,哥伦比亚大学伊万杰琳医学中心及数据科学研究院,哥伦比亚大学)
;
Department of Electrical and Electronic Engineering and School of Biomedical Engineering, The University of Hong Kong(电气电子工程系和生物医学工程学院,香港大学)
;
Department of Computer Science, City University of Hong Kong(计算机科学系,城市大学)
;
Department of Artificial Intelligence, Westlake University(人工智能系,西湖大学)
GeoDiT: A Diffusion-based Vision-Language Model for Geospatial Understanding
GeoDiT:一种基于扩散的视觉-语言模型,用于地理空间理解
Jiaqi Liu, Ronghao Fu, Haoran Liu, Lang Sun, Bo Yang
机构
*
College of Computer Science and Technology, Jilin University, Changchun 130012, China(吉林大学计算机科学与技术学院,长春 130012,中国)
;
Key Laboratory of Symbolic Computation and Knowledge Engineering of Ministry of Education(教育部符号计算与知识工程重点实验室)
Vision Language Model-based Testing of Industrial Autonomous Mobile Robots
基于视觉语言模型的工业自主移动机器人测试
Jiahui Wu, Chengjie Lu, Aitor Arrieta, Shaukat Ali, Thomas Peyrucain
机构
*
Simula Research Laboratory and University of Oslo(Simula研究实验室和奥斯陆大学)
;
Mondragon University(蒙dragon大学)
;
Simula Research Laboratory(Simula研究实验室)
;
PAL Robotics(PAL机器人技术)
ProFound: A moderate-sized vision foundation model for multi-task prostate imaging
ProFound:一种中等规模的视觉基础模型用于多任务前列腺成像
Yipei Wang, Yinsong Xu, Weixi Yi, Shaheer Ullah Saeed, Natasha Thorley, Alexander Ng, Yukun Zhou, Wen Yan, Dean Barratt, Shonit Punwani, Veeru Kasivisvanathan, Mark Emberton, Daniel C. Alexander, Yipeng Hu
机构
*
UCL Hawkes Institute, University College London, London, UK(伦敦大学霍克斯研究所,大学学院伦敦)
;
University College London, London, UK(大学学院伦敦)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications, Beijing, China(北京邮电大学人工智能学院)
;
Centre for Bioengineering, Queen Mary University of London, London, UK(伦敦女王学院生物工程中心)
;
School of Engineering and Materials Science, Queen Mary University of London, London, UK(伦敦女王学院工程与材料科学学院)
;
Digital Environment Research Institute, Queen Mary University of London, London, UK(伦敦女王学院数字环境研究所)
;
Centre for Medical Imaging, University College London, London, UK(伦敦大学医学成像中心)
;
Department of Radiology, University College London Hospital NHS Foundation Trust, London, UK(大学学院伦敦医院 NHS 基础信托放射科)
;
Centre for Urology Imaging, Prostate, AI and Surgical Studies (COMPASS) Research Group, Division of Surgery and Interventional Science, University College London, London, UK(泌尿科成像中心、前列腺、AI 和手术研究组,手术与介入科学系,大学学院伦敦)
;
Institute of Ophthalmology, University College London, London, UK(伦敦大学眼科研究所)
;
Department of Urology, University College London Hospital, London, UK(大学学院伦敦医院泌尿科)
;
Division of Surgery and Interventional Science, University College London, London, UK(手术与介入科学系,大学学院伦敦)
;
Department of Urology, Comprehensive Cancer Center, Medical University of Vienna, Vienna, Austria(维也纳医学大学综合癌症中心泌尿科)
;
Department of Computer Science, University College London, London, UK(计算机科学系,大学学院伦敦)
AgenticTagger: Structured Item Representation for Recommendation with LLM Agents
AgenticTagger: 基于LLM代理的结构化物品表示用于推荐
Zhouhang Xie, Bo Peng, Zhankui He, Ziqi Chen, Alice Han, Isabella Ye, Benjamin Coleman, Noveen Sachdeva, Fernando Pereira, Julian McAuley, Wang-Cheng Kang, Derek Zhiyuan Cheng, Beidou Wang, Randolph Brown