Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images
医学视觉语言模型 HuluMed 和 MedGemma 以及通用聊天机器人 Gemma 3、ChatGPT Plus 和 Claude Pro 在真实未见伤口图像上的评估
Yunzhe Xue, Mohammed Saim Ahmed Quadri, Neal Panse, Justin W. Ady, Usman Roshan
机构
*
Department of Computer Science, New Jersey Institute of Technology(新泽西理工学院计算机科学系)
;
Vascular and Endovascular Surgery, Robert Wood Johnson Hospital(罗伯特·伍德·约翰逊医院血管外科)
;
Department of Data Science, New Jersey Institute of Technology(新泽西理工学院数据科学系)
专题命中
领域大模型
:language model(title,abstract)
AI总结
本研究评估了六种视觉语言模型在慢性伤口分析任务上的表现,发现通用模型 ChatGPT 和 Claude 显著优于医学专用模型,表明广泛的多模态推理能力比领域知识更重要。
SAVER: Mitigating Hallucinations in Large Vision-Language Models via Style-Aware Visual Early Revision
SAVER:通过风格感知视觉早期修正减轻大型视觉语言模型中的幻觉
Zhaoxu Li, Chenqi Kong, Yi Yu, Qiangqiang Wu, Xinghao Jiang, Ngai-Man Cheung, Bihan Wen, Alex Kot, Xudong Jiang
机构
*
ROSE Lab, Interdisciplinary Graduate Programme, Nanyang Technological University, Singapore(南洋理工大学罗思实验室,跨学科研究生项目,新加坡)
;
ROSE Lab, School of Electrical and Electronic Engineering, Nanyang Technological University, Singapore(南洋理工大学罗思实验室,电子与电气工程学院,新加坡)
;
City University of Hong Kong, Hong Kong SAR(香港城市大学,香港特别行政区)
;
Shanghai Jiao Tong University, China(上海交通大学,中国)
;
Singapore University of Technology and Design, Singapore(新加坡科技设计大学,新加坡)
;
VinUniversity, Hanoi, Vietnam(越南文大学,河内,越南)
机构
*
Corporate Research, Robert Bosch GmbH(罗伯特·博世有限公司企业研究部)
;
Otto-von-Guericke-University Magdeburg(马格德堡奥托·冯·格里克大学)
;
University of Southampton(南安普顿大学)
CheXanatomy: Anatomy-Aware Vision-Language Modeling for Chest Radiographs
CheXanatomy: 面向胸部X光片的解剖感知视觉-语言建模
Sergios Gatidis, Curtis Langlotz, Christian Bluethgen
机构
*
Stanford Center for Artificial Intelligence in Medicine and Imaging, Stanford University(斯坦福大学医学与影像人工智能中心)
;
Department of Radiology, Stanford University(斯坦福大学放射学系)
$μ$Match: Foundation Models for Semi-supervised Learning and Domain Adaptation in EM
$\mu$Match:电子显微镜中半监督学习和领域适应的基础模型
Marei Freitag, Olesia Korchevaia, Luca Freckmann, Anwai Archit, Constantin Pape
机构
*
Life and Medical Sciences Institute (LIMES), University of Bonn, Germany(波恩大学生命与医学科学研究所(LIMES))
;
Institute of Computer Science, Georg-August-University Göttingen, Germany(哥廷根大学计算机科学研究所)
Predicting Immune Biomarkers with MultiModal Mixture-of-Expert Pathology Foundation Models Empowers Precision Oncology
使用多模态混合专家病理基础模型预测免疫生物标志物,赋能精准肿瘤学
Tianyu Liu, Ziqing Wang, Zhaokang Liang, Tong Ding, Peter Humphrey, Lorraine Colón-Cartagena, Emily Ling-Lin Pai, Kenneth Tou En Chang, Mohamed Kahila, Jonathan Chong Kai Liew, Tinglin Huang, Rex Ying, Kaize Ding, Faisal Mahmood, Wengong Jin
机构
*
Program of Computational Biology and Bioinforamtics, Yale University(耶鲁大学计算生物学与生物信息学项目)
;
Broad Institute of MIT and Harvard(麻省理工学院与哈佛大学博德研究所)
;
Department of Statistics and Data Science, Northwestern University(西北大学统计与数据科学系)
;
Department of Computer Science, Northeastern University(东北大学计算机科学系)
;
Department of Computer Science, Harvard University(哈佛大学计算机科学系)
;
Department of Pathology, Yale University(耶鲁大学病理学系)
;
Department of Anatomic Pathology and Laboratory Medicine, Hospital of the University of Pennsylvania(宾夕法尼亚大学医院解剖病理学与检验医学系)
;
Department of Pathology and Laboratory Medicine, University of California, San Francisco(加州大学旧金山分校病理学与检验医学系)
;
Department of Pathology and Laboratory Medicine, KK Women’s and Children’s Hospital(竹脚妇幼医院病理学与检验医学系)
;
Department of Biostatistics, Epidemiology and Informatics, Perelman School of Medicine, University of Pennsylvania(宾夕法尼亚大学佩雷尔曼医学院生物统计学、流行病学与信息学系)
Revisiting 2D Foundation Models for Scalable 3D Medical Image Classification
重新审视用于可扩展3D医学图像分类的2D基础模型
Han Liu, Bogdan Georgescu, Yanbo Zhang, Youngjin Yoo, Michael Baumgartner, Riqiang Gao, Jianing Wang, Gengyan Zhao, Eli Gibson, Dorin Comaniciu, Sasa Grbic
机构
*
Digital Technology and Innovation, Siemens Healthineers, Princeton NJ, USA(西门子医疗数字技术与创新,普林斯顿新泽西州,美国)
;
Digital Technology and Innovation, Siemens Healthineers, Erlangen, Germany(西门子医疗数字技术与创新,埃尔兰根,德国)
MedFM-Robust: Benchmarking Robustness of Medical Foundation Models
MedFM-Robust:医学基础模型的鲁棒性基准测试
Xiangxiang Cui, Tianjin Huang, Yifang Wang, Lijie Hu, Lu Yin
机构
*
Beijing Normal University, China(北京师范大学)
;
University of Exeter, United Kingdom(埃克塞特大学)
;
University College London, United Kingdom(伦敦大学学院)
;
Mohamed bin Zayed University of Artificial Intelligence, United Arab Emirates(穆罕默德·本·扎耶德人工智能大学)
;
University of Surrey, United Kingdom(萨里大学)
机构
*
School of Computer Science and Technique(计算机科学与技术学院)
;
Tongji University(同济大学)
;
School of Communications and Electronic Engineering(通讯与电子工程学院)
;
East China Normal University(华东师范大学)
Med-StepBench: A Hierarchical Reasoning Framework for Evaluating Hallucinations in Medical Vision-Language Models
Med-StepBench:一种用于评估医学视觉-语言模型幻觉的分层推理框架
Minh Khoi Nguyen, Dai Lam Le, Amir Reza Jafari, Tuan Dung Nguyen, Mai Hong Son, Mai Huy Thong, Quang Huy Nguyen, Thanh Trung Nguyen, Reza Farahbakhsh, Noel Crespi, Phi Le Nguyen
机构
*
AI4LIFE, Hanoi University of Science and Technology, Vietnam(AI4LIFE,越南科学与技术大学)
;
SAMOVAR, Télécom SudParis, Institut Polytechnique de Paris, France(SAMOVAR,法国电信南巴黎学院,巴黎理工学院)
;
Military Central Hospital, Vietnam(越南108军中心医院)
机构
*
Department of Computer Science and Engineering, The Hong Kong University of Science and Technology, Hong Kong SAR, China(香港科技大学计算机科学与工程系)
;
Department of Pathology, Nanfang Hospital, School of Basic Medical Sciences, Southern Medical University, Guangzhou, China(南方医科大学基础医学学院病理学系,广州医院)
;
Department of Pathology, The First Affiliated Hospital, School of Medicine, Zhejiang University, Hangzhou, China(浙江大学医学院第一附属医院病理学系,杭州)
;
State Key Laboratory of Holistic Integrative Management of Gastrointestinal Cancers, Department of Pathology, School of Basic Medicine and Xijing Hospital, Fourth Military Medical University, Xi’an, China(胃肠癌整体整合管理国家重点实验室,第四军医大学基础医学学院病理学系,西京医院)
机构
*
School of Biomedical Engineering (Suzhou), Division of Life Science and Medicine, University of Science and Technology of China, Hefei, China(生物医学工程学院(苏州),生命科学与医学系,中国科学技术大学,合肥,中国)
;
Suzhou Institute of Biomedical Engineering and Technology, Chinese Academy of Sciences, Suzhou, China(苏州生物医学工程与技术研究所,中国科学院,苏州,中国)
;
Shanghai Innovation Institute, Shanghai, China(上海创新研究院,上海,中国)
;
Medical School of Tianjin University, Tianjin, China(天津大学医学院,天津,中国)
;
Jinan Guoke Medical and Technology Development Co., Ltd., Pharmaceutical Valley New Drug Creation Platform, Jinan, China(济南国科医药科技发展有限公司,药谷新药创制平台,济南,中国)
Plug-and-play Class-aware Knowledge Injection for Prompt Learning with Visual-Language Model
即插即用的类感知知识注入用于具有视觉-语言模型的提示学习
Junhui Yin, Nan Pu, Xinyu Zhang, Lingfeng Yang, Lin Wu, Xiaojie Wang, Zhun Zhong
机构
*
University of Science and Technology Beijing(北京科技大学)
;
Hefei University of Technology(合肥工业大学)
;
University of Auckland(奥克兰大学)
;
Nanjing University of Science and Technology(南京理工大学)
;
Swansea University(斯旺西大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
Vision Foundation Models for Domain Generalisable Cross-View Localisation in Planetary Ground-Aerial Robotic Teams
行星地面-空中机器人团队中的域可推广跨视图局部化视觉基础模型
Lachlan Holden, Feras Dayoub, Alberto Candela, David Harvey, Tat-Jun Chin
机构
*
AI for Space Group and 3 Andy Thomas Centre for Space Resources, The University of Adelaide(AI空间组和安迪·托马斯太空资源中心,阿德莱德大学)
;
Jet Propulsion Laboratory, California Institute of Technology, Pasadena, CA 91109, USA(喷气推进实验室,加州理工学院,帕萨迪纳,CA 91109,美国)
;
California Institute of Technology(加州理工学院)
Comments7 pages, 10 figures. Presented at the International Conference on Space Robotics (iSpaRo) 2025 in Sendai, Japan. Dataset available: https://doi.org/10.5281/zenodo.17364038
机构
*
Artificial Intelligence Department, Faculty of Computer Engineering, University of Isfahan, Isfahan, Iran
;
Department of Computer Engineering Techniques, Mazaya University College, Nasiriyah, Iraq
;
Department of Medicine, Danube Private University, Krems an der Donau, Austria
;
Austrian Center for Medical Innovation
;
Research Center for Medical Image Analysis
;
Artificial Intelligence, Department of Medicine, Danube Private University, Krems an der Donau, Austria