RadVLM: A Multitask Conversational Vision-Language Model for Radiology
机构 * Department of Radiology, Kobe University(金泽大学放射科) ; Department of Computer Science, ETH Zurich(苏黎世联邦理工学院计算机科学系) ; Department of Advanced Imaging in Medical Magnetic Resonance, Kyoto University(京都大学医学磁共振高级成像部门) ; Department of Quantitative Biomedicine, University of Zurich(苏黎世大学定量生物医学系) ; Diagnostic and Interventional Radiology, University Hospital Zurich(苏黎世大学医院诊断与介入放射科)
专题命中 视觉定位与Grounding :vision-language model(title,abstract);grounding(abstract);分类 cs.CV、cs.AI
Comments 21 pages, 15 figures