Measuring and Aligning Abstraction in Vision-Language Models with Medical Taxonomies
利用医学分类法测量和对齐视觉-语言模型中的抽象能力
机构 * 1 School of Computation, Information ; Technology, Technical University of Munich, Germany 2 Institute of Machine Learning in Biomedical Imaging, Helmholtz Munich, Germany 3 LTCI, Télécom Paris, Institut Polytechnique de Paris, France 4 Munich Center for Machine Learning (MCML) 5 School of Biomedical Engineering ; Imaging Sciences, King's College London, UK 6 Department of Artificial Intelligence in Biomedical Imaging, FAU Erlangen-Nuremberg, Germany 7 Department of Computing, Imperial College London, UK
专题命中 VLM训练与架构 :vision-language model(title,abstract);分类 cs.AI
AI总结 本文提出通过医学分类法量化和缓解视觉-语言模型中的抽象错误,引入灾难性抽象错误概念,并通过风险约束阈值和分类法感知微调减少严重错误至2%以下。