Mammo-CLIP Dissect: A Framework for Analysing Mammography Concepts in Vision-Language Models
Suaiba Amina Salahuddin, Teresa Dorszewski, Marit Almenning Martiniussen, Tone Hovda, Antonio Portaluri, Solveig Thrun, Michael Kampffmeyer, Elisabeth Wetzer, Kristoffer Wickstrøm, Robert Jenssen
机构
*
UiT The Arctic University of Norway(乌塔大学极地大学)
;
Technical University of Denmark(技术大学)
;
Østfold Hospital Trust(奥斯fold医院信托)
;
Vestre Viken Hospital Trust(维斯特维肯医院信托)
;
Radboud University Nijmegen Medical Centre(拉德堡德大学奈梅亨医疗中心)
;
The Netherlands Cancer Institute(荷兰癌症研究所)
;
Antoni van Leeuwenhoek Hospital(安东尼·弗莱明医院)
;
University of Copenhagen(哥本哈根大学)
Large Vision-Language Model Alignment and Misalignment: A Survey Through the Lens of Explainability
Dong Shu, Haiyan Zhao, Jingyu Hu, Weiru Liu, Ali Payani, Lu Cheng, Mengnan Du
机构
*
Northwestern University(西北大学)
;
New Jersey Institute of Technology(新泽西理工学院)
;
University of Bristol(布里斯托大学)
;
Cisco Research(思科研究)
;
University of Illinois Chicago(伊利诺伊大学芝加哥分校)
机构
*
University of Trento(特伦托大学)
;
University of Bergamo(贝拉姆奥大学)
;
Indian Institute of Technology Bombay(印度班加罗尔理工学院)
;
LNMIIT Jaipur(斋普尔LNMIIT)
;
The University of Queensland(昆士兰大学)
;
Fondazione Bruno Kessler(布鲁诺·凯塞勒基金会)
机构
*
National Engineering Laboratory for Integrated Aero-Space-Ground-Ocean Big Data Application Technology(集成空天地海大数据应用技术国家工程实验室)
;
Northwestern Polytechnical University(西北工业大学)
;
Huiying Medical Technology Company Ltd.(慧影医疗技术有限公司)
;
The School of Computer Science(计算机学院)
;
The University of Sydney(悉尼大学)
;
Department of Computer Science and Engineering(计算机科学与工程系)
;
The Chinese University of Hong Kong(香港中文大学)
Fine-Grained VLM Fine-tuning via Latent Hierarchical Adapter Learning
Yumiao Zhao, Bo Jiang, Yuhe Ding, Xiao Wang, Jin Tang, Bin Luo
机构
*
Information Materials and Intelligent Sensing Laboratory of Anhui Province(安徽省信息材料与智能感知实验室)
;
Anhui Provincial Key Laboratory of Multimodal Cognitive Computation(安徽省多模态认知计算重点实验室)
;
School of Computer Science and Technology, Anhui University(安徽大学计算机科学与技术学院)
Vision Language Models Know Law of Conservation without Understanding More-or-Less
Dezhi Luo, Haiyun Lyu, Qingying Gao, Haoran Sun, Yijiang Li, Hokin Deng
机构
*
University of Michigan(密歇根大学)
;
University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
;
Johns Hopkins University(约翰霍普金斯大学)
;
University of California, San Diego(加州大学圣地亚哥分校)
;
Carnegie Mellon University(卡内基梅隆大学)
专题命中
VLM训练与架构
:vision language model(title,abstract);分类 cs.AI
CommentsPublished at the ICLR 2025 Workshop on Bidirectional Human-AI Alignment (BiAlign)
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
Dept. of Comp. Sci. and Tech., Institute for AI, Tsinghua University(计算机科学与技术系,人工智能研究院,清华大学)
;
Gaoling School of Artificial Intelligence, Renmin University of China(人工智能学院,中国人民大学)
;
Kuaishou Technology Inc.(快手科技有限公司)
;
Shanghai Key Laboratory of Multi. Info. Processing, East China Normal University(多信息处理重点实验室,华东师范大学)
Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models
Hyundong Jin, Hyung Jin Chang, Eunwoo Kim
机构
*
School of Computer Science and Engineering, Chung-Ang University(Chung-Ang 大学计算机科学与工程学院)
;
School of Computer Science, University of Birmingham(布拉德福德大学计算机科学学院)