机构
*
School of Airspace and Engineering, Shandong University(山东大学航空航天与工程学院)
;
School of Computer Science and Technology, Shandong University of Finance and Economics(山东财经大学计算机科学与技术学院)
;
School of Psychology and Neuroscience, University of Glasgow(格拉斯哥大学心理学与神经科学学院)
;
Faculty of Information Science and Engineering, Ocean University of China(中国海洋大学信息科学与工程学院)
Naming the Concepts Classifiers Rely On: Language-Anchored Decomposition for Faithful Explanation
命名概念分类器所依赖的内容:用于忠实解释的语言锚定分解
Ahsan Habib Akash, Dipkamal Bhusal, Stacey Jones, Donald A. Adjeroh, Binod Bhattarai, Prashnna Kumar Gyawali
机构
*
West Virginia University(西弗吉尼亚大学)
;
Rochester Institute of Technology(罗彻斯特理工学院)
;
O Analytics(O分析公司)
;
University of Aberdeen(阿伯丁大学)
;
Fogsphere (Redev.AI Ltd, UK)(Fogsphere(Redev.AI有限公司,英国))
;
University College London(伦敦大学学院)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract)
Jolia: Concept-Level Vision-Language Alignment for 3D CT Contrastive Learning
Jolia: 用于3D CT对比学习的概念级视觉-语言对齐
Julien Khlaut, Charles Corbière, Baptiste Callard, Amaury Prat, Leo Butsanets, Antoine Saporta, Théo Danielou, Leo Machado, Korentin Le Floch, Tom Boeken, Pierre Manceron, Corentin Dancette
机构
*
Raidium
;
Department of Vascular and Oncological Interventional Radiology, Hôpital Européen Georges Pompidou, AP-HP(欧洲乔治·蓬皮杜医院血管与肿瘤介入放射科,AP-HP)
;
Faculté de Santé, Université Paris-Cité(巴黎西岱大学健康学院)
;
HEKA, INRIA(HEKA,法国国家信息与自动化研究所)
;
Imaging Department, Fondation Ophtalmologique Adolphe de Rothschild(阿道夫·罗斯柴尔德眼科基金会影像科)
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
State Key Lab of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences, Beijing, China(中国科学院人工智能安全国家重点实验室,计算技术研究所,北京,中国)
;
Harbin Institute of Technology (Weihai)(哈尔滨工业大学(威海))
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract)
PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scene Understanding
PAR3D: 一种用于场景理解的统一部件感知3D多模态大语言模型
Shaohui Dai, Yansong Qu, You Shen, Shengchuan Zhang, Liujuan Cao
机构
*
Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(教育部多媒体可信感知与高效计算重点实验室,厦门大学)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract)
Improving Semantic Uncertainty Quantification in LVLMs with Semantic Gaussian Processes
利用语义高斯过程改进LVLM中的语义不确定性量化
Joseph Hoche, Andrei Bursuc, David Brellmann, Gilles Louppe, Pavel Izmailov, Angela Yao, Gianni Franchi
机构
*
AMIAD, Pôle Recherche, Palaiseau(AMIAD研究部,Palaiseau)
;
valeo.ai
;
Safran Tech
;
University of Liège(利耶大学)
;
New York University(纽约大学)
;
National University of Singapore(新加坡国立大学)
;
ENSTA Paris(巴黎ENSTA)
MM-Snowball: Evaluating and Mitigating Hallucination Snowballing in Multimodal Multi-Turn Dialogue
MM-Snowball:多模态多轮对话中的幻觉雪崩评估与缓解
Yue Jiang, Xue Jiang, Lihua Zhang, Zhiqiang Wang, Yuhang Lu, Peng Wang, Bo Han, Feng Zheng, Dingkang Yang
机构
*
College of Intelligent Robotics and Advanced Manufacturing, Fudan University(复旦大学智能机器人与先进制造学院)
;
Southern University of Science and Technology(南方科技大学)
;
TMLR Group, Hong Kong Baptist University(香港 Baptist 大学 TMLR 团体)
;
MM Lab, CUHK(CUHK 多模态实验室)
;
RAMS Lab, Huawei Technologies Co., Ltd.(华为技术有限公司 RAMS 实验室)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract)
COMET: Concept Space Dissection of the Modality Gap in Audio-Text Multimodal Contrastive Embeddings
COMET:音频-文本多模态对比嵌入中模态间隙的概念空间剖析
Yonggang Zhu, Liting Gao, Aidong Men, Wenwu Wang
机构
*
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
Centre for Vision, Speech, and Signal Processing (CVSSP), University of Surrey(Surrey 大学视觉、语音和信号处理中心)
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
;
Shandong University(山东大学)
;
Tongji University(同济大学)
;
City University of Hong Kong(香港城市大学)
机构
*
Department of Data Science \& AI, Faculty of Information Technology, Monash University, Melbourne, VIC 3800, Australia Alfred Health Radiology, Alfred Health, Melbourne, VIC 3004, Australia School of Translational Medicine, Faculty of Medicine, Nursing
;
Health Sciences, Monash University, Melbourne, VIC 3800, Australia Hong Kong Polytechnic University, Hong Kong SAR, China
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract)
CommentsEarly accepted by MICCAI 2026. This version of the contribution has been accepted for publication, after peer review (when applicable) but is not the Version of Record and does not reflect post-acceptance improvements, or any corrections