Hallucination Filtering in Radiology Vision-Language Models Using Discrete Semantic Entropy
在放射学视觉-语言模型中使用离散语义熵过滤幻觉
Patrick Wienholt, Sophie Caselitz, Robert Siepmann, Philipp Bruners, Keno Bressem, Christiane Kuhl, Jakob Nikolas Kather, Sven Nebelung, Daniel Truhn
机构
*
Lab for Artificial Intelligence in Medicine, Department of Diagnostic and Interventional Radiology, University Hospital RWTH Aachen(医学人工智能实验室,诊断与介入放射学部,RWTH亚琛大学医院)
;
Department of Diagnostic and Interventional Radiology, University Hospital RWTH Aachen(诊断与介入放射学部,RWTH亚琛大学医院)
;
Department of Diagnostic and Interventional Radiology, Technical University of Munich, School of Medicine and Health, Klinikum rechts der Isar, TUM University Hospital(诊断与介入放射学部,慕尼黑技术大学,医学院与健康学院,Klinikum rechts der Isar,TUM大学医院)
;
Department of Cardiovascular Radiology and Nuclear Medicine, Technical University of Munich, School of Medicine and Health, German Heart Center, TUM University Hospital(心血管放射学与核医学部,慕尼黑技术大学,医学院与健康学院,德国心脏中心,TUM大学医院)
;
Else Kroener Fresenius Center for Digital Health, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD Dresden University of Technology(数字健康中心,医学院与卡尔·古斯塔夫·卡尔斯大学医院,德累斯顿技术大学)
;
Department of Medicine I, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD Dresden University of Technology(第一医学部,医学院与卡尔·古斯塔夫·卡尔斯大学医院,德累斯顿技术大学)
;
Pathology & Data Analytics, Leeds Institute of Medical Research at St James’s, University of Leeds(病理学与数据分析,圣詹姆斯医院医学研究所,利兹大学)
;
Medical Oncology, National Center for Tumor Diseases (NCT), University Hospital Heidelberg(医学肿瘤学,国家肿瘤疾病中心(NCT),海德堡大学医院)
MINAR: Mechanistic Interpretability for Neural Algorithmic Reasoning
MINAR: 图神经网络中神经算法推理的机制可解释性
Jesse He, Helen Jenne, Max Vargas, Davis Brown, Gal Mishne, Yusu Wang, Henry Kvinge
机构
*
Pacific Northwest National Laboratory, Richland, WA(太平洋西北国家实验室)
;
Halıcıoğlu Data Science Institute, University of California, San Diego, San Diego, CA(哈利奇奥格鲁数据科学研究所,加州大学圣地亚哥分校)
;
Department of Computer and Information Science, University of Pennsylvania, Pennsylvaina, PA(计算机与信息科学系,宾夕法尼亚大学)
;
Department of Mathematics, University of Washington, Seattle, WA(数学系,华盛顿大学)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Probabilistic distances-based hallucination detection in LLMs with RAG
基于概率距离的LLM中幻觉检测方法(RAG)
Rodion Oblovatny, Alexandra Kuleshova, Konstantin Polev, Alexey Zaytsev
机构
*
Markov Lab, Department of Mathematics(马尔可夫实验室,数学系)
;
Computer Science, Saint-Petersburg University(计算机科学,圣彼得堡大学)
;
AI Center, Skoltech(人工智能中心,斯克里普丘克技术学院)
;
SB AI Lab(SB人工智能实验室)
;
AI Center, Skoltech, Risk department, Sber(人工智能中心,斯克里普丘克技术学院,风险部门)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
CommentsUpdated approach to constructing a hallucination detection score. Added results from experiments with the NLI task. The approach with trainable deep kernels has been removed, with a focus on the unsupervised approach
Learning What Matters: Prioritized Concept Learning via Relative Error-driven Sample Selection
学习关键要素:通过相对误差驱动的样本选择进行优先概念学习
Shivam Chandhok, Qian Yang, Oscar Manas, Kanishk Jain, Leonid Sigal, Aishwarya Agrawal
机构
*
Mila - Québec AI Institute(魁北克人工智能研究所)
;
University of British Columbia(不列颠哥伦比亚大学)
;
Université de Montréal(蒙特利尔大学)
;
Vector Institute for AI(人工智能矢量研究所)
;
Canada CIFAR AI Chair(加拿大CIFAR人工智能主席)
DynamicGTR: Leveraging Graph Topology Representation Preferences to Boost VLM Capabilities on Graph QAs
DynamicGTR: 利用图拓扑表示偏好提升视觉语言模型在图问答中的能力
Yanbin Wei, Jiangyue Yan, Chun Kang, Yang Chen, Hua Liu, James Kwok, Yu Zhang
机构
*
Southern University of Science and Technology(南方科技大学)
;
Hong Kong University of Science and Technology(香港理工大学)
;
Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))
;
Beihang University(北京航空航天大学)
机构
*
Vision and AI Lab, Indian Institute of Science, Bangalore, India(印度科学院视觉与人工智能实验室)
;
Chair for Machine Learning, University of Mannheim(曼海姆大学机器学习主任)
;
Max Planck Institute for Informatics, Saarland Informatics Campus, Saarbrücken, Germany(马克斯·普朗克信息研究所,萨尔兰州信息学院,萨尔布吕肯,德国)
SigVLP: Sigmoid Volume-Language Pre-Training for Self-Supervised CT-Volume Adaptive Representation Learning
SigVLP:基于sigmoid体积-语言预训练的自监督CT体积自适应表示学习
Jiayi Wang, Hadrien Reynaud, Ibrahim Ethem Hamamci, Sezgin Er, Suprosanna Shit, Bjoern Menze, Bernhard Kainz
机构
*
Friedrich-Alexander University Erlangen-Nürnberg(弗里德里希-亚历山大大学埃尔兰根-纽伦堡)
;
Department of Quantitative Biomedicine, University of Zurich(苏黎世大学定量生物医学系)
;
ETH AI Center, ETH Zurich(苏黎世联邦理工学院AI中心)
;
International School of Medicine, Istanbul Medipol University(伊斯坦布尔梅迪波尔大学国际医学院)
;
Department of Computing, Imperial College London(伦敦帝国学院计算机系)
Multi-Head RAG: Solving Multi-Aspect Problems with LLMs
多头RAG:利用LLMs解决多方面问题
Maciej Besta, Ales Kubicek, Robert Gerstenberger, Marcin Chrapek, Roman Niggli, Patrik Okanovic, Yi Zhu, Patrick Iff, Michal Podstawski, Lucas Weitzendorf, Mingyuan Chi, Joanna Gajda, Piotr Nyczyk, Jürgen Müller, Hubert Niewiadomski, Torsten Hoefler
机构
*
Department of Computer Science, ETH Zurich(苏黎世联邦理工学院计算机科学系)
;
IDEAS Research Institute(IDEAS研究 institute)
;
NASK National Research Institute(国家研究 institute)
专题命中
其他LLM
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Are Foundation Models the Route to Full-Stack Transfer in Robotics?
基础模型是否是机器人领域全栈迁移的途径?
Freek Stulp, Samuel Bustamante, João Silvério, Alin Albu-Schäffer, Jeannette Bohg, Shuran Song
机构
*
Institute of Robotics and Mechatronics, German Aerospace Center (DLR)(机器人与机电研究所,德国航空航天中心(DLR))
;
Stanford AI Lab, Stanford University(斯坦福大学人工智能实验室)