CommentsMICCAI 2026 Early Accept; Project Page: https://tahakoleilat.github.io/Evi-Steer. This preprint has not undergone peer review or any post-submission improvements or corrections. The Version of Record of this contribution will be published as part of the MICCAI 2026 proceedings in October
Grounding or Guessing? Visual Signals for Detecting Hallucinations in Sign Language Translation
基于视觉线索检测手语翻译中的幻觉:是依据视觉信息还是猜测?
Yasser Hamidullah, Koel Dutta Chowdhury, Yusser Al Ghussin, Shakib Yazdani, Cennet Oguz, Josef van Genabith, Cristina España-Bonet
机构
*
German Research Center for Artificial Intelligence (DFKI GmbH)(德国人工智能研究中心(DFKI GmbH))
;
Saarland Informatics Campus(萨尔兰州信息学校园)
;
Barcelona Supercomputing Center (BSC-CNS)(巴塞罗那超级计算中心(BSC-CNS))
What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucination Robustness
什么使LVLMs更少产生幻觉?揭示影响幻觉鲁棒性的架构因素
Yusheng He, Jizhe Zhou, Xia Du, Zheng Lin, Jun Luo, Jiancheng Lv
机构
*
School of Computer Science, Engineering Research Center of Machine Learning and Industry Intelligence, Sichuan University(计算机科学学院,机器学习与产业智能工程研究中心,四川大学)
;
School of Computer and Information Engineering, Xiamen University of Technology(计算机与信息工程学院,厦门理工大学)
;
Department of Electrical and Computer Engineering, University of Hong Kong(电气与计算机工程系,香港大学)
;
College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学)
Benchmarking Uncertainty and its Disentanglement in multi-label Chest X-Ray Classification
多标签胸部X光分类中的不确定性及其解缠基准测试
Simon Baur, Wojciech Samek, Jackie Ma
机构
*
Fraunhofer Heinrich-Hertz-Institut(弗劳恩霍夫海因里希-赫兹研究所)
;
Technische Universität Berlin(柏林技术大学)
;
The Berlin Institute for the Foundations of Learning and Data (BIFOLD)(柏林学习与数据基础研究所)
Physics Is All You Need? A Case Study in Physicist-Supervised AI Development of Scientific Software
物理学就是一切?物理学家监督人工智能开发科学软件的案例研究
Nhat-Minh Nguyen
机构
*
Kavli IPMU (WPI), UTIAS, The University of Tokyo(Kavli研究所(WPI)、UTIAS、东京大学)
;
Center for Data-Driven Discovery(数据驱动发现中心)
;
Institute For Interdisciplinary Research in Science(科学跨学科研究中心)
Comments10 pages, 2 figures, 2 tables, 1 physicist and a few AI agents. Accepted by ICML 2026 AI for Science Workshop. Code and development log are available at this repo: https://github.com/MinhMPA/clax-pt
Calibration Is Not Enough: Evaluating Confidence Estimation Under Language Variations
校准还不够:评估语言变化下的置信度估计
Yuxi Xia, Dennis Ulmer, Terra Blevins, Yihong Liu, Hinrich Schütze, Benjamin Roth
机构
*
Faculty of Computer Science, UniVie Doctoral School Computer Science(计算机科学系,维也纳大学计算机科学博士学院)
;
Faculty of Philological and Cultural Studies, University of Vienna, Austria(文学与文化研究系,维也纳大学,奥地利)
;
ILLC, University of Amsterdam, Netherlands(阿姆斯特丹大学ILLC,荷兰)
;
Khoury College of Computer Sciences, Northeastern University, USA(东北大学计算机科学学院,美国)
;
LMU Munich, Munich Center for Machine Learning (MCML), Germany(慕尼黑大学,慕尼黑机器学习中心(MCML),德国)
Mary Chriselda Antony Oliver, Lan Jiang, Aaron Bundi Anampiu, Elaf Almahmoud, Francesco Quinzan, Umang Bhatt
机构
*
Department of Applied Mathematics and Theoretical Physics, University of Cambridge(应用数学与理论物理系,剑桥大学)
;
Centre for Human-Inspired Artificial Intelligence, University of Cambridge(启发式人工智能中心,剑桥大学)
;
African Institute for Mathematical Sciences, South Africa(南非数学科学研究所)
;
Department of Engineering Science, University of Oxford(工程科学系,牛津大学)
An uncertainty-aware Bayesian framework for machine learning classification models: A case study in land cover classification
一种不确定性感知的贝叶斯机器学习分类模型框架:以土地覆盖分类为例
Samuel Bilson, Miles McCrory, Anna Pustogvar
机构
*
National Physical Laboratory, Teddington, UK(英国国家物理实验室,Teddington)
;
Department of Data Science(数据科学系)
;
Department of Thermal & Radiometric Metrology(热学与辐射计量学系)
;
School of Geography, Geology & the Environment(地理、地质与环境学院)
Universal Boosts, Specific Suppressors: Sparse Autoencoder Steering of Medical Vision-Language Models
通用增强,特定抑制:基于稀疏自编码器引导的医学视觉语言模型
Farhad Nooralahzadeh, Benjamin Gundersen, Nicolas Deperrois, Hidetoshi Matsuom, Mizuho Nishio, Thomas Frauenfelder, Ahmed Allam, Christian Blüthgen, Michael Moor, Michael Krauthammer
机构
*
University of Zurich and University Hospital of Zurich(苏黎世大学及苏黎世大学医院)
;
Kobe University(Kobe大学)
;
ETH AI Center(苏黎世联邦理工学院人工智能中心)
;
ETH Zurich(苏黎世联邦理工学院)
;
Stanford University(斯坦福大学)
;
Zurich University of Applied Sciences(苏黎世应用科学大学)
HawkesLLM: Semantic Uncertainty Propagation in Agentic Text Simulation
HawkesLLM:智能体文本模拟中的语义不确定性传播
Zewei Deng, Tinghan Ye, Liyan Xie
机构
*
Department of Industrial and Systems Engineering, University of Minnesota(工业与系统工程系,明尼苏达大学)
;
H. Milton Stewart School of Industrial and Systems Engineering, Georgia Institute of Technology(H. Milton Stewart工业与系统工程学院,佐治亚理工学院)
Task-Awareness Improves LLM Generations and Uncertainty
任务感知提升大语言模型生成与不确定性
Tim Tomov, Dominik Fuchsgruber, Stephan Günnemann
机构
*
School of Computation, Information \& Technology, Technical University of Munich
;
Munich Data Science Institute
;
Munich Center for Machine Learning
Testable and Actionable Calibration for Full Swap Regret
可检验且可操作的全面交换懊悔校准
Konstantina Bairaktari, Lunjia Hu, Huy L. Nguyen, Jonathan Ullman
机构
*
Department of Computer Science, Aarhus University(阿arhus大学计算机科学系)
;
Khoury College of Computer Sciences, Northeastern University(东北大学计算机科学学院)
;
Northeastern University(东北大学)