Priors in Time: Missing Inductive Biases for Language Model Interpretability
时间中的先验:语言模型可解释性中缺失的归纳偏置
Ekdeep Singh Lubana, Can Rager, Sai Sumedh R. Hindupur, Valerie Costa, Greta Tuckute, Oam Patel, Sonia Krishna Murthy, Thomas Fel, Daniel Wurgaft, Eric J. Bigelow, Johnny Lin, Demba Ba, Martin Wattenberg, Fernanda Viegas, Melanie Weber, Aaron Mueller
机构
*
Goodfire AI
;
Independent(独立)
;
SEAS, Harvard University(哈佛大学SEAS学院)
;
EPFL(瑞士联邦理工学院)
;
Kempner Institute at Harvard University(哈佛大学凯普内研究所)
;
Department of Psychology, Stanford University(斯坦福大学心理学系)
;
Department of Psychology, Harvard University(哈佛大学心理学系)
;
Decode Research
;
Boston University(波士顿大学)
HEDGE: Hallucination Estimation via Dense Geometric Entropy for VQA with Vision-Language Models
Sushant Gautam, Michael A. Riegler, Pål Halvorsen
机构
*
Simula Metropolitan Center for Digital Engineering (SimulaMet)(Simula数字工程研究中心)
;
Oslo Metropolitan University (OsloMet)(奥斯陆 Metropolitan 大学)
;
Simula Research Laboratory(Simula研究实验室)
Trust in foundation models and GenAI: A geographic perspective
Grant McKenzie, Krzysztof Janowicz, Carsten Kessler
机构
*
McGill University, Canada(麦吉尔大学,加拿大)
;
University of Vienna, Austria(维也纳大学,奥地利)
;
Bochum University of Applied Sciences, Germany(波鸿应用科学大学,德国)
;
Aalborg University Copenhagen, Denmark(奥胡斯大学哥本哈根分校,丹麦)
CommentsWithdrawn due to an accidental duplicate submission. This paper (arXiv:2510.15430) was unintentionally submitted as a new entry instead of a new version of our previous work (arXiv:2508.09201)
SS-DPPN: A self-supervised dual-path foundation model for the generalizable cardiac audio representation
Ummy Maria Muna, Md Mehedi Hasan Shawon, Md Jobayer, Sumaiya Akter, Md Rakibul Hasan, Md. Golam Rabiul Alam
机构
*
Department of Computer Science and Engineering(计算机科学与工程系)
;
BRAC University(布拉克大学)
;
Department of Electricial and Electronic Engineering(电气与电子工程系)
;
Department of Biomedical Engineering(生物医学工程系)
;
Linköping University(林肯堡大学)
;
Department of Electrical and Computer Engineering(电气与计算机工程系)
;
University of Maryland(马里兰大学)
;
School of Electrical Engineering, Computing and Mathematical Sciences(电气工程、计算与数学科学学院)
Unsupervised Conformal Inference: Bootstrapping and Alignment to Control LLM Uncertainty
Lingyou Pang, Lei Huang, Jianyu Lin, Tianyu Wang, Akira Horiguchi, Alexander Aue, Carey E. Priebe
机构
*
Department of Statistics, University of California, Davis(加州大学戴维斯分校统计系)
;
Department of Applied Mathematics and Statistics, Johns Hopkins University(约翰霍普金斯大学应用数学与统计学系)
专题命中
知识编辑与模型理解
:LLM(title,abstract);分类 cs.LG
Comments26 pages including appendix; 3 figures and 5 tables. Under review for ICLR 2026
Rethinking Circuit Completeness in Language Models: AND, OR, and ADDER Gates
Hang Chen, Jiaying Zhu, Xinyu Yang, Wenya Wang
机构
*
Hang Chen School of Computer Science and Technology Xi’an Jiaotong University(Hang Chen 计算机科学与技术学院 西安交通大学)
;
Jiaying Zhu School of Computer Science and Engineering The Chinese University of Hong Kong(Jiaying Zhu 计算科学与工程学院 香港中文大学)
;
Xinyu Yang School of Computer Science and Technology Xi’an Jiaotong University(Xinyu Yang 计算机科学与技术学院 西安交通大学)
;
Wenya Wang School of Computer Science and Engineering Nanyang Technological University(Wenya Wang 计算科学与工程学院 新加坡国立大学)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
Yifan Lu, Ziqi Zhang, Chunfeng Yuan, Jun Gao, Congxuan Zhang, Xiaojuan Qi, Bing Li, Weiming Hu
机构
*
Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information, CASIA(北京多模态信息超级智能安全重点实验室,中国科学院自动化所)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems, CASIA(多模态人工智能系统国家重点实验室,中国科学院自动化所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Hello Group(Hello集团)
;
Nanchang Hangkong University(南昌航空大学)
;
The University of Hong Kong(香港大学)
;
School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)