arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-08 至 2025-10-08 共收录 14 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 14 篇

2410.15460 2025-10-08 cs.AI cs.CL math.SP 88%

Hallucination Detox: Sensitivity Dropout (SenD) for Large Language Model Training

Shahrad Mohammadzadeh, Juan David Guerra, Marco Bonizzato, Reihaneh Rabbany, Golnoosh Farnadi

机构 * McGill University(麦吉尔大学) Polytechnique Montréal(蒙特利尔理工学院) Université de Montréal(蒙特利尔大学) Mila - Quebec Artificial Intelligence Institute(魁北克人工智能研究所) CIFAR AI Chair(CIFAR人工智能主席)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted to ACL 2025, accepted to Safe Generative AI Workshop @ NeurIPS 2024. Camera-ready version for ACL 2025 (to appear). Submitted July 2025

Journal ref Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 5538-5554, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05664 2025-10-08 cs.AI 84%

Large Language Model-Based Uncertainty-Adjusted Label Extraction for Artificial Intelligence Model Development in Upper Extremity Radiography

Hanna Kreutzer, Anne-Sophie Caselitz, Thomas Dratsch, Daniel Pinto dos Santos, Christiane Kuhl, Daniel Truhn, Sven Nebelung

机构 * Lab for Artificial Intelligence in Medicine, Department of Diagnostic and Interventional Radiology, University Hospital Aachen(人工智能医学实验室,诊断与介入放射科,亚琛大学医院) Department of Diagnostic and Interventional Radiology, University Hospital Aachen(诊断与介入放射科,亚琛大学医院) Institute for Diagnostic and Interventional Radiology, Faculty of Medicine and University Hospital Cologne, University of Cologne(诊断与介入放射学研究所,医学学院和科隆大学医院,科隆大学) Department of Diagnostic and Interventional Radiology, University Medical Center Mainz(诊断与介入放射科,马因茨大学医学中心)

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.AI

Comments 28 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10513 2025-10-08 cs.AI cs.LG 81%

Extracting PAC Decision Trees from Black Box Binary Classifiers: The Gender Bias Case Study on BERT-based Language Models

Ana Ozaki, Roberto Confalonieri, Ricardo Guimarães, Anders Imenes

机构 * Universitetet i Oslo, Norway(奥斯陆大学) Universita degli Studi di Padova, Italy(帕多瓦大学) Universitetet i Bergen, Norway(卑尔根大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI、cs.LG

Comments This is a revision of the version published at AAAI 2025. We fixed an issue in Theorem 8 and run again all the experiments. We also fixed small grammar mistakes found while producing this revised version

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22945 2025-10-08 cs.CL cs.AI 79%

OWL: Probing Cross-Lingual Recall of Memorized Texts via World Literature

Alisha Srivastava, Emir Korukluoglu, Minh Nhat Le, Duyen Tran, Chau Minh Pham, Marzena Karpinska, Mohit Iyyer

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);pretraining(abstract);分类 cs.CL、cs.AI

Comments Accepted to EMNLP 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05189 2025-10-08 cs.CL cs.AI 79%

A novel hallucination classification framework

Maksym Zavhorodnii, Dmytro Dehtiarov, Anna Konovalenko

机构 * Instituto Superior Técnico, Universidade de Lisboa(里斯本大学技术学院) Molde University College(莫尔德大学学院)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 15 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02476 2025-10-08 cs.LG q-bio.QM 78%

Uncertainty-Guided Model Selection for Tabular Foundation Models in Biomolecule Efficacy Prediction

Jie Li, Andrew McCarthy, Zhizhuo Zhang, Stephen Young

机构 * AIML GSK

专题命中 知识编辑与模型理解 :foundation model(title,comments);分类 cs.LG;large language model(comments);language model(comments)

Comments Accepted by NeurIPS 2025 workshop: 2nd Workshop on Multi-modal Foundation Models and Large Language Models for Life Sciences

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14420 2025-10-08 q-fin.CP cs.CL cs.LG 73%

SAE-FiRE: Enhancing Earnings Surprise Predictions Through Sparse Autoencoder Feature Selection

Huopu Zhang, Yanguang Liu, Miao Zhang, Zirui He, Mengnan Du

机构 * Georgia Institute of Technology(佐治亚理工学院) New Jersey Institute of Technology(新泽西理工学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.12324 2025-10-08 cs.CL cs.AI 73%

Cross-Document Cross-Lingual NLI via RST-Enhanced Graph Fusion and Interpretability Prediction

Mengying Yuan, Wenhao Wang, Zixuan Wang, Yujie Huang, Kangli Wei, Fei Li, Chong Teng, Donghong Ji

机构 * Key Laboratory of Aerospace Information Security and Trusted Computing, Ministry of Education, School of Cyber Science and Engineering, Wuhan University(航空信息安全与可信计算重点实验室,教育部,网络安全与工程学院,武汉大学) Zhejiang University(浙江大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments EMNLP 2025 Main (Camera Ready)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05116 2025-10-08 cs.CL cs.AI 73%

Hallucination is Inevitable for LLMs with the Open World Assumption

Bowen Xu

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17744 2025-10-08 cs.LG 70%

Randomly Removing 50% of Dimensions in Text Embeddings has Minimal Impact on Retrieval and Classification Tasks

Sotaro Takeshita, Yurina Takeshita, Daniel Ruffinelli, Simone Paolo Ponzetto

机构 * Data and Web Science Group, University of Mannheim(曼海姆大学数据与网络科学组)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted to EMNLP 2025 Main Conference (Oral), camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18427 2025-10-08 q-fin.CP q-fin.RM 67%

Tracing Positional Bias in Financial Decision-Making: Mechanistic Insights from Qwen2.5

Fabrizio Dimino, Krati Saxena, Bhaskarjit Sarmah, Stefano Pasquali

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24492 2025-10-08 cs.LG cs.AI 62%

Object Centric Concept Bottlenecks

David Steinmann, Wolfgang Stammer, Antonia Wüst, Kristian Kersting

机构 * Computer Science Department, TU Darmstadt(图宾根大学计算机科学系) Hessian Center for AI (hessian.AI)(海德堡人工智能中心) German Research Center for AI (DFKI)(德国人工智能研究中心) Centre for Cognitive Science, TU Darmstadt(图宾根大学认知科学中心)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11846 2025-10-08 eess.IV cs.AI cs.CV cs.LG q-bio.QM 62%

A Graph-Based Framework for Interpretable Whole Slide Image Analysis

Alexander Weers, Alexander H. Berger, Laurin Lux, Peter Schüffler, Daniel Rueckert, Johannes C. Paetzold

机构 * School of Computation, Information and Technology, Technical University of Munich(慕尼黑技术大学计算、信息与技术学院) Department of Computing, Imperial College London(伦敦帝国学院计算机系) Munich Center for Machine Learning(慕尼黑机器学习中心) Munich Data Science Institute, Technical University of Munich(慕尼黑数据科学研究所) Institute of Pathology, TUM School of Medicine and Health, Technical University of Munich(慕尼黑技术大学医学与健康学院病理学研究所) Weill Cornell Medicine, Cornell University(韦尔医学院,康奈尔大学)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

Comments 15 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05586 2025-10-08 cs.CV 50%

CalibCLIP: Contextual Calibration of Dominant Semantics for Text-Driven Image Retrieval

Bin Kang, Bin Chen, Junjie Wang, Yulin Li, Junzhi Zhao, Zhuotao Tian

机构 * Chengdu Institute of Computer Applications, Chinese Academy of Sciences(成都计算机应用研究所,中国科学院) University of Chinese Academy of Sciences(中国科学院大学) International Research Institute for Artificial Intelligence, Harbin Institute of Technology (Shenzhen)(人工智能国际研究院,哈尔滨工业大学(深圳)) Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)) Southwest Jiaotong University(西南交通大学) Tencent(腾讯)

专题命中 知识编辑与模型理解 :language model(abstract)

Comments ACMMM2025(oral)

详情

展开后加载摘要…

URL PDF HTML 收藏