arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-25 至 2025-09-25 共收录 11 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 11 篇

2509.19375 2025-09-25 cs.LG cs.AI stat.ML 88%

Uncertainty Quantification of Large Language Models using Approximate Bayesian Computation

Mridul Sharma, Adeetya Patel, Zaneta D' Souza, Samira Abbasgholizadeh Rahimi, Siva Reddy, Sreenath Madathil

机构 * Faculty of Dental Medicine and Oral Health Sciences, McGill University(牙医学院与口腔健康科学学院,麦吉尔大学) McGill University(麦吉尔大学) Mila–Quebec Artificial Intelligence Institute(魁北克人工智能研究所) School of Computer Science, McGill University(计算机科学学院,麦吉尔大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19563 2025-09-25 cs.CL cs.LG 81%

Uncertainty in Semantic Language Modeling with PIXELS

Stefania Radu, Marco Zullich, Matias Valdenegro-Toro

机构 * Department of Artificial Intelligence, Bernoulli Institute, University of Groningen(人工智能系、伯努利研究所、 Groningen大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

Comments 9 pages, 6 figures, UncertaiNLP 2025 Workshop @ EMNLP Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18485 2025-09-25 q-bio.NC cs.CV 78%

Deciphering Functions of Neurons in Vision-Language Models

Jiaqi Xu, Cuiling Lan, Yan Lu

机构 * University of Science and Technology of China(中国科学技术大学) Microsoft Research Asia(微软亚洲研究院)

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments Accepted by the 31st ACM International Conference on Multimedia (ACM MM 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19839 2025-09-25 cs.AI 77%

LatentGuard: Controllable Latent Steering for Robust Refusal of Attacks and Reliable Response Generation

Huizhen Shu, Xuying Li, Zhuo Li

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 9-page NeurIPS 2025 preprint including 3 figures and 1 table, with additional appendix material. Prepared using the NeurIPS 2025 preprint template and compiled with pdfLaTeX. All references are included via the provided .bbl file. Figures are in PDF format. No external supplementary files. All necessary style files and images are included

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19745 2025-09-25 cs.CL cs.SD 77%

PART: Progressive Alignment Representation Training for Multilingual Speech-To-Text with LLMs

Pei Zhang, Andong Chen, Xi Chen, Baosong Yang, Derek F. Wong, Fei Huang

机构 * Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团) The Chinese University of Hong Kong(香港中文大学) NLP 2 CT Lab, University of Macau(自然语言处理2CT实验室,澳门大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03090 2025-09-25 cs.CL cs.LG 73%

UNComp: Can Matrix Entropy Uncover Sparsity? -- A Compressor Design from an Uncertainty-Aware Perspective

Jing Xiong, Jianghan Shen, Fanghua Ye, Chaofan Tao, Zhongwei Wan, Jianqiao Lu, Xun Wu, Chuanyang Zheng, Zhijiang Guo, Min Yang, Lingpeng Kong, Ngai Wong

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Accepted at EMNLP 2025 (Main Conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20168 2025-09-25 cs.CL 70%

Probing Gender Bias in Multilingual LLMs: A Case Study of Stereotypes in Persian

Ghazal Kalhor, Behnam Bahrak

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted and forthcoming at the Widening Natural Language Processing Workshop (WiNLP 2025) at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20088 2025-09-25 cs.CL cs.AI 62%

Causal Understanding by LLMs: The Role of Uncertainty

Oscar Lithgow-Serrano, Vani Kanjirangat, Alessandro Antonucci

机构 * SUPSI, IDSIA(瑞士SUPSI和IDSIA)

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.CL、cs.AI

Comments Accepted in second UncertaiNLP workshop at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20065 2025-09-25 cs.CL 57%

From Input Perception to Predictive Insight: Modeling Model Blind Spots Before They Become Errors

Maggie Mi, Aline Villavicencio, Nafise Sadat Moosavi

机构 * University of Sheffield(谢菲尔德大学) University of Exeter(埃克塞特大学) The Alan Turing Institute(艾伦·图灵研究所) UFRN, Brazil(巴西UFRN)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13390 2025-09-25 cs.CL 57%

Aligned Probing: Relating Toxic Behavior and Model Internals

Andreas Waldis, Vagrant Gautam, Anne Lauscher, Dietrich Klakow, Iryna Gurevych

机构 * Ubiquitous Knowledge Processing Lab (UKP Lab)(通用知识处理实验室) Technical University of Darmstadt(德累斯顿技术大学) Information Systems Research Lab(信息系统研究实验室) Lucerne University of Applied Sciences and Arts(卢塞恩应用科学与艺术大学) Spoken Language Systems(语音语言系统) Saarland University(萨尔兰大学) Data Science Group(数据科学组) University of Hamburg(汉堡大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19562 2025-09-25 cs.CV 50%

CURE: Centroid-guided Unsupervised Representation Erasure for Facial Recognition Systems

Fnu Shivam, Nima Najafzadeh, Yenumula Reddy, Prashnna Gyawali

机构 * West Virginia University(西弗吉尼亚大学)

专题命中 知识编辑与模型理解 :prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏