arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-30 至 2025-09-30 共收录 30 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 30 篇

2509.24202 2025-09-30 cs.CL cs.AI 91%

Can Large Language Models Express Uncertainty Like Human?

Linwei Tao, Yi-Fan Yeh, Bo Kai, Minjing Dong, Tao Huang, Tom A. Lamb, Jialin Yu, Philip H. S. Torr, Chang Xu

机构 * School of Computer Science, University of Sydney(悉尼大学计算机科学学院) City University of Hong Kong(香港城市大学) Shanghai Jiao Tong University(上海交通大学) Department of Engineering Science, University of Oxford(牛津大学工程科学系)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15229 2025-09-30 cs.CL cs.CY 89%

Multilingual Prompting for Improving LLM Generation Diversity

Qihan Wang, Shidong Pan, Tal Linzen, Emily Black

机构 * New York University(纽约大学) Columbia University(哥伦比亚大学)

专题命中 知识编辑与模型理解 :prompting(title,abstract);LLM(title);large language model(abstract);language model(abstract)

Comments Accepted by EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18460 2025-09-30 cs.SE 89%

ExeCoder: Empowering Large Language Models with Executability Representation for Code Translation

Minghua He, Yue Chen, Fangkai Yang, Pu Zhao, Wenjie Yin, Yu Kang, Qingwei Lin, Saravan Rajmohan, Dongmei Zhang

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract)

Comments EMNLP 2025 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24338 2025-09-30 cs.CL 88%

AlignX: Advancing Multilingual Large Language Models with Multilingual Representation Alignment

Mengyu Bu, Shaolei Zhang, Zhongjun He, Hua Wu, Yang Feng

机构 * Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, Chinese Academy of Sciences (ICT/CAS)(智能信息处理重点实验室,计算技术研究所,中国科学院) Key Laboratory of AI Safety, Chinese Academy of Sciences(人工智能安全重点实验室,中国科学院) University of Chinese Academy of Sciences, Beijing, China(中国科学院大学,北京,中国) Baidu Inc.(百度公司)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

Comments Accepted to EMNLP 2025 Main Conference. The code will be available at https://github.com/ictnlp/AlignX

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.21372 2025-09-30 cs.CL cs.AI 88%

Retrieval-Enhanced Few-Shot Prompting for Speech Event Extraction

Máté Gedeon

机构 * Budapest University of Technology and Economics(布达佩斯技术与经济大学)

专题命中 知识编辑与模型理解 :prompting(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22831 2025-09-30 cs.AI cs.CL 86%

Toward a Theory of Generalizability in LLM Mechanistic Interpretability Research

Sean Trott

机构 * Department of Cognitive Science University of California, San Diego(认知科学系加州大学圣地亚哥分校)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);pretraining(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.20079 2025-09-30 cs.CL cs.AI 86%

Molecular Facts: Desiderata for Decontextualization in LLM Fact Verification

Anisha Gunjal, Greg Durrett

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Journal ref Findings of the Association for Computational Linguistics: EMNLP 2024 (2024) 3751-3768

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22735 2025-09-30 cs.CY cs.AI 85%

Regulating the Agency of LLM-based Agents

Seán Boddy, Joshua Joseph

机构 * Berkman Klein Center for Internet & Society(互联网与社会伯克曼克莱因中心) Harvard University(哈佛大学)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 4 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23002 2025-09-30 stat.ML cs.LG 79%

Unsupervised Conformal Inference: Bootstrapping and Alignment to Control LLM Uncertainty

Lingyou Pang, Lei Huang, Jianyu Lin, Tianyu Wang, Akira Horiguchi, Alexander Aue, Carey E. Priebe

机构 * Department of Statistics, University of California, Davis(加州大学戴维斯分校统计系) Department of Applied Mathematics and Statistics, Johns Hopkins University(约翰霍普金斯大学应用数学与统计学系)

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.LG

Comments 26 pages including appendix; 3 figures and 5 tables. Under review for ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24298 2025-09-30 cs.HC cs.AI cs.CL cs.CY cs.MM 79%

Bridging the behavior-neural gap: A multimodal AI reveals the brain's geometry of emotion more accurately than human self-reports

Changde Du, Yizhuo Lu, Zhongyu Huang, Yi Sun, Zisen Zhou, Shaozheng Qin, Huiguang He

机构 * State Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology, Institute of Automation, Chinese Academy of Sciences(脑认知与脑启发智能技术重点实验室,自动化研究所,中国科学院) School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学) School of Future Technology, University of Chinese Academy of Sciences(未来技术学院,中国科学院大学) State Key Laboratory of Cognitive Neuroscience and Learning, Beijing Normal University(认知神经科学与学习国家重点实验室,北京师范大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24171 2025-09-30 cs.LG 70%

Model Correlation Detection via Random Selection Probing

Ruibo Chen, Sheng Zhang, Yihan Wu, Tong Zheng, Peihua Mai, Heng Huang

机构 * University of Maryland, College Park(马里兰大学学院公园分校)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23684 2025-09-30 cs.LG 70%

Hedonic Neurons: A Mechanistic Mapping of Latent Coalitions in Transformer MLPs

Tanya Chowdhury, Atharva Nijasure, Yair Zick, James Allan

机构 * Center for Intelligent Information Retrieval(智能信息检索中心) University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23434 2025-09-30 cs.HC cs.AI 70%

NeuroBridge: Using Generative AI to Bridge Cross-neurotype Communication Differences through Neurotypical Perspective-taking

Rukhshan Haroon, Kyle Wigdor, Katie Yang, Nicole Toumanios, Eileen T. Crehan, Fahad Dogar

机构 * Computer Science Tufts University(计算机科学 华盛顿大学) Human Development, Cognitive Science Tufts University(人类发展与认知科学 华盛顿大学) Eunice Kennedy Shriver Center UMass Chan Medical School(欧尼丝·凯瑟琳·施里弗中心 马萨诸塞大学医学学院) Tufts University(华盛顿大学) UMass Chan Medical School(马萨诸塞大学医学学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.13111 2025-09-30 cs.LG cs.RO 70%

Uncertainty-Aware Trajectory Prediction via Rule-Regularized Heteroscedastic Deep Classification

Kumar Manas, Christian Schlauch, Adrian Paschke, Christian Wirth, Nadja Klein

机构 * Department of Mathematics and Computer Science, Freie Universität Berlin(自由大学柏林数学与计算机科学系) Continental Automotive Technologies GmbH, AI Lab Berlin(大陆汽车技术有限公司柏林AI实验室) Karlsruhe Institute of Technology, Scientific Computing Center, Methods for Big Data(卡尔斯鲁厄理工学院科学计算中心、大数据方法) Fraunhofer Institute for Open Communication Systems, Berlin, Germany(弗劳恩霍夫开放通信系统研究所,柏林德国)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments 17 Pages, 9 figures. Accepted to Robotics: Science and Systems(RSS), 2025

Journal ref Robotics: Science and Systems (RSS), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25177 2025-09-30 cs.CV 67%

Mitigating Hallucination in Multimodal LLMs with Layer Contrastive Decoding

Bingkui Tong, Jiaer Xia, Kaiyang Zhou

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德·本·扎耶德人工智能大学) Hong Kong Baptist University(香港 Baptist大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03271 2025-09-30 cs.HC 67%

Beyond Quantification: Navigating Uncertainty in Professional AI Systems

Sylvie Delacroix, Diana Robinson, Umang Bhatt, Jacopo Domenicucci, Jessica Montgomery, Gael Varoquaux, Carl Henrik Ek, Vincent Fortuin, Yulan He, Tom Diethe, Neill Campbell, Mennatallah El-Assady, Soren Hauberg, Ivana Dusparic, Neil Lawrence

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Journal ref RSS Data Science (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17550 2025-09-30 cs.CV 67%

T2VUnlearning: A Concept Erasing Method for Text-to-Video Diffusion Models

Xiaoyu Ye, Songjie Cheng, Yongtao Wang, Yajiao Xiong, Yishen Li

机构 * Peking University(北京大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23457 2025-09-30 cs.CV 67%

No Concept Left Behind: Test-Time Optimization for Compositional Text-to-Image Generation

Mohammad Hossein Sameti, Amir M. Mansourian, Arash Marioriyad, Soheil Fadaee Oshyani, Mohammad Hossein Rohban, Mahdieh Soleymani Baghshah

机构 * Sharif University of Technology(沙里夫技术大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 8 pages, 8 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17858 2025-09-30 cs.IR 67%

LexSemBridge: Fine-Grained Dense Representation Enhancement through Token-Aware Embedding Augmentation

Shaoxiong Zhan, Hai Lin, Hongming Tan, Xiaodong Cai, Hai-Tao Zheng, Xin Su, Zifei Shan, Ruitong Liu, Hong-Gee Kim

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 8 pages, 4 figures. Accepted to ECAI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14233 2025-09-30 cs.CL cs.AI cs.LG 67%

Mechanistic Fine-tuning for In-context Learning

Hakaze Cho, Peng Luo, Mariko Kato, Rin Kaenbyou, Naoya Inoue

机构 * Japan Advanced Institute of Science and Technology(日本科学技术先进研究院) Beijing Institute of Technology(北京理工大学) RIKEN(日本研究机构)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 28 pages, 31 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09129 2025-09-30 cs.LG cs.AI cs.CL 67%

NextLocLLM: Location Semantics Modeling and Coordinate-Based Next Location Prediction with LLMs

Shuai Liu, Ning Cao, Yile Chen, Yue Jiang, George Rosario Jagadeesh, Gao Cong

机构 * Nanyang Technological University(南洋理工大学)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments STIntelligence in CIKM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08885 2025-09-30 cs.CL cs.LG 62%

AdversariaL attacK sAfety aLIgnment(ALKALI): Safeguarding LLMs through GRACE: Geometric Representation-Aware Contrastive Enhancement- Introducing Adversarial Vulnerability Quality Index (AVQI)

Danush Khanna, Gurucharan Marthi Krishna Kumar, Basab Ghosh, Yaswanth Narsupalli, Vinija Jain, Vasu Sharma, Aman Chadha, Amitava Das

专题命中 知识编辑与模型理解 :preference optimization(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.01651 2025-09-30 q-bio.QM cs.AI cs.LG q-bio.BM 62%

FusionDTI: Fine-grained Binding Discovery with Token-level Fusion for Drug-Target Interaction

Zhaohan Meng, Zaiqiao Meng, Ke Yuan, Iadh Ounis

机构 * School of Computing Science(计算科学学院) School of Cancer Sciences(癌症科学学院) Cancer Research UK Scotland Institute(英国癌症研究苏格兰研究所) University of Glasgow(格拉斯哥大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

Comments Findings of EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24431 2025-09-30 cs.LG 57%

Semantic Compression via Multimodal Representation Learning

Eleonora Grassucci, Giordano Cicchetti, Aurelio Uncini, Danilo Comminiello

机构 * Dept. of Information Engineering, Electronics, and Telecomm.(信息工程、电子与电信系)

专题命中 知识编辑与模型理解 :post-training(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00310 2025-09-30 cs.RO cs.AI 57%

TReF-6: Inferring Task-Relevant Frames from a Single Demonstration for One-Shot Skill Generalization

Yuxuan Ding, Shuangge Wang, Tesca Fitzgerald

机构 * Yale University(耶鲁大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23717 2025-09-30 cs.AI 57%

Measuring Sparse Autoencoder Feature Sensitivity

Claire Tian, Katherine Tian, Nathan Hu

机构 * The Harker School(哈克尔学校) Stanford University(斯坦福大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

Comments NeurIPS 2025 Workshop on Mechanistic Interpretability Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12341 2025-09-30 cs.CV cs.AI 57%

Semantic Discrepancy-aware Detector for Image Forgery Identification

Ziye Wang, Minghang Yu, Chunyan Xu, Zhen Cui

机构 * Nanjing University of Science and Technology(南京理工大学) Beijing Normal University(北京师范大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15963 2025-09-30 cs.CV cs.CL 57%

OViP: Online Vision-Language Preference Learning for VLM Hallucination

Shujun Liu, Siyuan Wang, Zejun Li, Jianxiang Wang, Cheng Zeng, Zhongyu Wei

机构 * Fudan University(复旦大学) University of Southern California(南加州大学) ByteDance(字节跳动)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23849 2025-09-30 cs.CV 50%

CE-FAM: Concept-Based Explanation via Fusion of Activation Maps

Michihiro Kuroki, Toshihiko Yamasaki

机构 * The University of Tokyo(东京大学)

专题命中 知识编辑与模型理解 :language model(abstract)

Comments This paper has been accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23838 2025-09-30 cs.CV 50%

2nd Place Report of MOSEv2 Challenge 2025: Concept Guided Video Object Segmentation via SeC

Zhixiong Zhang, Shuangrui Ding, Xiaoyi Dong, Yuhang Zang, Yuhang Cao, Jiaqi Wang

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai Innovation Institute(上海创新研究院) Shanghai AI Laboratory(上海人工智能实验室) The Chinese University of Hong Kong(香港中文大学)

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏