arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-07 至 2025-11-07 共收录 10 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 10 篇

2402.18397 2025-11-07 cs.CL 90%

Decomposed Prompting: Probing Multilingual Linguistic Structure Knowledge in Large Language Models

Ercong Nie, Shuzhou Yuan, Bolei Ma, Helmut Schmid, Michael Färber, Frauke Kreuter, Hinrich Schütze

机构 * Center for Information and Language Processing (CIS)(信息与语言处理中心) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) ScaDS.AI and TU Dresden(ScaDS.AI 和 梅克伦堡-前波美拉尼亚技术大学) Department of Statistics, LMU Munich(统计学系,慕尼黑大学) University of Maryland, College Park(马里兰大学 College Park 分校)

专题命中 知识编辑与模型理解 :prompting(title,abstract);large language model(title);language model(title);分类 cs.CL

Comments Accepted to AACL-IJCNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04179 2025-11-07 cs.SE cs.AI 88%

Explaining Software Vulnerabilities with Large Language Models

Oshando Johnson, Alexandra Fomina, Ranjith Krishnamurthy, Vaibhav Chaudhari, Rohith Kumar Shanmuganathan, Eric Bodden

机构 * Chapman University(查普曼大学) Paderborn University(帕德博恩大学) University of Oldenburg(奥尔登堡大学) Paderborn University and Fraunhofer IEM(帕德博恩大学和弗劳恩霍夫IEM研究所)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03878 2025-11-07 cs.AI cs.IR cs.LG cs.MA 86%

KnowThyself: An Agentic Assistant for LLM Interpretability

Suraj Prasai, Mengnan Du, Ying Zhang, Fan Yang

机构 * Wake Forest University(威克森林大学) New Jersey Institute of Technology(新泽西理工学院)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 5 pages, 1 figure, Accepted for publication at the Demonstration Track of the 40th AAAI Conference on Artificial Intelligence (AAAI 26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04527 2025-11-07 cs.CL cs.AI 81%

Are language models aware of the road not taken? Token-level uncertainty and hidden state dynamics

Amir Zur, Atticus Geiger, Ekdeep Singh Lubana, Eric Bigelow

机构 * Department of Linguistics, Stanford University(斯坦福大学语言学系) Department of Psychology, Harvard University(哈佛大学心理学系) Center for Brain Science, Harvard University(哈佛大学脑科学中心) Physics of Intelligence Group, NTT Research(NTT研究物理智能小组)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03900 2025-11-07 cs.CL cs.LG 79%

GRAD: Graph-Retrieved Adaptive Decoding for Hallucination Mitigation

Manh Nguyen, Sunil Gupta, Dai Do, Hung Le

机构 * Applied Artificial Intelligence Initiative(应用人工智能计划) Deakin University(德金大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17796 2025-11-07 cs.CL 77%

Findings of the Fourth Shared Task on Multilingual Coreference Resolution: Can LLMs Dethrone Traditional Approaches?

Michal Novák, Miloslav Konopík, Anna Nedoluzhko, Martin Popel, Ondřej Pražák, Jakub Sido, Milan Straka, Zdeněk Žabokrtský, Daniel Zeman

机构 * Charles University(查理大学) University of West Bohemia(西波西米亚大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to CODI-CRAC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04514 2025-11-07 cs.CL cs.CV 77%

DAMRO: Dive into the Attention Mechanism of LVLM to Reduce Object Hallucination

Xuan Gong, Tianshi Ming, Xinpeng Wang, Zhihua Wei

机构 * Department of Computer Science and Technology, Tongji University(计算机科学与技术系,同济大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted by EMNLP2024 (Main Conference), add GitHub link

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24630 2025-11-07 cs.CL cs.AI 73%

Reasoning Models Hallucinate More: Factuality-Aware Reinforcement Learning for Large Reasoning Models

Junyi Li, Hwee Tou Ng

机构 * Department of Computer Science, National University of Singapore(计算机科学系,新加坡国立大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02850 2025-11-07 cs.CL cs.AI cs.CY cs.DB 62%

Harnessing Structured Knowledge: A Concept Map-Based Approach for High-Quality Multiple Choice Question Generation with Effective Distractors

Nicy Scaria, Silvester John Joseph Kennedy, Diksha Seth, Ananya Thakur, Deepak Subramani

机构 * Indian Institute of Science(印度科学研究院)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI

Comments Accepted to ECAI 2025

Journal ref The European Conference on Artificial Intelligence. 413 (2025). pp 4089--4096. IOS Press

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04078 2025-11-07 cs.CV 50%

Unveiling Deep Semantic Uncertainty Perception for Language-Anchored Multi-modal Vision-Brain Alignment

Zehui Feng, Chenqi Zhang, Mingru Wang, Minuo Wei, Shiwei Cheng, Cuntai Guan, Ting Han

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments 30 pages, 16 figures, under review as a conference paper

详情

展开后加载摘要…

URL PDF HTML 收藏