arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-03 至 2025-12-03 共收录 12 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 12 篇

2512.02772 2025-12-03 cs.CL cs.IR 89%

Towards Unification of Hallucination Detection and Fact Verification for Large Language Models

迈向大型语言模型中幻觉检测与事实验证的统一

Weihang Su, Jianming Long, Changyue Wang, Shiyu Lin, Jingyan Xu, Ziyi Ye, Qingyao Ai, Yiqun Liu

机构 * DCST, Tsinghua University(清华大学数据科学研究院) Fudan University(复旦大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 本文提出UniFact框架,通过统一评估方法揭示幻觉检测与事实验证的互补性,并证明混合方法在LLM中的优越性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13514 2025-12-03 cs.CL cs.AI 88%

Induction Head Toxicity Mechanistically Explains Repetition Curse in Large Language Models

诱导头毒性机制解释了大语言模型中的重复诅咒

Shuxun Wang, Qingyu Yin, Chak Tou Leong, Qiang Zhang, Linyi Yang

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文揭示诱导头的毒性机制导致大语言模型重复诅咒,并提出通过注意力头正则化技术缓解该问题的方法。

Comments Need to be refined

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.05036 2025-12-03 cs.CL 86%

From Word Vectors to Multimodal Embeddings: Techniques, Applications, and Future Directions For Large Language Models

从词向量到多模态嵌入:大型语言模型的技术、应用与未来方向

Charles Zhang, Benji Peng, Xintian Sun, Qian Niu, Junyu Liu, Keyu Chen, Ming Li, Pohsun Feng, Ziqian Bi, Ming Liu, Yichao Zhang, Xinyuan Song, Cheng Fei, Caitlyn Heqi Yin, Lawrence KQ Yan, Hongyang He, Tianyang Wang

机构 * Georgia Institute of Technology(佐治亚理工学院) Simon Fraser University(西蒙弗雷泽大学) Kyoto University(京都大学) National Taiwan Normal University(台湾师范大学) Purdue University(普渡大学) The University of Texas at Dallas(德克萨斯大学达拉斯分校) Emory University(埃默里大学) Cornell University(康奈尔大学) University of Wisconsin-Madison(威斯康星大学麦迪逊分校) The Hong Kong University of Science(香港科学大学) University of Liverpool(利物浦大学) University of Warwick(沃里克大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(title);分类 cs.CL

AI总结 本文综述了从词向量到多模态嵌入的发展,探讨了大型语言模型的技术、应用及未来方向。

Comments 21 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02072 2025-12-03 hep-ph physics.data-an 82%

QCD in Language Models: What do they really know about QCD?

语言模型中的QCD:它们真的了解QCD吗?

Antonin Sulc, Patrick L. S. Connor

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract)

AI总结 研究分析了语言模型对QCD的理解,揭示其在参数空间中对QCD概念的嵌入模式,并指出模型在高级量子场论表示上的局限性。

Comments 6 pages, 4 figures, presented at EPS HEP 2025 by Patrick L.S. Connor as Oral Presentation

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21214 2025-12-03 cs.CV cs.CL 79%

VoxRep: Enhancing 3D Spatial Understanding in 2D Vision-Language Models via Voxel Representation

VoxRep:通过体素表示增强2D视觉-语言模型的3D空间理解

Alan Dao, Norapat Buppodom

机构 * Menlo Research(Menlo研究)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

AI总结 本文提出VoxRep方法,通过将体素空间切分为2D切片并输入预训练的视觉-语言模型,实现对3D环境的高效语义理解。

Journal ref Proc. APSIPA ASC 2025, pp. 1464-1469

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02726 2025-12-03 cs.AI 77%

AuditCopilot: Leveraging LLMs for Fraud Detection in Double-Entry Bookkeeping

AuditCopilot: 利用大语言模型在复式记账中进行欺诈检测

Md Abdul Kadir, Sai Suresh Macharla Vasu, Sidharth S. Nair, Daniel Sonntag

机构 * German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心) Oldenburg University(奥尔登堡大学) Saarland University(萨尔兰州立大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);foundation model(abstract);分类 cs.AI

AI总结 AuditCopilot利用大语言模型在复式记账中检测欺诈,优于传统方法并提供自然语言解释,提升审计的可解释性和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02619 2025-12-03 quant-ph 75%

Quantum LLMs Using Quantum Computing to Analyze and Process Semantic Information

利用量子计算分析和处理语义信息的量子大语言模型

Timo Aukusti Laine

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出利用量子计算分析大语言模型语义信息的方法,通过量子电路与LLM语义空间的映射,实验验证了量子计算在语义相似性计算中的应用潜力。

Comments 18 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01797 2025-12-03 cs.AI cs.CL cs.CY 73%

H-Neurons: On the Existence, Impact, and Origin of Hallucination-Associated Neurons in LLMs

H-Neurons: 关于大语言模型中与幻觉相关的神经元的存在、影响及其起源

Cheng Gao, Huimin Chen, Chaojun Xiao, Zhiyi Chen, Zhiyuan Liu, Maosong Sun

机构 * Tsinghua University(清华大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文研究了大语言模型中与幻觉相关的神经元的存在、影响及起源,揭示了这些神经元在预测幻觉和因果行为中的作用,为提升模型可靠性提供新视角。

Comments 20 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00663 2025-12-03 cs.CL cs.AI 73%

Graphing the Truth: Structured Visualizations for Automated Hallucination Detection in LLMs

图示真相:用于自动幻觉检测的结构化可视化

Tanmay Agrawal

机构 * Department of Computer Science, University of Arizona(计算机科学系,亚利桑那大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出通过结构化可视化技术,帮助自动检测大型语言模型中的幻觉问题,提升模型可靠性和响应质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13813 2025-12-03 cs.CL 70%

Geometric Uncertainty for Detecting and Correcting Hallucinations in LLMs

几何不确定性用于检测和纠正大语言模型中的幻觉

Edward Phillips, Sean Wu, Soheila Molaei, Danielle Belgrave, Anshul Thakur, David Clifton

机构 * Department of Engineering Science, University of Oxford(牛津大学工程科学系) GlaxoSmithKline(葛兰素史克) Oxford Suzhou Centre for Advanced Research(牛津苏黎世高级研究中心)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出几何框架用于检测和纠正大语言模型中的幻觉,通过几何体积和几何怀疑方法提升响应可靠性。

Comments Revision. Clarified positioning as a unified geometric framework for global and local uncertainty in LLMs. Added baselines (Degree, Eccentricity) and expanded comparison to related methods. Included ablations (PCA dimension, number of archetypes, number of samples) and complexity analysis. Extended discussion of medical QA results and model-specific behaviour

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02981 2025-12-03 cs.CV 67%

InEx: Hallucination Mitigation via Introspection and Cross-Modal Multi-Agent Collaboration

InEx:通过内省与跨模态多智能体协作缓解幻觉

Zhongyu Yang, Yingfang Yuan, Xuanming Jiang, Baoyi An, Wei Pang

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

AI总结 InEx通过内省推理和跨模态多智能体协作,自主缓解大型语言模型的幻觉问题,实验表明其在多个基准上表现优异。

Comments Published in AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03079 2025-12-03 cs.CV 50%

Bias Beyond Demographics: Probing Decision Boundaries in Black-Box LVLMs via Counterfactual VQA

偏见超越人口统计:通过反事实视觉问答探测黑盒大视觉-语言模型的决策边界

Zaiying Zhao, Toshihiko Yamasaki

机构 * The University of Tokyo(东京大学)

专题命中 知识编辑与模型理解 :language model(abstract)

AI总结 本文通过反事实VQA基准探测黑盒LVLMs的决策边界,揭示非人口属性对决策的更大影响,并展示人类规范验证示例对提升模型响应一致性和公平性的作用。

详情

展开后加载摘要…

URL PDF HTML 收藏