arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-27 至 2025-10-27 共收录 13 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 13 篇

2505.13763 2025-10-27 cs.AI cs.CL q-bio.NC 84%

Language Models Are Capable of Metacognitive Monitoring and Control of Their Internal Activations

Li Ji-An, Hua-Dong Xiong, Robert C. Wilson, Marcelo G. Mattar, Marcus K. Benna

机构 * Neurosciences Graduate Program University of California San Diego(加州大学圣地亚哥分校神经科学研究生项目) School of Psychology Georgia Tech(佐治亚理工学院心理学系) Department of Psychology New York University(纽约大学心理学系) Department of Neurobiology University of California San Diego(加州大学圣地亚哥分校神经生物学系)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08966 2025-10-27 cs.CL cs.LG cs.NE 81%

Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers

Marek Kadlčík, Michal Štefánik, Timothee Mickus, Michal Spiegel, Josef Kuchař

机构 * Faculty of Informatics, Masaryk University(马萨里克大学信息学院) University of Helsinki(赫尔辛基大学) Kempelen Institute of Intelligent Technologies(凯姆佩尔智能技术研究所)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18512 2025-10-27 cs.IR cs.AI cs.CL cs.LG 80%

AcuRank: Uncertainty-Aware Adaptive Computation for Listwise Reranking

Soyoung Yoon, Gyuwan Kim, Gyu-Hwung Cho, Seung-won Hwang

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted at NeurIPS 2025. The first two authors contributed equally. Author order is randomly determined via coin toss

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21664 2025-10-27 cs.CV q-bio.QM 78%

Foundation Models in Dermatopathology: Skin Tissue Classification

Riya Gupta, Yiwei Zong, Dennis H. Murphree

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19661 2025-10-27 cs.AI 77%

AgentSense: LLMs Empower Generalizable and Explainable Web-Based Participatory Urban Sensing

Xusen Guo, Mingxing Peng, Xixuan Hao, Xingchen Zou, Qiongyan Wang, Sijie Ruan, Yuxuan Liang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Beijing Institute of Technology(北京理工大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 13 pages, 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17477 2025-10-27 cs.CL cs.AI cs.LG 75%

Do Robot Snakes Dream like Electric Sheep? Investigating the Effects of Architectural Inductive Biases on Hallucination

Jerry Huang, Prasanna Parthasarathi, Mehdi Rezagholizadeh, Boxing Chen, Sarath Chandar

机构 * Mila & Université de Montréal(Mila与蒙特利尔大学) Noah’s Ark Lab(Noah’s Ark实验室) Advanced Micro Devices Chandar Research Lab(Chandar研究实验室) Polytechnique Montréal(蒙特利尔理工学院) CIFAR AI Chair(CIFAR人工智能主席)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted to Findings of The 63rd Annual Meeting of the Association for Computational Linguistics (ACL) 2025. Official proceedings version available at https://aclanthology.org/2025.findings-acl.60/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07036 2025-10-27 cs.LG cs.AI 73%

Methodological Insights into Structural Causal Modelling and Uncertainty-Aware Forecasting for Economic Indicators

Federico Cerutti

机构 * University of Brescia, Italy(意大利布雷西亚大学) Imperial College London, UK(伦敦帝国理工学院) Cardiff University, UK(卡迪夫大学) University of Southampton, UK(南安普顿大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments Accepted at the 2nd edition of the Workshop in AI and Finance at ECAI-2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21121 2025-10-27 cs.RO cs.AI 70%

Generalizable Hierarchical Skill Learning via Object-Centric Representation

Haibo Zhao, Yu Qi, Boce Hu, Yizhe Zhu, Ziyan Chen, Heng Tian, Xupeng Zhu, Owen Howell, Haojie Huang, Robin Walters, Dian Wang, Robert Platt

专题命中 知识编辑与模型理解 :language model(abstract);foundation model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13737 2025-10-27 cs.AI 70%

Causal Head Gating: A Framework for Interpreting Roles of Attention Heads in Transformers

Andrew Nam, Henry Conklin, Yukang Yang, Thomas Griffiths, Jonathan Cohen, Sarah-Jane Leslie

机构 * Princeton Laboratory for AI Natural and Artificial Minds(普林斯顿人工智能实验室) Princeton University(普林斯顿大学) Department of Electrical and Computer Engineering(电气与计算机工程系) Department of Psychology(心理学系) Princeton Neuroscience Institute(普林斯顿神经科学研究所) Department of Philosophy(哲学系) Center for Statistics and Machine Learning(统计与机器学习中心)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 10 pages, 5 figures, 2 tables. The Thirty-Ninth Annual Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21119 2025-10-27 stat.ME stat.ML 67%

Leveraging semantic similarity for experimentation with AI-generated treatments

Lei Shi, David Arbour, Raghavendra Addanki, Ritwik Sinha, Avi Feller

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 31 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13345 2025-10-27 cs.CY cs.AI cs.CL cs.HC cs.LG 67%

Beyond Accuracy: Rethinking Hallucination and Regulatory Response in Generative AI

Zihao Li, Weiwei Yi, Jiahong Chen

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21323 2025-10-27 cs.CV cs.LG 57%

VL-SAE: Interpreting and Enhancing Vision-Language Alignment with a Unified Concept Set

Shufan Shen, Junshu Sun, Qingming Huang, Shuhui Wang

机构 * Key Lab of Intell. Info. Process., Inst. of Comput. Tech., CAS(智能信息处理重点实验室,计算技术研究所,中国科学院) University of Chinese Academy of Sciences(中国科学院大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02964 2025-10-27 cs.CV cs.LG 57%

FORLA: Federated Object-centric Representation Learning with Slot Attention

Guiqiu Liao, Matjaz Jogan, Eric Eaton, Daniel A. Hashimoto

机构 * PCASO Laboratory, Dept. of Surgery, University of Pennsylvania(宾夕法尼亚大学外科部PCASO实验室) Dept. of Computer and Information Science, University of Pennsylvania(宾夕法尼亚大学计算机与信息科学系)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

Comments Accepted by Neurips2025

详情

展开后加载摘要…

URL PDF HTML 收藏