arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-30 至 2025-10-30 共收录 164 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 9 篇

2505.22586 2025-10-30 cs.CL 89%

Precise In-Parameter Concept Erasure in Large Language Models

Yoav Gur-Arieh, Clara Suslik, Yihuai Hong, Fazl Barez, Mor Geva

机构 * Blavatnik School of Computer Science and AI, Tel Aviv University(塔尔斯基大学计算机科学与人工智能学院) New York University(纽约大学) University of Oxford & WhiteBox(牛津大学及WhiteBox)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);pretraining(abstract);分类 cs.CL

Comments Accepted to EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24966 2025-10-30 cs.LG cs.AI cs.CL stat.ML 85%

Sequences of Logits Reveal the Low Rank Structure of Language Models

Noah Golowich, Allen Liu, Abhishek Shetty

机构 * Microsoft Research(微软研究院) UC Berkeley(加州大学伯克利分校) MIT(麻省理工学院)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25368 2025-10-30 cs.LG cs.AI cs.NE 79%

Position: Biology is the Challenge Physics-Informed ML Needs to Evolve

Julien Martinelli

机构 * ELLIS Institute Finland(芬兰ELLIS研究所) Department of Computer Science, Aalto University(艾尔沃斯大学计算机科学系)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);foundation model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25175 2025-10-30 cs.CV 78%

Test-Time Adaptive Object Detection with Foundation Model

Yingjie Gao, Yanan Zhang, Zhi Cai, Di Huang

机构 * State Key Laboratory of Complex and Critical Software Environment, Beihang University(复杂与关键软件环境国家重点实验室,北京航空航天大学) School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院) School of Computer Science and Information Engineering, Hefei University of Technology(合肥工业大学计算机科学与信息工程学院)

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15356 2025-10-30 cs.CL 77%

NL-Debugging: Exploiting Natural Language as an Intermediate Representation for Code Debugging

Weiming Zhang, Qingyao Li, Xinyi Dai, Jizheng Chen, Kounianhua Du, Weiwen Liu, Yasheng Wang, Ruiming Tang, Yong Yu, Weinan Zhang

机构 * Shanghai Jiao Tong University(上海交通大学) Huawei Noah’s Ark Lab Shanghai(华为诺亚实验室)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.05583 2025-10-30 cs.CL 77%

OpenFactCheck: Building, Benchmarking Customized Fact-Checking Systems and Evaluating the Factuality of Claims and LLMs

Yuxia Wang, Minghan Wang, Hasan Iqbal, Georgi Georgiev, Jiahui Geng, Preslav Nakov

机构 * MBZUAI Monash University(墨尔本大学) Sofia University(索菲亚大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments 23 pages, 8 tables, 11 figures, Published In Proceedings of the 31st International Conference on Computational Linguistics 2025

Journal ref In Proceedings of the 31st International Conference on Computational Linguistics 2025, pages 11399-11421, Abu Dhabi, UAE. Association for Computational Linguistics

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25488 2025-10-30 cs.IR 75%

Generalized Pseudo-Relevance Feedback

Yiteng Tu, Weihang Su, Yujia Zhou, Yiqun Liu, Fen Lin, Qin Liu, Qingyao Ai

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18376 2025-10-30 cs.LG cs.SI 70%

GnnXemplar: Exemplars to Explanations -- Natural Language Rules for Global GNN Interpretability

Burouj Armgaan, Eshan Jain, Harsh Pandey, Mahesh Chandran, Sayan Ranu

机构 * Dept. of CSE, IIT Delhi(印度德里理工学院计算机科学与工程系) Fujitsu Research of India, Bangalore(印度班加罗尔富士通印度研究机构) Dept. of CSE and Yardi ScAI, IIT Delhi(印度德里理工学院计算机科学与工程系)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments 38 pages, 20 figures, NeurIPS 2025 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24813 2025-10-30 cs.CV cs.AI 57%

DualCap: Enhancing Lightweight Image Captioning via Dual Retrieval with Similar Scenes Visual Prompts

Binbin Li, Guimiao Yang, Zisen Qi, Haiping Wang, Yu Ding

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 其他LLM 5 篇

2510.14205 2025-10-30 cs.CL cs.AI 86%

DPRF: A Generalizable Dynamic Persona Refinement Framework for Optimizing Behavior Alignment Between Personalized LLM Role-Playing Agents and Humans

Bingsheng Yao, Bo Sun, Yuanzhe Dong, Yuxuan Lu, Dakuo Wang

机构 * Northeastern University(东北大学) Stanford University(斯坦福大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments In Submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09158 2025-10-30 cs.CL 80%

Augmenting Dialog with Think-Aloud Utterances for Modeling Individual Personality Traits by LLM

Seiya Ishikura, Hiroaki Yamada, Tatsuya Hiraoka, Hiroaki Yamada, Takenobu Tokunaga

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

Comments 8 pages, 1 figure. Accepted at the First Workshop on Tailoring AI: Exploring Active and Passive LLM Personalization (PALS2025@EMNLP2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24920 2025-10-30 cs.CR 67%

S3C2 Summit 2025-03: Industry Secure Supply Chain Summit

Elizabeth Lin, Jonah Ghebremichael, William Enck, Yasemin Acar, Michel Cukier, Alexandros Kapravelos, Christian Kastner, Laurie Williams

专题命中 其他LLM :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24934 2025-10-30 cs.CL 57%

Disaggregation Reveals Hidden Training Dynamics: The Case of Agreement Attraction

James A. Michaelov, Catherine Arnett

机构 * MIT(麻省理工学院) EleutherAI

专题命中 其他LLM :language model(abstract);分类 cs.CL

Comments Accepted to the First Workshop on Interpreting Cognition in Deep Learning Models (CogInterp @ NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24785 2025-10-30 eess.IV cs.IT math.IT 50%

Semantic Communications with World Models

Peiwen Jiang, Jiajia Guo, Chao-Kai Wen, Shi Jin, Jun Zhang

专题命中 其他LLM :foundation model(abstract)

Comments This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏