arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-04 至 2025-11-04 共收录 17 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 17 篇

2510.05702 2025-11-04 q-fin.CP cs.AI 90%

Uncovering Representation Bias for Investment Decisions in Open-Source Large Language Models

Fabrizio Dimino, Krati Saxena, Bhaskarjit Sarmah, Stefano Pasquali

机构 * Domyn

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00136 2025-11-04 cs.LG cs.AI 90%

A Dual Large Language Models Architecture with Herald Guided Prompts for Parallel Fine Grained Traffic Signal Control

Qing Guo, Xinhang Li, Junyu Chen, Zheng Guo, Xiaocong Li, Lin Zhang, Lei Li

机构 * School of Artificial Intelligence, Beijing University of Posts(人工智能学院,北京邮电大学) Beijing Big Data Center(北京大数据中心) Eastern Institute of Technology(东部技术研究所)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08809 2025-11-04 cs.LG 89%

Decoupling Contrastive Decoding: Robust Hallucination Mitigation in Multimodal Large Language Models

Wei Chen, Xin Yan, Bin Wen, Fan Yang, Tingting Gao, Di Zhang, Long Chen

机构 * HKUST(香港科技大学) University of Waterloo(滑铁卢大学) Kuaishou Technology(快手科技)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);preference optimization(abstract);分类 cs.LG

Comments 17 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10613 2025-11-04 cs.CL cs.AI 88%

Dynamic Topic Evolution with Temporal Decay and Attention in Large Language Models

Di Wu, Shuaidong Pan

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01363 2025-11-04 cs.AI 88%

Automatic Minds: Cognitive Parallels Between Hypnotic States and Large Language Model Processing

Giuseppe Riva, Brenda K. Wiederhold, Fabrizia Mantovani

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

Comments 4 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21590 2025-11-04 cs.CL cs.LG 88%

Representation Consistency for Accurate and Coherent LLM Answer Aggregation

Junqi Jiang, Tom Bewley, Salim I. Amoukou, Francesco Leofante, Antonio Rago, Saumitra Mishra, Francesca Toni

机构 * Imperial College London(帝国理工学院伦敦分校) J.P. Morgan AI Research(摩根大通人工智能研究) King’s College London(国王学院伦敦分校)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Accepted at NeurIPS 2025. Camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14314 2025-11-04 cs.CL cs.AI cs.LG 87%

Zero-knowledge LLM hallucination detection and mitigation through fine-grained cross-model consistency

Aman Goel, Daniel Schwartz, Yanjun Qi

机构 * Amazon Web Services, USA(亚马逊网络服务)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01311 2025-11-04 cs.AI 85%

llmSHAP: A Principled Approach to LLM Explainability

Filip Naudot, Tobias Sundqvist, Timotheus Kampik

机构 * Umeå University(乌梅大学) Tietoevry(蒂奥埃弗里) SAP(SAP公司)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14077 2025-11-04 cs.CL 85%

ERGO: Entropy-guided Resetting for Generation Optimization in Multi-turn Language Models

Haziq Mohammad Khalid, Athikash Jeyaganthan, Timothy Do, Yicheng Fu, Sean O'Brien, Vasu Sharma, Kevin Zhu

机构 * Algoverse AI Research(Algoverse AI研究院)

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL

Comments 14 pages, 5 figures

Journal ref Proceedings of the 2nd Workshop on Uncertainty Aware NLP (UncertaiNLP 2025), Suzhou, China, Association for Computational Linguistics, pp. 273--286, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00443 2025-11-04 cs.LG cs.AI cs.CV 84%

Region-Aware Reconstruction Strategy for Pre-training fMRI Foundation Model

Ruthwik Reddy Doodipala, Pankaj Pandey, Carolina Torres Rojas, Manob Jyoti Saikia, Ranganatha Sitaram

机构 * St. Jude Children’s Research Hospital(圣犹大儿童研究医院) The University of Memphis(密苏里大学梅尔斯分校)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);pretraining(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00620 2025-11-04 cs.CL 83%

Certain but not Probable? Differentiating Certainty from Probability in LLM Token Outputs for Probabilistic Scenarios

Autumn Toney-Wails, Ryan Wails

机构 * SciTech Strategies, Inc.(SciTech Strategies公司) Georgetown University(乔治城大学)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

Comments To appear at the Second Workshop on Uncertainty-Aware NLP @EMNLP 2025 (UncertaiNLP '25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00180 2025-11-04 cs.CL cs.LG 81%

ParaScopes: What do Language Models Activations Encode About Future Text?

Nicky Pochinkov, Yulia Volkova, Anna Vasileva, Sai V R Chereddy

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

Comments Main paper: 9 pages, 10 figures. Total 24 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00419 2025-11-04 cs.CV cs.AI 77%

LGCA: Enhancing Semantic Representation via Progressive Expansion

Thanh Hieu Cao, Trung Khang Tran, Gia Thinh Pham, Tuong Nghiem Diep, Thanh Binh Nguyen

机构 * University of Science, Vietnam National University Ho Chi Minh City, Vietnam(越南胡志明市国家大学科学大学) National University of Singapore, Singapore(新加坡国立大学) AISIA Lab, Ho Chi Minh City, Vietnam(越南胡志明市AISIA实验室)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);pretraining(abstract);分类 cs.AI

Comments 15 pages, 5 figures, to appear in SoICT 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01610 2025-11-04 cs.CV cs.AI 57%

DINO-MX: A Modular & Flexible Framework for Self-Supervised Learning

Mahmut Selman Gokmen, Cody Bumgardner

机构 * Computer Science Department University of Kentucky(卡内基梅隆大学计算机科学系)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26466 2025-11-04 cs.CV cs.LG 57%

Representation-Level Counterfactual Calibration for Debiased Zero-Shot Recognition

Pei Peng, MingKun Xie, Hang Hao, Tong Jin, ShengJun Huang

机构 * Nanjing University of Aeronautics and Astronautics(南京航空航天大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03690 2025-11-04 cs.LG 57%

Graph Neural Networks for Electricity Load Forecasting

Eloi Campagne, Yvenn Amara-Ouali, Yannig Goude, Itai Zehavi, Argyris Kalogeratos

机构 * Centre Borelli, Ecole Normale Supérieure Paris-Saclay, Gif-sur-Yvette, France(巴黎萨克雷大学Borelli中心) EDF Lab, Palaiseau, France(法国电力公司实验室)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

Comments 22 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00785 2025-11-04 cs.CV cs.AI 57%

Class-agnostic 3D Segmentation by Granularity-Consistent Automatic 2D Mask Tracking

Juan Wang, Yasutomo Kawanishi, Tomo Miyazaki, Zhijie Wang, Shinichiro Omachi

机构 * Graduate School of Engineering(工程研究生院) Multimodal Data Recognition Research Team(多模态数据识别研究团队) RIKEN GRP(日本理化学研究所)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI

Comments Under review in Pattern Recognition

详情

展开后加载摘要…

URL PDF HTML 收藏