arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7552 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7552 篇

2509.23434 2025-09-30 cs.HC cs.AI 70%

NeuroBridge: Using Generative AI to Bridge Cross-neurotype Communication Differences through Neurotypical Perspective-taking

Rukhshan Haroon, Kyle Wigdor, Katie Yang, Nicole Toumanios, Eileen T. Crehan, Fahad Dogar

机构 * Computer Science Tufts University(计算机科学 华盛顿大学) Human Development, Cognitive Science Tufts University(人类发展与认知科学 华盛顿大学) Eunice Kennedy Shriver Center UMass Chan Medical School(欧尼丝·凯瑟琳·施里弗中心 马萨诸塞大学医学学院) Tufts University(华盛顿大学) UMass Chan Medical School(马萨诸塞大学医学学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.13111 2025-09-30 cs.LG cs.RO 70%

Uncertainty-Aware Trajectory Prediction via Rule-Regularized Heteroscedastic Deep Classification

Kumar Manas, Christian Schlauch, Adrian Paschke, Christian Wirth, Nadja Klein

机构 * Department of Mathematics and Computer Science, Freie Universität Berlin(自由大学柏林数学与计算机科学系) Continental Automotive Technologies GmbH, AI Lab Berlin(大陆汽车技术有限公司柏林AI实验室) Karlsruhe Institute of Technology, Scientific Computing Center, Methods for Big Data(卡尔斯鲁厄理工学院科学计算中心、大数据方法) Fraunhofer Institute for Open Communication Systems, Berlin, Germany(弗劳恩霍夫开放通信系统研究所,柏林德国)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments 17 Pages, 9 figures. Accepted to Robotics: Science and Systems(RSS), 2025

Journal ref Robotics: Science and Systems (RSS), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17445 2025-09-29 cs.CL 70%

Semantic Reformulation Entropy for Robust Hallucination Detection in QA Tasks

Chaodong Tong, Qi Zhang, Lei Jiang, Yanbing Liu, Nannan Sun, Wei Li

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments 5pages, 5 figures, submitted to ICASSP 2026,

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11029 2025-09-29 cs.LG 70%

Exploiting the Asymmetric Uncertainty Structure of Pre-trained VLMs on the Unit Hypersphere

Li Ju, Max Andersson, Stina Fredriksson, Edward Glöckner, Andreas Hellander, Ekta Vats, Prashant Singh

机构 * Department of Information Technology, Uppsala University(信息科技系,乌普萨拉大学) Science for Life Laboratory, Uppsala University(生命科学实验室,乌普萨拉大学)

专题命中 知识编辑与模型理解 :language model(abstract);foundation model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20168 2025-09-25 cs.CL 70%

Probing Gender Bias in Multilingual LLMs: A Case Study of Stereotypes in Persian

Ghazal Kalhor, Behnam Bahrak

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted and forthcoming at the Widening Natural Language Processing Workshop (WiNLP 2025) at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18568 2025-09-24 cs.LG 70%

Explainable Graph Neural Networks: Understanding Brain Connectivity and Biomarkers in Dementia

Niharika Tewari, Nguyen Linh Dan Le, Mujie Liu, Jing Ren, Ziqi Xu, Tabinda Sarwar, Veeky Baths, Feng Xia

机构 * School of Computing Technologies RMIT University Melbourne VIC Australia Department of Biological Sciences Department of Computer Science \& Information Systems Birla Institute of Technology Institute of Innovation, Science Sustainability Federation University Australia Ballarat VIC Australia RMIT University Birla Institute of Technology Federation University Australia

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15655 2025-09-22 cs.CL eess.AS 70%

Layer-wise Minimal Pair Probing Reveals Contextual Grammatical-Conceptual Hierarchy in Speech Representations

Linyang He, Qiaolin Wang, Xilin Jiang, Nima Mesgarani

机构 * Department of Electrical Engineering, Columbia University(哥伦比亚大学电气工程系)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments EMNLP 2025 Main Conference (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21436 2025-09-22 cs.CL 70%

Discovering Semantic Subdimensions through Disentangled Conceptual Representations

Yunhao Zhang, Shaonan Wang, Nan Lin, Xinyi Dong, Chong Li, Chengqing Zong

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, CAS(多模态人工智能系统国家重点实验室,自动化研究所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) State Key Laboratory of Cognitive Science and Mental Health, Institute of Psychology, CAS(认知科学与心理健康国家重点实验室,心理研究所) Department of Psychology, University of Chinese Academy of Sciences(中国科学院大学心理学系) State Key Laboratory of Cognitive Neuroscience and Learning, Beijing Normal University(认知神经科学与学习国家重点实验室,北京师范大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14956 2025-09-19 cs.AI cs.MA 70%

Sentinel Agents for Secure and Trustworthy Agentic AI in Multi-Agent Systems

Diego Gosmar, Deborah A. Dahl

机构 * Head of AI Tesisquare(Tesisquare人工智能负责人) Voiceinteroperability.ai Initiative Member(Voiceinteroperability.ai 创始成员) Linux Foundation AI & Data(Linux基金会人工智能与数据) AI & Data Torino, TO 10100, Italy(人工智能与数据Torino, TO 10100, Italy) Principal Conversational Technologies(对话技术首席专家) Plymouth Meeting, Pennsylvania, USA(美国宾夕法尼亚州Plymouth Meeting)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 25 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11277 2025-09-19 cs.CV cs.LG 70%

Probing the Representational Power of Sparse Autoencoders in Vision Models

Matthew Lyle Olson, Musashi Hinck, Neale Ratzlaff, Changbai Li, Phillip Howard, Vasudev Lal, Shao-Yen Tseng

机构 * Oracle Intel Labs(英特尔实验室) Oregon State University(俄勒冈州立大学) Thoughtworks

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments ICCV 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19721 2025-09-19 cs.CL cs.CY 70%

Unsupervised Concept Vector Extraction for Bias Control in LLMs

Hannah Cyberey, Yangfeng Ji, David Evans

机构 * Department of Computer Science University of Virginia(计算机科学系 首都维吉尼亚大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20409 2025-09-17 cs.CL 70%

TAPS: Tool-Augmented Personalisation via Structured Tagging

Ekaterina Taktasheva, Jeff Dalton

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to EMNLP 2026 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11915 2025-09-16 cs.CL 70%

Uncertainty in Authorship: Why Perfect AI Detection Is Mathematically Impossible

Aadil Gani Ganie

机构 * UNIVERSITAT POLITECNICA DE VALENCIA(瓦伦西亚理工大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11569 2025-09-16 cs.CL 70%

D$^2$HScore: Reasoning-Aware Hallucination Detection via Semantic Breadth and Depth Analysis in LLMs

Yue Ding, Xiaofang Zhu, Tianze Xia, Junfei Wu, Xinlong Chen, Qiang Liu, Liang Wang

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07617 2025-09-10 cs.AI 70%

Transferable Direct Prompt Injection via Activation-Guided MCMC Sampling

Minghui Li, Hao Zhang, Yechao Zhang, Wei Wan, Shengshan Hu, pei Xiaobing, Jing Wang

机构 * School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件学院) School of Cyber Science and Engineering, Huazhong University of Science and Technology(华中科技大学网络安全学院) College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院) Faculty of Data Science, City University of Macau(澳门城市大学数据科学学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09297 2025-09-10 cs.LG 70%

When Do Neural Networks Learn World Models?

Tianren Zhang, Guanyu Chen, Feng Chen

机构 * Department of Automation, Tsinghua University, Beijing, China(自动化系,清华大学,北京,中国)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments ICML 2025; ICLR 2025 World Models Workshop (oral, outstanding paper award)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05714 2025-09-09 cs.AI cs.CV 70%

Towards Meta-Cognitive Knowledge Editing for Multimodal LLMs

Zhaoyu Fan, Kaihang Pan, Mingze Zhou, Bosheng Qin, Juncheng Li, Shengyu Zhang, Wenqiao Zhang, Siliang Tang, Fei Wu, Yueting Zhuang

机构 * Zhejiang University(浙江大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 15 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04905 2025-09-08 cs.LG 70%

Revolution or Hype? Seeking the Limits of Large Models in Hardware Design

Qiang Xu, Leon Stok, Rolf Drechsler, Xi Wang, Grace Li Zhang, Igor L. Markov

机构 * The Chinese University of Hong Kong(香港中文大学) Southeast University(东南大学) IBM TU Darmstadt(图尔纳大学) University of Bremen/DFKI(不莱梅大学/DFKI) Synopsys

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments Invited paper to appear at ICCAD'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03626 2025-09-05 cs.AI 70%

Explainable Knowledge Graph Retrieval-Augmented Generation (KG-RAG) with KG-SMILE

Zahra Zehtabi Sabeti Moghaddam, Zeinab Dehghani, Maneeha Rani, Koorosh Aslansefat, Bhupesh Kumar Mishra, Rameez Raja Kureshi, Dhavalkumar Thakker

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.03951 2025-09-05 cs.LG 70%

Uncertainty-Guided Likelihood Tree Search

Julia Grosse, Ruotian Wu, Ahmad Rashid, Cheng Zhang, Philipp Hennig, Pascal Poupart, Agustinus Kristiadi

机构 * University of Tübingen(图宾根大学) Tübingen AI Center(图宾根人工智能中心) University of Waterloo(滑铁卢大学) Vector Institute(向量研究所) GenAI at Meta(Meta的生成式人工智能部门)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21484 2025-09-03 q-bio.QM cs.LG stat.ML 70%

Data-driven Discovery of Digital Twins in Biomedical Research

Clémence Métayer, Annabelle Ballesta, Julien Martinelli

机构 * Inserm U1331, Institut Curie Saint-Cloud, France(法国国家医学研究院U1331,圣克鲁医院) Aalto University Espoo, Finland(芬兰艾尔托大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05200 2025-08-29 cs.LG math.ST stat.ML stat.TH 70%

Transformers Meet In-Context Learning: A Universal Approximation Theory

Gen Li, Yuchen Jiao, Yu Huang, Yuting Wei, Yuxin Chen

机构 * Department of Statistics and Data Science, Chinese University of Hong Kong(统计与数据科学系,香港中文大学) the Wharton School, University of Pennsylvania(宾夕法尼亚大学沃顿商学院) University of Pennsylvania(宾夕法尼亚大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19322 2025-08-28 eess.IV cs.AI cs.CV 70%

AT-CXR: Uncertainty-Aware Agentic Triage for Chest X-rays

Xueyang Li, Mingze Jiang, Gelei Xu, Jun Xia, Mengzhao Jia, Danny Chen, Yiyu Shi

机构 * University of Notre Dame(诺丁汉大学)

专题命中 知识编辑与模型理解 :LLM(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19099 2025-08-27 cs.CL 70%

Beyond the Black Box: Integrating Lexical and Semantic Methods in Quantitative Discourse Analysis with BERTopic

Thomas Compton

机构 * University of York(约克大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments 5 pages conference paper, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19096 2025-08-27 cs.AI 70%

Trustworthy Agents for Electronic Health Records through Confidence Estimation

Yongwoo Song, Minbyul Jeong, Mujeen Sung

机构 * Kyung Hee University(庆熙大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18953 2025-08-27 cs.AI 70%

Novel Approaches to Artificial Intelligence Development Based on the Nearest Neighbor Method

I. I. Priezzhev, D. A. Danko, A. V. Shubin

机构 * National University of Oil and Gas «Gubkin University»(石油国家大学「古比金大学」) IPLab LLC(IPLab公司)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 18 pages, 6 figures. Novel hierarchical neural networks based on k-nearest neighbors method for addressing hallucination effects, training complexity, and catastrophic forgetting in modern AI systems. Includes mathematical formulations using Kohonen self-organizing maps and experimental validation on MNIST handwritten digit recognition and machine translation tasks

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18001 2025-08-26 cs.LG stat.ML 70%

A Novel Framework for Uncertainty Quantification via Proper Scores for Classification and Beyond

Sebastian G. Gruber

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments PhD Thesis (cumulative, spanning 6 peer-reviewed publications)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17760 2025-08-26 cs.CV cs.CL 70%

CEIDM: A Controlled Entity and Interaction Diffusion Model for Enhanced Text-to-Image Generation

Mingyue Yang, Dianxi Shi, Jialu Zhou, Xinyu Wei, Leqian Li, Shaowu Yang, Chunping Qiu

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15977 2025-08-25 cs.CL 70%

Dancing with Deer: A Constructional Perspective on MWEs in the Era of LLMs

Claire Bonial, Julia Bonn, Harish Tayyar Madabushi

机构 * U.S. Army Research Lab(美国陆军研究实验室) University of Colorado Boulder(科罗拉多大学博尔德分校) University of Bath(巴斯大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Chapter in Phraseology and Multiword Expressions, Language Science Press (to appear)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15685 2025-08-22 cs.AR cs.AI 70%

Row-Column Hybrid Grouping for Fault-Resilient Multi-Bit Weight Representation on IMC Arrays

Kang Eun Jeon, Sangheum Yeon, Jinhee Kim, Hyeonsu Bang, Johnny Rhe, Jong Hwan Ko

机构 * Department of Electrical and Computer Engineering, Sungkyunkwan University(电气与计算机工程系,成均馆大学) Department of Semiconductor Engineering, Sungkyunkwan University(半导体工程系,成均馆大学) Department of Electrical and Computer Engineering, Duke University(电气与计算机工程系,杜克大学) Kim Jaechul Graduate School of AI, Korea Advanced Institute of Science and Technology(金 Jaechul人工智能研究生院,韩国科学技术院)

专题命中 知识编辑与模型理解 :language model(abstract);small language model(abstract);分类 cs.AI

Comments Accepted to appear at ICCAD'25 (Munich, Germany)

详情

展开后加载摘要…

URL PDF HTML 收藏