arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-11 至 2025-11-11 共收录 24 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 24 篇

2511.04053 2025-11-11 cs.AI 89%

Interpreting Multi-Attribute Confounding through Numerical Attributes in Large Language Models

Hirohane Takagi, Gouki Minegishi, Shota Kizawa, Issey Sukeda, Hitomi Yanaka

机构 * The University of Tokyo(东京大学) RIKEN(日本研究机构) Tohoku University(东北大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

Comments Accepted to IJCNLP-AACL 2025 (Main). Code available at https://github.com/htkg/num_attrs

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.08878 2025-11-11 cs.NI cs.IT cs.LG eess.SP math.IT 89%

Large Language Model Empowered Next-Generation MIMO Networks: Fundamentals, Challenges, and Visions

Zhe Wang, Jiayi Zhang, Hongyang Du, Ruichen Zhang, Dusit Niyato, Bo Ai, Khaled B. Letaief

机构 * State Key Laboratory of Advanced Rail Autonomous Operation(先进轨道交通自主运行状态关键实验室) Beijing Jiaotong University(北京交通大学) School of Electronics and Information Engineering(电子与信息工程学院) University of Hong Kong(香港大学) College of Computing & Data Science(计算与数据科学学院) Nanyang Technological University(南洋理工大学) Hong Kong University of Science and Technology(香港科技大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

Comments 9 pages, 4 figures, 1 table, to appear in Digital Communications and Networks

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02904 2025-11-11 cs.CL cs.AI cs.LG 88%

How Post-Training Reshapes LLMs: A Mechanistic View on Knowledge, Truthfulness, Refusal, and Confidence

Hongzhe Du, Weikai Li, Min Cai, Karim Saraipour, Zimin Zhang, Himabindu Lakkaraju, Yizhou Sun, Shichang Zhang

机构 * University of California, Los Angeles(加州大学洛杉矶分校) University of Alberta(阿尔伯塔大学) University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Harvard University(哈佛大学)

专题命中 知识编辑与模型理解 :post-training(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

Comments COLM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06475 2025-11-11 cs.CV 88%

NOAH: Benchmarking Narrative Prior driven Hallucination and Omission in Video Large Language Models

Kyuho Lee, Euntae Kim, Jinwoo Choi, Buru Chang

机构 * Korea University(韩国大学) Kyung Hee University(庆熙大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

Comments 18 pages, 9 figures. Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06700 2025-11-11 cs.CY cs.AI cs.CL 86%

Place Matters: Comparing LLM Hallucination Rates for Place-Based Legal Queries

Damian Curran, Vanessa Sporne, Lea Frermann, Jeannie Paterson

机构 * School of Computing and Information Systems, The University of Melbourne, Australia(计算与信息系统学院,墨尔本大学,澳大利亚) The Centre for Artificial Intelligence and Digital Ethics(人工智能与数字伦理中心) Melbourne Law School, The University of Melbourne, Australia(墨尔本法学院,墨尔本大学,澳大利亚)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07083 2025-11-11 cs.AI 85%

Increasing AI Explainability by LLM Driven Standard Processes

Marc Jansen, Marcel Pehlke

机构 * Computer Science Institute University of Applied Sciences Ruhr West Bottrop(应用科学鲁尔西贝特大学计算机科学研究所)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14496 2025-11-11 cs.AI 85%

LLM-Powered Swarms: A New Frontier or a Conceptual Stretch?

Muhammad Atta Ur Rahman, Melanie Schranz, Samira Hayat

机构 * Lakeside Labs GmbH(莱克西德实验室)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments This is the author's version of a paper submitted to IEEE Intelligent Systems. 2 Tables, 2 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06073 2025-11-11 cs.CL cs.AI cs.LG cs.LO 83%

Stemming Hallucination in Language Models Using a Licensing Oracle

Simeon Emanuilov, Richard Ackermann

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments 23 pages, 4 figures, 8 tables. Introduces the Licensing Oracle, an architectural solution for eliminating hallucinations in language models through formal SHACL validation against knowledge graphs. All datasets and models are available at https://huggingface.co/collections/s-emanuilov/licensing-oracle-experiments

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05516 2025-11-11 cs.CL cs.AI cs.SD eess.AS 81%

Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation

Canxiang Yan, Chunxiang Jin, Dawei Huang, Haibing Yu, Han Peng, Hui Zhan, Jie Gao, Jing Peng, Jingdong Chen, Jun Zhou, Kaimeng Ren, Ming Yang, Mingxue Yang, Qiang Xu, Qin Zhao, Ruijie Xiong, Shaoxiong Lin, Xuezhi Wang, Yi Yuan, Yifei Wu, Yongjie Lyu, Zhengyu He, Zhihao Qiu, Zhiqiang Fang, Ziyuan Huang

机构 * Inclusion AI Ant Group(Inclusion AI Ant集团)

专题命中 知识编辑与模型理解 :LLM(title);language model(abstract);分类 cs.CL、cs.AI

Comments 32 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16795 2025-11-11 cs.LG cs.AI cs.CL 80%

Steering Out-of-Distribution Generalization with Concept Ablation Fine-Tuning

Helena Casademunt, Caden Juang, Adam Karvonen, Samuel Marks, Senthooran Rajamanoharan, Neel Nanda

机构 * Harvard University(哈佛大学) Northeastern University(东北大学) Anthropic ML Alignment & Theory Scholars (MATS) program(ML对齐与理论学者(MATS)计划)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06496 2025-11-11 cs.RO cs.AI cs.CV 79%

A Low-Rank Method for Vision Language Model Hallucination Mitigation in Autonomous Driving

Keke Long, Jiacheng Guo, Tianyun Zhang, Hongkai Yu, Xiaopeng Li

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11881 2025-11-11 cs.CL 79%

Evaluating Human-LLM Representation Alignment: A Case Study on Affective Sentence Generation for Augmentative and Alternative Communication

Shadab Choudhury, Asha Kumar, Lara J. Martin

专题命中 知识编辑与模型理解 :LLM(title);language model(abstract);分类 cs.CL

Comments Published at IJCNLP-AACL 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05498 2025-11-11 cs.IR cs.AI 77%

Biomedical Hypothesis Explainability with Graph-Based Context Retrieval

Ilya Tyagin, Saeideh Valipour, Aliaksandra Sikirzhytskaya, Michael Shtutman, Ilya Safro

机构 * Center for Bioinformatics and Computational Biology University of Delaware(生物信息学与计算生物学中心 大田纳西大学) Computer and Information Sciences University of Delaware(计算机与信息科学 大田纳西大学) Drug Discovery and Biomedical Sciences (DDBS) College of Pharmacy University of South Carolina(药物发现与生物医学科学(DDBS)药学院 美国南卡罗来纳大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 30 pages, 10 figures,

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06048 2025-11-11 cs.CL cs.LG 73%

Visual Exploration of Feature Relationships in Sparse Autoencoders with Curated Concepts

Xinyuan Yan, Shusen Liu, Kowshik Thopalli, Bei Wang

机构 * University of Utah(犹他大学) Lawrence Livermore National Laboratory(劳伦斯利弗莫尔国家实验室)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments 8 pages (5 main paper+3 refernce), 2 figures, pulished at Mechanistic Interpretability Workshop at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20264 2025-11-11 cs.CL 70%

EMBRACE: Shaping Inclusive Opinion Representation by Aligning Implicit Conversations with Social Norms

Abeer Aldayel, Areej Alokaili

机构 * King Saud University, College of Computer and Information Sciences(沙特王后大学,计算机与信息科学学院)

专题命中 知识编辑与模型理解 :language model(abstract);post-training(abstract);分类 cs.CL

Comments Accepted, to appear IJCNLP-AACL 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07863 2025-11-11 cs.LG 70%

Robust Hallucination Detection in LLMs via Adaptive Token Selection

Mengjia Niu, Hamed Haddadi, Guansong Pang

机构 * Imperial College London, UK(伦敦帝国学院) Singapore Management University(新加坡管理大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16678 2025-11-11 cs.CL 70%

Mechanisms vs. Outcomes: Probing for Syntax Fails to Explain Performance on Targeted Syntactic Evaluations

Ananth Agarwal, Jasper Jian, Christopher D. Manning, Shikhar Murty

机构 * Stanford University(斯坦福大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Journal ref Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21412 2025-11-11 cs.CV 67%

Bridging the gap to real-world language-grounded visual concept learning

Whie Jung, Semin Kim, Junee Kim, Seunghoon Hong

机构 * School of Computing, KAIST(计算机学院,韩国科学技术院)

专题命中 知识编辑与模型理解 :language model(abstract);prompting(abstract)

Journal ref Advances in Neural Information Processing Systems (NeurIPS), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07006 2025-11-11 cs.LG cs.AI 62%

S$^2$Drug: Bridging Protein Sequence and 3D Structure in Contrastive Representation Learning for Virtual Screening

Bowei He, Bowen Gao, Yankai Chen, Yanyan Lan, Chen Ma, Philip S. Yu, Ya-Qin Zhang, Wei-Ying Ma

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.AI、cs.LG

Comments Accepted by AAAI 2026 Main Technical Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17959 2025-11-11 astro-ph.IM cs.AI cs.LG 62%

Universal Spectral Tokenization via Self-Supervised Panchromatic Representation Learning

Jeff Shen, Francois Lanusse, Liam Holden Parker, Ollie Liu, Tom Hehir, Leopoldo Sarra, Lucas Meyer, Micah Bowles, Sebastian Wagner-Carena, Sebastian Wagner-Carena, Helen Qu, Siavash Golkar, Alberto Bietti, Hatim Bourfoune, Nathan Cassereau, Pierre Cornette, Keiya Hirashima, Geraud Krawezik, Ruben Ohana, Nicholas Lourie, Michael McCabe, Rudy Morel, Payel Mukhopadhyay, Mariel Pettee, Bruno Régaldo-Saint Blancard, Kyunghyun Cho, Miles Cranmer, Shirley Ho

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

Comments Accepted at NeurIPS 2025 Machine Learning and the Physical Sciences Workshop; v2: added collaboration

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01334 2025-11-11 cs.CL cs.AI 62%

Skill Path: Unveiling Language Skills from Circuit Graphs

Hang Chen, Jiaying Zhu, Xinyu Yang, Wenya Wang

机构 * Hang Chen School of Computer Science and Technology Xi’an Jiaotong University(陈hang 王校计算机科学与技术学院 西安交通大学) Jiaying Zhu School of Computer Science and Engineering The Chinese University of Hong Kong(朱继燕 王校计算机科学与工程学院 香港中文大学) Xinyu Yang School of Computer Science and Technology Xi’an Jiaotong University(杨新宇 王校计算机科学与技术学院 西安交通大学) Wenya Wang School of Computer Science and Engineering Nanyang Technological University(王文雅 王校计算机科学与工程学院 新加坡国立大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments accepted by AAAI 2026 (oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07068 2025-11-11 cs.CV cs.LG 57%

ClusterMine: Robust Label-Free Visual Out-Of-Distribution Detection via Concept Mining from Text Corpora

Nikolas Adaloglou, Diana Petrusheva, Mohamed Asker, Felix Michels, Markus Kollmann

机构 * Heinrich Heine University of Düsseldorf(海因里希-海涅大学杜塞尔多夫分校)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

Comments Accepted in WACV 2026. Code in https://github.com/HHU-MMBS/clustermine_wacv_official 9 Tables, 11 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06653 2025-11-11 cs.CV cs.CL 57%

HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment

Ruijia Wu, Ping Chen, Fei Shen, Shaoan Zhao, Qiang Hui, Huanlin Gao, Ting Lu, Zhaoxiang Liu, Fang Zhao, Kai Wang, Shiguo Lian

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments Accepted by AAAI 2026 as an Oral Presentation (13 pages, 7 figures, 7 tables)

Journal ref AAAI2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05680 2025-11-11 cs.RO 50%

VLM-driven Skill Selection for Robotic Assembly Tasks

Jeong-Jung Kim, Doo-Yeol Koh, Chang-Hyun Kim

机构 * Department of AI Machinery, Korea Institute of Machinery & Materials(人工智能机械系,韩国机械材料研究院)

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏