arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7552 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7552 篇

2512.06020 2025-12-09 cs.CV cs.AI 70%

PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation

PrefGen: 多模态偏好学习用于偏好条件下的图像生成

Wenyi Mo, Tianyu Zhang, Yalong Bai, Ligong Han, Ying Ba, Dimitris N. Metaxas

机构 * Rutgers University(罗格斯大学) iN2X MIT-IBM Watson AI Lab(麻省理工-IBM沃森人工智能实验室) Red Hat AI Innovation(红帽人工智能创新)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 PrefGen通过多模态大语言模型提取用户偏好并注入扩散模型,实现个性化图像生成,优于现有方法。

Comments Project Page: \href{https://prefgen.github.io/}{\texttt{https://prefgen.github.io}}

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04309 2025-12-05 cs.CV cs.CL 70%

Text-Only Training for Image Captioning with Retrieval Augmentation and Modality Gap Correction

仅文本训练的图像描述生成:结合检索增强与模态差距校正

Rui Fonseca, Bruno Martins, Gil Rocha

机构 * INESC-ID, Instituto Superior Tecnico, University of Lisbon(INESC-ID,理工学院,里斯本大学)

专题命中 知识编辑与模型理解 :language model(abstract);prompting(abstract);分类 cs.CL

AI总结 TOMCap通过检索增强和模态差距校正,实现无需对齐图像-文本对的纯文本训练图像描述生成。

Comments Submitted to CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13813 2025-12-03 cs.CL 70%

Geometric Uncertainty for Detecting and Correcting Hallucinations in LLMs

几何不确定性用于检测和纠正大语言模型中的幻觉

Edward Phillips, Sean Wu, Soheila Molaei, Danielle Belgrave, Anshul Thakur, David Clifton

机构 * Department of Engineering Science, University of Oxford(牛津大学工程科学系) GlaxoSmithKline(葛兰素史克) Oxford Suzhou Centre for Advanced Research(牛津苏黎世高级研究中心)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出几何框架用于检测和纠正大语言模型中的幻觉,通过几何体积和几何怀疑方法提升响应可靠性。

Comments Revision. Clarified positioning as a unified geometric framework for global and local uncertainty in LLMs. Added baselines (Degree, Eccentricity) and expanded comparison to related methods. Included ablations (PCA dimension, number of archetypes, number of samples) and complexity analysis. Extended discussion of medical QA results and model-specific behaviour

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00706 2025-12-02 cs.CV cs.AI 70%

Optimizing LVLMs with On-Policy Data for Effective Hallucination Mitigation

利用策略数据优化LVLMs以实现有效的幻觉缓解

Chengzhi Yu, Yifan Xu, Yifan Chen, Wenyi Zhang

机构 * University of Science and Technology of China(中国科学技术大学) Hong Kong Baptist University(香港 Baptist 大学)

专题命中 知识编辑与模型理解 :language model(abstract);preference optimization(abstract);分类 cs.AI

AI总结 本文提出利用策略数据优化LVLMs,通过幻觉分类器和动态加权DPO算法,显著降低幻觉率并提升模型性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19382 2025-12-02 cs.CL 70%

Measuring and Guiding Monosemanticity

测量和引导单义性

Ruben Härle, Felix Friedrich, Manuel Brack, Stephan Wäldchen, Björn Deiseroth, Patrick Schramowski, Kristian Kersting

机构 * Computer Science Department, TU Darmstadt(图宾根大学计算机科学系) Lab1141 Aleph Alpha Research(Aleph Alpha研究) German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心) CERTAIN Centre of Cognitive Science, TU Darmstadt(图宾根大学认知科学中心)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出G-SAE方法,通过在训练期间将潜在表示条件化于标记概念,以提高大型语言模型的可解释性和可控性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.23174 2025-12-01 cs.CL 70%

Are LLMs Good Safety Agents or a Propaganda Engine?

LLMs是安全代理还是宣传引擎?

Neemesh Yadav, Francesco Ortu, Jiarui Liu, Joeun Yook, Bernhard Schölkopf, Rada Mihalcea, Alberto Cazzaniga, Zhijing Jin

机构 * SMU University of Trieste(特里este大学) AREA Science Park(AREA科学园) CMU(卡内基梅隆大学) University of Toronto(多伦多大学) Vector Institute(向量研究所) MPI for Intelligent Systems(智能系统研究所) University of Michigan(密歇根大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文通过PSP数据集研究LLMs在政治敏感内容上的拒绝行为,发现大多数模型存在审查倾向,并分析了影响拒绝分布的关键因素。

Comments 15 pages, 7 tables, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21389 2025-11-27 cs.IR cs.AI 70%

FITRep: Attention-Guided Item Representation via MLLMs

FITRep: 通过大语言模型实现的注意力引导的项目表示

Guoxiao Zhang, Ao Li, Tan Qu, Qianlong Xie, Xingxing Wang

机构 * Meituan(美团)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 FITRep通过引入注意力引导的白盒表示框架,利用多模态大语言模型实现细粒度项目去重,提升了广告点击率和每千次展示成本。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12964 2025-11-26 cs.CL 70%

MA-COIR: Leveraging Semantic Search Index and Generative Models for Ontology-Driven Biomedical Concept Recognition

MA-COIR: 利用语义搜索索引和生成模型进行面向本体的生物医学概念识别

Shanshan Liu, Noriki Nishida, Rumana Ferdous Munne, Narumi Tokunaga, Yuki Yamagata, Kouji Kozaki, Yuji Matsumoto

机构 * RIKEN AIP University of Tsukuba(茨川大学) RIKEN R-IH RIKEN BRC Osaka Electro-Communication University(大阪电讯大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 MA-COIR通过结合语义搜索索引和生成模型,提升生物医学领域本体驱动的概念识别能力。

Comments preprint

Journal ref https://aclanthology.org/2025.acl-srw.39/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18301 2025-11-25 cs.CL 70%

"AGI" team at SHROOM-CAP: Data-Centric Approach to Multilingual Hallucination Detection using XLM-RoBERTa

SHROOM-CAP团队的AGI:基于XLM-RoBERTa的多语言幻觉检测数据导向方法

Harsh Rathva, Pruthwik Mishra, Shrikant Malviya

机构 * Sardar Vallabhbhai National Institute of Technology (SVNIT)(萨达尔·瓦拉布尔·尼蒂国家理工学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 SHROOM-CAP团队提出基于XLM-RoBERTa的多语言幻觉检测方法,通过数据集统一和平衡提升模型性能,尤其在印地语等低资源语言中取得显著成果。

Comments Accepted to the 1st Workshop on Confabulation, Hallucinations & Overgeneration in Multilingual and Practical Settings (CHOMPS) at AACL-IJCNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00230 2025-11-25 cs.HC cs.AI 70%

Neural Transparency: Mechanistic Interpretability Interfaces for Anticipating Model Behaviors for Personalized AI

神经透明:用于预测模型行为的机制可解释性接口以实现个性化AI

Sheer Karny, Anthony Baez, Pat Pataranutaporn

机构 * MIT Media Lab(麻省理工学院媒体实验室) Massachusetts Institute of Technology(麻省理工学院)

专题命中 知识编辑与模型理解 :LLM(abstract);language model(abstract);分类 cs.AI

AI总结 本研究提出神经透明接口,通过可视化语言模型内部结构帮助用户预测AI行为,提升信任并促进更安全的人机交互。

Comments SK and AB are co-first authors

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17747 2025-11-21 cs.CL 70%

Discriminating Form and Meaning in Multilingual Models with Minimal-Pair ABX Tasks

通过最小对ABX任务区分多语言模型中的形式与意义

Maureen de Seyssel, Jie Chi, Skyler Seto, Maartje ter Hoeve, Masha Fedzechkina, Natalie Schluter

机构 * Apple(苹果公司)

专题命中 知识编辑与模型理解 :language model(abstract);pretraining(abstract);分类 cs.CL

AI总结 通过最小对ABX任务研究多语言模型中形式与意义的区分能力,揭示了语言识别和语义识别在训练过程中的变化规律。

Comments Comments: Published in EMNLP 2025. https://aclanthology.org/2025.emnlp-main.1210.pdf

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.03628 2025-11-21 cs.CL 70%

GPTopic: Dynamic and Interactive Topic Representations

GPTopic: 动态和交互式主题表示

Arik Reuter, Bishnu Khadka, Anton Thielmann, Christoph Weisser, Sebastian Fischer, Benjamin Säfken

机构 * University of Cambridge(剑桥大学) LMU Munich(慕尼黑大学) TU Clausthal(Clausthal技术大学) BASF(巴斯夫) Tribhuvan University(特里布文大学) MCML

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 GPTopic利用大型语言模型创建动态交互式主题表示,通过直观界面提升主题建模的可访问性和全面性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22720 2025-11-20 cs.CL 70%

Investigating Hallucination in Conversations for Low Resource Languages

Amit Das, Md. Najib Hasan, Souvika Sarkar, Zheng Zhang, Fatemeh Jamshidi, Tathagata Bhattacharya, Nilanjana Raychawdhury, Dongji Feng, Vinija Jain, Aman Chadha

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15202 2025-11-20 cs.AI 70%

SOLID: a Framework of Synergizing Optimization and LLMs for Intelligent Decision-Making

Yinsheng Wang, Tario G You, Léonard Boussioux, Shan Liu

机构 * Department of Industrial & Systems Engineering, University of Washington(工业与系统工程系,华盛顿大学) College of Engineering, University of Washington(工程学院,华盛顿大学) Michael G. Foster School of Business, University of Washington(Michael G. Foster商学院,华盛顿大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments NeurIPS 2025 WORKSHOP ML*OR Workshop: Mathematical Foundations and Operational Integration of Machine Learning for Uncertainty-Aware Decision-Making

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00447 2025-11-19 cs.CR cs.AI 70%

DRIP: Defending Prompt Injection via Token-wise Representation Editing and Residual Instruction Fusion

Ruofan Liu, Yun Lin, Zhiyong Huang, Jin Song Dong

机构 * National University of Singapore(新加坡国立大学) Shanghai Jiao Tong University(上海交通大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24796 2025-11-18 cs.CY cs.AI 70%

Mutual Wanting in Human--AI Interaction: Empirical Evidence from Large-Scale Analysis of GPT Model Transitions

HaoYang Shang, Xuan Liu

机构 * BreathingCORE

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.10811 2025-11-17 cs.LG 70%

Transformers know more than they can tell -- Learning the Collatz sequence

François Charton, Ashvni Narayanan

机构 * Axiom Math CERMICS, Ecole Nationale des Ponts et Chaussées(Axiom数学 CERMICS,国立桥梁与道路学院) Sydney Mathematical Research Institute, University of Sydney(悉尼数学研究所,悉尼大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.21239 2025-11-13 cs.CL 70%

Semantic Volume: Quantifying and Detecting both External and Internal Uncertainty in LLMs

Xiaomin Li, Zhou Yu, Ziji Zhang, Yingying Zhuang, Swair Shah, Narayanan Sadagopan, Anurag Beniwal

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08440 2025-11-12 cs.LG 70%

Coherence Mechanisms for Provable Self-Improvement

Mehryar Mohri, Jon Schneider, Yifan Wu

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17120 2025-11-12 cs.CL 70%

Self-Interpretability: LLMs Can Describe Complex Internal Processes that Drive Their Decisions

Dillon Plunkett, Adam Morris, Keerthi Reddy, Jorge Morales

机构 * Northeastern University(东北大学) Princeton University(普林斯顿大学) Independent Researcher(独立研究者)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20264 2025-11-11 cs.CL 70%

EMBRACE: Shaping Inclusive Opinion Representation by Aligning Implicit Conversations with Social Norms

Abeer Aldayel, Areej Alokaili

机构 * King Saud University, College of Computer and Information Sciences(沙特王后大学,计算机与信息科学学院)

专题命中 知识编辑与模型理解 :language model(abstract);post-training(abstract);分类 cs.CL

Comments Accepted, to appear IJCNLP-AACL 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07863 2025-11-11 cs.LG 70%

Robust Hallucination Detection in LLMs via Adaptive Token Selection

Mengjia Niu, Hamed Haddadi, Guansong Pang

机构 * Imperial College London, UK(伦敦帝国学院) Singapore Management University(新加坡管理大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16678 2025-11-11 cs.CL 70%

Mechanisms vs. Outcomes: Probing for Syntax Fails to Explain Performance on Targeted Syntactic Evaluations

Ananth Agarwal, Jasper Jian, Christopher D. Manning, Shikhar Murty

机构 * Stanford University(斯坦福大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Journal ref Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09360 2025-11-10 cs.CL 70%

MetaRAG: Metamorphic Testing for Hallucination Detection in RAG Systems

Channdeth Sok, David Luz, Yacine Haddam

机构 * Forvia Paris Tech Center, GIT, Immeuble Lumière, 40 avenue des Terroirs de France, 75012 Paris, France(巴黎Forvia技术中心,GIT,Lumière大厦,法国巴黎75012)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Identity-Aware AI workshop at 28th European Conference on Artificial Intelligence, October 25, 2025, Bologna, Italy

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02720 2025-11-05 cs.CV cs.AI 70%

LLEXICORP: End-user Explainability of Convolutional Neural Networks

Vojtěch Kůr, Adam Bajger, Adam Kukučka, Marek Hradil, Vít Musil, Tomáš Brázdil

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02243 2025-11-05 cs.AI 70%

When Modalities Conflict: How Unimodal Reasoning Uncertainty Governs Preference Dynamics in MLLMs

Zhuoran Zhang, Tengyue Wang, Xilin Gong, Yang Shi, Haotian Wang, Di Wang, Lijie Hu

机构 * Peking University(北京大学) Provable Responsible AI and Data Analytics (PRADA) Lab(可证明负责任的人工智能和数据分析实验室) King Abdullah University of Science and Technology(卡布斯大学) South China University of Technology(华南理工大学) Tsinghua University(清华大学) University of Georgia(佐治亚大学) MBZUAI(马克斯·普朗克人工智能研究所)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 19 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18376 2025-10-30 cs.LG cs.SI 70%

GnnXemplar: Exemplars to Explanations -- Natural Language Rules for Global GNN Interpretability

Burouj Armgaan, Eshan Jain, Harsh Pandey, Mahesh Chandran, Sayan Ranu

机构 * Dept. of CSE, IIT Delhi(印度德里理工学院计算机科学与工程系) Fujitsu Research of India, Bangalore(印度班加罗尔富士通印度研究机构) Dept. of CSE and Yardi ScAI, IIT Delhi(印度德里理工学院计算机科学与工程系)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments 38 pages, 20 figures, NeurIPS 2025 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24259 2025-10-29 cs.CL cs.RO 70%

Can LLMs Translate Human Instructions into a Reinforcement Learning Agent's Internal Emergent Symbolic Representation?

Ziqi Ma, Sao Mai Nguyen, Philippe Xu

机构 * U2IS, ENSTA, IP-Paris(U2IS、ENSTA、IP-巴黎)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21059 2025-10-28 cs.CL 70%

Dynamic Retriever for In-Context Knowledge Editing via Policy Optimization

Mahmud Wasif Nafee, Maiqi Jiang, Haipeng Chen, Yanfu Zhang

机构 * Rensselaer Polytechnic Institute(拉特格斯理工学院) Bangladesh University of Engineering and Technology(孟加拉工程与技术大学) College of William & Mary(威廉与玛丽学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted at EMNLP 2025. Copyright 2025 Association for Computational Linguistics (CC BY 4.0). 12 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17394 2025-10-28 cs.CV cs.AI 70%

HiProbe-VAD: Video Anomaly Detection via Hidden States Probing in Tuning-Free Multimodal LLMs

Zhaolin Cai, Fan Li, Ziwei Zheng, Yanjun Qin

机构 * Xinjiang University(新疆大学) Xi'an Jiaotong University(西安交通大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments Accepted by ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏