arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7596 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7596 篇

2310.03691 2025-02-25 cs.HC 88%

DirectGPT: A Direct Manipulation Interface to Interact with Large Language Models

Damien Masson, Sylvain Malacria, Géry Casiez, Daniel Vogel

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

Comments In Proceedings of the 2024 CHI Conference on Human Factors in Computing Systems (CHI '24)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.15389 2025-02-24 cs.CV 88%

The Role of Background Information in Reducing Object Hallucination in Vision-Language Models: Insights from Cutoff API Prompting

Masayo Tomita, Katsuhiko Hayashi, Tomoyuki Kaneko

专题命中 知识编辑与模型理解 :language model(title,abstract);prompting(title,abstract)

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.13459 2025-02-18 cs.CV 88%

Adapting Multi-modal Large Language Model to Concept Drift From Pre-training Onwards

Xiaoyu Yang, Jie Lu, En Yu

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

Comments ICLR 2025 Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06332 2025-02-18 cs.CV 88%

X-VARS: Introducing Explainability in Football Refereeing with Multi-Modal Large Language Model

Jan Held, Hani Itani, Anthony Cioppa, Silvio Giancola, Bernard Ghanem, Marc Van Droogenbroeck

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.04131 2024-10-08 cs.CL cs.AI cs.LG 88%

Towards Interpretable Sequence Continuation: Analyzing Shared Circuits in Large Language Models

Michael Lan, Philip Torr, Fazl Barez

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.18753 2024-09-30 cs.CV 88%

Enhancing Explainability in Multimodal Large Language Models Using Ontological Context

Jihen Amara, Birgitta König-Ries, Sheeba Samuel

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01355 2024-08-06 cs.CV cs.MM 88%

Hallu-PI: Evaluating Hallucination in Multi-modal Large Language Models within Perturbed Inputs

Peng Ding, Jingyu Wu, Jun Kuang, Dan Ma, Xuezhi Cao, Xunliang Cai, Shi Chen, Jiajun Chen, Shujian Huang

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

Comments Acccepted by ACM MM 2024, 14 pages, 11 figures, 9 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.02987 2024-06-06 cs.CV 88%

Enhancing Multimodal Large Language Models with Multi-instance Visual Prompt Generator for Visual Representation Enrichment

Wenliang Zhong, Wenyi Wu, Qi Li, Rob Barton, Boxin Du, Shioulin Sam, Karim Bouyarmane, Ismail Tutar, Junzhou Huang

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.00974 2024-06-04 eess.SY cs.SY 88%

Large Language Model Assisted Optimal Bidding of BESS in FCAS Market: An AI-agent based Approach

Borui Zhang, Chaojie Li, Guo Chen, Zhaoyang Dong

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.12540 2024-05-22 cs.CV cs.MM 88%

Context-Enhanced Video Moment Retrieval with Large Language Models

Weijia Liu, Bo Miao, Jiuxin Cao, Xuelin Zhu, Bo Liu, Mehwish Nasim, Ajmal Mian

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.08046 2024-04-08 cs.CV 88%

Chat-UniVi: Unified Visual Representation Empowers Large Language Models with Image and Video Understanding

Peng Jin, Ryuichi Takanobu, Wancai Zhang, Xiaochun Cao, Li Yuan

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

Comments Accepted by CVPR 2024 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17911 2024-03-13 cs.CV 88%

OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allocation

Qidong Huang, Xiaoyi Dong, Pan Zhang, Bin Wang, Conghui He, Jiaqi Wang, Dahua Lin, Weiming Zhang, Nenghai Yu

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

Comments CVPR 2024, code is available at https://github.com/shikiw/OPERA

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.06968 2024-02-27 cs.CV 88%

Hallucination Augmented Contrastive Learning for Multimodal Large Language Model

Chaoya Jiang, Haiyang Xu, Mengfan Dong, Jiaxing Chen, Wei Ye, Ming Yan, Qinghao Ye, Ji Zhang, Fei Huang, Shikun Zhang

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.04206 2024-02-07 cs.RO 88%

Explaining Autonomy: Enhancing Human-Robot Interaction through Explanation Generation with Large Language Models

David Sobrín-Hidalgo, Miguel A. González-Santamarta, Ángel M. Guerrero-Higueras, Francisco J. Rodríguez-Lera, Vicente Matellán-Olivera

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

Comments 26 pages, 15 Figures, 11 Tables. This paper is a preprint of an article submitted to the International Journal of Social Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.03284 2024-02-06 cs.CL cs.AI cs.LG 88%

Deal, or no deal (or who knows)? Forecasting Uncertainty in Conversations using Large Language Models

Anthony Sicilia, Hyunwoo Kim, Khyathi Raghavi Chandu, Malihe Alikhani, Jack Hessel

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(title);分类 cs.CL、cs.AI、cs.LG

Comments 2 Figures; 7 Tables; 27 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.16367 2023-05-29 cs.CL cs.AI cs.LG 88%

Role-Play with Large Language Models

Murray Shanahan, Kyle McDonell, Laria Reynolds

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.12487 2023-05-23 cs.AI cs.CL cs.LG 88%

Augmenting Autotelic Agents with Large Language Models

Cédric Colas, Laetitia Teodorescu, Pierre-Yves Oudeyer, Xingdi Yuan, Marc-Alexandre Côté

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.05077 2022-11-10 cs.CV 88%

Prompting Large Pre-trained Vision-Language Models For Compositional Concept Learning

Guangyue Xu, Parisa Kordjamshidi, Joyce Chai

专题命中 知识编辑与模型理解 :language model(title,abstract);prompting(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.24807 2026-08-26 cs.SE cs.AI 新提交 87%

Automatic Model Card Generation Using an LLM

基于大语言模型的自动模型卡片生成

Tajkia Rahman Toma, Balreet Grewal, Cor-Paul Bezemer

专题命中 知识编辑与模型理解 :LLM(title,summary_cn);分类 cs.AI

AI总结 本研究提出基于LLM的MCTidy与MCGenie,前者重组模型卡片为标准化模板,后者从仓库数据直接生成,经48个Hugging Face模型验证,可实现标准化、可扩展的模型卡片文档生成。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.24758 2026-08-26 cs.AI 新提交 87%

RACE: Scalable Statistical Estimation of Functional Consistency in LLM Neurons

RACE:LLM神经元中功能一致性的可扩展统计估计

Runyu Wang, Bo Liu, Xiaxin Zhang, Yu Han, Jiawei Cao, Xiaoye Zhang, Zhe Zhang, Yifan Yang, Peng Ping

机构 * Nantong University(南通大学) Chongqing University of Post and Telecommunications(重庆邮电大学) China Southern Power Grid Company Limited(中国南方电网有限责任公司) Meituan(美团)

专题命中 知识编辑与模型理解 :LLM(title,title_cn);分类 cs.AI

AI总结 针对LLM神经元功能一致性的可扩展统计估计难题,提出RACE框架,其领域特异性更优、计算开销低两个数量级,可有效评估Transformer神经元的全领域功能一致性。

Comments EMNLP-26 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.12851 2026-08-14 cs.AI 新提交 87%

Practice Makes Unsafe: Skill Misevolution in Self-Improving LLM Agents

熟能生险:自我改进的大语言模型智能体中的技能误演化

Xutao Mao, Liangjie Zhao, Xiang Zheng, Cong Wang

机构 * Adelaide University(阿德莱德大学) City University of Hong Kong(香港城市大学)

专题命中 知识编辑与模型理解 :LLM(title,summary_cn);分类 cs.AI

AI总结 该研究针对自我改进LLM智能体的技能误演化问题,构建SkillMisevo-Gym与SkillMisevo-Bench基准,提出SafeEvolve方法,可显著降低不安全检索与新会话危害,同时对良性效用影响极小。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.10836 2026-08-12 cs.SD cs.AI 新提交 87%

Whisper-Aware LLM: Self-Supervised Uncertainty Learning for Robust Whispered Speech Recognition

感知耳语的大语言模型:用于鲁棒耳语语音识别的自监督不确定性学习

Gaopeng Xu, Zhenyu Wang, Zheng Xue, Yinfeng Xia, Haitao Yao

机构 * Qwen Business Unit of Alibaba(阿里巴巴通义千问业务部)

专题命中 知识编辑与模型理解 :LLM(title,summary_cn);分类 cs.AI

AI总结 本文提出Whisper-Aware LLM框架,通过自监督学习量化声学信号缺陷并结合置信融合解码机制,在AISHELL6-Whisper数据集上将耳语语音识别的CER相对降低17%,幻觉率降至4.5%,达到最优性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.00997 2026-08-11 cs.CL 版本更新 87%

Uncertainty-Aware Variational Reward Factorization via Probabilistic Preference Bases for LLM Personalization

基于概率偏好基的不确定性感知变分奖励分解用于大语言模型个性化

Gyuseok Lee, Wonbin Kweon, Zhenrui Yue, SeongKu Kang, Jiawei Han, Dong Wang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Korea University(高丽大学)

专题命中 知识编辑与模型理解 :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出变分奖励分解框架,通过概率偏好基和变分分布实现用户偏好建模,提升LLM个性化效果。

Comments COLM'26

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.07251 2026-08-10 econ.GN cs.AI q-fin.EC 新提交 87%

Reading Copom's Tone: A Weighted LLM Framework for Hawkish-Dovish Sentiment, Forward Guidance, and Uncertainty

解读巴西央行货币政策委员会(Copom)的语气:用于判断鹰派-鸽派情绪、前瞻性指引及不确定性的加权大语言模型(LLM)框架

Gabriel de Macedo Santos

专题命中 知识编辑与模型理解 :LLM(title,title_cn);分类 cs.AI

AI总结 本研究提出一种加权大语言模型框架,用于分析巴西央行Copom声明的鹰派-鸽派情绪、前瞻性指引及不确定性,通过实证样本验证了方法的有效性,核心贡献为构建了可分离语气与政策指引的可审计系统。

Comments 12 pages, 8 tables, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26599 2026-07-30 cs.LG 新提交 87%

Uncertainty-Guided LLM Semantic Augmentation for Heterogeneous Treatment Effect Estimation

面向异质处理效应估计的不确定性引导大语言模型语义增强

Jialu Xu, Mengkun Liang, Guannan Liu, Xiaojie Mao, Junjie Wu

机构 * Beihang University(北京航空航天大学) School of Economics and Management, Tsinghua University(清华大学经济管理学院)

专题命中 知识编辑与模型理解 :LLM(title,summary_cn);分类 cs.LG

AI总结 本研究针对异质处理效应估计的局部不稳定性,提出CURL方法,通过LLM生成分离的分配与异质性导向表示,在四个基准上提升了10种主学习者的性能。

Comments 17 pages, 12 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.22661 2026-07-28 cs.AI 新提交 87%

TRE: Training-Free Hallucination Detection for Diffusion Language Models

TRE:用于扩散语言模型的无训练幻觉检测

Pengcheng Weng, Yanyu Qian, Yue Tan, Yixin Liu

机构 * University of Bern(伯尔尼大学) Nanyang Technological University(南洋理工大学) Griffith University(格里菲斯大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract,abstract_cn);large language model(abstract);分类 cs.AI

AI总结 针对扩散语言模型的幻觉问题,现有方法有局限。本文提出无训练的TRE指标,通过在模型解码过程中沿时空维度提取熵信号来估计幻觉风险,经实验验证其性能有竞争力,具有泛化性、效率和鲁棒性等优点。

Comments 25 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09001 2026-07-28 cs.CL 版本更新 87%

Entropy Sentinel: Probing Entropy Traces for LLM Monitoring

熵哨兵:基于STEM解码熵迹的连续LLM准确性监控

Pedro Memoli Buffa, Luciano Del Corro

机构 * Departamento de Matematica, FCEyN Universidad de Buenos Aires(数学系,布宜诺斯艾利斯大学) ELIAS Lab, Departamento de Ingeniería Universidad de San Andres(ELIAS实验室,圣安德烈斯大学)

专题命中 知识编辑与模型理解 :LLM(title,title_cn);分类 cs.CL

AI总结 提出利用输出熵迹作为推理时信号,通过轻量分类器预测实例正确性并聚合为领域级准确性估计,在STEM推理基准上验证了其用于监控和数据采集的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.18259 2026-07-22 cs.AI 新提交 87%

Probabilistic Concept-Aware Steering for Trustworthy LLM Inference

用于可信大语言模型推理的概率概念感知引导

Brian Becker, Rui Chu, Yingjie Lao

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究针对大语言模型推理中现有引导向量方法的问题,提出概率概念感知引导框架,通过概念驱动检索和概率校准,在保留模型能力的同时实现可控、安全的语义偏差引导。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.07903 2026-07-10 cs.CR cs.AI 新提交 87%

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

通过内部归因图对大语言模型越狱进行机制性可解释性研究

Anupam Wagle, Ifrat Ikhtear Uddin, Chaowei Zhang, Longwei Wang

机构 * Department of Computer Science, University of South Dakota(南达科他大学计算机科学系) School of Information and Artificial Intelligence, Yangzhou University(扬州大学信息与人工智能学院)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究大语言模型越狱漏洞,通过构建内部计算图揭示对抗攻击引发的内部推理转变,提出统一框架实现因果诊断,实验证明计算图偏差与不安全行为相关,干预可提升模型鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24251 2026-06-24 cs.AI 新提交 87%

Probing the Misaligned Thinking Process of Language Models

探究语言模型的错误对齐思维过程

Kaiwen Zhou, Constantin Venhoff, Jonathan Michala, Xin Eric Wang, William Saunders

机构 * University of Oxford(牛津大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract,abstract_cn);large language model(abstract);分类 cs.AI

AI总结 提出通过线性探针检测模型内部激活中的18种错误对齐指标,以可靠识别策略欺骗等行为,在分布外基准上达到0.935 AUROC。

详情

展开后加载摘要…

URL PDF HTML 收藏