arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-02-04 至 2026-02-04 共收录 18 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 18 篇

2602.02537 2026-02-04 cs.CV cs.LG 88%

WorldVQA: Measuring Atomic World Knowledge in Multimodal Large Language Models

WorldVQA:评估多模态大语言模型的原子视觉世界知识

Runjie Zhou, Youbo Shao, Haoyu Lu, Bowei Xing, Tongtong Bai, Yujie Chen, Jie Zhao, Lin Sui, Haotian Yao, Zijia Zhao, Hao Yang, Haoning Wu, Zaida Zhou, Jinguo Zhu, Zhiqi Huang, Yiping Bao, Yangyang Liu, Y. Charles, Xinyu Zhou

机构 * Moonshot AI

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.LG

AI总结 WorldVQA通过评估多模态大语言模型的原子视觉知识,建立衡量其事实性和百科全书广度的基准测试。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21129 2026-02-04 cs.AI 88%

Measuring and Analyzing Intelligence via Contextual Uncertainty in Large Language Models using Information-Theoretic Metrics

通过信息论度量在大语言模型中通过上下文不确定性测量和分析智能

Jae Wan Shim

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文提出通过熵衰减曲线和信息增益跨度分析大语言模型的内部动态,揭示模型规模与文本复杂性对认知档案的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02582 2026-02-04 cs.AI cs.CL cs.CY cs.IR cs.LG cs.SE 87%

Uncertainty and Fairness Awareness in LLM-Based Recommendation Systems

大语言模型推荐系统中的不确定性与公平性意识

Chandan Kumar Sah, Xiaoli Lian, Li Zhang, Tony Xu, Syed Shazaib Shah

机构 * Beihang University(北京航空航天大学) McGill University(麦吉尔大学)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文研究了大语言模型推荐系统中不确定性与公平性的影响,提出了一种新的评估方法并揭示了个性化与公平性之间的权衡。

Comments Accepted at the Second Conference of the International Association for Safe and Ethical Artificial Intelligence, IASEAI26, 14 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03132 2026-02-04 cs.LG cs.AI cs.NE 86%

Contrastive Concept-Tree Search for LLM-Assisted Algorithm Discovery

对比概念树搜索用于大语言模型辅助算法发现

Timothee Leleu, Sudeera Gunathilaka, Federico Ghimenti, Surya Ganguli

机构 * NTT Research(NTT研究院) Stanford University(斯坦福大学)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本研究提出对比概念树搜索(CCTS)方法,通过提取层次化概念表示和对比学习模型,提升LLM辅助算法发现的搜索效率和可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03271 2026-02-04 cs.CR 85%

LogicScan: An LLM-driven Framework for Detecting Business Logic Vulnerabilities in Smart Contracts

LogicScan: 一种基于大语言模型的智能合约业务逻辑漏洞检测框架

Jiaqi Gao, Zijian Zhang, Yuqiang Sun, Ye Liu, Chengwei Liu, Han Liu, Yi Li, Yang Liu

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 LogicScan利用大语言模型和业务规范语言,通过挖掘链上协议的业务不变量,实现对智能合约业务逻辑漏洞的高效检测。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21984 2026-02-04 cs.CV cs.CL 83%

Beyond the Vision Encoder: Identifying and Mitigating Spatial Bias in Large Vision-Language Models

超越视觉编码器:识别和缓解大视觉-语言模型中的空间偏差

Yingjie Zhu, Xuefeng Bai, Kehai Chen, Yang Xiang, Youcheng Pan, Yongshuai Hou, Weili Guan, Jun Yu, Min Zhang

机构 * Harbin Institute of Technology, Shenzhen, China(哈尔滨工业大学深圳学院) Peng Cheng Laboratory, Shenzhen, China(鹏城实验室)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

AI总结 本文提出AGCI机制,通过动态注入全局视觉上下文缓解大视觉-语言模型中的空间偏差问题,提升模型的空间鲁棒性和下游任务表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18179 2026-02-04 cs.CL cs.AI 82%

Problem Solved? Information Extraction Design Space for Layout-Rich Documents using LLMs

问题已解决?利用LLMs处理布局丰富文档的信息提取设计空间

Gaye Colakoglu, Gürkan Solmaz, Jonathan Fürst

机构 * Zurich University of Applied Sciences(苏黎世应用科学大学) NEC Laboratories Europe(NEC欧洲实验室)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本文通过LayIE-LLM测试套件研究了利用LLMs处理布局丰富文档的信息提取设计空间,证明通用LLMs在优化配置下可媲美专用模型,提供低成本无微调方案。

Comments accepted at EMNLP'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02558 2026-02-04 cs.LG cs.AI q-bio.QM 81%

PA-MIL: Phenotype-Aware Multiple Instance Learning Guided by Language Prompting and Genotype-to-Phenotype Relationships

PA-MIL:基于语言提示和基因型-表型关系的表型感知多实例学习

Zekang Yang, Hong Liu, Xiangdong Wang

机构 * Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)

专题命中 知识编辑与模型理解 :prompting(title,abstract);分类 cs.AI、cs.LG

AI总结 PA-MIL通过结合语言提示和基因型-表型关系,实现对癌症表型的前瞻性可解释性学习,提升全切片图像分析的可靠性与可问责性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.20834 2026-02-04 cs.CL cs.LG 81%

Linear representations in language models can change dramatically over a conversation

语言模型中的线性表示在对话中可以剧烈变化

Andrew Kyle Lampinen, Yuxuan Li, Eghbal Hosseini, Sangnie Bhardwaj, Murray Shanahan

机构 * Google DeepMind(谷歌DeepMind)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

AI总结 语言模型中的线性表示在对话中会随内容和角色变化而剧烈改变,影响可解释性和引导方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01769 2026-02-04 cs.LG cs.AI 79%

IRIS: Implicit Reward-Guided Internal Sifting for Mitigating Multimodal Hallucination

IRIS: 隐式奖励引导的内部筛选以缓解多模态幻觉

Yuanshuai Li, Yuping Yan, Jirui Han, Fei Ming, Lingjuan Lv, Yaochu Jin

机构 * Department of Artificial Intelligence, Westlake University, Hangzhou, China(人工智能系,西湖大学,杭州,中国) Sony Research, Sony(索尼研究,索尼)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);preference optimization(abstract);分类 cs.AI、cs.LG

AI总结 IRIS通过隐式奖励引导内部筛选,有效缓解多模态大语言模型的幻觉问题,无需外部反馈且性能优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04136 2026-02-04 cs.CV cs.AI 77%

UniFGVC: Universal Training-Free Few-Shot Fine-Grained Vision Classification via Attribute-Aware Multimodal Retrieval

UniFGVC: 一种通用无训练少样本细粒度视觉分类方法通过属性感知多模态检索

Hongyu Guo, Xiangzhao Hao, Jiarui Guo, Haiyun Guo, Jinqiao Wang, Tat-Seng Chua

机构 * School of Traffic and Transportation, Beijing Jiaotong University(交通与运输学院,北京交通大学) Foundation Modal Research Center, Institute of Automation, Chinese Academy of Sciences(基础模态研究中心,自动化研究所) Queen Mary School Hainan, Beijing University of Posts and Telecommunications(海南女王学院,北京邮电大学) Department of Computer Science, NUS School of Computing(计算机科学系,国立大学计算机学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.AI

AI总结 UniFGVC通过属性感知多模态检索方法,实现无训练少样本细粒度视觉分类,优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03696 2026-02-04 cs.LG cs.CL 73%

Conflict-Resolving and Sharpness-Aware Minimization for Generalized Knowledge Editing with Multiple Updates

冲突解决与尖锐性感知最小化:多更新通用知识编辑

Duy Nguyen, Hanqi Xiao, Archiki Prasad, Elias Stengel-Eskin, Hyunji Lee, Mohit Bansal

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 CoRSA通过冲突解决与尖锐性感知最小化方法,提升多更新知识编辑的泛化能力和稳定性,实现比基线方法更高的性能。

Comments 22 pages, 8 figures. Code link: https://github.com/duykhuongnguyen/CoRSA

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.16540 2026-02-04 cs.SD cs.AI eess.AS 70%

Do Models Hear Like Us? Probing the Representational Alignment of Audio LLMs and Naturalistic EEG

模型是像我们一样听吗?探查音频大语言模型与自然EEG的表征对齐

Haoyun Yang, Xin Xiao, Jiang Zhong, Yu Tian, Dong Xiaohua, Yu Mao, Hao Wu, Kaiwen Wei

机构 * School of Computer Science, Chongqing University(重庆大学计算机科学学院) Dept. of Comp. Sci. and Tech., Institute for AI, Tsinghua University(清华大学计算机科学与技术系、人工智能研究院) School of Economics and Business Administration, Chongqing University(重庆大学经济与商业管理学院) School of Artificial Intelligence, Southwest University(西南大学人工智能学院) The First Affiliated Hospital of Chongqing Medical University(重庆医科大学第一附属医院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究通过比较音频大语言模型与EEG信号,揭示了模型在自然聆听中的表征对齐特性,发现排名依赖分裂、时空对齐模式及情感分离现象。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21875 2026-02-04 cs.CL 70%

LUMINA: Detecting Hallucinations in RAG System with Context-Knowledge Signals

LUMINA:通过上下文-知识信号检测RAG系统中的幻觉

Samuel Yeh, Sharon Li, Tanwi Mallick

机构 * Department of Computer Science, University of Wisconsin-Madison(威斯康星大学麦迪逊分校计算机科学系) Argonne National Laboratory(阿贡国家实验室)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 LUMINA通过上下文-知识信号检测RAG系统中的幻觉,利用分布距离和token演变测量,实现高准确率和实用性。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12387 2026-02-04 cs.LG cond-mat.dis-nn cond-mat.stat-mech math-ph math.MP q-bio.NC stat.ML 70%

Neural Thermodynamics: Entropic Forces in Deep and Universal Representation Learning

神经热力学:深度和通用表征学习中的熵力

Liu Ziyin, Yizhou Xu, Isaac Chuang

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出神经热力学理论,揭示深度学习中熵力与对称性打破对表征学习和优化行为的调控作用。

Comments Published at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03781 2026-02-04 cs.RO 67%

A Scene Graph Backed Approach to Open Set Semantic Mapping

基于场景图的开放集合语义映射方法

Martin Günther, Felix Igelbrink, Oscar Lima, Lennart Niecksch, Marian Renz, Martin Atzmueller

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

AI总结 本文提出基于场景图的开放集合语义映射方法,通过实时更新三维语义场景图,实现大规模环境中的稳定、可验证的映射结构,提升感知与高层推理的一致性与效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02521 2026-02-04 cs.LG cs.AI eess.SP 62%

Scaled Dot-Product Attention implements projection of inputs onto a common surface

缩放点积注意力实现了对输入向量在共同表面上的投影

Terence D Sanger

机构 * Department of Electrical Engineering and Computer Science(电气工程与计算机科学系) University of California, Irvine(加州大学伊文斯顿分校)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出SDPA可重写为输入向量在共同表面的投影,揭示其在时间依赖性上下文中的作用,为非线性时间序列处理提供新视角。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16936 2026-02-04 cs.LG 57%

SPAR: Self-supervised Placement-Aware Representation Learning for Distributed Sensing

SPAR:用于分布式传感的自监督位置感知表示学习

Yizhuo Chen, Tianchen Wang, You Lyu, Yanlan Hu, Jinyang Li, Tomoyoshi Kimura, Hongjue Zhao, Yigong Hu, Denizhan Kara, Tarek Abdelzaher

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Stanford University(斯坦福大学)

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.LG

AI总结 SPAR通过信号与位置的二元性原则,提出了一种自监督位置感知表示学习框架,提升分布式传感在多种模态和任务中的鲁棒性和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏