arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-02-04 至 2026-02-04 共收录 299 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 18 篇

2602.03696 2026-02-04 cs.LG cs.CL 73%

Conflict-Resolving and Sharpness-Aware Minimization for Generalized Knowledge Editing with Multiple Updates

冲突解决与尖锐性感知最小化:多更新通用知识编辑

Duy Nguyen, Hanqi Xiao, Archiki Prasad, Elias Stengel-Eskin, Hyunji Lee, Mohit Bansal

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 CoRSA通过冲突解决与尖锐性感知最小化方法,提升多更新知识编辑的泛化能力和稳定性,实现比基线方法更高的性能。

Comments 22 pages, 8 figures. Code link: https://github.com/duykhuongnguyen/CoRSA

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.16540 2026-02-04 cs.SD cs.AI eess.AS 70%

Do Models Hear Like Us? Probing the Representational Alignment of Audio LLMs and Naturalistic EEG

模型是像我们一样听吗?探查音频大语言模型与自然EEG的表征对齐

Haoyun Yang, Xin Xiao, Jiang Zhong, Yu Tian, Dong Xiaohua, Yu Mao, Hao Wu, Kaiwen Wei

机构 * School of Computer Science, Chongqing University(重庆大学计算机科学学院) Dept. of Comp. Sci. and Tech., Institute for AI, Tsinghua University(清华大学计算机科学与技术系、人工智能研究院) School of Economics and Business Administration, Chongqing University(重庆大学经济与商业管理学院) School of Artificial Intelligence, Southwest University(西南大学人工智能学院) The First Affiliated Hospital of Chongqing Medical University(重庆医科大学第一附属医院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究通过比较音频大语言模型与EEG信号,揭示了模型在自然聆听中的表征对齐特性,发现排名依赖分裂、时空对齐模式及情感分离现象。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21875 2026-02-04 cs.CL 70%

LUMINA: Detecting Hallucinations in RAG System with Context-Knowledge Signals

LUMINA:通过上下文-知识信号检测RAG系统中的幻觉

Samuel Yeh, Sharon Li, Tanwi Mallick

机构 * Department of Computer Science, University of Wisconsin-Madison(威斯康星大学麦迪逊分校计算机科学系) Argonne National Laboratory(阿贡国家实验室)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 LUMINA通过上下文-知识信号检测RAG系统中的幻觉,利用分布距离和token演变测量,实现高准确率和实用性。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12387 2026-02-04 cs.LG cond-mat.dis-nn cond-mat.stat-mech math-ph math.MP q-bio.NC stat.ML 70%

Neural Thermodynamics: Entropic Forces in Deep and Universal Representation Learning

神经热力学:深度和通用表征学习中的熵力

Liu Ziyin, Yizhou Xu, Isaac Chuang

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出神经热力学理论,揭示深度学习中熵力与对称性打破对表征学习和优化行为的调控作用。

Comments Published at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03781 2026-02-04 cs.RO 67%

A Scene Graph Backed Approach to Open Set Semantic Mapping

基于场景图的开放集合语义映射方法

Martin Günther, Felix Igelbrink, Oscar Lima, Lennart Niecksch, Marian Renz, Martin Atzmueller

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

AI总结 本文提出基于场景图的开放集合语义映射方法,通过实时更新三维语义场景图,实现大规模环境中的稳定、可验证的映射结构,提升感知与高层推理的一致性与效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02521 2026-02-04 cs.LG cs.AI eess.SP 62%

Scaled Dot-Product Attention implements projection of inputs onto a common surface

缩放点积注意力实现了对输入向量在共同表面上的投影

Terence D Sanger

机构 * Department of Electrical Engineering and Computer Science(电气工程与计算机科学系) University of California, Irvine(加州大学伊文斯顿分校)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出SDPA可重写为输入向量在共同表面的投影,揭示其在时间依赖性上下文中的作用,为非线性时间序列处理提供新视角。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16936 2026-02-04 cs.LG 57%

SPAR: Self-supervised Placement-Aware Representation Learning for Distributed Sensing

SPAR:用于分布式传感的自监督位置感知表示学习

Yizhuo Chen, Tianchen Wang, You Lyu, Yanlan Hu, Jinyang Li, Tomoyoshi Kimura, Hongjue Zhao, Yigong Hu, Denizhan Kara, Tarek Abdelzaher

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Stanford University(斯坦福大学)

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.LG

AI总结 SPAR通过信号与位置的二元性原则,提出了一种自监督位置感知表示学习框架,提升分布式传感在多种模态和任务中的鲁棒性和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 其他LLM 22 篇

2601.11544 2026-02-04 cs.HC cs.AI cs.CL 88%

Medication counseling with large language models: balancing flexibility and rigidity

利用大语言模型进行用药咨询:在灵活性与严谨性之间取得平衡

Joar Sabel, Mattias Wingren, Andreas Lundell, Sören Andersson, Sara Rosenberg, Susanne Hägglund, Linda Estman, Malin Andtfolk

机构 * Department of Engineering and Information Technology Åbo Akademi University(工程与信息科技系阿博阿卡迪米大学) Experience Lab Åbo Akademi University(经验实验室阿博阿卡迪米大学) Department of Natural and Health Sciences Åbo Akademi University(自然与健康科学系阿博阿卡迪米大学) Department of Caring and Ethics University of Stavanger(护理与伦理系斯塔万格大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种利用大语言模型在药房环境中提供用药咨询的原型系统,旨在平衡对话要求与灵活性,减少幻觉并提高响应质量。

Comments Accepted for 2025 IEEE International Conference on Agentic AI (ICA). 14 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02178 2026-02-04 cs.CL 88%

AR-MAP: Are Autoregressive Large Language Models Implicit Teachers for Diffusion Large Language Models?

AR-MAP:自回归大语言模型是否是扩散大语言模型的隐式教师?

Liang Lin, Feng Xiong, Zengbin Wang, Kun Wang, Junhao Dong, Xuecai Hu, Yong Wang, Xiangxiang Chu

机构 * AMAP, Alibaba Group(AMAP,阿里巴巴集团) Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 AR-MAP通过利用自回归大语言模型作为隐式教师,有效提升扩散大语言模型的偏好对齐性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02630 2026-02-04 cs.MM cs.AI 85%

Trailer Reimagined: An Innovative, Llm-DRiven, Expressive Automated Movie Summary framework (TRAILDREAMS)

Trailer Reimagined: 一种创新的、基于大语言模型的、富有表现力的自动电影预告片生成框架 (TRAILDREAMS)

Roberto Balestri, Pasquale Cascarano, Mirko Degli Esposti, Guglielmo Pescatore

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 TRAILDREAMS利用大语言模型自动生成电影预告片,通过选择关键视觉片段和对话生成音频元素,提升预告片的吸引力和视觉效果,但在质量上仍需进一步改进以接近人工创作的预告片。

Journal ref OJCMT, 15(3), e202524 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05052 2026-02-04 cs.CR cs.CL 85%

Proactive defense against LLM Jailbreak

主动防御对抗大语言模型劫持

Weiliang Zhao, Jinjun Peng, Daniel Ben-Levi, Zhou Yu, Junfeng Yang

机构 * Department of Computer Science, Columbia University, US(计算机科学系,哥伦比亚大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 ProAct通过主动误导劫持方法,显著降低攻击成功率,提升大语言模型的安全性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03478 2026-02-04 cs.AI 79%

When Routing Collapses: On the Degenerate Convergence of LLM Routers

当路由崩溃:关于LLM路由的退化收敛

Guannan Lai, Han-Jia Ye

机构 * School of Artificial Intelligence, Nanjing University, China(人工智能学院,南京大学,中国) National Key Laboratory for Novel Software Technology, Nanjing University, China(新型软件技术国家重点实验室,南京大学,中国)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 本文提出EquiRouter,一种决策意识的路由器,通过直接学习模型排名来缓解LLM路由中的路由崩溃问题,有效降低计算和货币成本。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22513 2026-02-04 cs.AI 79%

Why Self-Rewarding Works: Theoretical Guarantees for Iterative Alignment of Language Models

为何自我奖励有效:语言模型迭代对齐的理论保证

Shi Fu, Yingjie Wang, Shengchao Hu, Peng Wang, Dacheng Tao

机构 * Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 本文通过理论分析揭示了自我奖励语言模型在迭代对齐中的有效性,证明了其性能随迭代次数呈指数衰减,并为线性softmax模型提供了定制化的理论保证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03468 2026-02-04 cs.AI cs.LG 79%

IntentRL: Training Proactive User-intent Agents for Open-ended Deep Research via Reinforcement Learning

IntentRL: 通过强化学习训练主动的用户意图代理以进行开放性深度研究

Haohao Luo, Zexi Li, Yuexiang Xie, Wenhao Zhang, Yaliang Li, Ying Shen

机构 * Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 IntentRL通过强化学习训练主动用户意图代理,提升开放性深度研究的意图识别与任务性能。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05742 2026-02-04 cs.HC 78%

Vipera: Blending Visual and LLM-Driven Guidance for Systematic Auditing of Text-to-Image Generative AI

Vipera:融合视觉与LLM驱动的指导以系统性审计文本到图像生成AI

Yanwei Huang, Wesley Hanwen Deng, Sijia Xiao, Motahhare Eslami, Jason I. Hong, Arpit Narechania, Adam Perer

专题命中 其他LLM :LLM(title,abstract)

AI总结 Vipera通过融合视觉与LLM驱动的指导,为文本到图像生成AI提供系统性审计方法,提升审计效率与准确性。

Comments 17 pages, 8 figures; Accepted by CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02232 2026-02-04 cs.HC 78%

Eye2Recall: Exploring the Design of Enhancing Reminiscence Activities via Eye Tracking-Based LLM-Powered Interaction Experience for Older Adults

Eye2Recall:探索通过基于眼动的LLM交互体验增强回忆活动的设计

Lei Han, Mingnan Wei, Qiongyan Chen, Anqi Wang, Rong Pang, Kefei Liu, Rongrong Chen, David Yip

专题命中 其他LLM :LLM(title,abstract)

AI总结 Eye2Recall通过结合眼动追踪和自然语言交互,设计了一种增强老年人回忆活动的系统,通过用户研究验证其有效性,旨在提升老年人的福祉和积极老龄化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23440 2026-02-04 cs.AI 77%

Self-Foveate: Enhancing Diversity and Difficulty of Synthesized Instructions from Unsupervised Text via Multi-Level Foveation

自聚焦:通过多级聚焦从无监督文本中增强合成指令的多样性和难度

Mingzhe Li, Xin Lu, Yanyan Zhao

机构 * Research Center for Social Computing and Interactive Robotics(社会计算与交互机器人研究中心) Harbin Institute of Technology(哈尔滨工业大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 Self-Foveate通过多级聚焦方法提升无监督文本合成指令的多样性和难度,采用细粒度到整体模式的提取策略,并结合重合成模块提高指令质量。

Comments Accepted to ACL 2025 (Findings). 23 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03155 2026-02-04 cs.HC 75%

Is It Possible to Make Chatbots Virtuous? Investigating a Virtue-Based Design Methodology Applied to LLMs

能否使聊天机器人变得有德性?探讨一种基于德性的设计方法应用于大语言模型

Matthew P. Lad, Louisa Conwill, Megan Levis Scheirer

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文探讨了基于德性的设计方法在大语言模型中的应用,通过伦理设计模式的编目和评估,提出了一种提升LLM伦理性的设计框架,并指出其在准确性和安全性方面的优势及潜在实施挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02766 2026-02-04 cs.LG cs.CL cs.CR 73%

Privately Fine-Tuned LLMs Preserve Temporal Dynamics in Tabular Data

私有微调的LLM在表格数据中保留时间动态

Lucas Rosenblatt, Peihan Liu, Ryan McKenna, Natalia Ponomareva

机构 * NYU(纽约大学) Columbia University(哥伦比亚大学) Google Research(谷歌研究)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出PATH框架,利用私有微调的LLM保留表格数据的时间动态,有效捕捉长距离依赖并提升合成数据质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03580 2026-02-04 cs.CR cs.AI 70%

Don't believe everything you read: Understanding and Measuring MCP Behavior under Misleading Tool Descriptions

不要轻信你所读到的:在误导性工具描述下理解和测量MCP行为

Zhihao Li, Boyang Ma, Xuelong Dai, Minghui Xu, Yue Zhang, Biwei Yan, Kun Li

机构 * School of Computer Science and Technology(计算机科学与技术学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究揭示MCP协议中描述与代码不一致的问题,发现13%的服务器存在潜在安全风险,强调需加强透明度和审计

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02961 2026-02-04 cs.AI 70%

Generative Engine Optimization: A VLM and Agent Framework for Pinterest Acquisition Growth

生成引擎优化:一个面向Pinterest获取增长的VLM和代理框架

Faye Zhang, Qianyu Cheng, Jasmine Wan, Vishwakarma Singh, Jinfeng Rao, Kofi Boakye

机构 * Stanford University(斯坦福大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 Pinterest提出GEO框架,通过微调VLM和AI代理预测用户搜索需求,构建语义连贯的集合页面,提升生成搜索时代的视觉平台流量增长。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03620 2026-02-04 physics.soc-ph 67%

Toward a new AI winter? How diffusion of technological innovation on networks leads to chaotic boom-bust cycles

迈向新的人工智能寒冬?技术扩散网络如何导致混沌繁荣-衰退周期

Sabin Roman, Francesco Bertolotti

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出一个数学模型,通过技术扩散和投资因素,揭示技术发展中的混沌繁荣-衰退周期,并指出可能引发新的人工智能寒冬。

Journal ref Frontiers in Artificial Intelligence, Vol. 8, 2025, Article 1671917

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24521 2026-02-04 cs.NE 67%

More than MACs: Exploring the Role of Neuromorphic Engineering in the Age of LLMs

超越MACs:探索在LLMs时代神经形态工程的作用

Wilkie Olin-Ammentorp

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨了在LLMs时代神经形态工程对AI系统扩展能力的贡献,通过分析生物与AI计算系统的差异,提出NI启发机制在AI硬件和软件中的应用机遇。

Comments 36 pages, 11 figures, review

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03334 2026-02-04 cs.CY 67%

The Personality Trap: How LLMs Embed Bias When Generating Human-Like Personas

人格陷阱:大语言模型在生成类人人格时嵌入偏见的方式

Jacopo Amidei, Gregorio Ferreira, Mario Muñoz Serrano, Rubén Nieto, Andreas Kaltenbrunner

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了大语言模型在生成类人人格时嵌入WEIRD偏见的问题,揭示了LLMs在生成合成人口时可能带来的刻板印象和风险。

Comments 26 pages, 2 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03227 2026-02-04 cs.CV 67%

Spiral RoPE: Rotate Your Rotary Positional Embeddings in the 2D Plane

Spiral RoPE: 在二维平面上旋转你的旋转位置嵌入

Haoyu Liu, Sucheng Ren, Tingyu Zhu, Peng Wang, Cihang Xie, Alan Yuille, Zeyu Zheng, Feng Wang

机构 * University of California, Berkeley(加州大学伯克利分校) University of California, Santa Cruz(加州大学圣克ruz分校) John Hopkins University(约翰霍普金斯大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Spiral RoPE通过多方向位置编码改进视觉变换器在分类、分割和生成任务中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02684 2026-02-04 cs.HC 67%

ADx3: A Collaborative Workflow for High-Quality Accessible Audio Description

ADx3:高质量可及音频描述的协作流程

Lana Do, Shasta Ihorn, Charity Pitcher-Cooper, Juvenal Francisco Barajas, Gio Jung, Xuan Duy Anh Nguyen, Sanjay Mirani, Ilmi Yoon

专题命中 其他LLM :language model(abstract);prompting(abstract)

AI总结 ADx3通过整合GenAD、RefineAD和AdaptAD模块,实现高质量可及音频描述的协作流程,提升描述质量和用户交互体验。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00755 2026-02-04 cs.MA cs.AI cs.NE 57%

Evolving Interpretable Constitutions for Multi-Agent Coordination

为多智能体协调演化可解释的宪法

Ujwal Kumar, Alice Saito, Hershraj Niranjani, Rayan Yessou, Phan Xuan Tan

机构 * College of Engineering Shibaura Institute of Technology(Shibaura Institute of Technology 工程学院) Faculty of Arts and Sciences The University of Tokyo(东京大学 文理学部) Department of EECS University of California, Berkeley(加州大学伯克利分校 电子工程与计算机科学系) Department of Informatics Università degli Studi di Milano-Bicocca(米兰-比科卡大学 信息学系)

专题命中 其他LLM :LLM(abstract);分类 cs.AI

AI总结 本文提出通过进化算法自动发现多智能体协调的可解释宪法,通过实验展示进化宪法在提升社会稳定性方面的有效性。

Comments 23 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02785 2026-02-04 cs.HC 50%

Smell with Genji: Rediscovering Human Perception through an Olfactory Game with AI

用Genji嗅觉:通过AI嗅觉游戏重新发现人类感知

Awu Chen, Vera Yu Wu, Yunge Wen, Yaluo Wang, Jiaxuan Olivia Yin, Yichen Wang, Qian Xiang, Richard Zhang, Paul Pu Liang, Hiroshi Ishii

专题命中 其他LLM :LLM(abstract)

AI总结 Smell with Genji通过AI嗅觉游戏重新发现人类感知,结合人机协作体验促进嗅觉反思与交互。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02534 2026-02-04 cs.SI 50%

DualMind: Towards Understanding Cognitive-Affective Cascades in Public Opinion Dissemination via Multi-Agent Simulation

DualMind: 通过多智能体模拟理解公共意见传播中的认知-情感 cascades

Enhao Huang, Tongtong Pan, Shuhuai Zhang, Qishu Jin, Liheng Zheng, Kaichun Hu, Yiming Li, Zhan Qin, Kui Ren

专题命中 其他LLM :LLM(abstract)

AI总结 DualMind通过多智能体模拟,建模公共意见传播中认知与情感的相互作用,提升危机管理的预测能力。

Comments Accepted as a demo paper at TheWebConf (WWW) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏