arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-19 至 2025-12-19 共收录 8 信号源:cs.CL, cs.AI, cs.LG

1. 长上下文与记忆 8 篇

2512.16532 2025-12-19 cs.AI cs.IR 77%

From Personalization to Prejudice: Bias and Discrimination in Memory-Enhanced AI Agents for Recruitment

从个性化到偏见:记忆增强型AI招聘代理中的偏见与歧视

Himanshu Gharat, Himanshi Agrawal, Gourab K. Patro

机构 * Phi Labs, Quantiphi Inc.(Phi实验室、Quantiphi公司)

专题命中 长上下文与记忆 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究揭示了记忆增强型AI招聘代理中偏见的系统性引入与强化,强调了对LLM基础AI代理的额外防护措施的必要性。

Comments In Proceedings of the Nineteenth ACM International Conference on Web Search and Data Mining (WSDM '26)

Journal ref In Proceedings of the Nineteenth ACM International Conference on Web Search and Data Mining (WSDM '26), 2026, Boise, ID, USA. ACM, New York, NY, USA

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.15790 2025-12-19 cs.CR 75%

Bilevel Optimization for Covert Memory Tampering in Heterogeneous Multi-Agent Architectures (XAMT)

针对异构多智能体架构中的隐蔽内存篡改的双层优化

Akhil Sharma, Shaikh Yaser Arafat, Jai Kumar Sharma, Ken Huang

专题命中 长上下文与记忆 :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 XAMT通过双层优化方法针对异构多智能体架构中的隐蔽内存篡改问题,提出了一种新的训练时威胁模型,以提升系统安全性。

Comments 10 pages, 5 figures, 4 tables. Conference-style paper (IEEEtran). Proposes unified bilevel optimization framework for covert memory poisoning attacks in heterogeneous multi-agent systems (MARL + RAG)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.14516 2025-12-19 cs.CR cs.CV cs.LG 70%

Memory Backdoor Attacks on Neural Networks

神经网络中的记忆后门攻击

Eden Luzon, Guy Amit, Roy Weiss, Torsten Kraub, Alexandra Dmitrienko, Yisroel Mirsky

机构 * Ben-Gurion University, Institute of Software Systems and Security(本·加隆大学软件系统与安全研究所) University of Würzburg(魏泽堡大学)

专题命中 长上下文与记忆 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出了一种神经网络的记忆后门攻击方法,通过恶意服务器系统性地提取客户端训练样本,揭示了联邦学习中的隐私安全漏洞。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.16262 2025-12-19 cs.AI 70%

Learning to Wait: Synchronizing Agents with the Physical World

学习等待:与物理世界同步智能体

Yifei She, Ping Zhang, He Liu, Yanmin Jia, Yang Jing, Zijun Liu, Peng Sun, Xiangbin Li, Xiaohe Hu

机构 * Infrawaves Shanghai Qiji Zhifeng Co., Ltd.(上海启智风有限公司) Tsinghua University(清华大学)

专题命中 长上下文与记忆 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出了一种智能体侧方法,使大语言模型通过预测等待时间实现与物理世界的同步,解决了异步环境中动作延迟的问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02262 2025-12-19 cs.CV 67%

From Frames to Clips: Training-free Adaptive Key Clip Selection for Long-Form Video Understanding

从帧到片段:训练免费的自适应关键片段选择用于长形式视频理解

Guangyu Sun, Archit Singhal, Burak Uzkent, Mubarak Shah, Chen Chen, Garin Kessler

机构 * Amazon(亚马逊公司) University of Central Florida(中央佛罗里达大学)

专题命中 长上下文与记忆 :large language model(abstract);language model(abstract)

AI总结 F2C通过自适应关键片段选择提升长视频理解,比均匀采样在多个基准上表现更优。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24320 2025-12-19 cs.LG stat.ML 57%

AuON: A Linear-time Alternative to Orthogonal Momentum Updates

AuON: 一种线性时间的替代正交动量更新方法

Dipan Maity

专题命中 长上下文与记忆 :language model(abstract);分类 cs.LG

AI总结 AuON是一种线性时间优化器,通过归一化非线性缩放实现替代单位范数动量更新,无需近似正交矩阵,有效处理爆炸性注意力logits并在语言建模任务中优于Muon

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.16285 2025-12-19 cs.HC 50%

Machines, AI and the past//future of things

机器、人工智能与事物的过去//未来

Karola Köpferl, Albrecht Kurze

专题命中 长上下文与记忆 :language model(abstract)

AI总结 本文通过重新激活1980年代东德打字机,探讨人工智能与历史技术的交互,挑战进步叙事,强调慢速与物质性在设计中的价值。

Comments State of Responsible Technology 2025 - Generative Things. pp 13-19. Stichting ThingsCon Amsterdam

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.15309 2025-12-19 cs.RO 50%

DynamicGSG: Dynamic 3D Gaussian Scene Graphs for Environment Adaptation

DynamicGSG: 动态3D高斯场景图用于环境适应

Luzhou Ge, Xiangyu Zhu, Zhuo Yang, Xuesong Li

机构 * School of Computer Science, Beijing Institute of Technology, China(计算机学院,北京理工大学)

专题命中 长上下文与记忆 :language model(abstract)

AI总结 DynamicGSG通过动态高斯场景图实现环境适应,利用高斯溅射技术构建层次化场景图,提升环境理解和适应能力。

Journal ref 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)

详情

展开后加载摘要…

URL PDF HTML 收藏