arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4786 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4786 篇

2512.19154 2025-12-23 cs.LG cs.AI 62%

Beyond Sliding Windows: Learning to Manage Memory in Non-Markovian Environments

超越滑动窗口:在非马尔可夫环境中学习管理内存

Geraud Nangue Tasse, Matthew Riemer, Benjamin Rosman, Tim Klinger

机构 * CSAM School, University of the Witwatersrand(瓦茨堡大学CSAM学院) IBM Research(IBM研究院) Mila, Université de Montréal(蒙特利尔大学Mila)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出自适应堆叠元算法,通过维护较小的记忆堆栈,在非马尔可夫环境中减少计算和内存需求,同时有效管理记忆以避免过度删除重要经验。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18940 2025-12-23 cs.CL cs.SE 62%

FASTRIC: Prompt Specification Language for Verifiable LLM Interactions

FASTRIC:用于可验证大语言模型交互的提示规范语言

Wen-Long Jin

机构 * Department of Civil and Environmental Engineering(土木与环境工程系) California Institute for Telecommunications and Information Technology(电信与信息科技学院) Institute of Transportation Studies(交通研究学院) University of California, Irvine, CA 92697-3600(加州大学伊维德分校)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL、cs.SE

AI总结 FASTRIC通过显式化有限状态机构建可验证的大语言模型交互协议,揭示模型特定的规范正式程度范围,实现交互设计的系统化工程。

Comments 13 pages, 3 figures. Supplementary materials at https://doi.org/10.17605/OSF.IO/PV6R3

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04680 2025-12-05 cs.SE cs.AI cs.HC 62%

Generative AI for Self-Adaptive Systems: State of the Art and Research Roadmap

生成式人工智能在自适应系统中的应用:现状与研究路线图

Jialong Li, Mingyue Zhang, Nianyu Li, Danny Weyns, Zhi Jin, Kenji Tei

机构 * Waseda University(早稻田大学) Southwest University(西南大学) Zhongguancun Laboratory(中关村实验室) Peking University(北京大学) Tokyo Institute of Technology(东京技术大学)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI、cs.SE

AI总结 本文探讨生成式人工智能在自适应系统中的应用现状与研究方向,分析其提升系统自主性和人机交互的潜力及挑战。

Comments Accepted by ACM Transactions on Autonomous and Adaptive Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01167 2025-12-03 cs.LG cs.AI cs.SY eess.SY 62%

A TinyML Reinforcement Learning Approach for Energy-Efficient Light Control in Low-Cost Greenhouse Systems

为低成本温室系统设计一种 TinyML 强化学习方法以实现节能照明控制

Mohamed Abdallah Salem, Manuel Cuevas Perez, Ahmed Harb Rabia

机构 * North Dakota State University(北达科他州立大学) Biosystems Engineering(生物系统工程)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种基于 TinyML 的强化学习方法,用于低成本温室系统的节能照明控制,通过 Q 学习算法实现动态亮度调节,有效稳定不同光照水平。

Comments Copyright 2025 IEEE. This is the author's version of the work that has been accepted for publication in Proceedings of the 5. Interdisciplinary Conference on Electrics and Computer (INTCEC 2025) 15-16 September 2025, Chicago-USA. The final version of record is available at: https://doi.org/10.1109/INTCEC65580.2025.11256135

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07338 2025-12-02 cs.AI cs.LG 62%

DeepPersona: A Generative Engine for Scaling Deep Synthetic Personas

DeepPersona: 一个用于扩展深度合成人设的生成引擎

Zhen Wang, Yufan Zhou, Zhongyan Luo, Lyumanshan Ye, Adam Wood, Man Yao, Saab Mansour, Luoshang Pan

机构 * UCSD(加州大学圣地亚哥分校) KU Leuven(鲁汶大学) SJTU(上海交通大学) University of Michigan(密歇根大学) Denison University(德尼森大学) Amazon(亚马逊) Meta

专题命中 记忆与上下文管理 :agentic(abstract);分类 cs.AI、cs.LG

AI总结 DeepPersona通过双阶段分类学引导方法生成深度合成人设,提升LLM个性化和人类模拟的准确性与多样性。

Comments add an author[Update], 12 pages, 5 figures, accepted at LAW 2025 Workshop (NeurIPS 2025) Project page: https://deeppersona-ai.github.io/

Journal ref LAW 2025 Workshop, NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20870 2025-11-27 cs.LG cs.AI stat.ML 62%

Selecting Belief-State Approximations in Simulators with Latent States

在具有潜在状态的模拟器中选择信念状态近似

Nan Jiang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种在具有潜在状态的模拟器中选择信念状态近似的方法,通过归约条件分布选择任务,探讨了两种不同的形式化方法及其在不同roll-out方法下的表现差异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18319 2025-11-25 cs.AI cs.LG cs.SY eess.SY 62%

Weakly-supervised Latent Models for Task-specific Visual-Language Control

弱监督潜在模型用于任务特定的视觉语言控制

Xian Yeow Lee, Lasitha Vidyaratne, Gregory Sin, Ahmed Farahat, Chetan Gupta

机构 * Industrial AI Lab, Hitachi America, Ltd.(日立美国有限公司工业人工智能实验室)

专题命中 记忆与上下文管理 :AI agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种任务特定的潜在动态模型,利用目标状态监督学习动作诱导位移,以提高空间定位任务中的视觉语言控制性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17671 2025-11-25 cs.CR cs.AI cs.CL 62%

MURMUR: Using cross-user chatter to break collaborative language agents in groups

利用跨用户交流打破协作语言代理组

Atharv Singh Patlan, Peiyao Sheng, S. Ashwin Hebbar, Prateek Mittal, Pramod Viswanath

机构 * Princeton University(普林斯顿大学) Sentient

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.CL

AI总结 MURMUR通过生成真实用户交互,揭示了跨用户污染攻击对多用户语言代理的威胁,并提出基于任务的聚类作为初步防御措施。

Comments 20 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13746 2025-11-19 eess.SY cs.AI cs.LG cs.SY 62%

Deep reinforcement learning-based spacecraft attitude control with pointing keep-out constraint

Juntang Yang, Mohamed Khalil Ben-Larbi

机构 * University of Würzburg(乌尔姆大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11362 2025-11-17 cs.LG cs.CL 62%

On-Device Fine-Tuning via Backprop-Free Zeroth-Order Optimization

Prabodh Katti, Sangwoo Park, Bipin Rajendran, Osvaldo Simeone

机构 * Institute for Intelligent Networked Systems, Northeastern University London(智能网络系统研究所,东北大学伦敦分校) Department of Engineering, King’s College London(工程系,伦敦大学国王学院)

专题命中 记忆与上下文管理 :agentic(abstract);分类 cs.CL、cs.LG

Comments Conference submission; Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07701 2025-11-12 cs.LG cs.AI 62%

Diffusion Guided Adversarial State Perturbations in Reinforcement Learning

Xiaolin Sun, Feidi Liu, Zhengming Ding, ZiZhan Zheng

机构 * Department of Computer Science, Tulane University(Tulane大学计算机科学系) Shanghai Center for Mathematical Science, Fudan University(复旦大学上海数学科学中心)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Journal ref NeurIPS 2025 Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03237 2025-11-12 q-bio.QM cs.AI cs.LG q-bio.BM 62%

UniSite: The First Cross-Structure Dataset and Learning Framework for End-to-End Ligand Binding Site Detection

Jigang Fan, Quanlin Wu, Shengjie Luo, Liwei Wang

机构 * Center for Data Science, Peking University(数据科学中心,北京大学) State Key Laboratory of General Artificial Intelligence, Peking University(通用人工智能国家重点实验室,北京大学) Center for Machine Learning Research, Peking University(机器学习研究中心,北京大学)

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.AI、cs.LG

Comments Accepted by NeurIPS 2025 as a Spotlight paper

Journal ref NeurIPS 2025 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17830 2025-11-05 cs.LG cs.AI 62%

Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning

Nicolas Castanet, Olivier Sigaud, Sylvain Lamprier

机构 * Sorbonne Université, CNRS, ISIR(索邦大学、国家科学研究中心、信息科学研究所)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.10712 2025-11-04 cs.LG cs.AI 62%

Neighboring State-based Exploration for Reinforcement Learning

Yu-Teng Li, Justin Lin, Jeffery Cheng, Pedro Pachuca

机构 * UC Berkeley(伯克利大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.12308 2025-10-30 cs.AI cs.LG cs.NE cs.RO 62%

SNN-Based Online Learning of Concepts and Action Laws in an Open World

Christel Grimaud, Dominique Longin, Andreas Herzig

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21107 2025-10-27 cs.LG cs.AI cs.RO 62%

ESCORT: Efficient Stein-variational and Sliced Consistency-Optimized Temporal Belief Representation for POMDPs

Yunuo Zhang, Baiting Luo, Ayan Mukhopadhyay, Gabor Karsai, Abhishek Dubey

机构 * Vanderbilt University(范德比大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Proceeding of the 39th Conference on Neural Information Processing Systems (NeurIPS'25). Code would be available at https://github.com/scope-lab-vu/ESCORT

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16313 2025-10-23 cs.LG cs.AI 62%

Improved Exploration in GFlownets via Enhanced Epistemic Neural Networks

Sajan Muhammad, Salem Lahlou

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted to the EXAIT Workshop at ICML 2025, and ICoIAS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17874 2025-10-22 cs.SE cs.AI 62%

Repairing Tool Calls Using Post-tool Execution Reflection and RAG

Jason Tsay, Zidane Wright, Gaodan Fang, Kiran Kate, Saurabh Jha, Yara Rizk

机构 * IBM Research(IBM研究院)

专题命中 记忆与上下文管理 :agentic(abstract);分类 cs.AI、cs.SE

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16816 2025-10-21 cs.LG cs.AI math-ph math.MP physics.comp-ph 62%

Efficient High-Accuracy PDEs Solver with the Linear Attention Neural Operator

Ming Zhong, Zhenya Yan

机构 * School of Advanced Interdisciplinary Sciences, University of Chinese Academy of Sciences(中国科学院大学先进交叉学科学院) State Key Laboratory of Mathematical Sciences, Academy of Mathematics and Systems Science, Chinese Academy of Sciences(中国科学院数学科学国家重点实验室) School of Mathematics and Information Science, Zhongyuan University of Technology(中原工学院数学与信息科学学院)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 31 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13366 2025-10-16 cs.CL cs.AI 62%

Document Intelligence in the Era of Large Language Models: A Survey

Weishi Wang, Hengchang Hu, Zhijie Zhang, Zhaochen Li, Hongxin Shao, Daniel Dahlmeier

机构 * SAP, Singapore(新加坡SAP)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00895 2025-10-14 cs.LG cs.AI 62%

State-Covering Trajectory Stitching for Diffusion Planners

Kyowoon Lee, Jaesik Choi

机构 * KAIST(韩国科学技术院)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI、cs.LG

Comments Accepted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04901 2025-10-07 cs.LG cs.AI 62%

Focused Skill Discovery: Learning to Control Specific State Variables while Minimizing Side Effects

Jonathan Colaço Carr, Qinyi Sun, Cameron Allen

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Reinforcement Learning Journal 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02484 2025-10-06 cs.LG cs.AI 62%

From Pixels to Factors: Learning Independently Controllable State Variables for Reinforcement Learning

Rafael Rodriguez-Sanchez, Cameron Allen, George Konidaris

机构 * Brown University(布朗大学) UC Berkeley(伯克利大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11679 2025-10-02 cs.CL cs.LG 62%

Ambiguity in LLMs is a concept missing problem

Zhibo Hu, Chen Wang, Yanfeng Shu, Hye-Young Paik, Liming Zhu

机构 * The University of New South Wales(新南威尔士大学) CSIRO Data61(澳大利亚联邦科学与工业研究组织数据61)

专题命中 记忆与上下文管理 :agentic(abstract);分类 cs.CL、cs.LG

Comments 17 pages, 11 figures, title updated

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24556 2025-09-30 cs.LG cs.AI physics.flu-dyn 62%

Deep Reinforcement Learning in Action: Real-Time Control of Vortex-Induced Vibrations

Hussam Sababha, Bernat Font, Mohammed Daqaq

机构 * Department of Mechanical Engineering, Tandon School of Engineering, New York, USA(机械工程系,工程学院,美国纽约) Department of Mechanical Engineering, Delft University of Technology, Delft, Netherlands(机械工程系,代尔夫特理工大学,荷兰代尔夫特)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04495 2025-09-26 cs.LG cs.CL 62%

Causal Reflection with Language Models

Abi Aryan, Zac Liu

机构 * Abide AI

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.14596 2025-09-19 cs.CV cs.AI cs.LG 62%

VLM Agents Generate Their Own Memories: Distilling Experience into Embodied Programs of Thought

Gabriel Sarch, Lawrence Jang, Michael J. Tarr, William W. Cohen, Kenneth Marino, Katerina Fragkiadaki

机构 * Carnegie Mellon University(卡内基梅隆大学) Google DeepMind(谷歌DeepMind)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Project website: https://ical-learning.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11367 2025-09-16 cs.LG cs.AI 62%

Detecting Model Drifts in Non-Stationary Environment Using Edit Operation Measures

Chang-Hwan Lee, Alexander Shim

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 28 pages, 3 figures, 17 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10531 2025-09-16 cs.LG cs.AI 62%

FinXplore: An Adaptive Deep Reinforcement Learning Framework for Balancing and Discovering Investment Opportunities

Himanshu Choudhary, Arishi Orra, Manoj Thakur

机构 * School of Mathematical & \& Statistical Sciences(数学与统计科学学院) Indian Institute of Technology Mandi(印度理工学院曼迪分校)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10277 2025-09-10 cs.CY cs.AI cs.CL cs.CR 62%

RealHarm: A Collection of Real-World Language Model Application Failures

Pierre Le Jeune, Jiaen Liu, Luca Rossi, Matteo Dora

专题命中 记忆与上下文管理 :AI agent(abstract);分类 cs.AI、cs.CL

Journal ref ACL Proceedings of the The First Workshop on LLM Security (LLMSEC), pp. 87-100, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏