arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2025-12-03 至 2025-12-03 共收录 2 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 2 篇

2512.01167 2025-12-03 cs.LG cs.AI cs.SY eess.SY 62%

A TinyML Reinforcement Learning Approach for Energy-Efficient Light Control in Low-Cost Greenhouse Systems

为低成本温室系统设计一种 TinyML 强化学习方法以实现节能照明控制

Mohamed Abdallah Salem, Manuel Cuevas Perez, Ahmed Harb Rabia

机构 * North Dakota State University(北达科他州立大学) Biosystems Engineering(生物系统工程)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种基于 TinyML 的强化学习方法,用于低成本温室系统的节能照明控制,通过 Q 学习算法实现动态亮度调节,有效稳定不同光照水平。

Comments Copyright 2025 IEEE. This is the author's version of the work that has been accepted for publication in Proceedings of the 5. Interdisciplinary Conference on Electrics and Computer (INTCEC 2025) 15-16 September 2025, Chicago-USA. The final version of record is available at: https://doi.org/10.1109/INTCEC65580.2025.11256135

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19686 2025-12-03 cs.AI 57%

From Memories to Maps: Mechanisms of In-Context Reinforcement Learning in Transformers

从记忆到地图:变换器中上下文强化学习的机制

Ching Fang, Kanaka Rajan

机构 * Harvard Medical School(哈佛医学院) Harvard University(哈佛大学)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI

AI总结 研究通过变换器模型探索上下文强化学习的机制,发现记忆在存储经验与缓存计算中扮演关键角色,支持灵活行为。

Comments Revised to around 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏