AI 中文总结
研究编码代理记忆问题,提出基于认知文献的两层设计理论及线索锚定记忆模型,通过实际编码任务评估,表明传递而非存储是关键,可靠记忆通道应让代理无需思考,给出相关实验结果及贡献。
AI 中文摘要
编码代理有一种记忆:文档。指令文件、计划工件和自动编写的内存目录是特意编写和检索的,代理必须选择写入和读取它们。人类专业知识存在于第二层,永远不会被记录下来:情境相关的操作事实(陷阱、位置、本地约定)作为工作的副作用被编码,并在情境提示时自动检索。我们认为第二层是长期运行代理的承载层,必须是一种约束属性,而非代理选择。我们的贡献包括:(1)基于记忆卸载、偶然编码和基于事件的前瞻记忆的认知文献的两层设计理论,各映射到一个架构要求;(2)线索锚定记忆模型,其中记忆在可组合词汇表(路径、符号、语义、事件、时间)上携带一流触发条件,由约束确定性评估,这是现有学术或系统未提供的组合;(3)对实际编码任务的控制评估,表明即使有预播种存储,自愿记忆使用也接近零(114轮中0次内存操作),确定性注入在每次播种运行中都能实现且零误报,39%的会话内重新读取会重新购买压缩边界前付费的内容;(4)重复压缩衰减探测:仅在对话中持有的十个事实在第一次总结时消失,在108次压缩中的106次中都不存在,被剥夺的代理会搜索约束自身的会话文件来重建它们,而从约束拥有的存储中注入的相同事实在所有138次压缩恢复中都完整到达,因为最终总结中没有这些事实。传递而非存储才是产品:代理的可靠记忆通道是代理无需思考的通道。
英文摘要
Coding agents ship with one kind of memory: documents. Instruction files, plan artifacts, and auto-written memory directories are deliberately authored and deliberately retrieved: the agent must choose to write them and choose to read them back. Human expertise runs on a second tier that never gets written down: situationally-bound operational facts (gotchas, locations, local conventions) encoded as a side effect of the work and retrieved involuntarily when the situation cues them. We argue this second tier is the load-bearing one for long-running agents and must be a harness property, not an agent choice. We contribute: (1) a two-tier design theory grounded in the cognitive literature on memory offloading, incidental encoding, and event-based prospective memory, each mapped to an architectural requirement; (2) a cue-anchored memory model where memories carry first-class trigger conditions over a composable vocabulary (path, symbol, semantic, event, temporal), evaluated deterministically by the harness, a composition no surveyed academic or shipped system provides; (3) a controlled evaluation on a real coding task showing that voluntary memory use is near zero even with a pre-seeded store (0 memory operations in 114 turns), that deterministic injection delivered in every seeded run with zero false alarms, and that 39% of intra-session re-reads re-buy content paid for before a compaction boundary; (4) a repeated-compaction decay probe: ten facts held only in conversation vanish at the first summary and stay absent from 106 of 108 compactions, and the deprived agent greps the harness's own session files to rebuild them, while the same facts injected from a harness-owned store arrive intact through all 138 compact-resumes as the final summary carries none. Delivery, not storage, is the product: the reliable memory channel for agents is the one the agent never has to think about.