LLMs Are Not Good Strategists, Yet Memory-Enhanced Agency Boosts Reasoning
大型语言模型(LLMs)并非优秀的战略家,然而记忆增强的智能体可提升推理能力
机构 * University of Chicago(芝加哥大学) ; University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
AI总结 该研究针对LLM在长时环境中战略推理的缺陷,提出EpicStar框架,结合跨回合记忆与动态门控机制,在星际争霸II测试中实现更高胜率且令牌消耗大幅减少。
Journal ref Published at Reasoning and Planning for LLMs at ICLR 2025