Combining Tree-Search, Generative Models, and Nash Bargaining Concepts in Game-Theoretic Reinforcement Learning
将树搜索、生成模型和纳什谈判概念结合在博弈论强化学习中
机构 * Google DeepMind(谷歌DeepMind) ; University of Waterloo(滑铁卢大学) ; University of Michigan(密歇根大学)
AI总结 本文提出基于深度博弈论强化学习的多智能体训练框架,通过生成最佳响应算法GenBR和纳什谈判理论构建对手混合策略,提升在不完全信息域中的对手建模能力。
Comments Accepted by IJCAI'25 main track
Journal ref Proc. 34th Int. Joint Conf. Artif. Intell. (IJCAI 2025), pp. 161-169