Reinforcement Learning with Exogenous States and Rewards
具有外源状态和奖励的强化学习
机构 * Intercom Dublin(Intercom 布鲁姆斯贝克) ; Collaborative Robotics and Intelligent Systems (CoRIS) Institute(协同机器人与智能系统研究所) ; Oregon State University(俄勒冈州立大学)
AI总结 本文提出了一种分解外源与内源状态和奖励的方法,通过分离外源噪声以提升强化学习效率。
Comments Substantial rewrite to improve rigor and clarity in response to referee reports at JMLR