MVR: Multi-view Video Reward Shaping for Reinforcement Learning
MVR:多视图视频奖励塑造用于强化学习
机构 * School of Intelligence Science and Technology, Peking University(北京大学智能科学与技术学院) ; State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室)
专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.CV、cs.LG
AI总结 MVR通过多视角视频和视觉语言模型提升强化学习中的奖励塑造,有效解决复杂动态任务中的状态相关性和视角偏见问题。
Comments ICLR 2026