WM-R1: Training GUI Agents to Reason and leverage World Models with Reinforcement Learning
WM-R1:用强化学习训练GUI智能体以推理并利用世界模型
机构 * School of Computer Science and Technology, East China Normal University(华东师范大学计算机科学与技术学院)
专题命中 通用世界模型 :world model(title,abstract);world models(title,abstract);world model(title,abstract);world models(title,abstract)
AI总结 该研究提出首个用世界模型替代真实环境训练GUI智能体的强化学习框架WM-R1,在Android基准测试中其性能显著优于仅GRPO的基线方法和推理时模拟方法。