Scaling Offline Model-Based RL via Jointly-Optimized World-Action Model Pretraining
通过联合优化的世界-动作模型扩展离线模型基于的强化学习
机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院) ; School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) ; Alibaba Group(阿里巴巴集团)
专题命中 模型式强化学习 :model-based RL(title,abstract);world model(abstract);world model(abstract);分类 cs.AI、cs.LG
AI总结 JOWA通过联合优化的世界-动作模型扩展离线RL,实现高效泛化和高性能
Comments Accepted by ICLR 2025