ARROW: Augmented Replay for RObust World models
ARROW:增强重放用于鲁棒世界模型
机构 * Imam Mohammad Ibn Saud Islamic University (IMSIU)(伊玛姆·穆罕默德·本·沙特伊斯兰大学) ; Monash University(莫纳什大学) ; University of New South Wales, Sydney(新南威尔士大学,悉尼) ; Cerenaut
AI总结 本文提出ARROW算法,一种基于模型的持续强化学习方法,通过高效的重放缓冲区减少灾难性遗忘,提升在无共享结构任务和有共享结构任务中的表现。
Comments 36 pages and 11 figures (includes Appendix)
Journal ref Transactions on Machine Learning Research, 2026