ReFORM: Reflected Flows for On-support Offline RL via Noise Manipulation
ReFORM:通过噪声操控实现支持下的离线强化学习
机构 * MIT(麻省理工学院) ; Boston University(波士顿大学) ; MIT Lincoln Laboratory(麻省理工学院林伍德实验室)
专题命中 模仿学习与强化学习 :manipulation(title);分类 cs.RO、cs.AI、cs.LG
AI总结 ReFORM通过反射流策略和噪声操控,在离线强化学习中实现更宽松的支持约束,从而在多模态分布下提升策略性能。
Comments 24 pages, 17 figures; Accepted by the fourteenth International Conference on Learning Representations (ICLR 2026)