Online Action-Stacking Improves Reinforcement Learning Performance for Air Traffic Control
在线动作堆叠提升空交通管制的强化学习性能
机构 * Project Bluebird, NATS(Project Bluebird,NATS) ; Department of Computer Science, University of Exeter(计算机科学系,埃克塞特大学) ; Project Bluebird, The Alan Turing Institute(Project Bluebird,艾伦·图灵研究所) ; The Alan Turing Institute and the University of Exeter(艾伦·图灵研究所和埃克塞特大学)
专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO、cs.AI、cs.LG
AI总结 在线动作堆叠通过简化动作空间提升空交通管制的强化学习性能,有效减少指令数量并实现与复杂动作空间相当的控制效果。