Rainbow Delay Compensation: A Multi-Agent Reinforcement Learning Framework for Mitigating Delayed Observation
机构 * Laboratory of Speech and Intelligent Information Processing, Institute of Acoustics, CAS(语音与智能信息处理实验室,声学研究所,中国科学院) ; University of Chinese Academy of Sciences(中国科学院大学) ; Department of Electronic Engineering, Tsinghua University(清华大学电子工程系)
Comments The code has been open-sourced in the RDC-pymarl project under https://github.com/linkjoker1006