Breaking the Performance Ceiling in Reinforcement Learning requires Inference Strategies
在强化学习中突破性能极限需要推理策略
Felix Chalumeau, Daniel Rajaonarivonivelomanantsoa, Ruan de Kock, Claude Formanek, Sasha Abramowitz, Oumayma Mahjoub, Wiem Khlifi, Simon Du Toit, Louay Ben Nessir, Refiloe Shabe, Noah De Nicola, Arnol Fokam, Siddarth Singh, Ulrich Mbou Sob, Arnu Pretorius
机构
*
School of Intelligence Science and Engineering, Harbin Institute of Technology, Shenzhen(哈尔滨工业大学智能科学与工程学院)
;
Baidu Inc.(百度公司)
;
Leiden University(莱顿大学)
;
Pengcheng Laboratory(鹏城实验室)
机构
*
The University of Tokyo(东京大学)
;
Institute of Science Tokyo(东京科学研究所)
;
RIKEN Center for Advanced Intelligence Project(理化学研究所先进情报项目中心)
;
Griffith University(格里菲斯大学)