$K$-Level Policy Gradients for Multi-Agent Reinforcement Learning
Aryaman Reddi, Gabriele Tiboni, Jan Peters, Carlo D'Eramo
机构
*
Department of Computer Science, TU Darmstadt(图恩-达姆施塔特技术大学计算机科学系)
;
Hessian Center for Artificial Intelligence (Hessian.ai)(黑森人工智能中心)
;
German Research Center for AI (DFKI)(德国人工智能研究中心)
;
Center for Cognitive Science, TU Darmstadt(图恩-达姆施塔特技术大学认知科学中心)
;
Center for Artificial Intelligence and Data Science, University of Würzburg(弗赖堡大学人工智能与数据科学中心)
A learning-driven automatic planning framework for proton PBS treatments of H&N cancers
Qingqing Wang, Liqiang Xiao, Chang Chang
机构
*
Department of Radiation Medicine and Applied Sciences, University of California at San Diego(放射医学与应用科学系,加州大学圣地亚哥分校)
;
Amazon Inc.(亚马逊公司)
;
California Protons Cancer Therapy Center(加利福尼亚质子癌症治疗中心)
CORB-Planner: Corridor as Observations for RL Planning in High-Speed Flight
Yechen Zhang, Bin Gao, Gang Wang, Jian Sun, Zhuo Li
机构
*
State Key Laboratory of Autonomous Intelligent Unmanned Systems, School of Automation, Beijing Institute of Technology(自主智能无人系统国家重点实验室,自动化学院,北京理工大学)
机构
*
School of Automation, Northwestern Polytechnical University(自动化学院,西北工业大学)
;
The University of Hong Kong(香港大学)
;
School of Software, Northwestern Polytechnical University(软件学院,西北工业大学)
;
Unmanned System Research Institute, Northwestern Polytechnical University(无人系统研究院,西北工业大学)
Hi-DARTS: Hierarchical Dynamically Adapting Reinforcement Trading System
Hoon Sagong, Heesu Kim, Hanbeen Hong
机构
*
Dept. of Computer Engineering Hongik University Seoul, South Korea
;
Dept. of Applied Data Science Sungkyunkwan University Seoul, South Korea
;
Dept. of Economics Hankuk University of Foreign Studies Seoul, South Korea