Advancing Robustness in Deep Reinforcement Learning with an Ensemble Defense Approach
Adithya Mohan, Dominik Rößle, Daniel Cremers, Torsten Schön
机构
*
AImotion Bavaria, Technische Hochschule Ingolstadt(巴伐利亚AImotion、英格尔施泰特技术大学)
;
School of Computation, Information and Technology, TU Munich(计算、信息与技术学院,慕尼黑技术大学)
Shared Control of Holonomic Wheelchairs through Reinforcement Learning
Jannis Bähler, Diego Paez-Granados, Jorge Peña-Queralta
机构
*
Swiss Paraplegic Research, SPF(瑞士瘫痪研究机构)
;
SCAI Lab, D-HEST, Swiss Federal School of Technology in Zurich - ETH Zurich(SCAI实验室,瑞士联邦理工学院-苏黎世-ETH Zurich)
;
Centre for Artificial Ingelligece, Zurich University of Applied Sciences - ZHAW. Switzerland(人工智能中心,瑞士应用科学大学-ZHAW)
SPLASH! Sample-efficient Preference-based inverse reinforcement learning for Long-horizon Adversarial tasks from Suboptimal Hierarchical demonstrations
Peter Crowley, Zachary Serlin, Tyler Paine, Makai Mann, Michael Benjamin, Calin Belta
机构
*
Boston University(波士顿大学)
;
MIT Lincoln Laboratory(麻省理工学院林肯实验室)
;
MIT Pavlab(麻省理工学院 Pavlab 实验室)
;
Woods Hole Oceanographic Institution(伍兹霍尔海洋研究所)
;
University of Maryland, College Park(马里兰大学学院公园分校)
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Tsinghua University(清华大学)
;
Tencent AI Lab(腾讯AI实验室)
机构
*
Centre for Cognitive Modelling, Moscow Institute of Physics and Technology(认知建模中心,莫斯科物理技术学院)
;
AIRI, the Artificial Intelligence Research Institute(人工智能研究所)
机构
*
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Hong Kong University of Science and Technology(香港科技大学)
;
Department of Thoracic Surgery, the Seventh Affiliated Hospital, Sun Yat-sen University(中山大学第七附属医院胸外科部门)
;
Bioengineering/Imperial-X, Imperial College London(生物工程/Imperial-X,帝国理工学院伦敦分校)
;
ROAS Thrust, Hong Kong University of Science and Technology (Guangzhou)(ROAS项目,香港科技大学(广州))
;
Department of Electronic and Computer Engineering, Hong Kong SAR(香港特别行政区电子与计算机工程系)