LPPG-RL: Lexicographically Projected Policy Gradient Reinforcement Learning with Subproblem Exploration
专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG
视觉与机器人
机器人、具身智能、机器人学习、操作、导航和具身世界模型。
专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG
机构 * University of Southern California(南加州大学)
专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG
机构 * Dalle Molle Institute for Artificial Intelligence (IDSIA) - USI/SUPSI(达摩信息技术研究所(IDSIA)- USI/SUPSI) ; Center of Excellence for Generative AI, King Abdullah University of Science and Technology(生成人工智能卓越中心,国王阿卜杜勒阿齐兹大学科学与技术学院) ; NNAISENSE
专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG
Comments 85 pages in main text + 4 pages of references + 26 pages of appendices, 12 figures in main text + 2 figures in appendices; source code available at https://github.com/struplm/eUDRL-GCSL-ODT-Convergence-public
机构 * Fudan University(复旦大学) ; Ant Group(蚂蚁集团)
专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG
Comments Preprint