arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 4134 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 模仿学习与强化学习 4134 篇

2012.04210 2020-12-09 cs.LG cs.AR 57%

The Architectural Implications of Distributed Reinforcement Learning on CPU-GPU Systems

Ahmet Inci, Evgeny Bolotin, Yaosheng Fu, Gal Dalal, Shie Mannor, David Nellans, Diana Marculescu

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

Comments To appear in the proceedings of the 6th Workshop on Energy Efficient Machine Learning and Cognitive Computing (EMC2) 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.03240 2020-12-08 cs.LG stat.ML 57%

Multi-task Reinforcement Learning with a Planning Quasi-Metric

Vincent Micheli, Karthigan Sinnathamby, François Fleuret

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.LG

Comments Deep RL Workshop, NeurIPS 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.12574 2020-11-26 cs.LG stat.ML 57%

Enhanced Scene Specificity with Sparse Dynamic Value Estimation

Jaskirat Singh, Liang Zheng

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.11293 2020-11-24 cs.LG cs.NE 57%

Evolutionary Planning in Latent Space

Thor V. A. N. Olesen, Dennis T. T. Nguyen, Rasmus Berg Palm, Sebastian Risi

专题命中 模仿学习与强化学习 :world model(abstract);分类 cs.LG

Comments Code to reproduce the experiments are available at https://github.com/two2tee/WorldModelPlanning Video of driving performance is available at https://youtu.be/3M39QgeF27U

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.09445 2020-11-19 cs.RO 57%

Cautious Bayesian Optimization for Efficient and Scalable Policy Search

Lukas P. Fröhlich, Melanie N. Zeilinger, Edgar D. Klenske

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.08272 2020-11-18 cs.CL cs.AI 57%

NLPGym -- A toolkit for evaluating RL agents on Natural Language Processing Tasks

Rajkumar Ramamurthy, Rafet Sifa, Christian Bauckhage

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI

Comments Accepted at Wordplay: When Language Meets Games Workshop @ NeurIPS 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.06116 2020-11-13 cs.RO cs.SY eess.SY 57%

A Data-Driven Reinforcement Learning Solution Framework for Optimal and Adaptive Personalization of a Hip Exoskeleton

Xikai Tu, Minhan Li, Ming Liu, Jennie Si, He, Huang

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO

Comments 7 pages, 9 figures, ICRA 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.00397 2020-11-03 cs.RO 57%

APPLR: Adaptive Planner Parameter Learning from Reinforcement

Zifan Xu, Gauraang Dhamankar, Anirudh Nair, Xuesu Xiao, Garrett Warnell, Bo Liu, Zizhao Wang, Peter Stone

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.13056 2020-10-27 cs.RO 57%

Proactive Action Visual Residual Reinforcement Learning for Contact-Rich Tasks Using a Torque-Controlled Robot

Yunlei Shi, Zhaopeng Chen, Hongxu Liu, Sebastian Riedel, Chunhui Gao, Qian Feng, Jun Deng, Jianwei Zhang

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO

Comments 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.12142 2020-10-26 cs.LG 57%

Bridging Imagination and Reality for Model-Based Deep Reinforcement Learning

Guangxiang Zhu, Minghao Zhang, Honglak Lee, Chongjie Zhang

专题命中 模仿学习与强化学习 :world model(abstract);分类 cs.LG

Comments Published on 34th Conference on Neural Information Processing Systems (NeurIPS 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.10691 2020-10-22 cs.LG stat.ML 57%

Safe Reinforcement Learning with Nonlinear Dynamics via Model Predictive Shielding

Osbert Bastani

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.08443 2020-10-19 cs.LG cs.SY eess.SY 57%

Policy Gradient for Continuing Tasks in Non-stationary Markov Decision Processes

Santiago Paternain, Juan Andres Bazerque, Alejandro Ribeiro

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.09762 2020-10-14 cs.RO 57%

TTR-Based Reward for Reinforcement Learning with Implicit Model Priors

Xubo Lyu, Mo Chen

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO

Comments This is serving as the full version of the paper that is accepted by IEEE / RSJ International Conference on Intelligent Robots and Systems, 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.02506 2020-10-07 cs.LG stat.ML 57%

Interactive Reinforcement Learning for Feature Selection with Decision Tree in the Loop

Wei Fan, Kunpeng Liu, Hao Liu, Yong Ge, Hui Xiong, Yanjie Fu

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

Comments arXiv admin note: substantial text overlap with arXiv:2008.12001

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.09361 2020-09-22 eess.SY cs.LG cs.SY 57%

Lyapunov-Based Reinforcement Learning for Decentralized Multi-Agent Control

Qingrui Zhang, Hao Dong, Wei Pan

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

Comments Accepted to The 2nd International Conference on Distributed Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.08749 2020-09-17 stat.ML cs.LG math.OC math.PR math.ST stat.TH 57%

Instance-dependent $\ell_\infty$-bounds for policy evaluation in tabular reinforcement learning

Ashwin Pananjady, Martin J. Wainwright

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

Comments Version v2 is consistent with manuscript to appear in IEEE Transactions on Information Theory

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.04777 2020-09-11 cs.LG stat.ML 57%

A framework for reinforcement learning with autocorrelated actions

Marcin Szulc, Jakub Łyskawa, Paweł Wawrzyński

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

Comments The 27th International Conference on Neural Information Processing (ICONIP2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.03349 2020-09-10 cs.LG cs.SY eess.SY stat.ML 57%

Deep Learning and Reinforcement Learning for Autonomous Unmanned Aerial Systems: Roadmap for Theory to Deployment

Jithin Jagannath, Anu Jagannath, Sean Furman, Tyler Gwin

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

Comments Preprint of Book Chapter to be published in Springer

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.05441 2020-09-01 cs.LG cs.MA stat.ML 57%

Delay-Aware Multi-Agent Reinforcement Learning for Cooperative and Competitive Environments

Baiming Chen, Mengdi Xu, Zuxin Liu, Liang Li, Ding Zhao

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.11811 2020-08-28 cs.LG math.OC stat.ML 57%

Constrained Markov Decision Processes via Backward Value Functions

Harsh Satija, Philip Amortila, Joelle Pineau

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.11543 2020-08-28 cs.RO cs.MA 57%

Continuous Deep Hierarchical Reinforcement Learning for Ground-Air Swarm Shepherding

Hung The Nguyen, Tung Duy Nguyen, Vu Phi Tran, Matthew Garratt, Kathryn Kasmarik, Sreenatha Anavatti, Michael Barlow, Hussein A. Abbass

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.09062 2020-08-18 stat.ML cs.CR cs.LG 57%

Adversarial Reinforcement Learning under Partial Observability in Autonomous Computer Network Defence

Yi Han, David Hubczenko, Paul Montague, Olivier De Vel, Tamas Abraham, Benjamin I. P. Rubinstein, Christopher Leckie, Tansu Alpcan, Sarah Erfani

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.LG

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.02521 2020-08-07 cs.RO 57%

Deep Reinforcement Learning based Local Planner for UAV Obstacle Avoidance using Demonstration Data

Lei He, Nabil Aouf, James F. Whidborne, Bifeng Song

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO

Comments Please find video demos at https://www.youtube.com/watch?v=4Zj49QtDdks

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.02435 2020-08-04 cs.LG stat.ML 57%

A Nonparametric Off-Policy Policy Gradient

Samuele Tosatto, Joao Carvalho, Hany Abdulsamad, Jan Peters

专题命中 模仿学习与强化学习 :robot learning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.03827 2020-07-22 cs.LG cs.CR cs.SY eess.SY math.OC stat.ML 57%

Manipulating Reinforcement Learning: Poisoning Attacks on Cost Signals

Yunhan Huang, Quanyan Zhu

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.LG

Comments This chapter is written for the forthcoming book "Game Theory and Machine Learning for Cyber Security" (Wiley-IEEE Press), edited by Charles Kamhoua et. al. arXiv admin note: text overlap with arXiv:1906.10571

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.09540 2020-07-21 cs.AI 57%

Multi-Principal Assistance Games

Arnaud Fickinger, Simon Zhuang, Dylan Hadfield-Menell, Stuart Russell

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.06123 2020-07-14 cs.SD cs.LG eess.AS stat.ML 57%

OtoWorld: Towards Learning to Separate by Learning to Move

Omkar Ranadive, Grant Gasser, David Terpay, Prem Seetharaman

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

Comments Published in Self Supervision in Audio and Speech Workshop, 37th International Conference on Machine Learning, Vienna, Austria (ICML 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.00463 2020-07-02 cs.AI 57%

A Generalized Reinforcement Learning Algorithm for Online 3D Bin-Packing

Richa Verma, Aniruddha Singhal, Harshad Khadilkar, Ansuma Basumatary, Siddharth Nayak, Harsh Vardhan Singh, Swagat Kumar, Rajesh Sinha

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI

Comments 9 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.00425 2020-07-02 cs.LG stat.ML 57%

Interaction-limited Inverse Reinforcement Learning

Martin Troussard, Emmanuel Pignat, Parameswaran Kamalaruban, Sylvain Calinon, Volkan Cevher

专题命中 模仿学习与强化学习 :robot learning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.06800 2020-06-30 cs.LG stat.ML 57%

Context-aware Dynamics Model for Generalization in Model-Based Reinforcement Learning

Kimin Lee, Younggyo Seo, Seunghyun Lee, Honglak Lee, Jinwoo Shin

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

Comments Accepted in ICML2020. First two authors contributed equally, website: https://sites.google.com/view/cadm code: https://github.com/younggyoseo/CaDM

详情

展开后加载摘要…

URL PDF HTML 收藏