arXivDaily arXiv每日学术速递 周一至周五更新

作者

Pieter Abbeel

Robotics

共收录 415
1902.04198 2019-04-22 cs.LG cs.AI stat.ML

Preferences Implicit in the State of the World

Rohin Shah, Dmitrii Krasheninnikov, Jordan Alexander, Pieter Abbeel, Anca Dragan

Comments Published at ICLR 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.03815 2019-04-15 cs.RO

Quasi-Direct Drive for Low-Cost Compliant Robotic Manipulation

David V. Gealy, Stephen McKinley, Brent Yi, Philipp Wu, Phillip R. Downey, Greg Balke, Allan Zhao, Menglong Guo, Rachel Thomasson, Anthony Sinclair, Peter Cuellar, Zoe McCarthy, Pieter Abbeel

Comments This is our long version - 8 pages. Our 6 page version without a discussion of thermal limits was accepted to ICRA 2019. 11 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.08894 2019-03-22 cs.LG cs.AI

Towards Characterizing Divergence in Deep Q-Learning

Joshua Achiam, Ethan Knight, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.01066 2019-03-21 cs.RO

Reinforcement Learning on Variable Impedance Controller for High-Precision Robotic Assembly

Jianlan Luo, Eugen Solowjow, Chengtao Wen, Juan Aparicio Ojea, Alice M. Agogino, Aviv Tamar, Pieter Abbeel

Comments ICRA 2019. More video results at https://sites.google.com/berkeley.edu/rl-robotic-assembly/home

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.03953 2019-03-12 cs.CV

Domain Randomization for Active Pose Estimation

Xinyi Ren, Jianlan Luo, Eugen Solowjow, Juan Aparicio Ojea, Abhishek Gupta, Aviv Tamar, Pieter Abbeel

Comments Accepted at International Conference on Robotics and Automation (ICRA) 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.11347 2019-03-01 cs.LG cs.RO stat.ML

Learning to Adapt in Dynamic, Real-World Environments Through Meta-Reinforcement Learning

Anusha Nagabandi, Ignasi Clavera, Simin Liu, Ronald S. Fearing, Pieter Abbeel, Sergey Levine, Chelsea Finn

Comments First 2 authors contributed equally. Website: https://sites.google.com/berkeley.edu/metaadaptivecontrol

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.03701 2019-02-12 cs.LG cs.RO stat.ML

Generalization through Simulation: Integrating Simulated and Real Data into Deep Reinforcement Learning for Vision-Based Autonomous Flight

Katie Kang, Suneel Belkhale, Gregory Kahn, Pieter Abbeel, Sergey Levine

Comments First three authors contributed equally. Accepted to ICRA 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.07882 2019-01-30 cs.LG cs.AI cs.CL cs.HC

Guiding Policies with Language via Meta-Learning

John D. Co-Reyes, Abhishek Gupta, Suvansh Sanjeev, Nick Altieri, Jacob Andreas, John DeNero, Pieter Abbeel, Sergey Levine

Comments Accepted at ICLR 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.01118 2019-01-15 cs.AI

Some Considerations on Learning to Explore via Meta-Reinforcement Learning

Bradly C. Stadie, Ge Yang, Rein Houthooft, Xi Chen, Yan Duan, Yuhuai Wu, Pieter Abbeel, Ilya Sutskever

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.02811 2019-01-14 cs.LG cs.AI cs.DC

Accelerated Methods for Deep Reinforcement Learning

Adam Stooke, Pieter Abbeel

Comments v2: -Added game performance statistics summary for algorithm scaling across full Atari game set. -Added full set of learning curves (appendix). -Fixed images to remove phantom borders. -Streamlined some discussion, moved some details to appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.06711 2018-11-19 cs.RO cs.LG

An Algorithmic Perspective on Imitation Learning

Takayuki Osa, Joni Pajarinen, Gerhard Neumann, J. Andrew Bagnell, Pieter Abbeel, Jan Peters

Comments 187 pages. Published in Foundations and Trends in Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.03555 2018-11-09 cs.AI

Modular Architecture for StarCraft II with Deep Reinforcement Learning

Dennis Lee, Haoran Tang, Jeffrey O Zhang, Huazhe Xu, Trevor Darrell, Pieter Abbeel

Comments Accepted to The 14th AAAI Conference on Artificial Intelligence and Interactive Digital Entertainment (AIIDE'18)

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.11043 2018-10-29 cs.LG cs.AI cs.CV cs.RO stat.ML

One-Shot Hierarchical Imitation Learning of Compound Visuomotor Tasks

Tianhe Yu, Pieter Abbeel, Sergey Levine, Chelsea Finn

Comments Video results available at https://sites.google.com/view/one-shot-hil

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.08174 2018-10-19 cs.RO

Establishing Appropriate Trust via Critical States

Sandy H. Huang, Kush Bhatia, Pieter Abbeel, Anca D. Dragan

Comments IROS 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1702.03465 2018-10-19 cs.RO cs.LG

Enabling Robots to Communicate their Objectives

Sandy H. Huang, David Held, Pieter Abbeel, Anca D. Dragan

Comments RSS 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.07167 2018-10-17 cs.RO cs.AI cs.LG

Composable Action-Conditioned Predictors: Flexible Off-Policy Learning for Robot Navigation

Gregory Kahn, Adam Villaflor, Pieter Abbeel, Sergey Levine

Comments Accepted to the Conference on Robot Learning (CoRL) 2018. Video at https://youtu.be/lOLT7zifEkg

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.03599 2018-10-16 cs.GR cs.CV cs.LG

SFV: Reinforcement Learning of Physical Skills from Videos

Xue Bin Peng, Angjoo Kanazawa, Jitendra Malik, Pieter Abbeel, Sergey Levine

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.06440 2018-10-16 cs.LG

Equivalence Between Policy Gradients and Soft Q-Learning

John Schulman, Xi Chen, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1802.10592 2018-10-08 cs.LG cs.AI cs.RO

Model-Ensemble Trust-Region Policy Optimization

Thanard Kurutach, Ignasi Clavera, Yan Duan, Aviv Tamar, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.04326 2018-09-21 cs.AI cs.GT

Learning with Opponent-Learning Awareness

Jakob N. Foerster, Richard Y. Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, Igor Mordatch

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.05214 2018-09-17 cs.LG cs.AI stat.ML

Model-Based Reinforcement Learning via Meta-Policy Optimization

Ignasi Clavera, Jonas Rothfuss, John Schulman, Yasuhiro Fujita, Tamim Asfour, Pieter Abbeel

Comments First 2 authors contributed equally. Accepted for Conference on Robot Learning (CoRL)

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.02808 2018-09-05 cs.LG cs.AI stat.ML

Latent Space Policies for Hierarchical Reinforcement Learning

Tuomas Haarnoja, Kristian Hartikainen, Pieter Abbeel, Sergey Levine

Comments ICML 2018; Videos: https://sites.google.com/view/latent-space-deep-rl Code: https://github.com/haarnoja/sac

详情

展开后加载摘要…

URL PDF HTML 收藏
1808.07804 2018-08-24 stat.ML cs.AI cs.LG stat.AP

Transfer Learning for Estimating Causal Effects using Neural Networks

Sören R. Künzel, Bradly C. Stadie, Nikita Vemuri, Varsha Ramakrishnan, Jasjeet S. Sekhon, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1801.01290 2018-08-10 cs.LG cs.AI stat.ML

Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor

Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, Sergey Levine

Comments ICML 2018 Videos: sites.google.com/view/soft-actor-critic Code: github.com/haarnoja/sac

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.02717 2018-08-07 cs.GR cs.AI cs.LG

DeepMimic: Example-Guided Deep Reinforcement Learning of Physics-Based Character Skills

Xue Bin Peng, Pieter Abbeel, Sergey Levine, Michiel van de Panne

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.10299 2018-07-30 cs.AI

Variational Option Discovery Algorithms

Joshua Achiam, Harrison Edwards, Dario Amodei, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.09341 2018-07-27 cs.LG cs.AI cs.CV cs.NE cs.RO stat.ML

Learning Plannable Representations with Causal InfoGAN

Thanard Kurutach, Aviv Tamar, Ge Yang, Stuart Russell, Pieter Abbeel

Comments ICML / IJCAI / AAMAS 2018 Workshop on Planning and Learning (PAL-18)

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.07635 2018-07-26 cs.RO cs.LG

Learning Robotic Assembly from CAD

Garrett Thomas, Melissa Chien, Aviv Tamar, Juan Aparicio Ojea, Pieter Abbeel

Comments In the proceedings of the IEEE International Conference on Robotics and Automation (ICRA), Brisbane, Australia, May 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.07280 2018-07-26 cs.AI

Learning Generalized Reactive Policies using Deep Neural Networks

Edward Groshev, Maxwell Goldstein, Aviv Tamar, Siddharth Srivastava, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.08534 2018-07-25 cs.LG cs.AI stat.ML

Safer Classification by Synthesis

William Wang, Angelina Wang, Aviv Tamar, Xi Chen, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏