arXivDaily arXiv每日学术速递 周一至周五更新

作者

Pieter Abbeel

Robotics

共收录 415
1701.00867 2017-01-05 cs.AI

A K-fold Method for Baseline Estimation in Policy Gradient Algorithms

Nithyanand Kota, Abhishek Mishra, Sunil Srinivasa, Xi, Chen, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.03852 2016-11-28 cs.LG cs.AI

A Connection between Generative Adversarial Networks, Inverse Reinforcement Learning, and Energy-Based Models

Chelsea Finn, Paul Christiano, Pieter Abbeel, Sergey Levine

Comments NIPS 2016 Workshop on Adversarial Training. First two authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.02779 2016-11-11 cs.AI cs.LG cs.NE stat.ML

RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

Yan Duan, John Schulman, Xi Chen, Peter L. Bartlett, Ilya Sutskever, Pieter Abbeel

Comments 14 pages. Under review as a conference paper at ICLR 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.01112 2016-10-07 cs.LG cs.RO

Reset-Free Guided Policy Search: Efficient Deep Reinforcement Learning with Stochastic Initial States

William Montgomery, Anurag Ajay, Chelsea Finn, Pieter Abbeel, Sergey Levine

详情

展开后加载摘要…

URL PDF HTML 收藏
1506.04304 2016-09-27 cs.CV

Combinatorial Energy Learning for Image Segmentation

Jeremy Maitin-Shepard, Viren Jain, Michal Januszewski, Peter Li, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.07088 2016-09-23 cs.LG cs.RO

Learning Modular Neural Network Policies for Multi-Task and Multi-Robot Transfer

Coline Devin, Abhishek Gupta, Trevor Darrell, Pieter Abbeel, Sergey Levine

Comments Under review at the International Conference on Robotics and Automation (ICRA) 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.05814 2016-09-20 cs.CY cs.RO

Toward a Science of Autonomy for Physical Systems: Paths

Pieter Abbeel, Ken Goldberg, Gregory Hager, Julie Shah

Comments A Computing Community Consortium (CCC) white paper, 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1509.06841 2016-08-12 cs.LG cs.RO

One-Shot Learning of Manipulation Skills with Online Dynamics Adaptation and Neural Network Priors

Justin Fu, Sergey Levine, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1606.03657 2016-06-14 cs.LG stat.ML

InfoGAN: Interpretable Representation Learning by Information Maximizing Generative Adversarial Nets

Xi Chen, Yan Duan, Rein Houthooft, John Schulman, Ilya Sutskever, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1604.06778 2016-05-30 cs.LG cs.AI cs.RO

Benchmarking Deep Reinforcement Learning for Continuous Control

Yan Duan, Xi Chen, Rein Houthooft, John Schulman, Pieter Abbeel

Comments 14 pages, ICML 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1603.00448 2016-05-30 cs.LG cs.AI cs.RO

Guided Cost Learning: Deep Inverse Optimal Control via Policy Optimization

Chelsea Finn, Sergey Levine, Pieter Abbeel

Comments International Conference on Machine Learning (ICML), 2016, to appear

详情

展开后加载摘要…

URL PDF HTML 收藏
1504.00702 2016-04-20 cs.LG cs.CV cs.RO

End-to-End Training of Deep Visuomotor Policies

Sergey Levine, Chelsea Finn, Trevor Darrell, Pieter Abbeel

Comments updating with revisions for JMLR final version

详情

展开后加载摘要…

URL PDF HTML 收藏
1509.06824 2016-03-16 cs.LG cs.RO

Model-based Reinforcement Learning with Parametrized Physical Models and Optimism-Driven Exploration

Christopher Xie, Sachin Patil, Teodor Moldovan, Sergey Levine, Pieter Abbeel

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1509.06113 2016-03-02 cs.LG cs.CV cs.RO

Deep Spatial Autoencoders for Visuomotor Learning

Chelsea Finn, Xin Yu Tan, Yan Duan, Trevor Darrell, Sergey Levine, Pieter Abbeel

Comments Published in the International Conference on Robotics and Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
1509.06791 2016-02-17 cs.LG cs.RO

Learning Deep Control Policies for Autonomous Aerial Vehicles with MPC-Guided Policy Search

Tianhao Zhang, Gregory Kahn, Sergey Levine, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1506.05254 2016-01-06 cs.LG

Gradient Estimation Using Stochastic Computation Graphs

John Schulman, Nicolas Heess, Theophane Weber, Pieter Abbeel

Comments Advances in Neural Information Processing Systems 28 (NIPS 2015)

详情

展开后加载摘要…

URL PDF HTML 收藏
1507.00814 2015-11-23 cs.AI cs.LG stat.ML

Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models

Bradly C. Stadie, Sergey Levine, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1507.01273 2015-09-24 cs.LG cs.RO

Learning Deep Neural Network Policies with Continuous Memory States

Marvin Zhang, Zoe McCarthy, Chelsea Finn, Sergey Levine, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1109.1966 2015-03-19 cs.AI

The path inference filter: model-based low-latency map matching of probe vehicle data

Timothy Hunter, Pieter Abbeel, Alexandre Bayen

Comments Preprint, 23 pages and 23 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1501.05611 2015-02-27 cs.RO

Learning Contact-Rich Manipulation Skills with Guided Policy Search

Sergey Levine, Nolan Wagener, Pieter Abbeel

Journal ref S. Levine, N. Wagener, P. Abbeel, "Learning Contact-Rich Manipulation Skills with Guided Policy Search," in International Conference on Robotics and Automation (ICRA), 2015

详情

展开后加载摘要…

URL PDF HTML 收藏
1302.6617 2013-02-28 cs.LG cs.AI

Arriving on time: estimating travel time distributions on large-scale road networks

Timothy Hunter, Aude Hofleitner, Jack Reilly, Walid Krichene, Jerome Thai, Anastasios Kouvelas, Pieter Abbeel, Alexandre Bayen

详情

展开后加载摘要…

URL PDF HTML 收藏
1301.0604 2013-01-07 cs.LG cs.AI stat.ML

Discriminative Probabilistic Models for Relational Data

Ben Taskar, Pieter Abbeel, Daphne Koller

Comments Appears in Proceedings of the Eighteenth Conference on Uncertainty in Artificial Intelligence (UAI2002)

详情

展开后加载摘要…

URL PDF HTML 收藏
1212.3393 2012-12-17 cs.RO cs.SE

Large Scale Estimation in Cyberphysical Systems using Streaming Data: a Case Study with Smartphone Traces

Timothy Hunter, Tathagata Das, Matei Zaharia, Pieter Abbeel, Alexandre M. Bayen

详情

展开后加载摘要…

URL PDF HTML 收藏
1205.4810 2012-07-10 cs.LG

Safe Exploration in Markov Decision Processes

Teodor Mihai Moldovan, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1207.1366 2012-07-09 cs.LG stat.ML

Learning Factor Graphs in Polynomial Time & Sample Complexity

Pieter Abbeel, Daphne Koller, Andrew Y. Ng

Comments Appears in Proceedings of the Twenty-First Conference on Uncertainty in Artificial Intelligence (UAI2005)

详情

展开后加载摘要…

URL PDF HTML 收藏