arXivDaily arXiv每日学术速递 周一至周五更新

作者

Sergey Levine

Robotics / Reinforcement Learning

共收录 578
1611.04201 2017-06-09 cs.LG cs.CV cs.RO

CAD2RL: Real Single-Image Flight without a Single Real Image

Fereshteh Sadeghi, Sergey Levine

Comments To appear at Robotics: Science and Systems Conference (R:SS), 2017. Supplementary video: https://www.youtube.com/watch?v=nXBWmzFrj5s

详情

展开后加载摘要…

URL PDF HTML 收藏
1706.00387 2017-06-02 cs.LG cs.AI cs.RO

Interpolated Policy Gradient: Merging On-Policy and Off-Policy Gradient Estimation for Deep Reinforcement Learning

Shixiang Gu, Timothy Lillicrap, Zoubin Ghahramani, Richard E. Turner, Bernhard Schölkopf, Sergey Levine

详情

展开后加载摘要…

URL PDF HTML 收藏
1703.01260 2017-05-30 cs.LG

EX2: Exploration with Exemplar Models for Deep Reinforcement Learning

Justin Fu, John D. Co-Reyes, Sergey Levine

详情

展开后加载摘要…

URL PDF HTML 收藏
1511.07111 2017-05-29 cs.CV

Adapting Deep Visuomotor Representations with Weak Pairwise Constraints

Eric Tzeng, Coline Devin, Judy Hoffman, Chelsea Finn, Pieter Abbeel, Sergey Levine, Kate Saenko, Trevor Darrell

详情

展开后加载摘要…

URL PDF HTML 收藏
1502.05477 2017-04-24 cs.LG

Trust Region Policy Optimization

John Schulman, Sergey Levine, Philipp Moritz, Michael I. Jordan, Pieter Abbeel

Comments 16 pages, ICML 2015

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.09001 2017-03-22 cs.RO cs.AI cs.LG

Learning from the Hindsight Plan -- Episodic MPC Improvement

Aviv Tamar, Garrett Thomas, Tianhao Zhang, Sergey Levine, Pieter Abbeel

Comments Additional experiments for neural network generalization and for varying the planning horizon. Paper accepted to ICRA 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1602.02867 2017-03-22 cs.AI cs.LG cs.NE stat.ML

Value Iteration Networks

Aviv Tamar, Yi Wu, Garrett Thomas, Sergey Levine, Pieter Abbeel

Comments Fixed missing table values

Journal ref Advances in Neural Information Processing Systems 29 pages 2154--2162, 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1603.06348 2017-03-21 cs.LG cs.RO

Learning Dexterous Manipulation for a Soft Robotic Hand from Human Demonstration

Abhishek Gupta, Clemens Eppner, Sergey Levine, Pieter Abbeel

Comments Accepted at International Conference on Intelligent Robots and Systems(IROS) 2016. Pdf file updated for stylistic consistency

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.00696 2017-03-14 cs.LG cs.AI cs.CV cs.RO

Deep Visual Foresight for Planning Robot Motion

Chelsea Finn, Sergey Levine

Comments ICRA 2017. Supplementary video: https://sites.google.com/site/robotforesight/

详情

展开后加载摘要…

URL PDF HTML 收藏
1612.00429 2017-03-13 cs.LG cs.AI cs.RO

Generalizing Skills with Semi-Supervised Reinforcement Learning

Chelsea Finn, Tianhe Yu, Justin Fu, Pieter Abbeel, Sergey Levine

Comments ICLR 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1703.02949 2017-03-09 cs.AI cs.RO

Learning Invariant Feature Spaces to Transfer Skills with Reinforcement Learning

Abhishek Gupta, Coline Devin, YuXuan Liu, Pieter Abbeel, Sergey Levine

Comments Published as a conference paper at ICLR 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.09049 2017-03-09 cs.RO cs.LG

Deep Reinforcement Learning for Tensegrity Robot Locomotion

Marvin Zhang, Xinyang Geng, Jonathan Bruce, Ken Caluwaerts, Massimo Vespignani, Vytas SunSpiral, Pieter Abbeel, Sergey Levine

Comments International Conference on Robotics and Automation (ICRA), 2017. Project website link is http://rll.berkeley.edu/drl_tensegrity

详情

展开后加载摘要…

URL PDF HTML 收藏
1703.02018 2017-03-07 cs.CV cs.LG cs.RO

Combining Self-Supervised Learning and Imitation for Vision-Based Rope Manipulation

Ashvin Nair, Dian Chen, Pulkit Agrawal, Phillip Isola, Pieter Abbeel, Jitendra Malik, Sergey Levine

Comments 8 pages, accepted to International Conference on Robotics and Automation (ICRA) 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.01283 2017-03-07 cs.LG cs.AI cs.RO

EPOpt: Learning Robust Neural Network Policies Using Model Ensembles

Aravind Rajeswaran, Sarvjeet Ghotra, Balaraman Ravindran, Sergey Levine

Comments Accepted for publication at the International Conference on Learning Representations (ICLR) 2017. Supplementary video: https://youtu.be/w1YJ9vwaoto

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.02247 2017-03-01 cs.LG

Q-Prop: Sample-Efficient Policy Gradient with An Off-Policy Critic

Shixiang Gu, Timothy Lillicrap, Zoubin Ghahramani, Richard E. Turner, Sergey Levine

Comments Conference Paper at the International Conference on Learning Representations (ICLR) 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1603.00622 2017-02-28 cs.LG

PLATO: Policy Learning using Adaptive Trajectory Optimization

Gregory Kahn, Tianhao Zhang, Sergey Levine, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1606.07419 2017-02-17 cs.CV cs.AI cs.RO

Learning to Poke by Poking: Experiential Learning of Intuitive Physics

Pulkit Agrawal, Ashvin Nair, Pieter Abbeel, Jitendra Malik, Sergey Levine

Journal ref NIPS 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1702.01182 2017-02-07 cs.LG cs.RO

Uncertainty-Aware Reinforcement Learning for Collision Avoidance

Gregory Kahn, Adam Villaflor, Vitchyr Pong, Pieter Abbeel, Sergey Levine

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.03852 2016-11-28 cs.LG cs.AI

A Connection between Generative Adversarial Networks, Inverse Reinforcement Learning, and Energy-Based Models

Chelsea Finn, Paul Christiano, Pieter Abbeel, Sergey Levine

Comments NIPS 2016 Workshop on Adversarial Training. First two authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.00633 2016-11-24 cs.RO cs.AI cs.LG

Deep Reinforcement Learning for Robotic Manipulation with Asynchronous Off-Policy Updates

Shixiang Gu, Ethan Holly, Timothy Lillicrap, Sergey Levine

详情

展开后加载摘要…

URL PDF HTML 收藏
1605.07157 2016-10-19 cs.LG cs.AI cs.CV cs.RO

Unsupervised Learning for Physical Interaction through Video Prediction

Chelsea Finn, Ian Goodfellow, Sergey Levine

Comments To appear in NIPS '16; Video results, code, and data available at: http://www.sites.google.com/site/robotprediction

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.01112 2016-10-07 cs.LG cs.RO

Reset-Free Guided Policy Search: Efficient Deep Reinforcement Learning with Stochastic Initial States

William Montgomery, Anurag Ajay, Chelsea Finn, Pieter Abbeel, Sergey Levine

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.07088 2016-09-23 cs.LG cs.RO

Learning Modular Neural Network Policies for Multi-Task and Multi-Robot Transfer

Coline Devin, Abhishek Gupta, Trevor Darrell, Pieter Abbeel, Sergey Levine

Comments Under review at the International Conference on Robotics and Automation (ICRA) 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1603.02199 2016-08-30 cs.LG cs.AI cs.CV cs.RO

Learning Hand-Eye Coordination for Robotic Grasping with Deep Learning and Large-Scale Data Collection

Sergey Levine, Peter Pastor, Alex Krizhevsky, Deirdre Quillen

Comments This is an extended version of "Learning Hand-Eye Coordination for Robotic Grasping with Large-Scale Data Collection," ISER 2016. Draft modified to correct typo in Algorithm 1 and add a link to the publicly available dataset

详情

展开后加载摘要…

URL PDF HTML 收藏
1509.06841 2016-08-12 cs.LG cs.RO

One-Shot Learning of Manipulation Skills with Online Dynamics Adaptation and Neural Network Priors

Justin Fu, Sergey Levine, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1607.04614 2016-07-18 cs.LG cs.RO

Guided Policy Search as Approximate Mirror Descent

William Montgomery, Sergey Levine

详情

展开后加载摘要…

URL PDF HTML 收藏
1603.00448 2016-05-30 cs.LG cs.AI cs.RO

Guided Cost Learning: Deep Inverse Optimal Control via Policy Optimization

Chelsea Finn, Sergey Levine, Pieter Abbeel

Comments International Conference on Machine Learning (ICML), 2016, to appear

详情

展开后加载摘要…

URL PDF HTML 收藏
1504.00702 2016-04-20 cs.LG cs.CV cs.RO

End-to-End Training of Deep Visuomotor Policies

Sergey Levine, Chelsea Finn, Trevor Darrell, Pieter Abbeel

Comments updating with revisions for JMLR final version

详情

展开后加载摘要…

URL PDF HTML 收藏
1509.06824 2016-03-16 cs.LG cs.RO

Model-based Reinforcement Learning with Parametrized Physical Models and Optimism-Driven Exploration

Christopher Xie, Sachin Patil, Teodor Moldovan, Sergey Levine, Pieter Abbeel

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1509.06113 2016-03-02 cs.LG cs.CV cs.RO

Deep Spatial Autoencoders for Visuomotor Learning

Chelsea Finn, Xin Yu Tan, Yan Duan, Trevor Darrell, Sergey Levine, Pieter Abbeel

Comments Published in the International Conference on Robotics and Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏