arXivDaily arXiv每日学术速递 周一至周五更新

作者

Pieter Abbeel

Robotics

共收录 415
2104.02180 2022-05-13 cs.GR cs.LG

AMP: Adversarial Motion Priors for Stylized Physics-Based Character Control

Xue Bin Peng, Ze Ma, Pieter Abbeel, Sergey Levine, Angjoo Kanazawa

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.01455 2022-05-05 cs.CV cs.AI cs.GR cs.LG

Zero-Shot Text-Guided Object Generation with Dream Fields

Ajay Jain, Ben Mildenhall, Jonathan T. Barron, Pieter Abbeel, Ben Poole

Comments CVPR 2022. 13 pages. Website: https://ajayj.com/dreamfields

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.12471 2022-05-03 cs.RO cs.AI cs.CV cs.LG

Coarse-to-fine Q-attention with Tree Expansion

Stephen James, Pieter Abbeel

Comments Project page and code: https://sites.google.com/view/q-attention-qte

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.03597 2022-04-19 cs.LG cs.AI

Imitating, Fast and Slow: Robust learning from demonstrations via decision-time planning

Carl Qi, Pieter Abbeel, Aditya Grover

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.13425 2022-04-07 cs.LG cs.AI

Don't Change the Algorithm, Change the Data: Exploratory Data for Offline Reinforcement Learning

Denis Yarats, David Brandfonbrener, Hao Liu, Michael Laskin, Pieter Abbeel, Alessandro Lazaric, Lerrel Pinto

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.01571 2022-04-05 cs.RO cs.AI cs.CV cs.LG

Coarse-to-Fine Q-attention with Learned Path Ranking

Stephen James, Pieter Abbeel

Comments Project page and code: https://sites.google.com/view/q-attention-lpr

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.15913 2022-04-04 cs.LG cs.AI

Pretraining Graph Neural Networks for few-shot Analog Circuit Modeling and Design

Kourosh Hakhamaneshi, Marcel Nassar, Mariano Phielipp, Pieter Abbeel, Vladimir Stojanović

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.00161 2022-03-31 cs.LG cs.AI

CIC: Contrastive Intrinsic Control for Unsupervised Skill Discovery

Michael Laskin, Hao Liu, Xue Bin Peng, Denis Yarats, Aravind Rajeswaran, Pieter Abbeel

Comments Project website: https://sites.google.com/view/cicrl/

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.15103 2022-03-30 cs.AI cs.RO

Adversarial Motion Priors Make Good Substitutes for Complex Reward Functions

Alejandro Escontrela, Xue Bin Peng, Wenhao Yu, Tingnan Zhang, Atil Iscen, Ken Goldberg, Pieter Abbeel

Comments 8 pages, 6 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.12462 2022-03-22 cs.LG cs.AI cs.HC cs.RO

Explaining Reinforcement Learning Policies through Counterfactual Trajectories

Julius Frost, Olivia Watkins, Eric Weiner, Pieter Abbeel, Trevor Darrell, Bryan Plummer, Kate Saenko

Comments Accepted at ICML HILL 2021 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.10050 2022-03-21 cs.LG cs.AI

SURF: Semi-supervised Reward Learning with Data Augmentation for Feedback-efficient Preference-based Reinforcement Learning

Jongjin Park, Younggyo Seo, Jinwoo Shin, Honglak Lee, Pieter Abbeel, Kimin Lee

Comments Accepted to ICLR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.08981 2022-03-11 cs.LG cs.AI cs.RO

Hierarchical Few-Shot Imitation with Skill Transition Models

Kourosh Hakhamaneshi, Ruihan Zhao, Albert Zhan, Pieter Abbeel, Michael Laskin

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.00942 2022-03-11 cs.LG cs.AI

JUMBO: Scalable Multi-task Bayesian Optimization using Offline Data

Kourosh Hakhamaneshi, Pieter Abbeel, Vladimir Stojanovic, Aditya Grover

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.07207 2022-03-09 cs.LG cs.AI cs.CL cs.CV cs.RO

Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents

Wenlong Huang, Pieter Abbeel, Deepak Pathak, Igor Mordatch

Comments Project website at https://huangwl18.github.io/language-planner

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.10608 2022-02-23 cs.LG cs.AI cs.MA

It Takes Four to Tango: Multiagent Selfplay for Automatic Curriculum Generation

Yuqing Du, Pieter Abbeel, Aditya Grover

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.06784 2022-02-14 cs.LG stat.ML

ProMP: Proximal Meta-Policy Search

Jonas Rothfuss, Dennis Lee, Ignasi Clavera, Tamim Asfour, Pieter Abbeel

Comments The first three authors contributed equally. Published at ICLR 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.03957 2022-02-09 cs.RO cs.AI cs.CV cs.LG

Bingham Policy Parameterization for 3D Rotations in Reinforcement Learning

Stephen James, Pieter Abbeel

Comments Project page and code: https://sites.google.com/view/rl-bpp

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.13423 2022-02-09 cs.RO cs.AI cs.LG

Towards More Generalizable One-shot Visual Imitation Learning

Zhao Mandi, Fangchen Liu, Kimin Lee, Pieter Abbeel

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.01686 2022-01-28 cs.LG stat.ML

Likelihood Contribution based Multi-scale Architecture for Generative Flows

Hari Prasanna Das, Pieter Abbeel, Costas J. Spanos

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.04973 2022-01-10 cs.LG cs.AI cs.PL cs.SE stat.ML

Contrastive Code Representation Learning

Paras Jain, Ajay Jain, Tianjun Zhang, Pieter Abbeel, Joseph E. Gonzalez, Ion Stoica

Comments In Proceedings of EMNLP 2021. 19 pages, 16 figures, 9 tables. Code available at https://github.com/parasj/contracode

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.00210 2021-12-14 cs.LG cs.AI cs.CV cs.RO

Mastering Atari Games with Limited Data

Weirui Ye, Shaohuai Liu, Thanard Kurutach, Pieter Abbeel, Yang Gao

Comments Published at NeurIPS 2021; Homepage: https://yewr.github.io/projects/efficientzero/

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.02852 2021-12-07 cs.LG cs.AI

Target Entropy Annealing for Discrete Soft Actor-Critic

Yaosheng Xu, Dailin Hu, Litian Liang, Stephen McAleer, Pieter Abbeel, Roy Fox

Journal ref neurips 2021 deep rl workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.00901 2021-12-03 cs.AI cs.LG

Hindsight Task Relabelling: Experience Replay for Sparse Reward Meta-RL

Charles Packer, Pieter Abbeel, Joseph E. Gonzalez

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.14204 2021-11-30 cs.LG cs.AI

Count-Based Temperature Scheduling for Maximum Entropy Reinforcement Learning

Dailin Hu, Pieter Abbeel, Roy Fox

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.03062 2021-11-05 cs.RO cs.AI cs.CV cs.LG cs.SY eess.SY

Generalization in Dexterous Manipulation via Geometry-Aware Multi-Task Learning

Wenlong Huang, Igor Mordatch, Pieter Abbeel, Deepak Pathak

Comments Website at https://huangwl18.github.io/geometry-dex

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.03026 2021-11-05 cs.LG cs.AI cs.HC

B-Pref: Benchmarking Preference-Based Reinforcement Learning

Kimin Lee, Laura Smith, Anca Dragan, Pieter Abbeel

Comments NeurIPS Datasets and Benchmarks Track 2021. Code is available at https://github.com/rll-research/B-Pref

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.00591 2021-11-02 cs.RO cs.LG

Offline-to-Online Reinforcement Learning via Balanced Replay and Pessimistic Q-Ensemble

Seunghyun Lee, Younggyo Seo, Kimin Lee, Pieter Abbeel, Jinwoo Shin

Comments CoRL 2021. First two authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.15191 2021-10-29 cs.LG cs.AI cs.RO

URLB: Unsupervised Reinforcement Learning Benchmark

Michael Laskin, Denis Yarats, Hao Liu, Kimin Lee, Albert Zhan, Kevin Lu, Catherine Cang, Lerrel Pinto, Pieter Abbeel

Comments Code for the Unsupervised Reinforcement Learning Benchmark is available at https://github.com/rll-research/url_benchmark

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.14818 2021-10-29 cs.LG cs.AI

Temporal-Difference Value Estimation via Uncertainty-Guided Soft Updates

Litian Liang, Yaosheng Xu, Stephen McAleer, Dailin Hu, Alexander Ihler, Pieter Abbeel, Roy Fox

Comments Accepted to Deep Reinforcement Learning Workshop @ NeurIPS 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.04551 2021-10-29 cs.LG

Behavior From the Void: Unsupervised Active Pre-Training

Hao Liu, Pieter Abbeel

Comments Advances in Neural Information Processing Systems(NeurIPS), 2021 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏