作者
Pieter Abbeel
Robotics
Zero-Shot Text-Guided Object Generation with Dream Fields
Comments CVPR 2022. 13 pages. Website: https://ajayj.com/dreamfields
Coarse-to-fine Q-attention with Tree Expansion
Comments Project page and code: https://sites.google.com/view/q-attention-qte
Imitating, Fast and Slow: Robust learning from demonstrations via decision-time planning
Don't Change the Algorithm, Change the Data: Exploratory Data for Offline Reinforcement Learning
Coarse-to-Fine Q-attention with Learned Path Ranking
Comments Project page and code: https://sites.google.com/view/q-attention-lpr
Pretraining Graph Neural Networks for few-shot Analog Circuit Modeling and Design
CIC: Contrastive Intrinsic Control for Unsupervised Skill Discovery
Comments Project website: https://sites.google.com/view/cicrl/
Adversarial Motion Priors Make Good Substitutes for Complex Reward Functions
Comments 8 pages, 6 figures, 3 tables
Explaining Reinforcement Learning Policies through Counterfactual Trajectories
Comments Accepted at ICML HILL 2021 Workshop
SURF: Semi-supervised Reward Learning with Data Augmentation for Feedback-efficient Preference-based Reinforcement Learning
Comments Accepted to ICLR 2022
Hierarchical Few-Shot Imitation with Skill Transition Models
JUMBO: Scalable Multi-task Bayesian Optimization using Offline Data
Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents
Comments Project website at https://huangwl18.github.io/language-planner
It Takes Four to Tango: Multiagent Selfplay for Automatic Curriculum Generation
ProMP: Proximal Meta-Policy Search
Comments The first three authors contributed equally. Published at ICLR 2019
Bingham Policy Parameterization for 3D Rotations in Reinforcement Learning
Comments Project page and code: https://sites.google.com/view/rl-bpp
Towards More Generalizable One-shot Visual Imitation Learning
Likelihood Contribution based Multi-scale Architecture for Generative Flows
Contrastive Code Representation Learning
Comments In Proceedings of EMNLP 2021. 19 pages, 16 figures, 9 tables. Code available at https://github.com/parasj/contracode
Mastering Atari Games with Limited Data
Comments Published at NeurIPS 2021; Homepage: https://yewr.github.io/projects/efficientzero/
Target Entropy Annealing for Discrete Soft Actor-Critic
Journal ref neurips 2021 deep rl workshop
Hindsight Task Relabelling: Experience Replay for Sparse Reward Meta-RL
Count-Based Temperature Scheduling for Maximum Entropy Reinforcement Learning
Generalization in Dexterous Manipulation via Geometry-Aware Multi-Task Learning
Comments Website at https://huangwl18.github.io/geometry-dex
B-Pref: Benchmarking Preference-Based Reinforcement Learning
Comments NeurIPS Datasets and Benchmarks Track 2021. Code is available at https://github.com/rll-research/B-Pref
Offline-to-Online Reinforcement Learning via Balanced Replay and Pessimistic Q-Ensemble
Comments CoRL 2021. First two authors contributed equally
URLB: Unsupervised Reinforcement Learning Benchmark
Comments Code for the Unsupervised Reinforcement Learning Benchmark is available at https://github.com/rll-research/url_benchmark
Temporal-Difference Value Estimation via Uncertainty-Guided Soft Updates
Comments Accepted to Deep Reinforcement Learning Workshop @ NeurIPS 2021
Behavior From the Void: Unsupervised Active Pre-Training
Comments Advances in Neural Information Processing Systems(NeurIPS), 2021 Spotlight