作者
Pieter Abbeel
Robotics
Teachable Reinforcement Learning via Advice Distillation
Comments Published at NeurIPS 2021
On the Effectiveness of Fine-tuning Versus Meta-reinforcement Learning
The Wisdom of Hindsight Makes Language Models Better Instruction Followers
Language Quantized AutoEncoders: Towards Unsupervised Text-Image Alignment
Comments Fixed typos
Towards Better Few-Shot and Finetuning Performance with Forgetful Causal Language Models
Comments Added T-FCM and better FCM results
Multi-Environment Pretraining Enables Transfer to Action Limited Datasets
VectorFusion: Text-to-SVG by Abstracting Pixel-Based Diffusion Models
Comments Project webpage: https://ajayj.com/vectorfusion
Fleet-DAgger: Interactive Robot Fleet Learning with Scalable Human Supervision
Comments CoRL 2022 Oral
StereoPose: Category-Level 6D Transparent Object Pose Estimation from Stereo Images via Back-View NOCS
Comments 7 pages, 6 figures, Project homepage: https://appsrv.cse.cuhk.edu.hk/~kaichen/stereopose.html
Sim-to-Real via Sim-to-Seg: End-to-end Off-road Autonomous Driving Without Real Data
Comments CoRL 2022 Paper
Dichotomy of Control: Separating What You Can Control from What You Cannot
Spending Thinking Time Wisely: Accelerating MCTS with Virtual Expansions
Journal ref Published at NeurIPS 2022
Multimodal Masked Autoencoders Learn Transferable Representations
Learning Visual Robotic Control Efficiently with Contrastive Pre-training and Data Augmentation
Autoregressive Uncertainty Modeling for 3D Bounding Box Prediction
Comments In ECCV 2022. Code and dataset are available at https://bbox.yuxuanliu.com
Real-World Robot Learning with Masked Visual Pre-training
Comments CoRL 2022; Project page: https://tetexiao.com/projects/real-mvp
Reducing Variance in Temporal-Difference Value Estimation via Ensemble of Deep Networks
Journal ref ICML 2022
HARP: Autoregressive Latent Video Prediction with High-Fidelity Image Generator
Comments Extended draft of the paper accepted to ICIP 2022 conference
Multi-Objective Policy Gradients with Topological Constraints
AdaCat: Adaptive Categorical Discretization for Autoregressive Models
Comments Uncertainty in Artificial Intelligence (UAI) 2022 13 pages, 4 figures
Sim-to-Real 6D Object Pose Estimation via Iterative Self-training for Robotic Bin Picking
Comments Accepted to ECCV 2022
DayDreamer: World Models for Physical Robot Learning
Comments Website: https://danijar.com/daydreamer
Patch-based Object-centric Transformers for Efficient Video Generation
Comments Project Website: https://sites.google.com/view/povt-public
Reinforcement Learning with Action-Free Pre-Training from Videos
Comments International Conference on Machine Learning (ICML 2022). Project page: https://sites.google.com/view/rl-apv
Deep Hierarchical Planning from Pixels
Comments Website: https://danijar.com/director
Reward Uncertainty for Exploration in Preference-based Reinforcement Learning
Comments ICLR 2022. Last two authors advised equally
DoorGym: A Scalable Door Opening Environment And Baseline Agent
Comments Accepted to NeurIPS2019 Deep Reinforcement Learning Workshop. Full version
Chain of Thought Imitation with Procedure Cloning
An Empirical Investigation of Representation Learning for Imitation
Comments Accepted to NeurIPS2021 Datasets and Benchmarks Track