作者
Sergey Levine
Robotics / Reinforcement Learning
Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Autonomous Evaluation and Refinement of Digital Agents
Comments Published at COLM 2024. Code at https://github.com/Berkeley-NLP/Agent-Eval-Refine
SELFI: Autonomous Self-Improvement with Reinforcement Learning for Social Navigation
Comments 20pages, 12 figures, 2 tables, Conference on Robot Learning 2024
LeLaN: Learning A Language-Conditioned Navigation Policy from In-the-Wild Videos
Comments 23 pages, 9 figures, 5 tables, Conference on Robot Learning 2024
KALIE: Fine-Tuning Vision-Language Models for Open-World Manipulation without Robot Data
Comments 8 pages, 7 figures
OpenVLA: An Open-Source Vision-Language-Action Model
Comments Website: https://openvla.github.io/
MOKA: Open-World Robotic Manipulation through Mark-Based Visual Prompting
FMB: a Functional Manipulation Benchmark for Generalizable Robotic Learning
Comments IJRR 2024
Unsupervised-to-Online Reinforcement Learning
Reinforcement Learning for Versatile, Dynamic, and Robust Bipedal Locomotion Control
Comments Accepted in International Journal of Robotics Research (IJRR) 2024. This is the author's version and will no longer be updated as the copyright may get transferred at anytime
Scaling Cross-Embodied Learning: One Policy for Manipulation, Navigation, Locomotion and Aviation
Comments Project website at https://crossformer-model.github.io/
D5RL: Diverse Datasets for Data-Driven Deep Reinforcement Learning
Comments RLC 2024
Chain of Code: Reasoning with a Language Model-Augmented Code Emulator
Comments ICML 2024 Oral; Project webpage: https://chain-of-code.github.io
Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review
Comments We plan to add more content/codes. Please let us know if there are any comments
Feedback Efficient Online Fine-Tuning of Diffusion Models
Comments Accepted at ICML 2024
Video Occupancy Models
Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs
HiLMa-Res: A General Hierarchical Framework via Residual RL for Combining Quadrupedal Locomotion and Manipulation
Comments IROS 2024
Commonsense Reasoning for Legged Robot Adaptation with Vision-Language Models
Comments 27 pages
AutoRT: Embodied Foundation Models for Large Scale Orchestration of Robotic Agents
Comments 26 pages, 9 figures, ICRA 2024 VLMNM Workshop
DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning
Comments 11 pages of main text, 28 pages in total
Strategically Conservative Q-Learning
Bridging Model-Based Optimization and Generative Modeling via Conservative Fine-Tuning of Diffusion Models
Comments Under review
Unfamiliar Finetuning Examples Control How Language Models Hallucinate
Octo: An Open-Source Generalist Robot Policy
Comments Project website: https://octo-models.github.io
Foundation Policies with Hilbert Representations
Comments ICML 2024
Learning Visuotactile Skills with Two Multifingered Hands
Comments Code and Project Website: https://toruowo.github.io/hato/