RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models
RLRC:基于强化学习的压缩视觉-语言-动作模型恢复
Yuxuan Chen, Yixin Han, Yize Huang, Xiao Li
机构
*
State Key Laboratory of Mechanical System and Vibration(机械系统与振动国家重点实验室)
;
Shanghai Key Laboratory of Intelligent Robotics(上海智能机器人重点实验室)
;
School of Mechanical Engineering, Shanghai Jiao Tong University(上海交通大学机械工程学院)
Roberto Capobianco, Harm van Seijen, Nolan D. Bard, Neil Burch, Fatima Davelouis, Josh Davidson, Alisa Devlic, Yunshu Du, Ishan Durugkar, Siddhant Gangapurwala, Daniel Hernandez, G. Zacharias Holland, Sahil Jain, Kenta Kawamoto, Raksha Kumaraswamy, Patrick MacAlpine, Dustin R. Morrill, Declan Oller, Francesco Riccio, Akanksha Saran, Craig Sherstan, Kaushik Subramanian, Thomas J. Walsh, Samuel Barrett, Kizza N. Frisbee, Mady Govil, Johannes Günther, Varun R. Kompella, James A. MacGlashan, Maxwell Svetlik, Michael D. Thomure, Jaden B. Travnik, Kevin Waugh, Elahe Aghapour, Florian Fuchs, Andreanne Lemay, Shruti Mishra, Takuma Seno, Peter Stone, Michael Spranger, Peter R. Wurman
机构
*
Sony AI, Zurich, Switzerland(索尼AI,苏黎世,瑞士)
;
Sony AI, North America (various locations)(索尼AI,北美(多地))
;
Sony AI, Tokyo, Japan(索尼AI,东京,日本)
机构
*
BioRobotics Institute, Scuola Superiore Sant’Anna, Pisa, Italy, and with the Department of Excellence in Robotics and AI, Scuola Superiore Sant’Anna, Pisa, Italy(生物机器人研究所,圣安娜高等学院,意大利比萨,以及与机器人与人工智能卓越部门,圣安娜高等学院,意大利比萨)
;
Institute for Systems and Robotics, Instituto Superior Técnico, Universidade de Lisboa(系统与机器人研究所,技术高等学院,里斯本大学)
The HydroGym Reinforcement Learning Platform for Fluid Dynamics
HydroGym:流体动力学的强化学习平台
Christian Lagemann, Sajeda Mokbel, Miro Gondrum, Mario Rüttgers, Yuning Wang, Pol Suárez, Ludger Paehler, Deniz A. Bezgin, Aaron B. Buhendwa, Jared L. Callaham, Samuel Ahnert, Nicholas Zolman, Xiao Shao, Jean-Christophe Loiseau, Nikolaus Adams, Matthias Meinke, Wolfgang Schröder, Kai Lagemann, Esther Lagemann, Ricardo Vinuesa, Steven L. Brunton
机构
*
University of Washington(华盛顿大学)
;
AI Institute in Dynamic Systems, University of Washington(华盛顿大学人工智能研究所)
;
RWTH Aachen University(亚琛工业大学)
;
Inha University(仁荷大学)
;
University of Michigan(密歇根大学)
;
KTH Royal Institute of Technology(瑞典皇家理工学院)
;
Technical University of Munich(慕尼黑工业大学)
;
Arts et Métiers Institute of Technology(国立高等工程技术学校)
;
CNAM(法国国立工艺学院)
;
DynFluid, HESAM Université(HESAM大学流体动力学实验室)
;
Munich Institute of Integrated Materials, Energy and Process Engineering, Technical University of Munich(慕尼黑工业大学综合材料、能源与工艺工程研究所)
;
JARA Center for Simulation and Data Science, RWTH Aachen University(亚琛工业大学JARA模拟与数据中心)
;
MediaTek Research(联发科技研究中心)
;
German Center for Neurodegenerative Diseases(德国神经退行性疾病中心)
Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL
Chain-of-Goals 分层策略用于长视界离线目标条件强化学习
Jinwoo Choi, Sang-Hyun Lee, Seung-Woo Seo
机构
*
Department of Electrical and Computer Engineering, Seoul National University, Seoul, South Korea(首尔国立大学电气与计算机工程系)
;
Department of Automotive Engineering, Ajou University, Gyeonggi-do, South Korea(全州大学汽车工程系)
机构
*
School of Artificial Intelligence, Shanghai Jiao Tong University, Shanghai, China(上海交通大学人工智能学院)
;
Zhongguancun Academy, Beijing, China(中关村学院)
;
School of Integrated Circuits, Shanghai Jiao Tong University, Shanghai, China(上海交通大学集成电路学院)
;
School of Computer Science, Shanghai Jiao Tong University, Shanghai, China(上海交通大学计算机科学学院)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University, Beijing, China(北京大学计算机科学学院多媒体信息处理国家重点实验室)
SHIELD: Safety on Humanoids via CBFs In Expectation on Learned Dynamics
SHIELD: 基于学习动力学期望的控制障碍函数实现人形机器人安全
Lizhi Yang, Blake Werner, Ryan K. Cosner, David Fridovich-Keil, Preston Culbertson, Aaron D. Ames
机构
*
Mechanical and Civil Engineering, California Institute of Technology(加州理工学院机械与土木工程系)
;
Aerospace Engineering and Engineering Mechanics, UT Austin(德克萨斯大学奥斯汀分校航空航天工程与工程力学系)
;
Computer Science, Cornell University(康奈尔大学计算机科学系)
CommentsAccepted to the 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025). Copyright transferred to IEEE. Video at https://youtu.be/-Qv1wR4jfj4
Training and Simulation of Quadrupedal Robot in Adaptive Stair Climbing and Descending for Indoor Firefighting: An End-to-End Reinforcement Learning Approach
四足机器人在室内灭火中的自适应爬楼梯训练与仿真:一种端到端强化学习方法
Baixiao Huang, Baiyu Huang, Yu Hou
机构
*
Independent Researcher(独立研究者)
;
Department of Construction Management, Western New England University(西雅图新英格兰大学建设管理系)
Dream-MPC: Gradient-Based Model Predictive Control with Latent Imagination
Dream-MPC:基于梯度与潜在想象的模型预测控制
Jonathan Spieler, Sven Behnke
机构
*
Autonomous Intelligent Systems, Computer Science Institute VI - Intelligent Systems(自主智能系统,计算机科学研究所VI - 智能系统)
;
Robotics, Center for Robotics(机器人学,机器人中心)
;
the Lamarr Institute for Machine Learning(拉马尔机器学习研究所)
;
Artificial Intelligence, University of Bonn, Germany(人工智能,波恩大学,德国)
机构
*
Department of Computing Science, University of Alberta(阿尔伯塔大学计算机科学系)
;
Alberta Machine Intelligence Institute(阿尔伯塔机器智能研究所)
;
Advanced Technology R&D Center, Mitsubishi Electric Corporation(三菱电机株式会社先进技术研发中心)
;
Information Technology R&D Center, Mitsubishi Electric Corporation(三菱电机株式会社信息技术研发中心)
机构
*
Institute of Cyber-Systems and Control, College of Control Science and Engineering, Zhejiang University(浙江大学控制科学与工程学院工业控制技术研究所)
;
Differential Robotics(微分机器人)
DeepForgeSeal: Latent Space-Driven Semi-Fragile Watermarking for Deepfake Detection Using Adversarial Reinforcement Learning
DeepForgeSeal:基于对抗强化学习的隐空间驱动半脆弱深度伪造检测水印技术
Tharindu Fernando, Clinton Fookes, Sridha Sridharan
机构
*
The Signal Processing, Artificial Intelligence and Vision Technologies (SAIVT), Queensland University of Technology, Australia(信号处理、人工智能与视觉技术研究所(SAIVT),昆士兰理工大学)
Imitation Learning from Human Motion Alone Does Not Guarantee Biomechanically Plausible Gait Kinetics
超越动作模仿:仅有人体动作数据是否足以解释步态控制与生物力学?
Xinyi Liu, Jangwhan Ahn, Edgar Lobaton, Jennie Si, He Huang
机构
*
Joint Department of Biomedical Engineering, University of North Carolina - Chapel Hill(北卡罗来纳大学教堂山分校联合生物医学工程系)
;
Department of Electrical and Computer Engineering, North Carolina State University(北卡罗来纳州立大学电气与计算机工程系)
;
School of Electrical, Computer and Energy Engineering, Arizona State University(亚利桑那州立大学电气、计算机与能源工程学院)