arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

2025-09-12 至 2025-09-12 共收录 5 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 模仿学习与强化学习 5 篇

2502.17813 2025-09-12 cs.RO cs.LG 81%

Safe Multi-Agent Navigation guided by Goal-Conditioned Safe Reinforcement Learning

Meng Feng, Viraj Parimi, Brian Williams

机构 * Computer Science and Artificial Intelligence Laboratory, Massachusetts Institute of Technology(计算机科学与人工智能实验室,麻省理工学院)

专题命中 模仿学习与强化学习 :navigation(title,abstract);分类 cs.RO、cs.LG

Comments Due to the limitation "The abstract field cannot be longer than 1,920 characters", the abstract here is shorter than that in the PDF file

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09674 2025-09-12 cs.RO cs.AI cs.CL cs.LG 75%

SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning

Haozhan Li, Yuxin Zuo, Jiale Yu, Yuhao Zhang, Zhaohui Yang, Kaiyan Zhang, Xuekai Zhu, Yuchen Zhang, Tianxing Chen, Ganqu Cui, Dehui Wang, Dingxiang Luo, Yuchen Fan, Youbang Sun, Jia Zeng, Jiangmiao Pang, Shanghang Zhang, Yu Wang, Yao Mu, Bowen Zhou, Ning Ding

机构 * Shanghai Jiao Tong University(上海交通大学) Peking University(北京大学) The University of Hong Kong(香港大学) Shanghai AI Lab(上海人工智能实验室)

专题命中 模仿学习与强化学习 :manipulation(abstract);robotic(abstract);分类 cs.RO、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09356 2025-09-12 cs.AI cs.RO 73%

Curriculum-Based Multi-Tier Semantic Exploration via Deep Reinforcement Learning

Abdel Hakim Drid, Vincenzo Suriani, Daniele Nardi, Abderrezzak Debilou

机构 * Department of Electrical Engineering - Mohamed Khider, University of Biskra, Biskra (Algeria)(巴尔克拉大学电子工程系) Department of Engineering - University of Basilicata, Potenza (Italy)(巴塞里卡大学工程系) Department of Computer, Control, and Management Engineering ``Antonio Ruberti'', Sapienza University of Rome, Rome (Italy)(罗马萨皮恩扎大学计算机、控制与管理工程系)

专题命中 模仿学习与强化学习 :robotics(abstract);embodied agent(abstract);分类 cs.RO、cs.AI

Comments The 19th International Conference on Intelligent Autonomous Systems (IAS 19), 2025, Genoa

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08775 2025-09-12 cs.RO 70%

Joint Model-based Model-free Diffusion for Planning with Constraints

Wonsuhk Jung, Utkarsh A. Mishra, Nadun Ranawaka Arachchige, Yongxin Chen, Danfei Xu, Shreyas Kousik

机构 * Georgia Institute of Technology(佐治亚理工学院)

专题命中 模仿学习与强化学习 :manipulation(abstract);robotic(abstract);分类 cs.RO

Comments The first two authors contributed equally. Last three authors advised equally. Accepted to CoRL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08257 2025-09-12 cs.RO cs.AI 62%

Symmetry-Guided Multi-Agent Inverse Reinforcement Learning

Yongkai Tian, Yirong Qi, Xin Yu, Wenjun Wu, Jie Luo

机构 * State Key Laboratory of Complex & Critical Software Environment, Beihang University, Beijing, China(复杂与关键软件环境国家重点实验室,北京航空航天大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.AI

Comments 8pages, 6 figures. Accepted for publication in the Proceedings of the 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025) as oral presentation

详情

展开后加载摘要…

URL PDF HTML 收藏