arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 4126 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 模仿学习与强化学习 4126 篇

2511.18617 2025-11-26 cs.RO cs.CV 62%

AutoFocus-IL: VLM-based Saliency Maps for Data-Efficient Visual Imitation Learning without Extra Human Annotations

AutoFocus-IL:基于视觉语言模型的数据高效视觉模仿学习中的显著性图

Litian Gong, Fatemeh Bahrani, Yutai Zhou, Amin Banayeeanzade, Jiachen Li, Erdem Bıyık

机构 * Department of Electrical and Computer Engineering, University of California, Riverside, USA(电气与计算机工程系,加州大学河滨分校) Thomas Lord Department of Computer Science, University of Southern California, USA(汤姆斯·劳德计算机科学系,南加州大学)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO、cs.CV

AI总结 AutoFocus-IL通过视觉语言模型自动生成显著性图,提升视觉模仿学习的数据效率和泛化能力,无需额外人类标注。

Comments 8 pages, 6 figures. Code and datasets available at http://autofocus-il.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19528 2025-11-26 cs.RO cs.AI 62%

Discover, Learn, and Reinforce: Scaling Vision-Language-Action Pretraining with Diverse RL-Generated Trajectories

发现、学习与强化:通过多样化强化学习生成轨迹扩展视觉-语言-动作预训练

Rushuai Yang, Zhiyuan Feng, Tianxiang Zhang, Kaixin Wang, Chuheng Zhang, Li Zhao, Xiu Su, Yi Chen, Jiang Bian

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) Tsinghua University(清华大学) Wuhan University(武汉大学) Central South University(中南大学) Microsoft Research(微软研究院)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO、cs.AI

AI总结 本文提出DLR框架,通过多样化强化学习生成轨迹,提升VLA预训练的多样性和扩展性,实现更广泛的状态-动作空间覆盖。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18606 2025-11-25 cs.RO cs.LG 62%

How to Train Your Latent Control Barrier Function: Smooth Safety Filtering Under Hard-to-Model Constraints

如何训练您的潜在控制障碍函数:在难以建模的约束下实现平滑的安全过滤

Kensuke Nakamura, Arun L. Bishop, Steven Man, Aaron M. Johnson, Zachary Manchester, Andrea Bajcsy

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO、cs.LG

AI总结 本文提出LatentCBF方法,通过梯度惩罚和价值训练解决潜在空间中不兼容价值函数的问题,实现平滑安全过滤并提升任务完成率。

Comments 3 figures, 10 tables, 22 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15284 2025-11-20 cs.RO cs.AI 62%

Path Planning through Multi-Agent Reinforcement Learning in Dynamic Environments

Jonas De Maeyer, Hossein Yarahmadi, Moharram Challenger

机构 * Department of Computer Science University of Antwerp (UA)(安特卫普大学计算机科学系) Department of Computer Engineering, Faculty of Engineering, Ayatollah Boroujerdi University(阿亚图拉·博鲁杰尔迪大学工程学院计算机工程系) Department of Computer Science University of Antwerp (UA) and Flanders Make(安特卫普大学计算机科学系和弗拉芒制造)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14153 2025-11-19 cs.CV cs.AI 62%

LENS: Learning to Segment Anything with Unified Reinforced Reasoning

Lianghui Zhu, Bin Ouyang, Yuxuan Zhang, Tianheng Cheng, Rui Hu, Haocheng Shen, Longjin Ran, Xiaoxin Chen, Li Yu, Wenyu Liu, Xinggang Wang

机构 * vivo Mobile Communication Co., Ltd.(vivo移动通信有限公司)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.CV

Comments Code is released at https://github.com/hustvl/LENS

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11992 2025-11-18 cs.MA cs.AI cs.LG 62%

Goal-Oriented Multi-Agent Reinforcement Learning for Decentralized Agent Teams

Hung Du, Hy Nguyen, Srikanth Thudumu, Rajesh Vasa, Kon Mouzakis

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments Accepted poster at the IEEE Consumer Communications & Networking Conference (CCNC) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.09010 2025-11-17 cs.RO cs.LG cs.SY eess.SY 62%

DiAReL: Reinforcement Learning with Disturbance Awareness for Robust Sim2Real Policy Transfer in Robot Control

Mohammadhossein Malmir, Josip Josifovski, Noah Klarmann, Alois Knoll

机构 * Department of Computer Engineering, School of Computation, Information and Technology, Technical University of Munich(计算机工程系,计算、信息与技术学院,慕尼黑技术大学) Rosenheim University of Applied Sciences(罗森海姆应用技术大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

Comments Accepted for publication in IEEE Transactions on Control Systems Technology (TCST)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.09852 2025-11-13 cs.RO cs.LG cs.MA 62%

Strategic Coordination of Drones via Short-term Distributed Optimization and Long-term Reinforcement Learning

Chuhao Qin, Evangelos Pournaras

机构 * School of Computer Science, University of Leeds(利兹大学计算机科学学院)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO、cs.LG

Comments 23 pages, 16 figures, accepted by Applied Soft Computing

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08339 2025-11-12 cs.LG cs.AI 62%

LPPG-RL: Lexicographically Projected Policy Gradient Reinforcement Learning with Subproblem Exploration

Ruiyu Qiu, Rui Wang, Guanghui Yang, Xiang Li, Zhijiang Shao

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07942 2025-11-12 cs.LG cs.AI 62%

Balance Equation-based Distributionally Robust Offline Imitation Learning

Rishabh Agrawal, Yusuf Alvi, Rahul Jain, Ashutosh Nayyar

机构 * University of Southern California(南加州大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05672 2025-11-12 stat.ML cs.AI cs.LG cs.NE cs.SY eess.SY 62%

On the Convergence and Stability of Upside-Down Reinforcement Learning, Goal-Conditioned Supervised Learning, and Online Decision Transformers

Miroslav Štrupl, Oleg Szehr, Francesco Faccio, Dylan R. Ashley, Rupesh Kumar Srivastava, Jürgen Schmidhuber

机构 * Dalle Molle Institute for Artificial Intelligence (IDSIA) - USI/SUPSI(达摩信息技术研究所(IDSIA)- USI/SUPSI) Center of Excellence for Generative AI, King Abdullah University of Science and Technology(生成人工智能卓越中心,国王阿卜杜勒阿齐兹大学科学与技术学院) NNAISENSE

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

Comments 85 pages in main text + 4 pages of references + 26 pages of appendices, 12 figures in main text + 2 figures in appendices; source code available at https://github.com/struplm/eUDRL-GCSL-ODT-Convergence-public

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07155 2025-11-11 cs.RO cs.LG 62%

Dynamics-Decoupled Trajectory Alignment for Sim-to-Real Transfer in Reinforcement Learning for Autonomous Driving

Thomas Steinecker, Alexander Bienemann, Denis Trescher, Thorsten Luettel, Mirko Maehlisch

机构 * University of the Bundeswehr Munich(联邦国防军慕尼黑大学)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06252 2025-11-11 cs.LG cs.AI 62%

MrCoM: A Meta-Regularized World-Model Generalizing Across Multi-Scenarios

Xuantang Xiong, Ni Mu, Runpeng Xie, Senhao Yang, Yaqing Wang, Lexiang Wang, Yao Luan, Siyuan Li, Shuang Xu, Yiqin Yang, Bo Xu

专题命中 模仿学习与强化学习 :world model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17830 2025-11-05 cs.LG cs.AI 62%

Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning

Nicolas Castanet, Olivier Sigaud, Sylvain Lamprier

机构 * Sorbonne Université, CNRS, ISIR(索邦大学、国家科学研究中心、信息科学研究所)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.05818 2025-11-05 cs.LG cs.RO 62%

Closing the Intent-to-Behavior Gap via Fulfillment Priority Logic

Bassel El Mabsout, Abdelrahman Abdelgawad, Renato Mancuso

机构 * Department of Computer Science, Boston University(波士顿大学计算机科学系) Systems Engineering Division, Boston University(波士顿大学系统工程分校)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00880 2025-11-04 cs.LG cs.AI 62%

KFCPO: Kronecker-Factored Approximated Constrained Policy Optimization

Joonyoung Lim, Younghwan Yoo

机构 * School of Computer Science and Engineering, Pusan National University, Busan, Korea(计算机科学与工程学院,釜山国立大学,韩国釜山)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

Comments 12 pages, 8 figures, submitted to ECAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00806 2025-11-04 cs.LG cs.AI 62%

Logic-informed reinforcement learning for cross-domain optimization of large-scale cyber-physical systems

Guangxi Wan, Peng Zeng, Xiaoting Dong, Chunhe Song, Shijie Cui, Dong Li, Qingwei Dong, Yiyang Liu, Hongfei Bai

机构 * State Key Laboratory of Robotics and Intelligent Systems(机器人与智能系统国家重点实验室) Shenyang Institute of Automation(沈阳自动化研究所) Chinese Academy of Sciences(中国科学院) University of Chinese Academy of Sciences(中国科学院大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00091 2025-11-04 cs.CV cs.RO 62%

Self-Improving Vision-Language-Action Models with Data Generation via Residual RL

Wenli Xiao, Haotian Lin, Andy Peng, Haoru Xue, Tairan He, Yuqi Xie, Fengyuan Hu, Jimmy Wu, Zhengyi Luo, Linxi "Jim" Fan, Guanya Shi, Yuke Zhu

机构 * NVIDIA(NVIDIA公司) CMU(卡内基梅隆大学) UC Berkeley(加州大学伯克利分校) UT Austin(德克萨斯大学奥斯汀分校)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO、cs.CV

Comments 26 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.11511 2025-11-04 cs.AI cs.CL cs.LG 62%

Multi-Step Reasoning with Large Language Models, a Survey

Aske Plaat, Annie Wong, Suzan Verberne, Joost Broekens, Niki van Stein, Thomas Back

机构 * LIACS(莱顿大学信息科学研究中心) Leiden University(莱顿大学)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments ACM Computing Surveys

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16724 2025-11-03 cs.AI cs.LG 62%

Towards Automated Semantic Interpretability in Reinforcement Learning via Vision-Language Models

Zhaoxin Li, Zhang Xi-Jia, Batuhan Altundas, Letian Chen, Rohan Paleja, Matthew Gombolay

机构 * Georgia Institute of Technology(佐治亚理工学院) Purdue University(普渡大学)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26089 2025-10-31 cs.LG cs.AI cs.MA 62%

Network-Constrained Policy Optimization for Adaptive Multi-agent Vehicle Routing

Fazel Arasteh, Arian Haghparast, Manos Papagelis

机构 * York University(约克大学)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments 29 pages, 12 figures. Fazel Arasteh and Arian Haghparast contributed equally to this research. Submitted to ACM Transactions on Spatial Algorithms and Systems (TSAS). The code for this work is publicly available at https://github.com/Arianhgh/HHAN

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20974 2025-10-29 cs.RO cs.LG 62%

Robust Point Cloud Reinforcement Learning via PCA-Based Canonicalization

Michael Bezick, Vittorio Giammarino, Ahmed H. Qureshi

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13414 2025-10-23 cs.LG cs.AI 62%

Provably Efficient Reward Transfer in Reinforcement Learning with Discrete Markov Decision Processes

Kevin Vora, Yu Zhang

机构 * School of Computing and Augmented Intelligence(计算与增强智能学院) Arizona State University(亚利桑那州立大学)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14318 2025-10-17 cs.CL cs.AI cs.LG 62%

Evaluating & Reducing Deceptive Dialogue From Language Models with Multi-turn RL

Marwa Abdulhai, Ryan Cheng, Aryansh Shrivastava, Natasha Jaques, Yarin Gal, Sergey Levine

机构 * UC Berkeley(伯克利大学) University of Oxford(牛津大学) University of Washington(华盛顿大学) UK AI Security Institute(英国人工智能安全研究所) Google DeepMind(谷歌DeepMind)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12072 2025-10-15 cs.AI cs.RO 62%

EmboMatrix: A Scalable Training-Ground for Embodied Decision-Making

Zixing Lei, Sheng Yin, Yichen Xiong, Yuanzhuo Ding, Wenhao Huang, Yuxi Wei, Qingyao Xu, Yiming Li, Weixin Li, Yunhong Wang, Siheng Chen

机构 * Shanghai Jiao Tong University(上海交通大学) Zhongguancun Academy(中关村学院) New York University(纽约大学)

专题命中 模仿学习与强化学习 :embodied agent(abstract);分类 cs.RO、cs.AI

Comments 10 pages 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23135 2025-10-14 cs.LG cs.AI 62%

Trust Region Reward Optimization and Proximal Inverse Reward Optimization Algorithm

Yang Chen, Menglin Zou, Jiaqi Zhang, Yitan Zhang, Junyi Yang, Gael Gendron, Libo Zhang, Jiamou Liu, Michael J. Witbrock

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) University of Auckland(奥克兰大学) Chongqing University(重庆大学)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments Accepted to NeurIPS 2025. Title used at submission and review: PIRO: Toward Stable Reward Learning for Inverse RL via Monotonic Policy Divergence Reduction

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11501 2025-10-14 cs.LG cs.RO 62%

Context-Aware Model-Based Reinforcement Learning for Autonomous Racing

Emran Yasser Moustafa, Ivana Dusparic

机构 * School of Computer Science and Statistics, Trinity College Dublin(计算机科学与统计学系,特里尼蒂学院都柏林)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

Comments Accepted to IEEE ICAR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10451 2025-10-14 cs.LG cs.AI 62%

Data-driven simulator of multi-animal behavior with unknown dynamics via offline and online reinforcement learning

Keisuke Fujii, Kazushi Tsutsui, Yu Teshima, Makoto Itoh, Naoya Takeishi, Nozomi Nishiumi, Ryoya Tanaka, Shunsuke Shigaki, Yoshinobu Kawahara

机构 * Graduate School of Informatics, Nagoya University, Japan(名古屋大学信息学研究科) RIKEN Center for Advanced Intelligence Project, Japan(RIKEN高级智能项目研究中心) Graduate School of Arts and Sciences, The University of Tokyo, Japan(东京大学文学系研究科) Project team for SIP, Japan(SIP项目团队) Japan Agency for Marine-Earth Science and Technology, Japan(日本海洋地球科学技术机构) Faculty of Education, Shitennoji University, Japan(世田谷大学教育学部) Graduate School of Engineering, The University of Tokyo, Japan(东京大学工学研究科) Graduate School of Science and Technology, Niigata University, Japan(新潟大学科学技术研究科) Graduate School of Science, Nagoya University, Japan(名古屋大学理学研究科) Principles of Informatics Research Division, National Institute of Informatics, Japan(信息学原理研究部门,日本信息处理技术研究所) Graduate School of Information Science, The University of Osaka, Japan(大阪大学信息科学研究科)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 21 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26605 2025-10-14 cs.AI cs.LG 62%

Fine-tuning Behavioral Cloning Policies with Preference-Based Reinforcement Learning

Maël Macuglia, Paul Friedrich, Giorgia Ramponi

机构 * Department of Informatics, University of Zurich(苏黎世大学信息学院) ETH AI Center(ETH人工智能中心)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 85 pages (11 + references and appendix), 9 figures. v2: added acknowledgements

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15327 2025-10-13 cs.AI cs.LG 62%

Search-Based Credit Assignment for Offline Preference-Based Reinforcement Learning

Xiancheng Gao, Yufeng Shi, Wengang Zhou, Houqiang Li

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

Comments 7 pages, 6 figures, under review

详情

展开后加载摘要…

URL PDF HTML 收藏