arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

2025-09-30 至 2025-09-30 共收录 123 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 具身导航 21 篇

1911.04449 2025-09-30 astro-ph.IM 50%

Autonomous Detection of Particles and Tracks in Optical Images

Andrew J. Liounis, Jeffrey L. Small, Jason C. Swenson, Joshua R. Lyzhoft, Benjamin W. Ashman, Kenneth M. Getzandanner, Michael C. Moreau, Coralie D. Adam, Jason M. Leonard, Derek S. Nelson, John Y. Pelgrift, Brent J. Bos, Steven R. Chesley, Carl W. Hergenrother, Dante S. Lauretta

专题命中 具身导航 :navigation(abstract)

Comments 23 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 具身推理 6 篇

2509.24527 2025-09-30 cs.AI cs.LG cs.RO stat.ML 85%

Training Agents Inside of Scalable World Models

Danijar Hafner, Wilson Yan, Timothy Lillicrap

专题命中 具身推理 :world model(title,abstract);robotics(abstract);分类 cs.RO、cs.AI、cs.LG

Comments Website: https://danijar.com/dreamer4/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24524 2025-09-30 cs.RO cs.AI cs.SY eess.SY 84%

PhysiAgent: An Embodied Agent Framework in Physical World

Zhihao Wang, Jianxiong Li, Jinliang Zheng, Wencong Zhang, Dongxiu Liu, Yinan Zheng, Haoyi Niu, Junzhi Yu, Xianyuan Zhan

机构 * AIR, Tsinghua University(清华大学) Peking University(北京大学) University of California, Berkeley(加州大学伯克利分校)

专题命中 具身推理 :embodied agent(title,abstract);robotic(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24804 2025-09-30 cs.LG 79%

DyMoDreamer: World Modeling with Dynamic Modulation

Boxuan Zhang, Runqing Wang, Wei Xiao, Weipu Zhang, Jian Sun, Gao Huang, Jie Chen, Gang Wang

专题命中 具身推理 :world model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23979 2025-09-30 cs.CL 78%

ByteSized32Refactored: Towards an Extensible Interactive Text Games Corpus for LLM World Modeling and Evaluation

Haonan Wang, Junfeng Sun, Xingdi Yuan, Ruoyao Wang, Ziang Xiao

机构 * Johns Hopkins University(约翰霍普金斯大学) Liaoning Technical University(辽宁技术大学) Microsoft Research Montréal(微软研究院蒙特利尔分校) Central University of Finance and Economics(中央财经大学)

专题命中 具身推理 :world model(title,abstract)

Comments 14 pages,15 figures, Accepted to the 5th Wordplay: When Language Meets Games Workshop, EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25161 2025-09-30 cs.CV 57%

Rolling Forcing: Autoregressive Long Video Diffusion in Real Time

Kunhao Liu, Wenbo Hu, Jiale Xu, Ying Shan, Shijian Lu

机构 * Nanyang Technological University(南洋理工大学) ARC Lab, Tencent PCG(腾讯PCG ARC实验室)

专题命中 具身推理 :world model(abstract);分类 cs.CV

Comments Project page: https://kunhao-liu.github.io/Rolling_Forcing_Webpage/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13888 2025-09-30 cs.RO 57%

InSpire: Vision-Language-Action Models with Intrinsic Spatial Reasoning

Ji Zhang, Shihan Wu, Xu Luo, Hao Wu, Lianli Gao, Heng Tao Shen, Jingkuan Song

机构 * Southwest Jiaotong University(西南交通大学) University of Electronic Science and Technology of China(电子科技大学) Tongji University(同济大学)

专题命中 具身推理 :robotic(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 机器人基础模型 4 篇

2509.23655 2025-09-30 cs.RO cs.AI cs.CV cs.LG 78%

Focusing on What Matters: Object-Agent-centric Tokenization for Vision Language Action models

Rokas Bendikas, Daniel Dijkman, Markus Peschl, Sanjay Haresh, Pietro Mazzaglia

机构 * Centre for Artificial Intelligence, UCL(人工智能中心,伦敦大学学院) Qualcomm AI Research(高通人工智能研究)

专题命中 机器人基础模型 :manipulation(abstract);robotic(abstract);分类 cs.RO、cs.AI、cs.CV;robot learning(comments)

Comments Presented at 9th Conference on Robot Learning (CoRL 2025), Seoul, Korea

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22199 2025-09-30 cs.RO cs.AI 73%

MimicDreamer: Aligning Human and Robot Demonstrations for Scalable VLA Training

Haoyun Li, Ivan Zhang, Runqi Ouyang, Xiaofeng Wang, Zheng Zhu, Zhiqin Yang, Zhentao Zhang, Boyuan Wang, Chaojun Ni, Wenkang Qin, Xinze Chen, Yun Ye, Guan Huang, Zhenbo Song, Xingang Wang

机构 * GigaAI CASIA NJUST(南京工业大学) Tsinghua University(清华大学)

专题命中 机器人基础模型 :manipulation(abstract);robotic(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24768 2025-09-30 cs.RO 57%

IA-VLA: Input Augmentation for Vision-Language-Action models in settings with semantically complex tasks

Eric Hannus, Miika Malin, Tran Nguyen Le, Ville Kyrki

机构 * Intelligent Robotics Group at the Department of Electrical Engineering and Automation, School of Electrical Engineering, Aalto University(Aalto大学电气工程学院电气工程与自动化系智能机器人组) Biomimetics and Intelligent Systems Group at the Faculty of Information Technology and Electrical Engineering, University of Oulu(奥卢大学信息科技与电气工程学院仿生学与智能系统组) Section of Mechanical Technology at the Department of Engineering Technology and Didactics, Technical University of Denmark(丹麦技术大学工程技术与教学系机械技术部门)

专题命中 机器人基础模型 :manipulation(abstract);分类 cs.RO

Comments Under review for ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24559 2025-09-30 cs.LG 57%

Emergent World Representations in OpenVLA

Marco Molinari, Leonardo Nevali, Saharsha Navani, Omar G. Younis

机构 * London School of Economics(伦敦经济学院) ETH Zurich(苏黎世联邦理工学院) Princeton University(普林斯顿大学) Department of Computer Science(计算机科学系) Mila - Quebec AI Institute(魁北克人工智能研究所)

专题命中 机器人基础模型 :world model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 模仿学习与强化学习 15 篇

2408.14183 2025-09-30 cs.RO cs.AI cs.LG 82%

Robot Navigation with Entity-Based Collision Avoidance using Deep Reinforcement Learning

Yury Kolomeytsev, Dmitry Golembiovsky

机构 * Lomonosov Moscow State University, Moscow, Russia(罗蒙诺索夫莫斯科国立大学,莫斯科,俄罗斯)

专题命中 模仿学习与强化学习 :navigation(title,abstract);分类 cs.RO、cs.AI、cs.LG

Comments 15 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23958 2025-09-30 cs.CV 79%

Reinforcement Learning with Inverse Rewards for World Model Post-training

Yang Ye, Tianyu He, Shuo Yang, Jiang Bian

专题命中 模仿学习与强化学习 :world model(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24219 2025-09-30 cs.RO cs.AI cs.LG 78%

ViReSkill: Vision-Grounded Replanning with Skill Memory for LLM-Based Planning in Lifelong Robot Learning

Tomoyuki Kagaya, Subramanian Lakshmi, Anbang Ye, Thong Jing Yuan, Jayashree Karlekar, Sugiri Pranata, Natsuki Murakami, Akira Kinose, Yang You

机构 * Panasonic Connect Co., Ltd.(松下电器(株式会社)) Panasonic R&D Center(松下研发中心) National University of Singapore(新加坡国立大学)

专题命中 模仿学习与强化学习 :robot learning(title);分类 cs.RO、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24539 2025-09-30 cs.RO 70%

Unlocking the Potential of Soft Actor-Critic for Imitation Learning

Nayari Marie Lessa, Melya Boukheddimi, Frank Kirchner

机构 * DFKI GmbH(德累斯顿夫琅和斐研究所) University of Bremen(不莱梅大学)

专题命中 模仿学习与强化学习 :robotics(abstract);robotic(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23802 2025-09-30 cs.LG 70%

STAIR: Addressing Stage Misalignment through Temporal-Aligned Preference Reinforcement Learning

Yao Luan, Ni Mu, Yiqin Yang, Bo Xu, Qing-Shan Jia

机构 * Beijing Key Laboratory of Embodied Intelligence Systems, Department of Automation, Tsinghua University, Beijing, China(北京智能体智能系统重点实验室,自动化系,清华大学,北京,中国) The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences, Beijing, China(复杂系统认知与决策智能重点实验室,自动化研究所,中国科学院,北京,中国)

专题命中 模仿学习与强化学习 :manipulation(abstract);navigation(abstract);分类 cs.LG

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20766 2025-09-30 cs.RO cs.LG 62%

Leveraging Temporally Extended Behavior Sharing for Multi-task Reinforcement Learning

Gawon Lee, Daesol Cho, H. Jin Kim

机构 * Department of Aerospace Engineering, Seoul National University(首尔国立大学航空航天工程系) Artificial Intelligence Institute, Seoul National University(首尔国立大学人工智能研究所)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

Comments Accepted for publication in the proceedings of the 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06775 2025-09-30 eess.SY cs.AI cs.IT cs.LG cs.NI cs.SY math.IT 62%

Agentic DDQN-Based Scheduling for Licensed and Unlicensed Band Allocation in Sidelink Networks

Po-Heng Chou, Pin-Qi Fu, Walid Saad, Li-Chun Wang

机构 * Department of Electronics and Electrical Engineering(电子与电气工程系) Institute of Communications Engineering(通讯工程研究所) National Yang Ming Chiao Tung University(阳明交通大学) Research Center for Information Technology Innovation(信息技术创新研究中心) Academia Sinica(台湾“中央研究院”) Bradley Department of Electrical and Computer Engineering(电气与计算机工程系)

专题命中 模仿学习与强化学习 :embodied agent(abstract);分类 cs.AI、cs.LG

Comments 6 pages, 3 figures, accepted by 2025 IEEE Globecom Workshops

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08440 2025-09-30 cs.RO cs.AI 62%

TGRPO :Fine-tuning Vision-Language-Action Model via Trajectory-wise Group Relative Policy Optimization

Zengjue Chen, Runliang Niu, He Kong, Qi Wang, Qianli Xing, Zipei Fan

机构 * School of Artificial Intelligence, Jilin University(人工智能学院,吉林大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10482 2025-09-30 cs.LG cs.AI 62%

Fine-tuning Diffusion Policies with Backpropagation Through Diffusion Timesteps

Ningyuan Yang, Jiaxuan Gao, Feng Gao, Yi Wu, Chao Yu

机构 * Tsinghua University(清华大学)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 9 pages for main text, 23 pages in total, submitted to Neurips, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12243 2025-09-30 cs.RO cs.AI 62%

RISE: Robust Imitation through Stochastic Encoding

Mumuksh Tayal, Manan Tayal, Ravi Prakash

机构 * Cyber Physical Systems, Indian Institute of Science (IISc), Bengaluru(印度科学研究院计算机物理系统,班加罗尔)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.AI

Comments 15 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19752 2025-09-30 cs.RO 57%

Beyond Human Demonstrations: Diffusion-Based Reinforcement Learning to Generate Data for VLA Training

Rushuai Yang, Hangxing Wei, Ran Zhang, Zhiyuan Feng, Xiaoyu Chen, Tong Li, Chuheng Zhang, Li Zhao, Jiang Bian, Xiu Su, Yi Chen

机构 * Hong Kong University of Science and Technology(香港科技大学) Microsoft Research Asia(微软亚洲研究院) Wuhan University(武汉大学) University of Chinese Academy of Sciences(中国科学院大学) Tsinghua University(清华大学) Big Data Institute, Central South University(中南大学大数据研究院)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23220 2025-09-30 cs.RO 57%

GLUE: Global-Local Unified Encoding for Imitation Learning via Key-Patch Tracking

Ye Chen, Zichen Zhou, Jianyu Dou, Te Cui, Yi Yang, Yufeng Yue

机构 * Beijing Institute of Technology(北京理工大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO

Comments 8 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23219 2025-09-30 cs.LG 57%

WirelessMathLM: Teaching Mathematical Reasoning for LLMs in Wireless Communications with Reinforcement Learning

Xin Li, Mengbing Liu, Yiyang Zhu, Wenhe Zhang, Li Wei, Jiancheng An, Chau Yuen

机构 * Nanyang Technological University(南洋理工大学)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.LG

Comments Project Homepage: https://lixin.ai/WirelessMathLM

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16757 2025-09-30 cs.RO 57%

HDMI: Learning Interactive Humanoid Whole-Body Control from Human Videos

Haoyang Weng, Yitang Li, Nikhil Sobanbabu, Zihan Wang, Zhengyi Luo, Tairan He, Deva Ramanan, Guanya Shi

机构 * Robotics Institute, Carnegie Mellon University(卡内基梅隆大学机器人研究所)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO

Comments website: hdmi-humanoid.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23425 2025-09-30 cs.MA 50%

Situational Awareness for Safe and Robust Multi-Agent Interactions Under Uncertainty

Benjamin Alcorn, Eman Hammad

专题命中 模仿学习与强化学习 :robotics(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 机器人数据与评测 28 篇

2509.23328 2025-09-30 cs.RO cs.AI cs.LG 89%

Space Robotics Bench: Robot Learning Beyond Earth

Andrej Orsula, Matthieu Geist, Miguel Olivares-Mendez, Carol Martinez

机构 * University of Luxembourg(卢森堡大学) Earth Species Project(地球物种计划)

专题命中 机器人数据与评测 :robotics(title,abstract);robot learning(title,abstract);分类 cs.RO、cs.AI、cs.LG

Comments The source code is available at https://github.com/AndrejOrsula/space_robotics_bench

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15635 2025-09-30 cs.CV cs.RO 88%

FindingDory: A Benchmark to Evaluate Memory in Embodied Agents

Karmesh Yadav, Yusuf Ali, Gunshi Gupta, Yarin Gal, Zsolt Kira

机构 * Georgia Tech(佐治亚理工学院) University of Oxford(牛津大学)

专题命中 机器人数据与评测 :embodied agent(title,abstract);robotics(abstract);manipulation(abstract);navigation(abstract)

Comments Our dataset and code can be found at: https://findingdory-benchmark.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24163 2025-09-30 cs.RO 83%

Preference-Based Long-Horizon Robotic Stacking with Multimodal Large Language Models

Wanming Yu, Adrian Röfer, Abhinav Valada, Sethu Vijayakumar

机构 * School of Informatics, University of Edinburgh(信息学院,爱丁堡大学) University of Freiburg(弗赖堡大学)

专题命中 机器人数据与评测 :robotic(title,abstract);manipulation(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25032 2025-09-30 cs.RO cs.AI cs.CV 82%

AIRoA MoMa Dataset: A Large-Scale Hierarchical Dataset for Mobile Manipulation

Ryosuke Takanami, Petr Khrapchenkov, Shu Morikuni, Jumpei Arima, Yuta Takaba, Shunsuke Maeda, Takuya Okubo, Genki Sano, Satoshi Sekioka, Aoi Kadoya, Motonari Kambara, Naoya Nishiura, Haruto Suzuki, Takanori Yoshimoto, Koya Sakamoto, Shinnosuke Ono, Hu Yang, Daichi Yashima, Aoi Horo, Tomohiro Motoda, Kensuke Chiyoma, Hiroshi Ito, Koki Fukuda, Akihito Goto, Kazumi Morinaga, Yuya Ikeda, Riko Kawada, Masaki Yoshikawa, Norio Kosuge, Yuki Noguchi, Kei Ota, Tatsuya Matsushima, Yusuke Iwasawa, Yutaka Matsuo, Tetsuya Ogata

机构 * The University of Tokyo(东京大学) AI Robot Association (AIRoA)(人工智能机器人协会) Toyota Motor Corporation(丰田汽车公司) Telexistence, Inc.(Telexistence公司) National Institute of Advanced Industrial Science and Technology (AIST)(国家先进工业科学与技术研究院) Waseda University(早稻田大学)

专题命中 机器人数据与评测 :manipulation(title,abstract);分类 cs.RO、cs.AI、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏