arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

International Conference on Robotics and Automation · 会议 · Robotics

2026-03-23 至 2026-03-23 共收录 5
2603.20164 2026-03-23 cs.RO cs.AI

The Robot's Inner Critic: Self-Refinement of Social Behaviors through VLM-based Replanning

机器人的内在批评者:通过基于视觉语言模型的重新规划实现社会行为的自我完善

Jiyu Lim, Youngwoo Yoon, Kwanghyun Park

机构 * KwangWoon University(匡文大学) ETRI(电子技术研究院)

AI总结 本文提出CRISP框架,通过基于视觉语言模型的重新规划,使机器人自主评估并优化其社会行为,提高灵活性和自主性,适用于多种平台。

Comments Accepted to ICRA 2026. 8 pages, 9 figures, Project page: https://limjiyu99.github.io/inner-critic/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.19628 2026-03-23 cs.CV cs.AI

Dual Prompt-Driven Feature Encoding for Nighttime UAV Tracking

双提示驱动的夜间无人机跟踪特征编码

Yiheng Wang, Changhong Fu, Liangliang Yao, Haobo Zuo, Zijie Zhang

机构 * Pratt School of Engineering, Duke University(杜克大学工程学院) School of Mechanical Engineering, Tongji University(同济大学机械工程学院) Shanghai Key Laboratory of Wearable Robotics and Human Machine Interaction, Tongji University(同济大学可穿戴机器人与人机交互重点实验室) School of Computing and Data Science, The University of Hong Kong(香港大学计算与数据科学学院)

AI总结 本文提出双提示驱动特征编码方法,通过多尺度频率感知光照提示和动态视角提示器提升夜间无人机跟踪的鲁棒性。

Comments Accepted to IEEE International Conference on Robotics and Automation 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02142 2026-03-23 cs.RO cs.CV

FD-VLA: Force-Distilled Vision-Language-Action Model for Contact-Rich Manipulation

FD-VLA:力感知的视觉-语言-动作模型用于接触密集的 manipulation

Ruiteng Zhao, Wenshuo Wang, Yicheng Ma, Xiaocong Li, Francis E. H. Tay, Marcelo H. Ang, Haiyue Zhu

机构 * Advanced Robotics Centre, National University of Singapore(新加坡国立大学先进机器人中心) School of Electrical & Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院) College of Information Science and Technology, Eastern Institute of Technology(东部技术学院信息科学与技术学院) John A. Paulson School of Engineering and Applied Sciences, Harvard University(哈佛大学约翰·A·保罗森工程与应用科学学院) Advanced Robotics Centre at National University of Singapore(新加坡国立大学先进机器人中心) Singapore Institute of Manufacturing Technology, Agency for Science, Technology and Research (A*STAR)(新加坡制造技术研究所,科技研究局(A*STAR))

AI总结 本文提出FD-VLA模型,通过力蒸馏模块在不依赖物理力传感器的情况下实现接触密集任务的力感知,提升机器人视觉-语言-动作的鲁棒性与实用性。

Comments ICRA 2026 Accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14608 2026-03-23 cs.RO

Latent Action Diffusion for Cross-Embodiment Manipulation

潜在动作扩散用于跨躯体操控

Erik Bauer, Elvis Nava, Robert K. Katzschmann

机构 * mimic robotics ETH AI Center Soft Robotics Lab Institute of Neuroinformatics

AI总结 本文提出通过潜在动作空间学习实现跨躯体操控,利用对比损失训练编码器统一不同末端执行器动作,提升多机器人控制效率和技能迁移效果。

Comments 8 pages, 5 figures. Accepted to the 2026 IEEE International Conference on Robotics & Automation (ICRA). Website: https://mimicrobotics.github.io/lad/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24129 2026-03-23 cs.RO cs.CV

Mash, Spread, Slice! Learning to Manipulate Object States via Visual Spatial Progress

Mash, Spread, Slice! 通过视觉空间进展学习操控物体状态

Priyanka Mandikal, Jiaheng Hu, Shivin Dass, Sagnik Majumder, Roberto Martín-Martín, Kristen Grauman

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

AI总结 SPARTA框架首次统一处理物体状态变化任务,通过空间进展和物体中心变化生成结构化策略观察和密集奖励,提升机器人操控效率和准确性。

Comments Accepted at ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏