arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

2026-02-24 至 2026-02-24 共收录 99 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 机器人操作 28 篇

2602.18681 2026-02-24 cs.CR cs.ET 50%

Media Integrity and Authentication: Status, Directions, and Futures

媒体完整性与认证:现状、方向与未来

Jessica Young, Sam Vaughan, Andrew Jenks, Henrique Malvar, Christian Paquin, Paul England, Thomas Roca, Juan LaVista Ferres, Forough Poursabzi, Neil Coles, Ken Archer, Eric Horvitz

专题命中 机器人操作 :manipulation(abstract)

AI总结 本文探讨了媒体完整性与认证的现状,分析了区分AI生成内容与真实内容的技术方法,并提出了增强边缘设备安全性的方向。

Comments 56 pages

Journal ref Microsoft Research Technical Report, January 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 具身导航 22 篇

2602.20041 2026-02-24 cs.RO cs.CV 86%

EEG-Driven Intention Decoding: Offline Deep Learning Benchmarking on a Robotic Rover

基于EEG的意图解码:在机器人越野车上的离线深度学习基准测试

Ghadah Alosaimi, Maha Alsayyari, Yixin Sun, Stamos Katsigiannis, Amir Atapour-Abarghouei, Toby P. Breckon

机构 * Department of Computer Science, Imam Mohammad Ibn Saud Islamic University(计算机科学系,伊玛目穆罕默德·伊本·沙特伊斯兰大学) Department of Computer Science, King Saud University(计算机科学系,国王沙特大学)

专题命中 具身导航 :robotic(title,abstract);robotics(abstract);navigation(abstract);分类 cs.RO、cs.CV

AI总结 本文提出了一种基于EEG的机器人越野车意图解码框架,通过离线深度学习模型评估,发现ShallowConvNet在动作和意图预测中表现最佳。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20048 2026-02-24 cs.AI cs.SE 79%

CodeCompass: Navigating the Navigation Paradox in Agentic Code Intelligence

CodeCompass: 在代理代码智能中导航悖论的探索

Tarakanath Paipuru

机构 * Independent Researcher(独立研究者)

专题命中 具身导航 :navigation(title,abstract);分类 cs.AI

AI总结 CodeCompass通过基于图的结构导航提升代码智能代理在隐藏依赖任务中的性能,揭示导航与检索本质不同,需显式引导代理利用结构上下文。

Comments 23 pages, 7 figures. Research study with 258 trials on SWE-bench-lite tasks. Code and data: https://github.com/tpaip607/research-codecompass

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18981 2026-02-24 cs.AI 79%

How Far Can We Go with Pixels Alone? A Pilot Study on Screen-Only Navigation in Commercial 3D ARPGs

仅凭像素能走多远?对商业3DARPG中纯屏幕导航的初步研究

Kaijie Xu, Mustafa Bugti, Clark Verbrugge

机构 * McGill University(麦吉尔大学) Connecticut College(康奈尔学院)

专题命中 具身导航 :navigation(title,abstract);分类 cs.AI

AI总结 本文提出纯屏幕导航代理,通过视觉可能性探索3DARPG关卡,初步实验显示其导航能力,但受限于视觉模型的不足,无法实现全面自动导航。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18941 2026-02-24 cs.CV 79%

Global Commander and Local Operative: A Dual-Agent Framework for Scene Navigation

全局指挥官与局部执行者:一种双智能体框架用于场景导航

Kaiming Jin, Yuefan Wu, Shengqiong Wu, Bobo Li, Shuicheng Yan, Tat-Seng Chua

机构 * National University of Singapore(新加坡国立大学) Simon Fraser University(西蒙弗雷泽大学) University of Oxford(牛津大学)

专题命中 具身导航 :navigation(title,abstract);分类 cs.CV

AI总结 DACo提出双智能体框架,通过解耦全局推理与局部执行,提升复杂环境中长时序导航的稳定性和性能。

Comments 18 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19346 2026-02-24 cs.RO cs.SY eess.SY 77%

Design and Control of Modular Magnetic Millirobots for Multimodal Locomotion and Shape Reconfiguration

模块化磁性微机器人多模态运动与形状重构设计与控制

Erik Garcia Oyono, Jialin Lin, Dandan Zhang

机构 * Department of Bioengineering, Imperial College London(帝国理工学院生物工程系)

专题命中 具身导航 :manipulation(abstract);navigation(abstract);robotic(abstract);分类 cs.RO

AI总结 本研究提出一种模块化磁性微机器人平台,通过多模块协同实现多模态运动与形状重构,展示了在受限环境中稳健控制的潜力。

Comments Accepted by 2026 ICRA

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19062 2026-02-24 cs.RO 74%

Path planning for unmanned surface vehicle based on predictive artificial potential field. International Journal of Advanced Robotic Systems

基于预测人工势场的无人水面车辆路径规划

Jia Song, Ce Hao, Jiangcheng Su

专题命中 具身导航 :robotic(title);分类 cs.RO

AI总结 本文提出一种结合时间信息和预测势场的改进方法,用于提升无人水面车辆路径规划的效率和避障能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19400 2026-02-24 cs.RO cs.AI cs.MA 73%

Hilbert-Augmented Reinforcement Learning for Scalable Multi-Robot Coverage and Exploration

Hilbert空间增强的强化学习用于可扩展的多机器人覆盖与探索

Tamil Selvan Gurunathan, Aryya Gangopadhyay

专题命中 具身导航 :robotics(abstract);robot learning(abstract);分类 cs.RO、cs.AI

AI总结 本文提出一种基于Hilbert空间的强化学习方法,用于提高多机器人覆盖与探索的效率和可扩展性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19308 2026-02-24 cs.RO cs.CV 73%

WildOS: Open-Vocabulary Object Search in the Wild

WildOS: 野外环境中的开放词汇物体搜索

Hardik Shah, Erica Tevere, Deegan Atha, Marcel Kaufmann, Shehryar Khattak, Manthan Patel, Marco Hutter, Jonas Frey, Patrick Spieler

机构 * Jet Propulsion Laboratory(喷气推进实验室) California Institute of Technology(加州理工学院) Swiss Federal Institute of Technology(瑞士联邦理工学院) ETH Zürich(苏黎世联邦理工学院) FieldAI Inc.(FieldAI公司) Stanford University(斯坦福大学) University of California Berkeley(加州大学伯克利分校)

专题命中 具身导航 :navigation(abstract);robotic(abstract);分类 cs.RO、cs.CV

AI总结 WildOS通过结合安全几何探索与语义视觉推理,实现野外环境中的开放词汇物体搜索,显著提升导航效率和自主性。

Comments 28 pages, 16 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18716 2026-02-24 cs.RO cs.AI 73%

Temporal Action Representation Learning for Tactical Resource Control and Subsequent Maneuver Generation

战术资源控制与后续机动生成的时序动作表示学习

Hoseong Jung, Sungil Son, Daesol Cho, Jonghae Park, Changhyun Choi, H. Jin Kim

机构 * Seoul National University(首尔国立大学) Georgia Institute of Technology(佐治亚理工学院)

专题命中 具身导航 :navigation(abstract);robotic(abstract);分类 cs.RO、cs.AI

AI总结 TART通过时序动作表示学习框架,有效整合资源使用与机动生成,提升有限资源下的战术决策能力。

Comments ICRA 2026, 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24243 2026-02-24 cs.RO cs.AI 73%

SafeFlowMatcher: Safe and Fast Planning using Flow Matching with Control Barrier Functions

SafeFlowMatcher: 基于流匹配与控制屏障函数的安全且快速规划

Jeongyong Yang, Seunghwan Jang, SooJean Han

机构 * Korea Advanced Institute of Science and Technology (KAIST)(韩国科学技术院)

专题命中 具身导航 :manipulation(abstract);navigation(abstract);分类 cs.RO、cs.AI

AI总结 SafeFlowMatcher通过结合流匹配与控制屏障函数,实现安全且高效的路径规划,在多个机器人任务中表现出更快、更平滑和更安全的性能。

Comments ICLR 2026(poster)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18976 2026-02-24 cs.RO 72%

Bumper Drone: Elastic Morphology Design for Aerial Physical Interaction

bumper无人机:用于空中物理交互的弹性形态设计

Pongporn Supa, Alex Dunnett, Feng Xiao, Rui Wu, Mirko Kovac, Basaran Bahadir Kocer

机构 * University of Bristol(布里斯托大学) School of Engineering Mathematics and Technology(工程数学与技术学院) School of Civil, Aerospace and Design Engineering(土木、航空航天与设计工程学院) Department of Aeronautics(航空系) Laboratory of Sustainability Robotics, EMPA(可持续机器人实验室,EMPA)

专题命中 具身导航 :manipulation(abstract);navigation(abstract);分类 cs.RO;robotics(comments)

AI总结 Bumper Drone通过弹性结构实现稳定的空中物理交互,减少pitch振荡并提升环境接触的稳定性与操控性。

Comments Accepted to the 9th IEEE-RAS International Conference on Soft Robotics (RoboSoft) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04847 2026-02-24 cs.LG 70%

Test-Time Adaptation for LLM Agents via Environment Interaction

通过环境交互实现LLM代理的测试时适应

Arthur Chen, Zuxin Liu, Jianguo Zhang, Akshara Prabhakar, Zhiwei Liu, Shelby Heinecke, Silvio Savarese, Victor Zhong, Caiming Xiong

机构 * University of Waterloo(滑铁卢大学) Salesforce AI Research(Salesforce AI研究)

专题命中 具身导航 :navigation(abstract);world model(abstract);分类 cs.LG

AI总结 通过环境交互实现LLM代理的测试时适应,利用语法对齐和动态接地策略提升代理在复杂环境中的泛化能力。

Comments Our code is available here: https://github.com/r2llab/GTTA

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18721 2026-02-24 cond-mat.mtrl-sci cs.RO 70%

A Machine Learning Approach Capturing Hidden Parameters in Autonomous Thin-Film Deposition

一种通过机器学习捕捉自主薄膜沉积中隐藏参数的方法

Yuanlong Zheng, Connor Blake, Layla Mravac, Fengxue Zhang, Yuxin Chen, Shuolong Yang

专题命中 具身导航 :robotics(abstract);robotic(abstract);分类 cs.RO

AI总结 本文提出一种结合原位光谱、机器人系统和高斯过程回归的自动化薄膜沉积方法,通过校准层和主动学习算法实现高精度薄膜制备,显著提升材料开发效率。

Journal ref npj Computational Materials 11, 327 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12846 2026-02-24 cs.RO cs.CV 62%

Unleashing the Power of Discrete-Time State Representation: Ultrafast Target-based IMU-Camera Spatial-Temporal Calibration

释放离散时间状态表示的潜力:超快基于目标的IMU-相机空间-时间校准

Junlin Song, Antoine Richard, Miguel Olivares-Mendez

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 本文提出了一种基于离散时间状态表示的高效IMU-相机空间-时间校准方法,以提高校准效率并解决连续时间表示的计算成本问题。

Comments Accepted by ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19698 2026-02-24 cs.DL cs.AI cs.CV cs.IR 62%

Iconographic Classification and Content-Based Recommendation for Digitized Artworks

图标分类与基于内容的数字艺术作品推荐

Krzysztof Kutt, Maciej Baczyński

机构 * Department of Human-Centered Artificial Intelligence(人中心人工智能系) Institute of Applied Computer Science(应用计算机科学研究所) Faculty of Physics, Astronomy and Applied Computer Science(物理、天文学与应用计算机科学学院) Jagiellonian University(雅盖隆大学)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV

AI总结 本文提出利用Iconclass词汇和AI方法,实现数字化艺术作品的图标分类与基于内容的推荐,通过四阶段工作流提升文化遗产库的导航效率。

Comments 14 pages, 7 figures; submitted to ICCS 2026 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19304 2026-02-24 cs.RO cs.AI cs.HC cs.MA 62%

Safe and Interpretable Multimodal Path Planning for Multi-Agent Cooperation

安全且可解释的多模态路径规划用于多智能体协作

Haojun Shi, Suyu Ye, Katherine M. Guerrerio, Jianzhi Shen, Yifan Yin, Daniel Khashabi, Chien-Ming Huang, Tianmin Shu

机构 * Johns Hopkins University(约翰霍普金斯大学)

专题命中 具身导航 :robotic(abstract);分类 cs.RO、cs.AI

AI总结 CaPE通过多模态路径规划实现安全且可解释的多智能体协作,利用视觉-语言模型和模型基于规划器确保路径调整的安全性和可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16874 2026-02-24 cs.MA cs.AI cs.RO 62%

Budget Allocation Policies for Real-Time Multi-Agent Path Finding

实时多智能体路径寻找中的预算分配策略

Raz Beck, Roni Stern

专题命中 具身导航 :robotics(abstract);分类 cs.RO、cs.AI

AI总结 本文提出了一种实时多智能体路径寻找中的预算分配策略,通过智能分配规划预算提升求解效率。

Comments 11 pages, 4 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16650 2026-02-24 eess.SY cs.LG cs.RO cs.SY math.DS math.OC 62%

Safe and Near-Optimal Control with Online Dynamics Learning

安全且近最优的控制与在线动力学学习

Manish Prajapat, Johannes Köhler, Melanie N. Zeilinger, Andreas Krause

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.LG

AI总结 本文提出了一种安全且近最优的在线控制方法,通过最大安全动态学习在有限时间内实现高精度动态建模,同时确保全程安全运行,适用于自动驾驶和无人机等对安全要求高的领域。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20093 2026-02-24 cs.IR 50%

ManCAR: Manifold-Constrained Latent Reasoning with Adaptive Test-Time Computation for Sequential Recommendation

ManCAR: 用于序列推荐的流形约束潜在推理与自适应测试时间计算

Kun Yang, Yuxuan Zhu, Yazhe Chen, Siyao Zheng, Bangyang Hong, Kangle Wu, Yabo Ni, Anxiang Zeng, Cong Fu, Hui Li

专题命中 具身导航 :navigation(abstract)

AI总结 ManCAR通过流形约束和自适应测试时间计算,提升序列推荐的推理精度和稳定性。

Comments 15 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19365 2026-02-24 physics.hist-ph 50%

Dutch Colonial Time: Time Signals in Paramaribo and the Dutch Caribbean

荷兰殖民时间:帕拉米波及荷兰加勒比地区的时信号

Richard de Grijs

专题命中 具身导航 :navigation(abstract)

AI总结 荷兰在殖民地建立了时信号系统,通过技术与仪式结合,将殖民地纳入全球导航体系,并反映了殖民管理的适应性与本地政治互动。

Comments 14 pages, incl. 5 figures. Accepted for publication in the Journal of Astronomical History and Heritage (December 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19364 2026-02-24 physics.hist-ph 50%

Marking Noon: The Time Balls and Time Flaps of the Netherlands

正午标记:荷兰的时间球与时间 flap

Richard de Grijs

专题命中 具身导航 :navigation(abstract)

AI总结 荷兰通过时间球和 flap 系统发展其航海时间信号,提升航海精度与科学现代化。

Comments 24 pages, incl. 15 figures. Accepted for publication in the Journal of Astronomical History and Heritage (December 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18978 2026-02-24 cs.HC 50%

Evaluating Replay Techniques for Asynchronous Task Handover in Immersive Analytics

评估异步任务交接在沉浸式分析中的回放技术

Zhengtai Gou, Junxiao Long, Tao Lu, Jian Zhao, Yalong Yang

专题命中 具身导航 :navigation(abstract)

AI总结 本文通过对比PC与VR平台,评估了沉浸式分析中异步任务交接的回放技术,发现VR环境下的沉浸式回放更有利于任务理解和流程重建。

Comments Accepted by IEEE VR 26

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 具身推理 6 篇

2602.19634 2026-02-24 cs.LG cs.AI stat.ML 86%

Compositional Planning with Jumpy World Models

基于跳跃世界模型的组合规划

Jesse Farebrother, Matteo Pirotta, Andrea Tirinzoni, Marc G. Bellemare, Alessandro Lazaric, Ahmed Touati

机构 * FAIR at Meta Mila -- Qu\'ebec AI Institute McGill University

专题命中 具身推理 :world model(title,abstract);manipulation(abstract);navigation(abstract);分类 cs.AI、cs.LG

AI总结 本文提出基于跳跃世界模型的组合规划方法,通过学习多步动态预测模型提升长时间任务的规划性能,实现复杂任务的解决。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19285 2026-02-24 cs.CV 79%

MRI Contrast Enhancement Kinetics World Model

MRI对比增强动力学世界模型

Jindi Kong, Yuting He, Cong Xia, Rongjun Ge, Shuo Li

机构 * Case Western Reserve University(凯斯西储大学) Jiangsu Cancer Hospital(江苏癌症医院) Southeast University(东南大学)

专题命中 具身推理 :world model(title,abstract);分类 cs.CV

AI总结 本文提出MRI CEKWorld模型,通过时空一致性学习解决MRI对比增强动力学建模中的时间连续性和内容一致性问题。

Comments Accepted by CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18739 2026-02-24 cs.LG 79%

When World Models Dream Wrong: Physical-Conditioned Adversarial Attacks against World Models

当世界模型做梦错误:针对世界模型的物理条件对抗攻击

Zhixiang Guo, Siyuan Liang, Andras Balogh, Noah Lunberry, Rong-Cheng Tu, Mark Jelasity, Dacheng Tao

机构 * Nanyang Technological University(南洋理工大学) University of Szeged(塞格德大学)

专题命中 具身推理 :world model(title,abstract);分类 cs.LG

AI总结 本文提出PhysCond-WMA攻击方法,通过扰动物理条件通道诱导世界模型的语义、逻辑或决策层面的扭曲,揭示生成式世界模型的安全性漏洞。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07854 2026-02-24 cs.CV 79%

Geometry-Aware Rotary Position Embedding for Consistent Video World Model

面向一致视频世界模型的几何感知旋转位置嵌入

Chendong Xiang, Jiajun Liu, Jintao Zhang, Xiao Yang, Zhengwei Fang, Shizun Wang, Zijun Wang, Yingtian Zou, Hang Su, Jun Zhu

机构 * Dept. of Comp. Sci. and Tech., Institute for AI, BNRist Center, THBI Lab, Tsinghua-Bosch Joint ML Center, Tsinghua University(清华大学计算机科学与技术系、人工智能研究院、BNRist中心、THBI实验室、清华-博世联合机器学习中心、清华大学) Gaoling School of Artificial Intelligence, Renmin University of China(北京人民大学人工智能学院) National University of Singapore(新加坡国立大学) Peking University(北京大学) School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院)

专题命中 具身推理 :world model(title,abstract);分类 cs.CV

AI总结 本文提出ViewRope和几何感知帧稀疏注意力,通过引入相机射线方向提升视频世界模型的长期一致性与效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19503 2026-02-24 cs.CV 57%

A Text-Guided Vision Model for Enhanced Recognition of Small Instances

一种基于文本引导的视觉模型,用于提升小实例的识别

Hyun-Ki Jung

专题命中 具身推理 :world model(abstract);分类 cs.CV

AI总结 本文提出了一种基于文本引导的改进YOLO-World模型,通过优化主干网络结构提升小物体检测精度和模型轻量化性能。

Comments Accepted for publication in Applied Computer Science (2026)

Journal ref Applied Computer Science, Vol. 22, No. 1, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.05744 2026-02-24 cs.NE 50%

Deep Learning: Our Miraculous Year 1990-1991

深度学习:我们的奇迹之年 1990-1991

Juergen Schmidhuber

专题命中 具身推理 :world model(abstract)

AI总结 本文回顾了1990-1991年作者在深度学习领域的突破性工作,奠定了生成式人工智能的基础,并对后续技术发展产生了深远影响。

Comments 52 pages, over 300 references, 38 illustrations, extending v1 of 4 Oct 2019

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 模仿学习与强化学习 11 篇

2602.18856 2026-02-24 cs.LG 85%

Issues with Measuring Task Complexity via Random Policies in Robotic Tasks

通过随机策略测量机器人任务复杂性的问题

Reabetswe M. Nkhumise, Mohamed S. Talamali, Aditya Gilra

机构 * University of Sheffield(谢菲尔德大学)

专题命中 模仿学习与强化学习 :robotic(title,abstract);robotics(abstract);manipulation(abstract);分类 cs.LG

AI总结 本文指出基于随机权重猜测的指标在测量机器人任务复杂性时存在矛盾,需发展更可靠的度量方法。

Comments 16 pages, 9 figures, The 25th International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏