arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

2026-01-15 至 2026-01-15 共收录 13 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 具身导航 13 篇

2601.08868 2026-01-15 cs.CV cs.AI cs.RO 85%

Residual Cross-Modal Fusion Networks for Audio-Visual Navigation

残差跨模态融合网络用于音频视觉导航

Yi Wang, Yinfeng Yu, Bin Ren

机构 * School of Computer Science and Technology, Xinjiang University, Urumqi, China(新疆大学计算机科学与技术学院) Joint International Research Laboratory of Silk Road Multilingual Cognitive(丝绸之路多语认知联合国际实验室) School of Mechatronic Engineering and Automation Shanghai University, Shanghai, China(上海大学机电工程与自动化学院)

专题命中 具身导航 :navigation(title,abstract);embodied agent(abstract);分类 cs.RO、cs.AI、cs.CV

AI总结 本文提出残差跨模态融合网络,通过双向残差交互实现音频视觉信息互补建模,提升跨域导航性能。

Comments Main paper (10 pages). Accepted for publication by the 14th international conference on Computational Visual Media (CVM 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08953 2026-01-15 cs.RO cs.AI 84%

Fairness risk and its privacy-enabled solution in AI-driven robotic applications

人工智能驱动机器人应用中的公平性风险及其隐私增强解决方案

Le Liu, Bangguo Yu, Nynke Vellinga, Ming Cao

专题命中 具身导航 :robotic(title,abstract);navigation(abstract);分类 cs.RO、cs.AI

AI总结 本文提出了一种基于效用的公平性度量标准,用于机器人决策,并分析公平性与隐私的联合关系,通过机器人导航任务验证了隐私预算对公平性的联合影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09318 2026-01-15 cs.RO cs.SY eess.SY 79%

Feedback-Based Mobile Robot Navigation in 3-D Environments Using Artificial Potential Functions Technical Report

基于反馈的三维环境移动机器人导航人工势场方法技术报告

Ro'i Lang, Elon Rimon

专题命中 具身导航 :navigation(title,abstract);分类 cs.RO

AI总结 本文提出了一种用于三维环境中移动机器人导航的人工势场方法,通过多项式函数构造并验证了在存在障碍物时的导航性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09248 2026-01-15 cs.CV cs.AI 73%

Hybrid guided variational autoencoder for visual place recognition

混合引导变分自编码器用于视觉地点识别

Ni Wang, Zihan You, Emre Neftci, Thorben Schoepe

机构 * Amazon Development Center Germany GmbH, Berlin, Germany(亚马逊德国开发中心) Southeast University, Nanjing, China(东南大学) Forschungszentrum Jülich GmbH, Aachen, Germany(尤利希研究中心) imec, Luven, Belgium(imec)

专题命中 具身导航 :robotics(abstract);navigation(abstract);分类 cs.AI、cs.CV

AI总结 本文提出一种混合引导变分自编码器用于视觉地点识别,结合事件视觉传感器和脉冲神经网络,实现紧凑、鲁棒且具有泛化能力的模型,提升移动机器人导航性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09578 2026-01-15 cs.RO cs.CV 62%

Multimodal Signal Processing For Thermo-Visible-Lidar Fusion In Real-time 3D Semantic Mapping

多模态信号处理用于热-可见-激光雷达融合的实时3D语义制图

Jiajun Sun, Yangyi Ou, Haoyuan Zheng, Chao yang, Yue Ma

机构 * College of Mechatronics and Control Engineering, Shenzhen University(深圳大学机械与控制工程学院) School of Robotics, Xi’an-Jiaotong Liverpool University(西安交通大学利物浦大学机器人学院)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 本文提出通过多模态信号处理融合热、可见和激光雷达数据,提升实时3D语义制图的精度与语义理解能力,适用于灾害评估和工业维护等场景。

Comments 5 pages,7 figures. Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07808 2026-01-15 cs.RO 61%

AcoustoBots: A swarm of robots for acoustophoretic multimodal interactions

AcoustoBots:用于声学电势多模交互的机器人群

Narsimlu Kemsaram, James Hardwick, Jincheng Wang, Bonot Gautam, Ceylan Besevli, Giorgos Christopoulos, Sourabh Dogra, Lei Gao, Akin Delibasi, Diego Martinez Plasencia, Orestis Georgiou, Marianna Obrist, Ryuji Hirayama, Sriram Subramanian

专题命中 具身导航 :robotic(abstract);分类 cs.RO;robotics(journal_ref)

AI总结 AcoustoBots通过可移动相位阵列和BeadDispenserBot实现声学电势多模交互,提升机器人群的灵活性和交互能力。

Journal ref Frontiers in Robotics and AI, 12:1537101, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09665 2026-01-15 cs.CV 57%

SCE-SLAM: Scale-Consistent Monocular SLAM via Scene Coordinate Embeddings

SCE-SLAM:通过场景坐标嵌入实现尺度一致的单目SLAM

Yuchen Wu, Jiahe Li, Xiaohan Yu, Lina Yu, Jin Zheng, Xiao Bai

机构 * School of Computer Science and Engineering, State Key Laboratory of Complex & Critical Software Environment, Jiangxi Research Institute, Beihang University(计算机科学与工程学院、复杂与关键软件环境国家重点实验室、江西研究院、北航) Macquarie University(麦考瑞大学) Beijing Key Laboratory of Semiconductor Neural Network Intelligent Sensing and Computing Technology(北京半导体神经网络智能感知与计算技术重点实验室)

专题命中 具身导航 :navigation(abstract);分类 cs.CV

AI总结 SCE-SLAM通过场景坐标嵌入实现单目SLAM的尺度一致性,实验显示在KITTI上轨迹误差减少8.36米,同时保持36 FPS。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03547 2026-01-15 cs.RO 57%

Shape-Space Graphs: Fast and Collision-Free Path Planning for Soft Robots

形状空间图:用于软机器人快速且无碰撞路径规划

Carina Veil, Moritz Flaschel, Ellen Kuhl

机构 * Department of Mechanical Engineering, Stanford University(机械工程系,斯坦福大学) Institute of Applied Mechanics, Friedrich-Alexander-Universität Erlangen–Nürnberg(应用力学研究所,弗赖堡-埃尔兰根-纽伦堡大学)

专题命中 具身导航 :robotics(abstract);分类 cs.RO

AI总结 本文提出了一种基于形状空间图的路径规划方法,用于软机器人在复杂环境中的快速且无碰撞导航,通过预计算形状库和多目标边成本实现高效路径生成。

Comments revised version

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20612 2026-01-15 cs.LG 57%

Policy Compatible Skill Incremental Learning via Lazy Learning Interface

通过懒惰学习接口实现策略兼容的技能增量学习

Daehee Lee, Dongsu Lee, TaeYoon Kwack, Wonje Choi, Honguk Woo

机构 * Sungkyunkwan University(成均馆大学) University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 具身导航 :embodied agent(abstract);分类 cs.LG

AI总结 本文提出SIL-C框架,通过懒惰学习接口实现技能与策略的兼容性,提升下游任务性能无需重新训练策略。

Comments NeurIPS 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09230 2026-01-15 cs.CV 57%

CLIDD: Cross-Layer Independent Deformable Description for Efficient and Discriminative Local Feature Representation

CLIDD: 跨层独立可变形描述用于高效且判别性的局部特征表示

Haodi Yao, Fenghua He, Ning Hao, Yao Su

机构 * School of Astronautics, Harbin Institute of Technology, China(航天学院,哈尔滨工业大学,中国) State Key Laboratory of General Artificial Intelligence, Beijing Institute for General Artificial Intelligence (BIGAI), China(通用人工智能国家重点实验室,北京通用人工智能研究院(BIGAI),中国)

专题命中 具身导航 :navigation(abstract);分类 cs.CV

AI总结 CLIDD通过跨层独立可变形描述实现高效且判别性的局部特征表示,提供高精度匹配与低计算开销的解决方案。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23459 2026-01-15 cs.IR cs.AI 57%

KLAN: Kuaishou Landing-page Adaptive Navigator

快手落地页自适应导航

Fan Li, Chang Meng, Jiaqi Fu, Shuchang Liu, Jiashuo Zhang, Tianke Zhang, Xueliang Wang, Xiaoqiang Feng

机构 * Duke University(杜克大学)

专题命中 具身导航 :navigation(abstract);分类 cs.AI

AI总结 KLAN通过分层框架实现个性化落地页导航,提升用户活跃度和满意度。

Comments We propose PLPM, a new task for selecting optimal landing pages upon user entry. Our solution, KLAN, models static and dynamic user interests and is successfully deployed on Kuaishou, improving DAU and user lifetime

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09632 2026-01-15 cs.HC 50%

Perceptually-Guided Adjusted Teleporting: Perceptual Thresholds for Teleport Displacements in Virtual Environments

基于感知的调整式传送:虚拟环境中传送位移的感知阈值

Rose Connolly, Victor Zordan, Rachel McDonnell

专题命中 具身导航 :navigation(abstract)

AI总结 本研究探讨了虚拟环境中传送位移的感知阈值,发现传送目的地可被无察觉调整,且向后调整容忍度更高,为适应性VR移动系统提供了新机遇。

Comments 9 pages. to be published in IEEE VR conference proceedings 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09359 2026-01-15 cond-mat.mtrl-sci 50%

Field report from Collaborative Research Center 1625: Heterogeneous research data management using ontology representations

协同研究中心1625的现场报告:使用本体表示进行异构研究数据管理

Doaa Mohamed, Samuel García Vázquez, Behnam Mardani, Victor Dudarev, Alfred Ludwig, Maribel Acosta, Markus Stricker

专题命中 具身导航 :navigation(abstract)

AI总结 本研究提出了一种基于本体表示的异构研究数据管理系统,用于管理多学科复杂材料数据,通过知识图谱和数据库结合实现数据的有效整合与应用。

Comments 22 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏