arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

2026-02-02 至 2026-02-02 共收录 15 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 机器人数据与评测 15 篇

2506.02883 2026-02-02 cs.LG 83%

A Continual Offline Reinforcement Learning Benchmark for Navigation Tasks

连续离线强化学习导航任务基准测试

Anthony Kobanda, Odalric-Ambrym Maillard, Rémy Portelas

机构 * Inria, Univ. Lille, CNRS, Centrale Lille, UMR 9198-CRIStAL, F-59000 Lille, France(法国里尔大学、法国国家科学研究中心、中央里尔学院、UMR 9198-CRIStAL)

专题命中 机器人数据与评测 :navigation(title,abstract);robotics(abstract);分类 cs.LG

AI总结 本文提出一个连续离线强化学习导航任务基准测试,通过视频游戏场景评估算法性能,解决灾难性遗忘、任务适应和内存效率等关键挑战。

Comments arXiv admin note: text overlap with arXiv:2412.14865

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22018 2026-02-02 cs.RO 70%

PocketDP3: Efficient Pocket-Scale 3D Visuomotor Policy

PocketDP3: 高效的微型3D视觉-运动政策

Jinhao Zhang, Zhexuan Zhou, Huizhe Li, Yichen Lai, Wenlong Xia, Haoming Song, Youmin Gong, Jie Mei

专题命中 机器人数据与评测 :manipulation(abstract);robotic(abstract);分类 cs.RO

AI总结 PocketDP3通过轻量级扩散混合器替代传统解码器,实现高效3D视觉-运动策略,减少参数消耗并提升实时部署性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10961 2026-02-02 cs.RO 70%

EquiContact: A Hierarchical SE(3) Vision-to-Force Equivariant Policy for Spatially Generalizable Contact-rich Tasks

EquiContact: 一种用于空间通用接触密集任务的层次SE(3)视觉到力等价策略

Joohwan Seo, Arvind Kruthiventy, Soomi Lee, Megan Teng, Seoyeon Choi, Xiang Zhang, Jongeun Choi, Roberto Horowitz

专题命中 机器人数据与评测 :manipulation(abstract);robotic(abstract);分类 cs.RO

AI总结 EquiContact通过分层策略和等价性设计,实现空间通用的接触密集任务视觉到力控制。

Comments Submitted to RSS

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12138 2026-02-02 cs.CV cs.LG cs.RO 67%

Monocular pose estimation of articulated open surgery tools -- in the wild

单目手术工具姿态估计——在真实环境中

Robert Spektor, Tom Friedman, Itay Or, Gil Bolotin, Shlomi Laufer

机构 * Faculty of Data and Decision Sciences, Technion - Israel Institute of Technology, Haifa, Israel(数据与决策科学学院,技术Ion-以色列理工学院,海法,以色列)

专题命中 机器人数据与评测 :robotic(abstract);分类 cs.RO、cs.CV、cs.LG

AI总结 本文提出了一种基于合成数据和真实数据的单目手术工具姿态估计框架,通过域适应和伪标签提升实际应用性能。

Comments Author Accepted Manuscript (AAM)

Journal ref Medical Image Analysis, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01987 2026-02-02 cs.RO cs.CV 62%

CaLiV: LiDAR-to-Vehicle Calibration of Arbitrary Sensor Setups

CaLiV:任意传感器配置的LiDAR到车辆校准

Ilir Tahiraj, Markus Edinger, Dominik Kulmer, Markus Lienkamp

机构 * Technical University of Munich(慕尼黑技术大学)

专题命中 机器人数据与评测 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 CaLiV提出了一种基于目标的校准技术,用于多LiDAR系统的传感器到传感器和传感器到车辆的校准,适用于非重叠视野范围,无需外部设备,实现了高精度的平移和旋转误差校正。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22467 2026-02-02 cs.RO cs.CV 62%

CARE: Multi-Task Pretraining for Latent Continuous Action Representation in Robot Control

CARE:用于机器人控制的潜在连续动作表示的多任务预训练

Jiaqi Shi, Xulong Zhang, Xiaoyang Qu, Jianzong Wang

机构 * Ping An Technology (Shenzhen) Co., Ltd.(平安科技(深圳)有限公司) University of Science and Technology of China(中国科学技术大学)

专题命中 机器人数据与评测 :robotic(abstract);分类 cs.RO、cs.CV

AI总结 CARE通过多任务预训练学习连续动作表示,提升机器人控制的可扩展性和可解释性。

Comments Accepted to 2026 IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22849 2026-02-02 cs.RO math.OC 61%

Robust Rigid Body Assembly via Contact-Implicit Optimal Control with Exact Second-Order Derivatives

通过接触隐式最优控制实现稳健刚体装配

Christian Dietz, Sebastian Albrecht, Gianluca Frison, Moritz Diehl, Armin Nurkanović

机构 * Autonomous Systems and Control, Siemens AG, Germany(自动化系统与控制,西门子股份有限公司,德国) Department of Microsystems Engineering (IMTEK), University of Freiburg, Germany(微系统工程系(IMTEK),弗赖堡大学,德国) Department of Mathematics, University of Freiburg, Germany(数学系,弗赖堡大学,德国)

专题命中 机器人数据与评测 :robotics(abstract,comments);分类 cs.RO

AI总结 本文提出了一种基于接触隐式最优控制的稳健刚体装配方法,通过高效利用导数信息减少物理模拟步骤,实现实验中99%的成功率。

Comments Submitted to Transactions on Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13644 2026-02-02 cs.RO 61%

On Your Own: Pro-level Autonomous Drone Racing in Uninstrumented Arenas

自主驾驶:在无仪器赛场上的专业级无人机竞速

Michael Bosello, Flavio Pinzarrone, Sara Kiade, Davide Aguiari, Yvo Keuter, Aaesha AlShehhi, Gyordan Caminati, Kei Long Wong, Ka Seng Chou, Junaid Halepota, Fares Alneyadi, Jacopo Panerati, Giovanni Pau

机构 * Autonomous Robotics Research Center, Technology Innovation Institute(自主机器人研究中心、技术创新院)

专题命中 机器人数据与评测 :navigation(abstract);分类 cs.RO;robotics(journal_ref)

AI总结 该研究提出了一种在无仪器赛场中实现专业级无人机竞速的自主系统,通过受控环境与挑战性环境的对比验证了其性能。

Journal ref IEEE Robotics and Automation Letters, vol. 11, no. 3, pp. 2674-2681, March 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.02667 2026-02-02 cs.RO 61%

Race Against the Machine: a Fully-annotated, Open-design Dataset of Autonomous and Piloted High-speed Flight

与机器赛跑:一个完全标注、开放设计的自主与驾驶高速飞行数据集

Michael Bosello, Davide Aguiari, Yvo Keuter, Enrico Pallotta, Sara Kiade, Gyordan Caminati, Flavio Pinzarrone, Junaid Halepota, Jacopo Panerati, Giovanni Pau

机构 * Autonomous Robotics Research Center of the Technology Innovation Institute(技术创新研究所自主机器人研究中心) University of Bologna(博洛尼亚大学)

专题命中 机器人数据与评测 :robotics(abstract,journal_ref);分类 cs.RO

AI总结 本文提出一个开放设计的自主与驾驶高速飞行数据集,用于支持无人机赛车研究,提供高精度飞行数据和标注信息,促进相关技术的发展和比较。

Comments 8 pages, 7 figures

Journal ref IEEE Robotics and Automation Letters, vol. 9, no. 4, pp. 3799-3806, April 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22878 2026-02-02 eess.IV cs.CV 57%

Development of Domain-Invariant Visual Enhancement and Restoration (DIVER) Approach for Underwater Images

水下图像领域不变视觉增强与恢复(DIVER)方法的开发

Rajini Makam, Sharanya Patil, Dhatri Shankari T M, Suresh Sundaram, Narasimhan Sundararajan

机构 * Department of Aerospace Engineering, Indian Institute of Science(航空航天工程系,印度科学院)

专题命中 机器人数据与评测 :robotic(abstract);分类 cs.CV

AI总结 DIVER通过整合经验校正与物理引导建模,实现水下图像的领域不变增强与恢复,显著提升视觉质量和机器人感知性能。

Comments Submitted to IEEE Journal of Oceanic Engineering

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13218 2026-02-02 cs.CV 57%

ObjectVisA-120: Object-based Visual Attention Prediction in Interactive Street-crossing Environments

ObjectVisA-120: 交互式过街环境中的基于物体的视觉注意力预测

Igor Vozniak, Philipp Mueller, Nils Lipp, Janis Sprenger, Konstantin Poddubnyy, Davit Hovhannisyan, Christian Mueller, Andreas Bulling, Philipp Slusallek

机构 * German Research Center for Artificial Intelligence (DFKI) GmbH(德国人工智能研究中心(DFKI)) Max Planck Institute for Intelligent Systems(智能系统马克斯·普朗克研究所) Institute for Visualization and Interactive Systems (VIS) at Stuttgart University(斯图加特大学可视化与交互系统研究所)

专题命中 机器人数据与评测 :navigation(abstract);分类 cs.CV

AI总结 ObjectVisA-120提出了一种新的虚拟现实数据集和基于物体的相似性指标,用于改进视觉注意力预测模型。

Comments Accepted for publication at the IEEE Intelligent Vehicles Symposium (IV), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22529 2026-02-02 cs.CV 57%

SHED Light on Segmentation for Dense Prediction

SHED:用于密集预测的分割光照

Seung Hyun Lee, Sangwoo Mo, Stella X. Yu

机构 * University of Michigan(密歇根大学)

专题命中 机器人数据与评测 :robotics(abstract);分类 cs.CV

AI总结 SHED通过整合分割到密集预测中,提出了一种新的编码器-解码器架构,以提升深度边界锐度、分割连贯性和3D重建质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22322 2026-02-02 cs.LG eess.SP 57%

Spatially-Adaptive Conformal Graph Transformer for Indoor Localization in Wi-Fi Driven Networks

空间自适应符合图变换器用于Wi-Fi驱动网络中的室内定位

Ayesh Abu Lehyeh, Anastassia Gharib, Safwan Wshah

机构 * Department of Computer Science, The University of Vermont(佛罗里达大学计算机科学系) Department of Computer Science & Engineering, American University of Sharjah(阿曼大学计算机科学与工程系)

专题命中 机器人数据与评测 :navigation(abstract);分类 cs.LG

AI总结 本文提出SAC-GT框架,通过空间自适应符合预测方法提升Wi-Fi驱动网络中室内定位的精度与可靠性。

Comments Accepted to IEEE ICC 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08198 2026-02-02 cs.GR cs.RO 57%

Emergent morphogenesis via planar fabrication enabled by a reduced model of composites

通过平面制造实现的新兴形态生成:由复合材料简化模型驱动

Yupeng Zhang, Adam Alon, M. Khalid Jawed

专题命中 机器人数据与评测 :robotics(abstract);分类 cs.RO

AI总结 本文提出了一种简化模型,通过平面制造实现可编程的三维形态生成,利用双层系统中的应变不匹配产生多种三维结构。

Comments GitHub repository: https://github.com/StructuresComp/discrete-shells-shrinky-dink/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15804 2026-02-02 cs.SD eess.AS 50%

CompSpoof: A Dataset and Joint Learning Framework for Component-Level Audio Anti-spoofing Countermeasures

CompSpoof:用于组件级音频反欺骗的数据库和联合学习框架

Xueping Zhang, Yechen Wang, Linxi Li, Liwei Jin, Ming Li

机构 * Suzhou Municipal Key Laboratory of Multimodal Intelligent Systems(多模态智能系统苏州市重点实验室) Digital Innovation Research Center(数字创新研究中心) Duke Kunshan University(杜克昆山大学) OfSpectrum, Inc.(OfSpectrum公司)

专题命中 机器人数据与评测 :manipulation(abstract)

AI总结 CompSpoof提出了一种用于组件级音频反欺骗的数据库和联合学习框架,通过分离音频组件并分别检测欺骗来提升反欺骗性能。

Comments accepted at ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏