arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Robotics and Automation · 会议 · Robotics

2026-05-26 至 2026-05-26 共收录 12
2605.25942 2026-05-26 cs.CV cs.RO

LRDDv3: High-Resolution Long-Range Drone Detection Dataset with Range Information and Thermal Data

LRDDv3:具有距离信息和热数据的高分辨率远程无人机检测数据集

Knut Peterson, Zaid Mayers, Azmain Yousuf, Priontu Chowdhury, Asher Zaczepinski, Solmaz Arezoomandan, Reihaneh Maarefdoust, David Han

机构 * iMaPLe Research Lab, Drexel University(Drexel大学iMaPLe研究实验室) University of Maine(缅因大学)

AI总结 提出LRDDv3数据集,包含102,532张高分辨率远程RGB图像和29,630张配对IR图像,支持远程无人机检测,提供距离信息。

Comments 8 pages, 5 figures. Accepted to the 2026 IEEE International Conference on Robotics and Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.25790 2026-05-26 cs.RO

HoLoArm: Deformable Arms for Collision-Tolerant Quadrotor Flight

HoLoArm: 用于碰撞容忍四旋翼飞行的可变形臂

Quang Ngoc Pham, Jonas Eschmann, Yang Zhou, Alejandro Ojeda Olarte, Giuseppe Loianno, Van Anh Ho

机构 * Japan Advanced Institute of Science and Technology(日本先进科学技术研究所) University of California Berkeley(加州大学伯克利分校) New York University(纽约大学)

AI总结 受蜻蜓翅膀结脉结构启发,提出具有柔性臂的四旋翼HoLoArm,结合强化学习控制策略实现被动变形与快速恢复,在高达7.6 m/s碰撞速度下保持稳定飞行。

Comments 8 pages, 15 figures, 1 table, Accepted at the IEEE Robotics and Automation Letters (RA-L) and the IEEE International Conference on Robotics and Automation (ICRA), 2026

Journal ref IEEE Robotics and Automation Letters, vol. 11, no. 3, pp. 3582-3589, March 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.25262 2026-05-26 cs.CV

Semantics-Guided Multimodal Masked Autoencoder Pretraining for 3D BEV Object Detection

语义引导的多模态掩码自编码器预训练用于3D BEV目标检测

Prabuddhi Wariyapperuma, Rajitha de Silva, Marc Hanheide, Thomas Bohné, Leonardo Guevara

机构 * University of Lincoln, Lincoln Centre for Autonomous Systems(林肯大学,林肯自主系统中心) University of Cambridge, Institute for Manufacturing, Department of Engineering(剑桥大学,制造研究所,工程系)

AI总结 提出语义引导的多模态掩码自编码器框架,通过语义引导的LiDAR体素掩码和辅助点语义解码分支,在预训练中注入语义信息,提升3D BEV目标检测性能。

Comments Accepted at the ICRA 2026 Workshop on Semantics for Reliable Robot Autonomy (SRRA) as a lightning talk and poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06218 2026-05-26 cs.RO

Few-Shot Neural Differentiable Simulator: Real-to-Sim Rigid-Contact Modeling

少样本神经可微模拟器:真实到模拟的刚体接触建模

Zhenhao Huang, Siyuan Luo, Bingyang Zhou, Ziqiu Zeng, Jason Pho, Fan Shi

机构 * National University of Singapore(新加坡国立大学) Prana Lab(Prana实验室)

AI总结 提出一种结合解析公式物理一致性与图神经网络表示能力的少样本真实到模拟方法,通过少量真实数据校准解析模拟器生成大规模合成数据集,并引入基于网格的图神经网络隐式建模刚体前向动力学及碰撞检测的代理梯度,实现完全可微性,从而提升模拟保真度和策略学习效率。

Comments Accepted in ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.20473 2026-05-26 cs.RO

Data-Driven Optimization of Tactile Sensor Configurations for Efficient Dexterous Manipulation

数据驱动的触觉传感器配置优化以实现高效灵巧操作

Haoran Guo, Haoyang Wang, Zhengxiong Li, He Bai, Lingfeng Tao

机构 * ShanghaiTech University, School of Information Science and Technology(上海科技大学信息科学与技术学院) University of Alberta(阿尔伯塔大学) Oklahoma State University(俄克拉荷马州立大学) University of Colorado Denver, Department of Computer Science and Engineering(科罗拉多大学丹佛分校计算机科学与工程系) Department of Robotics and Mechatronics Engineering, Kennesaw State University(凯斯西储大学机器人与机电工程系)

AI总结 提出两阶段框架量化触觉传感器对深度强化学习策略的贡献,将Shadow Hand传感器从92个减少至14个仍保持90%以上性能,并发现中指传感器具有负贡献。

Comments This work has been submitted to the ICRA for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24690 2026-05-26 cs.RO cs.LG

Sum of Costs Diffusion with Dynamic Guidance for Motion Planning

运动规划的动态引导代价和扩散模型

Aysu Aylin Kaplan, Özgür Erkent

机构 * Computer Engineering Department, Hacettepe University(哈切特佩大学计算机工程系)

AI总结 提出一种基于扩散模型的高泛化运动规划方法,通过总碰撞代价梯度引导去噪过程并动态选择引导起始步,在Mπnets数据集上取得最优性能。

Comments Accepted at the Frontiers of Optimization for Robotics Workshop at the IEEE International Conference of Robotics & Automation (ICRA), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24622 2026-05-26 cs.RO cs.CV

PoseRefer: Pathway-Local Parameters for Semantically Grounded Reference Resolution

PoseRefer: 用于语义基础指代消解的通路-局部参数

Anna Deichler

机构 * KTH Royal Institute of Technology(皇家理工学院)

AI总结 提出PoseRefer架构,通过解耦姿态和文本通路并冻结MiniLM类别嵌入,在MM-Conv数据集上实现31.9%的top-1准确率,并揭示融合准确性可能受类别表示伪影影响。

Comments ICRA 2026 Workshop on Semantics for Reliable Robot Autonomy: From Environment Understanding and Reasoning to Safe Interaction

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.17538 2026-05-26 cs.RO

Novel Algorithms for Smoothly Differentiable and Efficiently Vectorizable Contact Manifold Construction

用于光滑可微且高效可向量化的接触流形构建的新算法

Onur Beker, Andreas René Geist, Anselm Paulus, Georg Martius

机构 * University of Tübingen(图宾根大学)

AI总结 针对接触丰富场景中机器人行为优化,提出一种以光滑二次可微性和GPU大规模可向量化为优先的新碰撞检测流水线,包括可微SDF表示、宽/窄阶段例程和凸分解接触融合。

Comments This version adds late-breaking results in preparation for the CR2 workshop in ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24449 2026-05-26 cs.RO cs.LG

Vision-Guided Outdoor Flight and Obstacle Evasion via Reinforcement Learning

基于强化学习的视觉引导户外飞行与避障

Shiladitya Dutta, Aayush Gupta, Varun Saran, Avideh Zakhor

机构 * College of Engineering, Department of Electrical Engineering and Computer Science, University of California Berkeley(加州大学伯克利分校工程学院电气工程与计算机科学系)

AI总结 提出一种基于立体视觉深度和视觉惯性里程计的传感器运动策略,通过强化学习和特权学习在仿真中训练,实现零样本迁移到未知户外环境和无人机平台进行自主避障导航。

Comments Published in IEEE Robotics and Automation Letters, vol 11, no 2. Presented at the IEEE International Conference on Robotics and Automation 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24111 2026-05-26 cs.RO cs.AI

MASt3R-Nav: WayPixel Navigation in Relative 3D Maps

MASt3R-Nav: 相对3D地图中的WayPixel导航

Vansh Garg, Rohit Jayanti, Krish Pandya, Sarthak Chittawar, Siddharth Tourani, Muhammad Haris Khan, Sourav Garg, Madhava Krishna

机构 * Robotics Research Center, IIIT-Hyderabad, India(1 罗斯科技研究中心,IIIT-海得拉巴,印度) University of Heidelberg(2 海德堡大学) MBZUAI(3 MBZUAI)

AI总结 提出一种基于像素相对连接性的地图表示,通过相对3D坐标系中的像素对应构建地图,并利用像素级图进行全局路径规划,训练控制器预测轨迹,实现高精度导航。

Comments 2026 IEEE International Conference on Robotics & Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24074 2026-05-26 cs.CV cs.RO

WideDepth: Millimeter-Accurate Benchmark for Fisheye Depth Estimation

WideDepth: 用于鱼眼深度估计的毫米级精度基准

Ilia Indyk, Ignat Penshin, Ivan Sosin, Maxim Monastyrny, Aleksei Valenkov, Ilya Makarov

机构 * Robotics Center(机器人中心) AXXX Trusted AI Research Center, RAS(可信人工智能研究中心,俄罗斯科学院)

AI总结 提出首个室内鱼眼深度估计数据集WideDepth,包含101个场景的5K高分辨率立体对和毫米级真值,并引入基于LiDAR的立体鱼眼图像生成方法,评估多种模型,微调后性能提升高达62%。

Comments Accepted to IEEE International Conference on Robotics and Automation (ICRA) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09599 2026-05-26 cs.CV

BridgeTA: Bridging the Representation Gap in Knowledge Distillation via Teacher Assistant for Bird's Eye View Map Segmentation

BridgeTA: 通过教师助手弥合知识蒸馏中表示差距的鸟瞰图分割

Beomjun Kim, Suhan Woo, Sejong Heo, Euntai Kim

机构 * Yonsei University(延世大学) Hyundai Motor Company(现代汽车公司) Korea Institute of Science and Technology(韩国科学技术院)

AI总结 提出BridgeTA框架,利用教师助手网络在保持学生模型架构和推理成本不变的情况下,弥合激光雷达-相机融合与纯相机模型之间的表示差距,并通过Young不等式推导蒸馏损失实现稳定优化,在nuScenes数据集上mIoU提升4.2%。

Comments Accepted at ICRA 2026 (8 pages, 6 figures)

详情

展开后加载摘要…

URL PDF HTML 收藏