arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

2026-01-16 至 2026-01-16 共收录 5 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 5 篇

2601.09812 2026-01-16 cs.CV cs.RO 84%

LCF3D: A Robust and Real-Time Late-Cascade Fusion Framework for 3D Object Detection in Autonomous Driving

LCF3D: 一种用于自动驾驶中3D物体检测的鲁棒且实时的后期级联融合框架

Carlo Sgaravatti, Riccardo Pieroni, Matteo Corno, Sergio M. Savaresi, Luca Magri, Giacomo Boracchi

专题命中 感知 :autonomous driving(title,abstract);LiDAR(abstract);分类 cs.RO、cs.CV

AI总结 LCF3D通过结合RGB图像和LiDAR点云的后期级联融合方法,提升自动驾驶中3D物体检测的鲁棒性和实时性。

Comments 35 pages, 14 figures. Published at Pattern Recognition

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09952 2026-01-16 cs.CV cs.RO 62%

OT-Drive: Out-of-Distribution Off-Road Traversable Area Segmentation via Optimal Transport

OT-Drive: 基于最优传输的离群分布越野可通行区域分割

Zhihua Zhao, Guoqiang Li, Chen Min, Kangping Lu

机构 * School of Mechanical Engineering, Beijing Institute of Technology(北京理工大学机械工程学院) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) Shandong Pengxiang Automobile Co., Ltd(山东鹏翔汽车有限公司)

专题命中 感知 :autonomous driving(abstract);分类 cs.RO、cs.CV

AI总结 OT-Drive通过最优传输方法实现离群分布越野可通行区域分割,提升自动驾驶在复杂环境下的鲁棒性和泛化能力。

Comments 9 pages, 8 figures, 6 tables. This work has been submitted to the IEEE for possible publication. Code will be released upon acceptance

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08355 2026-01-16 cs.CV 57%

Semantic Misalignment in Vision-Language Models under Perceptual Degradation

视觉-语言模型在感知退化下的语义错位

Guo Cheng

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 本研究探讨了视觉-语言模型在感知退化下的语义错位问题,发现传统分割指标的下降并未影响下游行为,揭示了像素鲁棒性与多模态语义可靠性之间的脱节。

Comments 10 pages, 4 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22972 2026-01-16 cs.CV eess.SP 57%

Wavelet-based Multi-View Fusion of 4D Radar Tensor and Camera for Robust 3D Object Detection

基于小波的4D雷达张量与相机多视图融合用于鲁棒3D目标检测

Runwei Guan, Jianan Liu, Shaofeng Liang, Fangqiang Ding, Shanliang Yao, Xiaokai Bai, Daizong Liu, Tao Huang, Guoqiang Mao, Hui Xiong

机构 * Thrust of Artificial Intelligence, Hong Kong University of Science and Technology (Guangzhou)(人工智能 thrust,香港科技大学(广州)) Momoniai AI Department of Mechanical Engineering, Massachusetts Institute of Technology(机械工程系,麻省理工学院) School of Information Engineering, Yancheng Institute of Technology(信息工程学院,盐城科技学院) College of Information Science and Electronic Engineering, Zhejiang University(信息科学与电子工程学院,浙江大学) Institute for Math & AI, Wuhan University(数学与人工智能研究所,武汉大学) College of Science and Engineering and the Centre for AI and Data Science Innovation, James Cook University(科学与工程学院及人工智能与数据科学创新中心,詹姆斯库克大学) School of Transportation, Southeast University(交通运输学院,东南大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 WRCFormer通过小波注意力模块和几何引导渐进融合机制,高效融合4D雷达张量与相机图像,提升3D目标检测在恶劣天气下的鲁棒性。

Comments 10 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13601 2026-01-16 cs.CV 57%

Unleashing Semantic and Geometric Priors for 3D Scene Completion

释放语义和几何先验以实现3D场景补全

Shiyuan Chen, Wei Sui, Bohao Zhang, Zeyd Boukhers, John See, Cong Yang

机构 * D-Robotics

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 FoundationSSC通过双解耦机制和轴感知融合模块,提升3D场景补全的语义和几何指标表现。

Comments Accept by AAAI-2026

详情

展开后加载摘要…

URL PDF HTML 收藏